AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
84,548 results
4 Aug 2026

MIDAL: Math Image Descriptions for Accessible Learning

TutorialsDGX agent

arXiv:2608.00868v1 Announce Type: new Abstract: Many open educational resources are lacking in accessibility, especially in-depth image descriptions. In subjects like Science and Mathematics, however,

MIEScore: Human-Aligned Evaluation for Multi-Source Image Editing

Model ReleasesDGX agent

arXiv:2608.02059v1 Announce Type: new Abstract: Recent advances in unified multimodal models have significantly improved text-guided image editing abilities. In particular, models such as Nano-Banana-

Mind the Gap: Zero-Query Jailbreaks via Filter-Generator Discrepancy in Text-to-Image Systems

SafetyDGX agent

arXiv:2608.00973v1 Announce Type: new Abstract: Text-to-image (T2I) systems typically have prompt-level safety filters before the generator to block unsafe requests, yet such systems remain vulnerable

Content type
AllBlogX PostPaperYouTubeRedditGitHub

MiniWorld: Democratizing the Training of Video World Models from Scratch

HardwareDGX agent

arXiv:2608.01127v1 Announce Type: new Abstract: Video world models predict future observations conditioned on historical observations and control signals, enabling long-horizon generation through auto

Minute-Scale Training for Microrobot Navigation

Model ReleasesDGX agent

arXiv:2608.00854v1 Announce Type: new Abstract: Microrobots hold significant potential for various applications, where targeted navigation is a basic requirement. Deep reinforcement learning (DRL) has

Mistral releases Shieldstral, a 3B multimodal safety classifier that it says matches models up to 7x its size on text safety, available under Apache 2.0 (Mistral AI Blog)

Model ReleasesDGX agent

Mistral AI Blog: Mistral releases Shieldstral, a 3B multimodal safety classifier that it says matches models up to 7x its size on text safety, available under Apache 2.0 — Every product that ships a m

Mitigating Backdoors via Decoy Shortcuts and Knowledge Decoupling

Model ReleasesDGX agent

arXiv:2608.00732v1 Announce Type: cross Abstract: Backdoor attacks pose a serious threat to deep neural networks, especially when training relies on third-party data, allowing adversaries to inject ma

Mitigating Visual Degradation in MLLMs via Spatial-Spectral Visual Anchor Learning

SafetyDGX agent

arXiv:2608.01635v1 Announce Type: new Abstract: Despite the progress of multimodal large language models (MLLMs), they continue to exhibit deficiencies in visual perception. Following visual instructi

Mitigating Visual Hallucinations in Multimodal Systems through Retrieval-Augmented Reliability-Aware Inference

SafetyDGX agent

arXiv:2606.15782v2 Announce Type: replace-cross Abstract: Multimodal large language models (MLLMs) have demonstrated strong capabilities in vision-language understanding and natural-language response

MixedComplementarityProblems.jl: A Fast, Batched, Open-Source Interior Point Solver for Mixed Complementarity Problems

Model ReleasesDGX agent

arXiv:2608.00959v1 Announce Type: cross Abstract: Mixed complementarity problems (MCPs) arise as the first-order optimality conditions of nonlinear programs and noncooperative games, and provide a nat

MMPhysVideo: Physically Plausible Video Generation Through Joint RGB-Perception Modeling

SafetyDGX agent

arXiv:2604.02817v2 Announce Type: replace Abstract: Despite advancements in generating visually stunning content, video diffusion models (VDMs) often yield physically inconsistent results due to pixel

MoCRA: Mixture of Compositional Rank-1 Atoms for 4K All-in-One Video Restoration

Model ReleasesDGX agent

arXiv:2608.01829v1 Announce Type: new Abstract: Real-world video arrives hazy, rainy, dark, or noisy, and a deployable restorer faces three demands at once: no degradation label, native 4K output, and

Model-Agnostic FDR Control via Group Gaussian Mirror and Permutation SHAP

ApplicationsDGX agent

arXiv:2608.00989v1 Announce Type: cross Abstract: Most FDR-controlled feature selection methods are designed for coordinate-wise hypotheses, where each feature has a single weight or importance score.

Modeling Unknown Nonlocal PDE Systems via Flow Map Learning

TutorialsDGX agent

arXiv:2608.00400v1 Announce Type: new Abstract: Nonlocal partial differential equations arise in many applications but are often difficult to model and learn because of the presence of nonlocal operat

Models as Tools: An Agentic Coordination Framework for Unified Multimodal Visual Tracking

Model ReleasesDGX agent

arXiv:2608.00847v1 Announce Type: new Abstract: Most current visual trackers adopt a matching-based architecture trained exclusively on tracking datasets, whose performance gains depend heavily on the

MonitorVLM-v2: A Deployed Vision-Language Framework for Real-Time Safety Violation Detection

SafetyDGX agent

arXiv:2608.00975v1 Announce Type: new Abstract: Large vision--language models (VLMs) can reason step by step about complex visual scenes, but this open-ended, autoregressive chain-of-thought (CoT) app

MoRAL: Sensor-Grounded BEV Reasoning for Compact VLMs toward Edge-Oriented Autonomous Driving

Model ReleasesDGX agent

arXiv:2608.02449v1 Announce Type: new Abstract: Deploying vision-language models (VLMs) for safety-critical spatial reasoning on resource-constrained autonomous driving platforms requires both compact

Morphology Aware Reversible Semantic Tokenization and Hierarchical Word Composition for Tamil Language Models

Model ReleasesDGX agent

arXiv:2608.01153v1 Announce Type: new Abstract: Statistical subword tokenizers can process arbitrary text, but their units need not align with lexical or grammatical structure. This is especially impo

Motion Beyond Morphology: Bootstrapping Cross-Category Motion Transfer from Abstract Motion Representations

ResearchDGX agent

arXiv:2608.01628v1 Announce Type: new Abstract: Video motion transfer aims to animate a target object using dynamics from a reference video. Existing formulations largely rely on fixed structural corr

Motion Planning for Mobile Manipulators Navigating Doorways via Model Predictive Control

ResearchDGX agent

arXiv:2608.00206v1 Announce Type: new Abstract: Navigating doorways is a fundamental capability for mobile manipulators operating in human environments, requiring coordinated motion between the mobile

Move What Matters: Parameter-Efficient Domain Adaptation via Optimal Transport Flow for Collaborative Perception

Model ReleasesDGX agent

arXiv:2602.11565v5 Announce Type: replace Abstract: Efficient domain adaptation remains a fundamental challenge for deploying multi-agent systems across diverse environments in Vehicle-to-Everything (

Multi-Source Dynamic Graph Learning for Compound-Flood Forecasting in Managed Coastal Systems

SafetyDGX agent

arXiv:2608.01775v1 Announce Type: new Abstract: Compound flooding in managed coastal systems is influenced by hydrological conditions and water-management activity observed across multiple monitoring

Multi-View Unified Camera Fields: Geometry-Shaped Action-Facing Representations for RGB-Only Multi-Camera VLA Policies

ApplicationsDGX agent

arXiv:2608.01826v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have shown strong generalization in robotic manipulation, yet complex contact-rich tasks often benefit from multi-ca

Multiple result sets: How Database Migration Service automates SQL server to PostgreSQL translation

Model ReleasesDGX agent

In the Medium blog post, 'From MARS to SETOF REFCURSOR: Migrating Multi-Result Stored Procedures to PostgreSQL,' we explored the fundamental architectural differences between SQL Server and PostgreSQL

MUSS: Multilevel Subset Selection for Relevance and Diversity

ApplicationsDGX agent

arXiv:2503.11126v4 Announce Type: replace Abstract: The problem of relevant and diverse subset selection has a wide range of applications, including recommender systems and retrieval-augmented generat

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model

SafetyDGX agent

arXiv:2603.14686v2 Announce Type: replace Abstract: Human-Object Interaction (HOI) video reenactment aims to transfer the interaction dynamics of a source video to a novel target object while preservi

Native Multilingual Chain-of-Thought Reasoning in Low-Resource Southeast Asian Languages

SafetyDGX agent

arXiv:2608.00533v1 Announce Type: new Abstract: Large Language Models have achieved substantial progress in reasoning capabilities. Yet in low-resource native settings, many suffer from cross-lingual

Near-Optimal Reinforcement Learning for Constrained Recurrence Objectives

SafetyDGX agent

arXiv:2511.19849v2 Announce Type: replace-cross Abstract: Recurrence objectives, where a target region must be visited infinitely often, are a fundamental class of specifications for Markov decision p

NetDiff: Graph Diffusion with Improved Global Capabilities to Generate and Update Mobile Network Topologies

ResearchDGX agent

arXiv:2410.08238v2 Announce Type: replace-cross Abstract: We introduce NetDiff, a node-conditioned denoising diffusion model that generates directional link topologies and a two-slot transmit/receive

Network Information Enhances Unreliable News Domain Detection

ResearchDGX agent

arXiv:2608.02399v1 Announce Type: cross Abstract: Content-based detection of unreliable news is increasingly difficult, as low-reliability sources mimic credible journalism and generative AI makes fab

Neural Born Series Operator for Biomedical Ultrasound Computed Tomography

ResearchDGX agent

arXiv:2312.15575v2 Announce Type: replace-cross Abstract: Ultrasound Computed Tomography (USCT) provides a radiation-free option for high-resolution clinical imaging. Despite its potential, the comput

Neural Circuit Function Inference with LLMs

ResearchDGX agent

arXiv:2608.00059v1 Announce Type: new Abstract: The success of connectome mapping now shifts the challenge of understanding the nervous system to the interpretation of neural circuits. Here, we devise

Neural operator learning for collision-aware trajectory planning of spacecraft swarms

SafetyDGX agent

arXiv:2608.00320v1 Announce Type: new Abstract: Autonomous spacecraft swarms must plan fuel-efficient, collision-free maneuvers in increasingly congested orbits, yet classical trajectory optimization

Neural Surrogate HMC: On Using Neural Likelihoods for Hamiltonian Monte Carlo in Simulation-Based Inference

ResearchDGX agent

arXiv:2407.20432v3 Announce Type: replace Abstract: Bayesian inference methods such as Markov Chain Monte Carlo (MCMC) typically require repeated computations of the likelihood function, but in some s

New release of LLM adds support for reasoning traces, OpenAI Responses, server-side tools, and smarter logging

Model ReleasesDGX agent

I released LLM 0.32 this morning, the most significant new version of LLM since the initial launch of the project. The new version includes support for visible reasoning traces, server-side provider t

New York Smells: A Large Multimodal Dataset for Olfaction

Model ReleasesDGX agent

arXiv:2511.20544v2 Announce Type: replace Abstract: While olfaction is central to how animals perceive the world, this rich chemical sensory modality remains largely inaccessible to machines. One key

NISF++: Geometrically-grounded implicit representations of 3D+time cardiac function from 2D short- and long-axis MR views

SafetyDGX agent

arXiv:2608.00752v1 Announce Type: new Abstract: Clinical acquisition in cardiac magnetic resonance (CMR) imaging involves obtaining cross-sectional planes of the heart along the radial and longitudina

No One Wins in Nuclear War: A Social Simulation of Military Decision-making

SafetyDGX agent

arXiv:2608.01868v1 Announce Type: cross Abstract: WOPR is a social-simulation environment for studying how organizations make high-stakes decisions, built on a deterministic, replay-validated rules en

Noise-Robust Conditional Flow Matching: Generating Clean Samples from Noisy Datasets

TutorialsDGX agent

arXiv:2608.00064v1 Announce Type: new Abstract: Generative models learn the statistical properties of their training data, so high-quality generation depends on clean and representative datasets. In s

Non-KKT Accumulation in Entropic Mirror Descent

ResearchDGX agent

arXiv:2608.01658v1 Announce Type: cross Abstract: For mirror descent generated by a Legendre kernel, perhaps one of the most basic question in optimization is this: must every accumulation point of a

Nonlinear Laplacians Improve Signed-Directed Graph Learning

ResearchDGX agent

arXiv:2608.00836v1 Announce Type: new Abstract: While signed-directed graphs have been studied using linear Laplacians in the design of graph neural networks, relatively little research has focused on

Nonparametric Distribution Regression Re-calibration

SafetyDGX agent

arXiv:2602.13362v2 Announce Type: replace-cross Abstract: A key challenge in probabilistic regression is ensuring that predictive distributions accurately reflect true empirical uncertainty. Minimizin

Not All EEG Moments Are Equal: Position-Adaptive Time Scheduling for EEG Generation

SafetyDGX agent

arXiv:2608.00048v1 Announce Type: cross Abstract: Electroencephalography (EEG) generation is essential for alleviating data scarcity and enabling large scale neural modeling in brain computer interfac

‘Not healthy’ LLM use is more common than you think

ApplicationsDGX agent

Hank Green, a popular YouTuber and science communicator, said he is stepping back from production amid intense criticism over his use of AI. Green described his AI usage as 'not healthy,' but stressed

Not the Dimension, the Norm: What Matters in Gradient-Free Weight Perturbation of Language Models

Model ReleasesDGX agent

arXiv:2608.01624v1 Announce Type: new Abstract: Adapting a language model to a task no longer requires training all of its weights, and a line of parameter-efficient methods has driven the trainable c

Nova: An End-to-End MLIR Compiler for Deep Learning

Model ReleasesDGX agent

arXiv:2608.00029v1 Announce Type: cross Abstract: The performance of deep learning models at scale relies heavily on how effectively high-level mathematical operations are mapped to underlying physica

NVIDIA Alpamayo 2 Super, the Frontier Open Model for Robotaxis and Autonomous Vehicles, Now Available for Commercial Use

HardwareDGX agent

For robotaxis and other autonomous vehicles (AVs), the hardest problems aren’t the everyday scenarios. They’re the rare, complex situations that are difficult to anticipate and train for. Handling the

NVIDIA Joins NSF State and Regional AI Hubs Program to Expand AI Research and Education Across the US

HardwareDGX agent

NVIDIA is participating in the U.S. National Science Foundation’s (NSF) State and Regional Artificial Intelligence Infrastructure Hubs program, an effort launching today to expand access to the advanc

Nvidia open sources cuFile API, accelerating GPU read/write capability for high-speed storage

HardwareDGX agent

As artificial intelligence applications become ever hungrier for faster access to data, Nvidia Corp. today announced it is open-sourcing the application programming interface for its powerful cuFile v

Observatorio Lazaro: A self-populating database of anglicism usage in the Spanish press

ResearchDGX agent

arXiv:2608.00713v1 Announce Type: new Abstract: This paper describes Observatorio Lazaro, a language resource that monitors unassimilated lexical borrowings (predominantly English lexical borrowings o

Obshazard-bench: Benchmarking Multimodal Foundation Models for Real-Time Disaster Intelligence from Raw Earth Observation Streams

Model ReleasesDGX agent

arXiv:2608.00012v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) are increasingly used to interpret Earth observation data, yet their capability to support real-world disaster

Obsidian Security raises $85M as AI agents create cybersecurity’s next major attack surface

AgentsDGX agent

Obsidian Security Inc. has raised an 85 million Series D funding round at a post-money valuation of 1.1 billion as enterprises increasingly look to secure autonomous artificial intelligence agents acc

OC-VLA++: Monocular Geometry-Guided Cross-View Consistency for Viewpoint-Robust Robotic Manipulation

ApplicationsDGX agent

arXiv:2608.01066v1 Announce Type: new Abstract: We propose OC-VLA++, an extension of OC-VLA for viewpoint generalization under limited camera coverage. While OC-VLA grounds robot actions in the camera

OmniAI: A Surface-Adaptive Aerial Projection Interface for Human--Drone Interaction

AgentsDGX agent

arXiv:2608.00721v1 Announce Type: new Abstract: Drones in human environments often lack spatially grounded in- terfaces for situated communication. We present OmniAI, an em- bodied aerial agent that s

On the Identifiability of Masked Prediction: Mode Blindness and Mask Schedules

ResearchDGX agent

arXiv:2608.01383v1 Announce Type: new Abstract: Masked prediction learns representations by fitting a schedule-weighted family of conditional laws, but it remains unclear when near-optimal conditional

On the Limits of Machine-Learned Ranking for Modern Microarchitectural Policies

ResearchDGX agent

arXiv:2608.01041v1 Announce Type: cross Abstract: Machine-learning predictors estimate processor performance far faster than cycle-level simulation. For design-space exploration, however, the valuable

On the Limits of Support-Preserving Alignment and Bounded Filtering

SafetyDGX agent

arXiv:2607.18295v2 Announce Type: replace Abstract: We study whether alignment schemes that reshape a base model's output distribution, combined with bounded safety filters, can drive the probability

On the Viability of Semi-Supervised Segmentation Methods for Statistical Shape Modeling

Model ReleasesDGX agent

arXiv:2407.15260v3 Announce Type: replace Abstract: Statistical Shape Models (SSMs) excel at identifying population level anatomical variations, which is at the core of various clinical and biomedical

On the Wings of Imagination: Conflicting Script-based Multi-role Framework for Humor Caption Generation

ResearchDGX agent

arXiv:2602.06423v2 Announce Type: replace Abstract: Humor is a commonly used and intricate human language in daily life. Humor generation, especially in multi-modal scenarios, is a challenging task fo

Onboard Satellite Image Classification for Earth Observation: A Comparative Study of ViT Models

ResearchDGX agent

arXiv:2409.03901v4 Announce Type: replace Abstract: Remote sensing (RS) image classification is central to Earth observation, but onboard deployment requires models that are accurate, efficient, and r

← Previous
1…118119120121122…1410
Next →