AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,678
  • Agents7,513
  • Applications5,367
  • Concepts5
  • Hardware1,821
  • Industry6,154
  • Local Ai4,902
  • Model Releases23,619
  • Research19,969
  • Safety13,271
  • Syntheses17
  • Tools1,674
  • Tutorials3,366

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,678
  • Agents7,513
  • Applications5,367
  • Concepts5
  • Hardware1,821
  • Industry6,154
  • Local Ai4,902
  • Model Releases23,619
  • Research19,969
  • Safety13,271
  • Syntheses17
  • Tools1,674
  • Tutorials3,366

Source
Human
87,678Total entries
1Added by human
87,677Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
62,383 results
27 May 2026

MiRD: Reliable Set-Valued Prediction for Open-Ended Question Answering via Miscoverage Risk Decomposition

ResearchDGX agent

arXiv:2605.27091v1 Announce Type: cross Abstract: Reliable set-valued prediction provides a principled way to mitigate hallucinations in open-ended question answering (QA), yet existing conformal appr

MobileExplorer: Accelerating On-Device Inference for Mobile GUI Agents via Online Exploration

Model ReleasesDGX agent

arXiv:2605.26546v1 Announce Type: new Abstract: Mobile graphical user interface (GUI) agents enable AI models to autonomously operate smartphones on behalf of users. However, most existing systems foc

MobileMoE: Scaling On-Device Mixture of Experts

Model ReleasesDGX agent

arXiv:2605.27358v1 Announce Type: cross Abstract: Mixture-of-Experts (MoE) has become the de facto architecture for hundred-billion-parameter language models, yet its advantages at sub-billion scales

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Model discovery for dynamical systems with complex-valued product units

Model ReleasesDGX agent

arXiv:2605.27158v1 Announce Type: new Abstract: Discovering the governing equations of a dynamical system from observed trajectories provides deeper insight into its structure than mere prediction of

Model Merging on Loss Landscape: A Geometry Perspective

ResearchDGX agent

arXiv:2605.26693v1 Announce Type: cross Abstract: Model merging offers a promising avenue for knowledge integration and parallel development without retraining. Yet, existing methods either ignore the

Model Unlearning Objectives Vary for Distinct Language Functions

TutorialsDGX agent

arXiv:2605.26454v1 Announce Type: new Abstract: Large language models (LLMs) learn undesirable properties during pretraining, including dangerous knowledge and toxic text generation. Just as post-trai

Modeling Agentic Technical Debt and Stochastic Tax: A Standalone Framework for Measurement, Simulation, and Dashboarding

AgentsDGX agent

arXiv:2605.27320v1 Announce Type: new Abstract: Agentic AI systems combine probabilistic reasoning with delegated action through tools, context, memory, orchestration, and external workflow integratio

Modeling Dynamic Mixtures of Time-Delay Systems from Streaming Time Series

Model ReleasesDGX agent

arXiv:2605.26191v1 Announce Type: cross Abstract: This research addresses the problem of adaptive modeling in time-series data streams with clear input-output relationships. This problem is challengin

Modernising Reinforcement Learning-Based Navigation for Embodied Semantic Scene Graph Generation

SafetyDGX agent

arXiv:2603.25415v2 Announce Type: replace Abstract: Semantic world models enable embodied agents to reason about objects, relations, and spatial context beyond purely geometric representations. In Org

MolPIF: A Parameter Interpolation Flow Model for Molecule Generation

Model ReleasesDGX agent

arXiv:2507.13762v4 Announce Type: replace Abstract: Motivation: Structure-based drug design (SBDD) has advanced with deep generative models, but bridging the gap between continuous atomic coordinates

MONA: Muon Optimizer with Nesterov Acceleration for Scalable Language Model Training

Local AiDGX agent

arXiv:2605.26842v1 Announce Type: cross Abstract: The Muon optimizer has recently offered a promising alternative to AdamW for large language model training, leveraging matrix orthogonalization to pro

Monte Carlo Permutation Search

SafetyDGX agent

arXiv:2510.06381v2 Announce Type: replace-cross Abstract: We propose Monte Carlo Permutation Search (MCPS), a general-purpose Monte Carlo Tree Search (MCTS) algorithm that improves upon the GRAVE algo

More Expressive Feedforward Layers: Part I. Token-Adaptive Mixing of Activations

Model ReleasesDGX agent

arXiv:2605.26647v1 Announce Type: cross Abstract: Feedforward network (FFN) layers account for a large fraction of parameters and nonlinear expressivity in Transformer-based large language models (LLM

Morphling: Fast, Fused, and Flexible GNN Training at Scale

HardwareDGX agent

arXiv:2512.01678v5 Announce Type: replace Abstract: Graph Neural Networks (GNNs) present a fundamental hardware challenge by fusing irregular, memory-bound graph traversals with regular, compute-inten

MotionPRO: Exploring the Role of Pressure in Human MoCap and Beyond

ApplicationsDGX agent

arXiv:2504.05046v2 Announce Type: replace Abstract: Existing human Motion Capture (MoCap) methods mostly focus on the visual similarity while neglecting the physical plausibility. As a result, downstr

MRT: Masked Region Transformer for Layered Image Generation and Editing at Scale

Model ReleasesDGX agent

arXiv:2605.27235v1 Announce Type: new Abstract: Layered image generation and editing is a fundamental capability that enables layer-wise reuse, editing, and composition of generated visual content, an

MSCGC-KAN: Multi-scale Causal Graph Convolution and Kolmogorov-Arnold Feature Mapping for EEG Emotion Recognition

ResearchDGX agent

arXiv:2605.26624v1 Announce Type: new Abstract: Electroencephalogram (EEG)-based emotion recognition is an important affective computing task, and recent EEG foundation models provide useful generic r

MTL-FNO: A Lightweight Multi-Task Fourier Neural Operator for Sparse Field Reconstruction

Model ReleasesDGX agent

arXiv:2605.26718v1 Announce Type: new Abstract: Efficient onboard multi-field sparse reconstruction is essential for the autonomous operation of aerospace vehicles. While existing deep learning models

MuCon: Clipped Muon Updates for LLM Training

ResearchDGX agent

arXiv:2605.26459v1 Announce Type: new Abstract: Muon-style optimizers take a matrix-valued momentum or preconditioned update B = U operatorname{diag}(sigma_1,ldots,sigma_r) V^op and replace it with it

Multi-Agent Causal Discovery Using Large Language Models

Model ReleasesDGX agent

arXiv:2407.15073v4 Announce Type: replace Abstract: Causal discovery aims to identify causal relationships between variables and is a fundamental problem across the sciences. Traditional statistical c

Multi-Modal Building Inspection via Perceiver IO Fusion of Satellite and Street-Level Imagery

ResearchDGX agent

arXiv:2605.26381v1 Announce Type: new Abstract: We present a multi-modal classification framework that fuses satellite and street-level imagery through a Perceiver IO architecture operating on spatial

Multi-Robot Box Transport over Different Surfaces with Decentralized Role-based Proportional Control

HardwareDGX agent

arXiv:2605.26430v1 Announce Type: new Abstract: Collaborative transport of objects via pushing by multiple robots has many applications, ranging from construction and warehouse environments to post di

Multi-Stakeholder LLM Alignment: Decomposing Estimation from Aggregation

SafetyDGX agent

arXiv:2605.26878v1 Announce Type: new Abstract: Multi-stakeholder tasks require one output to satisfy users with conflicting preferences. Holistic LLM judges conflate utility estimation and utility ag

MULTISEISMO: A Multimodal Seismic Dataset and Model for Cross-Modal Seismic Understanding

Model ReleasesDGX agent

arXiv:2605.26320v1 Announce Type: cross Abstract: The application of generalist multimodal models (GMMs) to specialized scientific domains remains limited due to the scarcity of comprehensive domain-s

MUSE-Autoskill: Self-Evolving Agents via Skill Creation, Memory, Management, and Evaluation

AgentsDGX agent

arXiv:2605.27366v1 Announce Type: new Abstract: Large language model (LLM) agents rely on reusable skills to solve complex tasks. However, existing skill creation approaches treat skills as isolated a

MVISTA-4D: View-Consistent 4D World Model with Test-Time Action Inference for Robotic Manipulation

SafetyDGX agent

arXiv:2602.09878v2 Announce Type: replace Abstract: World-model-based imagine-then-act becomes a promising paradigm for robotic manipulation, yet existing approaches typically support either purely im

Natural Human Motion Recovery by Aligning High-Order Temporal Dynamics from Monocular Videos

ResearchDGX agent

arXiv:2605.26879v1 Announce Type: new Abstract: Human motion recovered from monocular videos often appears overly smooth or dynamically inconsistent, even when joint positions are numerically accurate

Natural Language Query to Configuration for Retrieval Agents

ResearchDGX agent

arXiv:2605.27361v1 Announce Type: new Abstract: Modern retrieval agents expose many configuration choices -- LLM, retriever, number of documents, number of hops, and synthesis strategy -- each shaping

Near-Optimal Regret in Adversarial Kernel Bandits

Model ReleasesDGX agent

arXiv:2605.26585v1 Announce Type: new Abstract: We study the adversarial kernel bandit problem, in which the loss at each round is induced by an arbitrary bounded element of a reproducing kernel Hilbe

Negligible in Size, Significant in Effect: On Scale Vectors in Large Language Models

Model ReleasesDGX agent

arXiv:2605.26895v1 Announce Type: cross Abstract: Normalization layers in modern large language models (LLMs) consist of a deterministic normalization operation and a learnable scale vector. While the

NeR-SC: Adapting Neural Video Representation to Screen Content

ApplicationsDGX agent

arXiv:2605.27024v1 Announce Type: new Abstract: Implicit neural representations have emerged as a promising paradigm for video compression, with recent methods achieving competitive performance on nat

NestedKV: Nested Memory Routing for Long-Context KV Cache Compression

Model ReleasesDGX agent

arXiv:2605.26678v1 Announce Type: new Abstract: Long-context language models are limited by the memory footprint of the key-value (KV) cache. Existing training-free KV compression methods usually rank

Neural Autoregressive Control Variates for the Quantum Monte Carlo Sign Problem

Model ReleasesDGX agent

arXiv:2605.26814v1 Announce Type: cross Abstract: We train a pair of autoregressive models to construct zero-mean control variates to mitigate the sign problem in quantum Monte Carlo simulations. The

Neural Bayesian Sequential Routing

AgentsDGX agent

arXiv:2605.26147v1 Announce Type: new Abstract: Human decision-making is sequential and uncertainty-aware, yet standard neural networks often rely on static, dense forward computation with limited vis

Neural Scaling Laws of Deep ReLU and Deep Operator Network: A Theoretical Study

ResearchDGX agent

arXiv:2410.00357v2 Announce Type: replace Abstract: Neural scaling laws play a pivotal role in the performance of deep neural networks and have been observed in a wide range of tasks. However, a compl

Neuro-Symbolic Verification of LLM Outputs for Data-Sensitive Domains (extended preprint)

SafetyDGX agent

arXiv:2605.26942v1 Announce Type: new Abstract: LLMs deployed in high-stakes domains face fundamental reliability challenges: hallucinations, inconsistencies, and privacy vulnerabilities introduce una

NightSight: Passive Computation for Navigation in Dark Using Events

HardwareDGX agent

arXiv:2605.26330v1 Announce Type: new Abstract: Small aerial robots are particularly well-suited for search and rescue in confined and hazardous environments due to their agility, low cost, and abilit

No Data? No Problem: Robust Vision-Tabular Learning with Missing Values

ApplicationsDGX agent

arXiv:2512.19602v2 Announce Type: replace Abstract: Large-scale medical biobanks provide imaging data complemented by extensive tabular information, such as clinical measurements or demographics. Howe

Nonlinear Data Integration via Kernel Methods for Data Collaboration Analysis

ResearchDGX agent

arXiv:2605.27219v1 Announce Type: new Abstract: Collaborative analysis of decentralized confidential datasets is important, but direct sharing of original datasets is often restricted by privacy and i

Normal Guidance is what Attention Needs

TutorialsDGX agent

arXiv:2605.27306v1 Announce Type: new Abstract: We consider training classifiers for 3D medical images using only one binary label for the entire volume rather than a label for each 2D slice. In such

Normalizing Flows on Quotient Manifolds via Boundary Quotients

ResearchDGX agent

arXiv:2511.22882v3 Announce Type: replace Abstract: We introduce boundary quotients and present a framework for learning densities on manifolds that arise as boundary quotients of simpler domains. We

Not All Disagreement Is Learnable: Token Teachability in On-Policy Distillation

Model ReleasesDGX agent

arXiv:2605.26844v1 Announce Type: new Abstract: On-policy distillation (OPD) trains a student on its own rollouts with token-level teacher supervision. Recent selective OPD methods exploit the non-uni

Not All Modalities Are Equal: Instruction-Aware Gating for Multimodal Videos

ResearchDGX agent

arXiv:2605.26232v1 Announce Type: new Abstract: Pre-trained video large language models excel at visual reasoning. However, they struggle when videos arrive with auxiliary streams, such as audio, dept

Not All Tokens Matter Equally: Dynamic In-context Vector Distillation with Decisive-Token Supervision for Long-form Medical Report Generation

ResearchDGX agent

arXiv:2605.27194v1 Announce Type: new Abstract: Distilling demonstration effects into hidden-space interventions offers a lightweight alternative to full finetuning. However, existing multimodal varia

NSF-SciFy: Mining the NSF Awards Database for Scientific Claims

ResearchDGX agent

arXiv:2503.08600v3 Announce Type: replace Abstract: We introduce NSF-SciFy, a comprehensive dataset of scientific claims and investigation proposals extracted from National Science Foundation award ab

O-MARC: Omni Memory-Augmented Compression Distillation for Efficient Video Understanding

Model ReleasesDGX agent

arXiv:2605.26584v1 Announce Type: new Abstract: Omnimodal large language models enable unified audio video understanding, but long joint token sequences make inference costly, and existing benchmarks

Object Pose and Shape Estimation for Grasping: Does it Work?

ResearchDGX agent

arXiv:2605.26944v1 Announce Type: cross Abstract: The problem of object pose and shape estimation has seen key advancements lately. Encoder-decoder (e.g., SAM3D, LRM, CRISP) and diffusion-based models

OCR-Reasoning Benchmark: Unveiling the True Capabilities of MLLMs in Complex Text-Rich Image Reasoning

Model ReleasesDGX agent

arXiv:2505.17163v2 Announce Type: replace-cross Abstract: Recent advancements in multimodal slow-thinking systems have demonstrated remarkable performance across various visual reasoning tasks. Howeve

ODOV: Benchmark the Open-Domain Open-Vocabulary Object Detection

Model ReleasesDGX agent

arXiv:2508.01253v2 Announce Type: replace Abstract: Existing studies typically investigate domain shift and category shift as independent problems, however, in real-world scenarios, the two types of s

Olaf-World: Orienting Latent Actions for Video World Modeling

SafetyDGX agent

arXiv:2602.10104v2 Announce Type: replace-cross Abstract: Scaling action-controllable world models is limited by the scarcity of action labels. While latent action learning promises to extract control

Omanic: Towards Step-wise Evaluation of Multi-hop Reasoning in Large Language Models

Model ReleasesDGX agent

arXiv:2603.16654v2 Announce Type: replace-cross Abstract: Evaluating the reasoning abilities of large language models (LLMs) solely from final answers can obscure failures in intermediate steps, espec

OMD-GraphRAG: Enhancing GraphRAG with Ontology-Guided Extraction, Multi-Dimensional Clustering and Dual-Channel Fusion

Model ReleasesDGX agent

arXiv:2603.25152v3 Announce Type: replace Abstract: Retrieval-Augmented Generation (RAG) systems face significant challenges in complex reasoning, multi-hop queries, and domain-specific QA. While exis

OmniGF: A Dual-Branch Vision-Language Framework for Unified Gaze Following

Local AiDGX agent

arXiv:2605.26399v1 Announce Type: new Abstract: Understanding human gaze behavior is essential for complex scene comprehension and human-computer interaction. Traditional gaze following models are typ

OmniInteract: Benchmarking Real-World Streaming Interaction for Real-Time Omnimodal Assistants

Model ReleasesDGX agent

arXiv:2605.26485v1 Announce Type: cross Abstract: We introduce OmniInteract, a streaming benchmark for real-time omnimodal large language models evaluated through native online inference over audio-vi

OmniRetriever: Any-to-Any Audio-Video-Text Retrieval via Fusion-as-Teacher Distillation

Model ReleasesDGX agent

arXiv:2605.26641v1 Announce Type: new Abstract: Unified multimodal embedding spaces have become the standard interface for cross-modal retrieval and multimodal RAG, and recent audio-video-text (AVT) e

OmniToM: Benchmarking Theory of Mind in LLMs via Explicit Belief Modeling

Model ReleasesDGX agent

arXiv:2605.26322v1 Announce Type: new Abstract: Theory of Mind (ToM), the ability to infer others' knowledge, intentions, and emotions, is commonly evaluated in large language models (LLMs) using end-

On the Detection of Commutative Factors in Factor Graphs: Necessary and Sufficient Conditions

ResearchDGX agent

arXiv:2605.26908v1 Announce Type: new Abstract: Exploiting the indistinguishability of objects in a probabilistic graphical model such as a factor graph is key to lifted probabilistic inference algori

On the Error-Correcting Effects of Stochasticity in Discrete Diffusion

ResearchDGX agent

arXiv:2605.26582v1 Announce Type: cross Abstract: Discrete diffusion models achieve strong performance in text and image generation, but their inference remains slow and must inherently balance sampli

On the Generalization Capabilities, Design Choices and Limitations of Keypoint Imitation Learning

TutorialsDGX agent

arXiv:2605.26649v1 Announce Type: new Abstract: RGB-based imitation learning requires many demonstrations to generalize to unseen objects or scenes, motivating research into intermediate representatio

On the Hidden Costs of Counterfactual Knowledge Training in LLM Unlearning

Model ReleasesDGX agent

arXiv:2605.27083v1 Announce Type: new Abstract: Counterfactual tuning (CFT) has emerged as a promising paradigm for Large Language Model (LLM) unlearning by training models to generate alternative fic

← Previous
1…582583584585586…1040
Next →