AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,343
  • Agents7,552
  • Applications5,409
  • Concepts5
  • Hardware1,835
  • Industry6,164
  • Local Ai4,928
  • Model Releases23,861
  • Research20,124
  • Safety13,369
  • Syntheses17
  • Tools1,677
  • Tutorials3,402

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,343
  • Agents7,552
  • Applications5,409
  • Concepts5
  • Hardware1,835
  • Industry6,164
  • Local Ai4,928
  • Model Releases23,861
  • Research20,124
  • Safety13,369
  • Syntheses17
  • Tools1,677
  • Tutorials3,402

Source
Human
88,343Total entries
1Added by human
88,342Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
62,897 results
2 Jun 2026

MLLM-Microscope: Unlocking Hidden Structure Within Multimodal Large Language Models

ResearchDGX agent

arXiv:2606.00909v1 Announce Type: cross Abstract: This work presents MLLM-Microscope, a novel system designed for analyzing the hidden representations within Multimodal Large Language Models (MLLMs).

MM-Snowball: Evaluating and Mitigating Hallucination Snowballing in Multimodal Multi-Turn Dialogue

Model ReleasesDGX agent

arXiv:2606.00622v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) demonstrate remarkable visual understanding, yet their reliability in interactive settings is severely undermin

MMDG-Bench: A Benchmark for Multimodal Domain Generalization

Model ReleasesDGX agent

arXiv:2606.00891v1 Announce Type: new Abstract: Multi-modal Domain Generalization (MMDG) seeks to leverage complementary modalities to enhance model robustness on unseen domains. Despite extensive pro

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

MMG2Skill: Can Agents Distill In-the-Wild Guides into Self-Evolving Skills?

Model ReleasesDGX agent

arXiv:2606.01993v1 Announce Type: cross Abstract: Abundant procedural knowledge on the Web holds great potential for helping agents solve long-horizon tasks. However, such knowledge is often multimoda

MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion

ResearchDGX agent

arXiv:2604.02941v2 Announce Type: replace Abstract: Speech-driven three-dimensional (3D) facial animation synthesis aims to build a mapping from one-dimensional (1D) speech signals to time-varying 3D

MobEvolve: An Agentic Self-Evolving Heuristic System for Interpretable Human Mobility Generation

SafetyDGX agent

arXiv:2606.01640v1 Announce Type: new Abstract: Human mobility generation aims to synthesize realistic trip chains for target populations based on individual features. Existing paradigms, including de

MOC: Multi-Order Communication in LLM-based Multi-Agent Systems

Model ReleasesDGX agent

arXiv:2606.02359v1 Announce Type: new Abstract: Despite the remarkable progress of Large Language Model (LLM) based Multi-Agent Systems, most research focuses on optimizing coordination topology while

Model-Based Quality Assessment for Massively Multilingual Parallel Data

Model ReleasesDGX agent

arXiv:2606.00285v1 Announce Type: new Abstract: Large-scale multilingual bitext often contains two distinct problems: non-parallel sentence pairs and low-quality translations. We decompose model-based

Model Multiplicity and Predictive Arbitrariness in Recidivism Risk Assessment

SafetyDGX agent

arXiv:2606.02198v1 Announce Type: new Abstract: Prediction tasks over individual futures, which are inherently noisy, often admit multiple similarly accurate models. When these models produce differen

Model-Native Computing Architecture: Envisioning Future System Architecture Through the Lens of Computer Architecture

Model ReleasesDGX agent

arXiv:2606.00288v1 Announce Type: new Abstract: Large language models are undergoing a transition from model technology to system technology. As developers use Codex, Claude Code, AutoGPT, and related

Model Parallelism With Subnetwork Data Parallelism

Model ReleasesDGX agent

arXiv:2507.09029v5 Announce Type: replace-cross Abstract: Pre-training large neural networks at scale imposes heavy memory demands on accelerators and often requires costly communication. We introduce

Modeling Depth Ambiguity: A Mixture-Density Representation for Flying-Point-Free Depth Estimation

ResearchDGX agent

arXiv:2606.02552v1 Announce Type: cross Abstract: Despite advances in depth estimation, flying points remain a persistent failure mode: near object boundaries, depth estimators often predict spurious

Modeling Distinct Human Interaction in Web Agents

AgentsDGX agent

arXiv:2602.17588v3 Announce Type: replace Abstract: Despite rapid progress in autonomous web agents, human involvement remains essential for shaping preferences and correcting agent behavior as tasks

Modeling Robotics Dataset Construction as an Artifact-Based Build Process

ResearchDGX agent

arXiv:2606.00162v1 Announce Type: cross Abstract: Robotic systems generate large volumes of multimodal sensor data, but converting ROS bag recordings into machine learning datasets is often handled by

Modeling Spectral Energy Shifts in Spatio-Temporal Graph Anomaly Detection

ResearchDGX agent

arXiv:2606.00304v1 Announce Type: new Abstract: Graph anomaly detection methods aim to distinguish anomalous nodes. While prior methods characterize anomalies through increased variation in the spectr

MoEIoU: Rethinking Bounding-Box Regression as a Mixture of Experts

SafetyDGX agent

arXiv:2606.00844v1 Announce Type: cross Abstract: Bounding-box regression is a fundamental component of object detection, playing a critical role in precise object localization. Existing Intersection-

Molecular Embedding-Based Algorithm Selection in Protein-Ligand Docking

Model ReleasesDGX agent

arXiv:2512.02328v2 Announce Type: replace-cross Abstract: Selecting an effective docking algorithm is highly context-dependent, and no single method performs reliably across structural, chemical, and

Moment-Video: Diagnosing Temporal Fidelity of Video MLLMs on Momentary Visual Events

Model ReleasesDGX agent

arXiv:2606.02522v1 Announce Type: cross Abstract: Video multimodal large language models (MLLMs) have made rapid progress on general and long-form video understanding, yet their ability to preserve br

MomentKV: Closing the Directional Gap in KV Cache Eviction for Long-Context Inference

Model ReleasesDGX agent

arXiv:2606.01563v1 Announce Type: new Abstract: Autoregressive decoding in Transformer-based language models relies on the KV cache, whose memory footprint grows linearly with sequence length and beco

Momento: Evaluating Persistent Memory and Reasoning with Multi-Session Agentic Conversations

Model ReleasesDGX agent

arXiv:2606.00832v1 Announce Type: new Abstract: Recent advances in agentic AI have enabled agents to complete complex tasks through tool use, reasoning, and multi-step planning. Yet existing benchmark

Monitoring Agentic Systems Before They're Reliable

AgentsDGX agent

arXiv:2606.02494v1 Announce Type: cross Abstract: Agentic systems entering production typically operate as partially integrated assemblies where structural defects, not task-level errors, dominate the

MORPHOS: Autoregressive 4D Generation with Temporal Structured Latents

ResearchDGX agent

arXiv:2606.02491v1 Announce Type: new Abstract: We present MORPHOS, a novel autoregressive framework that generates dynamic 3D assets from videos across diverse representations, including meshes, 3D G

Mos-Gen: A Generative Molecular Framework for Mosquito Insecticide Design

ResearchDGX agent

arXiv:2606.01846v1 Announce Type: new Abstract: Mosquito-borne infectious diseases cause more than 700000 deaths worldwide each year. The long-term use of conventional chemical insecticides has induce

MOSAIC: Modular Orchestration for Structured Agentic Intelligence and Composition

SafetyDGX agent

arXiv:2606.00708v1 Announce Type: new Abstract: Automated data science is a structured model-selection problem. A solution must choose data transformations, feature representations, architecture, trai

MOSS-Audio Technical Report

ResearchDGX agent

arXiv:2606.01802v1 Announce Type: cross Abstract: MOSS-Audio is a unified audio-language model for speech, environmental sound, and music understanding, supporting audio captioning, time-aware questio

Motif-based morphology signatures for interpretable ECG screening and monitoring

ResearchDGX agent

arXiv:2606.00107v1 Announce Type: cross Abstract: Electrocardiography (ECG) remains central to cardiovascular screening, yet interpretation remains largely manual and episodic. Clinical practice relie

Motion-aware Event Suppression for Event Cameras

Model ReleasesDGX agent

arXiv:2602.23204v3 Announce Type: replace Abstract: In this work, we introduce the first framework for Motion-aware Event Suppression, which learns to filter events triggered by IMOs and ego-motion in

MotionDreamer: Universal Skeletal Motion Generation for 3D Rigged Shapes

Model ReleasesDGX agent

arXiv:2606.01518v1 Announce Type: new Abstract: Motion generation for rigged shapes is vital for scalable 4D asset production. However, template-based methods are limited by specific topologies and fa

Move the Query, Not the Cache: Characterizing Cross-Instance Latent Attention Redistribution Across GPU Fabrics

Model ReleasesDGX agent

arXiv:2606.01502v1 Announce Type: cross Abstract: Frontier LLMs increasingly decide what a query attends to with a sparse-attention indexer that picks a few KV-cache blocks per query: attention's unit

MPMWorlds: Material-Point-Method Simulations for Inferring and Extrapolating Physical Dynamics

ResearchDGX agent

arXiv:2606.01538v1 Announce Type: cross Abstract: To study the ability to infer physical dynamics from videos and extrapolate them forward in time, we assemble a dataset of 2D Material Point Method (M

MT-EditFlow: Reinforcement Learning for Multi-Turn Image Editing with Flow Matching

Model ReleasesDGX agent

arXiv:2606.01985v1 Announce Type: new Abstract: Recent breakthroughs in instruction-based image editing have captured significant attention, as models are now capable of handling real-world editing de

MulFeRL: Enhancing Reinforcement Learning with Verbal Feedback in a Multi-turn Loop

TutorialsDGX agent

arXiv:2601.22900v2 Announce Type: replace Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) is widely used to improve reasoning across domains, but outcome-only scalar rewards are often

Multi-Agent Computer Use

Model ReleasesDGX agent

arXiv:2606.01533v1 Announce Type: cross Abstract: Computer use agents (CUAs) today are primarily deployed as single serial agents. This setup is suboptimal for complex long-horizon tasks that benefit

Multi-Agent Conformal Prediction with Personalized Statistical Validity

AgentsDGX agent

arXiv:2606.00717v1 Announce Type: cross Abstract: Uncertainty quantification is essential in high-stakes machine learning tasks. However, one of the principled solutions, conformal prediction, faces c

Multi-Contrast MRI Motion Correction via Parameter-Informed Disentanglement and Adaptive Experts

Model ReleasesDGX agent

arXiv:2606.00146v1 Announce Type: cross Abstract: Motion artifacts in magnetic resonance imaging (MRI) degrade diagnostic reliability. Existing deep learning methods are typically contrast-specific an

Multi-modal Video Representation Alignment for Robust Self-supervised Driver Distraction Detection

SafetyDGX agent

arXiv:2606.02352v1 Announce Type: new Abstract: Robust self-supervised learning of multi-modal video representations is critical for real-world applications such as driver distraction detection, where

Multi-Objective Reference-Aligned Machine Unlearning

SafetyDGX agent

arXiv:2606.00399v1 Announce Type: new Abstract: Machine unlearning aims to remove the influence of specific training samples while preserving the model's utility. Existing single-objective approaches,

Multi-Objective Reinforcement Learning for Tactical Decision Making for Trucks in Highway Traffic

SafetyDGX agent

arXiv:2601.18783v2 Announce Type: replace-cross Abstract: Balancing safety, efficiency, and operational costs in highway driving poses a challenging decision-making problem for heavy-duty vehicles. A

Multi-view Pyramid Transformer: Look Coarser to See Broader

Local AiDGX agent

arXiv:2512.07806v2 Announce Type: replace Abstract: We propose Multi-view Pyramid Transformer (MVP), a scalable multi-view transformer architecture that directly reconstructs large 3D scenes from tens

Multigrade Neural Network Approximation

ResearchDGX agent

arXiv:2601.16884v3 Announce Type: replace Abstract: We study multigrade deep learning (MGDL) as a principled framework for structured error refinement in deep neural networks. While the approximation

Multilingual Idioms in Sentences and Conversations Across High-, Medium-, and Low-Resource Languages

ResearchDGX agent

arXiv:2606.02147v1 Announce Type: cross Abstract: Idiomatic expressions pose a major challenge for multilingual NLP because their meanings shift between figurative and literal usage, often requiring c

Multilinguality of Large Language Models From a Structural Perspective

ResearchDGX agent

arXiv:2606.01800v1 Announce Type: cross Abstract: Large language models (LLMs) have excelled in processing multiple languages through pre- and post-training on multilingual data, even though English d

Multimodal Action Diffusion for Robust End-to-End Autonomous Driving

Model ReleasesDGX agent

arXiv:2606.02105v1 Announce Type: new Abstract: End-to-End Autonomous Driving (E2E-AD) systems have largely converged on predicting intermediate trajectory waypoints, delegating final control to hand-

Multimodal Approaches for Visually-Rich Document Type Classification: A Comparative Analysis

Model ReleasesDGX agent

arXiv:2606.02162v1 Announce Type: cross Abstract: Document type classification in visually rich documents remains challenging, as relevant information is distributed across textual, visual, and layout

Multimodal Function Vectors for Visual Relations

Local AiDGX agent

arXiv:2510.02528v2 Announce Type: replace Abstract: Large Multimodal Models (LMMs) demonstrate impressive in-context learning abilities from few multimodal demonstrations, yet the internal mechanisms

Multimodal Music Recommendation System using LLMs

Model ReleasesDGX agent

arXiv:2606.00125v1 Announce Type: cross Abstract: Music recommendation systems typically treat songs as opaque tokens, relying on collaborative interaction histories which overlooks semantic or acoust

MURMUR: An Efficient Inference System for Long-Form ASR

SafetyDGX agent

arXiv:2606.01483v1 Announce Type: cross Abstract: Long-form automatic speech recognition (ASR) requires both high accuracy and low latency, but existing systems force a trade-off between the two. Chun

MUSCLE-NET: Predicted-Multiscale-Aware Network for Pedestrian Trajectory Forecasting

AgentsDGX agent

arXiv:2606.00471v1 Announce Type: new Abstract: Accurate pedestrian trajectory prediction is essential for safe navigation in autonomous driving and intelligent transportation systems. Despite substan

MViewRouter: Internalizing Geometric Equivariance via Multi-view Alternating Attention for Combinatorial Routing

SafetyDGX agent

arXiv:2606.01084v1 Announce Type: cross Abstract: Combinatorial routing problems such as the Traveling Salesman Problem (TSP) and the Capacitated Vehicle Routing Problem (CVRP) are fundamental NP-hard

MyoSem: Aligning Electromyography to Natural-Language Action Semantics for Hand Action Understanding

SafetyDGX agent

arXiv:2606.00174v1 Announce Type: cross Abstract: Electromyography (EMG) directly reflects muscle activation and is a key sensing modality for gesture recognition, prosthetic control, and wearable int

naPINN: Noise-Adaptive Physics-Informed Neural Networks for Recovering Physics from Corrupted Measurement

Model ReleasesDGX agent

arXiv:2602.02547v2 Announce Type: replace-cross Abstract: Physics-Informed Neural Networks (PINNs) are effective methods for solving inverse problems and discovering governing equations from observati

NAPPure: Adversarial Purification for Robust Image Classification under Non-Additive Perturbations

ApplicationsDGX agent

arXiv:2510.14025v2 Announce Type: replace Abstract: Adversarial purification has achieved great success in combating adversarial image perturbations, which are usually assumed to be additive. However,

Navigating the Reality Gap: On-Device Continual Adaptation of ASR for Clinical Telephony

Model ReleasesDGX agent

arXiv:2512.16401v5 Announce Type: replace Abstract: Automatic Speech Recognition (ASR) can significantly reduce documentation burden in clinical workflows, but standard models degrade sharply in real-

NBQ: Next-Best-Question for Dynamic Profiling

ApplicationsDGX agent

arXiv:2606.00809v1 Announce Type: new Abstract: Many real-world conversational settings for knowledge discovery, including podcasts, hiring screens, and marketplaces, require a purpose-driven understa

NDPP-Grasp: Non-Differentiable Physical Plausibility Constraint-Guided Task-Oriented Dexterous Grasp Generation

SafetyDGX agent

arXiv:2606.02432v1 Announce Type: new Abstract: Task-oriented dexterous grasp generation aims to produce dexterous grasp poses that are both physically plausible and functionally suitable for specifie

Near-Optimal Private Tests for Simple and MLR Hypotheses

ResearchDGX agent

arXiv:2601.21959v2 Announce Type: replace-cross Abstract: We develop a near-optimal testing procedure under the framework of Gaussian differential privacy for simple as well as one- and two-sided test

Near-Optimal Pure Machine Unlearning for Smooth Strongly Convex Losses

ApplicationsDGX agent

arXiv:2606.01527v1 Announce Type: new Abstract: Machine unlearning is motivated by legal and user-facing requirements to remove the influence of individuals' data from trained models, such as the righ

Needles at Scale: LLM-Assisted Target Selection for Windows Vulnerability Research

AgentsDGX agent

arXiv:2606.01364v1 Announce Type: cross Abstract: The attack surface of a modern operating system is a haystack: thousands of signed binaries and millions of functions, almost none relevant to any giv

NestRL: A Nested Training Regime for Mutual Adaptation in Human-AI Teaming

AgentsDGX agent

arXiv:2602.17737v2 Announce Type: replace-cross Abstract: Mutual adaptation is a central challenge in human-AI teaming, as humans naturally adjust their strategies in response to an AI agent's behavio

Network Distributed Multi-Agent Reinforcement Learning for Consensus Control of Quadcopters

SafetyDGX agent

arXiv:2606.02107v1 Announce Type: cross Abstract: This paper proposes a Network Distributed Multi-Agent Reinforcement Learning (ND-MARL) framework for quadcopter consensus control. Compared to convent

← Previous
1…521522523524525…1049
Next →