AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
Human
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
59,295 results
14 Apr 2026

MARLIN: Multi-Agent Reinforcement Learning Guided by Language-Based Inter-Robot Negotiation

SafetyDGX agent

arXiv:2410.14383v4 Announce Type: replace Abstract: Multi-agent reinforcement learning is a key method for training multi-robot systems. Through rewarding or punishing robots over a series of episodes

MARS: Unleashing the Power of Speculative Decoding via Margin-Aware Verification

Local AiDGX agent

arXiv:2601.15498v2 Announce Type: replace Abstract: Speculative Decoding (SD) accelerates autoregressive large language model (LLM) inference by decoupling generation and verification. While recent me

MASH: Modeling Abstention via Selective Help-Seeking

AgentsDGX agent

arXiv:2510.01152v2 Announce Type: replace Abstract: LLMs cannot reliably recognize their parametric knowledge boundaries and often hallucinate answers to outside-of-boundary questions. In this paper,

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Masked Contrastive Pre-Training Improves Music Audio Key Detection

ResearchDGX agent

arXiv:2604.10021v1 Announce Type: cross Abstract: Self-supervised music foundation models underperform on key detection, which requires pitch-sensitive representations. In this work, we present the fi

Masked Training for Robust Arrhythmia Detection from Digitalized Multiple Layout ECG Images

ApplicationsDGX agent

arXiv:2508.09165v3 Announce Type: replace-cross Abstract: Background: Electrocardiograms are indispensable for diagnosing cardiovascular diseases, yet in many settings they exist only as paper printou

MASS: Motion-Aware Spatial-Temporal Grounding for Physics Reasoning and Comprehension in Vision-Language Models

Model ReleasesDGX agent

arXiv:2511.18373v2 Announce Type: replace Abstract: Vision Language Models (VLMs) perform well on standard video tasks but struggle with physics-related reasoning involving motion dynamics and spatial

MathAgent: Adversarial Evolution of Constraint Graphs for Mathematical Reasoning Data Synthesis

Model ReleasesDGX agent

arXiv:2604.11188v1 Announce Type: cross Abstract: Synthesizing high-quality mathematical reasoning data without human priors remains a significant challenge. Current approaches typically rely on seed

MatRes: Zero-Shot Test-Time Model Adaptation for Simultaneous Matching and Restoration

SafetyDGX agent

arXiv:2604.10081v1 Announce Type: cross Abstract: Real-world image pairs often exhibit both severe degradations and large viewpoint changes, making image restoration and geometric matching mutually in

MAVEN-T: Multi-Agent enVironment-aware Enhanced Neural Trajectory predictor with Reinforcement Learning

Model ReleasesDGX agent

arXiv:2604.10169v1 Announce Type: new Abstract: Trajectory prediction remains a critical yet challenging component in autonomous driving systems, requiring sophisticated reasoning capabilities while m

Maximum Entropy Relaxation of Multi-Way Cardinality Constraints for Synthetic Population Generation

SafetyDGX agent

arXiv:2603.22558v2 Announce Type: replace Abstract: Generating synthetic populations from aggregate statistics is a core component of microsimulation, agent-based modeling, policy analysis, and privac

MCAT: Scaling Many-to-Many Speech-to-Text Translation with MLLMs to 70 Languages

Model ReleasesDGX agent

arXiv:2512.01512v2 Announce Type: replace Abstract: Multimodal Large Language Models (MLLMs) have achieved great success in Speech-to-Text Translation (S2TT) tasks. However, current research is constr

MCERF: Advancing Multimodal LLM Evaluation of Engineering Documentation with Enhanced Retrieval

Model ReleasesDGX agent

arXiv:2604.09552v1 Announce Type: cross Abstract: Engineering rulebooks and technical standards contain multimodal information like dense text, tables, and illustrations that are challenging for retri

MCGA: A Multi-task Classical Chinese Literary Genre Audio Corpus

ResearchDGX agent

arXiv:2601.09270v3 Announce Type: replace Abstract: With the rapid advancement of Multimodal Large Language Models (MLLMs), their potential has gained significant attention in Chinese Classical Studie

MDP Planning as Policy Inference

SafetyDGX agent

arXiv:2602.17375v2 Announce Type: replace Abstract: We cast episodic Markov decision process (MDP) planning as Bayesian inference over policies. A policy is treated as the latent variable and is assig

Measuring and curing reasoning rigidity: from decorative chain-of-thought to genuine faithfulness

Model ReleasesDGX agent

arXiv:2603.22816v3 Announce Type: replace-cross Abstract: Language models increasingly show their work by writing step-by-step reasoning before answering. But are these steps genuinely used, or is the

Measuring the Authority Stack of AI Systems: Empirical Analysis of 366,120 Forced-Choice Responses Across 8 AI Models

Model ReleasesDGX agent

arXiv:2604.11216v1 Announce Type: new Abstract: What values, evidence preferences, and source trust hierarchies do AI systems actually exhibit when facing structured dilemmas? We present the first lar

Measuring What Matters!! Assessing Therapeutic Principles in Mental-Health Conversation

Model ReleasesDGX agent

arXiv:2604.05795v2 Announce Type: replace Abstract: The increasing use of large language models in mental health applications calls for principled evaluation frameworks that assess alignment with psyc

MedLVR: Latent Visual Reasoning for Reliable Medical Visual Question Answering

SafetyDGX agent

arXiv:2604.09757v1 Announce Type: cross Abstract: Medical vision--language models (VLMs) have shown strong potential for medical visual question answering (VQA), yet their reasoning remains largely te

MedP-CLIP: Medical CLIP with Region-Aware Prompt Integration

Local AiDGX agent

arXiv:2604.11197v1 Announce Type: new Abstract: Contrastive Language-Image Pre-training (CLIP) has demonstrated outstanding performance in global image understanding and zero-shot transfer through lar

MEDSYN: Benchmarking Multi-EviDence SYNthesis in Complex Clinical Cases for Multimodal Large Language Models

Model ReleasesDGX agent

arXiv:2602.21950v2 Announce Type: replace Abstract: Multimodal large language models (MLLMs) have shown great potential in medical applications, yet existing benchmarks inadequately capture real-world

MedVeriSeg: Teaching MLLM-Based Medical Segmentation Models to Verify Query Validity Without Extra Training

Model ReleasesDGX agent

arXiv:2604.10242v1 Announce Type: new Abstract: Despite recent advances in MLLM-based medical image segmentation, existing LISA-like methods cannot reliably reject false queries and often produce hall

MegaFake: A Theory-Driven Dataset of Fake News Generated by Large Language Models

ResearchDGX agent

arXiv:2408.11871v4 Announce Type: replace-cross Abstract: Fake news significantly influences decision-making processes by misleading individuals, organizations, and even governments. Large language mo

MeloTune: On-Device Arousal Learning and Peer-to-Peer Mood Coupling for Proactive Music Curation

Local AiDGX agent

arXiv:2604.10815v1 Announce Type: cross Abstract: MeloTune is an iPhone-deployed music agent that instantiates the Mesh Memory Protocol (MMP) and Symbolic-Vector Attention Fusion (SVAF) as a productio

Mem^2Evolve: Towards Self-Evolving Agents via Co-Evolutionary Capability Expansion and Experience Distillation

AgentsDGX agent

arXiv:2604.10923v1 Announce Type: cross Abstract: While large language model--powered agents can self-evolve by accumulating experience or by dynamically creating new assets (i.e., tools or expert age

Membership Inference Attacks Expose Participation Privacy in ECG Foundation Encoders

ResearchDGX agent

arXiv:2604.10424v1 Announce Type: new Abstract: Foundation-style ECG encoders pretrained with self-supervised learning are increasingly reused across tasks, institutions, and deployment contexts, ofte

MemDLM: Memory-Enhanced DLM Training

Model ReleasesDGX agent

arXiv:2603.22241v2 Announce Type: replace Abstract: Diffusion Language Models (DLMs) offer attractive advantages over Auto-Regressive (AR) models, such as full-attention parallel decoding and flexible

MEMENTO: Teaching LLMs to Manage Their Own Context

Model ReleasesDGX agent

arXiv:2604.09852v1 Announce Type: new Abstract: Reasoning models think in long, unstructured streams with no mechanism for compressing or organizing their own intermediate state. We introduce MEMENTO:

Merging Triggers, Breaking Backdoors: Defensive Poisoning for Instruction-Tuned Language Models

ResearchDGX agent

arXiv:2601.04448v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have greatly advanced Natural Language Processing (NLP), particularly through instruction tuning, which enables b

MERMAID: Memory-Enhanced Retrieval and Reasoning with Multi-Agent Iterative Knowledge Grounding for Veracity Assessment

Model ReleasesDGX agent

arXiv:2601.22361v2 Announce Type: replace-cross Abstract: Assessing the veracity of online content has become increasingly critical. Large language models (LLMs) have recently enabled substantial prog

MerNav: A Highly Generalizable Memory-Execute-Review Framework for Zero-Shot Object Goal Navigation

ApplicationsDGX agent

arXiv:2602.05467v2 Announce Type: replace-cross Abstract: Visual Language Navigation (VLN) is one of the fundamental capabilities for embodied intelligence and a critical challenge that urgently needs

METER: Evaluating Multi-Level Contextual Causal Reasoning in Large Language Models

Model ReleasesDGX agent

arXiv:2604.11502v1 Announce Type: cross Abstract: Contextual causal reasoning is a critical yet challenging capability for Large Language Models (LLMs). Existing benchmarks, however, often evaluate th

METRO: Towards Strategy Induction from Expert Dialogue Transcripts for Non-collaborative Dialogues

ResearchDGX agent

arXiv:2604.11427v1 Announce Type: cross Abstract: Developing non-collaborative dialogue agents traditionally requires the manual, unscalable codification of expert strategies. We propose ours, a metho

MetroGS: Efficient and Stable Reconstruction of Geometrically Accurate High-Fidelity Large-Scale Scenes

TutorialsDGX agent

arXiv:2511.19172v4 Announce Type: replace Abstract: Recently, 3D Gaussian Splatting and its derivatives have achieved significant breakthroughs in large-scale scene reconstruction. However, how to eff

MGA: Memory-Driven GUI Agent for Observation-Centric Interaction

SafetyDGX agent

arXiv:2510.24168v2 Announce Type: replace Abstract: Multimodal Large Language Models (MLLMs) have significantly advanced GUI agents, yet long-horizon automation remains constrained by two critical bot

Micro-Dexterity in Biological Micromanipulation: Embodiment, Perception, and Control

AgentsDGX agent

arXiv:2604.11640v1 Announce Type: new Abstract: Microscale manipulation has advanced substantially in controlled locomotion and targeted transport, yet many biomedical applications require precise and

Mild Over-Parameterization Benefits Asymmetric Tensor PCA

ResearchDGX agent

arXiv:2604.10208v1 Announce Type: new Abstract: Asymmetric Tensor PCA (ATPCA) is a prototypical model for studying the trade-offs between sample complexity, computation, and memory. Existing algorithm

MimicLM: Zero-Shot Voice Imitation through Autoregressive Modeling of Pseudo-Parallel Speech Corpora

SafetyDGX agent

arXiv:2604.11552v1 Announce Type: cross Abstract: Voice imitation aims to transform source speech to match a reference speaker's timbre and speaking style while preserving linguistic content. A straig

Min-k Sampling: Decoupling Truncation from Temperature Scaling via Relative Logit Dynamics

Model ReleasesDGX agent

arXiv:2604.11012v1 Announce Type: new Abstract: The quality of text generated by large language models depends critically on the decoding sampling strategy. While mainstream methods such as Top-k, Top

Minimal Embodiment Enables Efficient Learning of Number Concepts in Robot

SafetyDGX agent

arXiv:2604.11373v1 Announce Type: cross Abstract: Robots are increasingly entering human-interactive scenarios that require understanding of quantity. How intelligent systems acquire abstract numerica

Minimizing classical resources in variational measurement-based quantum computation for generative modeling

Model ReleasesDGX agent

arXiv:2604.11578v1 Announce Type: cross Abstract: Measurement-based quantum computation (MBQC) is a framework for quantum information processing in which a computational task is carried out through on

Mining Attribute Subspaces for Efficient Fine-tuning of 3D Foundation Models

ResearchDGX agent

arXiv:2604.10095v1 Announce Type: new Abstract: With the emergence of 3D foundation models, there is growing interest in fine-tuning them for downstream tasks, where LoRA is the dominant fine-tuning p

Mirai: Autoregressive Visual Generation Needs Foresight

Model ReleasesDGX agent

arXiv:2601.14671v2 Announce Type: replace Abstract: Autoregressive (AR) visual generators model images as sequences of discrete tokens and are trained with a next-token likelihood objective. This stri

Mitigating Privacy Risk via Forget Set-Free Unlearning

ResearchDGX agent

arXiv:2604.10636v1 Announce Type: new Abstract: Training machine learning models requires the storage of large datasets, which often contain sensitive or private data. Storing data is associated with

MIXAR: Scaling Autoregressive Pixel-based Language Models to Multiple Languages and Scripts

ResearchDGX agent

arXiv:2604.11575v1 Announce Type: new Abstract: Pixel-based language models are gaining momentum as alternatives to traditional token-based approaches, promising to circumvent tokenization challenges.

Mixture of Cognitive Reasoners: Modular Reasoning with Brain-Like Specialization

SafetyDGX agent

arXiv:2506.13331v3 Announce Type: replace Abstract: Human cognitive behavior arises from the interaction of specialized brain networks dedicated to distinct functions, such as language, logic, and soc

ML-Based Real-Time Downlink Performance Prediction in Standalone 5G NR Using Smartphones

ApplicationsDGX agent

arXiv:2604.09632v1 Announce Type: cross Abstract: We propose a machine learning (ML)-based framework for downlink performance prediction in 5G networks using real-time measurements from commercial off

MLLM-as-a-Judge Exhibits Model Preference Bias

Model ReleasesDGX agent

arXiv:2604.11589v1 Announce Type: new Abstract: Automatic evaluation using multimodal large language models (MLLMs), commonly referred to as MLLM-as-a-Judge, has been widely used to measure model perf

MM-LIMA: Less Is More for Alignment in Multi-Modal Datasets

SafetyDGX agent

arXiv:2308.12067v3 Announce Type: replace-cross Abstract: Multimodal large language models are typically trained in two stages: first pre-training on image-text pairs, and then fine-tuning using super

MMR-AD: A Large-Scale Multimodal Dataset for Benchmarking General Anomaly Detection with Multimodal Large Language Models

Model ReleasesDGX agent

arXiv:2604.10971v1 Announce Type: cross Abstract: In the progress of industrial anomaly detection, general anomaly detection (GAD) is an emerging trend and also the ultimate goal. Unlike the conventio

MMRareBench: A Rare-Disease Multimodal and Multi-Image Medical Benchmark

Model ReleasesDGX agent

arXiv:2604.10755v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have advanced clinical tasks for common conditions, but their performance on rare diseases remains largely unte

MobiFlow: Real-World Mobile Agent Benchmarking through Trajectory Fusion

SafetyDGX agent

arXiv:2604.09587v1 Announce Type: new Abstract: Mobile agents can autonomously complete user-assigned tasks through GUI interactions. However, existing mainstream evaluation benchmarks, such as Androi

Mobile GUI Agent Privacy Personalization with Trajectory Induced Preference Optimization

Model ReleasesDGX agent

arXiv:2604.11259v1 Announce Type: new Abstract: Mobile GUI agents powered by Multimodal Large Language Models (MLLMs) can execute complex tasks on mobile devices. Despite this progress, most existing

Modeling, Analysis and Activation of Planar Viscoelastically-combined Rimless Wheels

ResearchDGX agent

arXiv:2604.11295v1 Announce Type: new Abstract: This paper proposes novel passive-dynamic walkers formed by two cross-shaped frames and eight viscoelastic elements. Since it is a combination of two fo

Modular Delta Merging with Orthogonal Constraints: A Scalable Framework for Continual and Reversible Model Composition

ApplicationsDGX agent

arXiv:2507.20997v4 Announce Type: replace-cross Abstract: In real-world machine learning deployments, models must be continually updated, composed, and when required, selectively undone. However, exis

MoEITS: A Green AI approach for simplifying MoE-LLMs

Model ReleasesDGX agent

arXiv:2604.10603v1 Announce Type: cross Abstract: Large language models are transforming all areas of academia and industry, attracting the attention of researchers, professionals, and the general pub

MonoEM-GS: Monocular Expectation-Maximization Gaussian Splatting SLAM

SafetyDGX agent

arXiv:2604.10593v1 Announce Type: new Abstract: Feed-forward geometric foundation models can infer dense point clouds and camera motion directly from RGB streams, providing priors for monocular SLAM.

MoRI: Mixture of RL and IL Experts for Long-Horizon Manipulation Tasks

SafetyDGX agent

arXiv:2604.10165v1 Announce Type: new Abstract: Reinforcement Learning (RL) and Imitation Learning (IL) are the standard frameworks for policy acquisition in manipulation. While IL offers efficient po

MorphoFlow: Sparse-Supervised Generative Shape Modeling with Adaptive Latent Relevance

TutorialsDGX agent

arXiv:2604.11636v1 Announce Type: new Abstract: Statistical shape modeling (SSM) is central to population level analysis of anatomical variability, yet most existing approaches rely on densely annotat

MOSAIC: Multi-Domain Orthogonal Session Adaptive Intent Capture for Prescient Recommendations

SafetyDGX agent

arXiv:2604.10147v1 Announce Type: cross Abstract: Capturing user intent across heterogeneous behavioral domains stands as a fundamental challenge in session-based recommender systems. Yet, existing mu

MosaicMRI: A Diverse Dataset and Benchmark for Raw Musculoskeletal MRI

Model ReleasesDGX agent

arXiv:2604.11762v1 Announce Type: new Abstract: Deep learning underpins a wide range of applications in MRI, including reconstruction, artifact removal, and segmentation. However, progress has been dr

← Previous
1…949950951952953…989
Next →