AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent
84,648Total entries
1Added by human
84,647Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
2 Jun 2026

Logit Distillation on Manifolds: Mapping by Learning

ResearchDGX agent

arXiv:2606.00771v1 Announce Type: cross Abstract: A simple way to improve the performance of almost any machine learning model is not to train a single but several models with diverse algorithms which

Lookahead Sample Reward Guidance for Test-Time Scaling of Diffusion Models

SafetyDGX agent

arXiv:2602.03211v2 Announce Type: replace-cross Abstract: Diffusion models have demonstrated strong generative performance; however, generated samples often fail to fully align with human intent. This

LookWise: Knowing When and Where to Look for Fine-Grained Visual Reasoning in Multimodal Large Language Models

Local AiDGX agent

arXiv:2603.00171v3 Announce Type: replace-cross Abstract: Multimodal Large Language Models (MLLMs) are shifting towards 'Thinking with Images' by actively exploring image details. While effective, lar


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Low-Resource Safety Failures Are Action Failures, Not Representation Failures

Model ReleasesDGX agent

arXiv:2606.01196v1 Announce Type: cross Abstract: Safety alignment learned in high-resource languages transfers poorly to low-resource languages. Models refuse harmful prompts in English but fail to r

LP5X-PIM Sim: A High-Fidelity HW/SW Integrated Simulator for LPDDR5X-PIM

ResearchDGX agent

arXiv:2606.00636v1 Announce Type: cross Abstract: This tech note describes the architecture and execution results of the LPDDR5X-PIM simulator, developed by Samsung Electronics. Based on the latest re

Make a Video Call with LLM: A Measurement Campaign over Six Mainstream Apps

Model ReleasesDGX agent

arXiv:2510.00481v2 Announce Type: replace-cross Abstract: In 2025, Large Language Model (LLM) services have launched a new feature -- AI video chat -- allowing users to interact with AI agents via rea

Make Mechanistic Interpretability Auditable: A Call to Develop Guidelines via Continuous Collaborative Reviewing

SafetyDGX agent

arXiv:2606.00033v1 Announce Type: cross Abstract: While mechanistic interpretability (MI) has produced important insights into neural network internals, the field has yet to establish a standardized s

MARFT: Multi-Agent Reinforcement Fine-Tuning

AgentsDGX agent

arXiv:2504.16129v5 Announce Type: replace-cross Abstract: Large Language Model (LLM)-based Multi-Agent Systems (LaMAS) have demonstrated strong capabilities on complex agentic tasks requiring multifac

Margin Adaptive DPO: Leveraging Reward Model for Granular Control in Preference Optimization

Model ReleasesDGX agent

arXiv:2510.05342v2 Announce Type: replace-cross Abstract: Direct Preference Optimization (DPO) has emerged as a simple and effective method for aligning large language models. However, its reliance on

MASCOT: Towards Multi-Agent Socio-Collaborative Companion Systems

SafetyDGX agent

arXiv:2601.14230v2 Announce Type: replace-cross Abstract: Multi-agent systems (MAS) are emerging as promising socio-collaborative companions for emotional and cognitive support. However, existing syst

MASER: Modality-Adaptive Specialist Routing for Embodied 3D Spatial Intelligence

Model ReleasesDGX agent

arXiv:2606.02463v1 Announce Type: cross Abstract: In 3D environments, Embodied Agents answer spatially relevant questions through reasoning from a mixture of modalities including natural language, RGB

Masking Stale Observations Helps Search Agents -- Until It Doesn't: A Regime Map and Its Mechanism

AgentsDGX agent

arXiv:2606.00408v1 Announce Type: cross Abstract: Long-horizon search agents accumulate large amounts of retrieved content across many tool calls, making context-budget efficiency increasingly importa

MCP-Persona: Benchmarking LLM Agents on Real-World Personal Applications via Environment Simulation

Model ReleasesDGX agent

arXiv:2606.02470v1 Announce Type: new Abstract: The Model Context Protocol (MCP) has emerged as a transformative standard for connecting large language models (LLMs) with external data sources and too

Measuring and Mitigating Bias in Code Generated by Large Language Models

Model ReleasesDGX agent

arXiv:2606.00049v1 Announce Type: cross Abstract: Large language models (LLMs) are widely recognised for their applications in natural language generation and are increasingly used for code generation

Med-Scout: Curing MLLMs' Geometric Blindness in Medical Perception via Geometry-Aware RL Post-Training

Model ReleasesDGX agent

arXiv:2601.23220v2 Announce Type: replace-cross Abstract: Despite recent Multimodal Large Language Models (MLLMs)' linguistic prowess in medical diagnosis, we find even state-of-the-art MLLMs suffer f

Medication-Aware Financial Exploitation Detection for Alzheimer's Patients Using Edge-Aware Interaction Risk Modeling

ResearchDGX agent

arXiv:2606.00672v1 Announce Type: new Abstract: Financial exploitation is a growing concern for people with Alzheimer's disease, especially during periods of reduced cognitive stability. Conventional

MemGraphRAG: Memory-based Multi-Agent System for Graph Retrieval-Augmented Generation

AgentsDGX agent

arXiv:2606.00610v1 Announce Type: cross Abstract: Retrieval-Augmented Generation (RAG) has become an essential method for mitigating hallucinations in Large Language Models (LLMs) by leveraging extern

Memory-Efficient LLM Training with Dynamic Sparsity: From Stability to Practical Scaling

ResearchDGX agent

arXiv:2606.00888v1 Announce Type: cross Abstract: Dynamic Sparse Training (DST) offers a promising paradigm for improving the training and inference efficiency of deep neural networks; however, we fin

MemPro: Agentic Memory Systems as Evolvable Programs

AgentsDGX agent

arXiv:2606.00619v1 Announce Type: cross Abstract: Long-horizon autonomous agents require memory systems to retain historical information, track evolving states, and reuse relevant knowledge beyond fin

MENTIS: What Belief Changes Under Alignment? Measuring Multi-Scale Latent Torsion in Language Models

Local AiDGX agent

arXiv:2606.01060v1 Announce Type: cross Abstract: Preference alignment has substantially improved the observable behavior of large language models, yet it remains unclear what alignment changes intern

MESA: Improving MoE Safety Alignment via Decentralized Expertise

SafetyDGX agent

arXiv:2606.00651v1 Announce Type: cross Abstract: Mixture-of-Experts (MoE) architectures scale Large Language Models (LLMs) efficiently, enabling greater capacity with reduced computational cost by dy

MGRegBench: A Novel Benchmark Dataset with Anatomical Landmarks for Mammography Image Registration

Model ReleasesDGX agent

arXiv:2512.17605v2 Announce Type: replace-cross Abstract: Robust mammography registration is essential for clinically relevant applications like tracking disease progression in breast tissue. However,

MiCU: End-to-End Smart Home Command Understanding with Large Language Model

ApplicationsDGX agent

arXiv:2606.01099v1 Announce Type: cross Abstract: Command understanding systems in smart home ecosystems can automate device control and substantially improve user experience. However, while they perf

MidSteer: Optimal Affine Framework for Steering Generative Models

SafetyDGX agent

arXiv:2605.05220v2 Announce Type: replace-cross Abstract: Steering intermediate representations has emerged as a powerful strategy for controlling generative models, particularly in post-deployment al

MindClaw: Closed-Loop Embodied Mental-State Reasoning for Precision Intervention

AgentsDGX agent

arXiv:2606.01063v1 Announce Type: new Abstract: Theory of Mind (ToM) enables an agent to reason about another actor's beliefs, goals, and intentions, which is essential for human-centered embodied ass

MindGames Arena Generalization Track: In2AI Solution with Delayed Per-Step Reward Attribution

Model ReleasesDGX agent

arXiv:2606.00017v1 Announce Type: new Abstract: Training language model agents for multi-agent strategic interaction presents a core difficulty: the quality of any action may depend on future events t

MindZero: Learning Online Mental Reasoning With Zero Annotations

ApplicationsDGX agent

arXiv:2606.00240v1 Announce Type: new Abstract: Effective real-world assistance requires AI agents with robust Theory of Mind (ToM): inferring human mental states from their behavior. Despite recent a

MineDraft: A Framework for Batch Parallel Speculative Decoding

ApplicationsDGX agent

arXiv:2603.18016v2 Announce Type: replace-cross Abstract: Speculative decoding (SD) accelerates large language model inference by using a smaller draft model to propose draft tokens that are subsequen

MINTS: Minimalist Thompson Sampling

ResearchDGX agent

arXiv:2606.01655v1 Announce Type: cross Abstract: The Bayesian paradigm offers principled tools for sequential decision-making under uncertainty, but its reliance on a probabilistic model for all para

Mitigating Hallucinations in Large Language Models Via Decoder Layer Skipping

ResearchDGX agent

arXiv:2606.00819v1 Announce Type: new Abstract: Large Language Models (LLMs) have achieved strong performance across diverse natural language tasks, yet their outputs often suffer from hallucinations

Mitigating Perceptual Judgment Bias in Multimodal LLM-as-a-Judge via Perceptual Perturbation and Reward Modeling

SafetyDGX agent

arXiv:2606.02578v1 Announce Type: cross Abstract: Recent multimodal large language models have demonstrated strong reasoning ability, yet their reliability as automated evaluators remains limited by a

Mitigating Reward Hacking in RLHF via Bayesian Non-negative Reward Modeling

ResearchDGX agent

arXiv:2602.10623v2 Announce Type: replace-cross Abstract: Reward models learned from human preferences are central to aligning large language models (LLMs) via reinforcement learning from human feedba

MLLM-Microscope: Unlocking Hidden Structure Within Multimodal Large Language Models

ResearchDGX agent

arXiv:2606.00909v1 Announce Type: cross Abstract: This work presents MLLM-Microscope, a novel system designed for analyzing the hidden representations within Multimodal Large Language Models (MLLMs).

MMG2Skill: Can Agents Distill In-the-Wild Guides into Self-Evolving Skills?

Model ReleasesDGX agent

arXiv:2606.01993v1 Announce Type: cross Abstract: Abundant procedural knowledge on the Web holds great potential for helping agents solve long-horizon tasks. However, such knowledge is often multimoda

MobEvolve: An Agentic Self-Evolving Heuristic System for Interpretable Human Mobility Generation

SafetyDGX agent

arXiv:2606.01640v1 Announce Type: new Abstract: Human mobility generation aims to synthesize realistic trip chains for target populations based on individual features. Existing paradigms, including de

MOC: Multi-Order Communication in LLM-based Multi-Agent Systems

Model ReleasesDGX agent

arXiv:2606.02359v1 Announce Type: new Abstract: Despite the remarkable progress of Large Language Model (LLM) based Multi-Agent Systems, most research focuses on optimizing coordination topology while

Model-Native Computing Architecture: Envisioning Future System Architecture Through the Lens of Computer Architecture

Model ReleasesDGX agent

arXiv:2606.00288v1 Announce Type: new Abstract: Large language models are undergoing a transition from model technology to system technology. As developers use Codex, Claude Code, AutoGPT, and related

Model Parallelism With Subnetwork Data Parallelism

Model ReleasesDGX agent

arXiv:2507.09029v5 Announce Type: replace-cross Abstract: Pre-training large neural networks at scale imposes heavy memory demands on accelerators and often requires costly communication. We introduce

Modeling Depth Ambiguity: A Mixture-Density Representation for Flying-Point-Free Depth Estimation

ResearchDGX agent

arXiv:2606.02552v1 Announce Type: cross Abstract: Despite advances in depth estimation, flying points remain a persistent failure mode: near object boundaries, depth estimators often predict spurious

MoEIoU: Rethinking Bounding-Box Regression as a Mixture of Experts

SafetyDGX agent

arXiv:2606.00844v1 Announce Type: cross Abstract: Bounding-box regression is a fundamental component of object detection, playing a critical role in precise object localization. Existing Intersection-

Moment-Video: Diagnosing Temporal Fidelity of Video MLLMs on Momentary Visual Events

Model ReleasesDGX agent

arXiv:2606.02522v1 Announce Type: cross Abstract: Video multimodal large language models (MLLMs) have made rapid progress on general and long-form video understanding, yet their ability to preserve br

Monitoring Agentic Systems Before They're Reliable

AgentsDGX agent

arXiv:2606.02494v1 Announce Type: cross Abstract: Agentic systems entering production typically operate as partially integrated assemblies where structural defects, not task-level errors, dominate the

MOSAIC: Modular Orchestration for Structured Agentic Intelligence and Composition

SafetyDGX agent

arXiv:2606.00708v1 Announce Type: new Abstract: Automated data science is a structured model-selection problem. A solution must choose data transformations, feature representations, architecture, trai

MOSS-Audio Technical Report

ResearchDGX agent

arXiv:2606.01802v1 Announce Type: cross Abstract: MOSS-Audio is a unified audio-language model for speech, environmental sound, and music understanding, supporting audio captioning, time-aware questio

Motif-based morphology signatures for interpretable ECG screening and monitoring

ResearchDGX agent

arXiv:2606.00107v1 Announce Type: cross Abstract: Electrocardiography (ECG) remains central to cardiovascular screening, yet interpretation remains largely manual and episodic. Clinical practice relie

Move the Query, Not the Cache: Characterizing Cross-Instance Latent Attention Redistribution Across GPU Fabrics

Model ReleasesDGX agent

arXiv:2606.01502v1 Announce Type: cross Abstract: Frontier LLMs increasingly decide what a query attends to with a sparse-attention indexer that picks a few KV-cache blocks per query: attention's unit

MulFeRL: Enhancing Reinforcement Learning with Verbal Feedback in a Multi-turn Loop

TutorialsDGX agent

arXiv:2601.22900v2 Announce Type: replace Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) is widely used to improve reasoning across domains, but outcome-only scalar rewards are often

Multi-Agent Conformal Prediction with Personalized Statistical Validity

AgentsDGX agent

arXiv:2606.00717v1 Announce Type: cross Abstract: Uncertainty quantification is essential in high-stakes machine learning tasks. However, one of the principled solutions, conformal prediction, faces c

Multi-Contrast MRI Motion Correction via Parameter-Informed Disentanglement and Adaptive Experts

Model ReleasesDGX agent

arXiv:2606.00146v1 Announce Type: cross Abstract: Motion artifacts in magnetic resonance imaging (MRI) degrade diagnostic reliability. Existing deep learning methods are typically contrast-specific an

Multi-Objective Reinforcement Learning for Tactical Decision Making for Trucks in Highway Traffic

SafetyDGX agent

arXiv:2601.18783v2 Announce Type: replace-cross Abstract: Balancing safety, efficiency, and operational costs in highway driving poses a challenging decision-making problem for heavy-duty vehicles. A

Multilingual Idioms in Sentences and Conversations Across High-, Medium-, and Low-Resource Languages

ResearchDGX agent

arXiv:2606.02147v1 Announce Type: cross Abstract: Idiomatic expressions pose a major challenge for multilingual NLP because their meanings shift between figurative and literal usage, often requiring c

Multilinguality of Large Language Models From a Structural Perspective

ResearchDGX agent

arXiv:2606.01800v1 Announce Type: cross Abstract: Large language models (LLMs) have excelled in processing multiple languages through pre- and post-training on multilingual data, even though English d

Multimodal Approaches for Visually-Rich Document Type Classification: A Comparative Analysis

Model ReleasesDGX agent

arXiv:2606.02162v1 Announce Type: cross Abstract: Document type classification in visually rich documents remains challenging, as relevant information is distributed across textual, visual, and layout

Multimodal Function Vectors for Visual Relations

Local AiDGX agent

arXiv:2510.02528v2 Announce Type: replace Abstract: Large Multimodal Models (LMMs) demonstrate impressive in-context learning abilities from few multimodal demonstrations, yet the internal mechanisms

Multimodal Music Recommendation System using LLMs

Model ReleasesDGX agent

arXiv:2606.00125v1 Announce Type: cross Abstract: Music recommendation systems typically treat songs as opaque tokens, relying on collaborative interaction histories which overlooks semantic or acoust

MURMUR: An Efficient Inference System for Long-Form ASR

SafetyDGX agent

arXiv:2606.01483v1 Announce Type: cross Abstract: Long-form automatic speech recognition (ASR) requires both high accuracy and low latency, but existing systems force a trade-off between the two. Chun

MViewRouter: Internalizing Geometric Equivariance via Multi-view Alternating Attention for Combinatorial Routing

SafetyDGX agent

arXiv:2606.01084v1 Announce Type: cross Abstract: Combinatorial routing problems such as the Traveling Salesman Problem (TSP) and the Capacitated Vehicle Routing Problem (CVRP) are fundamental NP-hard

MyoSem: Aligning Electromyography to Natural-Language Action Semantics for Hand Action Understanding

SafetyDGX agent

arXiv:2606.00174v1 Announce Type: cross Abstract: Electromyography (EMG) directly reflects muscle activation and is a key sensing modality for gesture recognition, prosthetic control, and wearable int

naPINN: Noise-Adaptive Physics-Informed Neural Networks for Recovering Physics from Corrupted Measurement

Model ReleasesDGX agent

arXiv:2602.02547v2 Announce Type: replace-cross Abstract: Physics-Informed Neural Networks (PINNs) are effective methods for solving inverse problems and discovering governing equations from observati

NBQ: Next-Best-Question for Dynamic Profiling

ApplicationsDGX agent

arXiv:2606.00809v1 Announce Type: new Abstract: Many real-world conversational settings for knowledge discovery, including podcasts, hiring screens, and marketplaces, require a purpose-driven understa

← Previous
1…176177178179180…358
Next →