AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlog
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
Model Releases

LARAG: Link-Aware Retrieval Strategy for RAG Systems in Hyperlinked Technical Documentation

DGX agent

arXiv:2605.07517v1 Announce Type: cross Abstract: Retrieval-Augmented Generation (RAG) enhances the factual grounding of Large Language Models by conditioning their outputs on external documents. Howe

model-releasesarxiv-cs-ai
11 May 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Research

Latent-Space Causal Discovery from Indirect Neuroimaging Observations

DGX agent

arXiv:2602.09034v2 Announce Type: replace-cross Abstract: Neuroimaging does not observe causal variables directly: hemodynamics and volume conduction distort signals so that statistical dependence nee

researcharxiv-cs-ai
11 May 2026
Model Releases

Learning and Reusing Policy Decompositions for Hierarchical Generalized Planning with LLM Agents

DGX agent

arXiv:2605.06957v1 Announce Type: new Abstract: We present a dynamic policy-learning approach that combines generalized planning and hierarchical task decomposition for LLM-based agents. Our method, H

model-releasesarxiv-cs-ai
11 May 2026
Agents

Learning CLI Agents with Structured Action Credit under Selective Observation

DGX agent

arXiv:2605.08013v1 Announce Type: new Abstract: Command line interface (CLI) agents are emerging as a practical paradigm for agent-computer interaction over evolving filesystems, executable command li

agentsarxiv-cs-ai
11 May 2026
Safety

Learning Cross-Atlas Consistent Brain Disorder Representations via Disentangled Multi-Atlas Functional Connectivity Learning

DGX agent

arXiv:2605.07026v1 Announce Type: cross Abstract: Functional connectivity (FC) derived from resting-state fMRI is widely used to characterize large-scale brain network alterations in neurological and

safetyarxiv-cs-ai
11 May 2026
Local Ai

Learning Multi-Relational Graph Representations for DNA Methylation-Based Biological Age Estimation

DGX agent

arXiv:2605.07175v1 Announce Type: cross Abstract: Aging clocks aim to estimate biological age, a measure of physiological state distinct from chronological age, from observable biomarkers, and are wid

local-aiarxiv-cs-ai
11 May 2026
Local Ai

Learning to Communicate Locally for Large-Scale Multi-Agent Pathfinding

DGX agent

arXiv:2605.07637v1 Announce Type: new Abstract: Multi-agent pathfinding (MAPF) is a widely used abstraction for multi-robot trajectory planning problems, where multiple homogeneous agents move simulta

local-aiarxiv-cs-ai
11 May 2026
Research

Learning to Pose Problems: Reasoning-Driven and Solver-Adaptive Data Synthesis

DGX agent

arXiv:2511.09907v5 Announce Type: replace Abstract: Data synthesis for training large reasoning models offers a scalable alternative to limited, human-curated datasets, enabling the creation of high-q

researcharxiv-cs-ai
11 May 2026
Safety

Learning Visual Feature-Based World Models via Residual Latent Action

DGX agent

arXiv:2605.07079v1 Announce Type: cross Abstract: World models predict future transitions from observations and actions. Existing works predominantly focus on image generation only. Visual feature-bas

safetyarxiv-cs-ai
11 May 2026
Research

LensVLM: Selective Context Expansion for Compressed Visual Representation of Text

DGX agent

arXiv:2605.07019v1 Announce Type: cross Abstract: Vision Language Models (VLMs) offer the exciting possibility of processing text as rendered images, bypassing the need for tokenizing the text into lo

researcharxiv-cs-ai
11 May 2026
Research

Limitations on Accurate, Trusted, Human-level Reasoning

DGX agent

arXiv:2509.21654v2 Announce Type: replace-cross Abstract: We identify a fundamental incompatibility between the goals of accuracy, trust, and human-level reasoning in artificial intelligence (AI) syst

researcharxiv-cs-ai
11 May 2026
Local Ai

LiteGUI: Distilling Compact GUI Agents with Reinforcement Learning

DGX agent

arXiv:2605.07505v1 Announce Type: new Abstract: Developing lightweight, on-device vision-language GUI agents is essential for efficient cross-platform automated interaction. However, current on-device

local-aiarxiv-cs-ai
11 May 2026
Model Releases

LithoBench: Benchmarking Large Multimodal Models for Remote-Sensing Lithology Interpretation

DGX agent

arXiv:2605.07640v1 Announce Type: cross Abstract: Remote sensing lithology interpretation is fundamental to geological surveys, mineral exploration, and regional geological mapping. Unlike general lan

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

LLM-Based Agents for Competitive Landscape Mapping in Drug Asset Due Diligence

DGX agent

arXiv:2508.16571v4 Announce Type: replace Abstract: In this paper, we describe and benchmark a competitor-discovery component used within an agentic AI system for fast drug asset due diligence. A comp

model-releasesarxiv-cs-ai
11 May 2026
Agents

LLM-Guided Open Hypothesis Learning from Autonomous Scanning Probe Microscopy Experiments

DGX agent

arXiv:2605.06839v1 Announce Type: cross Abstract: Autonomous experimentation has transformed microscopy and materials discovery by enabling closed-loop optimization including imaging and spectroscopy

agentsarxiv-cs-ai
11 May 2026
Applications

LLM hallucinations in the wild: Large-scale evidence from non-existent citations

DGX agent

arXiv:2605.07723v1 Announce Type: cross Abstract: Large language models (LLMs) are known to generate plausible but false information across a wide range of contexts, yet the real-world magnitude and c

applicationsarxiv-cs-ai
11 May 2026
Research

LR-SGS: Robust LiDAR-Reflectance-Guided Salient Gaussian Splatting for Self-Driving Scene Reconstruction

DGX agent

arXiv:2603.12647v2 Announce Type: replace-cross Abstract: Recent 3D Gaussian Splatting (3DGS) methods have demonstrated the feasibility of self-driving scene reconstruction and novel view synthesis. H

researcharxiv-cs-ai
11 May 2026
Research

MACS: Modality-Aware Capacity Scaling for Efficient Multimodal MoE Inference

DGX agent

arXiv:2605.05225v2 Announce Type: replace-cross Abstract: Mixture-of-Experts Multimodal Large Language Models (MoE MLLMs) suffer from a significant efficiency bottleneck during Expert Parallelism (EP)

researcharxiv-cs-ai
11 May 2026
Model Releases

Mage: Multi-Axis Evaluation of LLM-Generated Executable Game Scenes Beyond Compile-Pass Rate

DGX agent

arXiv:2605.07342v1 Announce Type: cross Abstract: Compile-pass rate is the dominant evaluation signal for LLM code generation, yet for multi-component domain-specific artifacts it can be actively misl

model-releasesarxiv-cs-ai
11 May 2026
Research

Making AI Evaluation Deployment Relevant Through Context Specification

DGX agent

arXiv:2603.06811v3 Announce Type: replace Abstract: With many organizations struggling to gain value from AI deployments, pressure to evaluate AI in an informed manner has intensified. Status quo AI e

researcharxiv-cs-ai
11 May 2026
Safety

MaPPO: Maximum a Posteriori Preference Optimization with Prior Knowledge

DGX agent

arXiv:2507.21183v5 Announce Type: replace-cross Abstract: As the era of large language models (LLMs) unfolds, Preference Optimization (PO) methods have become a central approach to aligning LLMs with

safetyarxiv-cs-ai
11 May 2026
Model Releases

MAS-Algorithm: A Workflow for Solving Algorithmic Programming Problems with a Multi-Agent System

DGX agent

arXiv:2605.05949v2 Announce Type: replace Abstract: Algorithmic problem solving serves as a rigorous testbed for evaluating structured reasoning in AI coding systems, as it directly reflects a model's

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Mask2Cause: Causal Discovery via Adjacency Constrained Causal Attention

DGX agent

arXiv:2605.07280v1 Announce Type: cross Abstract: Leveraging deep learning for causal discovery in time series remains challenging because existing neural methods predominantly rely on component-wise

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Mathematical Reasoning via Intervention-Based Time-Series Causal Discovery Using LLMs as Concept Mastery Simulators

DGX agent

arXiv:2605.07600v1 Announce Type: cross Abstract: Recent methods for improving LLM mathematical reasoning, whether through MCTS-based test-time search or causal graph-guided knowledge injection, canno

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

MathlibPR: Pull Request Merge-Readiness Benchmark for Formal Mathematical Libraries

DGX agent

arXiv:2605.07147v1 Announce Type: cross Abstract: The ecosystem of Lean and Mathlib has become the de facto standard for large language model (LLM) assisted formal reasoning with remarkable successes

model-releasesarxiv-cs-ai
11 May 2026
Agents

mathsf{VISTA}: Decentralized Machine Learning in Adversary Dominated Environments

DGX agent

arXiv:2605.07841v1 Announce Type: cross Abstract: Decentralized machine learning often relies on outsourcing computations, such as gradient evaluations, to untrusted worker nodes. Existing robust aggr

agentsarxiv-cs-ai
11 May 2026
Model Releases

MatryoshkaLoRA: Learning Accurate Hierarchical Low-Rank Representations for LLM Fine-Tuning

DGX agent

arXiv:2605.07850v1 Announce Type: cross Abstract: With the rise in scale for deep learning models to billions of parameters, the computational cost of fine-tuning remains a significant barrier to depl

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

MAVEN: Multi-Agent Verification-Elaboration Network with In-Step Epistemic Auditing

DGX agent

arXiv:2605.07646v1 Announce Type: cross Abstract: While explicit reasoning trajectories enhance model interpretability, existing paradigms often rely on monolithic chains that lack intermediate verifi

model-releasesarxiv-cs-ai
11 May 2026
Research

Mechanistic Interpretability with Sparse Autoencoder Neural Operators

DGX agent

arXiv:2509.03738v4 Announce Type: replace-cross Abstract: We introduce sparse autoencoder neural operators (SAE-NOs), a new class of sparse autoencoders that operate in function spaces rather than fix

researcharxiv-cs-ai
11 May 2026
Model Releases

MedAction: Towards Active Multi-turn Clinical Diagnostic LLMs

DGX agent

arXiv:2605.07305v1 Announce Type: cross Abstract: Most existing LLM diagnoses are evaluated on static, single-turn settings where complete patient information is provided upfront, an oversimplificatio

model-releasesarxiv-cs-ai
11 May 2026
Agents

MedExAgent: Training LLM Agents to Ask, Examine, and Diagnose in Noisy Clinical Environments

DGX agent

arXiv:2605.07058v1 Announce Type: cross Abstract: Real-world clinical diagnosis is a complex process in which the doctor is required to obtain information from both interaction with the patient and co

agentsarxiv-cs-ai
11 May 2026
Model Releases

MELD: Multi-Task Equilibrated Learning Detector for AI-Generated Text

DGX agent

arXiv:2605.06903v1 Announce Type: cross Abstract: Large language models are now embedded in everyday writing workflows, making reliable AI-generated text detection important for academic integrity, co

model-releasesarxiv-cs-ai
11 May 2026
Agents

MEMOREPAIR: Barrier-First Cascade Repair in Agentic Memory

DGX agent

arXiv:2605.07242v1 Announce Type: new Abstract: Agentic memory evolves across tasks into durable derived artifacts: summaries, cached outputs, embeddings, learned skills, and executable tool procedure

agentsarxiv-cs-ai
11 May 2026
Research

Memory-Efficient Looped Transformer: Decoupling Compute from Memory in Looped Language Models

DGX agent

arXiv:2605.07721v1 Announce Type: cross Abstract: Recurrent LLM architectures have emerged as a promising approach for improving reasoning, as they enable multi-step computation in the embedding space

researcharxiv-cs-ai
11 May 2026
Hardware

MemSearcher: Training LLMs to Reason, Search and Manage Memory via End-to-End Reinforcement Learning

DGX agent

arXiv:2511.02805v2 Announce Type: replace-cross Abstract: LLM-based search agents often concatenate the full interaction history into the context, producing long and noisy inputs, and increasing compu

hardwarearxiv-cs-ai
11 May 2026
Safety

Miner:Mining Intrinsic Mastery for Data-Efficient RL in Large Reasoning Models

DGX agent

arXiv:2601.04731v2 Announce Type: replace Abstract: Current critic-free RL methods for large reasoning models suffer from severe inefficiency when training on positive homogeneous prompts (where all r

safetyarxiv-cs-ai
11 May 2026
Model Releases

MiniAppBench: Evaluating the Shift from Text to Interactive HTML Responses in LLM-Powered Assistants

DGX agent

arXiv:2603.09652v3 Announce Type: replace Abstract: With the rapid advancement of Large Language Models (LLMs) in code generation, human-AI interaction is evolving from static text responses to dynami

model-releasesarxiv-cs-ai
11 May 2026
Applications

MinMax Recurrent Neural Cascades

DGX agent

arXiv:2605.06384v2 Announce Type: replace-cross Abstract: We show that the MinMax algebra provides a form of recurrence that is expressively powerful, efficiently implementable, and most importantly i

applicationsarxiv-cs-ai
11 May 2026
Model Releases

MISA: Mixture of Indexer Sparse Attention for Long-Context LLM Inference

DGX agent

arXiv:2605.07363v1 Announce Type: cross Abstract: DeepSeek Sparse Attention (DSA) sets the state of the art for fine-grained inference-time sparse attention by introducing a learned token-wise indexer

model-releasesarxiv-cs-ai
11 May 2026
Applications

MIST: Multimodal Interactive Speech-based Tool-calling Conversational Assistants for Smart Homes

DGX agent

arXiv:2605.06897v1 Announce Type: cross Abstract: The rise of Internet of Things (IoT) devices in the physical world necessitates voice-based interfaces capable of handling complex user experiences. W

applicationsarxiv-cs-ai
11 May 2026
Model Releases

Mitigating Cognitive Bias in RLHF by Altering Rationality

DGX agent

arXiv:2605.06895v1 Announce Type: new Abstract: How can we make models robust to even imperfect human feedback? In reinforcement learning from human feedback (RLHF), human preferences over model outpu

model-releasesarxiv-cs-ai
11 May 2026
Research

Mixture of Masters: Sparse Chess Language Models with Player Routing

DGX agent

arXiv:2602.04447v2 Announce Type: replace-cross Abstract: Modern chess language models are dense transformers trained on millions of games played by thousands of high-rated individuals. However, these

researcharxiv-cs-ai
11 May 2026
Safety

Modality Gap-Driven Subspace Alignment Training Paradigm For Multimodal Large Language Models

DGX agent

arXiv:2602.07026v2 Announce Type: replace-cross Abstract: Despite the success of multimodal contrastive learning in aligning visual and linguistic representations, a persistent geometric anomaly, the

safetyarxiv-cs-ai
11 May 2026
Model Releases

Model-Driven Policy Optimization in Differentiable Simulators via Stochastic Exploration

DGX agent

arXiv:2605.07520v1 Announce Type: new Abstract: Differentiable planning enables gradient-based optimization of decision-making problems by leveraging differentiable models of system dynamics. However,

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

More Thinking, More Bias: Length-Driven Position Bias in Reasoning Models

DGX agent

arXiv:2605.06672v1 Announce Type: new Abstract: Chain-of-thought (CoT) reasoning and reasoning-tuned models such as DeepSeek-R1 are commonly assumed to reduce shallow heuristic biases by thinking care

model-releasesarxiv-cs-ai
11 May 2026
Safety

MORPH-U: Multi-Objective Resilient Motion Planning for V2X-Enabled Autonomous Driving in High-Uncertainty Environments via Simulation

DGX agent

arXiv:2605.07370v1 Announce Type: cross Abstract: V2X can warn an autonomous vehicle about hazards beyond line-of-sight, but it also brings uncertainty: messages may be delayed, dropped, or even forge

safetyarxiv-cs-ai
11 May 2026
Research

Motion-o: Trajectory-Grounded Video Reasoning

DGX agent

arXiv:2603.18856v2 Announce Type: replace-cross Abstract: Recent video reasoning models increasingly produce spatio-temporal evidence chains that localize objects at specific timestamps. While these t

researcharxiv-cs-ai
11 May 2026
Safety

MPD^2-Router: Mask-aware Multi-expert Prior-regularized Dual-head Deferral Router in Glaucoma Screening and Diagnosis

DGX agent

arXiv:2605.08024v1 Announce Type: new Abstract: Learning-to-defer (L2D) can make glaucoma screening safer by routing difficult/uncertain cases to humans, yet standard formulations overlook expert avai

safetyarxiv-cs-ai
11 May 2026
← Previous
1…351352353354355…448
Next →