AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
12 May 2026

A Versatile AI Agent for Rare Disease Diagnosis and Risk Gene Prioritization

AgentsDGX agent

arXiv:2605.06226v2 Announce Type: replace Abstract: Accurate and timely diagnosis is essential for effective treatment, particularly in the context of rare diseases. However, current diagnostic workfl

Absurd World: A Simple Yet Powerful Method to Absurdify the Real-world for Probing LLM Reasoning Capabilities

ApplicationsDGX agent

arXiv:2605.09678v1 Announce Type: new Abstract: While extremely powerful and versatile at various tasks, the thinking capabilities of large language models (LLMs) are often put under scrutiny as they

Acceptance Cards:A Four-Diagnostic Standard for Safe Fine-Tuning Defense Claims

Model ReleasesDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2605.10575v1 Announce Type: cross Abstract: Safe fine-tuning defenses are often endorsed on the basis of a held-out gap reduction, but the same reduction can come from sampling noise, subject ar

Ace-Skill: Bootstrapping Multimodal Agents with Prioritized and Clustered Evolution

AgentsDGX agent

arXiv:2605.08887v1 Announce Type: new Abstract: Self-evolving agents present a promising path toward continual adaptation by distilling task interactions into reusable knowledge artifacts. In practice

ActivationReasoning: Logical Reasoning in Latent Activation Spaces

SafetyDGX agent

arXiv:2510.18184v3 Announce Type: replace-cross Abstract: Large language models (LLMs) excel at generating fluent text, but their internal reasoning remains opaque and difficult to control. Sparse aut

Active Learning for Gaussian Process Regression Under Self-Induced Boltzmann Weights

TutorialsDGX agent

arXiv:2605.10654v1 Announce Type: cross Abstract: We consider the active learning problem where the goal is to learn an unknown function with low prediction error under an unknown Boltzmann distributi

Active Tabular Augmentation via Policy-Guided Diffusion Inpainting

SafetyDGX agent

arXiv:2605.10315v1 Announce Type: cross Abstract: Generative tabular augmentation is appealing in data-scarce domains, yet the prevailing focus on distributional fidelity does not reliably translate i

Active Testing of Large Language Models via Approximate Neyman Allocation

ResearchDGX agent

arXiv:2605.10075v1 Announce Type: new Abstract: Large language models (LLMs) require reliable evaluation from pre-training to test-time scaling, making evaluation a recurring rather than one-off cost.

AdaPreLoRA: Adafactor Preconditioned Low-Rank Adaptation

Model ReleasesDGX agent

arXiv:2605.08734v1 Announce Type: cross Abstract: Low-Rank Adaptation (LoRA) reparameterizes a weight update as a product of two low-rank factors, but the Jacobian J_{G} of the generator mapping the f

Adaptive Action Chunking via Multi-Chunk Q Value Estimation

AgentsDGX agent

arXiv:2605.10044v1 Announce Type: cross Abstract: Action chunking emerged as a pivotal technique in imitation learning, enabling policies to predict cohesive action sequences rather than single action

Adaptive Data Harvesting for Efficient Neural Network Learning with Universal Constraints

SafetyDGX agent

arXiv:2605.09707v1 Announce Type: cross Abstract: Training neural networks to satisfy universal constraints over continuous domains poses unique challenges. Common examples include Lyapunov Neural Net

Adaptive DNN Partitioning and Offloading in Heterogeneous Edge-Cloud Continuum

ResearchDGX agent

arXiv:2605.09623v1 Announce Type: cross Abstract: In recent years, the use of artificial intelligence on resource-constrained IoT devices has grown significantly. However, existing approaches to DNN p

Agent-First Tool API: A Semantic Interface Paradigm for Enterprise AI Agent Systems

AgentsDGX agent

arXiv:2605.10555v1 Announce Type: new Abstract: As AI agents transition from research prototypes to enterprise production systems, the tool interfaces they consume remain rooted in human-oriented CRUD

Agent-Omit: Adaptive Context Omission for Efficient LLM Agents

SafetyDGX agent

arXiv:2602.04284v2 Announce Type: replace Abstract: Managing agent context (e.g., thought and observation) during multi-turn agent-environment interactions is an emerging strategy to improve agent eff

Agent-Sentry: Bounding LLM Agents via Execution Provenance

SafetyDGX agent

arXiv:2603.22868v2 Announce Type: replace-cross Abstract: Agentic computing systems, while immensely capable, raise serious security, privacy, and safety concerns. A key issue is that the full set of

Agent-ValueBench: A Comprehensive Benchmark for Evaluating Agent Values

Model ReleasesDGX agent

arXiv:2605.10365v1 Announce Type: new Abstract: Autonomous agents have rapidly matured as task executors and seen widespread deployment via harnesses such as OpenClaw. Safety concerns have rightly dra

Agent-X: Full Pipeline Acceleration of On-device AI Agents

Local AiDGX agent

arXiv:2605.10380v1 Announce Type: new Abstract: LLM-based agents deliver state-of-the-art performance across tasks but incur high end-to-end latency on edge devices. We introduce Agent-X, a software-o

AgentCollabBench: Diagnosing When Good Agents Make Bad Collaborators

Model ReleasesDGX agent

arXiv:2605.08647v1 Announce Type: cross Abstract: Multi-agent systems achieve state-of-the-art outcomes through peer collaboration. However, when an agent in the pipeline silently drops a constraint,

AgentForesight: Online Auditing for Early Failure Prediction in Multi-Agent Systems

Model ReleasesDGX agent

arXiv:2605.08715v1 Announce Type: cross Abstract: LLM-based multi-agent systems are increasingly deployed on long-horizon tasks, but a single decisive error is often accepted by downstream agents and

Agentic AI Scientists Are Not Built For Autonomous Scientific Discovery

AgentsDGX agent

arXiv:2605.08956v1 Announce Type: new Abstract: A growing body of work pursues AI scientists capable of end-to-end autonomous scientific discovery. This position paper argues that although they alread

Agentic MIP Research: Accelerated Constraint Handler Generation

Model ReleasesDGX agent

arXiv:2605.09186v1 Announce Type: new Abstract: Mixed-integer programming (MIP) research is both mathematically sophisticated and engineering-intensive: testing an algorithmic hypothesis within a bran

Agentic Performance at the Edge: Insights from Benchmarking

Model ReleasesDGX agent

arXiv:2605.10384v1 Announce Type: new Abstract: Agentic artificial intelligence (AI) is a natural fit for Internet of Things (IoT) and edge systems, but edge deployments are often constrained to model

AgentPSO: Evolving Agent Reasoning Skill via Multi-agent Particle Swarm Optimization

Model ReleasesDGX agent

arXiv:2605.08704v1 Announce Type: new Abstract: Multi-agent reasoning has shown promise for improving the problem-solving ability of large language models by allowing multiple agents to explore divers

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks

Model ReleasesDGX agent

arXiv:2605.10286v1 Announce Type: new Abstract: Building effective clinical decision support systems requires the synthesis of complex heterogeneous multimodal data. Such modalities include temporal e

AHD Agent: Agentic Reinforcement Learning for Automatic Heuristic Design

Model ReleasesDGX agent

arXiv:2605.08756v1 Announce Type: new Abstract: Automatic heuristic design (AHD) has emerged as a promising paradigm for solving NP-hard combinatorial optimization problems (COPs). Recent works show t

AI-Care: A Conversational Agentic System for Task Coordination in Alzheimer's Disease Care

SafetyDGX agent

arXiv:2605.08480v1 Announce Type: new Abstract: Individuals with Alzheimer's disease (AD) and Alzheimer's disease-related dementia (ADRD) experience memory and thinking changes that impact their abili

AI Native Asset Intelligence

ApplicationsDGX agent

arXiv:2605.09115v1 Announce Type: cross Abstract: Modern security environments generate fragmented signals across cloud resources, identities, configurations, and third-party security tools. Although

AIPO: : Learning to Reason from Active Interaction

SafetyDGX agent

arXiv:2605.08401v1 Announce Type: cross Abstract: Recent advances in large language models (LLMs) have demonstrated remarkable reasoning capabilities, largely stimulated by Reinforcement Learning with

ALAM: Algebraically Consistent Latent Transitions for Vision-Language-Action Models

SafetyDGX agent

arXiv:2605.10819v1 Announce Type: cross Abstract: Vision-language-action (VLA) models remain constrained by the scarcity of action-labeled robot data, whereas action-free videos provide abundant evide

Align and Shine: Building High-Quality Sentence-Aligned Corpora for Multilingual Text Simplification

SafetyDGX agent

arXiv:2605.09476v1 Announce Type: cross Abstract: Text simplification plays a crucial role in improving the accessibility and comprehensibility of written information for diverse audiences, including

Aligning Agents via Planning: A Benchmark for Trajectory-Level Reward Modeling

Model ReleasesDGX agent

arXiv:2604.08178v2 Announce Type: replace Abstract: In classical Reinforcement Learning from Human Feedback (RLHF), Reward Models (RMs) serve as the fundamental signal provider for model alignment. As

Alignment as Jurisprudence

SafetyDGX agent

arXiv:2605.08416v1 Announce Type: new Abstract: Jurisprudence, the study of how judges should properly decide cases, and alignment, the science of getting AI models to conform to human values, share a

AllocMV: Optimal Resource Allocation for Music Video Generation via Structured Persistent State

ResearchDGX agent

arXiv:2605.10723v1 Announce Type: cross Abstract: Generating long-horizon music videos (MVs) is frequently constrained by prohibitive computational costs and difficulty maintaining cross-shot consiste

AlpsBench: An LLM Personalization Benchmark for Real-Dialogue Memorization and Preference Alignment

Model ReleasesDGX agent

arXiv:2603.26680v2 Announce Type: replace-cross Abstract: As Large Language Models (LLMs) evolve into lifelong AI assistants, LLM personalization has become a critical frontier. However, progress is c

Ambig-DS: A Benchmark for Task-Framing Ambiguity in Data-Science Agents

Model ReleasesDGX agent

arXiv:2605.09698v1 Announce Type: new Abstract: As data-science agents shift from co-pilots to auto-pilots, silent misframing becomes a critical failure mode. Agents quietly commit to plausible but un

An agentic framework for gravitational-wave counterpart association in the multi-messenger era

AgentsDGX agent

arXiv:2605.10584v1 Announce Type: cross Abstract: With the detection of gravitational waves (GWs), multi-messenger astronomy has opened a new window for advancing our understanding of astrophysics, de

An Empirical Study of Multi-Agent Collaboration for Automated Research

Model ReleasesDGX agent

arXiv:2603.29632v2 Announce Type: replace-cross Abstract: As AI agents evolve, the community is rapidly shifting from single Large Language Models (LLMs) to Multi-Agent Systems (MAS) to overcome cogni

An Explainable Unsupervised-to-Supervised Machine Learning Framework for Dietary Pattern Discovery Using UK National Dietary Survey Data

ResearchDGX agent

arXiv:2605.08242v1 Announce Type: cross Abstract: Clinical dietary assessment can generate detailed but high-dimensional nutrient and food-group information that is difficult to translate quickly into

An Uncertainty-Aware Resilience Micro-Agent for Causal Observability in the Computing Continuum

Local AiDGX agent

arXiv:2605.10718v1 Announce Type: cross Abstract: Grey failures in the computing continuum produce ambiguous overlapping symptoms that existing approaches fail to diagnose reliably, either due to a la

AnomalyClaw: A Universal Visual Anomaly Detection Agent via Tool-Grounded Refutation

Model ReleasesDGX agent

arXiv:2605.10397v1 Announce Type: cross Abstract: Visual anomaly detection (VAD) is crucial in many real-world fields, such as industrial inspection, medical imaging, infrastructure monitoring, and re

Any2Any 3D Diffusion Models with Knowledge Transfer: A Radiotherapy Planning Study

ResearchDGX agent

arXiv:2605.09622v1 Announce Type: cross Abstract: Voxel-wise dose prediction is a critical yet challenging task in practical radiotherapy (RT) planning, as bespoke models trained from scratch often st

APCD: Adaptive Path-Contrastive Decoding for Reliable Large Language Model Generation

TutorialsDGX agent

arXiv:2605.09492v1 Announce Type: cross Abstract: Large language models (LLMs) often suffer from hallucinations due to error accumulation in autoregressive decoding, where suboptimal early token choic

AR-VLA: True Autoregressive Action Expert for Vision-Language-Action Models

ResearchDGX agent

arXiv:2603.10126v2 Announce Type: replace-cross Abstract: We propose a standalone autoregressive (AR) Action Expert that generates actions as a continuous causal sequence while conditioning on refresh

Arcane: An Assertion Reduction Framework through Semantic Clustering and MCTS-Guided Rule Exploring

Model ReleasesDGX agent

arXiv:2605.10107v1 Announce Type: new Abstract: Assertion-based Verification (ABV) is essential for ensuring that hardware designs conform to their intended specifications. However, existing automated

ArchRAG: Attributed Community-based Hierarchical Retrieval-Augmented Generation

ResearchDGX agent

arXiv:2502.09891v4 Announce Type: replace-cross Abstract: Retrieval-Augmented Generation (RAG) has proven effective in integrating external knowledge into large language models (LLMs) for solving ques

Artificial Intelligence in Number Theory: LLMs for Algorithm Generation and Ensemble Methods for Conjecture Verification

Model ReleasesDGX agent

arXiv:2504.19451v3 Announce Type: cross Abstract: This paper presents two concrete applications of Artificial Intelligence to algorithmic and analytic number theory. Recent benchmarks of large languag

ASIA: an Autonomous System Identification Agent

AgentsDGX agent

arXiv:2605.10480v1 Announce Type: new Abstract: Over the years, research in system identification has provided a rich set of methods for learning dynamical models, together with well-established theor

AssayBench: An Assay-Level Virtual Cell Benchmark for LLMs and Agents

Model ReleasesDGX agent

arXiv:2605.10876v1 Announce Type: cross Abstract: Recent advances in machine learning and large-scale biological data collections have revived the prospect of building a virtual cell, a computational

Assessing Trustworthiness of AI Training Dataset using Subjective Logic -- A Use Case on Bias

SafetyDGX agent

arXiv:2508.13813v2 Announce Type: replace-cross Abstract: As AI systems increasingly rely on training data, assessing dataset trustworthiness has become critical, particularly for properties like fair

Assessment of RAG and Fine-Tuning for Industrial Question-Answering-Applications

ApplicationsDGX agent

arXiv:2605.09533v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly employed in enterprise question-answering (QA) systems, requiring adaptation to domain-specific knowledg

AtteConDA: Attention-Based Conflict Suppression in Multi-Condition Diffusion Models and Synthetic Data Augmentation

AgentsDGX agent

arXiv:2605.09425v1 Announce Type: cross Abstract: Recent conditional image generation methods can improve controllability by generating images that are faithful to conditions such as sketches, human p

Attention-based graph neural networks: a survey

TutorialsDGX agent

arXiv:2605.08679v1 Announce Type: cross Abstract: Graph neural networks (GNNs) aim to learn well-trained representations in a lower-dimension space for downstream tasks while preserving the topologica

Attention Drift: What Autoregressive Speculative Decoding Models Learn

TutorialsDGX agent

arXiv:2605.09992v1 Announce Type: cross Abstract: Speculative decoding accelerates LLM inference by drafting future tokens with a small model, but drafter models degrade sharply under template perturb

Attention-Mamba: A Mamba-Enhanced Multi-Scale Parallel Inference Network for Medical Image Segmentation

SafetyDGX agent

arXiv:2402.02286v4 Announce Type: replace-cross Abstract: U-shaped architectures have long dominated the field of medical image segmentation, while Transformers are widely employed for modeling long-r

Attractor-Vascular Coupling Theory: Formal Grounding and Empirical Validation for AAMI-Standard Cuffless Blood Pressure Estimation from Smartphone Photoplethysmography

ResearchDGX agent

arXiv:2605.10871v1 Announce Type: cross Abstract: This work proposes Attractor-Vascular Coupling Theory (AVCT), a mathematical framework showing that cardiac attractor geometry encodes blood pressure

Attribution-based Explanations for Markov Decision Processes

ResearchDGX agent

arXiv:2605.09780v1 Announce Type: new Abstract: Attribution techniques explain the outcome of an AI model by assigning a numerical score to its inputs. So far, these techniques have mainly focused on

AU-Harness: An Open-Source Toolkit for Holistic Evaluation of Audio LLMs

ResearchDGX agent

arXiv:2509.08031v3 Announce Type: replace-cross Abstract: Large Audio Language Models (LALMs) are rapidly advancing, but evaluating them remains challenging due to inefficient and non-standardized too

Auditing Data Membership in Reinforcement Learning With Verifiable Rewards

SafetyDGX agent

arXiv:2511.14045v2 Announce Type: replace-cross Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has become a core training stage in recent large language models (LLMs). Its reliance on

Auto-Rubric as Reward: From Implicit Preferences to Explicit Multimodal Generative Criteria

SafetyDGX agent

arXiv:2605.08354v1 Announce Type: new Abstract: Aligning multimodal generative models with human preferences demands reward signals that respect the compositional, multi-dimensional structure of human

Automated Approach for Solving Infinite-state Polynomial Reachability Games

Model ReleasesDGX agent

arXiv:2605.10169v1 Announce Type: new Abstract: Reachability games are two-player games played on a graph, where the objective of exttt{REACH} player is to reach the target set whereas the objective o

← Previous
1…262263264265266…358
Next →