AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
Research

Accelerated and data-efficient flow prediction in stirred tanks via physics-informed learning

DGX agent

arXiv:2605.07444v1 Announce Type: cross Abstract: The simulation of fluid flows is computationally expensive due to the complexity of its governing partial differential equations. Machine learning mod

researcharxiv-cs-ai
11 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Activation Differences Reveal Backdoors: A Comparison of SAE Architectures

DGX agent

arXiv:2605.07324v1 Announce Type: cross Abstract: Backdoor attacks on language models pose a significant threat to AI safety, where models behave normally on most inputs but exhibit harmful behavior w

safetyarxiv-cs-ai
11 May 2026
Agents

Active Learning for Communication Structure Optimization in LLM-Based Multi-Agent Systems

DGX agent

arXiv:2605.05703v2 Announce Type: replace-cross Abstract: Optimizing the communication structure of large language model based multi-agent systems (LLM-MAS) has been shown to improve downstream perfor

agentsarxiv-cs-ai
11 May 2026
Tutorials

Active teacher selection for reward learning

DGX agent

arXiv:2310.15288v3 Announce Type: replace Abstract: Reward learning techniques enable machine learning systems to learn objectives from human feedback. A core limitation of these systems is their assu

tutorialsarxiv-cs-ai
11 May 2026
Research

AdaCorrection: Adaptive Offset Cache Correction for Accurate Diffusion Transformers

DGX agent

arXiv:2602.13357v2 Announce Type: replace-cross Abstract: Diffusion Transformers (DiTs) achieve state-of-the-art performance in high-fidelity image and video generation but suffer from expensive infer

researcharxiv-cs-ai
11 May 2026
Model Releases

Adapting Vision-Language Models for Neutrino Event Classification in High-Energy Physics

DGX agent

arXiv:2509.08461v3 Announce Type: replace-cross Abstract: Recent advances in Large Language Models (LLMs) have demonstrated their remarkable capacity to process and reason over structured and unstruct

model-releasesarxiv-cs-ai
11 May 2026
Research

Adaptive auditing of AI systems with anytime-valid guarantees

DGX agent

arXiv:2605.07002v1 Announce Type: new Abstract: A major bottleneck in characterizing the failure modes of generative AI systems is the cost and time of annotation and evaluation. Consequently, adaptiv

researcharxiv-cs-ai
11 May 2026
Model Releases

Adaptive Memory Decay for Log-Linear Attention

DGX agent

arXiv:2605.06946v1 Announce Type: cross Abstract: Sequence models face a fundamental tradeoff between memory capacity and computational efficiency. Transformers achieve expressive context modeling at

model-releasesarxiv-cs-ai
11 May 2026
Research

Adaptive Negative Reinforcement for LLM Reasoning:Dynamically Balancing Correction and Diversity in RLVR

DGX agent

arXiv:2605.07137v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) has become a highly effective method for improving the reasoning abilities of Large Language Mod

researcharxiv-cs-ai
11 May 2026
Research

AdaTKG: Adaptive Memory for Temporal Knowledge Graph Reasoning

DGX agent

arXiv:2605.07121v1 Announce Type: new Abstract: Temporal knowledge graphs (TKGs) represent time-stamped relational facts and support a wide range of reasoning tasks over evolving events. However, exis

researcharxiv-cs-ai
11 May 2026
Model Releases

AgentEscapeBench: Evaluating Out-of-Domain Tool-Grounded Reasoning in LLM Agents

DGX agent

arXiv:2605.07926v1 Announce Type: new Abstract: As LLM-based agents increasingly rely on external tools, it is important to evaluate their ability to sustain tool-grounded reasoning beyond familiar wo

model-releasesarxiv-cs-ai
11 May 2026
Agents

Agentic AI and the Industrialization of Cyber Offense: Forecast, Consequences, and Defensive Priorities for Enterprises and the Mittelstand

DGX agent

arXiv:2605.06713v1 Announce Type: cross Abstract: Agentic AI systems can plan, call tools, inspect code, interact with web applications, and coordinate multi-step workflows. These same capabilities ch

agentsarxiv-cs-ai
11 May 2026
Safety

Agentic Coding Needs Proactivity, Not Just Autonomy

DGX agent

arXiv:2605.06717v1 Announce Type: cross Abstract: Coding agents are rapidly changing the landscape of software development, moving from inline completion to autonomous systems that edit repositories,

safetyarxiv-cs-ai
11 May 2026
Model Releases

Agentick: A Unified Benchmark for General Sequential Decision-Making Agents

DGX agent

arXiv:2605.06869v1 Announce Type: new Abstract: AI agent research spans a wide spectrum: from RL agents that learn from scratch to foundation model agents that leverage pre-trained knowledge, yet no u

model-releasesarxiv-cs-ai
11 May 2026
Agents

AgentProg: Empowering Long-Horizon GUI Agents with Program-Guided Context Management

DGX agent

arXiv:2512.10371v2 Announce Type: replace Abstract: The rapid development of mobile GUI agents has stimulated growing research interest in long-horizon task automation. However, building agents for th

agentsarxiv-cs-ai
11 May 2026
Agents

AGWM: Affordance-Grounded World Models for Environments with Compositional Prerequisites

DGX agent

arXiv:2605.06841v1 Announce Type: new Abstract: In model-based learning, the agent learns behaviors by simulating trajectories based on world model predictions. Standard world models typically learn a

agentsarxiv-cs-ai
11 May 2026
Research

AI and Consciousness: Shifting Focus Towards Tractable Questions

DGX agent

arXiv:2605.06965v1 Announce Type: cross Abstract: As language-based AI systems become more anthropomorphic, the question of whether they can have subjective experience is increasingly pressing. I focu

researcharxiv-cs-ai
11 May 2026
Model Releases

AI CFD Scientist: Toward Open-Ended Computational Fluid Dynamics Discovery with Physics-Aware AI Agents

DGX agent

arXiv:2605.06607v2 Announce Type: replace-cross Abstract: Recent LLM-based agents have closed substantial portions of the scientific discovery loop in software-only machine-learning research, in chemi

model-releasesarxiv-cs-ai
11 May 2026
Agents

Alternating Target-Path Planning for Scalable Multi-Agent Coordination

DGX agent

arXiv:2605.07744v1 Announce Type: new Abstract: The concurrent target assignment and pathfinding (TAPF) problem extends multi-agent pathfinding (MAPF) by asking planners to allocate distinct targets a

agentsarxiv-cs-ai
11 May 2026
Research

Amortized-Precision Quantization for Early-Exit Vision Transformers

DGX agent

arXiv:2605.07317v1 Announce Type: cross Abstract: Vision Transformers (ViTs) achieve strong performance across vision tasks, yet their deployment with low-precision early exiting remains fragile. Exis

researcharxiv-cs-ai
11 May 2026
Hardware

An Efficient Hybrid Sparse Attention with CPU-GPU Parallelism for Long-Context Inference

DGX agent

arXiv:2605.07719v1 Announce Type: cross Abstract: Long-context inference increasingly operates over CPU-resident KV caches, either because decoding-time KV states exceed GPU memory capacity or because

hardwarearxiv-cs-ai
11 May 2026
Model Releases

An Embarrassingly Simple Graph Heuristic Reveals Shortcut-Solvable Benchmarks for Sequential Recommendation

DGX agent

arXiv:2605.07125v1 Announce Type: cross Abstract: Sequential recommendation has increasingly shifted toward generative recommenders that combine sequential patterns with semantic item information. Yet

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

An Interpretable and Scalable Framework for Evaluating Large Language Models

DGX agent

arXiv:2605.07046v1 Announce Type: cross Abstract: Evaluation of large language models (LLMs) is increasingly critical, yet standard benchmarking methods rely on average accuracy, overlooking both the

model-releasesarxiv-cs-ai
11 May 2026
Safety

APEX: Assumption-free Projection-based Embedding eXamination Metric for Image Quality Assessment

DGX agent

arXiv:2605.07786v1 Announce Type: cross Abstract: As generative models achieve unprecedented visual quality, the gold standard for image evaluation remains traditional feature-distribution metrics (e.

safetyarxiv-cs-ai
11 May 2026
Safety

Approximation-Free Differentiable Oblique Decision Trees

DGX agent

arXiv:2605.07837v1 Announce Type: cross Abstract: Decision Trees (DTs) are widely used in safety-critical domains such as medical diagnosis, valued for their interpretability and effectiveness on tabu

safetyarxiv-cs-ai
11 May 2026
Agents

Are LLM Agents Behaviorally Coherent? Latent Profiles for Social Simulation

DGX agent

arXiv:2509.03736v2 Announce Type: replace Abstract: The impressive capabilities of Large Language Models (LLMs) raise the possibility that synthetic agents can serve as substitutes for real participan

agentsarxiv-cs-ai
11 May 2026
Agents

ARMOR: An Agentic Framework for Reaction Feasibility Prediction via Adaptive Utility-aware Multi-tool Reasoning

DGX agent

arXiv:2605.07103v1 Announce Type: new Abstract: Reaction feasibility prediction, as a fundamental problem in computational chemistry, has benefited from diverse tools enabled by recent advances in art

agentsarxiv-cs-ai
11 May 2026
Safety

ART for Diffusion Sampling: A Reinforcement Learning Approach to Timestep Schedule

DGX agent

arXiv:2601.18681v2 Announce Type: replace-cross Abstract: We consider time discretization for score-based diffusion models to generate samples from a learned reverse-time dynamic on a finite grid. Uni

safetyarxiv-cs-ai
11 May 2026
Safety

ASPECT: Node-Level Adaptive Spectral Fusion for Graph Contrastive Learning

DGX agent

arXiv:2604.01878v2 Announce Type: replace-cross Abstract: Spectral graph contrastive learning often constructs low- and high-frequency views to capture complementary graph signals, but these views are

safetyarxiv-cs-ai
11 May 2026
Safety

Asymmetric On-Policy Distillation: Bridging Exploitation and Imitation at the Token Level

DGX agent

arXiv:2605.06387v2 Announce Type: replace-cross Abstract: On-policy distillation (OPD) trains a student on its own trajectories with token-level teacher feedback and often outperforms off-policy disti

safetyarxiv-cs-ai
11 May 2026
Agents

ATHENA: Agentic Team for Hierarchical Evolutionary Numerical Algorithms

DGX agent

arXiv:2512.03476v2 Announce Type: replace-cross Abstract: Bridging the gap between theoretical conceptualization and computational implementation is a major bottleneck in Scientific Computing (SciC) a

agentsarxiv-cs-ai
11 May 2026
Research

Automated Evaluation can Distinguish the Good and Bad AI Responses to Patient Questions about Hospitalization

DGX agent

arXiv:2510.00436v2 Announce Type: replace Abstract: Automated approaches to answer patient-posed health questions are rising, but selecting among systems requires reliable evaluation. The current gold

researcharxiv-cs-ai
11 May 2026
Research

Automatic Image-Level Morphological Trait Annotation for Organismal Images

DGX agent

arXiv:2604.01619v3 Announce Type: replace-cross Abstract: Morphological traits are physical characteristics of biological organisms that provide vital clues on how organisms interact with their enviro

researcharxiv-cs-ai
11 May 2026
Research

BalCapRL: A Balanced Framework for RL-Based MLLM Image Captioning

DGX agent

arXiv:2605.07394v1 Announce Type: cross Abstract: Image captioning is one of the most fundamental tasks in computer vision. Owing to its open-ended nature, it has received significant attention in the

researcharxiv-cs-ai
11 May 2026
Safety

BEAVER: An Efficient Deterministic LLM Verifier

DGX agent

arXiv:2512.05439v2 Announce Type: replace Abstract: As large language models (LLMs) transition from research prototypes to production systems, practitioners often need reliable methods to verify model

safetyarxiv-cs-ai
11 May 2026
Applications

BeeVe: Unsupervised Acoustic State Discovery in Honey Bee Buzzing

DGX agent

arXiv:2605.07903v1 Announce Type: cross Abstract: Discovering structure in biological signals without supervision is a fundamental problem in computational intelligence, yet existing bioacoustic metho

applicationsarxiv-cs-ai
11 May 2026
Model Releases

Behavior Cue Reasoning: Monitorable Reasoning Improves Efficiency and Safety through Oversight

DGX agent

arXiv:2605.07021v1 Announce Type: new Abstract: Reasoning in Large Language Models (LLMs) poses a challenge for oversight as many misaligned behaviors do not surface until reasoning concludes. To addr

model-releasesarxiv-cs-ai
11 May 2026
Agents

Belief Memory: Agent Memory Under Partial Observability

DGX agent

arXiv:2605.05583v2 Announce Type: replace Abstract: LLM agents that operate over long context depend on external memory to accumulate knowledge over time. However, existing methods typically store eac

agentsarxiv-cs-ai
11 May 2026
Model Releases

Benchmarking EngGPT2-16B-A3B against Comparable Italian and International Open-source LLMs

DGX agent

arXiv:2605.07731v1 Announce Type: cross Abstract: This report benchmarks the performance of ENGINEERING Ingegneria Informatica S.p.A.'s EngGPT2MoE-16B-A3B LLM, a 16B parameter Mixture of Experts (MoE)

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Benchmarking World-Model Learning with Environment-Level Queries

DGX agent

arXiv:2510.19788v4 Announce Type: replace Abstract: World models are central to building AI agents capable of flexible reasoning and planning. Yet current evaluations (i) test only properties measurab

model-releasesarxiv-cs-ai
11 May 2026
Safety

Beyond Confidence: Rethinking Self-Assessments for Performance Prediction in LLMs

DGX agent

arXiv:2605.07806v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly used in settings where reliable self-assessment is critical. Assessing model reliability has evolved fro

safetyarxiv-cs-ai
11 May 2026
Model Releases

Beyond Factor Aggregation: Gauge-Aware Low-Rank Server Representations for Federated LoRA

DGX agent

arXiv:2605.06733v1 Announce Type: cross Abstract: Federated LoRA enables parameter-efficient adaptation of large language models under decentralized data and limited client resources.However, directly

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Beyond LoRA vs. Full Fine-Tuning: Gradient-Guided Optimizer Routing for LLM Adaptation

DGX agent

arXiv:2605.07111v1 Announce Type: cross Abstract: Recent literature on fine-tuning Large Language Models highlights a fundamental debate. While Full Fine-Tuning (FFT) provides the representational pla

model-releasesarxiv-cs-ai
11 May 2026
Safety

Beyond Pairs: Your Language Model is Secretly Optimizing a Preference Graph

DGX agent

arXiv:2605.08037v1 Announce Type: cross Abstract: Direct Preference Optimization (DPO) aligns language models using pairwise preference comparisons, offering a simple and effective alternative to Rein

safetyarxiv-cs-ai
11 May 2026
Model Releases

Beyond Retrieval: A Multitask Benchmark and Model for Code Search

DGX agent

arXiv:2605.04615v2 Announce Type: replace-cross Abstract: Code search has usually been evaluated as first-stage retrieval, even though production systems rely on broader pipelines with reranking and d

model-releasesarxiv-cs-ai
11 May 2026
Safety

Beyond State-Wise Mirror Descent: Offline Policy Optimization with Parametric Policies

DGX agent

arXiv:2602.23811v4 Announce Type: replace-cross Abstract: We investigate the theoretical aspects of offline reinforcement learning (RL) under general function approximation. While prior works (e.g., X

safetyarxiv-cs-ai
11 May 2026
Model Releases

Beyond the Black Box: Interpretability of Agentic AI Tool Use

DGX agent

arXiv:2605.06890v1 Announce Type: new Abstract: AI agents are promising for high-stakes enterprise workflows, but dependable deployment remains limited because tool-use failures are difficult to diagn

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

BGM-IV: an AI-powered Bayesian generative modeling approach for instrumental variable analysis

DGX agent

arXiv:2605.07029v1 Announce Type: cross Abstract: Instrumental-variable (IV) regression enables causal estimation under endogeneity, but modern IV problems often involve nonlinear structural effects a

model-releasesarxiv-cs-ai
11 May 2026
← Previous
1…347348349350351…448
Next →