AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,236 results
24 Apr 2026

Multi-Agent Empowerment and Emergence of Complex Behavior in Groups

AgentsDGX agent

arXiv:2604.21155v1 Announce Type: new Abstract: Intrinsic motivations are receiving increasing attention, i.e. behavioral incentives that are not engineered, but emerge from the interaction of an agen

Multimodal Bayesian Network for Robust Assessment of Casualties in Autonomous Triage

AgentsDGX agent

arXiv:2512.18908v2 Announce Type: replace Abstract: Mass Casualty Incidents can overwhelm emergency medical systems and resulting delays or errors in the assessment of casualties can lead to preventab

Musical Score Understanding Benchmark: Evaluating Large Language Models' Comprehension of Complete Musical Scores

Model ReleasesDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2511.20697v4 Announce Type: replace-cross Abstract: Understanding complete musical scores entails integrated reasoning over pitch, rhythm, harmony, and large-scale structure, yet the ability of

Navigating the Clutter: Waypoint-Based Bi-Level Planning for Multi-Robot Systems

Model ReleasesDGX agent

arXiv:2604.21138v1 Announce Type: cross Abstract: Multi-robot control in cluttered environments is a challenging problem that involves complex physical constraints, including robot-robot collisions, r

Nemobot Games: Crafting Strategic AI Gaming Agents for Interactive Learning with Large Language Models

Model ReleasesDGX agent

arXiv:2604.21896v1 Announce Type: new Abstract: This paper introduces a new paradigm for AI game programming, leveraging large language models (LLMs) to extend and operationalize Claude Shannon's taxo

NPU Design for Diffusion Language Model Inference

ResearchDGX agent

arXiv:2601.20706v2 Announce Type: replace-cross Abstract: Diffusion-based LLMs (dLLMs) fundamentally depart from traditional autoregressive (AR) LLM inference: they leverage bidirectional attention, b

Omission Constraints Decay While Commission Constraints Persist in Long-Context LLM Agents

Model ReleasesDGX agent

arXiv:2604.20911v1 Announce Type: cross Abstract: LLM agents deployed in production operate under operator-defined behavioral policies (system-prompt instructions such as prohibitions on credential di

On Reasoning Behind Next Occupation Recommendation

ApplicationsDGX agent

arXiv:2604.21204v1 Announce Type: cross Abstract: In this work, we develop a novel reasoning approach to enhance the performance of large language models (LLMs) in future occupation prediction. In thi

On the Relationship between Bayesian Networks and Probabilistic Structural Causal Models

ResearchDGX agent

arXiv:2603.27406v2 Announce Type: replace Abstract: In this paper, the relationship between probabilistic graphical models, in particular Bayesian networks, and causal diagrams, also called structural

On the Role of Preprocessing and Memristor Dynamics in Reservoir Computing for Image Classification

Model ReleasesDGX agent

arXiv:2604.21602v1 Announce Type: cross Abstract: Reservoir computing (RC) is an emerging recurrent neural network architecture that has attracted growing attention for its low training cost and modes

Open-H-Embodiment: A Large-Scale Dataset for Enabling Foundation Models in Medical Robotics

Model ReleasesDGX agent

arXiv:2604.21017v1 Announce Type: cross Abstract: Autonomous medical robots hold promise to improve patient outcomes, reduce provider workload, democratize access to care, and enable superhuman precis

OpenEstimate: Evaluating LLMs on Reasoning Under Uncertainty with Real-World Data

Model ReleasesDGX agent

arXiv:2510.15096v2 Announce Type: replace Abstract: Real-world settings where language models (LMs) are deployed -- in domains spanning healthcare, finance, and other forms of knowledge work -- requir

OpInf-LLM: Parametric PDE Solving with LLMs via Operator Inference

AgentsDGX agent

arXiv:2602.01493v2 Announce Type: replace-cross Abstract: Solving diverse partial differential equations (PDEs) is fundamental in science and engineering. Large language models (LLMs) have demonstrate

Planetary Exploration 3.0: A Roadmap for Software-Defined, Radically Adaptive Space Systems

AgentsDGX agent

arXiv:2604.20910v1 Announce Type: cross Abstract: The surface and subsurface of worlds beyond Mars remain largely unexplored. Yet these worlds hold keys to fundamental questions in planetary science -

Planning Beyond Text: Graph-based Reasoning for Complex Narrative Generation

ResearchDGX agent

arXiv:2604.21253v1 Announce Type: cross Abstract: While LLMs demonstrate remarkable fluency in narrative generation, existing methods struggle to maintain global narrative coherence, contextual logica

Post-AGI Economies: Autonomy and the First Fundamental Theorem of Welfare Economics

AgentsDGX agent

arXiv:2604.21216v1 Announce Type: cross Abstract: The First Fundamental Theorem of Welfare Economics assumes that welfare-bearing agents are autonomous and implicitly relies on a binary distinction be

PosterForest: Hierarchical Multi-Agent Collaboration for Scientific Poster Generation

AgentsDGX agent

arXiv:2508.21720v2 Announce Type: replace Abstract: Automating scientific poster generation requires hierarchical document understanding and coherent content-layout planning. Existing methods often re

Pre-trained LLMs Meet Sequential Recommenders: Efficient User-Centric Knowledge Distillation

ResearchDGX agent

arXiv:2604.21536v1 Announce Type: cross Abstract: Sequential recommender systems have achieved significant success in modeling temporal user behavior but remain limited in capturing rich user semantic

Predicting Scale-Up of Metal-Organic Framework Syntheses with Large Language Models

ResearchDGX agent

arXiv:2604.20899v1 Announce Type: cross Abstract: Scalable synthesis remains the gate between MOF discovery and industrial deployment, as scale-up know-how is fragmented across disparate reports. We i

Preserving Decision Sovereignty in Military AI: A Trade-Secret-Safe Architectural Framework for Model Replaceability, Human Authority, and State Control

SafetyDGX agent

arXiv:2604.20867v1 Announce Type: cross Abstract: Recent events surrounding the relationship between frontier AI suppliers and national-security customers have made a structural problem newly visible:

Preserving Knowledge in Large Language Model with Model-Agnostic Self-Decompression

ResearchDGX agent

arXiv:2406.11354v3 Announce Type: replace-cross Abstract: Humans can retain old knowledge while learning new information, but Large Language Models (LLMs) often suffer from catastrophic forgetting whe

Probabilistic Verification of Neural Networks via Efficient Probabilistic Hull Generation

SafetyDGX agent

arXiv:2604.21556v1 Announce Type: new Abstract: The problem of probabilistic verification of a neural network investigates the probability of satisfying the safe constraints in the output space when t

Probably Approximately Consensus: On the Learning Theory of Finding Common Ground

ResearchDGX agent

arXiv:2604.21811v1 Announce Type: cross Abstract: A primary goal of online deliberation platforms is to identify ideas that are broadly agreeable to a community of users through their expressed prefer

Process Supervision via Verbal Critique Improves Reasoning in Large Language Models

Model ReleasesDGX agent

arXiv:2604.21611v1 Announce Type: cross Abstract: Inference-time scaling for LLM reasoning has focused on three axes: chain depth, sample breadth, and learned step-scorers (PRMs). We introduce a fourt

Promoting Simple Agents: Ensemble Methods for Event-Log Prediction

ApplicationsDGX agent

arXiv:2604.21629v1 Announce Type: cross Abstract: We compare lightweight automata-based models (n-grams) with neural architectures (LSTM, Transformer) for next-activity prediction in streaming event l

Propensity Inference: Environmental Contributors to LLM Behaviour

ResearchDGX agent

arXiv:2604.21098v1 Announce Type: new Abstract: Motivated by loss of control risks from misaligned AI systems, we develop and apply methods for measuring language models' propensity for unsanctioned b

Quotient-Space Diffusion Models

SafetyDGX agent

arXiv:2604.21809v1 Announce Type: cross Abstract: Diffusion-based generative models have reformed generative AI, and have enabled new capabilities in the science domain, for example, generating 3D str

ReaGeo: Reasoning-Enhanced End-to-End Geocoding with LLMs

ResearchDGX agent

arXiv:2604.21357v1 Announce Type: new Abstract: This paper proposes ReaGeo, an end-to-end geocoding framework based on large language models, designed to overcome the limitations of traditional multi-

RealRoute: Dynamic Query Routing System via Retrieve-then-Verify Paradigm

Model ReleasesDGX agent

arXiv:2604.20860v1 Announce Type: cross Abstract: Despite the success of Retrieval-Augmented Generation (RAG) in grounding LLMs with external knowledge, its application over heterogeneous sources (e.g

Reasoning Primitives in Hybrid and Non-Hybrid LLMs

ApplicationsDGX agent

arXiv:2604.21454v1 Announce Type: cross Abstract: Reasoning in large language models is often treated as a monolithic capability, but its observed gains may arise from more basic operations. We study

ReCAPA: Hierarchical Predictive Correction to Mitigate Cascading Failures

Local AiDGX agent

arXiv:2604.21232v1 Announce Type: new Abstract: Vision-Language-Action systems follow instructions to execute multi-step tasks in multimodal environments. Recent VLA approaches typically rely on post-

Rectified Schrodinger Bridge Matching for Few-Step Visual Navigation

Model ReleasesDGX agent

arXiv:2604.05673v2 Announce Type: replace-cross Abstract: Visual navigation is a core challenge in Embodied AI, requiring autonomous agents to translate high-dimensional sensory observations into cont

Reinforcement Learning with Foundation Priors: Let the Embodied Agent Efficiently Learn on Its Own

SafetyDGX agent

arXiv:2310.02635v5 Announce Type: replace-cross Abstract: Reinforcement learning (RL) is a promising approach for solving robotic manipulation tasks. However, it is challenging to apply the RL algorit

Reinforcing privacy reasoning in LLMs via normative simulacra from fiction

Model ReleasesDGX agent

arXiv:2604.20904v1 Announce Type: cross Abstract: Information handling practices of LLM agents are broadly misaligned with the contextual privacy expectations of their users. Contextual Integrity (CI)

RELOOP: Recursive Retrieval with Multi-Hop Reasoner and Planners for Heterogeneous QA

SafetyDGX agent

arXiv:2510.20505v4 Announce Type: replace-cross Abstract: Retrieval-augmented generation (RAG) remains brittle on multi-step questions and heterogeneous evidence sources, trading accuracy against late

Replay-buffer engineering for noise-robust quantum circuit optimization

ResearchDGX agent

arXiv:2604.21863v1 Announce Type: cross Abstract: Deep reinforcement learning (RL) for quantum circuit optimization faces three fundamental bottlenecks: replay buffers that ignore the reliability of t

Retrofit: Continual Learning with Controlled Forgetting for Binary Security Detection and Analysis

Model ReleasesDGX agent

arXiv:2511.11439v2 Announce Type: replace-cross Abstract: Binary security has increasingly relied on deep learning to reason about malware behavior and program semantics. However, the performance ofte

Reversible Deep Learning for 13C NMR in Chemoinformatics: On Structures and Spectra

ResearchDGX agent

arXiv:2602.03875v4 Announce Type: replace-cross Abstract: We introduce a reversible deep learning model for 13C NMR that uses a single conditional invertible neural network for both directions between

Revisiting Content-Based Music Recommendation: Efficient Feature Aggregation from Large-Scale Music Models

ResearchDGX agent

arXiv:2604.20847v1 Announce Type: cross Abstract: Music Recommendation Systems (MRSs) are a cornerstone of modern streaming platforms. Existing recommendation models, spanning both recall and ranking

RIFT: Repurposing Negative Samples via Reward-Informed Fine-Tuning

SafetyDGX agent

arXiv:2601.09253v2 Announce Type: replace-cross Abstract: While Supervised Fine-Tuning (SFT) and Rejection Sampling Fine-Tuning (RFT) are standard for LLM alignment, they either rely on costly expert

Robust Test-time Video-Text Retrieval: Benchmarking and Adapting for Query Shifts

Model ReleasesDGX agent

arXiv:2604.20851v1 Announce Type: cross Abstract: Modern video-text retrieval (VTR) models excel on in-distribution benchmarks but are highly vulnerable to real-world query shifts, where the distribut

Robustness Analysis of POMDP Policies to Observation Perturbations

SafetyDGX agent

arXiv:2604.21256v1 Announce Type: new Abstract: Policies for Partially Observable Markov Decision Processes (POMDPs) are often designed using a nominal system model. In practice, this model can deviat

RV-HATE: Reinforced Multi-Module Voting for Implicit Hate Speech Detection

ResearchDGX agent

arXiv:2510.10971v2 Announce Type: replace-cross Abstract: Hate speech remains prevalent in human society and continues to evolve in its forms and expressions. Modern advancements in internet and onlin

SafeMERGE: Preserving Safety Alignment in Fine-Tuned Large Language Models via Selective Layer-Wise Model Merging

SafetyDGX agent

arXiv:2503.17239v3 Announce Type: replace-cross Abstract: Fine-tuning large language models (LLMs) is a common practice to adapt generalist models to specialized domains. However, recent studies show

SafeRedirect: Defeating Internal Safety Collapse via Task-Completion Redirection in Frontier LLMs

SafetyDGX agent

arXiv:2604.20930v1 Announce Type: cross Abstract: Internal Safety Collapse (ISC) is a failure mode in which frontier LLMs, when executing legitimate professional tasks whose correct completion structu

Satisfying Rationality Postulates of Structured Argumentation Through Deductive Support -- Technical Report

ResearchDGX agent

arXiv:2604.21515v1 Announce Type: new Abstract: ASPIC-style structured argumentation frameworks provide a formal basis for reasoning in artificial intelligence by combining internal argument structure

Scaling of Gaussian Kolmogorov--Arnold Networks

Model ReleasesDGX agent

arXiv:2604.21174v1 Announce Type: cross Abstract: The Gaussian scale parameter (epsilon) is central to the behavior of Gaussian Kolmogorov--Arnold Networks (KANs), yet its role in deep edge-based arch

Schoenfeld's Anatomy of Mathematical Reasoning by Language Models

ResearchDGX agent

arXiv:2512.19995v2 Announce Type: replace-cross Abstract: Large language models increasingly expose reasoning traces, yet their underlying cognitive structure and steps remain difficult to identify an

Secure LLM Fine-Tuning via Safety-Aware Probing

Model ReleasesDGX agent

arXiv:2505.16737v2 Announce Type: replace-cross Abstract: Large language models (LLMs) have achieved remarkable success across many applications, but their ability to generate harmful content raises s

Seeing Fast and Slow: Learning the Flow of Time in Videos

TutorialsDGX agent

arXiv:2604.21931v1 Announce Type: cross Abstract: How can we tell whether a video has been sped up or slowed down? How can we generate videos at different speeds? Although videos have been central to

SemanticAgent: A Semantics-Aware Framework for Text-to-SQL Data Synthesis

ResearchDGX agent

arXiv:2604.21414v1 Announce Type: new Abstract: Existing text-to-SQL synthesis pipelines still conflate executability with semantic validity: syntactic checks and execution-based validation can retain

SemaPop: Semantic-Persona Conditioned and Controllable Population Synthesis

SafetyDGX agent

arXiv:2602.11569v2 Announce Type: replace Abstract: Population synthesis is essential for individual-level simulation in transport planning and socio-economic analysis, yet remains challenging due to

Separable Expert Architecture: Toward Privacy-Preserving LLM Personalization via Composable Adapters and Deletable User Proxies

Model ReleasesDGX agent

arXiv:2604.21571v1 Announce Type: new Abstract: Current model training approaches incorporate user information directly into shared weights, making individual data removal computationally infeasible w

Serialisation Strategy Matters: How FHIR Data Format Affects LLM Medication Reconciliation

Model ReleasesDGX agent

arXiv:2604.21076v1 Announce Type: cross Abstract: Medication reconciliation at clinical handoffs is a high-stakes, error-prone process. Large language models are increasingly proposed to assist with t

SGD at the Edge of Stability: The Stochastic Sharpness Gap

ResearchDGX agent

arXiv:2604.21016v1 Announce Type: cross Abstract: When training neural networks with full-batch gradient descent (GD) and step size eta, the largest eigenvalue of the Hessian -- the sharpness S(oldsym

SparKV: Overhead-Aware KV Cache Loading for Efficient On-Device LLM Inference

Local AiDGX agent

arXiv:2604.21231v1 Announce Type: cross Abstract: Efficient inference for on-device Large Language Models (LLMs) remains challenging due to limited hardware resources and the high cost of the prefill

Spatial Metaphors for LLM Memory: A Critical Analysis of the MemPalace Architecture

Model ReleasesDGX agent

arXiv:2604.21284v1 Announce Type: new Abstract: MemPalace is an open-source AI memory system that applies the ancient method of loci (memory palace) spatial metaphor to organize long-term memory for l

Speculative Actions: A Lossless Framework for Faster Agentic Systems

AgentsDGX agent

arXiv:2510.04371v2 Announce Type: replace Abstract: AI agents are increasingly deployed in complex, interactive environments, yet their runtime remains a major bottleneck for training, evaluation, and

SPIRE: Structure-Preserving Interpretable Retrieval of Evidence

ResearchDGX agent

arXiv:2604.20849v1 Announce Type: cross Abstract: Retrieval-augmented generation over semi-structured sources such as HTML is constrained by a mismatch between document structure and the flat, sequenc

SQLyzr: A Comprehensive Benchmark and Evaluation Platform for Text-to-SQL

Model ReleasesDGX agent

arXiv:2604.21214v1 Announce Type: cross Abstract: Text-to-SQL models have significantly improved with the adoption of Large Language Models (LLMs), leading to their increasing use in real-world applic

← Previous
1…309310311312313…354
Next →