AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
Safety

Intervention Complexity as a Canonical Reward and a Measure of Intelligence

DGX agent

arXiv:2605.02175v1 Announce Type: new Abstract: The Legg--Hutter universal intelligence measure provides a rigorous scalar assessment of general intelligence as expected reward across all computable e

safetyarxiv-cs-ai
6 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

Investigating the Effects of Different Levels of User Control in an Interactive Educational Recommender System

DGX agent

arXiv:2605.01400v1 Announce Type: cross Abstract: Educational recommender systems (ERSs) are becoming increasingly important in enhancing educational outcomes and personalizing learning experiences by

researcharxiv-cs-ai
6 May 2026
Research

Iterative Finetuning is Mostly Idempotent

DGX agent

arXiv:2605.01130v1 Announce Type: new Abstract: If a model has some behavioral tendency, such as sycophancy or misalignment, and it is trained on its own outputs, will the tendency be amplified in the

researcharxiv-cs-ai
6 May 2026
Applications

KG-First, LLM-Fallback: A Hybrid Microservice for Grounded Skill Search and Explanation

DGX agent

arXiv:2605.01582v1 Announce Type: cross Abstract: Authoritative competency frameworks such as ESCO, ROME, and O*NET are essential for aligning education with labor market needs, yet their technical co

applicationsarxiv-cs-ai
6 May 2026
Safety

Khala: Scaling Acoustic Token Language Models Toward High-Fidelity Music Generation

DGX agent

arXiv:2605.01790v1 Announce Type: cross Abstract: A common design pattern in high-quality music generation is to handle structure and fidelity in different representation spaces: a generator first mod

safetyarxiv-cs-ai
6 May 2026
Agents

Latent State Design for World Models under Sufficiency Constraints

DGX agent

arXiv:2605.01694v1 Announce Type: new Abstract: A world model matters to an agent only through the state it constructs. That state must preserve some information, discard other information, and suppor

agentsarxiv-cs-ai
6 May 2026
Agents

Learning Decentralized LLM Collaboration with Multi-Agent Actor Critic

DGX agent

arXiv:2601.21972v4 Announce Type: replace Abstract: Recent work has explored optimizing LLM collaboration through Multi-Agent Reinforcement Learning (MARL). However, most MARL fine-tuning approaches r

agentsarxiv-cs-ai
6 May 2026
Agents

Less Interaction But More Explanation: A Communication Perspective on Agentic AI Interfaces

DGX agent

arXiv:2605.01610v1 Announce Type: cross Abstract: AI systems have long been expected to interact with users, answering questions, generating content, and continuing (social) conversations. Agentic AI,

agentsarxiv-cs-ai
6 May 2026
Agents

Lifting Traces to Logic: Programmatic Skill Induction with Neuro-Symbolic Learning for Long-Horizon Agentic Tasks

DGX agent

arXiv:2605.01293v1 Announce Type: new Abstract: Foundation model-driven agents often struggle with long-horizon planning due to the transient nature of purely prompting-based reasoning. While existing

agentsarxiv-cs-ai
6 May 2026
Model Releases

LiveFMBench: Unveiling the Power and Limits of Agentic Workflows in Specification Generation

DGX agent

arXiv:2605.01394v1 Announce Type: cross Abstract: Formal specification is essential for rigorous program verification, yet writing correct specifications remains costly and difficult to automate. Alth

model-releasesarxiv-cs-ai
6 May 2026
Tutorials

LLM-Assisted Repository-Level Generation with Structured Spec-Driven Engineering

DGX agent

arXiv:2605.02455v1 Announce Type: cross Abstract: State-of-the-art Large Language Models (LLMs) excel in code generation at the function level. However, the output quality significantly declines when

tutorialsarxiv-cs-ai
6 May 2026
Agents

LLM-enabled Social Agents

DGX agent

arXiv:2605.02335v1 Announce Type: cross Abstract: Large Language Models (LLMs) have transformed agent-agent and human-agent interaction by enabling software, physical, and simulation agents to communi

agentsarxiv-cs-ai
6 May 2026
Agents

LLM-Powered AI Agent Systems and Their Applications in Industry

DGX agent

arXiv:2505.16120v2 Announce Type: replace Abstract: The emergence of Large Language Models (LLMs) has reshaped agent systems. Unlike traditional rule-based agents with limited task scope, LLM-powered

agentsarxiv-cs-ai
6 May 2026
Research

LLMs Should Not Yet Be Credited with Decision Explanation

DGX agent

arXiv:2605.01164v1 Announce Type: new Abstract: This position paper argues that LLMs should not yet be credited with decision explanation. This matters because recent work increasingly treats accurate

researcharxiv-cs-ai
6 May 2026
Safety

Logic-Constrained Shortest Paths for Flight Planning

DGX agent

arXiv:2412.13235v4 Announce Type: replace Abstract: The logic-constrained shortest path problem (LCSPP) combines a one-to-one shortest path problem with satisfiability constraints imposed on the routi

safetyarxiv-cs-ai
6 May 2026
Model Releases

MAP-Law: Coverage-Driven Retrieval Control for Multi-Turn Legal Consultation

DGX agent

arXiv:2605.01486v1 Announce Type: new Abstract: Legal consultation is a high-stakes, knowledge-intensive task that requires agents to identify relevant legal issues, retrieve authoritative support, an

model-releasesarxiv-cs-ai
6 May 2026
Model Releases

MCP-Atlas: A Large-Scale Benchmark for Tool-Use Competency with Real MCP Servers

DGX agent

arXiv:2602.00933v2 Announce Type: replace-cross Abstract: The Model Context Protocol (MCP) is rapidly becoming the standard interface for Large Language Models (LLMs) to discover and invoke external t

model-releasesarxiv-cs-ai
6 May 2026
Local Ai

MedGemma 1.5 Technical Report

DGX agent

arXiv:2604.05081v2 Announce Type: replace Abstract: We introduce MedGemma 1.5 4B, the latest model in the MedGemma collection. MedGemma 1.5 expands on MedGemma 1 by integrating additional capabilities

local-aiarxiv-cs-ai
6 May 2026
Research

MEMAUDIT: An Exact Package-Oracle Evaluation Protocol for Budgeted Long-Term LLM Memory Writing

DGX agent

arXiv:2605.02199v1 Announce Type: new Abstract: Long-term LLM agents must compress streams of past interactions into persistent memory before future queries are known. Existing evaluations usually mea

researcharxiv-cs-ai
6 May 2026
Safety

MILD: Mediator Agent System with Bidirectional Perception and Multi-Layered Alignment for Human-Vehicle Collaboration

DGX agent

arXiv:2605.01507v1 Announce Type: new Abstract: Prior studies report that partial driving automation can increase the cognitive demands on human drivers. This effect largely arises from human drivers'

safetyarxiv-cs-ai
6 May 2026
Local Ai

MindMelody: A Closed-Loop EEG-Driven System for Personalized Music Intervention

DGX agent

arXiv:2605.01235v1 Announce Type: cross Abstract: Driven by the escalating global burden of mental health conditions, music-based interventions have attracted significant attention as a non-invasive,

local-aiarxiv-cs-ai
6 May 2026
Safety

MINT: Minimal Information Neuro-Symbolic Tree for Objective-Driven Knowledge-Gap Reasoning and Active Elicitation

DGX agent

arXiv:2602.05048v2 Announce Type: replace Abstract: Joint planning through language-based interactions is a key area of human-AI teaming. Planning problems in the open world often involve various aspe

safetyarxiv-cs-ai
6 May 2026
Safety

Model Routing as a Trust Problem: Route Receipts for Adaptive AI Systems

DGX agent

arXiv:2605.01710v1 Announce Type: new Abstract: AI products often route requests through version aliases, service tiers, tool choices, regional endpoints, fallback rules, or safety handling before res

safetyarxiv-cs-ai
6 May 2026
Safety

Model Spec Midtraining: Improving How Alignment Training Generalizes

DGX agent

arXiv:2605.02087v1 Announce Type: new Abstract: Some frontier AI developers aim to align language models to a Model Spec or Constitution that describes the intended model behavior. However, standard a

safetyarxiv-cs-ai
6 May 2026
Model Releases

MSEarth: A Multimodal Benchmark for Earth Science Phenomenon Discovery with MLLMs

DGX agent

arXiv:2505.20740v3 Announce Type: replace Abstract: The rapid advancement of multimodal large language models (MLLMs) offers new opportunities for complex scientific challenges, yet their application

model-releasesarxiv-cs-ai
6 May 2026
Model Releases

Multi-Agent Reasoning Improves Compute Efficiency: Pareto-Optimal Test-Time Scaling

DGX agent

arXiv:2605.01566v1 Announce Type: new Abstract: Advances in inference methods have enabled language models to improve their predictions without additional training. These methods often prioritize raw

model-releasesarxiv-cs-ai
6 May 2026
Tutorials

Multi-modal Relational Item Representation Learning for Inferring Substitutable and Complementary Items

DGX agent

arXiv:2507.22268v3 Announce Type: replace-cross Abstract: We study the problem of inferring substitutable and complementary items, which underpins applications such as alternative and follow-up purcha

tutorialsarxiv-cs-ai
6 May 2026
Research

Music Interpretation and Emotion Perception: A Computational and Neurophysiological Investigation

DGX agent

arXiv:2506.01982v3 Announce Type: replace-cross Abstract: This study investigates emotional expression and perception in music performance using computational and neurophysiological methods. The influ

researcharxiv-cs-ai
6 May 2026
Agents

NaviGNN: Multi-Agent Reinforcement Learning and Graph Neural Network for Sustainable Mobility in Futuristic Smart Cities

DGX agent

arXiv:2507.15143v3 Announce Type: replace Abstract: This paper investigates the feasibility of human mobility in extreme urban morphologies characterized by high-density vertical structures and linear

agentsarxiv-cs-ai
6 May 2026
Tutorials

Neural Decision-Propagation for Answer Set Programming

DGX agent

arXiv:2605.01797v1 Announce Type: new Abstract: Integration of Answer Set Programming (ASP) with neural networks has emerged as a promising tool in Neuro-symbolic AI. While existing approaches extend

tutorialsarxiv-cs-ai
6 May 2026
Model Releases

Neuro-Symbolic Agents for Hallucination-Free Requirements Reuse

DGX agent

arXiv:2605.01562v1 Announce Type: cross Abstract: The Object-Oriented Method for Requirements Authoring and Management (OOMRAM) is a requirements reuse framework that relies on exact identifier matchi

model-releasesarxiv-cs-ai
6 May 2026
Research

NEURON: A Neuro-symbolic System for Grounded Clinical Explainability

DGX agent

arXiv:2605.01189v1 Announce Type: new Abstract: Clinical AI adoption is hindered by the black-box/grey-box nature of high-performing models, which lack the ontological grounding and narrative transpar

researcharxiv-cs-ai
6 May 2026
Model Releases

NeuroState-Bench: A Human-Calibrated Benchmark for Commitment Integrity in LLM Agent Profiles

DGX agent

arXiv:2605.01847v1 Announce Type: new Abstract: Outcome-only evaluation under-specifies whether an evaluated agent profile preserves the commitments required to solve a multi-turn task coherently. Neu

model-releasesarxiv-cs-ai
6 May 2026
Model Releases

New Bounds for Zarankiewicz Numbers via Reinforced LLM Evolutionary Search

DGX agent

arXiv:2605.01120v1 Announce Type: new Abstract: The Zarankiewicz number extbf{Z}(m, n, s, t) is the maximum number of edges in a bipartite graph G_{m, n} such that there is no complete K_{s, t} bipart

model-releasesarxiv-cs-ai
6 May 2026
Safety

NORA: A Harness-Engineered Autonomous Research Agent for End-to-End Spatial Data Science

DGX agent

arXiv:2605.02092v1 Announce Type: new Abstract: The automation of scientific research workflows has emerged as a transformative frontier in artificial intelligence, yet existing autonomous research ag

safetyarxiv-cs-ai
6 May 2026
Research

On the Privacy of LLMs: An Ablation Study

DGX agent

arXiv:2605.02255v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly deployed in interactive and retrieval-augmented settings, raising significant privacy concerns. While at

researcharxiv-cs-ai
6 May 2026
Research

Optimization of CV-QKD Under Practical Constraints

DGX agent

arXiv:2605.02045v1 Announce Type: cross Abstract: Using reinforcement learning, we optimize for practical hardware constraints, including limited FIR filter taps at the transmitter and receiver, mean

researcharxiv-cs-ai
6 May 2026
Model Releases

ORPilot: A Production-Oriented Agentic LLM-for-OR Tool for Optimization Modeling

DGX agent

arXiv:2605.02728v1 Announce Type: new Abstract: This paper presents ORPilot, an open-source agentic AI system that translates real-world business problems into solver-ready optimization models. Unlike

model-releasesarxiv-cs-ai
6 May 2026
Research

Partial-differential-algebraic equations of nonlinear dynamics by Physics-Informed Neural-Network: (I) Operator splitting and framework assessment

DGX agent

arXiv:2408.01914v4 Announce Type: replace-cross Abstract: Several forms for constructing novel physics-informed neural-networks (PINN) for the solution of partial-differential-algebraic equations base

researcharxiv-cs-ai
6 May 2026
Model Releases

PERSA: Reinforcement Learning for Professor-Style Personalized Feedback with LLMs

DGX agent

arXiv:2605.01123v1 Announce Type: new Abstract: Large language models (LLMs) can provide automated feedback in educational settings, but aligning an LLMs style with a specific instructors tone while m

model-releasesarxiv-cs-ai
6 May 2026
Applications

Personalized Digital Health Modeling with Adaptive Support Users

DGX agent

arXiv:2605.02004v1 Announce Type: new Abstract: Personalized models are essential in digital health because individuals exhibit substantial physiological and behavioral heterogeneity. Yet personalizat

applicationsarxiv-cs-ai
6 May 2026
Model Releases

PhysicianBench: Evaluating LLM Agents in Real-World EHR Environments

DGX agent

arXiv:2605.02240v1 Announce Type: new Abstract: We introduce PhysicianBench, a benchmark for evaluating LLM agents on physician tasks grounded in real clinical setting within electronic health record

model-releasesarxiv-cs-ai
6 May 2026
Safety

Poly-EPO: Training Exploratory Reasoning Models

DGX agent

arXiv:2604.17654v3 Announce Type: replace Abstract: Exploration is a cornerstone of learning from experience: it enables agents to find solutions to complex problems, generalize to novel ones, and sca

safetyarxiv-cs-ai
6 May 2026
Research

Position: How can Graphs Help Large Language Models?

DGX agent

arXiv:2605.02452v1 Announce Type: new Abstract: With the rapid advancement of large language models (LLMs), classic graph learning tasks have greatly benefited from LLMs, including improved encoding o

researcharxiv-cs-ai
6 May 2026
Research

Position: LLM Serving Needs Mathematical Optimization and Algorithmic Foundations, Not Just Heuristics

DGX agent

arXiv:2605.01280v1 Announce Type: cross Abstract: This position paper argues that LLM inference serving has outgrown generic heuristics and now demands mathematical optimization and algorithmic founda

researcharxiv-cs-ai
6 May 2026
Safety

Position: Safety and Fairness in Agentic AI Depend on Interaction Topology, Not on Model Scale or Alignment

DGX agent

arXiv:2605.01147v1 Announce Type: new Abstract: As large language models are increasingly deployed as interacting agents in high-stakes decisions, the AI safety community assumes that safety propertie

safetyarxiv-cs-ai
6 May 2026
Research

(POSTER) From Sensors to Insight: Rapid, Edge-to-Core Application Development for Sensor-Driven Applications

DGX agent

arXiv:2605.02844v1 Announce Type: cross Abstract: Scientists increasingly rely on sensor-based data; however transforming raw streams into insights across the edge-to-cloud continuum remains difficult

researcharxiv-cs-ai
6 May 2026
Agents

Practical Limits of Autonomous Test Repair: A Multi-Agent Case Study with LLM-Driven Discovery and Self-Correction

DGX agent

arXiv:2605.01471v1 Announce Type: cross Abstract: Maintaining reliable UI test suites in large-scale enterprise applications is a persistent and costly challenge. We present an industrial case study o

agentsarxiv-cs-ai
6 May 2026
← Previous
1…360361362363364…448
Next →