AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent
84,433Total entries
1Added by human
84,432Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
6 May 2026

LLM-Powered AI Agent Systems and Their Applications in Industry

AgentsDGX agent

arXiv:2505.16120v2 Announce Type: replace Abstract: The emergence of Large Language Models (LLMs) has reshaped agent systems. Unlike traditional rule-based agents with limited task scope, LLM-powered

LLMs Should Not Yet Be Credited with Decision Explanation

ResearchDGX agent

arXiv:2605.01164v1 Announce Type: new Abstract: This position paper argues that LLMs should not yet be credited with decision explanation. This matters because recent work increasingly treats accurate

Logic-Constrained Shortest Paths for Flight Planning

SafetyDGX agent

arXiv:2412.13235v4 Announce Type: replace Abstract: The logic-constrained shortest path problem (LCSPP) combines a one-to-one shortest path problem with satisfiability constraints imposed on the routi


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

MAP-Law: Coverage-Driven Retrieval Control for Multi-Turn Legal Consultation

Model ReleasesDGX agent

arXiv:2605.01486v1 Announce Type: new Abstract: Legal consultation is a high-stakes, knowledge-intensive task that requires agents to identify relevant legal issues, retrieve authoritative support, an

MCP-Atlas: A Large-Scale Benchmark for Tool-Use Competency with Real MCP Servers

Model ReleasesDGX agent

arXiv:2602.00933v2 Announce Type: replace-cross Abstract: The Model Context Protocol (MCP) is rapidly becoming the standard interface for Large Language Models (LLMs) to discover and invoke external t

MedGemma 1.5 Technical Report

Local AiDGX agent

arXiv:2604.05081v2 Announce Type: replace Abstract: We introduce MedGemma 1.5 4B, the latest model in the MedGemma collection. MedGemma 1.5 expands on MedGemma 1 by integrating additional capabilities

MEMAUDIT: An Exact Package-Oracle Evaluation Protocol for Budgeted Long-Term LLM Memory Writing

ResearchDGX agent

arXiv:2605.02199v1 Announce Type: new Abstract: Long-term LLM agents must compress streams of past interactions into persistent memory before future queries are known. Existing evaluations usually mea

MILD: Mediator Agent System with Bidirectional Perception and Multi-Layered Alignment for Human-Vehicle Collaboration

SafetyDGX agent

arXiv:2605.01507v1 Announce Type: new Abstract: Prior studies report that partial driving automation can increase the cognitive demands on human drivers. This effect largely arises from human drivers'

MindMelody: A Closed-Loop EEG-Driven System for Personalized Music Intervention

Local AiDGX agent

arXiv:2605.01235v1 Announce Type: cross Abstract: Driven by the escalating global burden of mental health conditions, music-based interventions have attracted significant attention as a non-invasive,

MINT: Minimal Information Neuro-Symbolic Tree for Objective-Driven Knowledge-Gap Reasoning and Active Elicitation

SafetyDGX agent

arXiv:2602.05048v2 Announce Type: replace Abstract: Joint planning through language-based interactions is a key area of human-AI teaming. Planning problems in the open world often involve various aspe

Model Routing as a Trust Problem: Route Receipts for Adaptive AI Systems

SafetyDGX agent

arXiv:2605.01710v1 Announce Type: new Abstract: AI products often route requests through version aliases, service tiers, tool choices, regional endpoints, fallback rules, or safety handling before res

Model Spec Midtraining: Improving How Alignment Training Generalizes

SafetyDGX agent

arXiv:2605.02087v1 Announce Type: new Abstract: Some frontier AI developers aim to align language models to a Model Spec or Constitution that describes the intended model behavior. However, standard a

MSEarth: A Multimodal Benchmark for Earth Science Phenomenon Discovery with MLLMs

Model ReleasesDGX agent

arXiv:2505.20740v3 Announce Type: replace Abstract: The rapid advancement of multimodal large language models (MLLMs) offers new opportunities for complex scientific challenges, yet their application

Multi-Agent Reasoning Improves Compute Efficiency: Pareto-Optimal Test-Time Scaling

Model ReleasesDGX agent

arXiv:2605.01566v1 Announce Type: new Abstract: Advances in inference methods have enabled language models to improve their predictions without additional training. These methods often prioritize raw

Multi-modal Relational Item Representation Learning for Inferring Substitutable and Complementary Items

TutorialsDGX agent

arXiv:2507.22268v3 Announce Type: replace-cross Abstract: We study the problem of inferring substitutable and complementary items, which underpins applications such as alternative and follow-up purcha

Music Interpretation and Emotion Perception: A Computational and Neurophysiological Investigation

ResearchDGX agent

arXiv:2506.01982v3 Announce Type: replace-cross Abstract: This study investigates emotional expression and perception in music performance using computational and neurophysiological methods. The influ

NaviGNN: Multi-Agent Reinforcement Learning and Graph Neural Network for Sustainable Mobility in Futuristic Smart Cities

AgentsDGX agent

arXiv:2507.15143v3 Announce Type: replace Abstract: This paper investigates the feasibility of human mobility in extreme urban morphologies characterized by high-density vertical structures and linear

Neural Decision-Propagation for Answer Set Programming

TutorialsDGX agent

arXiv:2605.01797v1 Announce Type: new Abstract: Integration of Answer Set Programming (ASP) with neural networks has emerged as a promising tool in Neuro-symbolic AI. While existing approaches extend

Neuro-Symbolic Agents for Hallucination-Free Requirements Reuse

Model ReleasesDGX agent

arXiv:2605.01562v1 Announce Type: cross Abstract: The Object-Oriented Method for Requirements Authoring and Management (OOMRAM) is a requirements reuse framework that relies on exact identifier matchi

NEURON: A Neuro-symbolic System for Grounded Clinical Explainability

ResearchDGX agent

arXiv:2605.01189v1 Announce Type: new Abstract: Clinical AI adoption is hindered by the black-box/grey-box nature of high-performing models, which lack the ontological grounding and narrative transpar

NeuroState-Bench: A Human-Calibrated Benchmark for Commitment Integrity in LLM Agent Profiles

Model ReleasesDGX agent

arXiv:2605.01847v1 Announce Type: new Abstract: Outcome-only evaluation under-specifies whether an evaluated agent profile preserves the commitments required to solve a multi-turn task coherently. Neu

New Bounds for Zarankiewicz Numbers via Reinforced LLM Evolutionary Search

Model ReleasesDGX agent

arXiv:2605.01120v1 Announce Type: new Abstract: The Zarankiewicz number extbf{Z}(m, n, s, t) is the maximum number of edges in a bipartite graph G_{m, n} such that there is no complete K_{s, t} bipart

NORA: A Harness-Engineered Autonomous Research Agent for End-to-End Spatial Data Science

SafetyDGX agent

arXiv:2605.02092v1 Announce Type: new Abstract: The automation of scientific research workflows has emerged as a transformative frontier in artificial intelligence, yet existing autonomous research ag

On the Privacy of LLMs: An Ablation Study

ResearchDGX agent

arXiv:2605.02255v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly deployed in interactive and retrieval-augmented settings, raising significant privacy concerns. While at

Optimization of CV-QKD Under Practical Constraints

ResearchDGX agent

arXiv:2605.02045v1 Announce Type: cross Abstract: Using reinforcement learning, we optimize for practical hardware constraints, including limited FIR filter taps at the transmitter and receiver, mean

ORPilot: A Production-Oriented Agentic LLM-for-OR Tool for Optimization Modeling

Model ReleasesDGX agent

arXiv:2605.02728v1 Announce Type: new Abstract: This paper presents ORPilot, an open-source agentic AI system that translates real-world business problems into solver-ready optimization models. Unlike

Partial-differential-algebraic equations of nonlinear dynamics by Physics-Informed Neural-Network: (I) Operator splitting and framework assessment

ResearchDGX agent

arXiv:2408.01914v4 Announce Type: replace-cross Abstract: Several forms for constructing novel physics-informed neural-networks (PINN) for the solution of partial-differential-algebraic equations base

PERSA: Reinforcement Learning for Professor-Style Personalized Feedback with LLMs

Model ReleasesDGX agent

arXiv:2605.01123v1 Announce Type: new Abstract: Large language models (LLMs) can provide automated feedback in educational settings, but aligning an LLMs style with a specific instructors tone while m

Personalized Digital Health Modeling with Adaptive Support Users

ApplicationsDGX agent

arXiv:2605.02004v1 Announce Type: new Abstract: Personalized models are essential in digital health because individuals exhibit substantial physiological and behavioral heterogeneity. Yet personalizat

PhysicianBench: Evaluating LLM Agents in Real-World EHR Environments

Model ReleasesDGX agent

arXiv:2605.02240v1 Announce Type: new Abstract: We introduce PhysicianBench, a benchmark for evaluating LLM agents on physician tasks grounded in real clinical setting within electronic health record

Poly-EPO: Training Exploratory Reasoning Models

SafetyDGX agent

arXiv:2604.17654v3 Announce Type: replace Abstract: Exploration is a cornerstone of learning from experience: it enables agents to find solutions to complex problems, generalize to novel ones, and sca

Position: How can Graphs Help Large Language Models?

ResearchDGX agent

arXiv:2605.02452v1 Announce Type: new Abstract: With the rapid advancement of large language models (LLMs), classic graph learning tasks have greatly benefited from LLMs, including improved encoding o

Position: LLM Serving Needs Mathematical Optimization and Algorithmic Foundations, Not Just Heuristics

ResearchDGX agent

arXiv:2605.01280v1 Announce Type: cross Abstract: This position paper argues that LLM inference serving has outgrown generic heuristics and now demands mathematical optimization and algorithmic founda

Position: Safety and Fairness in Agentic AI Depend on Interaction Topology, Not on Model Scale or Alignment

SafetyDGX agent

arXiv:2605.01147v1 Announce Type: new Abstract: As large language models are increasingly deployed as interacting agents in high-stakes decisions, the AI safety community assumes that safety propertie

(POSTER) From Sensors to Insight: Rapid, Edge-to-Core Application Development for Sensor-Driven Applications

ResearchDGX agent

arXiv:2605.02844v1 Announce Type: cross Abstract: Scientists increasingly rely on sensor-based data; however transforming raw streams into insights across the edge-to-cloud continuum remains difficult

Practical Limits of Autonomous Test Repair: A Multi-Agent Case Study with LLM-Driven Discovery and Self-Correction

AgentsDGX agent

arXiv:2605.01471v1 Announce Type: cross Abstract: Maintaining reliable UI test suites in large-scale enterprise applications is a persistent and costly challenge. We present an industrial case study o

Privacy Preserving Machine Learning Workflow: from Anonymization to Personalized Differential Privacy Budgets in Federated Learning

SafetyDGX agent

arXiv:2605.02372v1 Announce Type: cross Abstract: The growing development of artificial intelligence based solutions, together with privacy legislation, has driven the rise of the so-called privacy pr

Quality-Aware Exploration Budget Allocation for Cooperative Multi-Agent Reinforcement Learning

AgentsDGX agent

arXiv:2605.01865v1 Announce Type: cross Abstract: Cooperative multi-agent reinforcement learning (MARL) requires agents to discover joint strategies in a combinatorially large state-action space, yet

Re-Key-Free, Risky-Free: Adaptable Model Usage Control

ApplicationsDGX agent

arXiv:2511.18772v2 Announce Type: replace-cross Abstract: Deep neural networks (DNNs) have become valuable intellectual property of model owners, due to the substantial resources required for their de

Reasoning Models Can be Accurately Pruned Via Chain-of-Thought Reconstruction

Model ReleasesDGX agent

arXiv:2509.12464v2 Announce Type: replace Abstract: Reasoning language models such as DeepSeek-R1 produce long chain-of-thought traces during inference time which make them costly to deploy at scale.

Reinforcement Learning Trained Observer Control for Bearings-Only Tracking

SafetyDGX agent

arXiv:2605.02120v1 Announce Type: new Abstract: This paper develops a deep reinforcement learning based observer control policy for autonomous bearings-only tracking of a moving target. The observer m

Reliable AI Needs to Externalize Implicit Knowledge: A Human-AI Collaboration Perspective

ResearchDGX agent

arXiv:2605.02010v1 Announce Type: new Abstract: This position paper argues that reliable AI requires infrastructure for human validation of implicit knowledge. AI learns from both explicit knowledge (

Repurposing and Evaluating the (In)Feasibility of Dataset Poisoning enabled Watermarking for Contrastive Learning

ResearchDGX agent

arXiv:2605.01834v1 Announce Type: cross Abstract: Contrastive learning (CL) reduces annotation cost via auto-derived supervisory signals. Since large-scale in-house CL datasets are infeasible, relianc

Resource-Efficient Reinforcement for Reasoning Large Language Models via Dynamic One-Shot Policy Refinement

SafetyDGX agent

arXiv:2602.00815v2 Announce Type: replace Abstract: Large language models (LLMs) have exhibited remarkable performance on complex reasoning tasks, with reinforcement learning under verifiable rewards

Rethinking Explanations: Formalizing Contrast in Description Logics

ResearchDGX agent

arXiv:2605.01442v1 Announce Type: new Abstract: There has been a growing interest in explaining entailments over description logic (DL) knowledge bases. The existing explanation formalisms focus on ju

Retrieval and Multi-Hop Reasoning in 1M-Token Context Windows: Evaluating LLMs on Classical Chinese Text

Model ReleasesDGX agent

arXiv:2605.02173v1 Announce Type: new Abstract: We evaluate the long-context retrieval and reasoning capabilities of five frontier large language models with advertised 1M-token context windows on a c

Retrieval-Augmented LLMs for Security Incident Analysis

Model ReleasesDGX agent

arXiv:2603.18196v3 Announce Type: replace-cross Abstract: Investigating cybersecurity incidents requires collecting and analyzing evidence from multiple log sources, including intrusion detection aler

Runtime Evaluation of Procedural Content Generation in an Endless Runner Game Using Autonomous Agents

AgentsDGX agent

arXiv:2605.01783v1 Announce Type: new Abstract: Procedural Content Generation (PCG) enables game content to be created algorithmically without direct manual level-design effort, but it introduces a se

SCALER:Synthetic Scalable Adaptive Learning Environment for Reasoning

ApplicationsDGX agent

arXiv:2601.04809v5 Announce Type: replace Abstract: Reinforcement learning (RL) offers a principled way to enhance the reasoning capabilities of large language models, yet its effectiveness hinges on

SCGNN: Semantic Consistency enhanced Graph Neural Network Guided by Granular-ball Computing

Model ReleasesDGX agent

arXiv:2605.02617v2 Announce Type: new Abstract: Capturing semantic consistency among nodes is crucial for effective graph representation learning. Existing approaches typically rely on k-nearest neigh

SCION: Size-aware Policy Orchestration for Nonstationary Object Caches (Long Paper Version)

SafetyDGX agent

arXiv:2605.01055v1 Announce Type: cross Abstract: Object caches underpin cloud and edge services, but production workloads are heterogeneous, nonstationary, and throughput-constrained. Recent simple n

SCPRM: A Schema-aware Cumulative Process Reward Model for Knowledge Graph Question Answering

TutorialsDGX agent

arXiv:2605.02819v1 Announce Type: new Abstract: Large language models excel at complex reasoning, yet evaluating their intermediate steps remains challenging. Although process reward models provide st

Seeking Information with RAG-Assistants: Does Model Size Matter in Human-AI Collaborations?

Model ReleasesDGX agent

arXiv:2605.00964v1 Announce Type: cross Abstract: Much research on LLMs has focused on increasing benchmark performance. However, the evaluation of such models in real-world collaborative human-AI wor

Separating Intelligence from Execution: A Workflow Engine for the Model Context Protocol

AgentsDGX agent

arXiv:2605.00827v1 Announce Type: cross Abstract: Large Language Model (LLM) agents increasingly interact with external systems through tool-calling protocols such as the Model Context Protocol (MCP).

Set-Based Training of Neural Barrier Certificates for Safety Verification of Dynamical Systems

SafetyDGX agent

arXiv:2605.02526v1 Announce Type: cross Abstract: Barrier certificates are scalar functions over the state space of dynamical systems that separate all unsafe states from all reachable states. The exi

Sheaf-Theoretic Planning: A Categorical Foundation for Resilient Multi-Agent Autonomous Systems

AgentsDGX agent

arXiv:2605.01879v1 Announce Type: new Abstract: The challenge of engineering autonomous agents capable of navigating the stochastic and adversarial nature of the physical world has historically reside

SMoE: An Algorithm-System Co-Design for Pushing MoE to the Edge via Expert Substitution

SafetyDGX agent

arXiv:2508.18983v3 Announce Type: replace Abstract: The Mixture of Experts (MoE) architecture has emerged as a key technique for scaling Large Language Models by activating only a subset of experts pe

Spectral- and Energy-efficient Multi-BS Multi-RIS Pinching-antenna Systems: A GNN-based Approach

ResearchDGX agent

arXiv:2605.01307v1 Announce Type: cross Abstract: This paper investigates coordinated downlink transmission in a multi-base station (multi-BS) multi-reconfigurable intelligent surface (multi-RIS)-assi

Static Analysis of Recursive SHACL

ResearchDGX agent

arXiv:2605.02787v1 Announce Type: cross Abstract: SHACL (Shapes Constraint Language) expresses constraints on RDF data by means of so-called shapes. Its central service is validation: verifying whethe

Strategy-Aware Optimization Modeling with Reasoning LLMs

ApplicationsDGX agent

arXiv:2605.02545v1 Announce Type: new Abstract: Large language models (LLMs) can generate syntactically valid optimization programs, yet often struggle to reliably choose an effective modeling strateg

← Previous
1…288289290291292…358
Next →