AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent
84,433Total entries
1Added by human
84,432Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
6 May 2026

Cripping AI: Reimagining AI Through Lived Disability Experiences

TutorialsDGX agent

arXiv:2605.02080v1 Announce Type: cross Abstract: Drawing on crip theory, this paper proposes cripping AI as a guiding framework to center lived disability experiences in AI research and development.

CyberAId: AI-Driven Cybersecurity for Financial Service Providers

AgentsDGX agent

arXiv:2605.01892v1 Announce Type: new Abstract: European financial institutions face mounting regulatory pressure while their security operations centres remain constrained not by data or staffing but

Data driven approach for Outdoor Channel Prediction in 5G and Beyond

ResearchDGX agent

arXiv:2605.01777v1 Announce Type: cross Abstract: An evolution of Wireless Communications towards 5G and beyond provides improved user experience in terms of quality of services. Understanding and est


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

DataClaw: A Process-Oriented Agent Benchmark for Exploratory Real-World Data Analysis

Model ReleasesDGX agent

arXiv:2605.02503v1 Announce Type: new Abstract: Evaluating autonomous data analysis agents requires testing their ability to perform exploratory analysis in underexplored data environments. However, m

DataEvolver: Let Your Data Build and Improve Itself via Goal-Driven Loop Agents

Model ReleasesDGX agent

arXiv:2605.01789v1 Announce Type: new Abstract: Constructing controllable visual data is a major bottleneck for image editing and multimodal understanding. Useful supervision is rarely produced by a s

Deciphering Shortcut Learning from an Evolutionary Game Theory Perspective

SafetyDGX agent

arXiv:2605.02658v2 Announce Type: new Abstract: Shortcut learning causes deep learning models to rely on non-essential features within the data. However, its formation in deep neural network training

Design-OS: A Specification-Driven Framework for Engineering System Design with a Control-Systems Design Case

AgentsDGX agent

arXiv:2603.20151v2 Announce Type: replace-cross Abstract: Engineering system design -- whether mechatronic, control, or embedded -- often proceeds in an ad hoc manner, with requirements left implicit

DiagramNet: An End-to-End Recognition Framework and Dataset for Non-Standard System-Level Diagrams

Model ReleasesDGX agent

arXiv:2605.01338v1 Announce Type: new Abstract: System-level diagrams encode the architectural blueprint of chip design, specifying module functions, dataflows, and interface protocols. However, non-s

Discover Fast Power Allocation Solution for Multi-Target Tracking via AlphaEvolve Evolution

ResearchDGX agent

arXiv:2605.01794v1 Announce Type: cross Abstract: Efficient radar resource allocation is a fundamental yet computationally challenging problem, as optimal solutions typically require iterative optimiz

Disentangling Intent from Role: Adversarial Self-Play for Persona-Invariant Safety Alignment

SafetyDGX agent

arXiv:2605.01899v1 Announce Type: new Abstract: The growing capabilities of large language models (LLMs) have driven their widespread deployment across diverse domains, even in potentially high-risk s

Distilling Long-CoT Reasoning through Collaborative Step-wise Multi-Teacher Decoding

ResearchDGX agent

arXiv:2605.02290v1 Announce Type: new Abstract: Distilling large reasoning models is essential for making Long-CoT reasoning practical, as full-scale inference remains computationally prohibitive. Exi

DocSync: Agentic Documentation Maintenance via Critic-Guided Reflexion

Model ReleasesDGX agent

arXiv:2605.02163v1 Announce Type: cross Abstract: Software documentation frequently drifts from executable logic as codebases evolve, creating technical debt that degrades maintainability and causes d

Double Rectified Linear Unit-based Modular Semantics for Quantitative Bipolar Argumentation Framework

ResearchDGX agent

arXiv:2605.02551v1 Announce Type: new Abstract: Quantitative Bipolar Argumentation Frameworks (QBAFs) provide an alternative approach to computing argument acceptability in Bipolar Argumentation Frame

E-MIA: Exam-Style Black-Box Membership Inference Attacks against RAG Systems

ResearchDGX agent

arXiv:2605.00955v1 Announce Type: cross Abstract: Retrieval-Augmented Generation (RAG) equips large language models (LLMs) with external evidence by retrieving documents at inference time, but it also

Effect-Transparent Governance for AI Workflow Architectures: Semantic Preservation, Expressive Minimality, and Decidability Boundaries

ResearchDGX agent

arXiv:2605.01030v2 Announce Type: new Abstract: We present a machine-checked formalization of structurally governed AI workflow architectures and prove that effect-level governance can be imposed with

Efficient Temporal Datalog Materialisation for Composite Event Recognition

SafetyDGX agent

arXiv:2605.02488v1 Announce Type: new Abstract: Several applications demand the timely detection of critical situations, such as threats to safety and transparency, over high-velocity streams of symbo

Empowering LLM Agents with Geospatial Awareness: Toward Grounded Reasoning for Wildfire Response

ResearchDGX agent

arXiv:2510.12061v2 Announce Type: replace Abstract: Effective disaster response is essential for safeguarding lives and property. Existing statistical approaches often lack semantic context, generaliz

EngiAgent: Fully Connected Coordination of LLM Agents for Solving Open-ended Engineering Problems with Feasible Solutions

AgentsDGX agent

arXiv:2605.02289v1 Announce Type: new Abstract: Engineering problem solving is central to real-world decision-making, requiring mathematical formulations that not only represent complex problems but a

EngiBench: A Benchmark for Evaluating Large Language Models on Engineering Problem Solving

Model ReleasesDGX agent

arXiv:2509.17677v2 Announce Type: replace Abstract: Large language models (LLMs) have shown strong performance on mathematical reasoning under well-defined conditions. However, real-world engineering

Entanglement is Half the Story: Post-Selection vs. Partial Traces

ResearchDGX agent

arXiv:2605.02385v1 Announce Type: cross Abstract: While tensor networks have their traditional application in simulating quantum systems, in the recent decade they have gathered interest as machine le

EO-Gym: A Multimodal, Interactive Environment for Earth Observation Agents

Model ReleasesDGX agent

arXiv:2605.01250v1 Announce Type: new Abstract: Earth Observation (EO) analysis is inherently interactive: resolving uncertainty often requires expanding the region of interest, retrieving historical

Epistemic Reject Option Prediction

ResearchDGX agent

arXiv:2511.04855v2 Announce Type: replace Abstract: In high-stakes applications, predictive models must not only produce accurate predictions but also quantify and communicate their uncertainty. Rejec

Evaluating Agentic AI in the Wild: Failure Modes, Drift Patterns, and a Production Evaluation Framework

Model ReleasesDGX agent

arXiv:2605.01604v1 Announce Type: new Abstract: Existing evaluation frameworks for large language models -- including HELM, MT-Bench, AgentBench, and BIG-bench -- are designed for controlled, single-s

Explainable AI for Blind and Low-Vision Users: Navigating Trust, Modality, and Interpretability in the Agentic Era

AgentsDGX agent

arXiv:2604.00187v2 Announce Type: replace-cross Abstract: Explainable Artificial Intelligence (XAI) is critical for ensuring trust and accountability, yet its development remains predominantly visual.

Faithful Mobile GUI Agents with Guided Advantage Estimator

AgentsDGX agent

arXiv:2605.01208v1 Announce Type: new Abstract: Vision-language model based graphical user interface (GUI) agents have shown strong interaction capabilities. However, they often behave unfaithfully, r

False Friends in the Shell: Unveiling the Emoticon Semantic Confusion in Large Language Models

SafetyDGX agent

arXiv:2601.07885v2 Announce Type: replace-cross Abstract: Emoticons are widely used in digital communication to convey affective intent, yet their safety implications for Large Language Models (LLMs)

FEDIN: Frequency-Enhanced Deep Interest Network for Click-Through Rate Prediction

Model ReleasesDGX agent

arXiv:2605.01726v1 Announce Type: cross Abstract: Sequential recommendation models often struggle to capture latent periodic patterns in user interests, primarily due to the noise inherent in time-dom

First-Order Efficiency for Probabilistic Value Estimation via A Statistical Viewpoint

ResearchDGX agent

arXiv:2605.02827v1 Announce Type: new Abstract: Probabilistic values, including Shapley values and semivalues, provide a model-agnostic framework to attribute the behavior of a black-box model to data

Foundation-Model-Based Agents in Industrial Automation: Purposes, Capabilities, and Open Challenges

AgentsDGX agent

arXiv:2605.02592v1 Announce Type: new Abstract: Foundation models, particularly large language models, are increasingly integrated into agent architectures for industrial tasks such as decision suppor

From Experimental Limits to Physical Insight: A Retrieval-Augmented Multi-Agent Framework for Interpreting Searches Beyond the Standard Model

AgentsDGX agent

arXiv:2605.02491v1 Announce Type: cross Abstract: Modern searches for physics beyond the Standard Model produce rapidly expanding literature containing heterogeneous information, including textual ana

From Laboratory to Real-World Applications: Benchmarking Agentic Code Reasoning at the Repository Level

Model ReleasesDGX agent

arXiv:2601.03731v3 Announce Type: replace-cross Abstract: As large language models (LLMs) evolve into autonomous agents, evaluating repository-level reasoning, the ability to maintain logical consiste

From Sensors to Insight: Rapid, Edge-to-Core Application Development for Sensor-Driven Applications

ResearchDGX agent

arXiv:2605.02859v1 Announce Type: cross Abstract: Scientists increasingly rely on sensor-based data, yet transforming raw streams into insights across the edge-to-cloud continuum remains difficult. Pr

From SFT to RL: Demystifying the Post-Training Pipeline for LLM-based Vulnerability Detection

SafetyDGX agent

arXiv:2602.14012v2 Announce Type: replace-cross Abstract: The integration of LLMs into vulnerability detection (VD) has shifted the field toward more interpretable and context-aware analysis. While po

GDPR Auto-Formalization with AI Agents and Human Verification

AgentsDGX agent

arXiv:2604.14607v2 Announce Type: replace Abstract: We study the overall process of automatic formalization of GDPR provisions using large language models, within a human-in-the-loop verification fram

Generative-AI and the transformation of workforce. A job postings-driven analysis

ResearchDGX agent

arXiv:2605.00843v1 Announce Type: cross Abstract: This paper investigates how generative-artificial intelligence AI is reshaping job requirements, skill compositions and sectoral dynamics across globa

GhostServe: A Lightweight Checkpointing System in the Shadow for Fault-Tolerant LLM Serving

AgentsDGX agent

arXiv:2605.00831v1 Announce Type: cross Abstract: The rise of million-token, agent-based applications has placed unprecedented demands on large language model (LLM) inference services. The long-runnin

GISclaw: A Comprehensive Open-Source LLM Agent System for Realistic Multi-Step Geospatial Analysis

Local AiDGX agent

arXiv:2603.26845v2 Announce Type: replace-cross Abstract: Most LLM-driven GIS assistants solve narrow single-step tasks tightly coupled to proprietary platforms such as ArcGIS or QGIS, limiting their

GOAT: A Training Framework for Goal-Oriented Agent with Tools

Model ReleasesDGX agent

arXiv:2510.12218v2 Announce Type: replace Abstract: Current approaches rely on zero-shot evaluation due to the absence of training data; while proprietary models such as GPT-4 exhibit strong reasoning

Governing What the EU AI Act Excludes: Accountability for Autonomous AI Agents in Smart City Critical Infrastructure

SafetyDGX agent

arXiv:2605.01091v1 Announce Type: cross Abstract: When a traffic signal controller adjusts green phases and a grid manager curtails power on the same corridor, each system may comply with its own obli

Grounding Multi-Hop Reasoning in Structural Causal Models via Group Relative Policy Optimization

SafetyDGX agent

arXiv:2605.01482v1 Announce Type: new Abstract: Multi-Hop Fact Verification (MHFV) necessitates complex reasoning across disparate evidence, posing significant challenges for Large Language Models (LL

HAAS: A Policy-Aware Framework for Adaptive Task Allocation Between Humans and Artificial Intelligence Systems

Model ReleasesDGX agent

arXiv:2605.02832v1 Announce Type: new Abstract: Deciding how to distribute work between humans and AI systems is a central challenge in organisational design. Most approaches treat this as a binary ch

HeavySkill: Heavy Thinking as the Inner Skill in Agentic Harness

AgentsDGX agent

arXiv:2605.02396v1 Announce Type: new Abstract: Recent advances in agentic harness with orchestration frameworks that coordinate multiple agents with memory, skills, and tool use have achieved remarka

HepScript: A Dual-Use DSL for Human-AI Collaborative Data Analysis Workflows in High-Energy Physics

AgentsDGX agent

arXiv:2605.01423v1 Announce Type: cross Abstract: The escalating data scale in High-Energy Physics (HEP) fuels a growing aspiration for higher analytical efficiency. While Large Language Models (LLMs)

HistCAD: A Constraint-Aware Parametric History-Based CAD Representation, Dataset, and Benchmark with Industrial Complexity

Model ReleasesDGX agent

arXiv:2602.19171v2 Announce Type: replace-cross Abstract: Parametric CAD sequences are reusable because dimensional and geometric constraints govern how parameter changes propagate. Existing CAD gener

Hybrid Inspection and Task-Based Access Control in Zero-Trust Agentic AI

AgentsDGX agent

arXiv:2605.02682v1 Announce Type: new Abstract: Authorizing Large Language Model (LLM)-driven agents to dynamically invoke tools and access protected resources introduces significant security risks, a

'I Don't Know' -- Towards Appropriate Trust with Certainty-Aware Retrieval Augmented Generation

Model ReleasesDGX agent

arXiv:2605.00957v1 Announce Type: cross Abstract: Achieving the right amount of trust in AI systems is important, but challenging. The problem is exacerbated with the rise of Large Language Models (LL

InsTraj: Instructing Diffusion Models with Travel Intentions to Generate Real-world Trajectories

ApplicationsDGX agent

arXiv:2604.04106v2 Announce Type: replace Abstract: The generation of realistic and controllable GPS trajectories is a fundamental task for applications in urban planning, mobility simulation, and pri

Intelligent Agents with Emotional Intelligence: Current Trends, Challenges, and Future Prospects

AgentsDGX agent

arXiv:2511.20657v2 Announce Type: replace-cross Abstract: The development of agents with emotional intelligence is becoming increasingly vital due to their significant role in human-computer interacti

Intervention Complexity as a Canonical Reward and a Measure of Intelligence

SafetyDGX agent

arXiv:2605.02175v1 Announce Type: new Abstract: The Legg--Hutter universal intelligence measure provides a rigorous scalar assessment of general intelligence as expected reward across all computable e

Investigating the Effects of Different Levels of User Control in an Interactive Educational Recommender System

ResearchDGX agent

arXiv:2605.01400v1 Announce Type: cross Abstract: Educational recommender systems (ERSs) are becoming increasingly important in enhancing educational outcomes and personalizing learning experiences by

Iterative Finetuning is Mostly Idempotent

ResearchDGX agent

arXiv:2605.01130v1 Announce Type: new Abstract: If a model has some behavioral tendency, such as sycophancy or misalignment, and it is trained on its own outputs, will the tendency be amplified in the

KG-First, LLM-Fallback: A Hybrid Microservice for Grounded Skill Search and Explanation

ApplicationsDGX agent

arXiv:2605.01582v1 Announce Type: cross Abstract: Authoritative competency frameworks such as ESCO, ROME, and O*NET are essential for aligning education with labor market needs, yet their technical co

Khala: Scaling Acoustic Token Language Models Toward High-Fidelity Music Generation

SafetyDGX agent

arXiv:2605.01790v1 Announce Type: cross Abstract: A common design pattern in high-quality music generation is to handle structure and fidelity in different representation spaces: a generator first mod

Latent State Design for World Models under Sufficiency Constraints

AgentsDGX agent

arXiv:2605.01694v1 Announce Type: new Abstract: A world model matters to an agent only through the state it constructs. That state must preserve some information, discard other information, and suppor

Learning Decentralized LLM Collaboration with Multi-Agent Actor Critic

AgentsDGX agent

arXiv:2601.21972v4 Announce Type: replace Abstract: Recent work has explored optimizing LLM collaboration through Multi-Agent Reinforcement Learning (MARL). However, most MARL fine-tuning approaches r

Less Interaction But More Explanation: A Communication Perspective on Agentic AI Interfaces

AgentsDGX agent

arXiv:2605.01610v1 Announce Type: cross Abstract: AI systems have long been expected to interact with users, answering questions, generating content, and continuing (social) conversations. Agentic AI,

Lifting Traces to Logic: Programmatic Skill Induction with Neuro-Symbolic Learning for Long-Horizon Agentic Tasks

AgentsDGX agent

arXiv:2605.01293v1 Announce Type: new Abstract: Foundation model-driven agents often struggle with long-horizon planning due to the transient nature of purely prompting-based reasoning. While existing

LiveFMBench: Unveiling the Power and Limits of Agentic Workflows in Specification Generation

Model ReleasesDGX agent

arXiv:2605.01394v1 Announce Type: cross Abstract: Formal specification is essential for rigorous program verification, yet writing correct specifications remains costly and difficult to automate. Alth

LLM-Assisted Repository-Level Generation with Structured Spec-Driven Engineering

TutorialsDGX agent

arXiv:2605.02455v1 Announce Type: cross Abstract: State-of-the-art Large Language Models (LLMs) excel in code generation at the function level. However, the output quality significantly declines when

LLM-enabled Social Agents

AgentsDGX agent

arXiv:2605.02335v1 Announce Type: cross Abstract: Large Language Models (LLMs) have transformed agent-agent and human-agent interaction by enabling software, physical, and simulation agents to communi

← Previous
1…287288289290291…358
Next →