AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent
84,433Total entries
1Added by human
84,432Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
7 May 2026

Unifying Dynamical Systems and Graph Theory to Mechanistically Understand Computation in Neural Networks

SafetyDGX agent

arXiv:2605.03598v2 Announce Type: cross Abstract: Understanding how biological and artificial neural networks implement computation from connectivity is a central problem in neuroscience and machine l

VCBench: Benchmarking LLMs in Venture Capital

Model ReleasesDGX agent

arXiv:2509.14448v2 Announce Type: replace Abstract: Benchmarks such as SWE-bench and ARC-AGI demonstrate how shared datasets accelerate progress toward artificial general intelligence (AGI). We introd

What Happens Inside Agent Memory? Circuit Analysis from Emergence to Diagnosis

Model ReleasesDGX agent

arXiv:2605.03354v1 Announce Type: new Abstract: Agent memory failures are silent: an LLM-based agent can produce a fluent response even when it fails to extract, retain, or retrieve the information ne


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

What You Think is What You See: Driving Exploration in VLM Agents via Visual-Linguistic Curiosity

AgentsDGX agent

arXiv:2605.03782v1 Announce Type: new Abstract: To navigate partially observable visual environments, recent VLM agents increasingly internalize world modeling capabilities into their policies via exp

When Agents Handle Secrets: A Survey of Confidential Computing for Agentic AI

HardwareDGX agent

arXiv:2605.03213v1 Announce Type: cross Abstract: Agentic AI systems, specifically LLM-driven agents that plan, invoke tools, maintain persistent memory, and delegate tasks to peer agents via protocol

6 May 2026

12 Angry AI Agents: Evaluating Multi-Agent LLM Decision-Making Through Cinematic Jury Deliberation

Model ReleasesDGX agent

arXiv:2605.01986v1 Announce Type: new Abstract: What if the twelve jurors of Sidney Lumet's 12 Angry Men (1957) were not men, but large language models? Would the one juror who disagrees still be able

6G Needs Agents: Toward Agentic AI-Native Networks for Autonomous Intelligence

Model ReleasesDGX agent

arXiv:2605.01546v1 Announce Type: cross Abstract: Sixth-generation (6G) networks are increasingly envisioned as AI-native infrastructures integrating communication, sensing, and computing into a unifi

A Cellular Doctrine of Morality: Intrinsic Active Precision and the Mind-Reality Overload Dilemma

Local AiDGX agent

arXiv:2605.01376v1 Announce Type: new Abstract: Current AI systems, grounded in oversimplified neuroscience, risk eroding the distinction between truth and falsehood. They maximize reward by amplifyin

A Compound AI Agent for Conversational Grant Discovery

AgentsDGX agent

arXiv:2605.02366v1 Announce Type: new Abstract: Research funding discovery remains fundamentally fragmented: researchers navigate disparate agency portals (e.g., in the United States, NSF, NIH, DARPA,

A Knowledge-Driven LLM-Based Decision-Support System for Explainable Defect Analysis and Mitigation Guidance in Laser Powder Bed Fusion

SafetyDGX agent

arXiv:2605.01100v1 Announce Type: new Abstract: This work presents a knowledge-driven decision-support system that integrates structured defect knowledge with LLM-based reasoning to provide explainabl

A Low-Latency Fraud Detection Layer for Detecting Adversarial Interaction Patterns in LLM-Powered Agents

AgentsDGX agent

arXiv:2605.01143v1 Announce Type: new Abstract: Large Language Model (LLM)-powered agents demonstrate strong capabilities in autonomous task execution, tool use, and multi-step reasoning. However, the

A Neuro-Symbolic Framework for Accountability in Public-Sector AI

ApplicationsDGX agent

arXiv:2512.12109v3 Announce Type: replace-cross Abstract: Automated eligibility systems increasingly determine access to essential public benefits, but the explanations they generate often fail to ref

A Sentence Relation-Based Approach to Sanitizing Malicious Instructions

ResearchDGX agent

arXiv:2605.01078v1 Announce Type: cross Abstract: Retrieval-augmented generation and tool-integrated LLM agents increasingly depend on external textual sources. This reliance broadens the available at

A Study of Belief Revision Postulates in Multi-Agent Systems (Extended Version)

AgentsDGX agent

arXiv:2605.02249v1 Announce Type: new Abstract: We investigate the belief revision problem in epistemic planning, i.e., what will be the beliefs of all agents in a multi-agent system after an agent ga

A Target-Free Harmonization Method for MRI

ResearchDGX agent

arXiv:2605.01282v1 Announce Type: cross Abstract: In MRI, variations in scan parameters, sequence, or hardware can lead to discrepancies in image appearance, even for the same subject. These inconsist

ABD: Default Exception Abduction in Finite First Order Worlds

Model ReleasesDGX agent

arXiv:2602.18843v3 Announce Type: replace Abstract: We introduce ABD, a benchmark for default-exception abduction over finite first-order worlds. Given a background theory with an abnormality predicat

ABox Abduction for Inconsistent Knowledge Bases under Repair Semantics

TutorialsDGX agent

arXiv:2605.01341v1 Announce Type: cross Abstract: Given a knowledge base (KB) with a non-entailed fact, the ABox abduction problem asks for possible extensions of the KB that would entail this fact. T

AcademiClaw: When Students Set Challenges for AI Agents

Model ReleasesDGX agent

arXiv:2605.02661v1 Announce Type: new Abstract: Benchmarks within the OpenClaw ecosystem have thus far evaluated exclusively assistant-level tasks, leaving the academic-level capabilities of OpenClaw

Adaptive 3D-RoPE: Physics-Aligned Rotary Positional Encoding for Wireless Foundation Models

SafetyDGX agent

arXiv:2605.00968v1 Announce Type: cross Abstract: Positional encoding plays a pivotal role in determin?ing the extrapolation and generalization performance of wireless foundation models for channel st

Agentic AI Systems Should Be Designed as Marginal Token Allocators

AgentsDGX agent

arXiv:2605.01214v1 Announce Type: new Abstract: This position paper argues that agentic AI systems should be designed and evaluated as marginal token allocation economies rather than as text generator

Agentic Multi-Source Grounding for Enhanced Query Intent Understanding: A DoorDash Case Study

AgentsDGX agent

arXiv:2603.01486v2 Announce Type: replace Abstract: Accurately mapping user queries to business categories is a fundamental Information Retrieval challenge for multi-category marketplaces, where conte

AI Agents for Sustainable SMEs: A Green ESG Assessment Framework

AgentsDGX agent

arXiv:2605.00841v1 Announce Type: new Abstract: This study presents a novel, AI-driven framework for assessing Environmental, Social, and Governance (ESG) performance in European small and medium-size

AI and Open-data Driven Scalable Solar Power Profiling

Model ReleasesDGX agent

arXiv:2605.02738v1 Announce Type: new Abstract: Solar photovoltaic (PV) deployment is expanding rapidly, yet detailed, up-to-date information on the spatial distribution and capacity of rooftop PV rem

AI Expert Twin: Capturing Expert Cognition for Human-Centred, Practice-Based Learning

ApplicationsDGX agent

arXiv:2605.01401v1 Announce Type: cross Abstract: Tacit knowledge embedded in expert practice remains difficult to capture, formalise, and scale. While AI-driven educational systems have advanced pers

AI-Generated Smells: An Analysis of Code and Architecture in LLM and Agent-Driven Development

AgentsDGX agent

arXiv:2605.02741v1 Announce Type: cross Abstract: The promise of Large Language Models in automated software engineering is often measured by functional correctness, overlooking the critical issue of

AI Safety as Control of Irreversibility: A Systems Framework for Decision-Energy and Sovereignty Boundaries

Model ReleasesDGX agent

arXiv:2605.01415v1 Announce Type: new Abstract: Recent AI systems compress the distance between capability growth and capability deployment. Earlier high-risk technologies were slowed by capital inten

AIs and Humans with Agency

ApplicationsDGX agent

arXiv:2605.02810v1 Announce Type: new Abstract: This paper compares agency in humans with potential agency in AI programs. Human agency takes many years to develop, as the frontal lobe is activated. E

Algebraic Semantics of Governed Execution: Monoidal Categories, Effect Algebras, and Coterminous Boundaries

SafetyDGX agent

arXiv:2605.01032v2 Announce Type: new Abstract: We present an algebraic semantics for governed execution in which governance is axiomatized, compositional, and coterminous with expressibility. The fra

AMSnet-q: Unsupervised Circuit Identification and Performance Labeling for AMS Circuits

ResearchDGX agent

arXiv:2605.01404v1 Announce Type: cross Abstract: Analog and mixed-signal (AMS) circuit design remains heavily reliant on expert knowledge. While recent AI-driven automation tools can generate candida

An Empirical Study of Agent Skills for Healthcare: Practice, Gaps, and Governance

Local AiDGX agent

arXiv:2605.02709v1 Announce Type: new Abstract: Healthcare automation is shaped by local procedures and organizational constraints, so agent capabilities rarely transfer unchanged across settings. Age

An explainable hypothesis-driven approach to Drug-Induced Liver Injury with HADES

Model ReleasesDGX agent

arXiv:2605.02669v1 Announce Type: new Abstract: Drug-induced liver injury (DILI) remains a leading cause of late-stage clinical trial attrition. However, existing computational predictors primarily re

APIOT: Autonomous Vulnerability Management Across Bare-Metal Industrial OT Networks

AgentsDGX agent

arXiv:2605.02346v1 Announce Type: cross Abstract: Bare-metal operational technology (OT) devices -- especially the microcontrollers running Modbus/TCP and CoAP at the base of industrial control system

Architectural Obsolescence of Unhardened Agentic-AI Runtimes

SafetyDGX agent

arXiv:2605.01740v1 Announce Type: cross Abstract: An agentic-AI runtime issues tool calls, sends messages, and actuates devices on behalf of an LLM. Catching the four ways an action can diverge from i

Are LLMs More Skeptical of Entertainment News?

Model ReleasesDGX agent

arXiv:2605.01727v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used for automated news credibility assessment, yet it remains unclear whether they apply even-handed stan

Are we Doomed to an AI Race? Why Self-Interest Could Drive Countries Towards a Moratorium on Superintelligence

ResearchDGX agent

arXiv:2605.01297v1 Announce Type: cross Abstract: This paper uses game theory to argue that, contrary to the prevailing view, a moratorium on Artificial Superintelligence (ASI) can be in a state's sel

Artificial Jagged Intelligence as Uneven Optimization Energy Allocation Capability Concentration, Redistribution, and Optimization Governance

Model ReleasesDGX agent

arXiv:2605.01420v1 Announce Type: new Abstract: Artificial Jagged Intelligence (AJI) denotes a recurring pattern in which large learning systems exhibit strong local capabilities while remaining weak

Beyond Isolated Investor: Predicting Startup Success via Roleplay-Based Collective Agents

AgentsDGX agent

arXiv:2512.22608v3 Announce Type: replace Abstract: Due to the high value and high failure rates of startups, predicting their success is a critical challenge. Existing approaches typically model star

Beyond Scalars: Evaluating and Understanding LLM Reasoning via Geometric Progress and Stability

ResearchDGX agent

arXiv:2603.10384v2 Announce Type: replace Abstract: Evaluating LLM reliability via scalar probabilities often fails to capture the structural dynamics of reasoning. We introduce TRACED, a framework th

Beyond State Machines: Executing Network Procedures with Agentic Tool-Calling Sequences

AgentsDGX agent

arXiv:2605.02584v1 Announce Type: cross Abstract: Agentic AI will be an essential enabling technology for designing future mobile communication systems, which could provide flexible and customized ser

Caliper-in-the-Loop: Black-Box Optimization for Hyperledger Fabric Performance Tuning

ResearchDGX agent

arXiv:2605.02690v1 Announce Type: cross Abstract: Hyperledger Fabric performance depends on many interacting configuration parameters, making manual tuning difficult. We study automated throughput tun

Can Semantic Methods Enhance Team Sports Tactics? A Methodology for Football with Broader Applications

SafetyDGX agent

arXiv:2601.00421v2 Announce Type: replace Abstract: This paper explores how semantic-space reasoning, traditionally used in computational linguistics, can be extended to tactical decision-making in te

Catching the Infection Before It Spreads: Foresight-Guided Defense in Multi-Agent Systems

SafetyDGX agent

arXiv:2605.01758v1 Announce Type: new Abstract: Large multimodal model-based Multi-Agent Systems (MASs) enable collaborative complex problem solving through specialized agents. However, MASs are vulne

Causal Software Engineering: A Vision and Roadmap

Model ReleasesDGX agent

arXiv:2605.02454v1 Announce Type: cross Abstract: Software engineering increasingly involves making high-stakes decisions under uncertainty, using signals from code, field data, and socio-technical pr

CBV: Clean-label Backdoor Attacks on Vision Language Models via Diffusion Models

TutorialsDGX agent

arXiv:2605.02202v1 Announce Type: new Abstract: Vision-Language Models (VLMs) have achieved remarkable success in tasks such as image captioning and visual question answering (VQA). However, as their

CellxPert: Inference-Time MCMC Steering of a Multi-Omics Single-Cell Foundation Model for In-Silico Perturbation

ResearchDGX agent

arXiv:2605.00930v1 Announce Type: cross Abstract: In this work, we introduce CellxPert, a scalable multimodal foundation model that unifies single-cell and spatial multi-omics within a common represen

Certified Purity for Cognitive Workflow Executors: From Static Analysis to Cryptographic Attestation

ResearchDGX agent

arXiv:2605.01037v2 Announce Type: cross Abstract: We present a certified purity architecture that converts governance enforcement in cognitive workflow systems from a runtime convention into a structu

Claw-Eval: Towards Trustworthy Evaluation of Autonomous Agents

SafetyDGX agent

arXiv:2604.06132v2 Announce Type: replace Abstract: Large language models are increasingly deployed as autonomous agents for multi-step workflows in real-world software environments. However, existing

ClinicBot: A Guideline-Grounded Clinical Chatbot with Prioritized Evidence RAG and Verifiable Citations

AgentsDGX agent

arXiv:2605.00846v1 Announce Type: new Abstract: Clinical diagnosis requires answers that are accurate, verifiable, and explicitly grounded in official guidelines. While large language models excel at

Co-Generative De Novo Functional Protein Design

ResearchDGX agent

arXiv:2605.00948v1 Announce Type: cross Abstract: De novo functional protein design aims to generate protein sequences that realize specified biochemical functions without relying on evolutionary temp

Code World Model Preparedness Report

ResearchDGX agent

arXiv:2605.00932v1 Announce Type: cross Abstract: This report documents the preparedness assessment of Code World Model (CWM), a model for code generation and reasoning about code from Meta. We conduc

CoFlow: Coordinated Few-Step Flow for Offline Multi-Agent Decision Making

AgentsDGX agent

arXiv:2605.01457v1 Announce Type: new Abstract: Generative models have emerged as a major paradigm for offline multi-agent reinforcement learning (MARL), but existing approaches require many iterative

Coherent Hierarchical Multi-Label Learning to Defer for Medical Imaging

SafetyDGX agent

arXiv:2605.02734v1 Announce Type: new Abstract: Learning to Defer (L2D) enables a model to predict autonomously or defer to an expert, but prior work largely assumes flat label spaces. We study the fi

Complexity Horizons of Compressed Models in Analog Circuit Analysis

AgentsDGX agent

arXiv:2605.02285v1 Announce Type: new Abstract: The deployment of Large Language Models (LLMs) for specialized engineering domains, such as circuit analysis, often faces a trade-off between reasoning

Compress Then Adapt? No, Do It Together via Task-aware Union of Subspaces

Model ReleasesDGX agent

arXiv:2605.02829v1 Announce Type: new Abstract: Adapting large pretrained models to diverse tasks is now routine, yet the two dominant strategies of parameter-efficient fine-tuning (PEFT) and low-rank

Context-Aware Wireless Token Communication via Joint Token Masking and Detection

ResearchDGX agent

arXiv:2605.02123v1 Announce Type: cross Abstract: The increasing use of token-based representations in language-driven applications has motivated wireless token communication, where tokens are treated

ContextCov: Deriving and Enforcing Executable Constraints from Agent Instruction Files

Model ReleasesDGX agent

arXiv:2603.00822v2 Announce Type: replace-cross Abstract: As Large Language Model (LLM) agents increasingly execute complex, autonomous software engineering tasks, developers rely on natural language

Controllable and Verifiable Process Data Synthesis for Process Reward Models

ResearchDGX agent

arXiv:2605.02395v1 Announce Type: new Abstract: Process reward models (PRMs) rely on high-quality process supervision data, yet existing construction methods often provide limited control over error l

Conventional Commit Classification using Large Language Models and Prompt Engineering

Model ReleasesDGX agent

arXiv:2605.02033v1 Announce Type: cross Abstract: Conventional commits provide a structured format for writing commit messages, which improves readability, software maintenance, and enables automation

Counterfactual Reasoning in Automated Planning

TutorialsDGX agent

arXiv:2605.02603v1 Announce Type: new Abstract: Automated planning traditionally assumes that all aspects of a planning task (initial state, goals, and available actions) are fully specified in advanc

CoVSpec: Efficient Device-Edge Co-Inference for Vision-Language Models via Speculative Decoding

ResearchDGX agent

arXiv:2605.02218v1 Announce Type: new Abstract: Vision-language models (VLMs) have demonstrated strong capabilities in multimodal perception and reasoning. However, deploying large VLMs on mobile devi

← Previous
1…286287288289290…358
Next →