AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,289 results
Safety

Diffusion-Guided Cooperative Policy Learning for Target Tracking Based on Underwater Mobile Agent Networks

DGX agent

arXiv:2603.29426v2 Announce Type: replace-cross Abstract: Multi-agent reinforcement learning (MARL) provides a promising solution for cooperative target tracking in networks of autonomous underwater v

safetyarxiv-cs-lg
13 Aug 2026
Model Releases
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Governing Agentic AI in FinTech

DGX agent

arXiv:2608.11344v1 Announce Type: cross Abstract: Financial institutions are delegating consequential decisions to agentic AI systems that decompose goals, coordinate models and tools, and act with li

model-releasesarxiv-cs-ai
13 Aug 2026
Model Releases

MBA: Multimodal Benchmark and Agents for Real-World Business Ideation

DGX agent

arXiv:2608.11616v1 Announce Type: new Abstract: Agentic systems powered by large language models (LLMs) have opened new opportunities for business ideation. Yet existing approaches remain confined to

model-releasesarxiv-cs-ai
13 Aug 2026
Agents

Multi-Agent Target-Existence Verification and Learned Mask Geometry Refinement: Winning Report of the MeViS-Text Track at the 8th LSVOS Challenge 2026

DGX agent

arXiv:2608.11458v1 Announce Type: new Abstract: We present the first-place solution to the MeViS-Text track of the 8th Large-scale Video Object Segmentation (LSVOS) Challenge 2026: referring video obj

agentsarxiv-cs-cv
13 Aug 2026
Agents

The Sleeping Agent: What Gist-Based Context Compression Loses and Why

DGX agent

arXiv:2608.11775v1 Announce Type: new Abstract: Gist-based context compression---summarising older conversation history into compact representations---is a common approach in long-horizon language mod

agentsarxiv-cs-ai
13 Aug 2026
Model Releases

Actionable Hallucination Detection: Translating Latent Uncertainty into Agentic Critique

DGX agent

arXiv:2608.10430v1 Announce Type: cross Abstract: Large Language Models (LLMs) deployed as AI agents frequently exhibit user specification-grounding failures, executing hallucinated, undesired actions

model-releasesarxiv-cs-ai
12 Aug 2026
Model Releases

HoosierHelp: Benchmarking LLM Agents for Social Service Navigation

DGX agent

arXiv:2608.09946v1 Announce Type: cross Abstract: Social service navigation requires connecting help-seeking individuals to resources that satisfy their needs and specific constraints. Although LLM ag

model-releasesarxiv-cs-ai
12 Aug 2026
Hardware

TideRL: Boosting Agentic RL Goodput with Readiness-Aware Scheduling

DGX agent

arXiv:2608.10402v1 Announce Type: new Abstract: Reinforcement learning (RL) for large language models is moving toward multi-turn agentic workloads, where rollout tasks repeatedly pause for external e

hardwarearxiv-cs-lg
12 Aug 2026
Agents

Agentic AI-powered flexible fiber-bundle endoscopy for high-resolution NIR-II fluorescence imaging in vivo

DGX agent

arXiv:2608.08402v1 Announce Type: new Abstract: Fiber-bundle endoscopy offers a compact and flexible route for clinical fluorescence imaging through natural human orifices, but since its first report

agentsarxiv-cs-cv
11 Aug 2026
Safety

CyberAGENTS: Structured Autonomy for Agentic Gamified Learning in Cybersecurity

DGX agent

arXiv:2608.07965v1 Announce Type: new Abstract: Gamification is especially effective in learning domains requiring active problem-solving and iterative skill-building, such as cybersecurity education.

safetyarxiv-cs-ai
11 Aug 2026
Model Releases

Evo-Bench: Can Language Models Improve Agent Harness?

DGX agent

arXiv:2608.09096v1 Announce Type: new Abstract: Large Language Models (LLMs) have driven rapid progress in autonomous agents, yet standard evaluations remain confined to static task solving. An emergi

model-releasesarxiv-cs-cl
11 Aug 2026
Agents

Hierarchical Fast--Slow ReAct Agent for Zero-Shot Object-Goal Navigation

DGX agent

arXiv:2608.09816v1 Announce Type: new Abstract: Zero-shot object-goal navigation (ZSON) requires a robot to find a named object category in a building it has never entered. The prevailing approach sco

agentsarxiv-cs-ro
11 Aug 2026
Model Releases

MasDrift: Benchmarking Authorization Preservation Across Multi-Agent Architectures

DGX agent

arXiv:2608.07556v1 Announce Type: cross Abstract: Multi-agent systems (MAS) decompose long-horizon tasks across supervisors and subagents, but delegated goals do not necessarily carry their original a

model-releasesarxiv-cs-ai
11 Aug 2026
Safety

Multi-Agent AI Safety as an Institutional Design Problem

DGX agent

arXiv:2608.09828v1 Announce Type: cross Abstract: AI agents increasingly work inside systems that govern how they delegate tasks, move information, execute actions, and use shared resources. Recent wo

safetyarxiv-cs-ai
11 Aug 2026
Model Releases

OBLIVION: Workflow-Level Operational Skill Unlearning for Deployed Agents

DGX agent

arXiv:2608.08264v1 Announce Type: new Abstract: Large language model agents are becoming operational interfaces to files, memories, registries, and external tools. This deployment shift creates a new

model-releasesarxiv-cs-ai
11 Aug 2026
Agents

REVEAL: A Rubric-Guided Agent for Explicit Evidence Sufficiency Verificationin Long-Video Question Answering

DGX agent

arXiv:2608.08612v1 Announce Type: cross Abstract: Recently, retrieval-augmented and memory-augmented methods have emerged as two promising paradigms for long-video question answering. However, existin

agentsarxiv-cs-ai
11 Aug 2026
Model Releases

SuperLocalMemory 4.0: The Governed Memory Operating System for AI Agents

DGX agent

arXiv:2608.08253v1 Announce Type: new Abstract: AI agents are becoming shared infrastructure, yet durable memory is commonly assembled from separate retrieval, governance, and operational components.

model-releasesarxiv-cs-ai
11 Aug 2026
Safety

Gated-BEPO: Confidence-Gated Bellman Credit Assignment for Large Language Model Agents

DGX agent

arXiv:2608.06861v1 Announce Type: new Abstract: Training large language model agents in long-horizon environments requires assigning credit from sparse terminal outcomes to individual actions. Existin

safetyarxiv-cs-ai
10 Aug 2026
Research

PsychoAgent: An Affect-Sensitive Cognitive Architecture for Conflict-Aware Memory in LLM Agents

DGX agent

arXiv:2608.07438v1 Announce Type: new Abstract: Human-like cognition does not select past experience by topical similarity alone: affective significance and unresolved conflict also shape what becomes

researcharxiv-cs-ai
10 Aug 2026
Agents

Robot guide with multi-agent control and automatic scenario generation with LLM

DGX agent

arXiv:2509.10317v2 Announce Type: replace-cross Abstract: The article describes the development of a hybrid social robot control architecture to overcome the limitations of traditional approaches, whe

agentsarxiv-cs-lg
10 Aug 2026
Model Releases

Autonomous Research Agents: A Survey of AI Scientists and the Verification Gap

DGX agent

arXiv:2608.05179v1 Announce Type: cross Abstract: Large language model (LLM) agents are increasingly used across the scientific research lifecycle: ideation, literature search, experiment design and e

model-releasesarxiv-cs-ai
7 Aug 2026
Safety

DreamGuard: Efficient Runtime Guardrail for LLM Agents via Risk-Aware World Model

DGX agent

arXiv:2608.05695v1 Announce Type: new Abstract: As large language model (LLM) agents increasingly invoke external tools and interact with real-world systems, unsafe actions may cause irreversible cons

safetyarxiv-cs-ai
7 Aug 2026
Model Releases

EcoAgent-Bench: Evaluating Economic Decision-Making in Budget-Constrained LLM Agents

DGX agent

arXiv:2608.05519v1 Announce Type: new Abstract: Agent benchmarks usually measure task completion and treat resource use as an auxiliary statistic. In deployment, however, the choice among a local look

model-releasesarxiv-cs-ai
7 Aug 2026
Local Ai

From Passive Mirrors to Active Agents: Holonic Digital Twins for Physical AI over Networks

DGX agent

arXiv:2608.06227v1 Announce Type: cross Abstract: Despite advances in artificial intelligence (AI) across multiple sectors, today's AI tools, including deep learning and generative AI, still fail when

local-aiarxiv-cs-ai
7 Aug 2026
Model Releases

GrandCode: Achieving Grandmaster Level in Competitive Programming via Agentic Reinforcement Learning

DGX agent

arXiv:2604.02721v3 Announce Type: replace Abstract: Competitive programming remains one of the last few human strongholds in coding against AI. The best AI system to date still underperforms the best

model-releasesarxiv-cs-ai
7 Aug 2026
Agents

iARCS: Iterative Agentic RL for Controllable 3D Scene Generation

DGX agent

arXiv:2608.06161v1 Announce Type: new Abstract: Synthetic 3D scene generation is increasingly used as a data source for computer vision and embodied AI, but existing generators often optimize perceptu

agentsarxiv-cs-ai
7 Aug 2026
Model Releases

Matching Matters: A Fair Quality-Efficiency Benchmark for Command-Line Agents

DGX agent

arXiv:2606.21140v2 Announce Type: replace-cross Abstract: Rapid advances in large language models have improved the task-solving capabilities of command-line-interface (CLI)-based agents, whose CLIs d

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

SearchAuditor: Auditing and Attributing Failures in Long-Horizon Search Agents

DGX agent

arXiv:2608.05212v1 Announce Type: new Abstract: Deep search agents tackle challenging questions through long-horizon web interactions, a process that is both complex and fragile: small reasoning error

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

The Next Screenshot Knows: Gated Hindsight Distillation for Mobile GUI Agents

DGX agent

arXiv:2608.06065v1 Announce Type: new Abstract: GUI agents are commonly trained offline from successful interaction trajectories. Standard training decomposes each trajectory into prefix-action pairs:

model-releasesarxiv-cs-cv
7 Aug 2026
Model Releases

TRAJDEBUG: Tracing Error Lifecycle to Identify Critical Failures in Long-Horizon Agent Trajectories

DGX agent

arXiv:2608.06346v1 Announce Type: new Abstract: LLM-based agentic systems have shown remarkable capabilities in complex domains, while suffering from cascading errors and difficulty in debugging. Crit

model-releasesarxiv-cs-ai
7 Aug 2026
Agents

When Do Corrective Features Help? An Agent for Corrective Feature Discovery on Black-Box Forecasters

DGX agent

arXiv:2608.05207v1 Announce Type: new Abstract: Frozen pretrained forecasters often fail in structured, recurring ways that are costly to repair through fine-tuning. We study corrective feature discov

agentsarxiv-cs-lg
7 Aug 2026
Model Releases

Active-SWE: Benchmarking Coding Agents for Proactive Bug Fixing without Issue Reports

DGX agent

arXiv:2608.04682v1 Announce Type: cross Abstract: Coding agents powered by large language models (LLMs) are increasingly adopted in software engineering (SWE) scenarios, capable of fixing a specific b

model-releasesarxiv-cs-ai
6 Aug 2026
Agents

AutoProteinEngine: A Large Language Model Driven Agent Framework for Multimodal AutoML in Protein Engineering

DGX agent

arXiv:2411.04440v1 Announce Type: cross Abstract: Protein engineering is important for biomedical applications, but conventional approaches are often inefficient and resource-intensive. While deep lea

agentsarxiv-cs-ai
6 Aug 2026
Agents

Combating Knowledge Corruption in Agent Systems: A Byzantine-Tolerant Secure Collaborative RAG Framework

DGX agent

arXiv:2608.04366v1 Announce Type: cross Abstract: While retrieval-augmented generation systems partially address the hallucination issues in large language models, it also introduces new vulnerabiliti

agentsarxiv-cs-ai
6 Aug 2026
Model Releases

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables

DGX agent

arXiv:2608.04077v1 Announce Type: new Abstract: Evaluating financial AI agents requires criteria aligned with real professional work. Existing rubric methods typically derive criteria from task prompt

model-releasesarxiv-cs-ai
6 Aug 2026
Agents

Pun Intended: Multi-Agent Translation of Wordplay with Contrastive Learning and Phonetic-Semantic Embeddings for CLEF JOKER 2025 Task 2

DGX agent

arXiv:2507.06506v2 Announce Type: replace-cross Abstract: Translating wordplay across languages presents unique challenges that have long confounded both professional human translators and machine tra

agentsarxiv-cs-ai
6 Aug 2026
Model Releases

CUADebug: Diagnosing and Repairing Computer-Use Agent Failures

DGX agent

arXiv:2608.02643v1 Announce Type: cross Abstract: Computer-use agents (CUAs) operate real desktop and web interfaces through screenshots, mouse and keyboard actions, and stateful UI feedback, yet thei

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Cura 1T: Specialized Model for Agentic Healthcare

DGX agent

arXiv:2607.15314v2 Announce Type: replace Abstract: Healthcare AI agents handle patient consultation, clinical reasoning over text and images, interactive diagnosis, and electronic health record (EHR)

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

DiagChain: A Diagnostic Benchmark for Evaluating LLM Agents on Evidence-Grounded Attack Chain Reconstruction

DGX agent

arXiv:2608.03591v1 Announce Type: cross Abstract: Large Language Model (LLM) agents offer a promising approach to attack chain reconstruction by retrieving and interpreting heterogeneous telemetry to

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

From Social Coding to Agentic Coding: Productivity and Relational Reconfiguration in Open-Source Communities

DGX agent

arXiv:2608.03585v1 Announce Type: new Abstract: Open-source software communities are a form of digital public infrastructure that not only produces code, but also generates public knowledge and interp

model-releasesarxiv-cs-ai
5 Aug 2026
Agents

LACE: Large Language Model Aided Multi-Agent Framework for Agile RISC-V Instruction Extension

DGX agent

arXiv:2608.02915v1 Announce Type: cross Abstract: Domain-specific Instruction Set Architecture eXtensions (ISAX) are widely adopted in the RISC-V ecosystem to accelerate emerging workloads, but implem

agentsarxiv-cs-cl
5 Aug 2026
Model Releases

Screenshots or Tools? Eliciting Tool Use and Managing Multimodal Context in Hybrid GUI-MCP Computer-Use Agents

DGX agent

arXiv:2608.03327v1 Announce Type: new Abstract: Hybrid computer-use agents can act through screenshots or call text tools. We find that having a tool available does not settle which way the effect goe

model-releasesarxiv-cs-ai
5 Aug 2026
Agents

Towards Improving Sequential Decision-Making in LLM Agents via Experience Memory

DGX agent

arXiv:2608.03420v1 Announce Type: new Abstract: Large language models have improved substantially on single-shot reasoning tasks, but their performance in sequential decision-making is less well under

agentsarxiv-cs-ai
5 Aug 2026
Local Ai

Where Reasoning Diverges: Localized Multi-Agent Debate for Multi-Hop Question Answering

DGX agent

arXiv:2608.01463v2 Announce Type: replace Abstract: Multi-agent debate commonly exchanges complete rationales even when disagreements concern only a few intermediate claims. We introduce Localized Mul

local-aiarxiv-cs-ai
5 Aug 2026
Model Releases

Your Agentic LLMs Secretly Encode Latent Signals of Indirect Prompt-Injection Exposure

DGX agent

arXiv:2608.02657v1 Announce Type: cross Abstract: Agentic LLMs are vulnerable to indirect prompt injection (IPI) attacks, e.g., malicious side-tasks hidden in external tool results. While many efforts

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Cooperative Coevolution for Resource-Constrained Agentic LLM Post-Training

DGX agent

arXiv:2608.02391v1 Announce Type: cross Abstract: Tool-using large language model (LLM) agents produce long, multi-turn trajectories, making gradient-based post-training memory-intensive. Evolution st

model-releasesarxiv-cs-lg
4 Aug 2026
Agents

DeepVoyager-VL: Incentivizing Vision-in-the-Loop Search for Long-Horizon Multimodal Agents

DGX agent

arXiv:2608.01827v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have advanced visual understanding and reasoning, yet their static parametric knowledge limits their ability to

agentsarxiv-cs-cv
4 Aug 2026
Safety

Invisible Ink Threats: Adversarial Goals Behind Legitimate Tasks in Computer-Use Agents

DGX agent

arXiv:2608.02018v1 Announce Type: new Abstract: Computer-use agents (CUAs), which empower large language models to autonomously operate operating systems and the web, are increasingly vulnerable to in

safetyarxiv-cs-cv
4 Aug 2026
← Previous
1…7475767778…236
Next →