AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,153 results
Agents

ArenaRL: Scaling RL for Open-Ended Agents via Tournament-based Relative Ranking

DGX agent

arXiv:2601.06487v3 Announce Type: replace-cross Abstract: Reinforcement learning has substantially improved the performance of LLM agents on tasks with verifiable outcomes, but it still struggles on o

agentsarxiv-cs-ai
23 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

FORCE-Bench: A Benchmark, Dataset, and Evaluation Harness for Agentic AI in Enterprise Finance

DGX agent

arXiv:2607.19409v1 Announce Type: new Abstract: Recent advances in large language models have accelerated deployment of agentic systems in operational finance. Existing benchmarks emphasize measuring

model-releasesarxiv-cs-ai
23 Jul 2026
Safety

JANUS: Foreseeing Latent Risk for Long-Horizon Agent Safety

DGX agent

arXiv:2607.19913v1 Announce Type: new Abstract: Agent safety is moving from content moderation toward preventing operational failures before tool-using agents act. We propose Janus, a foresight-orient

safetyarxiv-cs-ai
23 Jul 2026
Agents

Automatic Ordinary Differential Equations Discovery For Biological Systems Using Large Language Model Powered Agentic System

DGX agent

arXiv:2607.13608v1 Announce Type: new Abstract: Automatic scientific discovery has long been a goal of computational scholars - a machine that can discover nature's secrets on its own, moving computat

agentsarxiv-cs-ai
16 Jul 2026
Model Releases

Baselines Before Architecture: Evaluating Coding Agents for Autonomous Penetration Testing

DGX agent

arXiv:2607.13085v1 Announce Type: cross Abstract: Recent autonomous penetration testing papers report high benchmark scores while adding multi-component security harnesses around frontier LLMs. Becaus

model-releasesarxiv-cs-ai
16 Jul 2026
Agents

Benefits and Limitations of Communication in Multi-Agent Reasoning

DGX agent

arXiv:2510.13903v2 Announce Type: replace-cross Abstract: Chain-of-thought prompting has popularized step-by-step reasoning in large language models, yet model performance still degrades as problem co

agentsarxiv-cs-ai
16 Jul 2026
Safety

Mind the Gap: Action Rebinding Attacks against Android GUI Agents

DGX agent

arXiv:2601.12349v3 Announce Type: replace-cross Abstract: Large multimodal model powered GUI agents are emerging as high-privilege operators on mobile platforms, entrusted to perceive screen content a

safetyarxiv-cs-ai
16 Jul 2026
Agents

Networked Intelligence: Active Shared Context Graphs for Human-AI Team Science

DGX agent

arXiv:2607.13220v1 Announce Type: new Abstract: Most AI-for-science systems focus on scaling a single reasoning process through better models, larger context windows, long-horizon agentic execution, o

agentsarxiv-cs-ai
16 Jul 2026
Agents

Social Simulations: from Agent-Based Modeling to Digital Twins

DGX agent

arXiv:2607.13693v1 Announce Type: cross Abstract: This book chapter covers the evolution of social simulation from classical agent-based models, in which agents interact according to explicitly define

agentsarxiv-cs-ai
16 Jul 2026
Model Releases

Declarative by Design, Assistable Only by Convention: Benchmarking Multi-Agent Frameworks for AI-Assistability

DGX agent

arXiv:2602.11198v2 Announce Type: replace-cross Abstract: Multi-agent frameworks (MAFs) promise to simplify LLM-driven software development, yet no principled metric captures how well AI coding assist

model-releasesarxiv-cs-ai
15 Jul 2026
Agents

Fine-Tuned Multi-Agent Framework for Detecting OCEAN in Life Narratives

DGX agent

arXiv:2607.12215v1 Announce Type: new Abstract: Accurately assessing personality from text is challenging because traits are latent, context-dependent, and often subtly expressed across long narrative

agentsarxiv-cs-cl
15 Jul 2026
Agents

ReflectVLN: Training Vision-Language Navigation Agents with Reflective Reasoning

DGX agent

arXiv:2607.12680v1 Announce Type: new Abstract: Existing vision-language navigation methods often couple a VLM with waypoint decoders to produce multi-step action plans, but they typically lack an exp

agentsarxiv-cs-cv
15 Jul 2026
Model Releases

SheetMind: An End-to-End LLM-Powered Multi-Agent Framework for Spreadsheet Automation

DGX agent

arXiv:2506.12339v2 Announce Type: replace-cross Abstract: We present SheetMind, a modular multi-agent framework powered by large language models (LLMs) for spreadsheet automation via natural language

model-releasesarxiv-cs-ai
15 Jul 2026
Agents

A Transdiagnostic Space of Disorder Like Phenotypes in Reinforcement Learning Agents

DGX agent

arXiv:2607.07753v1 Announce Type: cross Abstract: Modelling psychological disorders in artificial agents offers both a testbed for computational psychiatry and a lens on the failure modes of affective

agentsarxiv-cs-ai
10 Jul 2026
Agents

Agentic AI and Retrieval-Augmented Models in Straight-Through Underwriting

DGX agent

arXiv:2607.07858v1 Announce Type: new Abstract: Artificial intelligence (AI) is beginning to reshape actuarial practice, particularly in domains that require reasoning over unstructured documents, het

agentsarxiv-cs-ai
10 Jul 2026
Agents

SimRPD: Optimizing Recruitment Proactive Dialogue Agents through Simulator-Based Data Evaluation and Selection

DGX agent

arXiv:2601.02871v3 Announce Type: replace Abstract: Task-oriented proactive dialogue agents play a pivotal role in recruitment, particularly for steering conversations towards specific business outcom

agentsarxiv-cs-ai
10 Jul 2026
Model Releases

SolarChain-Eval: A Physics-Constrained Benchmark for Trustworthy Economic Agents in Decentralized Energy Markets

DGX agent

arXiv:2607.08681v1 Announce Type: new Abstract: As agentic AI systems are increasingly applied to cyber-physical environments, their evaluation requires assessment of both task performance and trustwo

model-releasesarxiv-cs-ai
10 Jul 2026
Local Ai

TRACE: A Two-Channel Robust Attribution Watermark via Complementary Embeddings for LLM-Agent Trajectories

DGX agent

arXiv:2607.08400v1 Announce Type: cross Abstract: LLM agents reach users through resellers, who may rebrand a developer's agent or substitute a cheaper model. When provenance is disputed, attribution

local-aiarxiv-cs-ai
10 Jul 2026
Local Ai

WebSwarm: Recursive Multi-Agent Orchestration for Deep-and-Wide Web Search

DGX agent

arXiv:2607.08662v1 Announce Type: cross Abstract: Large language model (LLM)-based web search agents are transforming information seeking from simple factoid question answering into complex, deep-and-

local-aiarxiv-cs-ai
10 Jul 2026
Agents

ContestTrade: A Multi-Agent Trading System Based on Internal Contest Mechanism

DGX agent

arXiv:2508.00554v4 Announce Type: replace-cross Abstract: In financial trading, large language model (LLM)-based agents demonstrate significant potential, but their decisions can be sensitive to noisy

agentsarxiv-cs-cl
9 Jul 2026
Agents

From Agentic to Autogenic Network Management for AI-Native 6G and Beyond: A Standards Perspective

DGX agent

arXiv:2607.06786v1 Announce Type: cross Abstract: Standards bodies, including TM Forum, 3GPP, and ETSI, are converging on Agentic AI as the foundation for next-generation network management, where Lar

agentsarxiv-cs-ai
9 Jul 2026
Agents

From Noisy Traces to Root Causes: Structural Trajectory Analysis and Causal Extraction for Agent Optimization

DGX agent

arXiv:2607.07702v1 Announce Type: new Abstract: The optimization of long-horizon agents increasingly relies on reflection-based mechanisms, where a large language model (LLM) acts as an optimizer to d

agentsarxiv-cs-cl
9 Jul 2026
Agents

AgoraSim: A Hybrid Agent-Based Modeling Framework

DGX agent

arXiv:2607.05999v1 Announce Type: new Abstract: LLM-agent simulations make natural-language social scenarios easy to instantiate, but their outputs can be overread as predictions and are often difficu

agentsarxiv-cs-ai
8 Jul 2026
Agents

CurateEvo: Data-Curation Evolving for Agentic Post-Training

DGX agent

arXiv:2607.06140v1 Announce Type: new Abstract: Large language model (LLM) agents require post-training methods that can improve long-horizon decision making from environment feedback. However, existi

agentsarxiv-cs-cl
8 Jul 2026
Agents

Proof of Execution: Runtime Verification for Governed AI Agent Actions

DGX agent

arXiv:2607.05397v1 Announce Type: cross Abstract: Agent systems increasingly execute rather than advise. When an AI agent queries regulated data, invokes effectful tools, and mutates persistent state,

agentsarxiv-cs-ai
8 Jul 2026
Agents

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI

DGX agent

arXiv:2607.04426v1 Announce Type: new Abstract: Embodied AI is moving from isolated perception or action modules toward physical agents that understand, plan under goals, act through robot bodies, mon

agentsarxiv-cs-ro
7 Jul 2026
Model Releases

AgentGym2: Benchmarking Large Language Model Agents in De-Idealized Real-World Environments

DGX agent

arXiv:2607.05174v1 Announce Type: new Abstract: Language agents, i.e., LLM agents, progress rapidly and are increasingly deployed in production environments. This trend underscores the urgent need for

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

APeB: Benchmarking Personalization Ability of Large Language Model Agents

DGX agent

arXiv:2607.03162v1 Announce Type: new Abstract: LLM-powered agents struggle with personalization when users issue raw, underspecified queries. In this setting, agents must infer latent intent, extract

model-releasesarxiv-cs-ai
7 Jul 2026
Safety

Attributing Emergence in Million-Agent Systems

DGX agent

arXiv:2605.11404v2 Announce Type: replace Abstract: Large language models (LLMs) can simulate human-like reasoning and decision-making in individual agents. LLM-powered multi-agent systems (MAS) combi

safetyarxiv-cs-ai
7 Jul 2026
Agents

CogRad: A Cognitively-Inspired Multi-Agent Framework for Radiology Report Generation

DGX agent

arXiv:2607.03853v1 Announce Type: new Abstract: Automated radiology report generation (RRG) can ease radiologist workload, yet most existing systems produce a report in a single forward pass, with no

agentsarxiv-cs-cv
7 Jul 2026
Agents

CompactionRL: Reinforcement Learning with Context Compaction for Long-Horizon Agents

DGX agent

arXiv:2607.05378v1 Announce Type: new Abstract: Long-horizon agentic LLMs are increasingly limited by finite context windows, as extended interaction trajectories can exceed the maximum context length

agentsarxiv-cs-lg
7 Jul 2026
Agents

Do GUI Agents Believe Their Eyes? Diagnosing State-Belief Reliance on Pixels versus Structure

DGX agent

arXiv:2607.04334v1 Announce Type: new Abstract: Multimodal GUI agents read an interface through two redundant channels: the rendered pixels of a screenshot and a serialized structure such as a DOM or

agentsarxiv-cs-ai
7 Jul 2026
Agents

Evaluating Generative Agents with Actions Grounded in Socially Distributed Task Environments using Incognita

DGX agent

arXiv:2607.02975v1 Announce Type: new Abstract: Effective agency in social environments depends on when an agent seeks knowledge, when it acts, and whether its actions are justified by acquired inform

agentsarxiv-cs-ai
7 Jul 2026
Model Releases

Governed Individuation: Cryptographically Decoupling an Agent's Learning from Its Authority

DGX agent

arXiv:2607.04613v1 Announce Type: new Abstract: Autonomous agents are moving from sandboxed text generators to operators of code, data, and physical infrastructure, and they increasingly learn while d

model-releasesarxiv-cs-ai
7 Jul 2026
Agents

MRMS: A Multi-Resolution Memory Substrate for Long-Lived AI Agents

DGX agent

arXiv:2607.04617v1 Announce Type: new Abstract: Long-lived AI agents require continuity across interactions, but continuity cannot be obtained by simply extending the prompt window. An agent must pres

agentsarxiv-cs-ai
7 Jul 2026
Agents

Rethinking Scientific Discovery in an Agentic Era

DGX agent

arXiv:2607.03863v1 Announce Type: new Abstract: Artificial intelligence has advanced scientific discovery, but most AI4Science systems remain fragmented tools that rely on humans to coordinate problem

agentsarxiv-cs-cl
7 Jul 2026
Agents

SelfMem: Self-Optimizing Memory for AI Agents

DGX agent

arXiv:2607.03726v1 Announce Type: new Abstract: While current AI agents support increasingly long context windows, tool use, and skill execution for long-horizon tasks, they still require memory syste

agentsarxiv-cs-cl
7 Jul 2026
Model Releases

SovereignNegotiation-Bench: Evaluating User-Owned Personal Agents In Delegated Bargaining Under Privacy, Consent, Evidence, And Institutional Pressure

DGX agent

arXiv:2607.02814v1 Announce Type: cross Abstract: Personal agents will increasingly negotiate on behalf of users: splitting costs with other personal agents, appealing platform decisions, escalating s

model-releasesarxiv-cs-ai
7 Jul 2026
Safety

The agent creates, we validate: A Lightweight Framework for Agentic Artifact Generation

DGX agent

arXiv:2607.02615v1 Announce Type: cross Abstract: Generating structured artifacts with Large Language Models - e.g. database queries, threat framework mappings, entity schemas - is relatively straight

safetyarxiv-cs-ai
7 Jul 2026
Agents

Untrusted Content Masking for Web Agents with Security Guarantees

DGX agent

arXiv:2607.05277v1 Announce Type: cross Abstract: Defenses that provide security guarantees against prompt injection attacks rely on strict isolation between trusted instructions and untrusted data. I

agentsarxiv-cs-lg
7 Jul 2026
Model Releases

When Claws Remember but Do Not Tell: Stealthy Memory Injection in Persistent Personal Agents

DGX agent

arXiv:2607.05189v1 Announce Type: cross Abstract: Persistent personal agents combine long-term memory with access to users' external environments, enabling personalized foreground assistance and proac

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

GroundEval: A Deterministic Replacement for LLM-as-Judge in Stateful Agent Evaluation

DGX agent

arXiv:2606.22737v2 Announce Type: replace Abstract: Before letting an agent operate over real context, can you prove it used the right evidence? GroundEval turns that question into a deterministic tes

model-releasesarxiv-cs-ai
3 Jul 2026
Agents

SkillCoach: Self-Evolving Rubrics for Evaluating and Enhancing Agentic Skill-Use

DGX agent

arXiv:2607.01874v1 Announce Type: new Abstract: Skills are becoming a reusable operational layer for LLM agents, encoding SOPs, domain rules, tool workflows, scripts, and validation routines. In reali

agentsarxiv-cs-ai
3 Jul 2026
Agents

SlowBA: An efficiency backdoor attack towards VLM-based GUI agents

DGX agent

arXiv:2603.08316v3 Announce Type: replace-cross Abstract: Modern vision-language-model (VLM) based graphical user interface (GUI) agents are expected not only to execute actions accurately but also to

agentsarxiv-cs-cl
2 Jul 2026
Model Releases

Beyond expert users: agents should help users construct preferences, not just elicit them

DGX agent

arXiv:2606.30863v1 Announce Type: new Abstract: Agents typically assume an expert user -- one with well-formed preferences about what they want -- and default to clarifying questions whenever the task

model-releasesarxiv-cs-ai
1 Jul 2026
Model Releases

Agent-Computer Observation Interfaces Enable Dynamic Computer Use

DGX agent

arXiv:2606.29472v1 Announce Type: new Abstract: SWE-agent established the action interface as an underexplored design axis for software-engineering agents; we make the analogous case for the observati

model-releasesarxiv-cs-ai
30 Jun 2026
Agents

Agentic AI for ISAC: Analysis, Framework, and Case Study

DGX agent

arXiv:2512.15044v2 Announce Type: replace Abstract: Integrated sensing and communication (ISAC) has emerged as a key development direction in the sixth-generation (6G) era, which provides essential su

agentsarxiv-cs-ai
30 Jun 2026
Agents

Capability Gates Are Not Authorization: Confused-Deputy Failures in LLM Agent Frameworks

DGX agent

arXiv:2606.28679v1 Announce Type: cross Abstract: Tool-using LLM agents increasingly read untrusted content while holding side-effecting tools such as payments, email, CRM, and infrastructure APIs, ye

agentsarxiv-cs-ai
30 Jun 2026
← Previous
1…3132333435…233
Next →