AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,040 results
Agents

MAGIC: Multi-Step Advantage-Gated Causal Influence for Multi-agent Reinforcement Learning

DGX agent

arXiv:2605.01805v1 Announce Type: cross Abstract: A key challenge in multi-agent reinforcement learning (MARL) lies in designing learning signals that effectively promote coordination among agents. De

agentsarxiv-cs-lg
5 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

A Survey of Safe Reinforcement Learning and Constrained MDPs: A Technical Survey on Single-Agent and Multi-Agent Safety

DGX agent

arXiv:2505.17342v2 Announce Type: replace Abstract: Safe Reinforcement Learning (SafeRL) is the subfield of reinforcement learning that explicitly deals with safety constraints during the learning and

safetyarxiv-cs-lg
30 Apr 2026
Agents

FGDM: Reasoning Aware Multi-Agentic Framework for Software Bug Detection using Chain of Thought and Tree of Thought Prompting

DGX agent

arXiv:2604.24831v1 Announce Type: cross Abstract: Deep Learning methods are becoming prominent in automated software bug detection; however, they lack the global understanding of the given code. Conse

agentsarxiv-cs-lg
29 Apr 2026
Local Ai

Code Broker: A Multi-Agent System for Automated Code Quality Assessment

DGX agent

arXiv:2604.23088v1 Announce Type: cross Abstract: We present Code Broker, a multi agent system built with Google Agent Development Kit ADK that analyses Python code from files, local directories, or G

local-aiarxiv-cs-ai
28 Apr 2026
Safety

DLM: Unified Decision Language Models for Offline Multi-Agent Sequential Decision Making

DGX agent

arXiv:2604.23557v1 Announce Type: cross Abstract: Building scalable and reusable multi-agent decision policies from offline datasets remains a challenge in offline multi-agent reinforcement learning (

safetyarxiv-cs-ai
28 Apr 2026
Safety

FastOMOP: A Foundational Architecture for Reliable Agentic Real-World Evidence Generation on OMOP CDM data

DGX agent

arXiv:2604.24572v1 Announce Type: new Abstract: The Observational Medical Outcomes Partnership Common Data Model (OMOP CDM), maintained by the Observational Health Data Sciences and Informatics (OHDSI

safetyarxiv-cs-ai
28 Apr 2026
Model Releases

SoccerRef-Agents: Multi-Agent System for Automated Soccer Refereeing

DGX agent

arXiv:2604.23392v1 Announce Type: new Abstract: Refereeing is vital in sports, where fair, accurate, and explainable decisions are fundamental. While intelligent assistant technologies are being widel

model-releasesarxiv-cs-ai
28 Apr 2026
Agents

HACHIMI: Scalable and Controllable Student Persona Generation via Orchestrated Agents

DGX agent

arXiv:2603.04855v3 Announce Type: replace Abstract: Student Personas (SPs) are emerging as infrastructure for educational LLMs, yet prior work often relies on ad-hoc prompting or hand-crafted profiles

agentsarxiv-cs-cl
27 Apr 2026
Agents

HiCrew: Hierarchical Reasoning for Long-Form Video Understanding via Question-Aware Multi-Agent Collaboration

DGX agent

arXiv:2604.21444v1 Announce Type: new Abstract: Long-form video understanding remains fundamentally challenged by pervasive spatiotemporal redundancy and intricate narrative dependencies that span ext

agentsarxiv-cs-ai
24 Apr 2026
Agents

The AI Criminal Mastermind

DGX agent

arXiv:2604.20868v1 Announce Type: cross Abstract: In this paper, I evaluate the risks of an AI criminal mastermind, an AI agent capable of planning, coordinating, and committing a crime through the on

agentsarxiv-cs-ai
24 Apr 2026
Agents

TriEx: A Game-based Tri-View Framework for Explaining Internal Reasoning in Multi-Agent LLMs

DGX agent

arXiv:2604.20043v1 Announce Type: cross Abstract: Explainability for Large Language Model (LLM) agents is especially challenging in interactive, partially observable settings, where decisions depend o

agentsarxiv-cs-ai
23 Apr 2026
Agents

A Multi-Agent Framework with Structured Reasoning and Reflective Refinement for Multimodal Empathetic Response Generation

DGX agent

arXiv:2604.18988v1 Announce Type: new Abstract: Multimodal empathetic response generation (MERG) aims to generate emotionally engaging and empathetic responses based on users' multimodal contexts. Exi

agentsarxiv-cs-cv
22 Apr 2026
Agents

Debating the Unspoken: Role-Anchored Multi-Agent Reasoning for Half-Truth Detection

DGX agent

arXiv:2604.19005v1 Announce Type: new Abstract: Half-truths, claims that are factually correct yet misleading due to omitted context, remain a blind spot for fact verification systems focused on expli

agentsarxiv-cs-cl
22 Apr 2026
Agents

From Craft to Kernel: A Governance-First Execution Architecture and Semantic ISA for Agentic Computers

DGX agent

arXiv:2604.18652v1 Announce Type: cross Abstract: The transition of agentic AI from brittle prototypes to production systems is stalled by a pervasive crisis of craft. We suggest that the prevailing o

agentsarxiv-cs-ai
22 Apr 2026
Model Releases

Towards Optimal Agentic Architectures for Offensive Security Tasks

DGX agent

arXiv:2604.18718v1 Announce Type: cross Abstract: Agentic security systems increasingly audit live targets with tool-using LLMs, but prior systems fix a single coordination topology, leaving unclear w

model-releasesarxiv-cs-ai
22 Apr 2026
Agents

Designing Explainable Conversational Agentic Systems for Guarani Speakers

DGX agent

arXiv:2603.05743v3 Announce Type: replace Abstract: Although artificial intelligence (AI) and Human-Computer Interaction (HCI) systems are often presented as universal solutions, their design remains

agentsarxiv-cs-cl
21 Apr 2026
Safety

End-to-End Optimization of LLM-Driven Multi-Agent Search Systems via Heterogeneous-Group-Based Reinforcement Learning

DGX agent

arXiv:2506.02718v2 Announce Type: replace Abstract: Large language models (LLMs) are versatile, yet their deployment in complex real-world settings is limited by static knowledge cutoffs and the diffi

safetyarxiv-cs-lg
21 Apr 2026
Local Ai

The Consensus Trap: Rescuing Multi-Agent LLMs from Adversarial Majorities via Token-Level Collaboration

DGX agent

arXiv:2604.17139v1 Announce Type: new Abstract: Multi-agent large language model (LLM) architectures increasingly rely on response-level aggregation, such as Majority Voting (MAJ), to raise reasoning

local-aiarxiv-cs-cl
21 Apr 2026
Agents

Discover and Prove: An Open-source Agentic Framework for Hard Mode Automated Theorem Proving in Lean 4

DGX agent

arXiv:2604.15839v1 Announce Type: new Abstract: Most ATP benchmarks embed the final answer within the formal statement -- a convention we call 'Easy Mode' -- a design that simplifies the task relative

agentsarxiv-cs-ai
20 Apr 2026
Agents

Eco-Bee: A Personalised Multi-Modal Agent for Advancing Student Climate Awareness and Sustainable Behaviour in Campus Ecosystems

DGX agent

arXiv:2604.15327v1 Announce Type: cross Abstract: Universities are microcosms of urban ecosystems, with concentrated consumption patterns in food, transport, energy, and product usage. These environme

agentsarxiv-cs-ai
20 Apr 2026
Model Releases

FS-Researcher: Test-Time Scaling for Long-Horizon Research Tasks with File-System-Based Agents

DGX agent

arXiv:2602.01566v2 Announce Type: replace Abstract: Deep research is emerging as a representative long-horizon task for large language model (LLM) agents. However, long trajectories in deep research o

model-releasesarxiv-cs-cl
20 Apr 2026
Agents

Just Type It in Isabelle! AI Agents Drafting, Mechanizing, and Generalizing from Human Hints

DGX agent

arXiv:2604.15713v1 Announce Type: cross Abstract: Type annotations are essential when printing terms in a way that preserves their meaning under reparsing and type inference. We study the problem of c

agentsarxiv-cs-ai
20 Apr 2026
Safety

Layered Mutability: Continuity and Governance in Persistent Self-Modifying Agents

DGX agent

arXiv:2604.14717v1 Announce Type: cross Abstract: Persistent language-model agents increasingly combine tool use, tiered memory, reflective prompting, and runtime adaptation. In such systems, behavior

safetyarxiv-cs-lg
17 Apr 2026
Agents

SciFi: A Safe, Lightweight, User-Friendly, and Fully Autonomous Agentic AI Workflow for Scientific Applications

DGX agent

arXiv:2604.13180v1 Announce Type: new Abstract: Recent advances in agentic AI have enabled increasingly autonomous workflows, but existing systems still face substantial challenges in achieving reliab

agentsarxiv-cs-ai
17 Apr 2026
Agents

A Multi-Agent Feedback System for Detecting and Describing News Events in Satellite Imagery

DGX agent

arXiv:2604.12772v1 Announce Type: new Abstract: Changes in satellite imagery often occur over multiple time steps. Despite the emergence of bi-temporal change captioning datasets, there is a lack of m

agentsarxiv-cs-cv
15 Apr 2026
Agents

LIFE -- an energy efficient advanced continual learning agentic AI framework for frontier systems

DGX agent

arXiv:2604.12874v1 Announce Type: new Abstract: The rapid advancement of AI has changed the character of HPC usage such as dimensioning, provisioning, and execution. Not only has energy demand been am

agentsarxiv-cs-ai
15 Apr 2026
Agents

Mathematics Teachers Interactions with a Multi-Agent System for Personalized Problem Generation

DGX agent

arXiv:2604.12066v1 Announce Type: new Abstract: Large language models can increasingly adapt educational tasks to learners characteristics. In the present study, we examine a multi-agent teacher-in-th

agentsarxiv-cs-ai
15 Apr 2026
Model Releases

Mobile GUI Agents under Real-world Threats: Are We There Yet?

DGX agent

arXiv:2507.04227v2 Announce Type: replace-cross Abstract: Recent years have witnessed a rapid development of mobile GUI agents powered by large language models (LLMs), which can autonomously execute d

model-releasesarxiv-cs-ai
15 Apr 2026
Agents

A collaborative agent with two lightweight synergistic models for autonomous crystal materials research

DGX agent

arXiv:2604.11540v1 Announce Type: new Abstract: Current large language models require hundreds of billions of parameters yet struggle with domain-specific reasoning and tool coordination in materials

agentsarxiv-cs-ai
14 Apr 2026
Agents

Beyond Offline A/B Testing: Context-Aware Agent Simulation for Recommender System Evaluation

DGX agent

arXiv:2604.09549v1 Announce Type: cross Abstract: Recommender systems are central to online services, enabling users to navigate through massive amounts of content across various domains. However, the

agentsarxiv-cs-ai
14 Apr 2026
Agents

Beyond RAG for Agent Memory: Retrieval by Decoupling and Aggregation

DGX agent

arXiv:2602.02007v3 Announce Type: replace-cross Abstract: Agent memory systems often adopt the standard Retrieval-Augmented Generation (RAG) pipeline, yet its underlying assumptions differ in this set

agentsarxiv-cs-ai
14 Apr 2026
Agents

Diagnosing Retrieval vs. Utilization Bottlenecks in LLM Agent Memory

DGX agent

arXiv:2603.02473v2 Announce Type: replace Abstract: Memory-augmented LLM agents store and retrieve information from prior interactions, yet the relative importance of how memories are written versus h

agentsarxiv-cs-ai
14 Apr 2026
Local Ai

Inside the Scaffold: A Source-Code Taxonomy of Coding Agent Architectures

DGX agent

arXiv:2604.03515v2 Announce Type: replace-cross Abstract: LLM-based coding agents can localize bugs, generate patches, and run tests with diminishing human oversight, yet the scaffolding code that sur

local-aiarxiv-cs-ai
14 Apr 2026
Agents

Problem Reductions at Scale: Agentic Integration of Computationally Hard Problems

DGX agent

arXiv:2604.11535v1 Announce Type: new Abstract: Solving an NP-hard optimization problem often requires reformulating it for a specific solver -- quantum hardware, a commercial optimizer, or a domain h

agentsarxiv-cs-ai
14 Apr 2026
Safety

Resilient Write: A Six-Layer Durable Write Surface for LLM Coding Agents

DGX agent

arXiv:2604.10842v1 Announce Type: cross Abstract: LLM-powered coding agents increasingly rely on tool-use protocols such as the Model Context Protocol~(MCP) to read and write files on a developer's wo

safetyarxiv-cs-ai
14 Apr 2026
Model Releases

WebForge: Breaking the Realism-Reproducibility-Scalability Trilemma in Browser Agent Benchmark

DGX agent

arXiv:2604.10988v1 Announce Type: new Abstract: Existing browser agent benchmarks face a fundamental trilemma: real-website benchmarks lack reproducibility due to content drift, controlled environment

model-releasesarxiv-cs-ai
14 Apr 2026
Agents

EpiAgent: An Agent-Centric System for Ancient Inscription Restoration

DGX agent

arXiv:2604.09367v1 Announce Type: new Abstract: Ancient inscriptions, as repositories of cultural memory, have suffered from centuries of environmental and human-induced degradation. Restoring their i

agentsarxiv-cs-cv
13 Apr 2026
Model Releases

Semantic Rate-Distortion for Bounded Multi-Agent Communication: Capacity-Derived Semantic Spaces and the Communication Cost of Alignment

DGX agent

arXiv:2604.09521v1 Announce Type: cross Abstract: When two agents of different computational capacities interact with the same environment, they need not compress a common semantic alphabet differentl

model-releasesarxiv-cs-ai
13 Apr 2026
Agents

SkillForge: Forging Domain-Specific, Self-Evolving Agent Skills in Cloud Technical Support

DGX agent

arXiv:2604.08618v1 Announce Type: cross Abstract: Deploying LLM-powered agents in enterprise scenarios such as cloud technical support demands high-quality, domain-specific skills. However, existing s

agentsarxiv-cs-ai
13 Apr 2026
Agents

Towards Context-Aware Image Anonymization with Multi-Agent Reasoning

DGX agent

arXiv:2603.27817v3 Announce Type: replace-cross Abstract: Street-level imagery contains personally identifiable information (PII), some of which is context-dependent. Existing anonymization methods ei

agentsarxiv-cs-ai
13 Apr 2026
Model Releases

When Identity Skews Debate: Anonymization for Bias-Reduced Multi-Agent Reasoning

DGX agent

arXiv:2510.07517v5 Announce Type: replace Abstract: Multi-agent debate (MAD) aims to improve large language model (LLM) reasoning by letting multiple agents exchange answers and then aggregate their o

model-releasesarxiv-cs-ai
13 Apr 2026
Agents

ORACLE-SWE: Quantifying the Contribution of Oracle Information Signals on SWE Agents

DGX agent

arXiv:2604.07789v1 Announce Type: cross Abstract: Recent advances in language model (LM) agents have significantly improved automated software engineering (SWE). Prior work has proposed various agenti

agentsarxiv-cs-cl
10 Apr 2026
Agents

RAGEN-2: Reasoning Collapse in Agentic RL

DGX agent

arXiv:2604.06268v1 Announce Type: new Abstract: RL training of multi-turn LLM agents is inherently unstable, and reasoning quality directly determines task performance. Entropy is widely used to track

agentsarxiv-cs-lg
10 Apr 2026
Agents

Robust Multi-Agent Target Tracking in Intermittent Communication Environments via Analytical Belief Merging

DGX agent

arXiv:2604.07575v1 Announce Type: new Abstract: Autonomous multi-agent target tracking in GPS-denied and communication-restricted environments (e.g., underwater exploration, subterranean search and re

agentsarxiv-cs-ro
10 Apr 2026
Model Releases

MAP-Graph: Provenance-Aware Shared Memory for Multi-Agent Workflows

DGX agent

arXiv:2608.10509v1 Announce Type: new Abstract: Shared memory helps language-model agents reuse information across long workflows, yet relevant evidence may not be admissible for a particular agent or

model-releasesarxiv-cs-ai
12 Aug 2026
Agents

ElasticBack: Stealthy Conditional Backdoor in LLM-Agent Skills via Coupled Trigger-Rule Optimization

DGX agent

arXiv:2608.09577v1 Announce Type: new Abstract: Agent skills, bundles of instructions and resources that an LLM agent loads on demand, form an emerging supply chain where a single poisoned skill can p

agentsarxiv-cs-ai
11 Aug 2026
Model Releases

From Product Search to Preference Articulation: The Economics of Agentic Commerce

DGX agent

arXiv:2608.08395v1 Announce Type: cross Abstract: Generative AI is shifting digital commerce from browsing toward agentic search, in which consumers delegate product discovery to AI agents. We compare

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Social Gym and SPaRTan: Benchmarking and Improving LLM Social Reasoning via Multi-Agent Game Tournaments

DGX agent

arXiv:2608.09128v1 Announce Type: cross Abstract: LLM agents are increasingly deployed in multi-agent social settings where they must cooperate, negotiate, and adapt to other agents. Measuring and imp

model-releasesarxiv-cs-ai
11 Aug 2026
← Previous
1…1920212223…230
Next →