AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,289 results
Safety

From Risk to Rescue: An Agentic Survival Analysis Framework for Liquidation Prevention

DGX agent

arXiv:2604.14583v1 Announce Type: new Abstract: Decentralized Finance (DeFi) lending protocols like Aave v3 rely on over-collateralization to secure loans, yet users frequently face liquidation due to

safetyarxiv-cs-lg
17 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

Hierarchical Reinforcement Learning with Augmented Step-Level Transitions for LLM Agents

DGX agent

arXiv:2604.05808v2 Announce Type: replace-cross Abstract: Large language model (LLM) agents have demonstrated strong capabilities in complex interactive decision-making tasks. However, existing LLM ag

researcharxiv-cs-lg
16 Apr 2026
Hardware

SemiFA: An Agentic Multi-Modal Framework for Autonomous Semiconductor Failure Analysis Report Generation

DGX agent

arXiv:2604.13236v1 Announce Type: new Abstract: Semiconductor failure analysis (FA) requires engineers to examine inspection images, correlate equipment telemetry, consult historical defect records, a

hardwarearxiv-cs-cv
16 Apr 2026
Model Releases

The cognitive companion: a lightweight parallel monitoring architecture for detecting and recovering from reasoning degradation in LLM agents

DGX agent

arXiv:2604.13759v1 Announce Type: cross Abstract: Large language model (LLM) agents on multi-step tasks suffer reasoning degradation, looping, drift, stuck states, at rates up to 30% on hard tasks. Cu

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

ToolOmni: Enabling Open-World Tool Use via Agentic learning with Proactive Retrieval and Grounded Execution

DGX agent

arXiv:2604.13787v1 Announce Type: new Abstract: Large Language Models (LLMs) enhance their problem-solving capability by utilizing external tools. However, in open-world scenarios with massive and evo

model-releasesarxiv-cs-cl
16 Apr 2026
Agents

A hierarchical spatial-aware algorithm with efficient reinforcement learning for human-robot task planning and allocation in production

DGX agent

arXiv:2604.12669v1 Announce Type: new Abstract: In advanced manufacturing systems, humans and robots collaborate to conduct the production process. Effective task planning and allocation (TPA) is cruc

agentsarxiv-cs-ai
15 Apr 2026
Local Ai

Local-Splitter: A Measurement Study of Seven Tactics for Reducing Cloud LLM Token Usage on Coding-Agent Workloads

DGX agent

arXiv:2604.12301v1 Announce Type: cross Abstract: We present a systematic measurement study of seven tactics for reducing cloud LLM token usage when a small local model can act as a triage layer in fr

local-aiarxiv-cs-ai
15 Apr 2026
Model Releases

QuarkMedSearch: A Long-Horizon Deep Search Agent for Exploring Medical Intelligence

DGX agent

arXiv:2604.12867v1 Announce Type: new Abstract: As agentic foundation models continue to evolve, how to further improve their performance in vertical domains has become an important challenge. To this

model-releasesarxiv-cs-ai
15 Apr 2026
Local Ai

Toward Efficient and Robust Behavior Models for Multi-Agent Driving Simulation

DGX agent

arXiv:2512.05812v5 Announce Type: replace-cross Abstract: Scalable multi-agent driving simulation requires behavior models that are both realistic and computationally efficient. We address this by opt

local-aiarxiv-cs-cv
15 Apr 2026
Model Releases

When Reasoning Models Hurt Behavioral Simulation: A Solver-Sampler Mismatch in Multi-Agent LLM Negotiation

DGX agent

arXiv:2604.11840v1 Announce Type: new Abstract: Large language models are increasingly used as agents in social, economic, and policy simulations. A common assumption is that stronger reasoning should

model-releasesarxiv-cs-lg
15 Apr 2026
Model Releases

Agentic Exploration of PDE Spaces using Latent Foundation Models for Parameterized Simulations

DGX agent

arXiv:2604.09584v1 Announce Type: new Abstract: Flow physics and more broadly physical phenomena governed by partial differential equations (PDEs), are inherently continuous, high-dimensional and ofte

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

From Perception to Autonomous Computational Modeling: A Multi-Agent Approach

DGX agent

arXiv:2604.06788v2 Announce Type: replace-cross Abstract: We present a solver-agnostic framework in which coordinated large language model (LLM) agents autonomously execute the complete computational

model-releasesarxiv-cs-cl
14 Apr 2026
Research

From Pixels to Digital Agents: An Empirical Study on the Taxonomy and Technological Trends of Reinforcement Learning Environments

DGX agent

arXiv:2603.23964v2 Announce Type: replace Abstract: The remarkable progress of reinforcement learning (RL) is intrinsically tied to the environments used to train and evaluate artificial agents. Movin

researcharxiv-cs-ai
14 Apr 2026
Model Releases

HearthNet: Edge Multi-Agent Orchestration for Smart Homes

DGX agent

arXiv:2604.09618v1 Announce Type: cross Abstract: Smart-home users increasingly want to control their homes in natural language rather than assemble rules, dashboards, and API integrations by hand. At

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Hunt Globally: Wide Search AI Agents for Drug Asset Scouting in Investing, Business Development, and Competitive Intelligence

DGX agent

arXiv:2602.15019v3 Announce Type: replace Abstract: Bio-pharmaceutical innovation has shifted: many new drug assets now originate outside the United States and are disclosed primarily via regional, no

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Infusing Theory of Mind into Socially Intelligent LLM Agents

DGX agent

arXiv:2509.22887v2 Announce Type: replace Abstract: Theory of Mind (ToM)-an understanding of the mental states of others-is a key aspect of human social intelligence, yet, chatbots and LLM-based socia

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

MERMAID: Memory-Enhanced Retrieval and Reasoning with Multi-Agent Iterative Knowledge Grounding for Veracity Assessment

DGX agent

arXiv:2601.22361v2 Announce Type: replace-cross Abstract: Assessing the veracity of online content has become increasingly critical. Large language models (LLMs) have recently enabled substantial prog

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Mobile GUI Agent Privacy Personalization with Trajectory Induced Preference Optimization

DGX agent

arXiv:2604.11259v1 Announce Type: new Abstract: Mobile GUI agents powered by Multimodal Large Language Models (MLLMs) can execute complex tasks on mobile devices. Despite this progress, most existing

model-releasesarxiv-cs-ai
14 Apr 2026
Safety

PEMANT: Persona-Enriched Multi-Agent Negotiation for Travel

DGX agent

arXiv:2604.10475v1 Announce Type: new Abstract: Modeling household-level trip generation is fundamental to accurate demand forecasting, traffic flow estimation, and urban system planning. Existing stu

safetyarxiv-cs-ai
14 Apr 2026
Model Releases

Pioneer Agent: Continual Improvement of Small Language Models in Production

DGX agent

arXiv:2604.09791v1 Announce Type: new Abstract: Small language models are attractive for production deployment due to their low cost, fast inference, and ease of specialization. However, adapting them

model-releasesarxiv-cs-ai
14 Apr 2026
Research

SynthAgent: Adapting Web Agents with Synthetic Supervision

DGX agent

arXiv:2511.06101v3 Announce Type: replace-cross Abstract: Web agents struggle to adapt to new websites due to the scarcity of environment specific tasks and demonstrations. Recent works have explored

researcharxiv-cs-ai
14 Apr 2026
Research

3D-VCD: Hallucination Mitigation in 3D-LLM Embodied Agents through Visual Contrastive Decoding

DGX agent

arXiv:2604.08645v1 Announce Type: cross Abstract: Large multimodal models are increasingly used as the reasoning core of embodied agents operating in 3D environments, yet they remain prone to hallucin

researcharxiv-cs-ai
13 Apr 2026
Model Releases

BadSkill: Backdoor Attacks on Agent Skills via Model-in-Skill Poisoning

DGX agent

arXiv:2604.09378v1 Announce Type: cross Abstract: Agent ecosystems increasingly rely on installable skills to extend functionality, and some skills bundle learned model artifacts as part of their exec

model-releasesarxiv-cs-ai
13 Apr 2026
Agents

Bayesian Social Deduction with Graph-Informed Language Models

DGX agent

arXiv:2506.17788v2 Announce Type: replace Abstract: Social reasoning - inferring unobservable beliefs and intentions from partial observations of other agents - remains a challenging task for large la

agentsarxiv-cs-ai
13 Apr 2026
Model Releases

CORA: Conformal Risk-Controlled Agents for Safeguarded Mobile GUI Automation

DGX agent

arXiv:2604.09155v1 Announce Type: cross Abstract: Graphical user interface (GUI) agents powered by vision language models (VLMs) are rapidly moving from passive assistance to autonomous operation. How

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

From Reasoning to Agentic: Credit Assignment in Reinforcement Learning for Large Language Models

DGX agent

arXiv:2604.09459v1 Announce Type: new Abstract: Reinforcement learning (RL) for large language models (LLMs) increasingly relies on sparse, outcome-level rewards -- yet determining which actions withi

model-releasesarxiv-cs-cl
13 Apr 2026
Model Releases

MONETA: Multimodal Industry Classification through Geographic Information with Multi Agent Systems

DGX agent

arXiv:2604.07956v2 Announce Type: replace Abstract: Industry classification schemes are integral parts of public and corporate databases as they classify businesses based on economic activity. Due to

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

StreamMeCo: Long-Term Agent Memory Compression for Efficient Streaming Video Understanding

DGX agent

arXiv:2604.09000v1 Announce Type: new Abstract: Vision agent memory has shown remarkable effectiveness in streaming video understanding. However, storing such memory for videos incurs substantial memo

model-releasesarxiv-cs-cv
13 Apr 2026
Safety

Front-End Ethics for Sensor-Fused Health Conversational Agents: An Ethical Design Space for Biometrics

DGX agent

arXiv:2604.06203v1 Announce Type: cross Abstract: The integration of continuous data from built-in sensors and Large Language Models (LLMs) has fueled a surge of 'Sensor-Fused LLM agents' for personal

safetyarxiv-cs-ai
10 Apr 2026
Model Releases

PASK: Toward Intent-Aware Proactive Agents with Long-Term Memory

DGX agent

arXiv:2604.08000v1 Announce Type: cross Abstract: Proactivity is a core expectation for AGI. Prior work remains largely confined to laboratory settings, leaving a clear gap in real-world proactive age

model-releasesarxiv-cs-cl
10 Apr 2026
Agents

@skills: Attention is all you have

DGX agent

arXiv:2608.12610v1 Announce Type: new Abstract: There are 56,804 public agent skills today, and teams write many more privately. The dominant delivery model is installation: once installed, a skill's

agentsarxiv-cs-ai
14 Aug 2026
Agents

SyncSBC: Decentralized Swarm Behavior Prediction for Synchronized Autonomous Control

DGX agent

arXiv:2608.06587v1 Announce Type: cross Abstract: Robot swarms utilize many independent limited-sensing agents to produce complex emergent behaviors without requiring centralized control. However, lit

agentsarxiv-cs-ai
10 Aug 2026
Agents

Dr. AGENTONOMICS: A Didactic Experiment of AGENTONOMICS

DGX agent

arXiv:2608.03524v1 Announce Type: new Abstract: AGENTONOMICS is a framework that treats AI agents as economic entities that can be designed, managed, and governed through an integrated management arch

agentsarxiv-cs-ai
5 Aug 2026
Agents

Mixed-Initiative Human-Robot Teaming under Suboptimality with Online Bayesian Adaptation

DGX agent

arXiv:2403.16178v2 Announce Type: replace-cross Abstract: For effective human-agent teaming, robots and other artificial intelligence (AI) agents must infer their human partner's abilities and behavio

agentsarxiv-cs-ai
5 Aug 2026
Agents

Human-LLM Collaborative Inductive Coding for Conceptualizing K-12 Educator AI Use

DGX agent

arXiv:2607.28889v1 Announce Type: cross Abstract: Qualitative researchers increasingly encounter interaction corpora whose scale exceeds what manual coding alone can address, and large language models

agentsarxiv-cs-ai
3 Aug 2026
Agents

Belief Coevolution in a Social Network of Generalist and Specialist Large Language Models

DGX agent

arXiv:2607.27512v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly deployed in multi-agent environments. However, the processes by which beliefs form and propagate among int

agentsarxiv-cs-cl
31 Jul 2026
Agents

Metis: Memory Foundation Model

DGX agent

arXiv:2607.26760v1 Announce Type: new Abstract: Recent advances in AI agents have increasingly internalized native capabilities into their underlying foundation models, giving rise to multimodal found

agentsarxiv-cs-cl
30 Jul 2026
Agents

Self-Configurable Mesh-Networks for Scalable Distributed Submodular Bandit Optimization

DGX agent

arXiv:2602.19366v2 Announce Type: replace-cross Abstract: We study how to scale distributed bandit submodular coordination under realistic communication constraints in bandwidth, data rate, and connec

agentsarxiv-cs-ro
30 Jul 2026
Agents

Which RAG Paradigm Wins at Scale? A Scaling Study of Retrieval-Augmented Generation Paradigms

DGX agent

arXiv:2607.26497v1 Announce Type: new Abstract: Retrieval-augmented generation (RAG) methods range from lexical and dense retrieval to graph-based indexing and agentic search. They are usually evaluat

agentsarxiv-cs-cl
30 Jul 2026
Agents

Compact Latent Coordination for Autonomous Vehicles at Unsignalized Intersections

DGX agent

arXiv:2607.21488v1 Announce Type: cross Abstract: Coordinating autonomous vehicles at unsignalized intersections remains a critical challenge for multi-agent reinforcement learning (MARL) systems, whi

agentsarxiv-cs-ai
24 Jul 2026
Agents

TTHE: Test-Time Harness Evolution

DGX agent

arXiv:2607.08124v1 Announce Type: cross Abstract: The behavior of an LLM agent is determined not only by the underlying model, but also by its harness: the executable program that constructs context,

agentsarxiv-cs-lg
10 Jul 2026
Agents

Progressive Disclosure for LLM-Maintained Wiki Knowledge Bases: a Preregistered Ablation

DGX agent

arXiv:2607.04576v1 Announce Type: new Abstract: LLM agents increasingly answer questions against knowledge bases they help maintain. A common intuition holds that progressive disclosure, a compact cat

agentsarxiv-cs-cl
7 Jul 2026
Agents

SWE-INTERACT: Reimagining SWE Benchmarks as User-Driven Long-Horizon Coding Sessions

DGX agent

arXiv:2606.30573v1 Announce Type: new Abstract: We introduce SWE-Interact, a new testbed for evaluating coding agents on multi-turn, interactive, user-driven software engineering tasks. Existing front

agentsarxiv-cs-lg
30 Jun 2026
Agents

CCKS: Consensus-based Communication and Knowledge Sharing

DGX agent

arXiv:2606.12281v1 Announce Type: cross Abstract: In Decentralized Training and Decentralized Execution (DTDE) for cooperative Multi-Agent Reinforcement Learning (MARL), action-advising-based knowledg

agentsarxiv-cs-ai
11 Jun 2026
Agents

A Motivational Architecture for Conversational AGI

DGX agent

arXiv:2606.05411v1 Announce Type: new Abstract: Motivational architectures in cognitive AI have largely been designed for physical agents regulating bodily needs. Conversational agents operate in a di

agentsarxiv-cs-ai
6 Jun 2026
Agents

OdysseyArena: Benchmarking Large Language Models For Long-Horizon, Active and Inductive Interactions

DGX agent

arXiv:2602.05843v2 Announce Type: replace Abstract: The rapid advancement of Large Language Models (LLMs) has catalyzed the development of autonomous agents capable of navigating complex environments.

agentsarxiv-cs-cl
5 Jun 2026
Safety

A Unified Framework for Locality in Scalable MARL

DGX agent

arXiv:2602.16966v2 Announce Type: replace-cross Abstract: Scalable methods for networked multi-agent reinforcement learning let each agent plan using only a small neighborhood of the agent graph. This

safetyarxiv-cs-ai
4 Jun 2026
Agents

MemCollab: Cross-Model Memory Collaboration via Contrastive Trajectory Distillation

DGX agent

arXiv:2603.23234v2 Announce Type: replace Abstract: LLM agents increasingly rely on memory mechanisms to reuse knowledge from past problem-solving experiences. However, existing methods typically cons

agentsarxiv-cs-ai
29 May 2026
← Previous
1…9495969798…236
Next →