AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,153 results
Model Releases

Beyond Pixels: Exploring DOM Downsampling for LLM-Based Web Agents

DGX agent

arXiv:2508.04412v3 Announce Type: replace Abstract: The advent of large language models (LLMs) has sparked an evolution of autonomous web browsing agents: given a web browsing task and serialised user

model-releasesarxiv-cs-ai
11 Aug 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Bidirectional Context Self-Distillation for Reinforcement Learning of Skill-Based LLM Agents

DGX agent

arXiv:2608.09555v1 Announce Type: new Abstract: External natural-language skills provide large language model (LLM) agents with reusable and editable guidance for solving complex tasks. Yet their effe

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

DarwinX: Evolving Agent Harnesses Through Natural Selection

DGX agent

arXiv:2608.07545v1 Announce Type: cross Abstract: An LLM agent's capability depends not only on model weights but on its harness: prompts, tools, skills, and control flow. Self-improvement loops alrea

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Discovering Diverse Planning Policies for Multimodal Embodied Agents with Quality-Diversity Optimization

DGX agent

arXiv:2608.08523v1 Announce Type: new Abstract: Multimodal embodied agents are increasingly required to solve long-horizon tasks by integrating visual observations, textual goals, and interaction hist

model-releasesarxiv-cs-ai
11 Aug 2026
Tutorials

Experience-Sensitive Game Learning: A Behavioral Study of Humans and Language Agents

DGX agent

arXiv:2608.07490v1 Announce Type: cross Abstract: Large language model agents are increasingly evaluated through games, but most benchmarks emphasize final outcomes rather than how players learn from

tutorialsarxiv-cs-ai
11 Aug 2026
Agents

IDRAAK: From Multi-Agent NLP to Few-Shot Prompting for Semantic Drift Detection in Technical Requirements

DGX agent

arXiv:2608.08801v1 Announce Type: new Abstract: Translating technical requirements across languages can introduce semantic drift, altering numerical constraints, polarities, modalities, or other speci

agentsarxiv-cs-cl
11 Aug 2026
Agents

InfMem: Learning System-2 Memory Control for Long-Context Agent

DGX agent

arXiv:2602.02704v2 Announce Type: replace Abstract: Reasoning over ultra-long documents requires synthesizing sparse evidence scattered across distant segments under strict memory constraints. While s

agentsarxiv-cs-cl
11 Aug 2026
Model Releases

Knowing You Is Everything: LLM Agents Achieve Near-Perfect Profile-Consistent Reaction Prediction in Social Media Simulation

DGX agent

arXiv:2608.07498v1 Announce Type: cross Abstract: Autonomous AI agents in social media present concrete risks to democratic discourse and platform governance, while also offering tools for pre-deploym

model-releasesarxiv-cs-ai
11 Aug 2026
Safety

Legal Responsibilities Using Autonomous Agents For Artificial Intelligence

DGX agent

arXiv:2608.08022v1 Announce Type: new Abstract: Recent incidents involving Artificial Intelligence (AI) agents, which were reported escaping their containment `unintentionally' to gain unauthorized ac

safetyarxiv-cs-ai
11 Aug 2026
Safety

Metanormative Theory for RL-Based Moral Agents

DGX agent

arXiv:2608.08220v1 Announce Type: new Abstract: The overlapping disciplines of machine ethics and value alignment are concerned with designing artificial agents that are aligned with human values and

safetyarxiv-cs-ai
11 Aug 2026
Model Releases

Open Evaluation Agent: Efficient and Promptable Evaluation of Visual Generative Models

DGX agent

arXiv:2608.09666v1 Announce Type: new Abstract: Recent advances in visual generative models have enabled high-quality image and video generation, but evaluating these models often demands sampling hun

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Ouroboros: A Self-Developing Frontier Coding Agent with Reviewed Core Evolution

DGX agent

arXiv:2608.08311v1 Announce Type: cross Abstract: We present Ouroboros, a self-developing agent harness whose tools, prompts, context assembly, and core implementation improve through reviewed commits

model-releasesarxiv-cs-ai
11 Aug 2026
Agents

QuantumMind: Constraint-Grounded Agentic Reasoning for Speedup Analysis in Quantum Computing

DGX agent

arXiv:2608.07743v1 Announce Type: new Abstract: Identifying a meaningful quantum speedup requires more than matching a classical problem to a familiar quantum primitive: the claim must preserve the ta

agentsarxiv-cs-ai
11 Aug 2026
Safety

Reading is not Reasoning: Bridging the Agentic Policy Gap in Vision-Text Compression

DGX agent

arXiv:2608.08960v1 Announce Type: new Abstract: Multi-step language-model agents repeatedly process growing interaction histories, leading to substantial context costs. Vision--text compression reduce

safetyarxiv-cs-ai
11 Aug 2026
Model Releases

Sparks of Cooperative Reasoning: LLMs as Strategic Hanabi Agents

DGX agent

arXiv:2601.18077v3 Announce Type: replace Abstract: Cooperative reasoning under incomplete information remains challenging for both humans and multi-agent systems. The card game Hanabi embodies this c

model-releasesarxiv-cs-cl
11 Aug 2026
Agents

STEMMA: An Adversarial Multi-Agent Framework for Evaluating Self-Identity Consistency in LLMs

DGX agent

arXiv:2608.08164v1 Announce Type: cross Abstract: Knowledge Distillation is a widely adopted technique in the training and fine-tuning of large language models (LLMs) enabling transfer of structured i

agentsarxiv-cs-ai
11 Aug 2026
Model Releases

SWE-Bench ProMax: Benchmarking Agents on Large-Scale Multilingual Code Refactoring

DGX agent

arXiv:2608.09802v1 Announce Type: new Abstract: As AI coding agents take on increasingly complex, long-horizon software engineering tasks, existing benchmarks are rapidly saturating and their evaluati

model-releasesarxiv-cs-cl
11 Aug 2026
Model Releases

WebChoreArena: Evaluating Web Browsing Agents on Realistic Tedious Web Tasks

DGX agent

arXiv:2506.01952v2 Announce Type: replace-cross Abstract: Powered by large language models (LLMs), web browsing agents operate graphical user interfaces in a human-like manner, offering a transparent

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

When LLM Agents Negotiate: Private Information and Dynamic Bargaining in Supply Chains

DGX agent

arXiv:2608.07538v1 Announce Type: new Abstract: As LLM agents move from decision support to autonomous procurement, firms need to know whether delegated negotiators create value, divide it predictably

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

You Don't Need To Stay in The Loop: An Agentic Robotics Loop for Robot-Policy Improvement

DGX agent

arXiv:2608.07555v1 Announce Type: new Abstract: Coding agents such as Claude Code and Codex close the software loop: a main agent manages the loop, subagents analyze and execute, tools do the work. We

model-releasesarxiv-cs-ro
11 Aug 2026
Model Releases

HLSmith: An Expert-Guided Agentic Framework for C/C++-to-HLS Translation

DGX agent

arXiv:2608.06791v1 Announce Type: cross Abstract: Application-specific FPGA accelerators offer substantial performance and energy-efficiency gains across many application domains, but developing them

model-releasesarxiv-cs-ai
10 Aug 2026
Local Ai

Homebot: A Personal AI Agent for Conversational Home Assistance and Automation

DGX agent

arXiv:2608.02254v2 Announce Type: replace Abstract: exttt{Homebot} is a locally deployable AI agent for conversational household assistance and automation. It accepts voice and instant-messaging reque

local-aiarxiv-cs-ai
10 Aug 2026
Safety

How Much, Then Where: Credit-Conserving Action-to-Token Allocation for Multi-Turn Agent Reinforcement Learning

DGX agent

arXiv:2608.07118v1 Announce Type: new Abstract: Credit assignment in multi-turn agent reinforcement learning operates at two levels: assigning trajectory-level credit to actions and distributing each

safetyarxiv-cs-ai
10 Aug 2026
Safety

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents

DGX agent

arXiv:2608.07068v1 Announce Type: new Abstract: Long-horizon agents accumulate growing contexts during interaction, impairing performance and stability. Compact memory mitigates this problem by compre

safetyarxiv-cs-ai
10 Aug 2026
Agents

APQF: Agentic Profiling-Guided Structured Pruning and Mixed-Precision Quantization with Adaptive Fine-Tuning

DGX agent

arXiv:2608.05499v1 Announce Type: cross Abstract: Modern deep neural networks achieve strong performance, but their scale makes them costly and slow, especially on resource-constrained edge devices. P

agentsarxiv-cs-ai
7 Aug 2026
Agents

DoctorAgents: an agentic framework to iteratively refine AutoML pipeline for small clinical temporal data

DGX agent

arXiv:2608.05375v1 Announce Type: new Abstract: Clinical machine learning (ML) has the potential to support high-stakes medical decision-making, but reliable deployment is often constrained by scarce,

agentsarxiv-cs-ai
7 Aug 2026
Safety

EnvACE: Internalizing Environment Dynamics via World Rehearsal for Agentic Reinforcement Learning

DGX agent

arXiv:2608.06197v1 Announce Type: new Abstract: Training large language model agents for long-horizon tool use typically relies on interactions with real or synthesized executable environments, whose

safetyarxiv-cs-ai
7 Aug 2026
Model Releases

FinEvo-Bench: A Longitudinal Benchmark for Self-Evolving Agents in Professional Financial Workflows

DGX agent

arXiv:2608.06144v1 Announce Type: new Abstract: Most agent benchmarks evaluate tasks independently and cannot measure whether experience from one task helps with later tasks. Existing self-evolution b

model-releasesarxiv-cs-ai
7 Aug 2026
Agents

Learning Context-Free Grammars for Grammar-Constrained Decoding via Declarative Agentic Programming with Guarantees

DGX agent

arXiv:2608.05493v1 Announce Type: cross Abstract: Language models (LMs) are increasingly used to interact with external services via programs written in domain-specific languages (DSLs). Unfortunately

agentsarxiv-cs-ai
7 Aug 2026
Agents

ArtAnno: Annotating Implicit Semantics in Artworks through LLM Agent-Driven Bidirectional Human-AI Augmentation

DGX agent

arXiv:2608.05026v1 Announce Type: cross Abstract: High-quality annotation of artworks is essential for computational art research, yet extracting implicit semantics remains challenging due to the reli

agentsarxiv-cs-ai
6 Aug 2026
Safety

Calibrating Artificial Guilt: Neurally Grounded Reward Shaping for Prosocial Multi-Agent Reinforcement Learning

DGX agent

arXiv:2608.04663v1 Announce Type: new Abstract: Cooperative multi-agent reinforcement learning often adds social terms to individual rewards, yet the scale of those terms is usually chosen by hand. We

safetyarxiv-cs-ai
6 Aug 2026
Model Releases

Diagnosing Tool-Selection Reasoning in LLM Agents with Canary Tools

DGX agent

arXiv:2608.04719v1 Announce Type: new Abstract: Agent evaluations tell us that a model picked the wrong tool, but rarely why. We introduce canary tools: diagnostic probe tools planted in an agent's Mo

model-releasesarxiv-cs-ai
6 Aug 2026
Safety

HALT: Verification-Aware Stopping for Retrieval-Augmented Search Agents

DGX agent

arXiv:2608.02009v2 Announce Type: replace Abstract: Retrieval-augmented search agents answer multi-hop questions by repeatedly issuing search queries and accumulating evidence. This creates a stopping

safetyarxiv-cs-ai
6 Aug 2026
Safety

SafeCommit: Certifying When Memory-Grounded Agents May Safely Act

DGX agent

arXiv:2608.04289v1 Announce Type: new Abstract: Long-horizon agents increasingly use persistent memory and tools to take actions with external side effects. A central failure mode is premature commitm

safetyarxiv-cs-ai
6 Aug 2026
Local Ai

AgenticVAU: Multi-Agent Explore-Verify Reasoning for Video Anomaly Understanding

DGX agent

arXiv:2608.03779v1 Announce Type: new Abstract: Video anomaly understanding (VAU) focuses on comprehensively interpreting abnormal events in videos, requiring models to identify anomalous occurrences,

local-aiarxiv-cs-cv
5 Aug 2026
Agents

Asking Questions the Right Way: A Multi-Agent Conversational System for Prompt Formulation in Complex Task Resolution

DGX agent

arXiv:2608.01366v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are integral to complex intellectual tasks, yet output quality remains constrained by user-provided prompts. Iter

agentsarxiv-cs-ai
5 Aug 2026
Agents

Don't Regenerate, Debug: A Domain-Specific Agent for Repairing Near-Miss Hardware Operators

DGX agent

arXiv:2608.02712v1 Announce Type: cross Abstract: Kernel generation for hardware accelerators such as GPUs and NPUs has become a proving ground for large language models (LLMs), and state-of-the-art s

agentsarxiv-cs-ai
5 Aug 2026
Model Releases

EduClaw-Bench: A Long-Horizon Benchmark for Pedagogical LLM Agents with Simulated Learners

DGX agent

arXiv:2608.03206v1 Announce Type: cross Abstract: Large language models (LLMs) power educational applications from tutoring to essay scoring, but each is a point solution to a single task, and only re

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

When Agents Learn to Be You: Benchmarking Privacy Leakage, Impersonation Risk, and Defenses in Persona Skills

DGX agent

arXiv:2608.03700v1 Announce Type: cross Abstract: Persona skills distill personal interaction histories into portable and executable artifacts for downstream agents. While enabling flexible personaliz

model-releasesarxiv-cs-cl
5 Aug 2026
Agents

Agentic Graph Token Reasoning

DGX agent

arXiv:2608.00542v1 Announce Type: new Abstract: Graphs model relational data throughout science and industry, from citation networks to product co-purchase graphs. Because the nodes of many such graph

agentsarxiv-cs-lg
4 Aug 2026
Agents

FactorJEPA: Factorizing Monolithic Futures into Layout-Agent-Interaction Channels for Crowded and Chaotic Global South Urban Worlds

DGX agent

arXiv:2608.01049v1 Announce Type: cross Abstract: World models have attracted significant attention for their ability to capture and predict the structure and dynamics of the physical world. In this e

agentsarxiv-cs-cv
4 Aug 2026
Model Releases

Long-Horizon Embodied Decision-Making via Multimodal Memory Compression

DGX agent

arXiv:2608.01456v1 Announce Type: cross Abstract: Agents are increasingly expected to act not only as task executors, but also as decision-makers on behalf of human users. This shift requires agents t

model-releasesarxiv-cs-cl
4 Aug 2026
Research

MemSIF: From Structured Interactions to Dual-Track Fact Memory for LLM Agents

DGX agent

arXiv:2608.01742v1 Announce Type: cross Abstract: Long-term memory is critical for LLM agents operating over long-horizon interactions. However, several persistent limitations of existing memory syste

researcharxiv-cs-cl
4 Aug 2026
Model Releases

Distilling Knowledge from Large Language Models into Lightweight Reinforcement Learning Agents for Autonomous Cyber Operations

DGX agent

arXiv:2607.28826v1 Announce Type: new Abstract: Autonomous Cyber Operations (ACO) are increasingly important for defending enterprise networks as cyber threats continue to evolve in sophistication. AC

model-releasesarxiv-cs-lg
3 Aug 2026
Model Releases

MMShopBench: A Real-Log Benchmark for Multimodal, Multi-Turn Shopping Agents

DGX agent

arXiv:2607.29002v1 Announce Type: new Abstract: Online shoppers increasingly turn to AI shopping assistants, using images and multi-turn dialogue to express and refine product needs that are difficult

model-releasesarxiv-cs-ai
3 Aug 2026
Model Releases

SeekBrain: An Autonomous Multi-Agent System for Accelerating Neuroscience Discovery

DGX agent

arXiv:2607.29347v1 Announce Type: cross Abstract: Modern neuroscience relies on integrating multi-scale, multimodal datasets to uncover the neural principles underlying intelligence. However, analytic

model-releasesarxiv-cs-ai
3 Aug 2026
Local Ai

Zero-Mem: Zero-Token Memory Operations for LLM Agents

DGX agent

arXiv:2607.29377v1 Announce Type: new Abstract: LLM agents need memory to act consistently over long interactions, yet many systems use additional LLM calls to operate that memory. Generating intermed

local-aiarxiv-cs-cl
3 Aug 2026
Safety

Harness-G: A Graph-Structured Harness for Search Agents

DGX agent

arXiv:2607.27652v1 Announce Type: new Abstract: Reinforcement learning (RL) search agents commonly model retrieval as free-form natural-language query generation and optimize multi-turn interactions u

safetyarxiv-cs-cl
31 Jul 2026
← Previous
1…6364656667…233
Next →