AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,761 results
Model Releases

Compaction as Epistemic Failure: How Agentic LLM Tools Fabricate Confirmed Results from Killed Processes

DGX agent

arXiv:2607.13071v1 Announce Type: cross Abstract: Agentic LLM coding tools compress long session histories into compaction summaries that subsequent sessions inherit as ground truth. This paper docume

model-releasesarxiv-cs-ai
16 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Set-shifting Behavioral Test for Harnessed Agents

DGX agent

arXiv:2607.13396v1 Announce Type: new Abstract: What happens to an LLM agent's tool choice when the reliable tool silently changes within an ongoing session? We borrow set-shifting from cognitive psyc

model-releasesarxiv-cs-ai
16 Jul 2026
Model Releases

Agents-A1-4B (Qwen3.7-4B ???) : Scaling the Horizon, Not the Parameters

DGX agent

MODEL + GGUF : https://huggingface.co/InternScience/models?search=a1-4b Technical Report Benchmark Qwen3.5-4B Agents-A1-4B Qwen3.5 Qwen3.6 Nex-N2-mini Agents-A1 🧠 Dense Models (~4B) 🔀 MoE Models (35B-

model-releasesr-localllama
15 Jul 2026
Model Releases

AutoTrace: From Patches to Triggers via Agentic Interprocedural Exploration

DGX agent

arXiv:2607.12058v1 Announce Type: cross Abstract: Given a vulnerability-fixing commit, trigger localization asks which specific statement turns the vulnerable program state into a concrete unsafe oper

model-releasesarxiv-cs-ai
15 Jul 2026
Agents

EFLUX: Elastic Multi-Robot Formation Navigation and Adaptation with Agentic LLMs

DGX agent

arXiv:2607.12050v1 Announce Type: new Abstract: Multi-robot teams operating in confined or cluttered environments must adapt both their formation geometry and group topology to navigate through comple

agentsarxiv-cs-ro
15 Jul 2026
Agents

Evidence-Grounded Verified Agentic Reasoning: A Path Toward Eliminating LLM Hallucination in Empirical Inference via Tool-Attested Kernel Proofs

DGX agent

arXiv:2607.12650v1 Announce Type: cross Abstract: Tool access alone does not make LLM empirical reasoning governable: accepted outputs need not descend from attested evidence, and accepted deductions

agentsarxiv-cs-ai
15 Jul 2026
Agents

SymbOmni: Evolving Agentic Omni Models via Symbolic Concept Learning

DGX agent

arXiv:2607.12042v1 Announce Type: new Abstract: Visual generation is increasingly ubiquitous in diverse domains, from text-to-image/video synthesis to multimodal interactive creation. Yet prevailing m

agentsarxiv-cs-cv
15 Jul 2026
Hardware

How to Run an Autoresearch Workflow with RL Agent Skills and NVIDIA NeMo

DGX agent

Autonomous coding agents such as Codex (GPT‑5.5) can fully automate reinforcement‑learning research workflows by provisioning GPU‑hosted environments, orchestrating experiments, and iteratively optimi

hardwarenvidia-developer
14 Jul 2026
Agents

Building an agentic AI solution at Bluesight with Amazon Bedrock

DGX agent

In this post, we describe how Bluesight used two AWS engagements and Amazon Bedrock AgentCore to evolve from a single-product AI prototype to Prism, a unified agentic AI solution spanning six healthca

agentsaws-ml-blog
13 Jul 2026
Agents

Calibrated Stackelberg Games: Learning Optimal Commitments Against Calibrated Agents

DGX agent

arXiv:2306.02704v2 Announce Type: replace-cross Abstract: We introduce Calibrated Stackelberg Games (CSGs), a generalization of the standard Stackelberg Games (SGs) framework. In CSGs, a principal rep

agentsarxiv-cs-lg
10 Jul 2026
Agents

The agentic marketing stack starts with the data layer

DGX agent

Agentic marketing leverages AI agents to automate marketing tasks, and this approach requires a robust data foundation as its core component. Databricks discusses how organizations need unified data p

agentsdatabricks
10 Jul 2026
Agents

Tool-Making and Self-Evolving LLM Agents in Low-Latency Systems

DGX agent

arXiv:2607.08010v1 Announce Type: new Abstract: Production LLM agents often waste latency and reliability by regenerating code for the same procedural steps on every request. We replace this inference

agentsarxiv-cs-cl
10 Jul 2026
Agents

Holy…? First public demo of an iPhone agent ordering me a donut by @agi_inc & @divgarg. Just late night roommate sessions at @AGIHouseSF

DGX agent

A demonstration of an AI agent with iPhone integration that autonomously ordered a donut, showcasing early progress in agentic AI capabilities. The demo was presented publicly by AGI Inc and Div Garg,

agentsdiv-garg--x
9 Jul 2026
Agents

TIL about soak testing my agents

DGX agent

Soak testing for AI agents involves running agents continuously over extended periods to identify performance degradation, resource leaks, stability issues, and edge cases that may not appear during s

agentsyohei-nakajima--x
9 Jul 2026
Agents

Towards Agentic AI Governance: A Preliminary Assessment

DGX agent

arXiv:2607.07612v1 Announce Type: cross Abstract: Artificial intelligence is rapidly evolving from generative systems to agentic AI capable of autonomously planning and executing tasks. Widely charact

agentsarxiv-cs-ai
9 Jul 2026
Model Releases

Beyond the Leaderboard: A Synthesis of Tool-Use, Planning, and Reasoning Failures in Large Language Model Agents

DGX agent

arXiv:2607.05775v1 Announce Type: new Abstract: Large language model (LLM) agents are increasingly evaluated on their ability to use tools, plan multi-step tasks, coordinate with other agents, and ope

model-releasesarxiv-cs-ai
8 Jul 2026
Safety

CONTRA: Red-Teaming Configurations of Personalizable Agents

DGX agent

arXiv:2607.03220v1 Announce Type: cross Abstract: Recent tools such as OpenClaw have extended the capabilities of LLM-based agents from simple dialog-based systems to fully autonomous agents. These sy

safetyarxiv-cs-ai
7 Jul 2026
Model Releases

DualView: Preventing Indirect Prompt Injection in Personal AI Agents

DGX agent

arXiv:2607.03821v1 Announce Type: cross Abstract: Personal AI agents that run on the user's local machine, such as OpenClaw, automate daily tasks including web search, email, and file management. Thei

model-releasesarxiv-cs-ai
7 Jul 2026
Agents

is anyone standardizing on ATIF as a format for agent traces? or a different format? or is it still wild wild west, roll your own

DGX agent

This post discusses the current state of standardization for agent execution trace formats, questioning whether ATIF (Agent Trace Interchange Format) or alternative formats have achieved adoption, or

agentsharrison-chase--x
7 Jul 2026
Agents

Latent Programming Horizons in Coding Agents

DGX agent

arXiv:2607.05188v1 Announce Type: new Abstract: A coding agent solving a software-engineering task spends dozens of steps reasoning, editing code, and running tests, yet little is known about what the

agentsarxiv-cs-lg
7 Jul 2026
Agents

Learning Task-Sufficient World Models by Synergizing Agentic Exploration and Structured Modeling

DGX agent

arXiv:2607.04409v1 Announce Type: new Abstract: Learning and planning in imagination using world models provides an effective paradigm for training agents for decision-making. However, existing approa

agentsarxiv-cs-lg
7 Jul 2026
Model Releases

Measuring Harness-Induced Belief Divergence in Multi-Step LLM Agents

DGX agent

arXiv:2607.04528v1 Announce Type: new Abstract: Software-agent benchmarks usually report whether an agent solves a task, but the agent reaches that outcome through a harness that controls what it sees

model-releasesarxiv-cs-ai
7 Jul 2026
Agents

Organizational Memory for Agentic Business Process Execution

DGX agent

arXiv:2607.03228v1 Announce Type: new Abstract: LLM-based agents offer new opportunities for automating business process execution beyond the limits of rule-based systems. However, general-purpose LLM

agentsarxiv-cs-ai
7 Jul 2026
Agents

Social Networks of LLM Agents

DGX agent

arXiv:2607.03695v1 Announce Type: new Abstract: Large language model (LLM) agents are increasingly deployed in interacting populations, raising the question of what such populations come to believe co

agentsarxiv-cs-lg
7 Jul 2026
Agents

We're hosting the LA Agentic AI Meetup this Thursday, July 9, 5 to 7pm at Gulp in Playa Vista. Come talk to engineers, founders, and builder…

DGX agent

We're hosting the LA Agentic AI Meetup this Thursday, July 9, 5 to 7pm at Gulp in Playa Vista. Come talk to engineers, founders, and builders working across RAG, agentic workflows, and the broader AI

agentspinecone--x
7 Jul 2026
Agents

@hwchase17 Yes using deep agents and okf both for my wiki. https://github.com/varunyn/wiki-langGraph

DGX agent

This post discusses using deep agents and OpenKeyFlow (OKF) for a wiki project built with LangGraph, a framework for constructing agent-based applications. The developer is leveraging LangGraph's capa

agentsharrison-chase--x
5 Jul 2026
Agents

Can coding-agents replicate scientific ML papers? We know this is possible because we can already do this @dair_ai. Still a great read. So t…

DGX agent

Can coding-agents replicate scientific ML papers? We know this is possible because we can already do this @dair_ai. Still a great read. So they try to replicate an ML paper from its materials alone. T

agentsdair-ai--x
3 Jul 2026
Model Releases

EO-Agents: A Three-Agent LLM Pipeline for Earth Observation Hypothesis Generation

DGX agent

arXiv:2607.01584v1 Announce Type: new Abstract: Large language models have recently been explored for scientific hypothesis generation, but most prior work relies on unstructured literature and free-f

model-releasesarxiv-cs-ai
3 Jul 2026
Safety

Safeguarding LLM Agents from Misalignment through Provenance Analysis

DGX agent

arXiv:2607.01236v1 Announce Type: cross Abstract: As LLM agents gain increasing access to powerful tools, ensuring that their actions are aligned with the user's intent becomes critical. When an agent

safetyarxiv-cs-ai
3 Jul 2026
Agents

The Agentic Garden of Forking Paths

DGX agent

arXiv:2607.01507v1 Announce Type: new Abstract: Empirical research rarely admits a unique analysis. Different analytical choices can lead to different conclusions from the same data, yet these hidden

agentsarxiv-cs-ai
3 Jul 2026
Model Releases

UA-ChatDev: Uncertainty-Aware Multi-Agent Collaboration for Reliable Software Development

DGX agent

arXiv:2607.02186v1 Announce Type: new Abstract: Software development is a complex task that demands cooperation among agents with diverse roles. Large language models (LLMs) have enabled autonomous mu

model-releasesarxiv-cs-ai
3 Jul 2026
Agents

WorkPods agent memory as wikis. @langchain #openwiki @cognition #deepwiki @karpathy #llmwiki @Factory #autowiki @hwchase17 @BraceSproul

DGX agent

This post discusses using WorkPods agent memory systems organized as wikis, likely exploring how LLM agents can maintain and access persistent knowledge repositories in wiki format for improved contex

agentsharrison-chase--x
3 Jul 2026
Agents

Conversable Complexity: Agentic LLM Collectives as Interpretable Substrates

DGX agent

arXiv:2607.01047v1 Announce Type: new Abstract: Complexity and interpretability rarely coincide: systems rich enough for complex behaviours to emerge are usually too opaque to question, while transpar

agentsarxiv-cs-cl
2 Jul 2026
Agents

How to evaluate AI agents, avoid reward hacking, and build better specs

DGX agent

Agent evals are repeatable tests that score whether AI agents completed a task correctly. Learn how to design rubrics, test suites, and trace-based evals that catch failures and prevent reward hacking

agentsarize-ai
2 Jul 2026
Model Releases

LLM Wikis are being slept on. I argue that creating knowledge bases with LLMs or coding agents is one of the most valuable applications of A…

DGX agent

LLM Wikis are being slept on. I argue that creating knowledge bases with LLMs or coding agents is one of the most valuable applications of AI today. It's about being intentional in building and scalin

model-releasesdair-ai--x
2 Jul 2026
Agents

Special episode with @bentannyhill on the Max Agency podcast. A month ago, his team shipped LangSmith Engine, our agent that hunts through y…

DGX agent

Special episode with @bentannyhill on the Max Agency podcast. A month ago, his team shipped LangSmith Engine, our agent that hunts through your agent's failures, prioritizes issues, and drafts the fix

agentsharrison-chase--x
2 Jul 2026
Agents

ACE: Pluggable Adaptive Context Elasticizer across Agents

DGX agent

arXiv:2606.31564v1 Announce Type: new Abstract: The increasing complexity of agentic tasks has led to rapidly growing trajectory lengths, which poses significant challenges for large language model (L

agentsarxiv-cs-ai
1 Jul 2026
Agents

An Agentic AI Framework to Accelerate Scientific Discovery in Plant Phenotyping

DGX agent

arXiv:2606.31831v1 Announce Type: new Abstract: High-throughput plant phenotyping now generates image derived datasets far faster than scientists can analyze them. At Oak Ridge National Laboratory's A

agentsarxiv-cs-ai
1 Jul 2026
Agents

InfiniteWeb: Scalable Web Environment Synthesis for GUI Agent Training

DGX agent

arXiv:2601.04126v3 Announce Type: replace-cross Abstract: GUI agents that interact with graphical interfaces on behalf of users represent a promising direction for practical AI assistants. However, tr

agentsarxiv-cs-ai
1 Jul 2026
Agents

ShopX: A Foundation Model for Intent-to-Item Fulfillment in Agentic Shopping

DGX agent

arXiv:2606.31693v1 Announce Type: cross Abstract: The wave of AI-native applications is moving shopping beyond page- and feed-based browsing toward intent-driven experiences orchestrated by LLM agents

agentsarxiv-cs-ai
1 Jul 2026
Agents

Agentic Coding on Supabase with OpenCode

DGX agent

The article covers the Supabase plugin for OpenCode, an AI coding agent, which bundles the Supabase MCP server and agent skills for direct database and project interaction . The plugin ships official

agentssupabase-blog
30 Jun 2026
Local Ai

AI Trading's Alpha Singularity: Emergent Market Reasoning through Agent-to-Agent Self-Evolution

DGX agent

arXiv:2606.29194v1 Announce Type: new Abstract: Automated alpha mining holds the scoring function fixed and varies the search algorithm over it. A search that converges against a fixed scorer overfits

local-aiarxiv-cs-ai
30 Jun 2026
Agents

Clarus: Coordinating Autonomous Research Agents toward Web-Scale Scientific Collaboration

DGX agent

arXiv:2606.30246v1 Announce Type: new Abstract: Existing autonomous research agents can support parts of the research process, but most systems still treat research as either an isolated assistant tas

agentsarxiv-cs-ai
30 Jun 2026
Agents

GUICrafter: Weakly-Supervised GUI Agent Leveraging Massive Unannotated Screenshots

DGX agent

arXiv:2606.29705v1 Announce Type: new Abstract: Data, as the fundamental substrate of modern intelligence, has greatly driven the development of current foundation models. Naturally, researchers aim t

agentsarxiv-cs-ai
30 Jun 2026
Agents

Hybrid Retriever Evolution for Multimodal Document Reasoning Agents

DGX agent

arXiv:2606.29648v1 Announce Type: cross Abstract: Different retrievers, including lexical, semantic, and multimodal approaches, provide highly complementary strengths for multimodal document understan

agentsarxiv-cs-ai
30 Jun 2026
Model Releases

It Lied to a Doctor to Buy Poison Ingredients: Quantifying Real-World Misuse of Phone-use Agents

DGX agent

arXiv:2606.27944v1 Announce Type: cross Abstract: Phone-use Agents can execute complex tasks end to end across real mobile applications. By operating a real device on the user's behalf, they reach far

model-releasesarxiv-cs-ai
30 Jun 2026
Agents

Mandol: An Agglomerative Agent Memory System for Long-Term Conversations

DGX agent

arXiv:2606.29778v1 Announce Type: cross Abstract: Long-term conversational agents need to remember and query cross-session, multi-typed information with complex correlations. Existing agent memory sys

agentsarxiv-cs-ai
30 Jun 2026
Agents

Modeling Earth-Scale Human-Like Societies with One Billion Agents

DGX agent

arXiv:2506.12078v2 Announce Type: replace-cross Abstract: Understanding the dynamic evolution of complex social phenomena requires both high-fidelity modeling of human behavior and large-scale simulat

agentsarxiv-cs-ai
30 Jun 2026
← Previous
1…5051525354…371
Next →