AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,770 results
Agents

The next scaling law is multi-agent swarms. Mixture of models.

DGX agent

The next scaling law is multi-agent swarms. Mixture of models. Introducing Sakana Fugu: A full multi-agent orchestration system accessible via a single model API. Our ‘Fugu Ultra’ model matches the pe

agentsdavid-ha--x
24 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Agents

ActiveGraph: 1 month in: 📄Paper #1: The Log is the Agent 🧠3 LongMemEval Experiments 🔄 Paper #2: Regimes, self-improvement loop 🎓 http://…

DGX agent

ActiveGraph: 1 month in: 📄Paper #1: The Log is the Agent 🧠3 LongMemEval Experiments 🔄 Paper #2: Regimes, self-improvement loop 🎓 http://learn.activegraph.ai 🗂️ 2 reference agents (code, research) 💾 co

agentsyohei-nakajima--x
23 Jun 2026
Agents

GLM-5.2 is available in Perplexity's Agent API. Just tested it, and it's powerful when paired with the Search SDK inside a sandbox. - Spin u…

DGX agent

GLM-5.2 is available in Perplexity's Agent API. Just tested it, and it's powerful when paired with the Search SDK inside a sandbox. - Spin up a sandbox environment - Call the web search tool (people s

agentszhipu-ai--x
23 Jun 2026
Agents

I'm digging the eve agentic framework from Vercel. I like that everything is files, from the tools to the skills to the evals. More importan…

DGX agent

I'm digging the eve agentic framework from Vercel. I like that everything is files, from the tools to the skills to the evals. More importantly, it's gets you building with agents fast. Very promising

agentsdair-ai--x
23 Jun 2026
Safety

Memory Contagion: Cross-Temporal Propagation of Evaluator Bias via Agent Memory

DGX agent

arXiv:2606.23195v1 Announce Type: new Abstract: Large Language Model (LLM) agents increasingly rely on memory systems to maintain long-term coherence. Recent work shows that agent memories degrade dur

safetyarxiv-cs-lg
23 Jun 2026
Model Releases

Steer, Don't Solve: Training Small Critic Models for Large Code Agents

DGX agent

arXiv:2606.21811v1 Announce Type: cross Abstract: End-to-end code agent training is resource-intensive and plateaus on the strategy-level reasoning needed to resolve code issues, since jointly optimiz

model-releasesarxiv-cs-lg
23 Jun 2026
Agents

Hermes Agent has a new Blank Slate setup mode. The default Quick/Full setup modes work great for most, but if you would rather build your ag…

DGX agent

Hermes Agent has a new Blank Slate setup mode. The default Quick/Full setup modes work great for most, but if you would rather build your agent from the ground up you can now start with just a provide

agentsnous-research--x
20 Jun 2026
Model Releases

AI Coding Agents Can Reproduce Social Science Findings

DGX agent

arXiv:2606.11447v1 Announce Type: new Abstract: Recent anecdotal evidence suggests that AI coding agents can reproduce published findings when provided with original data and code; yet systematic eval

model-releasesarxiv-cs-cl
11 Jun 2026
Agents

Knowing When to Ask: Self-Gated Clarification for Hierarchical Language Agents

DGX agent

arXiv:2606.11349v1 Announce Type: new Abstract: In hierarchical reasoning, failures often originate at intermediate decision points where the agent commits to a wrong branch without recognizing that i

agentsarxiv-cs-ai
11 Jun 2026
Agents

Organize then Retrieve: Hierarchical Memory Navigation for Efficient Agents

DGX agent

arXiv:2606.11680v1 Announce Type: new Abstract: Large language model (LLM) agents struggle with long-horizon tasks due to their inherent statelessness, requiring all task-relevant information to be en

agentsarxiv-cs-ai
11 Jun 2026
Agents

The Confident Liar: Diagnosing Multi-Agent Debate with Log-Probabilities and LLM-as-Judge

DGX agent

arXiv:2606.10296v1 Announce Type: cross Abstract: Multi-agent debate systems are typically evaluated only on whether the final answer is correct, overlooking the quality of the intermediate reasoning

agentsarxiv-cs-ai
10 Jun 2026
Agents

VISTA: A Versatile Interactive User Simulation Toolkit for Agent Evaluation

DGX agent

arXiv:2606.11079v1 Announce Type: new Abstract: Evaluation remains a critical bottleneck for interactive agent development. Existing evaluation methods often rely on static benchmarks, which fail to c

agentsarxiv-cs-cl
10 Jun 2026
Agents

How an Agent Built a 3D Paris Gallery by Chaining Two Hugging Face Spaces

DGX agent

This article describes how an AI agent was used to create a 3D virtual gallery of Paris by chaining together two Hugging Face Spaces applications. It demonstrates a practical example of using agents t

agentshugging-face
9 Jun 2026
Agents

If people only knew how much OpenMed runs on HF Stack like Buckets, Datasets, and Spaces, From datasets, to agent traces, to medical intelli…

DGX agent

If people only knew how much OpenMed runs on HF Stack like Buckets, Datasets, and Spaces, From datasets, to agent traces, to medical intelligent MCPs, to binary builds for OpenMed Agent, @huggingface

agentsclem-delangue--x
9 Jun 2026
Agents

MetaEvo: A Meta-Optimization Framework for Experience-Driven Agent Evolution

DGX agent

arXiv:2606.07603v1 Announce Type: cross Abstract: Large language models (LLMs) exhibit strong reasoning capabilities, yet most LLM-based agents are statically deployed and unable to improve through ta

agentsarxiv-cs-ai
9 Jun 2026
Agents

PACE: Anytime-Valid Acceptance Tests for Self-Evolving Agents

DGX agent

arXiv:2606.08106v1 Announce Type: new Abstract: Self-evolving agents improve by repeatedly proposing changes to their own prompts, skills, or workflows and keeping those that score higher on a small h

agentsarxiv-cs-ai
9 Jun 2026
Model Releases

SWE-Marathon: Can Agents Autonomously Complete Ultra-Long-Horizon Software Work?

DGX agent

arXiv:2606.07682v1 Announce Type: cross Abstract: AI agents are increasingly expected to complete long-horizon workflows that require sustained progress over hours, millions of tokens, and complex env

model-releasesarxiv-cs-ai
9 Jun 2026
Agents

The Token Not Taken: Sampling, State, and the Variability of AI Agent Outputs

DGX agent

arXiv:2606.08998v1 Announce Type: new Abstract: Agentic AI systems can behave differently across runs: the same request may produce a different plan, a different tool call, a different code edit, or a

agentsarxiv-cs-ai
9 Jun 2026
Agents

AdMem: Advanced Memory for Task-solving Agents

DGX agent

arXiv:2606.06787v1 Announce Type: new Abstract: Large Language Models (LLMs) show promise as tool-using agents but remain limited in long-horizon tasks that require remembering, organizing, and reusin

agentsarxiv-cs-ai
8 Jun 2026
Agents

Apple unveils an Apple Intelligence feature to automatically change compromised passwords, using agentic AI, and save them to the Passwords app (James Pero/Gizmodo)

DGX agent

James Pero / Gizmodo: Apple unveils an Apple Intelligence feature to automatically change compromised passwords, using agentic AI, and save them to the Passwords app — Agentic AI and security are norm

agentstechmeme
8 Jun 2026
Agents

At @tryramp, engineers use Devin Desktop to bring their favorite agents into one place. With Devin Desktop they can dispatch, monitor, and j…

DGX agent

At @tryramp, engineers use Devin Desktop to bring their favorite agents into one place. With Devin Desktop they can dispatch, monitor, and jump between agents from a single surface with shared context

agentscognition-ai--x
8 Jun 2026
Agents

AutoTool: Dynamic Tool Selection and Integration for Agentic Reasoning

DGX agent

arXiv:2512.13278v2 Announce Type: replace Abstract: Agentic reinforcement learning has advanced large language models (LLMs) to reason through long chain-of-thought trajectories while interleaving ext

agentsarxiv-cs-cl
8 Jun 2026
Model Releases

Declarative Skills for AI Agents in Knowledge-Grounded Tool-Use Workflows

DGX agent

arXiv:2606.06923v1 Announce Type: new Abstract: We study orchestration mechanisms for tool-using AI agents in realistic customer-service workflows over an unstructured knowledge base. We argue that de

model-releasesarxiv-cs-ai
8 Jun 2026
Agents

The Sim-to-Real Gap of Foundation Model Agents: A Unified MDP Perspective

DGX agent

arXiv:2606.07017v1 Announce Type: new Abstract: Foundation model agents are increasingly deployed for real-world decision-making, but suffer from the sim-to-real gap. While robotics and classical cont

agentsarxiv-cs-ai
8 Jun 2026
Safety

AdaMEM: Test-Time Adaptive Memory for Language Agents

DGX agent

arXiv:2606.05684v1 Announce Type: new Abstract: A central challenge for language agents is utilizing past experience to adapt to dynamic test-time conditions. While recent work demonstrates the promis

safetyarxiv-cs-ai
6 Jun 2026
Agents

Beyond Semantic Organization: Memory as Execution State Management for Long-Horizon Agents

DGX agent

arXiv:2606.06090v1 Announce Type: new Abstract: LLM-based agents increasingly tackle long-horizon tasks with interdependent decisions, where each action reshapes future constraints and intermediate er

agentsarxiv-cs-ai
6 Jun 2026
Agents

“de log is de agent”… activegraph for EU regulation? (just found this) https://djimit.nl/blog/activegraph-event-sourced-agents

DGX agent

This post explores the concept of using ActiveGraph with event sourcing for AI agents operating under EU regulations, suggesting a potential architecture where detailed logging and event histories ser

agentsyohei-nakajima--x
6 Jun 2026
Agents

Insurance of Agentic AI

DGX agent

arXiv:2606.05449v1 Announce Type: new Abstract: Agentic artificial intelligence (AI) systems are transforming the risk landscape by extending beyond information generation to autonomous planning, tool

agentsarxiv-cs-ai
6 Jun 2026
Agents

Knowledge Activation: AI Skills as the Institutional Knowledge Primitive for Agentic Software Development

DGX agent

arXiv:2603.14805v2 Announce Type: replace Abstract: Enterprise software organizations accumulate critical institutional knowledge - architectural decisions, deployment procedures, compliance policies,

agentsarxiv-cs-ai
6 Jun 2026
Agents

Personal AI Agent for Camera Roll VQA

DGX agent

arXiv:2606.05275v1 Announce Type: new Abstract: We study the personal camera roll visual question answering setting. In this setting, a conversational AI assistant can access a user's personal camera

agentsarxiv-cs-cv
5 Jun 2026
Agents

Retrospective Harness Optimization: Improving LLM Agents via Self-Preference over Trajectory Rollouts

DGX agent

arXiv:2606.05922v1 Announce Type: cross Abstract: AI agents rely on a harness of skills, tools, and workflows to solve complex problems. Continually improving this harness is essential for adapting to

agentsarxiv-cs-cl
5 Jun 2026
Model Releases

AIP: A Graph Representation for Learning and Governing Agent Skills

DGX agent

arXiv:2606.04781v1 Announce Type: new Abstract: Agent Skills today consist largely of free-form prose requiring the agent to read, interpret, and re-derive how to act in every session. This imposes tw

model-releasesarxiv-cs-ai
4 Jun 2026
Agents

Beyond Prompt-Based Planning: MCP-Native Graph Planning-based Biomedical Agent System

DGX agent

arXiv:2606.04494v1 Announce Type: new Abstract: Biomedical agents promise to automate complex biological workflows, yet current systems face two fundamental bottlenecks: bioinformatics tools are highl

agentsarxiv-cs-ai
4 Jun 2026
Model Releases

Exploring the Topology and Memory of Consensus: How LLM Agents Agree, Fragment, or Settle When Forming Conventions

DGX agent

arXiv:2606.04197v1 Announce Type: cross Abstract: How much should an LLM agent remember, and how should multi-agent systems be connected when trying to reach consensus? We show these two design choice

model-releasesarxiv-cs-cl
4 Jun 2026
Model Releases

I am hooked on Dynamic Workflows! The idea of generating harnesses on the fly is so compelling that I reverse-engineered it for my agent orc…

DGX agent

I am hooked on Dynamic Workflows! The idea of generating harnesses on the fly is so compelling that I reverse-engineered it for my agent orchestrator. And then I built a monitoring dashboard (as an HT

model-releasesdair-ai--x
4 Jun 2026
Agents

its been such fun befriending Pari and seeing him completely reinvent his company for the agentic era, WHILE having the most insanely stacke…

DGX agent

its been such fun befriending Pari and seeing him completely reinvent his company for the agentic era, WHILE having the most insanely stacked customer base I've ever seen in the hardest engineering do

agentsswyx--x
4 Jun 2026
Agents

MIRAGE: Mobile Agents with Implicit Reasoning and Generative World Models

DGX agent

arXiv:2606.04627v1 Announce Type: new Abstract: Mobile agents are increasingly expected to operate everyday applications from screenshots and language goals, where reliable control requires reasoning

agentsarxiv-cs-ai
4 Jun 2026
Agents

Strabo: Declarative Specification and Implementation of Agentic Interaction Protocols

DGX agent

arXiv:2606.05043v1 Announce Type: new Abstract: The last few years have witnessed major advances in the modeling and implementation of multiagent systems based on declarative interaction protocols. Ou

agentsarxiv-cs-ai
4 Jun 2026
Agents

Topology Matters: Measuring Memory Leakage in Multi-Agent LLMs

DGX agent

arXiv:2512.04668v4 Announce Type: replace-cross Abstract: Graph topology is a fundamental determinant of memory leakage in multi-agent LLM systems, yet its effects remain poorly quantified. We introdu

agentsarxiv-cs-ai
4 Jun 2026
Agents

Automata-Conditioned Cooperative Multi-Agent Reinforcement Learning

DGX agent

arXiv:2511.02304v2 Announce Type: replace-cross Abstract: We study learning multi-task, multi-agent policies for cooperative, temporal objectives, under centralized training, decentralized execution.

agentsarxiv-cs-ai
3 Jun 2026
Agents

Chopped it up with @swyx on @latentspacepod and we ran the gamut on this one. We talked platform, how roles are evolving, the agentic era, t…

DGX agent

Chopped it up with @swyx on @latentspacepod and we ran the gamut on this one. We talked platform, how roles are evolving, the agentic era, the future of open source, and what we’re building next. Spoi

agentsswyx--x
3 Jun 2026
Model Releases

RealClawBench: Live OpenClaw Benchmarks from Real Developer-Agent Sessions

DGX agent

arXiv:2606.03889v1 Announce Type: new Abstract: Agent benchmarks should reflect what users actually ask deployed agents to do, yet existing benchmarks often miss key realism properties of real develop

model-releasesarxiv-cs-cl
3 Jun 2026
Agents

Sema4.ai’s autonomous agent-building platform gets simpler to use, adds deeper business context and more

DGX agent

Sema4.ai Inc. a startup that provides tools for building and managing artificial intelligence agents, today announced a massive revamp of its platform, with big changes coming to every layer of the ag

agentssiliconangle
3 Jun 2026
Agents

ToolGate: Token-Efficient Pre-Call Control for Tool-Augmented Vision-Language Agents

DGX agent

arXiv:2606.03054v1 Announce Type: new Abstract: Tool-augmented vision-language agents can acquire external perceptual evidence through OCR, detection, segmentation, and other tools, but executing ever

agentsarxiv-cs-ai
3 Jun 2026
Model Releases

WRIT: Write-Read Intensive Trajectory Synthesis for Multi-Turn User-Facing Agents

DGX agent

arXiv:2606.02908v1 Announce Type: cross Abstract: Multi-turn user-facing agents must infer user intent from incomplete requests, collect missing information through dialogue and tools, and execute val

model-releasesarxiv-cs-ai
3 Jun 2026
Agents

Agentic Authoring of Interactive Multiview Visualizations in Genomics

DGX agent

arXiv:2606.00370v1 Announce Type: cross Abstract: Diverse genomics data, scientific questions, and analysis tasks typically demand highly specialized visualizations. Therefore, users often must custom

agentsarxiv-cs-ai
2 Jun 2026
Agents

AgentxGCore: Agentic AI for Next-Generation Mobile Core Network

DGX agent

arXiv:2606.00417v1 Announce Type: cross Abstract: To meet the stringent requirements of emerging applications and the increasingly complex network management and operation, the Next Generation Mobile

agentsarxiv-cs-ai
2 Jun 2026
Agents

Code2Math: Can Your Code Agent Effectively Evolve Math Problems Through Exploration?

DGX agent

arXiv:2603.03202v3 Announce Type: replace Abstract: As large language models (LLMs) advance their mathematical capabilities toward the IMO and research level, the scarcity of challenging, high-quality

agentsarxiv-cs-cl
2 Jun 2026
← Previous
1…6162636465…371
Next →