AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,736 results
Model Releases

the fact that http://pi.dev agent is so good, with virtually no sophisticated harness whatsoever, is a testament to the fact token vendor (c…

DGX agent

the fact that http://pi.dev agent is so good, with virtually no sophisticated harness whatsoever, is a testament to the fact token vendor (codex/claude) agents are overrated. highly. today's moat of c

model-releasesharrison-chase--x
11 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Agents

the team spent a lot of time totally revamping docs with care❤️ ofc for LangChain+deepagents DX but lots of ppl use them as general learning…

DGX agent

the team spent a lot of time totally revamping docs with care❤️ ofc for LangChain+deepagents DX but lots of ppl use them as general learning guides on patterns across Agents, Context Eng, Infra, Prod,

agentsharrison-chase--x
11 Apr 2026
Agents

Agent harnesses are spark LangSmith is databricks

DGX agent

Agent harnesses are spark LangSmith is databricks harnesses seem to be the abstraction that encapsulates all of the 'business logic' or 'business connections' into a coherent unit that you can iterate

agentsharrison-chase--x
10 Apr 2026
Safety

Android Coach: Improve Online Agentic Training Efficiency with Single State Multiple Actions

DGX agent

arXiv:2604.07277v1 Announce Type: cross Abstract: Online reinforcement learning (RL) serves as an effective method for enhancing the capabilities of Android agents. However, guiding agents to learn th

safetyarxiv-cs-ai
10 Apr 2026
Model Releases

GameWorld: Towards Standardized and Verifiable Evaluation of Multimodal Game Agents

DGX agent

arXiv:2604.07429v1 Announce Type: new Abstract: Towards an embodied generalist for real-world interaction, Multimodal Large Language Model (MLLM) agents still suffer from challenging latency, sparse f

model-releasesarxiv-cs-cv
10 Apr 2026
Safety

Governed Capability Evolution for Embodied Agents: Safe Upgrade, Compatibility Checking, and Runtime Rollback for Embodied Capability Modules

DGX agent

arXiv:2604.08059v1 Announce Type: new Abstract: Embodied agents are increasingly expected to improve over time by updating their executable capabilities rather than rewriting the agent itself. Prior w

safetyarxiv-cs-ro
10 Apr 2026
Agents

@hwchase17 middleware was the right abstraction for it too. way more adoptable than asking everyone to restructure their agent setup

DGX agent

LangChain's Middleware abstraction, introduced by Harrison Chase (@hwchase17) in LangChain 1.0 Alpha, addresses context engineering in AI agents by providing clean `before_model`, `after_model`, an...

agentsharrison-chase--x
10 Apr 2026
Safety

Karma Mechanisms for Decentralised, Cooperative Multi Agent Path Finding

DGX agent

arXiv:2604.07970v1 Announce Type: cross Abstract: Multi-Agent Path Finding (MAPF) is a fundamental coordination problem in large-scale robotic and cyber-physical systems, where multiple agents must co

safetyarxiv-cs-ro
10 Apr 2026
Safety

KD-MARL: Resource-Aware Knowledge Distillation in Multi-Agent Reinforcement Learning

DGX agent

arXiv:2604.06691v1 Announce Type: new Abstract: Real world deployment of multi agent reinforcement learning MARL systems is fundamentally constrained by limited compute memory and inference time. Whil

safetyarxiv-cs-ai
10 Apr 2026
Agents

MAT-Cell: A Multi-Agent Tree-Structured Reasoning Framework for Batch-Level Single-Cell Annotation

DGX agent

arXiv:2604.06269v1 Announce Type: cross Abstract: Automated cellular reasoning faces a core dichotomy: supervised methods fall into the Reference Trap and fail to generalize to out-of-distribution cel

agentsarxiv-cs-ai
10 Apr 2026
Safety

Multi-agent Reach-avoid MDP via Potential Games and Low-rank Policy Structure

DGX agent

arXiv:2410.17690v2 Announce Type: replace-cross Abstract: We optimize finite horizon multi-agent reach-avoid Markov decision process (MDP) via local feedback policies. The global feedback polic

safetyarxiv-cs-ro
10 Apr 2026
Safety

TwinLoop: Simulation-in-the-Loop Digital Twins for Online Multi-Agent Reinforcement Learning

DGX agent

arXiv:2604.06610v1 Announce Type: cross Abstract: Decentralised online learning enables runtime adaptation in cyber-physical multi-agent systems, but when operating conditions change, learned policies

safetyarxiv-cs-ai
10 Apr 2026
Model Releases

VenusBench-Mobile: A Challenging and User-Centric Benchmark for Mobile GUI Agents with Capability Diagnostics

DGX agent

arXiv:2604.06182v1 Announce Type: cross Abstract: Existing online benchmarks for mobile GUI agents remain largely app-centric and task-homogeneous, failing to reflect the diversity and instability of

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

I just gave a workshop at @aiDotEngineer in London on building real multi-agent systems. The best part was hearing people laugh, interrupt u…

DGX agent

I just gave a workshop at @aiDotEngineer in London on building real multi-agent systems. The best part was hearing people laugh, interrupt us with questions.. You could feel they were following, think

model-releasesswyx--x
9 Apr 2026
Agents

Another banger article from the @LangChain team! Harness evolution combined with specialist local models will be the way forward undoubtedly…

DGX agent

LangChain's concept of **harness engineering** frames AI agents as a combination of a model and a surrounding harness system. An agent equals a model plus a harness — harness engineering is how sy...

agentsharrison-chase--x
8 Apr 2026
Agents

Coming soon: WorldSim in your Hermes Agent?

DGX agent

Coming soon: WorldSim in your Hermes Agent? putting this skill out soon~ in hermes agent, i can create a new kind of worldsim-based instance for higher fidelity, narrower simulations here i had it try

agentsnous-research--x
7 Apr 2026
Agents

having an event stream like activegraph is the base of how we move to next level and that will probably be composition, which can let you ac…

DGX agent

having an event stream like activegraph is the base of how we move to next level and that will probably be composition, which can let you achieve better results with smaller models. - get the core eve

agentsyohei-nakajima--x
24 Jul 2026
Agents

Escher-Loop: Mutual Evolution by Closed-Loop Self-Referential Optimization

DGX agent

arXiv:2604.23472v1 Announce Type: new Abstract: While recent autonomous agents demonstrate impressive capabilities, they predominantly rely on manually scripted workflows and handcrafted heuristics, i

agentsarxiv-cs-ai
28 Apr 2026
Model Releases

Defending Your Enterprise When AI Models Can Find Vulnerabilities Faster Than Ever

DGX agent

Introduction Advances in AI model-powered exploitation have demonstrated that general-purpose AI models can excel at vulnerability discovery, even without being purpose-built for the task. Eventually,

model-releasesgoogle-cloud-ai
16 Apr 2026
Agents

Welcome to the singularity 😎

DGX agent

Welcome to the singularity 😎 hermes agent from @NousResearch is the fastest growing agent of all time. @OpenClaw went from 0 → 40K stars in 61 days. hermes did it in 45 days. in the past 7 days alone,

agentsnous-research--x
10 Apr 2026
Model Releases

Writer launches major agentic AI improvements with Palmyra X6 flagship model

DGX agent

Generative artificial intelligence startup Writer Inc. today announced the release of its next-generation flagship model, Palmyra X6, designed to deliver frontier-level performance for marketing and r

model-releasessiliconangle
13 Aug 2026
Agents

ChemWorld: Programmable Chemical Worlds for Controlled and Replayable Agent Experimentation

DGX agent

arXiv:2608.10792v1 Announce Type: new Abstract: Autonomous chemistry increasingly depends on environments in which agents can repeatedly act, observe, and adapt.Physical laboratories provide essential

agentsarxiv-cs-ai
12 Aug 2026
Model Releases

DSAgentBench: Can Agents Automate End-to-End Data-Science Workflows in Real Computer Environments?

DGX agent

arXiv:2608.10366v1 Announce Type: new Abstract: Real-world data science involves long-horizon workflows that span data wrangling, exploration, modeling, visualization, and validation, and require coor

model-releasesarxiv-cs-ai
12 Aug 2026
Agents

Pay with confidence: How Solv Labs built verifiable, auditable agent payments on Amazon Bedrock AgentCore payments

DGX agent

Solv Labs built a governed agent-payments workflow on Amazon Bedrock AgentCore payments, where every transaction is authorized, attested in an AWS Nitro Enclave, priced for risk, and anchored to a pub

agentsaws-ml-blog
12 Aug 2026
Agents

Very interesting new work from Anthropic. (bookmark it) They evolve mind viruses, ideas that spread through a multi-agent system by getting …

DGX agent

Very interesting new work from Anthropic. (bookmark it) They evolve mind viruses, ideas that spread through a multi-agent system by getting each host to pass them on, then measure what governs the spr

agentsdair-ai--x
12 Aug 2026
Agents

Business Arena: Benchmarking LLM Agents in a Realistic Marketplace

DGX agent

arXiv:2608.08621v1 Announce Type: new Abstract: Running a business is a challenging form of intelligent work. Operators must infer opportunities from partial signals, commit capital under uncertainty,

agentsarxiv-cs-ai
11 Aug 2026
Agents

FailForge: Distilling Procedural Competence from Persistent Failures into Code Agents

DGX agent

arXiv:2608.08570v1 Announce Type: new Abstract: Rejection sampling fine-tuning (RFT) is widely used to train code agents by generating trajectories on verifiable software engineering tasks, retaining

agentsarxiv-cs-ai
11 Aug 2026
Agents

Mendel Godel Machine: Recursive Self-Improving Coding Agents via Comparative Evolution

DGX agent

arXiv:2608.07645v1 Announce Type: new Abstract: Self-improving coding agents that iteratively rewrite their own source code have demonstrated impressive performance on coding tasks. However, existing

agentsarxiv-cs-ai
11 Aug 2026
Model Releases

MetaSpace: Metamorphic Testing for Spatial Cognition in Embodied Agents

DGX agent

arXiv:2608.07533v1 Announce Type: new Abstract: An embodied agent is an intelligent entity that interacts with its environment through a physical body. Currently, the evaluation of embodied agents pri

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

MoRSE: Task-Oriented Multi-Agent System with Mixture of Role-Subtask Experts

DGX agent

arXiv:2608.09251v1 Announce Type: cross Abstract: Large language model-based multi-agent systems have recently shown strong potential for complex, long-horizon tasks. However, existing methods mainly

model-releasesarxiv-cs-ai
11 Aug 2026
Agents

Not an A11y: How Android Accessibility Exposes Mobile AI Agents to Indirect Prompt Injection

DGX agent

arXiv:2608.08939v1 Announce Type: new Abstract: The rise of autonomous AI agents represents a major paradigm shift in how users interact with mobile devices. Frameworks such as MobileRun and Mobile-Us

agentsarxiv-cs-ai
11 Aug 2026
Model Releases

SkillConsist: Detecting Inconsistencies in Agent Skills via Bidirectional Graph Alignment

DGX agent

arXiv:2608.07639v1 Announce Type: cross Abstract: Agent Skills provide reusable capabilities to LLM agents. Agent Skill inconsistencies can expose undisclosed dangerous behavior or cause wrong Skill s

model-releasesarxiv-cs-ai
11 Aug 2026
Agents

CAi Copilot: Reducing Operational Workload in Molecular Design through Intent-Driven Agentic Workflows

DGX agent

arXiv:2608.06961v1 Announce Type: new Abstract: Early-stage molecular design is an iterative process, not just a task of generating molecules. Researchers turn broad goals into design strategies, refi

agentsarxiv-cs-ai
10 Aug 2026
Model Releases

Fisher-R1: Training LLM Agents for Reliable Hypothesis Testing

DGX agent

arXiv:2608.07437v1 Announce Type: new Abstract: Reliable hypothesis testing is the foundation of many empirical scientific claims. Large language model (LLM) agents are increasingly used to automate t

model-releasesarxiv-cs-ai
10 Aug 2026
Agents

Harnessing the Synergy between LLM Agents and Knowledge Graphs for Urban Socioeconomic Prediction

DGX agent

arXiv:2411.00028v3 Announce Type: replace-cross Abstract: Socioeconomic prediction aims to leverage various urban data to predict the socioeconomic indicators of regions such as population and commerc

agentsarxiv-cs-ai
10 Aug 2026
Agents

Online Security Learning in Cooperative Multi-Agent Systems under Hidden Byzantine Attacks

DGX agent

arXiv:2608.06520v1 Announce Type: new Abstract: We study online cooperative control of a multi-agent system under Byzantine attacks. Namely, an unknown, fixed subset of agents are Byzantine comprised

agentsarxiv-cs-lg
10 Aug 2026
Agents

Trajectory-Relative Hindsight Distillation for Agentic Reinforcement Learning

DGX agent

arXiv:2608.07371v1 Announce Type: cross Abstract: Recent agentic reinforcement learning methods use hindsight to complement sparse outcome rewards. However, a completed rollout can yield many such sig

agentsarxiv-cs-cl
10 Aug 2026
Agents

Imagine image 2.0, non-agentic yet, more to come in a week or two 💙

DGX agent

Imagine image 2.0, non-agentic yet, more to come in a week or two 💙 Announcing Imagine Image 2.0, our next generation image model with precision editing, crisp text rendering, improved factuality, and

agentselon-musk--x
8 Aug 2026
Agents

Agentic Nesting: A New Methodology for Existing Enterprise Application Integration and Services

DGX agent

arXiv:2608.05159v1 Announce Type: new Abstract: Enterprise operations extensively rely on multiple heterogeneous business systems and information applications, which also result in severe data silos a

agentsarxiv-cs-ai
7 Aug 2026
Agents

Beyond Relevance: Bayesian Evidence Acquisition for Agentic Whole-Slide Image Reasoning

DGX agent

arXiv:2608.05757v1 Announce Type: new Abstract: Whole-slide image (WSI) reasoning requires an agent to sequentially acquire visual evidence before answering a diagnostic question. Existing training-fr

agentsarxiv-cs-cv
7 Aug 2026
Agents

Hierarchical Server Architecture for Agentic Science

DGX agent

arXiv:2608.05332v1 Announce Type: cross Abstract: Agentic science is transforming the landscape of computational work, extending to scientific pipelines and workload managers. The workloads require sp

agentsarxiv-cs-ai
7 Aug 2026
Agents

Routing Is Least Learnable Where It Is Most Valuable: Bounds on Representation Routing for Web Agents

DGX agent

arXiv:2608.06171v1 Announce Type: new Abstract: Web agents observe a browser through text, pixels, or both, and the choice is usually fixed once for all tasks. We measure six observation modes across

agentsarxiv-cs-cl
7 Aug 2026
Agents

The Vulnerability With No CVE: Managing Persistent Gaps Between Mandate and Authority in AI Coding Agents

DGX agent

arXiv:2608.05884v1 Announce Type: cross Abstract: Existing guidance identifies excessive agency, excessive permission, weak task-bound authorization, and inadequate agent controls as important risks.

agentsarxiv-cs-cl
7 Aug 2026
Agents

Communication-Enhanced Tutoring for Efficient Decentralized Multi-Agent Reinforcement Learning

DGX agent

arXiv:2508.13661v4 Announce Type: replace Abstract: Centralized Training with Decentralized Execution (CTDE) is the dominant paradigm in multi-agent reinforcement learning (MARL), enabling agents to a

agentsarxiv-cs-lg
6 Aug 2026
Safety

Control agent behaviors and cost beyond a single action: new capabilities in Amazon Bedrock AgentCore

DGX agent

Learn about new capabilities in Amazon Bedrock AgentCore: temporal policies powered by Dogwood, a new open source policy language for AI agents, and rate limiting on the gateway. These features give y

safetyaws-ml-blog
6 Aug 2026
Agents

Fully onboard with productizing hillclimbing as an automated service for any agentic task. I've had the fortune of knowing @silennai since t…

DGX agent

Fully onboard with productizing hillclimbing as an automated service for any agentic task. I've had the fortune of knowing @silennai since the AutoGPT days, and I know that him and Kion are going to d

agentsjerry-liu--x
6 Aug 2026
Agents

State2State: Environment-Derived Mid-Training for LLM Agents

DGX agent

arXiv:2608.04934v1 Announce Type: new Abstract: Training LLM agents commonly relies on supervised fine-tuning from expert trajectories or online reinforcement learning over human-specified tasks with

agentsarxiv-cs-cl
6 Aug 2026
Agents

stratum: A System Infrastructure for Massive Agent-Centric ML Workloads

DGX agent

arXiv:2603.03589v3 Announce Type: replace-cross Abstract: Recent advances in large language models (LLMs) transform how machine learning (ML) pipelines are developed and evaluated. LLMs enable a new t

agentsarxiv-cs-lg
6 Aug 2026
← Previous
1…4849505152…370
Next →