AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,736 results
Model Releases

Social Gym and SPaRTan: Benchmarking and Improving LLM Social Reasoning via Multi-Agent Game Tournaments

DGX agent

arXiv:2608.09128v1 Announce Type: cross Abstract: LLM agents are increasingly deployed in multi-agent social settings where they must cooperate, negotiate, and adapt to other agents. Measuring and imp

model-releasesarxiv-cs-ai
11 Aug 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Wix launches Symphony, a new standalone multi-agent system built for business operations

DGX agent

Cloud-based website builder Wix Ltd. today announced the launch of Symphony, a new standalone agentic artificial intelligence platform that proactively learns business values, interests, needs, practi

model-releasessiliconangle
11 Aug 2026
Model Releases

StepJack: Benchmarking Computer-Use Agent Safety Against Multi-Step Indirect Prompt Injection

DGX agent

arXiv:2608.06477v1 Announce Type: cross Abstract: Computer-use agents (CUAs) face a growing threat from indirect prompt injection, where adversarial instructions are planted in the environment such as

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

Very cool idea to have agents design complex systems by searching over the model structure itself. New research from Sakana AI introduces CE…

DGX agent

Very cool idea to have agents design complex systems by searching over the model structure itself. New research from Sakana AI introduces CEDAR, which uses LLM agents to write, simulate, and refine sy

model-releasesdair-ai--x
10 Aug 2026
Agents

AI agentic Internet traffic will obviously VASTLY exceed human usage. Not a close call at all. Cloudflare’s forecast is accurate.

DGX agent

AI agentic Internet traffic will obviously VASTLY exceed human usage. Not a close call at all. Cloudflare’s forecast is accurate. For context, Global bandwidth is somewhere between 2-8 Pbps (2,000-8,0

agentselon-musk--x
9 Aug 2026
Agents

How cheap models changed multi-agent economics

DGX agent

Orchestrator-executor just became the smart default for production agents: an expensive model plans, cheap models execute, and cost per completed task decides the roster. The post How cheap models cha

agentsarize-ai
7 Aug 2026
Agents

Building an agentic app deployer with Amazon Bedrock and AWS Lambda

DGX agent

PDI Technologies built PDI Brew, an agentic platform on AWS where non-technical employees describe a tool in plain English and receive a fully provisioned, multi-tenant web application in seconds. See

agentsaws-ml-blog
6 Aug 2026
Agents

EASy: Towards Efficient LLM-Based Agentic System

DGX agent

arXiv:2608.04588v1 Announce Type: cross Abstract: Agentic systems have emerged as a promising paradigm for solving complex tasks by coordinating specialized LLM-based agents. However, most existing sy

agentsarxiv-cs-ai
6 Aug 2026
Agents

Meta takes on Anthropic and OpenAI with its first AI coding agent, Muse Code

DGX agent

Meta Platforms Inc. is getting more serious in its efforts to challenge leading artificial intelligence labs Anthropic PBC and OpenAI PBC with the release of its first AI coding agent, called Muse Cod

agentssiliconangle
6 Aug 2026
Agents

MetaVideoAgent: Automated Video-Agent Evolution for Long-Form Video Understanding

DGX agent

arXiv:2608.04587v1 Announce Type: new Abstract: Long-form video understanding requires locating sparse, question-relevant evidence in long, multimodal videos. Real-world video distributions differ in

agentsarxiv-cs-cv
6 Aug 2026
Agents

AgentStream: How Well Do Self-Evolving LLM Agents Perform Under Streaming Tasks?

DGX agent

arXiv:2608.00155v1 Announce Type: cross Abstract: Large language model (LLM) agents can self-evolve by continually improving from their own accumulated experience. However, existing studies predominan

agentsarxiv-cs-lg
5 Aug 2026
Safety

BAP-SQL: Budget-Aware Observation Planning for Agentic Text-to-SQL

DGX agent

arXiv:2608.02876v1 Announce Type: new Abstract: Tool-using agents do not merely consume observations: their actions determine what arrives next. In agentic text-to-SQL, a broad query can spend context

safetyarxiv-cs-ai
5 Aug 2026
Agents

Before Reasoning Can Fail: Pre-Evidence Procedural Failures in Agentic RAG

DGX agent

arXiv:2608.02011v2 Announce Type: replace Abstract: Agentic retrieval-augmented generation (RAG) systems can fail before evidence-conditioned reasoning is tested: an agent may retrieve candidate snipp

agentsarxiv-cs-ai
5 Aug 2026
Model Releases

Harness choice is a big deal. So much room to advance and improve results across the board with agent harnesses. Great paper highlighting th…

DGX agent

Harness choice is a big deal. So much room to advance and improve results across the board with agent harnesses. Great paper highlighting this. New research releases DataSpace, a benchmark where data

model-releasesdair-ai--x
5 Aug 2026
Local Ai

How we built an MCP bridge to give our AgentCore-hosted AI agent access to local MCP tools

DGX agent

AI agents on Amazon Bedrock AgentCore run in the cloud, but users' tools and files live on their laptops. Learn how to build a secure MCP bridge that lets a cloud-hosted agent call local MCP servers b

local-aiaws-ml-blog
5 Aug 2026
Agents

TraceCAD: Trace-Guided Repair for Agentic CAD Generation

DGX agent

arXiv:2608.03062v1 Announce Type: new Abstract: LLM-based CAD agents produce executable parametric programs, but their correction loops may lose evidence about satisfied requirements, faulty operation

agentsarxiv-cs-ai
5 Aug 2026
Safety

CRISP: Critical Step Perception for Training Efficient Deep Search Agents

DGX agent

arXiv:2608.01867v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly extended into deep search agents that solve complex questions through multi-step interaction with external

safetyarxiv-cs-cl
4 Aug 2026
Safety

DiffuseAgent-MI: Distributionally-Grounded,Tool-Integrated Self-Evolving Agents for Faithful Visual Reasoning

DGX agent

arXiv:2608.00540v1 Announce Type: new Abstract: Tool-integrated vision-language agents have made remarkable progress on compositional and multi-step visual reasoning. Yet their outputs frequently exhi

safetyarxiv-cs-cv
4 Aug 2026
Agents

MedTextWeaver: Procedural Knowledge Evolution in Agentic Medical Text Editing

DGX agent

arXiv:2602.00740v2 Announce Type: replace Abstract: Medical text editing is essential for improving communication among diverse stakeholders in clinical settings. However, adapting LLM agents to this

agentsarxiv-cs-cl
4 Aug 2026
Agents

When Prompts Control Robots: Prompt Injection Attacks in Multi-Agent Robotic Systems

DGX agent

arXiv:2608.00747v1 Announce Type: new Abstract: Large language models are increasingly integrated into autonomous robotic systems for task planning and control, but this integration exposes them to pr

agentsarxiv-cs-ro
4 Aug 2026
Model Releases

Agentic Harness for Real-World Compilers

DGX agent

arXiv:2603.20075v2 Announce Type: replace-cross Abstract: Compilers are critical to modern computing, yet fixing compiler bugs is difficult. While recent large language model (LLM) advancements enable

model-releasesarxiv-cs-ai
3 Aug 2026
Safety

CAGE: Certified Authorization under Typed-Return Uncertainty for Tool-Using Agents

DGX agent

arXiv:2607.29190v1 Announce Type: new Abstract: Tool-using LLM agents act on typed tool returns, records pairing provenance and categorical fields with numerical values. Runtime permission gates gener

safetyarxiv-cs-ai
3 Aug 2026
Local Ai

Multimodal Reinforcement Learning with Adaptive Verifier for AI Agents

DGX agent

arXiv:2512.03438v3 Announce Type: replace Abstract: Agentic reasoning models trained with multimodal reinforcement learning (MMRL) have become increasingly capable, yet they are almost universally opt

local-aiarxiv-cs-ai
3 Aug 2026
Agents

NeSyFS: A Neuro-symbolic Fast-Slow Thinking Framework for LLM Agent under Partial Observability

DGX agent

arXiv:2607.28942v1 Announce Type: new Abstract: Recently Large Language Models (LLMs) have been increasingly deployed as autonomous agents in applications such as self-reflection, retrieval-augmented

agentsarxiv-cs-ai
3 Aug 2026
Agents

datasette-agent 0.4a0

DGX agent

Release: datasette-agent 0.4a0 New await context.browser_task() mechanism allowing agent tools to run code directly in the user's browser. #33 This is an exciting new capability: it makes it easy for

agentssimon-willison
31 Jul 2026
Agents

FaithEyes: Towards Faithful Tool Use via Multi-Agent Process-Image Verification

DGX agent

arXiv:2607.28225v1 Announce Type: new Abstract: Agentic vision-language models (VLMs), which interleave textual reasoning with explicit tool calls such as cropping and code-based image manipulation, h

agentsarxiv-cs-cv
31 Jul 2026
Model Releases

ORCA-bench: How Ready Are Language Model Agents for Oncall?

DGX agent

arXiv:2607.28545v1 Announce Type: new Abstract: Large language models can write, patch, and search code, but oncall root cause analysis (RCA) demands something different: reasoning over noisy metrics,

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

SARC-DQ: Runtime Data-Quality Gating for Agentic AI: Silent Evidence Defects, the Incompetence Shield, and Downstream-Only Remediation

DGX agent

arXiv:2607.26313v1 Announce Type: cross Abstract: Agentic systems act, so a defect in the evidence they retrieve becomes a wrong action with a currency cost. The most dangerous enterprise defects are

model-releasesarxiv-cs-ai
31 Jul 2026
Model Releases

TREK: A Travel Reasoning and Evaluation Kit for LLM Agents in Complex Trip Planning

DGX agent

arXiv:2607.26977v1 Announce Type: new Abstract: Travel planning is a demanding stress test for tool-using LLM agents: a usable itinerary is a single artifact that must be right along many axes at once

model-releasesarxiv-cs-cl
30 Jul 2026
Agents

Voice Memory for Agentic Speech Recognition

DGX agent

arXiv:2607.26410v1 Announce Type: new Abstract: We present Voice Memory, a inference-only scheme for agentic speech recognition: at stream time, a frozen corrector reads a single per-domain memory.md

agentsarxiv-cs-cl
30 Jul 2026
Model Releases

Authoring Agent Skills: A Software-Engineering Approach

DGX agent

arXiv:2607.25032v1 Announce Type: cross Abstract: Agent Skills are an emerging way to extend large language model agents with reusable procedural knowledge that the agent loads on demand. Anthropic in

model-releasesarxiv-cs-ai
29 Jul 2026
Hardware

Super interesting new work from NVIDIA. (bookmark it) They suggest building agents as Python objects. Very cool idea and I think it could a …

DGX agent

Super interesting new work from NVIDIA. (bookmark it) They suggest building agents as Python objects. Very cool idea and I think it could a lot with agent reliability. More below: Agent development to

hardwaredair-ai--x
29 Jul 2026
Agents

A corrective agentic hybrid RAG and an operations-grounded evaluation for a scientific facility

DGX agent

arXiv:2607.24663v1 Announce Type: cross Abstract: Scientific user facilities accumulate decades of operational knowledge that no single search index covers: electronic logbooks, technical documents, i

agentsarxiv-cs-ai
28 Jul 2026
Model Releases

Beyond Sequential Interaction: Benchmarking Parallel Execution and Coordination for GUI Agents

DGX agent

arXiv:2607.22689v1 Announce Type: new Abstract: Graphical user interface (GUI) agents are systems powered by large multimodal models (LMMs). They perceive screen state and execute user instructions th

model-releasesarxiv-cs-ai
28 Jul 2026
Safety

Moral Hazard in Multi-Agent Language Models

DGX agent

arXiv:2607.23982v1 Announce Type: cross Abstract: Cooperation can fail when socially valuable effort is costly, weakly observable, and mainly benefits others. Drawing on Holmstrom's team moral-hazard

safetyarxiv-cs-ai
28 Jul 2026
Agents

Multi-agent DRL-based Lane Change Decision Model for Cooperative Platooning in Mixed Traffic

DGX agent

arXiv:2601.11809v2 Announce Type: replace Abstract: Connected automated vehicles (CAVs) possess the ability to communicate and coordinate with one another, enabling cooperative platooning that enhance

agentsarxiv-cs-ai
28 Jul 2026
Agents

New from @Reuters: the OpenAI agent that hacked into Hugging Face also compromised code that a customer was running on @Modal. “We’re aware …

DGX agent

New from @Reuters: the OpenAI agent that hacked into Hugging Face also compromised code that a customer was running on @Modal. “We’re aware a Modal customer published an unauthenticated endpoint that

agentsgary-marcus--x
28 Jul 2026
Model Releases

StateAct: Program State, before Pixels, for Long-Horizon Computer-Use Agents

DGX agent

arXiv:2607.22798v1 Announce Type: cross Abstract: Computer-use agents are usually improved by strengthening perception: better models for reading a screenshot and choosing where to click. Yet a screen

model-releasesarxiv-cs-cv
28 Jul 2026
Agents

Multi-Agent Debate and Visual Information Extraction for SeePhys Pro: A 1st-Place Technical Report from ICML 2026 AI4Math Track 3 Challenge

DGX agent

arXiv:2607.21946v1 Announce Type: new Abstract: This technical report presents our approach to Challenge Track~3: SeePhys Pro at the 3rd AI for Math Workshop, where the task is to answer college-level

agentsarxiv-cs-lg
27 Jul 2026
Agents

Powered by Cohere’s secure agentic platform North, Automations enable employees, regardless of their technical abilities, to: 1. Turn comple…

DGX agent

Powered by Cohere’s secure agentic platform North, Automations enable employees, regardless of their technical abilities, to: 1. Turn complex workflows into simple outputs 2. Stay in control at each s

agentscohere--x
27 Jul 2026
Model Releases

SwiftMem: Fast Agentic Memory via Query-aware Indexing

DGX agent

arXiv:2601.08160v2 Announce Type: replace Abstract: Agentic memory systems have become critical for enabling LLM agents to maintain long-term context and retrieve relevant information efficiently. How

model-releasesarxiv-cs-cl
27 Jul 2026
Agents

This release marks a significant step forward in making agentic AI practical, scalable, and deeply integrated into daily business operations…

DGX agent

This release marks a significant step forward in making agentic AI practical, scalable, and deeply integrated into daily business operations. Automations is available today to all North customers. Lea

agentscohere--x
27 Jul 2026
Agents

With Comfy MCP, workflows can be built from anywhere. Describe the idea, let an AI agent build it, and check back when it's done. To try thi…

DGX agent

Comfy MCP enables users to create ComfyUI workflows from any location by simply describing their ideas. An AI agent will automatically build the workflow, after which the user can monitor progress unt

agentscomfyui--x
27 Jul 2026
Agents

AWS EC2 compute evolves to meet agentic AI and physical AI demand

DGX agent

Twenty years on, AWS EC2 compute is meeting demand shaped by agentic AI, physical AI, and customers pushing general-purpose cloud infrastructure into new territory its earliest architects never antici

agentssiliconangle
25 Jul 2026
Agents

In the spirit of transparency, here’s what I asked @OpenAI: • Radical transparency: let’s release the traces from the “rogue” agents so the …

DGX agent

In the spirit of transparency, here’s what I asked @OpenAI: • Radical transparency: let’s release the traces from the “rogue” agents so the entire research community can study what happened. • More ca

agentsgary-marcus--x
25 Jul 2026
Safety

Workload-Aware Caching for Multi-Agent Systems

DGX agent

arXiv:2607.20495v1 Announce Type: new Abstract: Multi-agent systems decompose complex tasks into directed acyclic graphs (DAGs) of specialized agent executions, creating natural opportunities for cach

safetyarxiv-cs-ai
24 Jul 2026
Local Ai

A local-first harness for multi-agent workflows

DGX agent

Hey all, I’ve been working on this in my spare time and finally feel ready to share it outside my own circles. Arbiter is a single binary for running agents locally. I originally built it because I wa

local-air-ollama
23 Jul 2026
Hardware

Agentic AI compute reshapes cloud economics as AMD targets GPU and CPU demand surge

DGX agent

Agentic AI compute has become the defining workload of the cloud’s next era, reshaping how hyperscalers, neo-clouds and enterprises architect their infrastructure. As autonomous agents move into produ

hardwaresiliconangle
23 Jul 2026
← Previous
1…4243444546…370
Next →