AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,608 results
15 Jul 2026

Fin-Analyst at FinMMEval 2026 Task 3: A Live Hybrid Trading Agent with LLM Specialists and Rule-Based Signals

AgentsDGX agent

arXiv:2607.12233v1 Announce Type: cross Abstract: Large language model (LLM) trading agents show promising performance in equity markets, yet remain narrowly focused on US equities with little evidenc

14 Jul 2026

ICYMI: @LangChain Deep Agents on @nvidia Nemotron 3 Ultra. frontier open-model agents at ~10x lower cost than closed. Run on Fireworks, then…

Model ReleasesDGX agent

ICYMI: @LangChain Deep Agents on @nvidia Nemotron 3 Ultra. frontier open-model agents at ~10x lower cost than closed. Run on Fireworks, then post-train it into specialized intelligence you own. https:

10 Jul 2026
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Remember When It Matters: Proactive Memory Agent for Long-Horizon Agents

Model ReleasesDGX agent

arXiv:2607.08716v1 Announce Type: new Abstract: In long-horizon tasks, decision-relevant state is often scattered across an expanding trajectory, while the action agent must surface it and act. As tra

9 Jul 2026

When Agents Go Rogue: Activation-Based Detection of Malicious Behaviors in Multi-Agent Systems

Local AiDGX agent

arXiv:2607.06807v1 Announce Type: cross Abstract: While enabling effective collaboration on complex tasks, LLM-based Multi-Agent Systems (MAS) face critical security challenges due to vulnerabilities

8 Jul 2026

Enterprises don’t just need agents that perform. They need agents they can shape, govern and improve as their business evolves. That’s the i…

Model ReleasesDGX agent

Enterprises don’t just need agents that perform. They need agents they can shape, govern and improve as their business evolves. That’s the idea behind our work with LangChain. Read more: https://nvda.

7 Jul 2026

MechMath Agent Team: LLM Driven Agents for Mathematical Research

AgentsDGX agent

arXiv:2607.04394v1 Announce Type: new Abstract: AI reasoning has become a central focus in contemporary artificial intelligence, largely driven by the success of large language models. However, mathem

3 Jul 2026

Something about this year’s @aiDotEngineer World’s Fair just hit different. Last year was the year of “let the agents rip.” This year was th…

Model ReleasesDGX agent

Something about this year’s @aiDotEngineer World’s Fair just hit different. Last year was the year of “let the agents rip.” This year was the year of realizing that autonomy without structure creates

2 Jul 2026

LiteParse runs anywhere, including inside agent runtimes. To prove it, we built an email processing assistant using @flueai (the agentic fra…

Model ReleasesDGX agent

LiteParse runs anywhere, including inside agent runtimes. To prove it, we built an email processing assistant using @flueai (the agentic framework by @astrodotbuild), @resend webhooks, and @tursodatab

1 Jul 2026

Google named a Leader in 2026 Gartner® Magic Quadrant™ for Analytics and Business Intelligence Platforms for third year in a row

Model ReleasesDGX agent

For the third consecutive year, Google has been recognized as a Leader in the 2026 Gartner® Magic Quadrant™ for Analytics and Business Intelligence Platforms. This recognition comes on the heels of Go

30 Jun 2026

Introducing SWE-Together: a multi-turn benchmark built from real user–agent coding sessions. Coding agents are often benchmarked like exam-t…

Model ReleasesDGX agent

Introducing SWE-Together: a multi-turn benchmark built from real user–agent coding sessions. Coding agents are often benchmarked like exam-takers: given the full spec up front, then graded on the fina

Perforce launches Agentic Gateway to govern AI agents and cut token costs

Model ReleasesDGX agent

Perforce Software Inc. today expanded its Perforce Intelligence lineup with an agentic gateway for managing artificial intelligence agents, an autonomous testing platform driven by natural language an

PolicyGuard: A Dialogue-Grounded Sub-Agent Verifier for Policy Adherence in LLM Agents

Model ReleasesDGX agent

arXiv:2606.29225v1 Announce Type: new Abstract: LLM agents handle user requests on behalf of organizations through tool calls and must follow the company policies stated in their system prompts. Prior

Scaling the Horizon, Not the Parameters: Reaching Trillion-Parameter Performance with a 35B Agent

Model ReleasesDGX agent

arXiv:2606.30616v1 Announce Type: new Abstract: We introduce Agents-A1, a 35B Mixture-of-Experts Agentic Model that reaches trillion-parameter-level performance by scaling the agent horizon. We invest

SwarmX: Agentic Scheduling for Low-Latency Agentic Systems

HardwareDGX agent

arXiv:2606.21401v2 Announce Type: replace-cross Abstract: Agentic AI applications compose multiple model calls and tool executions, creating new scheduling challenges for GPU-CPU clusters. Their infer

29 Jun 2026

Your AI Travel Agent Would Book You a Bullfight: An Agentic Benchmark for Implicit Animal Welfare in Frontier AI Models

Model ReleasesDGX agent

arXiv:2606.18142v3 Announce Type: replace Abstract: AI agents are moving from advisors to actors, booking travel, planning menus, and running procurement on behalf of users. Existing benchmarks for AI

26 Jun 2026

New paper on giving LLM agents experience that improves the weights and stays readable at the same time. Agent-experience methods split into…

Model ReleasesDGX agent

New paper on giving LLM agents experience that improves the weights and stays readable at the same time. Agent-experience methods split into two camps. Externalized natural-language rules stay interpr

23 Jun 2026

BARD-MARL: Byzantine-Agent Detection for Learned Communication in Multi-Agent Reinforcement Learning

SafetyDGX agent

arXiv:2606.20701v1 Announce Type: cross Abstract: Learned communication improves coordination in cooperative multi-agent reinforcement learning, but it also creates a trust problem: a trained policy m

10 Jun 2026

Divide and Cooperate: Role-Decomposed Multi-Agent LLM Training with Cross-Agent Learning Signals

Model ReleasesDGX agent

arXiv:2606.10684v1 Announce Type: cross Abstract: Modern language agents which perform multi-step reasoning have shown strong performance in knowledge-intensive question answering. However, existing a

8 Jun 2026

It's a TRAP! Task-Redirecting Agent Persuasion Benchmark for Web Agents

Model ReleasesDGX agent

arXiv:2512.23128v2 Announce Type: replace-cross Abstract: Web-based agents powered by large language models are increasingly used for tasks such as email management or professional networking. Their r

6 Jun 2026

RAINO: Anchoring Agents in Reality, A Systematic Review and Conceptual Framework for Realism in Agent-Based Modelling

AgentsDGX agent

arXiv:2606.05167v1 Announce Type: cross Abstract: Realism is a central yet seemingly under-theorized concept in Agent-Based Modelling. This paper presents a Systematic Literature Review, aiming to ide

4 Jun 2026

Agent Planning Benchmark: A Diagnostic Framework for Planning Capabilities in LLM Agents

Model ReleasesDGX agent

arXiv:2606.04874v1 Announce Type: new Abstract: Planning is central to LLM agents: before acting, an agent must decompose goals, select tools, reason over constraints, and decide when a task is infeas

From Agent Traces to Trust: Evidence Tracing and Execution Provenance in LLM Agents

SafetyDGX agent

arXiv:2606.04990v1 Announce Type: cross Abstract: Large language model (LLM)-based agents increasingly solve complex tasks by interacting with external tools, retrieval systems, memory modules, enviro

Proof-Carrying Agent Actions: Model-Agnostic Runtime Governance for Heterogeneous Agent Systems

Model ReleasesDGX agent

arXiv:2606.04104v1 Announce Type: cross Abstract: Agent systems execute through runtimes with very different control points: local coding tools, framework SDKs, managed agent platforms, API gateways,

2 Jun 2026

Agentic-J: An AI Agent for Biological Microscopy Image Analysis

AgentsDGX agent

arXiv:2606.02080v1 Announce Type: cross Abstract: Biological image analysis increasingly demands integration across heterogeneous tools, programming environments, and domain knowledge that few researc

BAGEN: Are LLM Agents Budget-Aware?

AgentsDGX agent

arXiv:2606.00198v1 Announce Type: cross Abstract: While agents are increasingly spending more resources, today agent cost is mostly measured only after execution. A Budget-Aware Agent (BAGEN) should t

COMAP: Co-Evolving World Models and Agent Policies for LLM Agents

SafetyDGX agent

arXiv:2606.02372v1 Announce Type: new Abstract: Equipping language agents with world models enables them to anticipate environment dynamics and evaluate candidate actions before execution. However, ex

Where Do Deep-Research Agents Go Wrong? Span-Level Error Localization in Agent Trajectories

Model ReleasesDGX agent

arXiv:2606.02060v1 Announce Type: new Abstract: Deep-research agents solve tasks through long trajectories of search, tool use, evidence inspection, and answer synthesis. Evaluation based on final ans

31 May 2026

😂PewDiePie building his own agent orchestrator and releasing it was not on my 2026 bingo card. Own the agent. Own the harness. It's not tha…

AgentsDGX agent

PewDiePie has developed and released his own AI agent orchestrator, representing a trend toward independent creators building custom AI infrastructure rather than relying on existing platforms. The po

29 May 2026

Catalyst-Agent: Autonomous heterogeneous catalyst screening with an LLM Agent

AgentsDGX agent

arXiv:2603.01311v2 Announce Type: replace Abstract: The discovery of novel catalysts tailored for particular applications is a major challenge for the twenty-first century. Traditional methods for thi

28 May 2026

Long Live the Librarian! A Persistent Search Sub-Agent for Energy-Efficient Multi-Agent Software Engineering Systems

HardwareDGX agent

arXiv:2605.27787v1 Announce Type: cross Abstract: Multi-agent systems (MAS) have substantially advanced autonomous software engineering (SWE), but their growing inference energy demands raise sustaina

🆕The Age of Async Agents: Devin’s 7x PR growth, 80% AI commits, background agents, memory, testing, & Open-Inspect https://www.latent.space…

Local AiDGX agent

🆕The Age of Async Agents: Devin’s 7x PR growth, 80% AI commits, background agents, memory, testing, & Open-Inspect https://www.latent.space/p/cognition @cognition cofounder + CPO @walden_yan and Open-

26 May 2026

Agent-as-Peer-Debriefer: A Multi-Agent Framework with Perspective-Based Refinement for Qualitative Analysis

AgentsDGX agent

arXiv:2605.24600v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used for qualitative data analysis (QDA), yet their outputs often miss the depth and nuance of human analy

Agent World Model: Infinity Synthetic Environments for Agentic Reinforcement Learning

Model ReleasesDGX agent

arXiv:2602.10090v3 Announce Type: replace Abstract: Recent advances in large language model (LLM) have empowered autonomous agents to perform multi-turn interactions with tools and environments. Howev

Agent-X: Evaluating Deep Multimodal Reasoning in Vision-Centric Agentic Tasks

Model ReleasesDGX agent

arXiv:2505.24876v2 Announce Type: replace-cross Abstract: Deep reasoning is fundamental for solving complex tasks, especially in vision-centric scenarios that demand sequential, multimodal understandi

Building agents (like software) is a deeply iterative process --> this is why we try to supply as much easy to use tooling and infra as poss…

AgentsDGX agent

Building agents (like software) is a deeply iterative process --> this is why we try to supply as much easy to use tooling and infra as possible so builders can start on their agent improvement journe

25 May 2026

MARGIN: Runtime Confidence Calibration for Multi-Agent Foundation Model Coordination

AgentsDGX agent

arXiv:2605.22949v1 Announce Type: new Abstract: Foundation model agents increasingly operate in multi-agent deployments where a coordinator must decide which agent's response to trust. The standard ap

22 May 2026

The Blueprint: How Movix fills a gap in dental skills with specialized agentic AI

Model ReleasesDGX agent

Welcome to The Blueprint, a regular feature where we highlight how Google Cloud customers are tackling unique and common challenges across industries using the latest AI and cloud technologies. We hop

21 May 2026

Agent JIT Compilation for Latency-Optimizing Web Agent Planning and Scheduling

AgentsDGX agent

arXiv:2605.21470v1 Announce Type: new Abstract: Computer-use agents (CUA) automate tasks specified with natural language such as 'order the cheapest item from Taco Bell' by generating sequences of cal

20 May 2026

Agentic Trading: When LLM Agents Meet Financial Markets

AgentsDGX agent

arXiv:2605.19337v1 Announce Type: new Abstract: A growing body of work explores how Large Language Models (LLMs) can be embedded in trading systems as agents that perceive market information, retrieve

⚠️👇 🚨Breaking ⚠️ If we can’t make AI agents follow rules, we are screwed. New study from METR reports that “when the agents were faced wit…

SafetyDGX agent

⚠️👇 🚨Breaking ⚠️ If we can’t make AI agents follow rules, we are screwed. New study from METR reports that “when the agents were faced with hard tasks, they routinely violated constraints” This—routin

IR-Agent: Expert-Inspired LLM Agents for Structure Elucidation from Infrared Spectra

AgentsDGX agent

arXiv:2508.16112v2 Announce Type: replace Abstract: Spectral analysis provides crucial clues for the elucidation of unknown materials. Among various techniques, infrared spectroscopy (IR) plays an imp

19 May 2026

Agents for Experiments, Experiments for Agents: A Design Grammar for AI-Enabled Experimental Science

AgentsDGX agent

arXiv:2605.17746v1 Announce Type: new Abstract: AI systems are becoming active participants in organizational and knowledge work. They increasingly interact with humans, coordinate workflows, and oper

NVIDIA-Verified Agent Skills Provide Capability Governance for AI Agents

HardwareDGX agent

NVIDIA agent skills enable specialized agents to deliver targeted capabilities across enterprise workflows , with governance provided through verified, scanned, and signed skills. JFrog AI Catalog aut

Scalable voice agent design with Amazon Nova Sonic: multi-agent, tools, and session segmentation

AgentsDGX agent

In this post, you’ll learn how to use Amazon Nova Sonic, Amazon Bedrock AgentCore, and Strands BidiAgent to build scalable, maintainable voice agents that handle these challenges efficiently, resultin

17 May 2026

build v1 of agent ship it (dogfooding counts) ⭐️ collect tracing data ⭐️ ⭐️⭐️ point agentic compute at data ⭐️⭐️ understand failures at scal…

AgentsDGX agent

This post outlines a development roadmap for an AI agent system, prioritizing initial version building with internal testing (dogfooding), implementing tracing and observability for data collection, a

15 May 2026

Orchard: An Open-Source Agentic Modeling Framework

AgentsDGX agent

arXiv:2605.15040v1 Announce Type: new Abstract: Agentic modeling aims to transform LLMs into autonomous agents capable of solving complex tasks through planning, reasoning, tool use, and multi-turn in

14 May 2026

SHM-Agents: A Generalist-Specialist Integrated Agent System for Structural Health Monitoring

AgentsDGX agent

arXiv:2605.12916v1 Announce Type: cross Abstract: Artificial intelligence is increasingly used to simplify complex tasks. In engineering applications of structural health monitoring (SHM), existing sp

13 May 2026

When you want to move from single agent SQLite on something like QMD, PostgreSQL is a great choice for multi agent and production quality, b…

Model ReleasesDGX agent

When you want to move from single agent SQLite on something like QMD, PostgreSQL is a great choice for multi agent and production quality, but not as snappy. So we made it much more snappy with BM25 &

12 May 2026

Agent-Sentry: Bounding LLM Agents via Execution Provenance

SafetyDGX agent

arXiv:2603.22868v2 Announce Type: replace-cross Abstract: Agentic computing systems, while immensely capable, raise serious security, privacy, and safety concerns. A key issue is that the full set of

Insider Attacks in Multi-Agent LLM Consensus Systems

AgentsDGX agent

arXiv:2605.08268v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly deployed in multi-agent systems where agents communicate in natural language to solve tasks jointly. A k

11 May 2026

The best way to level up from 1 agent => many agents. No more cycling between terminal tabs

Model ReleasesDGX agent

This post discusses strategies for scaling from managing a single AI agent to coordinating multiple agents efficiently, likely addressing workflow challenges and tooling improvements that eliminate th

// The Memory Curse in LLM Agents // (bookmark it) Long histories apparently degrades agents as they become increasingly history-following a…

TutorialsDGX agent

// The Memory Curse in LLM Agents // (bookmark it) Long histories apparently degrades agents as they become increasingly history-following and risk-minimizing. Across 7 LLMs and 4 social dilemma games

9 May 2026

the best teams building agents ship early and iterate quickly you can't just ship an agent and then forget about the key to getting the best…

AgentsDGX agent

Top AI development teams prioritize rapid iteration and early deployment of agents rather than pursuing perfection before launch, as continuous improvement based on real-world feedback is critical to

30 Apr 2026

We built a community repo of AI agent configs (with Ollama setups included) — just hit 888 stars. What's your Ollama agent stack?

Local AiDGX agent

A community repository of AI agent configurations for Ollama has reached 888 GitHub stars, featuring pre-built setups and configurations for running AI agents locally. The project appears to focus on

29 Apr 2026

Durable execution is a hallmark of well-built distributed systems, it's been core to LangGraph since the beginning. In the agents-as-softwar…

AgentsDGX agent

Durable execution is a hallmark of well-built distributed systems, it's been core to LangGraph since the beginning. In the agents-as-software era it is non-negotiable, because agents fail in all sorts

Recursive Multi-Agent Systems

AgentsDGX agent

arXiv:2604.25917v1 Announce Type: cross Abstract: Recursive or looped language models have recently emerged as a new scaling axis by iteratively refining the same model computation over latent states

28 Apr 2026

Agentic Data Engineering with Genie Code and Lakeflow

AgentsDGX agent

This Databricks blog post discusses agentic data engineering capabilities enabled by Genie Code and Lakeflow, which are tools designed to automate and streamline data engineering workflows through AI-

Beyond Single-Agent Alignment: Preventing Context-Fragmented Violations in Multi-Agent Systems

Model ReleasesDGX agent

arXiv:2604.22879v1 Announce Type: cross Abstract: We identify and formalize a novel security risk: Context-Fragmented Violations (CFVs) - a class of policy breaches where individual agent actions appe

first up in the 'going to production' series: durable execution! long running agents need to be able to 1. survive crashes 2. resume after a…

AgentsDGX agent

first up in the 'going to production' series: durable execution! long running agents need to be able to 1. survive crashes 2. resume after an indefinite pause durable execution solves this by checkpoi

With agents on the rise, is the ‘modern’ data stack already legacy infrastructure?

AgentsDGX agent

The modern data stack might already be the new legacy. In response, Google Cloud is rebuilding for a world where AI agents — not humans — are the primary users of data infrastructure, unveiling an age

← Previous
1…1516171819…294
Next →