AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,608 results
7 Jul 2026

NKI-Agent: Domain-Specific Fine-Tuning and Agentic Tool Use for Neuron Kernel Generation

Model ReleasesDGX agent

arXiv:2607.04395v1 Announce Type: new Abstract: Recent agentic approaches to LLM-based kernel generation have achieved impressive results on CUDA. For emerging AI accelerators such as AWS Trainium and

3 Jul 2026

Power Systems Agent Benchmark: Executable Evaluation of AI Agents in Electric Power Engineering

Model ReleasesDGX agent

arXiv:2606.20950v2 Announce Type: replace Abstract: Executable evaluation -- checking the consequences of an agent's actions with a program rather than grading its prose -- has become a prominent way

1 Jul 2026

Position: Collaborative Agentic AI Needs Interoperability Across Ecosystems

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
AgentsDGX agent

arXiv:2505.21550v2 Announce Type: replace-cross Abstract: Collaborative agentic AI is projected to transform entire industries by enabling AI-powered agents to autonomously perceive, plan, and act wit

30 Jun 2026

Is Lying an Emergent Behaviour in LLMs? Evidence from Gaslighting AI agents in a Sustainability Game

AgentsDGX agent

arXiv:2606.28456v1 Announce Type: cross Abstract: LLMs agents are increasingly used in multi-agent settings, yet their behaviour in sustainability games remains largely unexplored. This work investiga

LLM agents security duality: a comprehensive survey of self-security and empowered cybersecurity

AgentsDGX agent

arXiv:2606.28450v1 Announce Type: cross Abstract: Large language model (LLM) agents are rapidly being integrated into real-world systems. Their autonomy and tool-use capabilities generate substantial

Why Trust Your Agent? Empirical Security Gains from TRiSM-Guided Agentic Workflows in Healthcare

Model ReleasesDGX agent

arXiv:2606.28666v1 Announce Type: cross Abstract: Agent-based AI has enabled the automation of tasks by exposing application tools and resources to large language models (LLMs). However, to improve sc

24 Jun 2026

IPO Finance Agent: Evaluation of LLM Financial Analysts beyond Finance Agent v2, with Automated Rubric Generation -- the Case of the SpaceX (SPCX) IPO

Model ReleasesDGX agent

arXiv:2606.23032v2 Announce Type: replace Abstract: Finance Agent v2 (by Vals AI) has emerged as the reference benchmark for evaluating both Anthropic Claude and OpenAI ChatGPT frontier language model

23 Jun 2026

Snyk launches Evo Agentic Development Security to police AI coding agents

Model ReleasesDGX agent

Cybersecurity company Snyk Ltd. today launched Evo Agentic Development Security, a new layer of its artificial intelligence security platform built to police the autonomous coding agents that increasi

11 Jun 2026

MoCA-Agent: A Market-of-Claims Code Agent for Financial and Numerical Reasoning

AgentsDGX agent

arXiv:2606.11537v1 Announce Type: new Abstract: Financial and tabular question answering requires more than fluent reasoning: answers must be grounded in the exact facts, formulas, units, signs, and s

9 Jun 2026

Detecting and containing AI-powered threats with Google Security Operations agents

Model ReleasesDGX agent

To defend against the growing range of AI-accelerated threat actors, organizations need to be able to respond faster to outpace the adversary.Recently, we announced Google AI Threat Defense, an automa

4 Jun 2026

Outstanding paper on long-horizon agents. (bookmark it) Similar to humans, how do you make agents persist on a difficult task, and how is th…

Model ReleasesDGX agent

Outstanding paper on long-horizon agents. (bookmark it) Similar to humans, how do you make agents persist on a difficult task, and how is that useful? And which models today work well on this? This ne

Skill issue: Lessons from skilling up coding agents Getting agents to actually use Langfuse was a 'skill issue' — literally. Marc Klingen fr…

ToolsDGX agent

Skill issue: Lessons from skilling up coding agents Getting agents to actually use Langfuse was a 'skill issue' — literally. Marc Klingen from Clickhouse on teaching coding agents to use new tools, an

3 Jun 2026

Agent libOS: A Library-OS-Inspired Runtime for Long-Running, Capability-Controlled LLM Agents

Local AiDGX agent

arXiv:2606.03895v1 Announce Type: cross Abstract: Large language model (LLM) agents are evolving from request-response assistants into long-running software actors: they maintain state across model ca

SAGE: A Quantitative Evaluation of Socialized Evolution in Agent Ecosystems

AgentsDGX agent

arXiv:2606.03544v1 Announce Type: new Abstract: Self-improving language agents are typically evaluated in isolation: an agent attempts a task, receives feedback, and iteratively refines its own behavi

2 Jun 2026

Big paper on AI coding agents using Github & other data The auto-complete tools (Copilot) led to 2.2x more code, local agents like original …

Model ReleasesDGX agent

Big paper on AI coding agents using Github & other data The auto-complete tools (Copilot) led to 2.2x more code, local agents like original Claude Code led to 7.4x, & current remote coding agents 17.3

1 Jun 2026

Very good advice on self-improving agents. (bookmark it) This is something I am seeing in my own experiments with coding agents and harnesse…

Model ReleasesDGX agent

Very good advice on self-improving agents. (bookmark it) This is something I am seeing in my own experiments with coding agents and harnesses for long-horizon tasks. What I have found is that stronger

29 May 2026

Distributed Non-Uniform Scaling Control of Multi-Agent Formation with Dynamic Agent Joining

AgentsDGX agent

arXiv:2605.29191v1 Announce Type: cross Abstract: Non-uniform scaling control of formation enables multi-agent systems to adjust their shape by scaling with different ratios along different coordinate

Paper Agents, Paper Gains: An Empirical Analysis of DeFi Investment Agents

SafetyDGX agent

arXiv:2605.29174v1 Announce Type: new Abstract: DeFi investment agents, systems that use AI for autonomous on-chain trading, have attained over USD 3 billion in combined token valuations since late 20

28 May 2026

Gamma-World: Generative Multi-Agent World Modeling Beyond Two Players

Model ReleasesDGX agent

arXiv:2605.28816v1 Announce Type: new Abstract: World models for interactive video generation have largely focused on single-agent settings, where future observations are generated from a single contr

21 May 2026

Multi-agent Collaboration with State Management

AgentsDGX agent

arXiv:2605.20563v1 Announce Type: cross Abstract: Recent advances in multi-agent systems have shown great potential for solving complex tasks. However, when multiple agents edit a shared codebase conc

20 May 2026

ARM: Discovering Agentic Reasoning Modules for Generalizable Multi-Agent Systems

AgentsDGX agent

arXiv:2510.05746v2 Announce Type: replace Abstract: Large Language Model (LLM)-powered Multi-agent systems (MAS) have achieved state-of-the-art results on various complex reasoning tasks. Recent works

19 May 2026

ADR: An Agentic Detection System for Enterprise Agentic AI Security

Model ReleasesDGX agent

arXiv:2605.17380v1 Announce Type: new Abstract: We present the Agentic AI Detection and Response (ADR) system, the first large-scale, production-proven enterprise framework for securing AI agents oper

Agent Bazaar: Enabling Economic Alignment in Multi-Agent Marketplaces

SafetyDGX agent

arXiv:2605.17698v1 Announce Type: new Abstract: The deployment of Large Language Models (LLMs) as autonomous economic agents introduces systemic risks that extend beyond individual capability failures

MementoGUI: Learning Agentic Multimodal Memory Control for Long-Horizon GUI Agents

Local AiDGX agent

arXiv:2605.18652v1 Announce Type: new Abstract: Recent GUI agents have made substantial progress in visual grounding and action prediction, yet they remain brittle in long-horizon tasks that require m

18 May 2026

maybe you can’t just tack statefulness onto an agent, you have to figure out how to represent the agent as a state

AgentsDGX agent

This post explores the architectural challenge of implementing state management in AI agents, arguing that statefulness cannot simply be added as an afterthought but requires fundamental redesign of h

What we announced in streaming AI at Next ‘26

Model ReleasesDGX agent

Every device, user, and microservice generates data. Ingesting this data, extracting meaning and insights, and driving business decisions in real time has the potential to deliver transformational bus

12 May 2026

SkillMaster: Toward Autonomous Skill Mastery in LLM Agents

AgentsDGX agent

arXiv:2605.08693v1 Announce Type: new Abstract: Skills provide an effective mechanism for improving LLM agents on complex tasks, yet in existing agent frameworks, their creation, refinement, and selec

9 May 2026

my fave point from here: the earlier you think about your agent as a system that can be measured & improved, the faster you can get a robust…

AgentsDGX agent

my fave point from here: the earlier you think about your agent as a system that can be measured & improved, the faster you can get a robust agent into production This isn’t just a technical thing, it

7 May 2026

Material Database Agent: A Multimodal Agentic Framework for Scientific Literature Mining

AgentsDGX agent

arXiv:2605.04278v1 Announce Type: new Abstract: Materials science workflows rely on structured and unstructured data from the vast body of available scientific literature. However, most of the experim

6 May 2026

A Study of Belief Revision Postulates in Multi-Agent Systems (Extended Version)

AgentsDGX agent

arXiv:2605.02249v1 Announce Type: new Abstract: We investigate the belief revision problem in epistemic planning, i.e., what will be the beliefs of all agents in a multi-agent system after an agent ga

Agentic-imodels: Evolving agentic interpretability tools via autoresearch

Model ReleasesDGX agent

arXiv:2605.03808v1 Announce Type: cross Abstract: Agentic data science (ADS) systems are rapidly improving their capability to autonomously analyze, fit, and interpret data, potentially moving towards

5 May 2026

'To get the most out of agent observability, store feedback with your traces. That is what turns agent traces from logs into a learning syst…

AgentsDGX agent

Agent observability is enhanced by storing feedback alongside execution traces, transforming them from simple logs into a learning system that enables continuous improvement. This approach allows deve

WSO2 launches Agent Manager to help enterprises tame AI agent sprawl

Model ReleasesDGX agent

Open-source technology provider WS02 LLC today announced the launch of WSO2 Agent Manager, an open control plane for artificial intelligence agents that gives enterprises a unified way to identify, go

4 May 2026

Agent Factories for High Level Synthesis: How Far Can General-Purpose Coding Agents Go in Hardware Optimization?

Model ReleasesDGX agent

arXiv:2603.25719v2 Announce Type: replace-cross Abstract: We present an empirical study of how far general-purpose coding agents -- without hardware-specific training -- can optimize hardware designs

Swarm management in agent harnesses: owning long-running agents

AgentsDGX agent

As we have built our own harness management tools internally at Arize, and watched external systems like Devin @cognition start managing other Devins, managed agents at @AnthropicAI and long running T

29 Apr 2026

agents are going and are already changing every industry & vertical, awesome blog from the @MadrigalPharma team on using Deep Agents, Skills…

SafetyDGX agent

agents are going and are already changing every industry & vertical, awesome blog from the @MadrigalPharma team on using Deep Agents, Skills, & LangSmith at the frontier of biopharma it’s an awesome t

Proactive agent that thinks and acts like you. Multiplayer AI Brain for teams. Proper GUI for commanding 50 agents. http://Sauna.ai is all t…

Model ReleasesDGX agent

Proactive agent that thinks and acts like you. Multiplayer AI Brain for teams. Proper GUI for commanding 50 agents. http://Sauna.ai is all three. Sauna goes live today. First 2000 people, use access c

28 Apr 2026

Agentic AI platforms for autonomous training and rule induction of human-human and virus-human protein-protein interactions

AgentsDGX agent

arXiv:2604.23924v1 Announce Type: new Abstract: We instruct an AI agent to construct two separate agentic AI platforms: one for autonomous training of predictive ML models for human-human and virus-hu

Appian adopts MCP protocol and partners with Snowflake to provide more structure and control for AI agents

AgentsDGX agent

Business process automation software company Appian Inc. is stepping up its game in agentic artificial intelligence, announcing major updates to its platform focused on AI-assisted application develop

Complete Cyclic Subtask Graphs for Tool-Using LLM Agents: Flexibility, Cost, and Bottlenecks in Multi-Agent Workflows

Model ReleasesDGX agent

arXiv:2604.22820v1 Announce Type: cross Abstract: Long-horizon tool-using tasks sometimes benefit from revisiting earlier subtasks for recovery and exploration, but added multi-agent workflow flexibil

The FIDO Alliance launches two working groups to establish industry standards for securing AI agent transactions; Google contributes the Agent Payments Protocol (Lily Hay Newman/Wired)

Model ReleasesDGX agent

Lily Hay Newman / Wired: The FIDO Alliance launches two working groups to establish industry standards for securing AI agent transactions; Google contributes the Agent Payments Protocol — AI agents ma

27 Apr 2026

Cost-Effective Communication: An Auction-based Method for Language Agent Interaction

AgentsDGX agent

arXiv:2511.13193v2 Announce Type: replace Abstract: Multi-agent systems (MAS) built on large language models (LLMs) often suffer from inefficient 'free-for-all' communication, leading to exponential t

Lately I've been having fun with running coding agents fully locally. The setup I landed on is: - Pi agent - Gemma 4 26B A4B - Server of cho…

Model ReleasesDGX agent

Lately I've been having fun with running coding agents fully locally. The setup I landed on is: - Pi agent - Gemma 4 26B A4B - Server of choice: LM Studio/Ollama/llama.cpp I wrote a step-by-step guide

25 Apr 2026

Inference that never sleeps, for agents that never stop. 'Why cowork when you can delegate?' That's @DhruvBatra_ on @yutori_ai's new Delegat…

AgentsDGX agent

Inference that never sleeps, for agents that never stop. 'Why cowork when you can delegate?' That's @DhruvBatra_ on @yutori_ai's new Delegate — an always-on agent that monitors, researches, and acts a

24 Apr 2026

Survey on Evaluation of LLM-based Agents

SafetyDGX agent

arXiv:2503.16416v2 Announce Type: replace Abstract: LLM-based agents represent a paradigm shift in AI, enabling autonomous systems to plan, reason, and use tools while interacting with dynamic environ

We’re launching the beta for our new commercial AI product: Sakana Fugu 🐡, a multi-agent orchestration system! Blog: https://sakana.ai/fugu…

AgentsDGX agent

We’re launching the beta for our new commercial AI product: Sakana Fugu 🐡, a multi-agent orchestration system! Blog: https://sakana.ai/fugu-beta Fugu hits SOTA on SWE-Pro, GPQA-D, and ALE-Bench, and h

23 Apr 2026

Forage V2: Knowledge Evolution and Transfer in Autonomous Agent Organizations

AgentsDGX agent

arXiv:2604.19837v1 Announce Type: new Abstract: Autonomous agents operating in open-world tasks -- where the completion boundary is not given in advance -- face denominator blindness: they systematica

Google Cloud databases are being rebuilt for the age of AI agents

AgentsDGX agent

AI agents depend on the quality and accessibility of the data behind them, placing Google Cloud databases at the center of that equation. The agentic data cloud introduced at Google Cloud Next 2026 un

22 Apr 2026

DR-MMSearchAgent: Deepening Reasoning in Multimodal Search Agents

AgentsDGX agent

arXiv:2604.19264v1 Announce Type: new Abstract: Agentic multimodal models have garnered significant attention for their ability to leverage external tools to tackle complex tasks. However, it is obser

Explicit Trait Inference for Multi-Agent Coordination

AgentsDGX agent

arXiv:2604.19278v1 Announce Type: new Abstract: LLM-based multi-agent systems (MAS) show promise on complex tasks but remain prone to coordination failures such as goal drift, error cascades, and misa

How far are we from agents that can self-generate world knowledge? The work proposes an outcome-based reward that measures how much an agent…

Model ReleasesDGX agent

How far are we from agents that can self-generate world knowledge? The work proposes an outcome-based reward that measures how much an agent's self-generated world knowledge actually improves its task

Introducing workspace agents in ChatGPT

Model ReleasesDGX agent

OpenAI introduced workspace agents in ChatGPT, a feature that enables AI assistants to perform tasks and take actions within connected workspace applications and tools. These agents can help automate

OMAC: A Holistic Optimization Framework for LLM-Based Multi-Agent Collaboration

AgentsDGX agent

arXiv:2505.11765v3 Announce Type: replace-cross Abstract: Agents powered by advanced large language models (LLMs) have demonstrated impressive capabilities across diverse complex applications. Recentl

Pay attention to this one, AI devs. This is particularly interesting if you work with long-horizon terminal agents that often drown in their…

AgentsDGX agent

Pay attention to this one, AI devs. This is particularly interesting if you work with long-horizon terminal agents that often drown in their own observations. TACO is a self-evolving framework that au

Rethinking Scale: Deployment Trade-offs of Small Language Models under Agent Paradigms

AgentsDGX agent

arXiv:2604.19299v1 Announce Type: cross Abstract: Despite the impressive capabilities of large language models, their substantial computational costs, latency, and privacy risks hinder their widesprea

Vibe coders are not going to like this. UC San Diego just published the first real field study of experienced developers using AI agents. Th…

AgentsDGX agent

Vibe coders are not going to like this. UC San Diego just published the first real field study of experienced developers using AI agents. They watched 13 of them code in the wild and surveyed 99 more.

21 Apr 2026

AppZen targets accounts payable bottleneck with new AI agent service center

AgentsDGX agent

Autonomous finance platform company AppZen Inc. today launched AP Inbox Service Center, a new accounts payable service that deploys eight prebuilt artificial intelligence agents to automate how financ

GenericAgent: A Token-Efficient Self-Evolving LLM Agent via Contextual Information Density Maximization (V1.0)

AgentsDGX agent

arXiv:2604.17091v1 Announce Type: new Abstract: Long-horizon large language model (LLM) agents are fundamentally limited by context. As interactions become longer, tool descriptions, retrieved memorie

Meet Replit Security Agent - providing comprehensive app security reviews in minutes And you get $5 in credits to try it for a limited time …

AgentsDGX agent

Meet Replit Security Agent - providing comprehensive app security reviews in minutes And you get $5 in credits to try it for a limited time Security Agent’s hybrid static analysis and AI-scanning appr

StepPO: Step-Aligned Policy Optimization for Agentic Reinforcement Learning

Model ReleasesDGX agent

arXiv:2604.18401v1 Announce Type: new Abstract: General agents have given rise to phenomenal applications such as OpenClaw and Claude Code. As these agent systems (a.k.a. Harnesses) strive for bolder

← Previous
1…1314151617…294
Next →