AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,724 results
17 Apr 2026

Replit Agent got even better at keeping you in your creative flow! It now suggests follow-up tasks using full context of your project to bui…

AgentsDGX agent

Replit Agent got even better at keeping you in your creative flow! It now suggests follow-up tasks using full context of your project to build on your ideas: • new features to build • performance impr

Theory of Mind in Action: The Instruction Inference Task in Dynamic Human-Agent Collaboration

Model ReleasesDGX agent

arXiv:2507.02935v2 Announce Type: replace Abstract: Successful human-agent teaming relies on an agent being able to understand instructions given by a (human) principal. In many cases, an instruction

16 Apr 2026

ExpSeek: Self-Triggered Experience Seeking for Web Agents

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model ReleasesDGX agent

arXiv:2601.08605v2 Announce Type: replace Abstract: Experience intervention in web agents emerges as a promising technical paradigm, enhancing agent interaction capabilities by providing valuable insi

pi-Play: Multi-Agent Self-Play via Privileged Self-Distillation without External Data

AgentsDGX agent

arXiv:2604.14054v1 Announce Type: cross Abstract: Deep search agents have emerged as a promising paradigm for addressing complex information-seeking tasks, but their training remains challenging due t

POINTS-Seeker: Towards Training a Multimodal Agentic Search Model from Scratch

AgentsDGX agent

arXiv:2604.14029v1 Announce Type: new Abstract: While Large Multimodal Models (LMMs) demonstrate impressive visual perception, they remain epistemically constrained by their static parametric knowledg

15 Apr 2026

Agentic Discovery with Active Hypothesis Exploration for Visual Recognition

AgentsDGX agent

arXiv:2604.12999v1 Announce Type: new Abstract: We introduce HypoExplore, an agentic framework that formulates neural architecture discovery for visual recognition as a hypothesis-driven scientific in

Data Fabric: Querying agent traces in BigQuery

AgentsDGX agent

How to join LLM traces with billing, infrastructure, and customer data using Iceberg and BigQuery If you run AI agents in production, you’ve probably run into a simple problem: you... The post Data Fa

Leapwork hands off code validation to AI agents to keep pace with automated software development

AgentsDGX agent

Danish agentic artificial intelligence startup Leapwork ApS has announced the launch of its fully automated Continuous Validation Platform, aiming to help enterprises keep up with the velocity of gene

M2HRI: An LLM-Driven Multimodal Multi-Agent Framework for Personalized Human-Robot Interaction

AgentsDGX agent

arXiv:2604.11975v1 Announce Type: new Abstract: Multi-robot systems hold significant promise for social environments such as homes and hospitals, yet existing multi-robot works treat robots as functio

Oracle says the agentic AI bottleneck isn’t the model — it’s the database

AgentsDGX agent

Enterprise AI deployments are stalling not because agents are hard to build, but because organizations lack the data infrastructure to run them reliably at scale. The shift from chatbots to autonomous

14 Apr 2026

Beyond RAG for Cyber Threat Intelligence: A Systematic Evaluation of Graph-Based and Agentic Retrieval

AgentsDGX agent

arXiv:2604.11419v1 Announce Type: new Abstract: Cyber threat intelligence (CTI) analysts must answer complex questions over large collections of narrative security reports. Retrieval-augmented generat

Building agents locally does not mean they’re ready to deploy in production LangSmith deployments helps with that

TutorialsDGX agent

Building agents locally does not mean they’re ready to deploy in production LangSmith deployments helps with that 🔐 One deployment, isolated data per user. Add custom auth so every user gets their own

ChatInject: Abusing Chat Templates for Prompt Injection in LLM Agents

AgentsDGX agent

arXiv:2509.22830v3 Announce Type: replace Abstract: The growing deployment of large language model (LLM) based agents that interact with external environments has created new attack surfaces for adver

How to make Codex (or any agent) do your work without any instructions (it learns by watching you!). Open-source

AgentsDGX agent

This Reddit post from r/ChatGPT discusses a technique for enabling OpenAI's Codex (or similar AI coding agents) to autonomously replicate a user's workflow by observing and learning from their actions

Tomorrow: Hermes Agent Jam. Nous Research team, presentations, Q&A. Tuesday April 14th, 4PM EST in the Nous Discord

AgentsDGX agent

Nous Research hosted the Hermes Agent Jam event on Tuesday, April 14th at 4PM EST within their Discord server. The event featured presentations and a Q&A session with the Nous Research team, centered

Verify Before You Fix: Agentic Execution Grounding for Trustworthy Cross-Language Code Analysis

AgentsDGX agent

arXiv:2604.10800v1 Announce Type: cross Abstract: Learned classifiers deployed in agentic pipelines face a fundamental reliability problem: predictions are probabilistic inferences, not verified concl

13 Apr 2026

Bayesian Ego-graph Inference for Networked Multi-Agent Reinforcement Learning

Local AiDGX agent

arXiv:2509.16606v5 Announce Type: replace-cross Abstract: In networked multi-agent reinforcement learning (Networked-MARL), decentralized agents must act under local observability and constrained comm

'DOOMSCROLL' By Hermes Agent

AgentsDGX agent

'DOOMSCROLL' is a creative work or project titled after the term for compulsive, anxiety-inducing social media consumption, produced by Hermes Agent, an AI agent associated with Nous Research's Hermes

Import AI 453: Breaking AI agents; MirrorCode; and ten views on gradual disempowerment

SafetyDGX agent

Import AI issue 453 covers research and developments around vulnerabilities in AI agent systems, including methods for breaking or adversarially manipulating AI agents. The issue also features MirrorC

Litmus (Re)Agent: A Benchmark and Agentic System for Predictive Evaluation of Multilingual Models

Model ReleasesDGX agent

arXiv:2604.08970v1 Announce Type: cross Abstract: We study predictive multilingual evaluation: estimating how well a model will perform on a task in a target language when direct benchmark results are

Multi-agent Adaptive Mechanism Design

AgentsDGX agent

arXiv:2512.21794v3 Announce Type: replace-cross Abstract: We study a sequential mechanism design problem in which a principal seeks to elicit truthful reports from multiple rational agents while start

This is one of the most important pieces written about agents this year. Your harness = your memory. If you don’t own the harness, you don’t…

AgentsDGX agent

Harrison Chase, co-founder of LangChain, shared a post emphasizing the critical importance of the 'harness' in AI agent systems, arguing that the harness essentially functions as the agent's memory an

12 Apr 2026

@hwchase17 is right, memory creates lock-in. But not just conversations. Tool registries, hooks, agent conventions -> that's harness memory …

AgentsDGX agent

@hwchase17 is right, memory creates lock-in. But not just conversations. Tool registries, hooks, agent conventions -> that's harness memory too. In my setup I define agent tools once in a shared regis

@hwchase17 is spot on: Agent harnesses are the real foundation now — and they're fused with memory. Give a closed/proprietary harness contro…

AgentsDGX agent

@hwchase17 is spot on: Agent harnesses are the real foundation now — and they're fused with memory. Give a closed/proprietary harness control of your agent's memory (context + personalization) and you

11 Apr 2026

Hot take: I can't see any startup building their critical core operations on Claude Managed Agents or any proprietary harness as investable.…

Model ReleasesDGX agent

Hot take: I can't see any startup building their critical core operations on Claude Managed Agents or any proprietary harness as investable. The past weeks have shown why it's critical to build on top

i agree there should be managed agents, i just think they should be built on open harnesses and open memory standards

AgentsDGX agent

i agree there should be managed agents, i just think they should be built on open harnesses and open memory standards @hwchase17 Good article! I half agree :) the problem is that as agents become used

I told my girlfriend I was talking with my Hermes agent on Whatsapp. She got really happy and a bit confused and asked me since when I have …

AgentsDGX agent

I told my girlfriend I was talking with my Hermes agent on Whatsapp. She got really happy and a bit confused and asked me since when I have a Hermes agent? I told her since this morning, at which poin

Relying on model providers' stateful APIs or harnesses creates lock-in: switching models means losing your agent's memory -- a cost that onl…

AgentsDGX agent

Relying on model providers' stateful APIs or harnesses creates lock-in: switching models means losing your agent's memory -- a cost that only grows as agents get better at learning a big part of agent

10 Apr 2026

AgentOpt v0.1 Technical Report: Client-Side Optimization for LLM-Based Agent

Model ReleasesDGX agent

arXiv:2604.06296v1 Announce Type: cross Abstract: AI agents are increasingly deployed in real-world applications, including systems such as Manus, OpenClaw, and coding agents. Existing research has pr

As agent systems scale, the control layer becomes the main source of complexity. What starts simple turns into coordination and composition …

AgentsDGX agent

As agent systems scale, the control layer becomes the main source of complexity. What starts simple turns into coordination and composition problems. I built langchain-middleware-stack to better struc

Learning Without Losing Identity: Capability Evolution for Embodied Agents

SafetyDGX agent

arXiv:2604.07799v1 Announce Type: new Abstract: Embodied agents are expected to operate persistently in dynamic physical environments, continuously acquiring new capabilities over time. Existing appro

Qualixar OS: A Universal Operating System for AI Agent Orchestration

SafetyDGX agent

arXiv:2604.06392v1 Announce Type: new Abstract: We present Qualixar OS, the first application-layer operating system for universal AI agent orchestration. Unlike kernel-level approaches (AIOS) or sing

Reasoning Graphs: Deterministic Agent Accuracy through Evidence-Centric Chain-of-Thought Feedback

AgentsDGX agent

arXiv:2604.07595v1 Announce Type: cross Abstract: Language model agents reason from scratch on every query: each time an agent retrieves evidence and deliberates, the chain of thought is discarded and

9 Apr 2026

One for product. One for research. II-Agent: prompt to UI, prototype, Figma. II-Commons: live arXiv + PubMed, Policy_CA, faster A2A. (note: …

AgentsDGX agent

The search results did not return any information about the specific tweet or the 'II-Agent' and 'II-Commons' products mentioned in the URL. The X (Twitter) post is not publicly indexed or accessib...

8 Apr 2026

Proud to power @NousResearch's Hermes Agent with MiniMax M2.7, and excited for what we're building together. Try MiniMax M2.7 in Hermes Agen…

AgentsDGX agent

Proud to power @NousResearch's Hermes Agent with MiniMax M2.7, and excited for what we're building together. Try MiniMax M2.7 in Hermes Agent today → https://portal.nousresearch.com #MiniMax #NousRese

12 Aug 2026

Agentic AI infrastructure shifts enterprise focus from model choice to platform control

AgentsDGX agent

As agentic AI infrastructure moves from experimentation into production, enterprises are confronting a more complex question than which model to use: how to control the cost, data exposure and infrast

Bandwidth-Efficient Multi-Agent Communication through Information Bottleneck and Vector Quantization

AgentsDGX agent

arXiv:2602.02035v2 Announce Type: replace-cross Abstract: Multi-agent reinforcement learning systems deployed in real-world robotics applications face severe communication constraints that significant

On Understanding, Identifying, and Mitigating Vulnerabilities in Agentic Large Language Models

AgentsDGX agent

arXiv:2608.10530v1 Announce Type: cross Abstract: Large Language Models (LLMs) have undergone a shift from stateless conversational interfaces to autonomous agents capable of multi-step planning, tool

VibeLifeBench: Can Your Life Agent Be Proactive and Persistent in a Living World?

Model ReleasesDGX agent

arXiv:2608.10875v1 Announce Type: cross Abstract: Large language model (LLM) agents are increasingly deployed as personal assistants. Existing evaluations, however, mostly use short, self-contained re

10 Aug 2026

Coupling Planning with Episodic Memory in LLM Agents for Software Issue Resolution

Model ReleasesDGX agent

arXiv:2608.06811v1 Announce Type: cross Abstract: Resolving a real software issue with a large language model (LLM) agent is a long repair episode, often tens to hundreds of steps spanning exploration

How nOps shipped FinOps agents 75% faster with Amazon Bedrock AgentCore

AgentsDGX agent

nOps rebuilt its Clara FinOps AI agent on Amazon Bedrock AgentCore, replacing a self-managed Amazon EKS stack running LangChain and LangGraph. The move cut time-to-production by 75% (from 10-12 months

7 Aug 2026

Now we have a timeline of the OpenAI accidental attack against Hugging Face

AgentsDGX agent

OpenAI gave a last-minute presentation at the Black Hat security on Wednesday about 'the Hugging Face Incident' (previously on this blog). The video was published yesterday. It's short and information

5 Aug 2026

HyperAgent: Planning and Acting over Tool-Schema Hypergraphs for Tool-Use LLM Agents

AgentsDGX agent

arXiv:2608.02650v1 Announce Type: new Abstract: Large language model (LLM) agents increasingly rely on external tools to complete complex real-world tasks. However, reliable tool-use planning remains

People continue to turn to outside AI education resources. Companies just aren't providing enough. I'm running a free AI Agent Workshop in a…

AgentsDGX agent

People continue to turn to outside AI education resources. Companies just aren't providing enough. I'm running a free AI Agent Workshop in a week with @mcuban for absolute beginners. 10,000+ people ha

Prime Agent - a new coding harness surpassing Codex/CC/PI

Model ReleasesDGX agent

Prime Agent is an open-source coding and research agent for general and long-running work. A self-improving RLM harness for coding and long-running autonomous tasks. Designed to be both token-efficien

The Agent Operating System (AOS): A Reference Operating Architecture for Distributed Agentic Systems

SafetyDGX agent

arXiv:2608.03214v1 Announce Type: new Abstract: Large language models have transformed artificial intelligence from isolated prediction services into components of long-running, distributed systems th

VeriTrace: Human-Like Temporal Exploration Completes Agentic Action Space

Model ReleasesDGX agent

arXiv:2608.02878v1 Announce Type: new Abstract: Large language models have shown promise for automated Verilog RTL generation, yet state-of-the-art multi-agent systems plateau at ~95% accuracy on stan

4 Aug 2026

Bayesian and Motivated Reasoning in AI Agents

AgentsDGX agent

arXiv:2608.00339v1 Announce Type: cross Abstract: AI agents increasingly perform open-ended tasks in settings where their conclusions can guide consequential decisions. We provide evidence that AI age

Progressive Agent Skill Generation via Reinforcement Learning

AgentsDGX agent

arXiv:2608.01678v1 Announce Type: cross Abstract: Existing skill generation methods largely rely on heuristics or pipeline-style consolidation, which must be specially designed for different evidence

RoMeRL: Balancing Feedback Coverage and the Memory-Reward Trap in Self-Evolving Agent Memory via Reduced-Order Utility States

AgentsDGX agent

arXiv:2608.02508v1 Announce Type: cross Abstract: Learning-based memory systems for self-evolving LLM agents face two tightly coupled challenges. First, trajectory-indexed utilities grow with the inte

ScrambleToolBench: Agents Search Exhaustively Even When Their Own Map Points to the Next Step

Model ReleasesDGX agent

arXiv:2608.02358v1 Announce Type: new Abstract: To operate robustly in open-world environments, autonomous agents should be able to infer the behavior of unfamiliar systems through interaction alone,

SWE-Touch: Benchmarking Coding Agents When Users Touch the Code

AgentsDGX agent

arXiv:2608.02499v1 Announce Type: cross Abstract: Real-world software development requires coding agents to operate in shared workspaces where users may inspect and modify code during an ongoing task,

V-Mem: Modality-Routed Retrieval for Long-Term Multimodal Agentic Memory

AgentsDGX agent

arXiv:2608.01543v1 Announce Type: cross Abstract: Interaction between users and LLM agents is increasingly multimodal: conversations interleave text with images, and a later question may target either

3 Aug 2026

AgenticRepair: Multi-Faceted Program Context Engineering for Agentic Vulnerability Repair

AgentsDGX agent

arXiv:2607.29422v1 Announce Type: cross Abstract: Automated vulnerability repair aims to reduce the time and effort required to patch security flaws from a vulnerability triage report. Recent agentic

From weeks to minutes: How Formula 1® uses agentic AI on AWS to accelerate data operations

AgentsDGX agent

Formula 1® partnered with AWS to build the Data Accelerator, using agentic AI on Amazon Bedrock AgentCore to transform its MarTech data platform. Learn how F1 cut data source onboarding from up to 8 w

31 Jul 2026

EMBL AI Librarian: Life-Sciences Knowledge Layer for AI Agents

Model ReleasesDGX agent

arXiv:2607.28229v1 Announce Type: new Abstract: The web is increasingly accessed by AI agents rather than humans. Every agent needs knowledge, especially in the life-sciences, where agentic pipelines

What’s new in AI infrastructure and orchestration this month

Model ReleasesDGX agent

At Google, AI is a soup-to-nuts endeavor. Obviously, we make leading AI models like Gemini and Nano Banana. We incorporate AI into the tools you use every day (think Gmail, BigQuery, AlloyDB, Google C

30 Jul 2026

Four Ways to Deploy More Secure AI Agents

HardwareDGX agent

NVIDIA’s AI Red Team found common failure modes in enterprise AI agents: weak access controls, unrestricted code execution via tools, unprotected network egress, and exposure of plaintext secrets. The

Living-Harness Is an Interactive-Agent Evolver

AgentsDGX agent

arXiv:2607.26598v1 Announce Type: cross Abstract: Large language model (LLM) agents may recover from a failure within an episode or after a retry, yet the same execution failure can recur in later tas

SkillCAT: Contrastive, Assessment-Augmented and Topology-AwareSkill Self-Evolution for LLM Agents

AgentsDGX agent

arXiv:2606.13317v2 Announce Type: replace Abstract: Skill self-evolution methods for LLM agents aim to turn execution trajectories into reusable skill documents. However, current pipelines typically d

← Previous
1…2223242526…296
Next →