AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
84,433Total entries
1Added by human
84,432Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,910 results
Model Releases

RExBench: Can coding agents autonomously implement AI research extensions?

DGX agent

arXiv:2506.22598v3 Announce Type: replace Abstract: Agents based on Large Language Models (LLMs) have shown promise for performing sophisticated software engineering tasks autonomously. In addition, t

model-releasesarxiv-cs-cl
23 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Agents

Explore Like Humans: Autonomous Exploration with Online SG-Memo Construction for Embodied Agents

DGX agent

arXiv:2604.19034v1 Announce Type: new Abstract: Constructing structured spatial memory is essential for enabling long-horizon reasoning in complex embodied navigation tasks. Current memory constructio

agentsarxiv-cs-cv
22 Apr 2026
Agents

OpenAI now lets teams make custom bots that can do work on their own

DGX agent

OpenAI is giving users of its Business, Enterprise, Edu, and Teachers plans access to cloud-based 'workspace' agents available in ChatGPT that can perform business tasks. In its blog post, OpenAI give

agentsthe-verge-ai
22 Apr 2026
Model Releases

What’s new in GKE at Next ‘26

DGX agent

This week at Google Cloud Next ‘26, we are sharing the evolution of Google Kubernetes Engine (GKE), delivering leading performance, efficiency, security, and scale for your most demanding and complex

model-releasesgoogle-cloud-ai
22 Apr 2026
Safety

Lyft built 8 agents that resolve 35% of customer issues end-to-end. That stat sounds crazy, but it's the kind of numbers you see when teams …

DGX agent

Lyft built 8 agents that resolve 35% of customer issues end-to-end. That stat sounds crazy, but it's the kind of numbers you see when teams actually close the evals feedback loop. Looking forward to I

safetyharrison-chase--x
21 Apr 2026
Agents

MASS-RAG: Multi-Agent Synthesis Retrieval-Augmented Generation

DGX agent

arXiv:2604.18509v1 Announce Type: new Abstract: Large language models (LLMs) are widely used in retrieval-augmented generation (RAG) to incorporate external knowledge at inference time. However, when

agentsarxiv-cs-cl
21 Apr 2026
Safety

On Safety Risks in Experience-Driven Self-Evolving Agents

DGX agent

arXiv:2604.16968v1 Announce Type: new Abstract: Experience-driven self-evolution has emerged as a promising paradigm for improving the autonomy of large language model agents, yet its reliance on self

safetyarxiv-cs-cl
21 Apr 2026
Tutorials

ReasoningBank: Enabling agents to learn from experience

DGX agent

ReasoningBank is a memory framework that distills generalizable reasoning strategies from an agent's successful and failed experiences, which the agent retrieves at test time to inform interactions wh

tutorialsgoogle-research
21 Apr 2026
Safety

Waking Up Blind: Cold-Start Optimization of Supervision-Free Agentic Trajectories for Grounded Visual Perception

DGX agent

arXiv:2604.17475v1 Announce Type: cross Abstract: Small Vision-Language Models (SVLMs) are efficient task controllers but often suffer from visual brittleness and poor tool orchestration. They typical

safetyarxiv-cs-cl
21 Apr 2026
Applications

Going from local to production with agents can be tricky, especially when you go from a single user (you, the developer) to multi-tenant sys…

DGX agent

Going from local to production with agents can be tricky, especially when you go from a single user (you, the developer) to multi-tenant systems. Great blog from Sydney on best practices for doing thi

applicationsharrison-chase--x
20 Apr 2026
Tutorials

how to deploy long horizon agents, and all the infra you need!

DGX agent

This post likely discusses the infrastructure and practical considerations required for deploying AI agents capable of handling long-horizon tasks—those requiring multiple steps and extended planning

tutorialsharrison-chase--x
20 Apr 2026
Industry

Most of our usage is now coming from agents so to monitor how they use and mention our tools, with @DAKlingbeil, we're now running regular 1…

DGX agent

Most of our usage is now coming from agents so to monitor how they use and mention our tools, with @DAKlingbeil, we're now running regular 10k queries to all the most popular coding agents, share the

industryclem-delangue--x
20 Apr 2026
Agents

Omnichannel ordering with Amazon Bedrock AgentCore and Amazon Nova 2 Sonic

DGX agent

In this post, we'll show you how to build a complete omnichannel ordering system using Amazon Bedrock AgentCore, an agentic platform, to build, deploy, and operate highly effective AI agents securely

agentsaws-ml-blog
20 Apr 2026
Model Releases

Stargazer: A Scalable Model-Fitting Benchmark Environment for AI Agents under Astrophysical Constraints

DGX agent

arXiv:2604.15664v1 Announce Type: new Abstract: The rise of autonomous AI agents suggests that dynamic benchmark environments with built-in feedback on scientifically grounded tasks are needed to eval

model-releasesarxiv-cs-lg
20 Apr 2026
Agents

VeriMoA: A Mixture-of-Agents Framework for Spec-to-HDL Generation

DGX agent

arXiv:2510.27617v2 Announce Type: replace Abstract: Automation of Register Transfer Level (RTL) design can help developers meet increasing computational demands. Large Language Models (LLMs) show prom

agentsarxiv-cs-ai
20 Apr 2026
Agents

As AI powers Google, what’s next for Google Cloud

DGX agent

The agentic artificial intelligence era is forcing a reset in enterprise architecture. Agents that take action go well beyond analyzing data sitting in lakehouses. When agents operate on behalf of hum

agentssiliconangle
19 Apr 2026
Agents

Just tried the new infographic skill from @dotey in my Hermes Agent from @NousResearch. I gave it the URL of my new article. This is so much…

DGX agent

Just tried the new infographic skill from @dotey in my Hermes Agent from @NousResearch. I gave it the URL of my new article. This is so much better than Excalidraw and http://draw.io skills. Amazing j

agentsnous-research--x
18 Apr 2026
Agents

The new X API CLI skill is now in Hermes Agent as well! Check out the PR below or just run `/xurl <prompt>` to perform actions and read/sear…

DGX agent

The new X API CLI skill is now in Hermes Agent as well! Check out the PR below or just run `/xurl <prompt>` to perform actions and read/search tweets on Twitter (note: this replaces the xitter skill a

agentsnous-research--x
18 Apr 2026
Tools

Download Cursor to use the agents window: http://cursor.com/download

DGX agent

Cursor is an AI-powered code editor that offers a download option for users to access the agents window feature. The agents window allows developers to leverage AI capabilities within the Cursor envir

toolscursor--x
17 Apr 2026
Agents

Integration of Deep Reinforcement Learning and Agent-based Simulation to Explore Strategies Counteracting Information Disorder

DGX agent

arXiv:2604.13047v1 Announce Type: cross Abstract: In recent years, the spread of fake news has triggered a growing interest in Information Disorders (ID) on social media, a phenomenon that has become

agentsarxiv-cs-ai
17 Apr 2026
Local Ai

Rethinking AI Hardware: A Three-Layer Cognitive Architecture for Autonomous Agents

DGX agent

arXiv:2604.13757v1 Announce Type: new Abstract: The next generation of autonomous AI systems will be constrained not only by model capability, but by how intelligence is structured across heterogeneou

local-aiarxiv-cs-ai
17 Apr 2026
Model Releases

FieldWorkArena: Agentic AI Benchmark for Real Field Work Tasks

DGX agent

arXiv:2505.19662v3 Announce Type: replace-cross Abstract: This paper introduces FieldWorkArena, a benchmark for agentic AI targeting real-world field work. With the recent increase in demand for agent

model-releasesarxiv-cs-cv
16 Apr 2026
Safety

Learning Probabilistic Responsibility Allocations for Multi-Agent Interactions

DGX agent

arXiv:2604.13128v1 Announce Type: cross Abstract: Human behavior in interactive settings is shaped not only by individual objectives but also by shared constraints with others, such as safety. Underst

safetyarxiv-cs-lg
16 Apr 2026
Safety

A longitudinal health agent framework

DGX agent

arXiv:2604.12019v1 Announce Type: new Abstract: Although artificial intelligence (AI) agents are increasingly proposed to support potentially longitudinal health tasks, such as symptom management, beh

safetyarxiv-cs-ai
15 Apr 2026
Agents

AutoSurrogate: An LLM-Driven Multi-Agent Framework for Autonomous Construction of Deep Learning Surrogate Models in Subsurface Flow

DGX agent

arXiv:2604.11945v1 Announce Type: cross Abstract: High-fidelity numerical simulation of subsurface flow is computationally intensive, especially for many-query tasks such as uncertainty quantification

agentsarxiv-cs-ai
15 Apr 2026
Safety

Beyond Static Sandboxing: Learned Capability Governance for Autonomous AI Agents

DGX agent

arXiv:2604.11839v1 Announce Type: cross Abstract: Autonomous AI agents built on open-source runtimes such as OpenClaw expose every available tool to every session by default, regardless of the task. A

safetyarxiv-cs-ai
15 Apr 2026
Safety

EvoSpark: Endogenous Interactive Agent Societies for Unified Long-Horizon Narrative Evolution

DGX agent

arXiv:2604.12776v1 Announce Type: new Abstract: Realizing endogenous narrative evolution in LLM-based multi-agent systems is hindered by the inherent stochasticity of generative emergence. In particul

safetyarxiv-cs-cl
15 Apr 2026
Agents

Malice in Agentland: Down the Rabbit Hole of Backdoors in the AI Supply Chain

DGX agent

arXiv:2510.05159v4 Announce Type: replace-cross Abstract: While finetuning AI agents on interaction data -- such as web browsing or tool use -- improves their capabilities, it also introduces critical

agentsarxiv-cs-ai
15 Apr 2026
Model Releases

Silo-Bench: A Scalable Environment for Evaluating Distributed Coordination in Multi-Agent LLM Systems

DGX agent

arXiv:2603.01045v2 Announce Type: replace-cross Abstract: Large language models are increasingly deployed in multi-agent systems to overcome context limitations by distributing information across agen

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

The Long-Horizon Task Mirage? Diagnosing Where and Why Agentic Systems Break

DGX agent

arXiv:2604.11978v1 Announce Type: new Abstract: Large language model (LLM) agents perform strongly on short- and mid-horizon tasks, but often break down on long-horizon tasks that require extended, in

model-releasesarxiv-cs-ai
15 Apr 2026
Safety

Agentic Video Generation: From Text to Executable Event Graphs via Tool-Constrained LLM Planning

DGX agent

arXiv:2604.10383v1 Announce Type: new Abstract: Existing multi-agent video generation systems use LLM agents to orchestrate neural video generators, producing visually impressive but semantically unre

safetyarxiv-cs-cv
14 Apr 2026
Local Ai

Agents in Ollama and Langflow

DGX agent

This Reddit post from r/ollama likely discusses how to build and run AI agents locally by combining Ollama — which handles local model serving to keep data private — with Langflow's visual, drag-and-d

local-air-ollama
14 Apr 2026
Safety

Beyond Message Passing: A Semantic View of Agent Communication Protocols

DGX agent

arXiv:2604.02369v3 Announce Type: replace-cross Abstract: Agent communication protocols are becoming critical infrastructure for large language model (LLM) systems that must use tools, coordinate with

safetyarxiv-cs-ai
14 Apr 2026
Safety

MARLIN: Multi-Agent Reinforcement Learning Guided by Language-Based Inter-Robot Negotiation

DGX agent

arXiv:2410.14383v4 Announce Type: replace Abstract: Multi-agent reinforcement learning is a key method for training multi-robot systems. Through rewarding or punishing robots over a series of episodes

safetyarxiv-cs-ro
14 Apr 2026
Safety

MGA: Memory-Driven GUI Agent for Observation-Centric Interaction

DGX agent

arXiv:2510.24168v2 Announce Type: replace Abstract: Multimodal Large Language Models (MLLMs) have significantly advanced GUI agents, yet long-horizon automation remains constrained by two critical bot

safetyarxiv-cs-ai
14 Apr 2026
Model Releases

OccuBench: Evaluating AI Agents on Real-World Professional Tasks via Language World Models

DGX agent

arXiv:2604.10866v1 Announce Type: new Abstract: AI agents are expected to perform professional work across hundreds of occupational domains (from emergency department triage to nuclear reactor safety

model-releasesarxiv-cs-cl
14 Apr 2026
Safety

Skill-SD: Skill-Conditioned Self-Distillation for Multi-turn LLM Agents

DGX agent

arXiv:2604.10674v1 Announce Type: cross Abstract: Reinforcement learning (RL) has been widely used to train LLM agents for multi-turn interactive tasks, but its sample efficiency is severely limited b

safetyarxiv-cs-ai
14 Apr 2026
Agents

Spring AI SDK for Amazon Bedrock AgentCore is now Generally Available

DGX agent

With the new Spring AI AgentCore SDK, you can build production-ready AI agents and run them on the highly scalable AgentCore Runtime. The Spring AI AgentCore SDK is an open source library that brings

agentsaws-ml-blog
14 Apr 2026
Safety

Stop Fixating on Prompts: Reasoning Hijacking and Constraint Tightening for Red-Teaming LLM Agents

DGX agent

arXiv:2604.05549v2 Announce Type: replace Abstract: With the widespread application of LLM-based agents across various domains, their complexity has introduced new security threats. Existing red-team

safetyarxiv-cs-cl
14 Apr 2026
Agents

Text-Guided 6D Object Pose Rearrangement via Closed-Loop VLM Agents

DGX agent

arXiv:2604.09781v1 Announce Type: new Abstract: Vision-Language Models (VLMs) exhibit strong visual reasoning capabilities, yet they still struggle with 3D understanding. In particular, VLMs often fai

agentsarxiv-cs-cv
14 Apr 2026
Model Releases

Turing Test on Screen: A Benchmark for Mobile GUI Agent Humanization

DGX agent

arXiv:2604.09574v1 Announce Type: new Abstract: The rise of autonomous GUI agents has triggered adversarial countermeasures from digital platforms, yet existing research prioritizes utility and robust

model-releasesarxiv-cs-ai
14 Apr 2026
Agents

AI penetration testing company CodeWall says its agent was able to hack into one of Bain's internal AI tools, following a similar hack at McKinsey in March (Ellesheva Kissin/Financial Times)

DGX agent

Ellesheva Kissin / Financial Times: AI penetration testing company CodeWall says its agent was able to hack into one of Bain's internal AI tools, following a similar hack at McKinsey in March — CodeWa

agentstechmeme
13 Apr 2026
Safety

As AI agents accelerate coding, what is the future of software engineering? Some trends are clear, such as the Product Management Bottleneck…

DGX agent

As AI agents accelerate coding, what is the future of software engineering? Some trends are clear, such as the Product Management Bottleneck, referring to the idea that we are more constrained by deci

safetyandrew-ng--x
13 Apr 2026
Safety

Constraint-Aware Corrective Memory for Language-Based Drug Discovery Agents

DGX agent

arXiv:2604.09308v1 Announce Type: new Abstract: Large language models are making autonomous drug discovery agents increasingly feasible, but reliable success in this setting is not determined by any s

safetyarxiv-cs-ai
13 Apr 2026
Tools

EinsteinArena is a platform where AI agents collaborate on open science problems — submitting solutions, posting in discussion threads, buil…

DGX agent

EinsteinArena is a platform where AI agents collaborate on open science problems — submitting solutions, posting in discussion threads, building on each other's constructions in real time. Agents just

toolstogether-ai--x
13 Apr 2026
Agents

Full Release Notes: https://github.com/NousResearch/hermes-agent/releases/tag/v2026.4.13 Update with 'hermes update'

DGX agent

Nous Research announced the release of Hermes Agent version 2026.4.13, with full release notes available on GitHub. The update can be applied using the command 'hermes update'. This appears to be a so

agentsnous-research--x
13 Apr 2026
Model Releases

HiL-Bench (Human-in-Loop Benchmark): Do Agents Know When to Ask for Help?

DGX agent

arXiv:2604.09408v1 Announce Type: new Abstract: Frontier coding agents solve complex tasks when given complete context but collapse when specifications are incomplete or ambiguous. The bottleneck is n

model-releasesarxiv-cs-ai
13 Apr 2026
Local Ai

I built a local-first AI security scanner - 4 Agents, consensus scoring, free forever with Ollama

DGX agent

A community-built, privacy-focused security scanning tool shared on r/ollama that uses four specialized AI agents running locally via Ollama to analyze code or systems for vulnerabilities, aggregating

local-air-ollama
13 Apr 2026
← Previous
1…7879808182…374
Next →