AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,771 results
15 Apr 2026

AutoSurrogate: An LLM-Driven Multi-Agent Framework for Autonomous Construction of Deep Learning Surrogate Models in Subsurface Flow

AgentsDGX agent

arXiv:2604.11945v1 Announce Type: cross Abstract: High-fidelity numerical simulation of subsurface flow is computationally intensive, especially for many-query tasks such as uncertainty quantification

Beyond Static Sandboxing: Learned Capability Governance for Autonomous AI Agents

SafetyDGX agent

arXiv:2604.11839v1 Announce Type: cross Abstract: Autonomous AI agents built on open-source runtimes such as OpenClaw expose every available tool to every session by default, regardless of the task. A

EvoSpark: Endogenous Interactive Agent Societies for Unified Long-Horizon Narrative Evolution

SafetyDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2604.12776v1 Announce Type: new Abstract: Realizing endogenous narrative evolution in LLM-based multi-agent systems is hindered by the inherent stochasticity of generative emergence. In particul

Malice in Agentland: Down the Rabbit Hole of Backdoors in the AI Supply Chain

AgentsDGX agent

arXiv:2510.05159v4 Announce Type: replace-cross Abstract: While finetuning AI agents on interaction data -- such as web browsing or tool use -- improves their capabilities, it also introduces critical

Silo-Bench: A Scalable Environment for Evaluating Distributed Coordination in Multi-Agent LLM Systems

Model ReleasesDGX agent

arXiv:2603.01045v2 Announce Type: replace-cross Abstract: Large language models are increasingly deployed in multi-agent systems to overcome context limitations by distributing information across agen

The Long-Horizon Task Mirage? Diagnosing Where and Why Agentic Systems Break

Model ReleasesDGX agent

arXiv:2604.11978v1 Announce Type: new Abstract: Large language model (LLM) agents perform strongly on short- and mid-horizon tasks, but often break down on long-horizon tasks that require extended, in

M^star: Every Task Deserves Its Own Memory Harness

AgentsDGX agent

arXiv:2604.11811v1 Announce Type: cross Abstract: Large language model agents rely on specialized memory systems to accumulate and reuse knowledge during extended interactions. Recent architectures ty

14 Apr 2026

Agentic Video Generation: From Text to Executable Event Graphs via Tool-Constrained LLM Planning

SafetyDGX agent

arXiv:2604.10383v1 Announce Type: new Abstract: Existing multi-agent video generation systems use LLM agents to orchestrate neural video generators, producing visually impressive but semantically unre

Agents in Ollama and Langflow

Local AiDGX agent

This Reddit post from r/ollama likely discusses how to build and run AI agents locally by combining Ollama — which handles local model serving to keep data private — with Langflow's visual, drag-and-d

Beyond Message Passing: A Semantic View of Agent Communication Protocols

SafetyDGX agent

arXiv:2604.02369v3 Announce Type: replace-cross Abstract: Agent communication protocols are becoming critical infrastructure for large language model (LLM) systems that must use tools, coordinate with

MARLIN: Multi-Agent Reinforcement Learning Guided by Language-Based Inter-Robot Negotiation

SafetyDGX agent

arXiv:2410.14383v4 Announce Type: replace Abstract: Multi-agent reinforcement learning is a key method for training multi-robot systems. Through rewarding or punishing robots over a series of episodes

MGA: Memory-Driven GUI Agent for Observation-Centric Interaction

SafetyDGX agent

arXiv:2510.24168v2 Announce Type: replace Abstract: Multimodal Large Language Models (MLLMs) have significantly advanced GUI agents, yet long-horizon automation remains constrained by two critical bot

OccuBench: Evaluating AI Agents on Real-World Professional Tasks via Language World Models

Model ReleasesDGX agent

arXiv:2604.10866v1 Announce Type: new Abstract: AI agents are expected to perform professional work across hundreds of occupational domains (from emergency department triage to nuclear reactor safety

Skill-SD: Skill-Conditioned Self-Distillation for Multi-turn LLM Agents

SafetyDGX agent

arXiv:2604.10674v1 Announce Type: cross Abstract: Reinforcement learning (RL) has been widely used to train LLM agents for multi-turn interactive tasks, but its sample efficiency is severely limited b

Spring AI SDK for Amazon Bedrock AgentCore is now Generally Available

AgentsDGX agent

With the new Spring AI AgentCore SDK, you can build production-ready AI agents and run them on the highly scalable AgentCore Runtime. The Spring AI AgentCore SDK is an open source library that brings

Stop Fixating on Prompts: Reasoning Hijacking and Constraint Tightening for Red-Teaming LLM Agents

SafetyDGX agent

arXiv:2604.05549v2 Announce Type: replace Abstract: With the widespread application of LLM-based agents across various domains, their complexity has introduced new security threats. Existing red-team

Text-Guided 6D Object Pose Rearrangement via Closed-Loop VLM Agents

AgentsDGX agent

arXiv:2604.09781v1 Announce Type: new Abstract: Vision-Language Models (VLMs) exhibit strong visual reasoning capabilities, yet they still struggle with 3D understanding. In particular, VLMs often fai

Turing Test on Screen: A Benchmark for Mobile GUI Agent Humanization

Model ReleasesDGX agent

arXiv:2604.09574v1 Announce Type: new Abstract: The rise of autonomous GUI agents has triggered adversarial countermeasures from digital platforms, yet existing research prioritizes utility and robust

13 Apr 2026

AI penetration testing company CodeWall says its agent was able to hack into one of Bain's internal AI tools, following a similar hack at McKinsey in March (Ellesheva Kissin/Financial Times)

AgentsDGX agent

Ellesheva Kissin / Financial Times: AI penetration testing company CodeWall says its agent was able to hack into one of Bain's internal AI tools, following a similar hack at McKinsey in March — CodeWa

As AI agents accelerate coding, what is the future of software engineering? Some trends are clear, such as the Product Management Bottleneck…

SafetyDGX agent

As AI agents accelerate coding, what is the future of software engineering? Some trends are clear, such as the Product Management Bottleneck, referring to the idea that we are more constrained by deci

Constraint-Aware Corrective Memory for Language-Based Drug Discovery Agents

SafetyDGX agent

arXiv:2604.09308v1 Announce Type: new Abstract: Large language models are making autonomous drug discovery agents increasingly feasible, but reliable success in this setting is not determined by any s

EinsteinArena is a platform where AI agents collaborate on open science problems — submitting solutions, posting in discussion threads, buil…

ToolsDGX agent

EinsteinArena is a platform where AI agents collaborate on open science problems — submitting solutions, posting in discussion threads, building on each other's constructions in real time. Agents just

Full Release Notes: https://github.com/NousResearch/hermes-agent/releases/tag/v2026.4.13 Update with 'hermes update'

AgentsDGX agent

Nous Research announced the release of Hermes Agent version 2026.4.13, with full release notes available on GitHub. The update can be applied using the command 'hermes update'. This appears to be a so

HiL-Bench (Human-in-Loop Benchmark): Do Agents Know When to Ask for Help?

Model ReleasesDGX agent

arXiv:2604.09408v1 Announce Type: new Abstract: Frontier coding agents solve complex tasks when given complete context but collapse when specifications are incomplete or ambiguous. The bottleneck is n

I built a local-first AI security scanner - 4 Agents, consensus scoring, free forever with Ollama

Local AiDGX agent

A community-built, privacy-focused security scanning tool shared on r/ollama that uses four specialized AI agents running locally via Ollama to analyze code or systems for vulnerabilities, aggregating

Interactive ASR: Towards Human-Like Interaction and Semantic Coherence Evaluation for Agentic Speech Recognition

AgentsDGX agent

arXiv:2604.09121v1 Announce Type: cross Abstract: Recent years have witnessed remarkable progress in automatic speech recognition (ASR), driven by advances in model architectures and large-scale train

SAGE: A Service Agent Graph-guided Evaluation Benchmark

Model ReleasesDGX agent

arXiv:2604.09285v1 Announce Type: new Abstract: The development of Large Language Models (LLMs) has catalyzed automation in customer service, yet benchmarking their performance remains challenging. Ex

10 Apr 2026

A Physical Agentic Loop for Language-Guided Grasping with Execution-State Monitoring

SafetyDGX agent

arXiv:2604.07395v1 Announce Type: cross Abstract: Robotic manipulation systems that follow language instructions often execute grasp primitives in a largely single-shot manner: a model proposes an act

ATBench: A Diverse and Realistic Agent Trajectory Benchmark for Safety Evaluation and Diagnosis

Model ReleasesDGX agent

arXiv:2604.02022v2 Announce Type: replace Abstract: Evaluating the safety of LLM-based agents is increasingly important because risks in realistic deployments often emerge over multi-step interactions

ClawBench: Can AI Agents Complete Everyday Online Tasks?

Model ReleasesDGX agent

arXiv:2604.08523v1 Announce Type: new Abstract: AI agents may be able to automate your inbox, but can they automate other routine aspects of your life? Everyday online tasks offer a realistic yet unso

EMSDialog: Synthetic Multi-person Emergency Medical Service Dialogue Generation from Electronic Patient Care Reports via Multi-LLM Agents

AgentsDGX agent

arXiv:2604.07549v1 Announce Type: new Abstract: Conversational diagnosis prediction requires models to track evolving evidence in streaming clinical conversations and decide when to commit to a diagno

For anyone building agentic workflows: the real bottleneck isn't the model, it's the harness. Open standards are a must, not just for flexib…

AgentsDGX agent

For anyone building agentic workflows: the real bottleneck isn't the model, it's the harness. Open standards are a must, not just for flexibility, but to ensure we actually own the long-term memory th

Incorporating Social Awareness into Control of Unknown Multi-Agent Systems: A Real-Time Spatiotemporal Tubes Approach

SafetyDGX agent

arXiv:2510.25597v2 Announce Type: replace-cross Abstract: This paper presents a decentralized control framework that incorporates social awareness into multi-agent systems with unknown dynamics to ach

Learning to Search: A Decision-Based Agent for Knowledge-Based Visual Question Answering

AgentsDGX agent

arXiv:2604.07146v2 Announce Type: replace Abstract: Knowledge-based visual question answering (KB-VQA) requires vision-language models to understand images and use external knowledge, especially for r

MemReader: From Passive to Active Extraction for Long-Term Agent Memory

SafetyDGX agent

arXiv:2604.07877v1 Announce Type: new Abstract: Long-term memory is fundamental for personalized and autonomous agents, yet populating it remains a bottleneck. Existing systems treat memory extraction

OrgForge: A Multi-Agent Simulation Framework for Verifiable Synthetic Corporate Corpora

AgentsDGX agent

arXiv:2603.14997v2 Announce Type: replace Abstract: Building and evaluating enterprise AI systems requires synthetic organizational corpora that are internally consistent, temporally structured, and c

PyFi: Toward Pyramid-like Financial Image Understanding for VLMs via Adversarial Agents

AgentsDGX agent

arXiv:2512.14735v2 Announce Type: replace-cross Abstract: This paper proposes PyFi, a novel framework for pyramid-like financial image understanding that enables vision language models (VLMs) to reaso

9 Apr 2026

Cursor can now attach demos and screenshots of its work to PRs it opens. Your team can review artifacts created by cloud agents directly in …

ToolsDGX agent

Cursor cloud agents now run in isolated virtual machines where they can build, test, and verify their own code — then automatically attach artifacts such as videos, screenshots, and logs to the pul...

Gotta love @LangChain! 🥹 Providing high quality open source agent infra DAILY Single handedly advancing the trajectory of AI

AgentsDGX agent

Gotta love @LangChain! 🥹 Providing high quality open source agent infra DAILY Single handedly advancing the trajectory of AI Open Harness, Model Choice, Open Memory (take it wherever you need), Open P

The memory ownership point is the real story here. Running a persistent agent 24/7, memory isn't a feature — it's the entire value layer. Lo…

AgentsDGX agent

The memory ownership point is the real story here. Running a persistent agent 24/7, memory isn't a feature — it's the entire value layer. Lose it and you're back to day zero. Open harness + portable m

The race to deploy AI agents is exposing a critical gap in enterprise data management

ApplicationsDGX agent

As AI agents demand real-time access to live, governed data, database lifecycle management has moved from a back-office task to a strategic imperative for enterprise infrastructure. The complexity of

Zencoder launches AI platform to automate the surrounding work that coding agents don’t handle

Model ReleasesDGX agent

Zencoder, a startup building an artificial intelligence orchestration layer for software development and business, today announced the launch of Zenflow Work, a major expansion of its platform that fo

8 Apr 2026

Improving the academic workflow: Introducing two AI agents for better figures and peer review

ResearchDGX agent

Google Research introduced two AI agents to streamline academic workflows: **PaperVizAgent**, a visualizer agent for drawing academic figures, and **ScholarPeer**, a reviewer agent that automatica...

Introducing the Hermes Ecosystem Map I was an early user of Hermes Agent from @NousResearch and have been a power user ever since But as the…

Model ReleasesDGX agent

Introducing the Hermes Ecosystem Map I was an early user of Hermes Agent from @NousResearch and have been a power user ever since But as the ecosystem has grown, its been hard to keep up, so I did som

11 Aug 2026

we recently trimmed the deepagents harness base prompt by 65% (including tool info) it shows — deepagents is cheap!

Model ReleasesDGX agent

we recently trimmed the deepagents harness base prompt by 65% (including tool info) it shows — deepagents is cheap! We ran DeepSeek V4 Flash through 4 more agent harnesses (Hermes Agent, Pi Agent, Pri

10 Aug 2026

Blast Radius

AgentsDGX agent

arXiv:2608.07440v1 Announce Type: new Abstract: Agentic coding faces growing problems of affordability and wasted tokens. We introduce Blast Radius, a predictive memory management layer that estimates

7 Aug 2026

How TReNDS automates root-cause analysis with Amazon Bedrock

AgentsDGX agent

TReNDS, a research center at Georgia State University, built an agentic AI pipeline on Amazon Bedrock and the open-source Strands Agents SDK that automatically investigates production errors in real t

4 Aug 2026

How Cloudflare enforces engineering standards using AI

AgentsDGX agent

We created the Cloudflare Codex, a governed body of engineering standards that AI agents consume across the development lifecycle. By pairing structured RFCs with agentic reviews, teams automatically

27 Jul 2026

Announcing general availability of SAP Business Data Cloud Connect for BigQuery

Model ReleasesDGX agent

Traditional data replication techniques often struggle to deliver the data freshness that modern workflows require. To help organizations overcome this challenge, SAP and Google Cloud are announcing t

23 Jul 2026

Knowledge-Centric Self-Improvement

AgentsDGX agent

arXiv:2607.19592v1 Announce Type: new Abstract: Self-improving AI systems typically treat the agent as the object that improves, by optimizing prompts, workflows, harnesses, or even the agent's own co

1 Jul 2026

When Regulation Has Memory: Hysteresis and Control Burden in Artificial Agency

AgentsDGX agent

arXiv:2606.30975v1 Announce Type: new Abstract: Adaptive agents are usually judged by what they do, but an agent can appear stable while the internal effort required to keep it stable is increasing. T

10 Jun 2026

Everyone Operating At The Frontier Satya Nadella, Chairman & CEO, Microsoft, interviewed by @saranormous & @eladgil (No Priors) and @swyx (L…

Model ReleasesDGX agent

Everyone Operating At The Frontier Satya Nadella, Chairman & CEO, Microsoft, interviewed by @saranormous & @eladgil (No Priors) and @swyx (Latent Space) Crossover special at Microsoft Build 2026. Summ

28 Apr 2026

The Chameleon's Limit: Investigating Persona Collapse and Homogenization in Large Language Models

AgentsDGX agent

arXiv:2604.24698v1 Announce Type: new Abstract: Applications based on large language models (LLMs), such as multi-agent simulations, require population diversity among agents. We identify a pervasive

27 Apr 2026

AI-native software engineering teams operate very differently than traditional teams. The obvious difference is that AI-native teams use cod…

AgentsDGX agent

AI-native software engineering teams operate very differently than traditional teams. The obvious difference is that AI-native teams use coding agents to build products much faster, but this leads to

QuantClaw: Precision Where It Matters for OpenClaw

AgentsDGX agent

arXiv:2604.22577v1 Announce Type: new Abstract: Autonomous agent systems such as OpenClaw introduce significant efficiency challenges due to long-context inputs and multi-turn reasoning. This results

22 Apr 2026

Commvault brings its full suite of data backup and resilience capabilities to Google Cloud

AgentsDGX agent

Commvault Systems Inc. is deepening its integration with Google LLC’s cloud platform in order to help bolster the defenses of modern “agentic enterprises” and enhance their data resilience. At Google

21 Apr 2026

AdaExplore: Failure-Driven Adaptation and Diversity-Preserving Search for Efficient Kernel Generation

AgentsDGX agent

arXiv:2604.16625v1 Announce Type: new Abstract: Recent large language model (LLM) agents have shown promise in using execution feedback for test-time adaptation. However, robust self-improvement remai

Amplitude introduces contextual AI assistant to guide users

AgentsDGX agent

Amplitude Inc. today introduced an embedded artificial intelligence support agent that helps users navigate digital products without leaving the application. Amplitude AI Assistant uses behavioral dat

20 Apr 2026

Aikido Security debuts Endpoint for AI-native developer security

AgentsDGX agent

Belgian cybersecurity company Aikido Security BV today launched Endpoint, a lightweight security agent designed to secure artificial intelligence use on developer workstations and address supply chain

16 Apr 2026

i feel no shame. zero. none.

AgentsDGX agent

i feel no shame. zero. none. You can now visualize Pi traces that you upload on @huggingface! Let's make sharing agent traces 10x more common to make agent AI more open and collaborative! Also, becaus

← Previous
1…6263646566…297
Next →