AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,919 results
Agents

Adaptive Auto-Harness: Sustained Self-Improvement for Agentic System Deployment on Open-Ended Task Streams

DGX agent

arXiv:2606.01770v1 Announce Type: cross Abstract: Auto-harness systems such as A-Evolve, GEPA, and Meta-Harness improve LLM agents by optimizing prompts, skills, tools, memories, and supporting infras

agentsarxiv-cs-ai
2 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Agentic Transformers Provably Learn to Search via Reinforcement Learning

DGX agent

arXiv:2606.00183v1 Announce Type: cross Abstract: Tree search is a central abstraction behind many language-agent reasoning and decision-making tasks: agents must explore actions, remember failures, a

safetyarxiv-cs-ai
2 Jun 2026
Agents

AI agents, open data and governance take center stage at Snowflake Summit

DGX agent

Snowflake Inc. is using its Summit 2026 conference today in San Francisco to present a vision of what it calls the “agentic enterprise,” unveiling a broad set of products and enhancements that it says

agentssiliconangle
2 Jun 2026
Model Releases

AutoMedBench: Towards Medical AutoResearch with Agentic AI Models

DGX agent

arXiv:2606.01961v1 Announce Type: new Abstract: Autonomous agents are increasingly expected to support end-to-end medical-AI research workflows, moving beyond isolated prediction tasks or short-form c

model-releasesarxiv-cs-ai
2 Jun 2026
Agents

AWS adds database features and license options aimed at simplifying agent deployment

DGX agent

Amazon Web Services Inc. today enhanced its database services to simplify the process of building and operating agentic artificial intelligence applications while also lowering the barriers to cloud m

agentssiliconangle
2 Jun 2026
Model Releases

BADGER: Bridging Agentic and Deterministic Evaluation for Generative Enterprise Reasoning

DGX agent

arXiv:2606.02109v1 Announce Type: new Abstract: Enterprise AI systems that translate natural language into SQL queries and orchestrate multi-step agentic reasoning pipelines require evaluation approac

model-releasesarxiv-cs-ai
2 Jun 2026
Safety

Beyond End-to-End Video Models: An LLM-Based Multi-Agent System for Educational Video Generation

DGX agent

arXiv:2602.11790v2 Announce Type: replace Abstract: Although recent end-to-end video generation models demonstrate impressive performance in visually oriented content creation, they remain limited in

safetyarxiv-cs-ai
2 Jun 2026
Model Releases

BraveGuard: From Open-World Threats to Safer Computer-Use Agents

DGX agent

arXiv:2606.01166v1 Announce Type: cross Abstract: Computer-use agents extend language models from text generation to sustained interaction with files, terminals, browsers, and external tools. This shi

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

Context Matters: Repository-Aware Security Analysis of the Agent Skill Ecosystem

DGX agent

arXiv:2603.16572v2 Announce Type: replace-cross Abstract: Agent skills extend local AI agents, such as Claude Code and OpenClaw, with additional functionality. Their growing popularity has led to dedi

model-releasesarxiv-cs-ai
2 Jun 2026
Agents

Diversity Over Frequency: Rethinking Tool Use in Visual Chain-of-Thought Agents

DGX agent

arXiv:2606.00096v1 Announce Type: cross Abstract: Visual agents employ external visual tools within visual chains of thought to incorporate fine-grained evidence. While prior work has mainly studied t

agentsarxiv-cs-ai
2 Jun 2026
Agents

Early research with @LangChain Labs on building more efficient verifiers for agent work product.

DGX agent

LangChain Labs is conducting research into developing more efficient verification systems for validating the outputs and work products of AI agents. This research likely focuses on improving the speed

agentsharrison-chase--x
2 Jun 2026
Agents

Efficient verification is what makes scaling legal agents practical. Excited to partner with @hwchase17 and the @LangChain Labs team on desi…

DGX agent

Efficient verification is what makes scaling legal agents practical. Excited to partner with @hwchase17 and the @LangChain Labs team on designing efficient verifiers - sharing early results showing op

agentsharrison-chase--x
2 Jun 2026
Agents

GitHub unveils a GitHub Copilot desktop app in technical preview, which introduces a new feature called canvases for bidirectional work between users and agents (Mario Rodriguez/The GitHub Blog)

DGX agent

Mario Rodriguez / The GitHub Blog: GitHub unveils a GitHub Copilot desktop app in technical preview, which introduces a new feature called canvases for bidirectional work between users and agents — At

agentstechmeme
2 Jun 2026
Model Releases

HLL: Can Agents Cross Humanity's Last Line of Verification?

DGX agent

arXiv:2606.02449v1 Announce Type: new Abstract: Multimodal agents are increasingly expected to operate interfaces on behalf of users, raising a central deployment question: can they truly substitute f

model-releasesarxiv-cs-ai
2 Jun 2026
Agents

How Generation Architecture Shapes Code Complexity in Multi-Agent LLM Systems: A Paired Study on HumanEval

DGX agent

arXiv:2606.00308v1 Announce Type: cross Abstract: Large-language-model code generation has shifted from single-shot prompting to multi-agent orchestrations - analyst, coder, tester, and debugger pipel

agentsarxiv-cs-ai
2 Jun 2026
Agents

Introducing Devin Desktop. Manage fleets of local and cloud agents from one surface. Plan, delegate, review, and ship without leaving your e…

DGX agent

Devin Desktop is a unified management interface from Cognition AI that enables users to coordinate multiple AI agents operating across local and cloud environments. The platform streamlines workflows

agentscognition-ai--x
2 Jun 2026
Agents

Microsoft debuts an expansion of its model families and agentic AI intelligence for developers

DGX agent

Microsoft Corp. announced an expansion to its artificial intelligence models and agentic AI infrastructure today that brings more data and context into the hands of developers and business users as th

agentssiliconangle
2 Jun 2026
Safety

Network Distributed Multi-Agent Reinforcement Learning for Consensus Control of Quadcopters

DGX agent

arXiv:2606.02107v1 Announce Type: cross Abstract: This paper proposes a Network Distributed Multi-Agent Reinforcement Learning (ND-MARL) framework for quadcopter consensus control. Compared to convent

safetyarxiv-cs-ai
2 Jun 2026
Agents

Rashomon Memory: Towards Argumentation-Driven Retrieval for Multi-Perspective Agent Memory

DGX agent

arXiv:2604.03588v3 Announce Type: replace Abstract: AI agents operating over extended time horizons accumulate experiences that serve multiple concurrent goals, and must often maintain conflicting int

agentsarxiv-cs-ai
2 Jun 2026
Agents

Recognize Your Orchestrator: An Entropy Dynamics Perspective for LLM Multi-Agent Systems

DGX agent

arXiv:2606.01351v1 Announce Type: new Abstract: The transition from single-turn models to Multi-Agent Systems (MAS) promises enhanced problem-solving capabilities, yet the centralized orchestration to

agentsarxiv-cs-ai
2 Jun 2026
Model Releases

SeClaw: Spec-Driven Security Task Synthesis for Evaluating Autonomous Agents

DGX agent

arXiv:2606.02302v1 Announce Type: cross Abstract: Autonomous LLM agents increasingly operate in stateful environments where they access tools, files, memory, and external services. While such capabili

model-releasesarxiv-cs-ai
2 Jun 2026
Research

Simulating Macroeconomic Expectations in Survey Experiments with LLM-based Economic Agents

DGX agent

arXiv:2505.17648v5 Announce Type: replace-cross Abstract: We introduce a framework for simulating macroeconomic expectations in survey experiments using LLM-based economic agents (LLM Agents). We cons

researcharxiv-cs-ai
2 Jun 2026
Model Releases

SPADE-Bench: Evaluating Spontaneous Strategic Deception in Agents via Plan-Action Divergence

DGX agent

arXiv:2606.02380v1 Announce Type: cross Abstract: As LLM-based agents expand their operational scope, reliability becomes a prerequisite for real-world deployment. However, in practical applications,

model-releasesarxiv-cs-ai
2 Jun 2026
Agents

Streaming tokens is now buttery smooth on @telegram with Hermes Agent.

DGX agent

Nous Research announced improvements to token streaming functionality for the Hermes Agent on Telegram, enhancing the smoothness and performance of real-time token generation. This update likely addre

agentsnous-research--x
2 Jun 2026
Agents

The agent development lifecycle has been manual for too long. We’re building a future where it runs continuously, without manual triggers. W…

DGX agent

The agent development lifecycle has been manual for too long. We’re building a future where it runs continuously, without manual triggers. Where well-understood issue types resolve without human revie

agentsharrison-chase--x
2 Jun 2026
Agents

The quarq agent is built on LangGraph! LangGraph makes it easy to build complex memory systems (quarq is now at the top of the LongMemEval l…

DGX agent

The Quarq agent is constructed using LangGraph, a framework that simplifies the development of complex memory systems. Quarq has achieved a top ranking on the LongMemEval benchmark, demonstrating the

agentsharrison-chase--x
2 Jun 2026
Model Releases

TimeSage-MT: A Multi-Turn Benchmark for Evaluating Agentic Time Series Reasoning

DGX agent

arXiv:2606.01498v1 Announce Type: cross Abstract: Time series data inform critical decisions across many real-world domains. While large language model (LLM) agents can analyze data through natural la

model-releasesarxiv-cs-ai
2 Jun 2026
Agents

Unified Context Evolution for LLM Agents

DGX agent

arXiv:2606.02304v1 Announce Type: new Abstract: LLM-based agents can solve multi-step interactive tasks by combining reasoning with environment feedback, yet each episode starts from the same fixed co

agentsarxiv-cs-cl
2 Jun 2026
Agents

Workday introduces new capabilities for building and verifying AI agents

DGX agent

Workday Inc. today announced new capabilities aimed at providing developers new ways to build on top of its platform using their own tools. During DevCon 2026, the company’s annual developer conferenc

agentssiliconangle
2 Jun 2026
Agents

all agents in the future are going to need to write and execute code LangSmith Sandboxes are GA - try them out today https://docs.langchain.…

DGX agent

all agents in the future are going to need to write and execute code LangSmith Sandboxes are GA - try them out today https://docs.langchain.com/langsmith/sandboxes .@MukilLoganathan’s Interrupt keynot

agentsharrison-chase--x
1 Jun 2026
Agents

Demo2: Multimodal Interactive Hybrid Agent

DGX agent

Demo2 showcases Qwen's multimodal interactive hybrid agent capabilities, likely demonstrating the integration of multiple data types (text, image, audio, video) with interactive features and hybrid pr

agentsqwen--x
1 Jun 2026
Agents

Demo3:Browser Agent

DGX agent

Demo3 showcases Qwen's browser agent capabilities, likely demonstrating an AI system's ability to autonomously interact with web browsers to perform tasks such as navigation, form filling, or informat

agentsqwen--x
1 Jun 2026
Safety

DiTTo: Scalable Order-aware All-in-One Image Restoration Agent

DGX agent

arXiv:2605.30915v1 Announce Type: new Abstract: Real-world images rarely suffer from a single degradation, and the order in which degradations are removed substantially affects the final restoration q

safetyarxiv-cs-cv
1 Jun 2026
Agents

Fleet computer use is now available in LangSmith's APAC instance! You can now give your Fleet agents access to a virtual computer if you're …

DGX agent

LangSmith's APAC instance now supports Fleet computer use, enabling Fleet agents to access virtual computers for enhanced functionality. This feature allows users in the Asia-Pacific region to leverag

agentsharrison-chase--x
1 Jun 2026
Model Releases

From Prompt Injection to Persistent Control: Defending Agentic Harness Against Trojan Backdoors

DGX agent

arXiv:2605.31042v1 Announce Type: cross Abstract: LLM agents are evolving from conversational chatbots to operational tools in real-world workspaces. In local agentic harnesses, an LLM can read and wr

model-releasesarxiv-cs-ai
1 Jun 2026
Agents

In @latentspacepod podcast, I shared my view on video generation, world models, LLMs, agents, continual learning and where the next frontier…

DGX agent

In @latentspacepod podcast, I shared my view on video generation, world models, LLMs, agents, continual learning and where the next frontier is. 1. Video models get most of their intelligence from lan

agentsswyx--x
1 Jun 2026
Agents

Industrializing Prediction-Powered Inference: The GLIDE Library for Reliable GenAI and Agentic Systems Evaluation

DGX agent

arXiv:2605.31278v1 Announce Type: new Abstract: Reliable evaluation of agentic systems requires unbiased estimates with valid uncertainty, but standard practice navigates between costly human annotati

agentsarxiv-cs-ai
1 Jun 2026
Agents

Introducing Search as Code, our new search architecture for AI agents. It writes Python that calls our search stack directly, instead of loo…

DGX agent

Introducing Search as Code, our new search architecture for AI agents. It writes Python that calls our search stack directly, instead of looping through function calls one at a time. Available in the

agentsperplexity--x
1 Jun 2026
Agents

Investigating Detection and Obfuscation of Prompt Injection Attacks Against Software Reverse Engineering AI Agents

DGX agent

arXiv:2605.30677v1 Announce Type: cross Abstract: Agentic software reverse engineering systems are vulnerable to prompt injection attacks placed into the source code of executable binary files. This r

agentsarxiv-cs-ai
1 Jun 2026
Model Releases

Learning Multi-Agent Coordination via Sheaf-ADMM

DGX agent

arXiv:2605.31005v1 Announce Type: new Abstract: We present a differentiable optimization framework for multi-agent coordination. An input is decomposed into overlapping local views, each processed by

model-releasesarxiv-cs-lg
1 Jun 2026
Model Releases

MAVEN: Improving Generalization in Agentic Tool Calling

DGX agent

arXiv:2605.30738v1 Announce Type: new Abstract: Generalization across agentic tool-calling environments remains a central challenge for reliable agentic reasoning systems. Although large language mode

model-releasesarxiv-cs-ai
1 Jun 2026
Agents

SAGE: A Novelty Gate for Efficient Memory Evolution in Agentic LLMs

DGX agent

arXiv:2605.30711v1 Announce Type: cross Abstract: Agentic LLMs must continuously decide whether newly extracted facts should be added, merged with existing memories, or ignored, yet prior work has foc

agentsarxiv-cs-ai
1 Jun 2026
Safety

Stateful Online Monitoring Catches Distributed Agent Attacks

DGX agent

arXiv:2605.31593v1 Announce Type: cross Abstract: Language models can find thousands of severe software vulnerabilities, and agents are increasingly being misused for cyberattacks. To avoid detection,

safetyarxiv-cs-ai
1 Jun 2026
Agents

Stop manually triaging agent failures. Let LangSmith Engine fix it.

DGX agent

LangSmith Engine is a tool designed to automatically diagnose and resolve agent failures, eliminating the need for manual troubleshooting and triage. The feature appears to leverage automated analysis

agentsharrison-chase--x
1 Jun 2026
Agents

The best eval harness for production AI and agents: A comparison

DGX agent

A practical comparison of production AI evaluation harnesses, including what to look for across instrumentation, evaluators, online evals, CI gates, and agent workflows. The post The best eval harness

agentsarize-ai
1 Jun 2026
Model Releases

AgentDoG 1.5: A Lightweight and Scalable Alignment Framework for AI Agent Safety and Security

DGX agent

arXiv:2605.29801v1 Announce Type: new Abstract: Modern open-world agents such as OpenClaw exhibit powerful cross-environment execution capabilities yet introduce broad new safety risk sources. Meanwhi

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

BenchTrace: A Benchmark for Testing Reflection Ability and Controlled Evolution in LLM Agents

DGX agent

arXiv:2605.29225v1 Announce Type: new Abstract: Self-evolving agents improve over time by reflecting on past failures, but existing evaluation is limited in two ways: it measures only task scores, lea

model-releasesarxiv-cs-ai
29 May 2026
Safety

GrepSeek: Training Search Agents for Direct Corpus Interaction

DGX agent

arXiv:2605.29307v1 Announce Type: cross Abstract: Large Language Model (LLM) search agents have shown strong promise for knowledge-intensive language tasks through multiple rounds of reasoning and inf

safetyarxiv-cs-ai
29 May 2026
← Previous
1…8788899091…374
Next →