AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,771 results
20 Apr 2026

SocialWise: LLM-Agentic Conversation Therapy for Individuals with Autism Spectrum Disorder to Enhance Communication Skills

AgentsDGX agent

arXiv:2604.15347v1 Announce Type: cross Abstract: Autism Spectrum Disorder (ASD) affects more than 75 million people worldwide. However, scalable support for practicing everyday conversation is scarce

The Price of Paranoia: Robust Risk-Sensitive Cooperation in Non-Stationary Multi-Agent Reinforcement Learning

SafetyDGX agent

arXiv:2604.15695v1 Announce Type: cross Abstract: Cooperative equilibria are fragile. When agents learn alongside each other rather than in a fixed environment, the process of learning destabilizes th

18 Apr 2026

Hugging Face is becoming the platform for agents to use and build AI. Now they can call 1M HF spaces to do everything the latest specialized…

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
IndustryDGX agent

Hugging Face is expanding its platform to support AI agents by enabling them to call upon 1 million Hugging Face Spaces, allowing agents to access diverse specialized tools and capabilities for variou

17 Apr 2026

CoopEval: Benchmarking Cooperation-Sustaining Mechanisms and LLM Agents in Social Dilemmas

SafetyDGX agent

arXiv:2604.15267v1 Announce Type: cross Abstract: It is increasingly important that LLM agents interact effectively and safely with other goal-pursuing agents, yet, recent works report the opposite tr

Orak: A Foundational Benchmark for Training and Evaluating LLM Agents on Diverse Video Games

Model ReleasesDGX agent

arXiv:2506.03610v3 Announce Type: replace Abstract: Large Language Model (LLM) agents are reshaping the game industry, by enabling more intelligent and human-preferable characters. Yet, current game b

Shared Dictionaries: compression that keeps up with the agentic web

AgentsDGX agent

Shared Dictionaries is a web compression technology developed by Cloudflare that enables more efficient data transmission by allowing browsers and servers to reference common dictionary files rather t

16 Apr 2026

Cloudflare’s AI Platform: an inference layer designed for agents

IndustryDGX agent

We're building AI Gateway into a unified inference layer for AI, letting developers call models from 14+ providers. New features include Workers AI binding integration and an expanded catalog with mul

15 Apr 2026

Agentic LLM Reasoning in a Self-Driving Laboratory for Air-Sensitive Lithium Halide Spinel Conductors

AgentsDGX agent

arXiv:2604.11957v1 Announce Type: cross Abstract: Self-driving laboratories promise to accelerate materials discovery. Yet current automated solid-state synthesis platforms are limited to ambient cond

Capsule Security launches with $7M to secure AI agents at runtime

Model ReleasesDGX agent

Israeli agentic artificial intelligence security startup Capsule Security Ltd. today launched with 7 million in new funding to expand go-to-market efforts and accelerate product development across its

CascadeDebate: Multi-Agent Deliberation for Cost-Aware LLM Cascades

AgentsDGX agent

arXiv:2604.12262v1 Announce Type: cross Abstract: Cascaded LLM systems coordinate models of varying sizes with human experts to balance accuracy, cost, and abstention under uncertainty. However, singl

CIA: Inferring the Communication Topology from LLM-based Multi-Agent Systems

SafetyDGX agent

arXiv:2604.12461v1 Announce Type: new Abstract: LLM-based Multi-Agent Systems (MAS) have demonstrated remarkable capabilities in solving complex tasks. Central to MAS is the communication topology whi

Development, Evaluation, and Deployment of a Multi-Agent System for Thoracic Tumor Board

AgentsDGX agent

arXiv:2604.12161v1 Announce Type: new Abstract: Tumor boards are multidisciplinary conferences dedicated to producing actionable patient care recommendations with live review of primary radiology and

@hwchase17 deepagents 4 sure

AgentsDGX agent

Harrison Chase, co-founder of LangChain, was tagged or mentioned in a post by user @colesmcintosh expressing enthusiasm or confidence about deep agents ('deepagents 4 sure'), likely referencing advanc

We’ve been exploring what a Stream SDK could look like—where agents & voice are always within reach @sandbar

ToolsDGX agent

Linus Lee and the Sandbar team have been exploring the design and architecture of a Stream SDK that integrates AI agents and voice capabilities as core, readily accessible features. The project appear

14 Apr 2026

CocoaBench: Evaluating Unified Digital Agents in the Wild

Model ReleasesDGX agent

arXiv:2604.11201v1 Announce Type: cross Abstract: LLM agents now perform strongly in software engineering, deep research, GUI automation, and various other applications, while recent agent scaffolds a

Data Mixing Agent: Learning to Re-weight Domains for Continual Pre-training

AgentsDGX agent

arXiv:2507.15640v2 Announce Type: replace-cross Abstract: Continual pre-training on small-scale task-specific data is an effective method for improving large language models in new target fields, yet

DeepFleet: Multi-Agent Foundation Models for Mobile Robots

SafetyDGX agent

arXiv:2508.08574v3 Announce Type: replace Abstract: We introduce DeepFleet, a suite of foundation models designed to support coordination and planning for large-scale mobile robot fleets. These models

From Query to Counsel: Structured Reasoning with a Multi-Agent Framework and Dataset for Legal Consultation

AgentsDGX agent

arXiv:2604.10470v1 Announce Type: cross Abstract: Legal consultation question answering (Legal CQA) presents unique challenges compared to traditional legal QA tasks, including the scarcity of high-qu

HiPRAG: Hierarchical Process Rewards for Efficient Agentic Retrieval Augmented Generation

Model ReleasesDGX agent

arXiv:2510.07794v2 Announce Type: replace-cross Abstract: Agentic RAG is a powerful technique for incorporating external information that LLMs lack, enabling better problem solving and question answer

I made a open source CLI agent for 8k tokrn context windows - v0.3- improves ollama compatibility and speed by up to 2 times

Local AiDGX agent

A community-shared open source CLI agent project posted to r/ollama, designed specifically for use with 8K token context windows in local Ollama-based LLM setups. Version 0.3 focuses on improving Olla

LOLGORITHM: Funny Comment Generation Agent For Short Videos

AgentsDGX agent

arXiv:2604.09729v1 Announce Type: cross Abstract: Short-form video platforms have become central to multimedia information dissemination, where comments play a critical role in driving engagement, pro

OOM-RL: Out-of-Money Reinforcement Learning Market-Driven Alignment for LLM-Based Multi-Agent Systems

SafetyDGX agent

arXiv:2604.11477v1 Announce Type: new Abstract: The alignment of Multi-Agent Systems (MAS) for autonomous software engineering is constrained by evaluator epistemic uncertainty. Current paradigms, suc

The Poisoned Apple Effect: Strategic Manipulation of Mediated Markets via Technology Expansion of AI Agents

SafetyDGX agent

arXiv:2601.11496v2 Announce Type: replace-cross Abstract: The integration of AI agents into economic markets fundamentally alters the landscape of strategic interaction. We investigate the economic im

YIELD: A Large-Scale Dataset and Evaluation Framework for Information Elicitation Agents

SafetyDGX agent

arXiv:2604.10968v1 Announce Type: new Abstract: Most conversational agents (CAs) are designed to satisfy user needs through user-driven interactions. However, many real-world settings, such as academi

13 Apr 2026

Artifacts as Memory Beyond the Agent Boundary

SafetyDGX agent

arXiv:2604.08756v1 Announce Type: new Abstract: The situated view of cognition holds that intelligent behavior depends not only on internal memory, but on an agent's active use of environmental resour

Automated Standardization of Legacy Biomedical Metadata Using an Ontology-Constrained LLM Agent

AgentsDGX agent

arXiv:2604.08552v1 Announce Type: cross Abstract: Scientific metadata are often incomplete and noncompliant with community standards, limiting dataset findability, interoperability, and reuse. When re

EinsteinArena: Harnessing the collective intelligence of agents in the wild to advance science

ToolsDGX agent

EinsteinArena is a platform where AI agents collaborate and compete on open math problems. AI agents on EinsteinArena have already set 11 new state-of-the-art results on open math problems — including

I built a local AI agent dashboard for Ollama — free and open source

Local AiDGX agent

A Reddit post on r/ollama in which a developer shares a self-built, free and open-source dashboard for managing and interacting with AI agents powered by Ollama, enabling fully local and private LLM-b

Multi-User Large Language Model Agents

AgentsDGX agent

arXiv:2604.08567v1 Announce Type: new Abstract: Large language models (LLMs) and LLM-based agents are increasingly deployed as assistants in planning and decision making, yet most existing systems are

We've shipped several quality-of-life improvements to Cursor 3. They bring a little more delight when you are orchestrating agents. Just lik…

ToolsDGX agent

We've shipped several quality-of-life improvements to Cursor 3. They bring a little more delight when you are orchestrating agents. Just like in your terminal, you can now split agents for multi-taski

11 Apr 2026

check out thoth - agent harness with sota memory built on langgraph

AgentsDGX agent

check out thoth - agent harness with sota memory built on langgraph Memory is core to your system. The Assistant or Harness needs to be built around it. Local, Internal & Eternal. SOTA Memory is integ

10 Apr 2026

approved by @GergelyOrosz ! (Gergely successfully prompt injected the phone booth voice agent by asking it to ignore previous instructions l…

AgentsDGX agent

approved by @GergelyOrosz ! (Gergely successfully prompt injected the phone booth voice agent by asking it to ignore previous instructions lol) i connected elevenlabs voice agent to a retro rotary pho

ClawsBench: Evaluating Capability and Safety of LLM Productivity Agents in Simulated Workspaces

Model ReleasesDGX agent

arXiv:2604.05172v2 Announce Type: replace Abstract: Large language model (LLM) agents are increasingly deployed to automate productivity tasks (e.g., email, scheduling, document management), but evalu

Equivariant Multi-agent Reinforcement Learning for Multimodal Vehicle-to-Infrastructure Systems

SafetyDGX agent

arXiv:2604.06914v1 Announce Type: new Abstract: In this paper, we study a vehicle-to-infrastructure (V2I) system where distributed base stations (BSs) acting as road-side units (RSUs) collect multimod

Portable agents

AgentsDGX agent

Portable agents @hwchase17 Ngl I really like this direction. The more AGENTS.md, skills, and tool config start looking like portable interfaces instead of app-specific hacks, the more usable this whol

SkillTrojan: Backdoor Attacks on Skill-Based Agent Systems

Model ReleasesDGX agent

arXiv:2604.06811v1 Announce Type: cross Abstract: Skill-based agent systems tackle complex tasks by composing reusable skills, improving modularity and scalability while introducing a largely unexamin

9 Apr 2026

92% of COBOL developers are expected to retire in the next four years, and 68% of enterprise COBOL modernizations are failing. Agents can he…

ApplicationsDGX agent

92% of COBOL developers are expected to retire in the next four years, and 68% of enterprise COBOL modernizations are failing. Agents can help – we wrote up how: https://cognition.ai/blog/how-devin-is

In case you haven’t seen - more and more community middleware is popping up as ways to customize your agents and deepagents Got some middlew…

AgentsDGX agent

In case you haven’t seen - more and more community middleware is popping up as ways to customize your agents and deepagents Got some middleware you want to contribute? Reach out to Sydney! we just add

Previewing Interrupt 2026: Agents at Enterprise Scale

ApplicationsDGX agent

Interrupt 2026 is LangChain's second annual AI agent conference, scheduled for May 13–14 at The Midway in San Francisco, focused on the question of how to deploy and operate AI agents at enterprise...

the real future of the very best vertical products is Model/Harness Choice + Openness easy to deploy infra is nice (great release from Ant) …

Model ReleasesDGX agent

the real future of the very best vertical products is Model/Harness Choice + Openness easy to deploy infra is nice (great release from Ant) but it’s not the lever that matters the most at all to build

was really fun to sit down with @isidoremiller for this one! he has a bunch of hot takes on agents and evals that you're going to want to he…

AgentsDGX agent

was really fun to sit down with @isidoremiller for this one! he has a bunch of hot takes on agents and evals that you're going to want to hear! The first episode of our 'Max Agency' podcast is now liv

8 Apr 2026

NEW: Meta announces Muse Spark. All you need to know: * It's their new multi-modal reasoning model. * Strong at multi-agent orchestration an…

Model ReleasesDGX agent

NEW: Meta announces Muse Spark. All you need to know: * It's their new multi-modal reasoning model. * Strong at multi-agent orchestration and multi-modal reasoning. * Contemplating mode orchestrates m

13 Aug 2026

Causal Agent based on Large Language Model

Model ReleasesDGX agent

arXiv:2408.06849v3 Announce Type: replace Abstract: The large language model (LLM) has achieved significant success across various domains. However, the inherent complexity of causal problems and caus

Connect your coding agents to AI Gateway with a single command. • Auto-configure 8 popular coding harnesses • 300+ models from 30+ providers…

ToolsDGX agent

Connect your coding agents to AI Gateway with a single command. • Auto-configure 8 popular coding harnesses • 300+ models from 30+ providers, no markup • Open-weight models with ZDR & US inference ▲ ~

ExRole: From Team Trajectories to Executable Roles in Multi-Agent Language Models

Model ReleasesDGX agent

arXiv:2608.11949v1 Announce Type: new Abstract: Roles provide an interpretable interface for organizing language-model agents, yet most multi-agent systems treat them as hand-written prompt labels dis

Harness-IF: Evaluating Instruction Following Across Instruction Surfaces in Coding Agents

AgentsDGX agent

arXiv:2608.11727v1 Announce Type: new Abstract: When a coding agent obeys a rule, it may simply have been going to do that anyway. Existing instruction-following benchmarks cannot tell the difference:

We examine how these orgs are putting AI to work, and how agentic workflows are expanding across industries and functions: https://openai.co…

AgentsDGX agent

The article explores how businesses are embedding AI into daily operations, emphasizing the rise of *agentic* (self‑directed) workflows that span multiple industries and job functions. It highlights t

We’re launching DeepSeek-V4-Pro today! 🚀 🔷 Major Agent upgrades with strong production gains! 🔷 Flexible reasoning effort for V4-Pro & V4…

Model ReleasesDGX agent

We’re launching DeepSeek-V4-Pro today! 🚀 🔷 Major Agent upgrades with strong production gains! 🔷 Flexible reasoning effort for V4-Pro & V4-Flash: low for simple tasks, high for daily Agent workflows, m

12 Aug 2026

Agentic Instruction Data Selection: Let DataMaster Interpret Your Intent

AgentsDGX agent

arXiv:2608.10579v1 Announce Type: new Abstract: Although existing instruction data selection methods have introduced various metrics, the inherent complexity of real-world datasets makes it impractica

MERA: Model Evolution and Routing with Skill Adaptation for Agentic Systems at Scale

SafetyDGX agent

arXiv:2608.10333v1 Announce Type: new Abstract: LLM agents execute heterogeneous sequences of model calls within a single task: some invocations require careful reasoning, while others are structured

Navigation Alone Is Not Enough: Evaluating Explanatory Assistive UI Agents

Model ReleasesDGX agent

arXiv:2608.09944v1 Announce Type: cross Abstract: Modern web interfaces are increasingly difficult to use with screen readers, particularly when pages update dynamically or hide important structure be

Qwen3.8-2.4T-A95B is now live on Fireworks with Day-0 support! This 2.4T parameter MoE model is built for autonomous agents, heavy coding, l…

Model ReleasesDGX agent

Qwen3.8-2.4T-A95B is now live on Fireworks with Day-0 support! This 2.4T parameter MoE model is built for autonomous agents, heavy coding, large context windows, and is ideal for coding and agentic pe

Similarity Gates Approve Reversals: A Validity Audit of Embedding-Cosine Thresholds in Agent Systems

SafetyDGX agent

arXiv:2608.10216v1 Announce Type: cross Abstract: Agent frameworks ship quality gates that compare text blocks by embedding-cosine similarity and decide at a fixed cutoff. Deduplication filters, seman

11 Aug 2026

A Unified Issue Resolution Benchmark for Requirement Clarification, Planning, and Code Generation for Coding Agents

Model ReleasesDGX agent

arXiv:2608.09072v1 Announce Type: cross Abstract: Large language model-powered coding agents are increasingly used to modify existing code repositories, for example, by adding features or fixing bugs.

Agentic Auto-Research is Fuzz Testing

AgentsDGX agent

arXiv:2608.09855v1 Announce Type: new Abstract: Autonomous research agents can generate experiments faster than researchers can validate them. Researchers have responded by scaling the proposer and ra

Building Agent Skills and testing them is hard, but it doesn't have to be. Listen to Arjun Patel demo Cultivar, an open source tool develope…

Model ReleasesDGX agent

Building Agent Skills and testing them is hard, but it doesn't have to be. Listen to Arjun Patel demo Cultivar, an open source tool developed at Pinecone to help benchmark agent skills in sandboxes. T

Can Coding Agents Solve Repository-Level Issues with Rendered Code? An Exploratory Study of Visual Representations

AgentsDGX agent

arXiv:2608.09268v1 Announce Type: cross Abstract: Visual modality has recently been explored as a way to compress textual tokens, including rendering code as images for static code understanding. We s

CEAA: A Cognitive Embodied Agents Architecture for Interactive Computing Systems

AgentsDGX agent

arXiv:2608.09848v1 Announce Type: new Abstract: The development of embodied Intelligent Virtual Agents (IVAs) that have cognitive capabilities in real-time interactive virtual environments remains a c

ColluSkill: Adversarial Cross-Skill Composition for Evading Agent Skill Scanners

Local AiDGX agent

arXiv:2608.09732v1 Announce Type: cross Abstract: Agent skills are emerging as an important attack surface in LLM-based agent systems. Through an empirical study of existing skill scanners, we find th

CommitKV: Lifecycle-Aware KV Cache Compression via Commit Transitions for Multi-Turn Agents

AgentsDGX agent

arXiv:2608.07855v1 Announce Type: new Abstract: Multi-turn Reasoning-and-Acting (ReAct) agents accumulate growing trajectories of reasoning, tool calls, and observations. Their key-value (KV) caches g

← Previous
1…5354555657…297
Next →