AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,762 results
21 Apr 2026

A Multi-Agent Approach for Claim Verification from Tabular Data Documents

AgentsDGX agent

arXiv:2604.17225v1 Announce Type: new Abstract: We present a novel approach for claim verification from tabular data documents. Recent LLM-based approaches either employ complex pretraining/fine-tunin

BrainMem: Brain-Inspired Evolving Memory for Embodied Agent Task Planning

AgentsDGX agent

arXiv:2604.16331v1 Announce Type: cross Abstract: Embodied task planning requires agents to execute long-horizon, goal-directed actions in complex 3D environments, where success depends on both immedi

🏆LangChain was just named 2026 @googlecloud Partner of the Year Award in the Marketplace: Agent Platform category. ☁️ @gcloudpartners #Goog…

AgentsDGX agent

LangChain received Google Cloud's 2026 Partner of the Year Award in the Marketplace: Agent Platform category, recognizing its contributions to the cloud ecosystem. The award highlights LangChain's pro

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Why Agents Compromise Safety Under Pressure

SafetyDGX agent

arXiv:2603.14975v2 Announce Type: replace-cross Abstract: Large Language Model agents deployed in complex environments frequently encounter a conflict between maximizing goal achievement and adhering

20 Apr 2026

DeepKeep rolls out Vibe AI Red Teaming to test AI apps and agents

TutorialsDGX agent

Artificial intelligence security platform company DeepKeep Ltd. today announced the launch of Vibe AI Red Teaming, a capability designed for human-steered, dynamic testing and attack simulation on AI

LinuxArena: A Control Setting for AI Agents in Live Production Software Environments

Model ReleasesDGX agent

arXiv:2604.15384v1 Announce Type: cross Abstract: We introduce LinuxArena, a control setting in which agents operate directly on live, multi-service production environments. LinuxArena contains 20 env

SocialGrid: A Benchmark for Planning and Social Reasoning in Embodied Multi-Agent Systems

Model ReleasesDGX agent

arXiv:2604.16022v1 Announce Type: new Abstract: As Large Language Models (LLMs) transition from text processors to autonomous agents, evaluating their social reasoning in embodied multi-agent settings

17 Apr 2026

Will agentic AI governance run amok? The lesson of Asimov’s Three Laws

AgentsDGX agent

Asimov’s Three Laws of Robotics: A robot may not injure a human being or, through inaction, allow a human being to come to harm. A robot must obey the orders given it by human beings except where such

16 Apr 2026

Adaptive Memory Crystallization for Autonomous AI Agent Learning in Dynamic Environments

AgentsDGX agent

arXiv:2604.13085v1 Announce Type: new Abstract: Autonomous AI agents operating in dynamic environments face a persistent challenge: acquiring new capabilities without erasing prior knowledge. We prese

15 Apr 2026

Agentic Insight Generation in VSM Simulations

AgentsDGX agent

arXiv:2604.12421v1 Announce Type: new Abstract: Extracting actionable insights from complex value stream map simulations can be challenging, time-consuming, and error-prone. Recent advances in large l

DRPG (Decompose, Retrieve, Plan, Generate): An Agentic Framework for Academic Rebuttal

AgentsDGX agent

arXiv:2601.18081v2 Announce Type: replace Abstract: Despite the growing adoption of large language models (LLMs) in scientific research workflows, automated support for academic rebuttal, a crucial st

I've commented that 'this is the year of subagents', but that is largely an optimization problem. the inverse problem - having agents that c…

ToolsDGX agent

I've commented that 'this is the year of subagents', but that is largely an optimization problem. the inverse problem - having agents that compose and boss agents that manage/query them - is a capabil

New in LangSmith Fleet: Tool access controls and usage tracking. 📊 Track cost and usage by user, agent, and tool from a single dashboard 💳…

AgentsDGX agent

New in LangSmith Fleet: Tool access controls and usage tracking. 📊 Track cost and usage by user, agent, and tool from a single dashboard 💳 Set spend limits per user or team to prevent surprises ✅ Cont

The next evolution of the Agents SDK

Model ReleasesDGX agent

OpenAI's Agents SDK represents an evolution of their framework for building agentic AI applications, providing developers with tools to create, orchestrate, and deploy AI agents that can perform multi

14 Apr 2026

ACE-TA: An Agentic Teaching Assistant for Grounded Q&A, Quiz Generation, and Code Tutoring

AgentsDGX agent

arXiv:2604.09572v1 Announce Type: cross Abstract: We introduce ACE-TA, the Agentic Coding and Explanations Teaching Assistant framework, that autonomously routes conceptual queries drawn from programm

AgencyBench: Benchmarking the Frontiers of Autonomous Agents in 1M-Token Real-World Contexts

Model ReleasesDGX agent

arXiv:2601.11044v3 Announce Type: replace Abstract: Large Language Models (LLMs) based autonomous agents demonstrate multifaceted capabilities to contribute substantially to economic production. Howev

ArtiCAD: Articulated CAD Assembly Design via Multi-Agent Code Generation

AgentsDGX agent

arXiv:2604.10992v1 Announce Type: new Abstract: Parametric Computer-Aided Design (CAD) of articulated assemblies is essential for product development, yet generating these multi-part, movable models f

ClawGuard: A Runtime Security Framework for Tool-Augmented LLM Agents Against Indirect Prompt Injection

Local AiDGX agent

arXiv:2604.11790v1 Announce Type: cross Abstract: Tool-augmented Large Language Model (LLM) agents have demonstrated impressive capabilities in automating complex, multi-step real-world tasks, yet rem

🚀 deepagents 0.5 release 👉 Async subagents - kick off background tasks on any Agent Protocol backed server while you continue to interact …

Model ReleasesDGX agent

🚀 deepagents 0.5 release 👉 Async subagents - kick off background tasks on any Agent Protocol backed server while you continue to interact with the main agent. Start multiple background tasks in parall

It looks like everyone is finally catching up with the fact that agent sessions in CLI mode can only get you so far. It makes sense that the…

Model ReleasesDGX agent

It looks like everyone is finally catching up with the fact that agent sessions in CLI mode can only get you so far. It makes sense that the new Codex app, Cursor, and Claude Code (desktop) feel and l

Learning to Assist: Physics-Grounded Human-Human Control via Multi-Agent Reinforcement Learning

AgentsDGX agent

arXiv:2603.11346v2 Announce Type: replace Abstract: Humanoid robotics has strong potential to transform daily service and caregiving applications. Although recent advances in general motion tracking w

Manifest platform from Manifold targets AI agent supply chain security gaps

AgentsDGX agent

Artificial intelligence detection and response platform startup Manifold Security Inc. today announced the launch of Manifest, a new supply chain intelligence platform that is designed to map and anal

MiniMax M2.7 is now available in LM Studio. This model excels at agentic tool calling 🛠️ Requires at least ~138GB to run locally https://lm…

AgentsDGX agent

MiniMax M2.7 is a large language model now available for local deployment through LM Studio, notable for its strong performance in agentic tool calling tasks. The model requires a substantial minimum

RPA-Check: A Multi-Stage Automated Framework for Evaluating Dynamic LLM-based Role-Playing Agents

Local AiDGX agent

arXiv:2604.11655v1 Announce Type: cross Abstract: The rapid adoption of Large Language Models (LLMs) in interactive systems has enabled the creation of dynamic, open-ended Role-Playing Agents (RPAs).

13 Apr 2026

Alleviating Community Fear in Disasters via Multi-Agent Actor-Critic Reinforcement Learning

AgentsDGX agent

arXiv:2604.08802v1 Announce Type: new Abstract: During disasters, cascading failures across power grids, communication networks, and social behavior amplify community fear and undermine cooperation. E

Asked 5.4 if any AI agent apps had a “tag response” or “copy to folder” option.

AgentsDGX agent

A Reddit user on r/ChatGPT asked GPT-5.4 whether any AI agent apps offer organizational features such as 'tag response' or 'copy to folder' — functions for saving and categorizing AI-generated outputs

cuba-memorys v0.7.0 — Persistent Memory for AI Agents

Local AiDGX agent

**cuba-memorys v0.7.0** is a persistent memory library for AI agents, shared in the r/ollama community, designed to give local LLM-powered agents the ability to retain information across sessions. It

Self-Sovereign Agent

AgentsDGX agent

arXiv:2604.08551v1 Announce Type: cross Abstract: We investigate the emerging prospect of self-sovereign agents -- AI systems that can economically sustain and extend their own operation without human

SPEAR: An Engineering Case Study of Multi-Agent Coordination for Smart Contract Auditing

SafetyDGX agent

arXiv:2602.04418v3 Announce Type: replace-cross Abstract: We present SPEAR, a multi-agent coordination framework for smart contract auditing that applies established MAS patterns in a realistic securi

V-CAGE: Vision-Closed-Loop Agentic Generation Engine for Robotic Manipulation

AgentsDGX agent

arXiv:2604.09036v1 Announce Type: new Abstract: Scaling Vision-Language-Action (VLA) models requires massive datasets that are both semantically coherent and physically feasible. However, existing sce

12 Apr 2026

the most important abstraction in AI agents isnt the model — its the harness it orchestrates tools, memory, prompts. this is where all the a…

AgentsDGX agent

the most important abstraction in AI agents isnt the model — its the harness it orchestrates tools, memory, prompts. this is where all the alpha is deepagents is our take: built-in tools, memory, smar

10 Apr 2026

AEROS: A Single-Agent Operating Architecture with Embodied Capability Modules

SafetyDGX agent

arXiv:2604.07039v1 Announce Type: cross Abstract: Robotic systems lack a principled abstraction for organizing intelligence, capabilities, and execution in a unified manner. Existing approaches either

ForkKV: Scaling Multi-LoRA Agent Serving via Copy-on-Write Disaggregated KV Cache

HardwareDGX agent

arXiv:2604.06370v1 Announce Type: cross Abstract: The serving paradigm of large language models (LLMs) is rapidly shifting towards complex multi-agent workflows where specialized agents collaborate ov

great question! knowledge (system prompt, skills) and tools (apis, clis, whatnot) will always be needed, even as the best way to build agent…

AgentsDGX agent

great question! knowledge (system prompt, skills) and tools (apis, clis, whatnot) will always be needed, even as the best way to build agents changes i would focus on those and processes around those

Harnessing Embodied Agents: Runtime Governance for Policy-Constrained Execution

SafetyDGX agent

arXiv:2604.07833v1 Announce Type: new Abstract: Embodied agents are evolving from passive reasoning systems into active executors that interact with tools, robots, and physical environments. Once gran

How Much LLM Does a Self-Revising Agent Actually Need?

AgentsDGX agent

arXiv:2604.07236v2 Announce Type: new Abstract: Recent LLM-based agents often place world modeling, planning, and reflection inside a single language model loop. This can produce capable behavior, but

@IeloEmanuele is on a roll! you can now use claude code's advisor strategy with @LangChain agents because langchain/deepagents are provider …

Model ReleasesDGX agent

@IeloEmanuele is on a roll! you can now use claude code's advisor strategy with @LangChain agents because langchain/deepagents are provider agnostic, you can use different providers for your advisor a

TurboAgent: An LLM-Driven Autonomous Multi-Agent Framework for Turbomachinery Aerodynamic Design

AgentsDGX agent

arXiv:2604.06747v2 Announce Type: new Abstract: The aerodynamic design of turbomachinery is a complex and tightly coupled multi-stage process involving geometry generation, performance prediction, opt

9 Apr 2026

Claude Agent SDK tracing in LangSmith just got an upgrade. Now you can trace: → Subagents → Child runs inside MCP tools → Cost tracking + mo…

Model ReleasesDGX agent

Claude Agent SDK tracing in LangSmith just got an upgrade. Now you can trace: → Subagents → Child runs inside MCP tools → Cost tracking + more Update to the latest Python SDK to try it out. Docs: http

we just added a langchain-task-steering middleware to our community registry, s/o @EHallvaxhiu for the contribution! keep your agents on tra…

AgentsDGX agent

we just added a langchain-task-steering middleware to our community registry, s/o @EHallvaxhiu for the contribution! keep your agents on track w/ ordered task pipelines. this is great for structured d

8 Apr 2026

Curriculum Learning for harnesses - should we teach agents how we teach kids? start small and easy and progressively get harder for my resea…

AgentsDGX agent

Curriculum Learning for harnesses - should we teach agents how we teach kids? start small and easy and progressively get harder for my research friends here's an under-explored area we're thinking abo

Hermes Agent v0.8.0 is here. Full changelog below ↓

AgentsDGX agent

Nous Research released Hermes Agent v0.8.0 (tagged v2026.4.8) on April 8, 2026, describing it as 'the intelligence release' and encompassing 209 merged pull requests with 82 resolved issues. Key ad...

7 Apr 2026

Visually rich documents are especially challenging for agents. Tables, charts, and images often break traditional document pipelines, making…

Model ReleasesDGX agent

Visually rich documents are especially challenging for agents. Tables, charts, and images often break traditional document pipelines, making complex reasoning difficult📄 So we teamed up with @lancedb

13 Aug 2026

AVA-Encoder: Towards Agent-Native Video Representation Learning

Model ReleasesDGX agent

arXiv:2608.12313v1 Announce Type: cross Abstract: Creative agents still lack an effective way to learn from high-quality human films, limiting their ability to produce cinematic-grade videos. A key ch

Benchmarking LLM Judges for Mobile Agent Evaluation

Model ReleasesDGX agent

arXiv:2608.11434v1 Announce Type: new Abstract: Mobile agent benchmarks increasingly rely on LLM-based judges to evaluate task completion, yet the reliability of these judges on mobile agent trajector

12 Aug 2026

Blacksmith raises $45M to aid AI code validation as agentic development grows

AgentsDGX agent

Blacksmith Software Inc. today announced it has raised 45 million in new funding for its continuous integration service, which combines code development with cloud-based testing instead of on the deve

From Faulty Memories to Corrected Actions: Dependency-Guided Rollback Repair for Memory-Augmented Agents

Model ReleasesDGX agent

arXiv:2608.10502v1 Announce Type: new Abstract: Persistent memory lets language-model agents reuse information across sessions, but it also makes errors durable: a poisoned, stale, or misattributed re

Muse Glimmer is live on Fireworks. The new open-weight model from Meta Superintelligence Labs is a 30B dense model built for always-on agent…

AgentsDGX agent

Muse Glimmer is live on Fireworks. The new open-weight model from Meta Superintelligence Labs is a 30B dense model built for always-on agents that reason across many sequential tool calls and can reco

11 Aug 2026

Agent-MD: Selective LLM Intervention with Event-Driven Escalation for Stateful GCMC--MD Campaigns

AgentsDGX agent

arXiv:2608.07637v1 Announce Type: new Abstract: Long-running molecular simulation campaigns require repeated continuation from saved states, provenance-aware progression, adaptive assessment, and occa

Autonomy Reshapes How Personalization Affects Privacy Concerns and Trust in LLM Agents

SafetyDGX agent

arXiv:2510.04465v3 Announce Type: replace-cross Abstract: LLM agents require personal information for personalization in order to effectively act on users' behalf, but this raises privacy concerns tha

Guixu: Valuation-Driven Data Discovery for Autonomous AI Agents with On-Chain Attestation

AgentsDGX agent

arXiv:2608.07949v1 Announce Type: new Abstract: Autonomous agents increasingly rely on external data to complete downstream tasks such as model training and decision support. However, existing data di

Multilingual Agent-Based World Modeling for Social Science

Model ReleasesDGX agent

arXiv:2512.07195v2 Announce Type: replace-cross Abstract: Multi-agent role-playing has recently shown promise for studying social behavior with language agents, but existing simulations are mostly mon

NeuroPilot: An Agent-Driven Smart Pipeline for Processing, Quality Control, and Managing Neuroimages

AgentsDGX agent

arXiv:2608.07541v1 Announce Type: cross Abstract: Transforming raw neuroimage archives into analysis-ready derivatives relies on three brittle stages: data standardization, modality-specific preproces

Route AI Agent Workloads Across Models with NVIDIA NeMo Switchyard

HardwareDGX agent

NVIDIA NeMo Switchyard is a routing platform that directs AI agent workloads to the most suitable specialized or frontier model for each step of a task, balancing performance, cost, and latency. It of

SkillsMetric: Mapping the Detection Boundary of Static Analysis for Malicious Agent Skills

AgentsDGX agent

arXiv:2608.08468v1 Announce Type: cross Abstract: Agent Skills---structured packages of instructions and scripts that augment LLM-based agents---are rapidly proliferating, yet their security propertie

The Replay Gap: Static Evaluation of Model Switching in LLM Agents Scores the Wrong World

AgentsDGX agent

arXiv:2608.08239v1 Announce Type: cross Abstract: LLM routers promise efficiency by matching each request to the cheapest adequate model, and are increasingly applied per step inside multi-step agents

10 Aug 2026

BONSAI: Evolvability-Guided Tree Search over Skills

AgentsDGX agent

arXiv:2608.07056v1 Announce Type: new Abstract: A skill is a naturallanguage document that steers a frozen agent whose weights cannot be updated so any capability the agent lacks must be supplied in p

CyberForge: Verified Vulnerability Injection at Repository Level for Cybersecurity Agent Training

Model ReleasesDGX agent

arXiv:2608.06471v1 Announce Type: cross Abstract: Despite recent advances, frontier large language model (LLM) agents remain limited in discovering and patching complex vulnerabilities in real-world s

Do AI Personas Grow? Analyzing and Benchmarking Personality Evolution in LLM Agents After Life Events

Model ReleasesDGX agent

arXiv:2608.06485v1 Announce Type: cross Abstract: Personality-conditioned LLM agents (PC-Agents) are increasingly used in emotional support, social simulation, and role-playing, motivating the develop

EMAS: Stabilizing Multi-Agent System Evolution through Evidence-Guided Revision

Model ReleasesDGX agent

arXiv:2608.07196v1 Announce Type: new Abstract: Many methods for automated multi-agent system design optimize prompts and topologies during an initial design stage and then deploy the resulting system

← Previous
1…4546474849…297
Next →