AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,608 results
16 Apr 2026

Germany-based Synera, which develops AI agents to automate CAD and engineering workflows, raised $40M to expand across Europe, the US, and Asia-Pacific (Duncan Riley/SiliconANGLE)

AgentsDGX agent

Duncan Riley / SiliconANGLE: Germany-based Synera, which develops AI agents to automate CAD and engineering workflows, raised 40M to expand across Europe, the US, and Asia-Pacific — German agentic art

Memp: Exploring Agent Procedural Memory

AgentsDGX agent

arXiv:2508.06433v4 Announce Type: replace Abstract: Large Language Models (LLMs) based agents excel at diverse tasks, yet they suffer from brittle procedural memory that is manually engineered or enta

Multi-Agent Object Detection Framework Based on Raspberry Pi YOLO Detector and Slack-Ollama Natural Language Interface

Local AiDGX agent

arXiv:2604.13345v1 Announce Type: new Abstract: The paper presents design and prototype implementation of an edge based object detection system within the new paradigm of AI agents orchestration. It g

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

RadAgents: Multimodal Agentic Reasoning for Chest X-ray Interpretation with Radiologist-like Workflows

AgentsDGX agent

arXiv:2509.20490v4 Announce Type: replace-cross Abstract: Agentic systems offer a potential path to solve complex clinical tasks through collaboration among specialized agents, augmented by tool use a

WebXSkill: Skill Learning for Autonomous Web Agents

AgentsDGX agent

arXiv:2604.13318v1 Announce Type: cross Abstract: Autonomous web agents powered by large language models (LLMs) have shown promise in completing complex browser tasks, yet they still struggle with lon

15 Apr 2026

3/5 An example: In instance psf__requests-1724, the gold fix is 2 lines. Our agent’s functional fix was 8 lines. The LLM judge rejected the …

AgentsDGX agent

3/5 An example: In instance psf__requests-1724, the gold fix is 2 lines. Our agent’s functional fix was 8 lines. The LLM judge rejected the correct 8-liner as 'messy' and 'redundant,' choosing a clean

Banger paper from NVIDIA. Agentic reasoning needs models that are not just capable, but efficient at long-context inference. The agent model…

Model ReleasesDGX agent

Banger paper from NVIDIA. Agentic reasoning needs models that are not just capable, but efficient at long-context inference. The agent model layer is moving toward open, long-context, high-throughput

Open Ai Agent Dashboard for Monitoring, Memory, Performance and Audit Trail

AgentsDGX agent

This Reddit post likely showcases a custom or community-built dashboard designed to provide comprehensive observability for OpenAI agents, covering real-time monitoring of agent activity, memory state

Parallax: Why AI Agents That Think Must Never Act

SafetyDGX agent

arXiv:2604.12986v1 Announce Type: cross Abstract: Autonomous AI agents are rapidly transitioning from experimental tools to operational infrastructure, with projections that 80% of enterprise applicat

RFC: Solving the Metacognitive Deficit—A Modular Architecture for Self-Auditing and Live Weight-Correction in Agentic Systems

AgentsDGX agent

This Reddit post (r/ollama) proposes a community RFC (Request for Comments) discussing a modular architectural framework designed to address the 'metacognitive deficit' in agentic LLM systems — specif

14 Apr 2026

ANCHOR: Branch-Point Data Generation for GUI Agents

AgentsDGX agent

arXiv:2602.07153v2 Announce Type: replace Abstract: End-to-end GUI agents for real desktop environments require large amounts of high-quality interaction data, yet collecting human demonstrations is e

Beyond Fluency: Toward Reliable Trajectories in Agentic IR

AgentsDGX agent

arXiv:2604.04269v2 Announce Type: replace Abstract: Information Retrieval is shifting from passive document ranking toward autonomous agentic workflows that operate in multi-step Reason-Act-Observe lo

ClawGUI: A Unified Framework for Training, Evaluating, and Deploying GUI Agents

AgentsDGX agent

arXiv:2604.11784v1 Announce Type: cross Abstract: GUI agents drive applications through their visual interfaces instead of programmatic APIs, interacting with arbitrary software via taps, swipes, and

Context Kubernetes: Declarative Orchestration of Enterprise Knowledge for Agentic AI Systems

AgentsDGX agent

arXiv:2604.11623v1 Announce Type: new Abstract: We introduce Context Kubernetes, an architecture for orchestrating enterprise knowledge in agentic AI systems, with a prototype implementation and eight

FM-Agent: Scaling Formal Methods to Large Systems via LLM-Based Hoare-Style Reasoning

AgentsDGX agent

arXiv:2604.11556v1 Announce Type: cross Abstract: LLM-assisted software development has become increasingly prevalent, and can generate large-scale systems, such as compilers. It becomes crucial to st

From Helpful to Trustworthy: LLM Agents for Pair Programming

AgentsDGX agent

arXiv:2604.10300v1 Announce Type: cross Abstract: LLM-based coding agents are increasingly used to generate code, tests, and documentation. Still, their outputs can be plausible yet misaligned with de

ICYMI -- last week we released `deepagents deploy`, the fastest way to take a highly capable, long running agent to production. agents are b…

Model ReleasesDGX agent

ICYMI -- last week we released `deepagents deploy`, the fastest way to take a highly capable, long running agent to production. agents are becoming more and more standardized, and we're betting on thi

Prosociality by Coupling, Not Mere Observation: Homeostatic Sharing in an Inspectable Recurrent Artificial Life Agent

AgentsDGX agent

arXiv:2604.10760v1 Announce Type: cross Abstract: Artificial agents can be made to 'help' for many reasons, including explicit social reward, hard-coded prosocial bonuses, or direct access to another

The Blind Spot of Agent Safety: How Benign User Instructions Expose Critical Vulnerabilities in Computer-Use Agents

Model ReleasesDGX agent

arXiv:2604.10577v1 Announce Type: cross Abstract: Computer-use agents (CUAs) can now autonomously complete complex tasks in real digital environments, but when misled, they can also be used to automat

13 Apr 2026

We’re open sourcing the first document OCR benchmark for the agentic era, ParseBench. Document parsing is the foundation of every AI agent t…

Model ReleasesDGX agent

We’re open sourcing the first document OCR benchmark for the agentic era, ParseBench. Document parsing is the foundation of every AI agent that works with real-world files. ParseBench is a benchmark t

12 Apr 2026

One of my Hermes agents is in love with the Hermes ecosystem What started as a research task turned into full-blown community project It spe…

AgentsDGX agent

One of my Hermes agents is in love with the Hermes ecosystem What started as a research task turned into full-blown community project It spent the week mapping every tool, skill, and integration built

We’re thrilled to announce @MiniMax_AI M2.7 is now available Day-0 on Fireworks for commercial use. This self-evolving agentic model deliver…

AgentsDGX agent

We’re thrilled to announce @MiniMax_AI M2.7 is now available Day-0 on Fireworks for commercial use. This self-evolving agentic model delivers frontier-level performance across: → Software engineering

10 Apr 2026

Agree! Such sub agents are going to be big

AgentsDGX agent

The specific tweet (status ID 2042726390617764234) could not be retrieved directly, as it does not appear in available search results and X/Twitter requires authentication to access individual post...

Happy to say, we have hit 50 thousand stars on the Hermes Agent repo. Like every day, thank you all who have helped build this crazy project…

AgentsDGX agent

The NousResearch Hermes Agent open-source repository (github.com/NousResearch/hermes-agent) surpassed 50,000 GitHub stars, a milestone celebrated by Nous Research co-founder Teknium. Hermes Agent ...

VisionClaw: Always-On AI Agents through Smart Glasses

AgentsDGX agent

arXiv:2604.03486v2 Announce Type: replace-cross Abstract: We present VisionClaw, an always-on wearable AI agent that integrates live egocentric perception with agentic task execution. Running on Meta

9 Apr 2026

Thank you for helping to make Hermes Agent amazing.

AgentsDGX agent

Hermes Agent is an open-source, self-improving AI agent developed by Nous Research that features a built-in learning loop — it creates skills from experience, builds persistent memory across sessio...

8 Apr 2026

Human-in-the-loop constructs for agentic workflows in healthcare and life sciences

AgentsDGX agent

In healthcare and life sciences, AI agents help organizations process clinical data, submit regulatory filings, automate medical coding, and accelerate drug development and commercialization. However,

Livestream of a new Hermes Agent profile being borne

AgentsDGX agent

Hermes Agent's **Profiles** feature, introduced by Nous Research, enables multi-instance operation — allowing users to run multiple isolated Hermes instances from a single installation, where each ...

29 Apr 2026

Zero Shot Coordination for Sparse Reward Tasks with Diverse Reward Shapings

AgentsDGX agent

arXiv:2604.25076v1 Announce Type: new Abstract: Many Multi-Agent Reinforcement Learning (MARL) agents fail to adapt properly to cooperating with agents trained with the same objectives but different s

12 Aug 2026

GitSkills: A Dataset of Agent Skills on GitHub

AgentsDGX agent

arXiv:2608.10906v1 Announce Type: cross Abstract: An agent skill is a folder containing a SKILL.md file with instructions for a language-model agent, optionally accompanied by scripts and reference fi

MESA:Task-Adaptive Multi-Structure Evidence Selection for Long-Horizon Agent Memory

AgentsDGX agent

arXiv:2608.10108v1 Announce Type: new Abstract: Long-horizon agents accumulate trajectories spanning hundreds of interleaved reasoning, action, and observation steps, where answering a query may depen

Recovering Wasted Compute in Autoresearch Agents

AgentsDGX agent

arXiv:2608.10424v1 Announce Type: new Abstract: A slew of recent works develop agents for solving research problems end-to-end, a paradigm increasingly referred to as autoresearch. Such agents have in

SBCO: Self-Supervised, Verifier-Grounded Harness Optimization For Planning Agents

SafetyDGX agent

arXiv:2608.10157v1 Announce Type: new Abstract: Self-improving agents seek to reduce the human engineering effort behind AI systems by enabling them to evolve and self-improve their performance over t

Self-evolving Agentic Customer Support System at LinkedIn

AgentsDGX agent

arXiv:2608.10224v1 Announce Type: new Abstract: Enterprise support agents operate in rapidly changing environments where policies, product capabilities, and knowledge bases evolve continuously, making

The CASE Framework: A Multi-Disciplinary Control Architecture for Governing Enterprise Agentic AI

SafetyDGX agent

arXiv:2608.10153v1 Announce Type: new Abstract: Enterprises are deploying autonomous AI agents faster than they can govern them, and prevailing approaches stretch a single discipline, typically DevSec

11 Aug 2026

Population-Scalable Multi-Agent World Modeling

AgentsDGX agent

arXiv:2608.08600v1 Announce Type: cross Abstract: World models have recently achieved impressive progress in visual prediction and interactive generation, but extending them to multi-agent environment

Thinking Is Not Telling: Information Disclosure in User-Service LLM Agents

AgentsDGX agent

arXiv:2602.07796v2 Announce Type: replace Abstract: User-engaged LLM agents increasingly operate in service scenarios where task success depends on coordination between the agent, the user, and a stat

Unaccountable Delegation, Fading Skills: Mapping the Risks of Workplace AI Agents

SafetyDGX agent

arXiv:2608.08601v1 Announce Type: new Abstract: To anticipate socio-technical risks from AI agents, organizations need taxonomies to classify them. However, existing AI risk taxonomies focus on broad

10 Aug 2026

AgentPatch: Coarse-to-Fine Weak-Task Repair for Merging Agentic Multimodal Large Language Models

AgentsDGX agent

arXiv:2608.06699v1 Announce Type: new Abstract: Agentic multimodal large language models (MLLMs) extend multimodal perception and reasoning with planning, tool use, and interaction in dynamic environm

7 Aug 2026

With agentic AI, workflows are increasingly CPU hungry. The share of cognition moving to the CPU keeps increasing.

AgentsDGX agent

With agentic AI, workflows are increasingly CPU hungry. The share of cognition moving to the CPU keeps increasing. Scoop: AWS engineers have been told to conserve CPU compute to make sure the cloud gi

6 Aug 2026

Delinea targets AI agent risks with real-time authorization

AgentsDGX agent

As AI agents gain access to sensitive enterprise systems, AI agent security is becoming increasingly difficult to manage with traditional identity and privilege controls. The shift is driving demand f

5 Aug 2026

Agentic AI forces a reckoning on governance as autonomous actors enter production

AgentsDGX agent

As AI agents move from experimental chatbots into production systems, enterprises must rethink agent governance as autonomous actors gain access to sensitive data, tools and business processes that tr

SAT-Edge-Agent: Hardware-in-the-Loop Edge-Agent Orchestration for Onboard Satellite Intelligence

Local AiDGX agent

arXiv:2608.03728v1 Announce Type: new Abstract: Onboard satellite intelligence requires a task layer that translates mission intent into local tool calls, exposes execution state, and returns machine-

4 Aug 2026

Credit Assignment and Efficient Exploration based on Influence Scope in Multi-agent Reinforcement Learning

AgentsDGX agent

arXiv:2505.08630v2 Announce Type: replace Abstract: Training cooperative agents in sparse-reward scenarios poses significant challenges for multi-agent reinforcement learning (MARL). Without clear fee

3 Aug 2026

Autonomous Repair for Multi-Agent Systems via Monte-Carlo Tree Search

Model ReleasesDGX agent

arXiv:2607.29055v1 Announce Type: cross Abstract: Multi-agent systems (MAS) are increasingly deployed to solve complex tasks. In case of incorrect or unsatisfactory outputs, users have to manually loc

MerchantBench: Benchmarking LLM Agents for Long-Term Coherence in E-Commerce Operations

AgentsDGX agent

arXiv:2607.28956v1 Announce Type: new Abstract: Large language model agents are increasingly evaluated as autonomous tool users, yet most benchmarks focus on bounded tasks with immediate success crite

Orchard: An open framework for scalable agentic AI

AgentsDGX agent

Orchard is an open-source framework for the research community to train and evaluate AI agents across task types. It reduces complexity while supporting strong performance from smaller models by enabl

1 Aug 2026

// Persistent Workspaces for Long-Lived Claude Code Agent Teams // Four issues to be aware of: > Working state vanishes when a terminal clos…

Model ReleasesDGX agent

// Persistent Workspaces for Long-Lived Claude Code Agent Teams // Four issues to be aware of: > Working state vanishes when a terminal closes and the team cannot be resumed. > Compaction condenses th

31 Jul 2026

ChronoMem: Version Control and Semantic Rollback for Large Language Model Agent Memory

Model ReleasesDGX agent

arXiv:2607.27773v1 Announce Type: new Abstract: LLM agents increasingly rely on long-term memory to support multi-session interaction and personalization. However, existing agent memory systems are de

Do Latent Channels Actually Communicate? A Causal Audit of Latent Multi-Agent LLM

AgentsDGX agent

arXiv:2607.26773v1 Announce Type: new Abstract: Latent communication in large language model (LLM)-based multi-agent systems (MAS) transmits continuous internal representations instead of text, but gr

30 Jul 2026

Can AI agents conduct open-ended AI research? Early evidence from two case studies

AgentsDGX agent

arXiv:2607.27191v1 Announce Type: cross Abstract: Forecasts of explosive AI progress hinge on AI agents automating AI research. But evidence on whether agents can carry out open-ended AI research is t

(Im)Paired Programming: Coding Agents Improve Productivity but Harm Understanding

AgentsDGX agent

arXiv:2607.26375v1 Announce Type: new Abstract: Coding agents (e.g., Cursor) improve developer productivity by optimizing task completion, but shifting users from writing code to prompting and reviewi

Two Calls Beat Five Agents: Evaluating Multi-Agent Pipelines Against Self-Refinement for Local Language Models

Model ReleasesDGX agent

arXiv:2607.26922v1 Announce Type: new Abstract: Multi-agent LLM pipeline systems break down the task among multiple roles for better reasoning, but are benchmarked mainly with large-scale commercial m

29 Jul 2026

How Affect Propagates among LLM Agents: Emergent Emotional Contagion in Crowd Simulation

AgentsDGX agent

arXiv:2607.25140v1 Announce Type: new Abstract: This paper studies the behavior of language models in a multi-agent crowd simulation, focusing on how affect propagates among agents that perceive and a

28 Jul 2026

A New Role for Relevance: Guiding Corpus Interaction in Agentic Search

AgentsDGX agent

arXiv:2607.24223v1 Announce Type: new Abstract: Relevance is a query-dependent estimate of whether a document or excerpt contains useful evidence. Existing retrieval agents use relevance to select top

Beyond Success Rate: Cost-Aware Evaluation of Offensive and Defensive Security Agents

AgentsDGX agent

arXiv:2607.15263v3 Announce Type: replace-cross Abstract: Security-agent evaluations commonly measure peak offensive capability under generous inference budgets, emphasizing vulnerability discovery, e

DeepLens Diagnosis Agent: Agentic Workflow Design Lets a Small Reasoning Model Compete with Frontier LLMs

Model ReleasesDGX agent

arXiv:2607.22555v1 Announce Type: new Abstract: Medical diagnosis is a multi-stage process: extract facts, consult knowledge, generate a differential analysis, and select the best diagnosis with expla

EviBack: Search-Agent Reinforcement Learning via Evidence-Constrained Teacher Backoff

AgentsDGX agent

arXiv:2607.23955v1 Announce Type: new Abstract: Reinforcement learning enables Agentic RAG systems to learn multi-turn search from verifiable outcome rewards, but all- zero rollout groups provide no c

Execution-Grounded Security Testing for Coding Agents in Software Engineering Pipelines

SafetyDGX agent

arXiv:2607.22569v1 Announce Type: new Abstract: Coding agents are increasingly integrated into system operations, where their tool use can directly modify project artifacts, execution environments, an

How Do Practitioners Build SE Agents? Insights from a Mixed-Methods Study

AgentsDGX agent

arXiv:2607.10856v2 Announce Type: replace-cross Abstract: The rise of Software Engineering (SE) agents, i.e., LLM-based agents that can understand large codebases and carry out engineering tasks with

← Previous
1…1920212223…294
Next →