AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,607 results
Agents

Do Latent Channels Actually Communicate? A Causal Audit of Latent Multi-Agent LLM

DGX agent

arXiv:2607.26773v1 Announce Type: new Abstract: Latent communication in large language model (LLM)-based multi-agent systems (MAS) transmits continuous internal representations instead of text, but gr

agentsarxiv-cs-ai
31 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Agents

Can AI agents conduct open-ended AI research? Early evidence from two case studies

DGX agent

arXiv:2607.27191v1 Announce Type: cross Abstract: Forecasts of explosive AI progress hinge on AI agents automating AI research. But evidence on whether agents can carry out open-ended AI research is t

agentsarxiv-cs-lg
30 Jul 2026
Agents

(Im)Paired Programming: Coding Agents Improve Productivity but Harm Understanding

DGX agent

arXiv:2607.26375v1 Announce Type: new Abstract: Coding agents (e.g., Cursor) improve developer productivity by optimizing task completion, but shifting users from writing code to prompting and reviewi

agentsarxiv-cs-cl
30 Jul 2026
Model Releases

Two Calls Beat Five Agents: Evaluating Multi-Agent Pipelines Against Self-Refinement for Local Language Models

DGX agent

arXiv:2607.26922v1 Announce Type: new Abstract: Multi-agent LLM pipeline systems break down the task among multiple roles for better reasoning, but are benchmarked mainly with large-scale commercial m

model-releasesarxiv-cs-lg
30 Jul 2026
Agents

How Affect Propagates among LLM Agents: Emergent Emotional Contagion in Crowd Simulation

DGX agent

arXiv:2607.25140v1 Announce Type: new Abstract: This paper studies the behavior of language models in a multi-agent crowd simulation, focusing on how affect propagates among agents that perceive and a

agentsarxiv-cs-ai
29 Jul 2026
Agents

A New Role for Relevance: Guiding Corpus Interaction in Agentic Search

DGX agent

arXiv:2607.24223v1 Announce Type: new Abstract: Relevance is a query-dependent estimate of whether a document or excerpt contains useful evidence. Existing retrieval agents use relevance to select top

agentsarxiv-cs-cl
28 Jul 2026
Agents

Beyond Success Rate: Cost-Aware Evaluation of Offensive and Defensive Security Agents

DGX agent

arXiv:2607.15263v3 Announce Type: replace-cross Abstract: Security-agent evaluations commonly measure peak offensive capability under generous inference budgets, emphasizing vulnerability discovery, e

agentsarxiv-cs-ai
28 Jul 2026
Model Releases

DeepLens Diagnosis Agent: Agentic Workflow Design Lets a Small Reasoning Model Compete with Frontier LLMs

DGX agent

arXiv:2607.22555v1 Announce Type: new Abstract: Medical diagnosis is a multi-stage process: extract facts, consult knowledge, generate a differential analysis, and select the best diagnosis with expla

model-releasesarxiv-cs-ai
28 Jul 2026
Agents

EviBack: Search-Agent Reinforcement Learning via Evidence-Constrained Teacher Backoff

DGX agent

arXiv:2607.23955v1 Announce Type: new Abstract: Reinforcement learning enables Agentic RAG systems to learn multi-turn search from verifiable outcome rewards, but all- zero rollout groups provide no c

agentsarxiv-cs-ai
28 Jul 2026
Safety

Execution-Grounded Security Testing for Coding Agents in Software Engineering Pipelines

DGX agent

arXiv:2607.22569v1 Announce Type: new Abstract: Coding agents are increasingly integrated into system operations, where their tool use can directly modify project artifacts, execution environments, an

safetyarxiv-cs-ai
28 Jul 2026
Agents

How Do Practitioners Build SE Agents? Insights from a Mixed-Methods Study

DGX agent

arXiv:2607.10856v2 Announce Type: replace-cross Abstract: The rise of Software Engineering (SE) agents, i.e., LLM-based agents that can understand large codebases and carry out engineering tasks with

agentsarxiv-cs-ai
28 Jul 2026
Agents

‼️ Hugging Face built an interactive replay of the OpenAI agent that breached them. It includes 17,613 logged attacker actions across the 4.…

DGX agent

‼️ Hugging Face built an interactive replay of the OpenAI agent that breached them. It includes 17,613 logged attacker actions across the 4.5-day campaign, with the live command stream and more. https

agentsclem-delangue--x
28 Jul 2026
Agents

ReasFlow: Assisting Reasoning-Centric Scientific Discovery in Applied Mathematics via a Knowledge-Based Multi-Agent System

DGX agent

arXiv:2607.14178v2 Announce Type: replace Abstract: Recent advances in Large Language Models have fueled autonomous AI agents capable of tackling complex scientific tasks, yet existing automated resea

agentsarxiv-cs-ai
28 Jul 2026
Agents

SF-AMS: Strategic Forgetting for Structured Memory in LLM Agent

DGX agent

arXiv:2607.22562v1 Announce Type: new Abstract: Managing long-context dependencies remains a primary bottleneck in LLM agents, as redundant and irrelevant information can degrade multi-step reasoning.

agentsarxiv-cs-ai
28 Jul 2026
Agents

Snowflake debuts Cortex AI Gateway to govern and monitor enterprise AI agents

DGX agent

Snowflake Inc. today introduced Cortex AI Gateway, a centralized control layer that lets enterprises connect, govern and monitor artificial intelligence agents as they reach into models, tools, Model

agentssiliconangle
28 Jul 2026
Agents

Are agent skills always worth using? The answer is no? This paper provides some important insights to understand this more. (bookmark it) Pa…

DGX agent

Are agent skills always worth using? The answer is no? This paper provides some important insights to understand this more. (bookmark it) Paper summary: Adding procedural skills to an agent is usually

agentsdair-ai--x
27 Jul 2026
Safety

Building the enterprise environment for agentic AI

DGX agent

For the enterprise, the promise of agentic AI is much more than just a better chatbot. It is software agents that execute business tasks end-to-end across people, business workflows, data, and systems

safetymit-tech-review
27 Jul 2026
Agents

here’s a video on my approach at applying this to agents w @activegraphai:

DGX agent

here’s a video on my approach at applying this to agents w @activegraphai: 🆕 ActiveGraph: The Log is the Agent my talk from AI Engineer is live!!! 😆 https://www.youtube.com/watch?v=khVX_BUnEwU it's ab

agentsyohei-nakajima--x
27 Jul 2026
Model Releases

Toward User-Conditioned Evaluation of Personal LLM Agents under Temporal Interventions

DGX agent

arXiv:2607.21635v1 Announce Type: new Abstract: Personal agents maintain memories, learned skills, tool configurations, and policy state that evolve with each user. Existing agent benchmarks often eva

model-releasesarxiv-cs-lg
27 Jul 2026
Agents

Streaming Multi-Agent Autoregressive Diffusion Model with World State Registers

DGX agent

arXiv:2607.21594v1 Announce Type: new Abstract: Multi-agent interactive world models should not only generate consistent observations, but also maintain world states that persist across agents and evo

agentsarxiv-cs-cv
24 Jul 2026
Agents

A Survey on Hypergame Theory: Modelling Misaligned Perceptions and Nested Beliefs for Multi-Agent Systems

DGX agent

arXiv:2507.19593v3 Announce Type: replace Abstract: Classical game-theoretic models typically assume rational agents, complete information, and common knowledge of payoffs - assumptions that are often

agentsarxiv-cs-ai
16 Jul 2026
Model Releases

Do Agent Optimizers Compound? A Continual-Learning Evaluation on Terminal-Bench 2.0

DGX agent

arXiv:2607.14004v1 Announce Type: new Abstract: Most reported gains from agent-optimization methods are one-shot: an agent is optimized against a fixed benchmark and the resulting improvement is repor

model-releasesarxiv-cs-ai
16 Jul 2026
Agents

AgentCheck: A Reproduce-Intervene-Mitigate Workbench for LLM Agents over MCP

DGX agent

arXiv:2607.11098v2 Announce Type: replace-cross Abstract: Tool-using LLM agents are mostly evaluated assuming all tools work. When a tool times out, returns a week-stale value, or has its description

agentsarxiv-cs-ai
15 Jul 2026
Agents

Kiro CLI observability: trace and evaluate agent changes with Arize Skills

DGX agent

Use Arize Skills with Kiro CLI to trace coding-agent changes, build datasets from failures, run experiments, and validate prompts before shipping. The post Kiro CLI observability: trace and evaluate a

agentsarize-ai
15 Jul 2026
Agents

WANDR tests how well agents discover large sets of entities and verify specific facts about each one. It provides a dense, interpretable eva…

DGX agent

WANDR tests how well agents discover large sets of entities and verify specific facts about each one. It provides a dense, interpretable eval signal that reveals whether an agent fails, and where. The

agentsperplexity--x
14 Jul 2026
Agents

GPUs for all 🤗 Creating ZeroGPU demos or apps is now available to ALL @huggingface users tell your agent: 'Build a HF ZeroGPU demo for this…

DGX agent

GPUs for all 🤗 Creating ZeroGPU demos or apps is now available to ALL @huggingface users tell your agent: 'Build a HF ZeroGPU demo for this model' New Space: https://huggingface.co/new-space Agent ski

agentsclem-delangue--x
13 Jul 2026
Agents

Pinecone has introduced Nexus, a knowledge engine built specifically for AI agents. Instead of repeatedly searching documents with tradition…

DGX agent

Pinecone has introduced Nexus, a knowledge engine built specifically for AI agents. Instead of repeatedly searching documents with traditional RAG, Nexus compiles knowledge once from sources like data

agentspinecone--x
13 Jul 2026
Agents

Agents That Teach: Towards Designing Incidental Learning Back into AI-Assisted Software Development

DGX agent

arXiv:2607.06101v1 Announce Type: cross Abstract: AI coding agents are rapidly reshaping how software is built, with developers increasingly delegating substantial coding tasks to autonomous agents in

agentsarxiv-cs-ai
8 Jul 2026
Model Releases

Introducing the NemoClaw Deep Agents Blueprint, a reference architecture for building open agent systems developed with @NVIDIA ✅ A fully op…

DGX agent

Introducing the NemoClaw Deep Agents Blueprint, a reference architecture for building open agent systems developed with @NVIDIA ✅ A fully open stack enterprises can own and customize ✅ Benchmark-leadi

model-releasesharrison-chase--x
8 Jul 2026
Agents

Explicit Credit Assignment through Local Rewards and Dependence Graphs in Multi-Agent Reinforcement Learning

DGX agent

arXiv:2601.21523v2 Announce Type: replace Abstract: To promote cooperation in Multi-Agent Reinforcement Learning, the reward signals of all agents can be aggregated together, forming global rewards th

agentsarxiv-cs-lg
7 Jul 2026
Local Ai

How a local-first personal AI agent uses orchestration, memory, tools, LangGraph, background workflows, and child agents

DGX agent

This post describes architectural patterns for building a local-first personal AI agent, covering key components including orchestration frameworks, memory systems, tool integration, LangGraph for wor

local-aiharrison-chase--x
7 Jul 2026
Safety

No Time Like the Present: Agentic Test-Time Training for LLM Agents

DGX agent

arXiv:2607.03441v1 Announce Type: cross Abstract: LLM agents often degrade over long episodes: as trajectories grow, they revisit explored states, repeat failed actions, and lose strategies that previ

safetyarxiv-cs-ai
7 Jul 2026
Agents

Multimodal prompting is clearly the future. I love experimenting with new ways to interact with agents. As a researcher and engineer, I've f…

DGX agent

Multimodal prompting is clearly the future. I love experimenting with new ways to interact with agents. As a researcher and engineer, I've found that the richer the inputs to the agent and the richer

agentsdair-ai--x
3 Jul 2026
Model Releases

PACE: A Proxy for Agentic Capability Evaluation

DGX agent

arXiv:2607.02032v1 Announce Type: new Abstract: Evaluating LLM agents on benchmarks like SWE-Bench and GAIA can be expensive, time-consuming, and requires complex infrastructure. A single evaluation c

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

A Systematic Approach to Multi-Agent AI from Advanced Regulatory Control Theory: Safe and Auditable LLM Operator Agents for Process Control

DGX agent

arXiv:2606.30877v1 Announce Type: cross Abstract: Recent literature shows that large language models (LLMs) are useful for general-purpose tasks yet perform poorly on specific domain ones. One reason

model-releasesarxiv-cs-lg
1 Jul 2026
Model Releases

Claude Fable 5 is available in Devin. You can use Fable 5 in Devin Cloud’s Ultra agent, our smartest and most capable agent, which excels at…

DGX agent

Claude Fable 5 is available in Devin. You can use Fable 5 in Devin Cloud’s Ultra agent, our smartest and most capable agent, which excels at long-horizon tasks and debugging. Claude Fable 5 is also av

model-releasescognition-ai--x
1 Jul 2026
Agents

Hierarchical Message-Passing Policies for Multi-Agent Reinforcement Learning

DGX agent

arXiv:2507.23604v2 Announce Type: replace Abstract: Decentralized Multi-Agent Reinforcement Learning (MARL) methods allow for learning scalable multi-agent policies, but suffer from partial observabil

agentsarxiv-cs-lg
1 Jul 2026
Model Releases

Some of LangSmith's biggest and most engaged customers trace with: - @aisdk - Claude Agents SDK - OpenAI Agent SDK - No framework at all And…

DGX agent

Some of LangSmith's biggest and most engaged customers trace with: - @aisdk - Claude Agents SDK - OpenAI Agent SDK - No framework at all And I'm sure we'll see tons of http://pi.dev in the near future

model-releasesharrison-chase--x
1 Jul 2026
Agents

HyphaeDB: A Living Knowledge Topology for Agent-First Memory

DGX agent

arXiv:2606.28781v1 Announce Type: new Abstract: Every existing vector database and agent memory framework treats memory as passive storage that agents query explicitly. No system propagates knowledge

agentsarxiv-cs-ai
30 Jun 2026
Agents

Linguistic Firewall: Geometry as Defense in Multi-Agent Systems Routing

DGX agent

arXiv:2606.30555v1 Announce Type: new Abstract: The rapid integration of Large Language Models (LLMs) has driven the evolution of Multi-Agent Systems (MAS), where specialized agents collaborate to exe

agentsarxiv-cs-ai
30 Jun 2026
Model Releases

SEO Agent. Security Agent. Design in Claude. Replit Desktop. Shopify on Replit. Skills + Custom Instructions. 450+ integrations. Package Fir…

DGX agent

SEO Agent. Security Agent. Design in Claude. Replit Desktop. Shopify on Replit. Skills + Custom Instructions. 450+ integrations. Package Firewall. And more! ✨ All shipped in June! 🤯 And there's a good

model-releasesreplit--x
30 Jun 2026
Tools

Introducing Cursor for iOS. Build from anywhere by launching always-on cloud agents. Or remotely control agents running on your computer fro…

DGX agent

Introducing Cursor for iOS. Build from anywhere by launching always-on cloud agents. Or remotely control agents running on your computer from the app. Composer 2.5 is 75% off in the app now through Ju

toolscursor--x
29 Jun 2026
Model Releases

Supercharging the agentic era with Spanner’s multi-model architecture

DGX agent

In the agentic era, the role of the database has fundamentally changed. It is no longer a passive repository; it’s a critical context engine designed to ground generative AI apps, models and power aut

model-releasesgoogle-cloud-ai
29 Jun 2026
Model Releases

we just released dynamic subagents, which let your agent programmatically orchestrate subagents in a code interpreter. this lets agents do w…

DGX agent

we just released dynamic subagents, which let your agent programmatically orchestrate subagents in a code interpreter. this lets agents do work at scale that tool calls can't reliably handle, like pro

model-releasesharrison-chase--x
29 Jun 2026
Tutorials

🧑‍🏫3hr Long Deep Agents Course Great course from a member of the community on Deep Agents Covers task planning, file systems for context m…

DGX agent

🧑‍🏫3hr Long Deep Agents Course Great course from a member of the community on Deep Agents Covers task planning, file systems for context management, subagent-spawning, and long-term memory https://www

tutorialsharrison-chase--x
27 Jun 2026
Applications

Agents are easy to demo locally. The hard part is shipping them inside a real app. We published a deployment cookbook for @LangChain agents:…

DGX agent

Agents are easy to demo locally. The hard part is shipping them inside a real app. We published a deployment cookbook for @LangChain agents: full-stack examples with streaming UI, subagents, thread hi

applicationsharrison-chase--x
25 Jun 2026
Agents

MAS-PromptBench: When Does Prompt Optimization Improve Multi-Agent LLM Systems?

DGX agent

arXiv:2606.23664v1 Announce Type: new Abstract: Multi-agent systems (MAS) offer a scalable path forward for agentic AI, comprising multiple LLM-based agents, each assigned a system prompt and a positi

agentsarxiv-cs-lg
23 Jun 2026
Model Releases

Bittensor Agent Arenas as a Trajectory Primitive: Distilling a Shopping Agent from ShoppingBench Subnet Traces

DGX agent

arXiv:2606.10064v1 Announce Type: cross Abstract: Small-model agentic post-training is bottlenecked less by the algorithm than by the trajectory substrate it consumes. Leading recipes (RLVR, group-rel

model-releasesarxiv-cs-ai
10 Jun 2026
← Previous
1…2526272829…367
Next →