AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,761 results
Agents

Evaluating Generative Agents with Actions Grounded in Socially Distributed Task Environments using Incognita

DGX agent

arXiv:2607.02975v1 Announce Type: new Abstract: Effective agency in social environments depends on when an agent seeks knowledge, when it acts, and whether its actions are justified by acquired inform

agentsarxiv-cs-ai
7 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Governed Individuation: Cryptographically Decoupling an Agent's Learning from Its Authority

DGX agent

arXiv:2607.04613v1 Announce Type: new Abstract: Autonomous agents are moving from sandboxed text generators to operators of code, data, and physical infrastructure, and they increasingly learn while d

model-releasesarxiv-cs-ai
7 Jul 2026
Agents

MRMS: A Multi-Resolution Memory Substrate for Long-Lived AI Agents

DGX agent

arXiv:2607.04617v1 Announce Type: new Abstract: Long-lived AI agents require continuity across interactions, but continuity cannot be obtained by simply extending the prompt window. An agent must pres

agentsarxiv-cs-ai
7 Jul 2026
Agents

Rethinking Scientific Discovery in an Agentic Era

DGX agent

arXiv:2607.03863v1 Announce Type: new Abstract: Artificial intelligence has advanced scientific discovery, but most AI4Science systems remain fragmented tools that rely on humans to coordinate problem

agentsarxiv-cs-cl
7 Jul 2026
Agents

SelfMem: Self-Optimizing Memory for AI Agents

DGX agent

arXiv:2607.03726v1 Announce Type: new Abstract: While current AI agents support increasingly long context windows, tool use, and skill execution for long-horizon tasks, they still require memory syste

agentsarxiv-cs-cl
7 Jul 2026
Model Releases

SovereignNegotiation-Bench: Evaluating User-Owned Personal Agents In Delegated Bargaining Under Privacy, Consent, Evidence, And Institutional Pressure

DGX agent

arXiv:2607.02814v1 Announce Type: cross Abstract: Personal agents will increasingly negotiate on behalf of users: splitting costs with other personal agents, appealing platform decisions, escalating s

model-releasesarxiv-cs-ai
7 Jul 2026
Safety

The agent creates, we validate: A Lightweight Framework for Agentic Artifact Generation

DGX agent

arXiv:2607.02615v1 Announce Type: cross Abstract: Generating structured artifacts with Large Language Models - e.g. database queries, threat framework mappings, entity schemas - is relatively straight

safetyarxiv-cs-ai
7 Jul 2026
Agents

Untrusted Content Masking for Web Agents with Security Guarantees

DGX agent

arXiv:2607.05277v1 Announce Type: cross Abstract: Defenses that provide security guarantees against prompt injection attacks rely on strict isolation between trusted instructions and untrusted data. I

agentsarxiv-cs-lg
7 Jul 2026
Model Releases

When Claws Remember but Do Not Tell: Stealthy Memory Injection in Persistent Personal Agents

DGX agent

arXiv:2607.05189v1 Announce Type: cross Abstract: Persistent personal agents combine long-term memory with access to users' external environments, enabling personalized foreground assistance and proac

model-releasesarxiv-cs-ai
7 Jul 2026
Agents

Most agentic retrieval demos assume clean, well-structured documents. Enterprise reality is often different, consisting of messy PDFs where …

DGX agent

Most agentic retrieval demos assume clean, well-structured documents. Enterprise reality is often different, consisting of messy PDFs where critical information is buried across tables, figures, and c

agentsjerry-liu--x
6 Jul 2026
Agents

We're teaming up with Hugging Face to make open agent traces the fuel for the next generation of open coding models. You can now upload @Dro…

DGX agent

Hugging Face is partnering to make open agent traces available as training data for developing next-generation open-source coding models. This collaboration enables developers to upload agent traces,

agentsclem-delangue--x
6 Jul 2026
Agents

Active Graphの「The Log Is the Agent」で何が変わる? AIエージェントを“会話”ではなく“ログ”で動かす発想です。 監査・差し戻し・再実行が重いチームほど、この設計が効きます。

DGX agent

Active Graph proposes a paradigm shift in AI agent architecture by using logs as the primary mechanism for agent operations rather than conversation-based interactions. This design approach is particu

agentsyohei-nakajima--x
5 Jul 2026
Agents

Sakana AI is heading to #ICML2026 in Seoul (July 6–11)! 🐟🇰🇷 Our team will present 11 papers spanning multi-agent coordination, sparse and…

DGX agent

Sakana AI is heading to #ICML2026 in Seoul (July 6–11)! 🐟🇰🇷 Our team will present 11 papers spanning multi-agent coordination, sparse and efficient LLMs, test-time scaling, long-term memory, and agent

agentsdavid-ha--x
4 Jul 2026
Model Releases

GroundEval: A Deterministic Replacement for LLM-as-Judge in Stateful Agent Evaluation

DGX agent

arXiv:2606.22737v2 Announce Type: replace Abstract: Before letting an agent operate over real context, can you prove it used the right evidence? GroundEval turns that question into a deterministic tes

model-releasesarxiv-cs-ai
3 Jul 2026
Agents

SkillCoach: Self-Evolving Rubrics for Evaluating and Enhancing Agentic Skill-Use

DGX agent

arXiv:2607.01874v1 Announce Type: new Abstract: Skills are becoming a reusable operational layer for LLM agents, encoding SOPs, domain rules, tool workflows, scripts, and validation routines. In reali

agentsarxiv-cs-ai
3 Jul 2026
Agents

Vercel's Andrew Qu on why agents are a new kind of software

DGX agent

Andrew Qu from Vercel discusses how AI agents represent a fundamentally different category of software compared to traditional applications, exploring their unique characteristics and implications for

agentslatent-space
3 Jul 2026
Agents

Multi-Agent Teams Hold Experts Back

DGX agent

Multi-agent LLM systems are increasingly deployed as autonomous collaborators, where agents interact freely rather than execute fixed, pre-specified workflows. In such settings, effective coordination

agentsapple-ml-research
2 Jul 2026
Agents

SlowBA: An efficiency backdoor attack towards VLM-based GUI agents

DGX agent

arXiv:2603.08316v3 Announce Type: replace-cross Abstract: Modern vision-language-model (VLM) based graphical user interface (GUI) agents are expected not only to execute actions accurately but also to

agentsarxiv-cs-cl
2 Jul 2026
Model Releases

Beyond expert users: agents should help users construct preferences, not just elicit them

DGX agent

arXiv:2606.30863v1 Announce Type: new Abstract: Agents typically assume an expert user -- one with well-formed preferences about what they want -- and default to clarifying questions whenever the task

model-releasesarxiv-cs-ai
1 Jul 2026
Agents

the log is the agent!

DGX agent

This post explores the concept that an agent's log or execution history serves as the primary mechanism for decision-making and learning, suggesting that the sequential record of actions and outcomes

agentsyohei-nakajima--x
1 Jul 2026
Model Releases

Agent-Computer Observation Interfaces Enable Dynamic Computer Use

DGX agent

arXiv:2606.29472v1 Announce Type: new Abstract: SWE-agent established the action interface as an underexplored design axis for software-engineering agents; we make the analogous case for the observati

model-releasesarxiv-cs-ai
30 Jun 2026
Agents

Agentic AI for ISAC: Analysis, Framework, and Case Study

DGX agent

arXiv:2512.15044v2 Announce Type: replace Abstract: Integrated sensing and communication (ISAC) has emerged as a key development direction in the sixth-generation (6G) era, which provides essential su

agentsarxiv-cs-ai
30 Jun 2026
Agents

Agentic AI Health Assistants for Proactive Patient Care and Scalable Risk Management

DGX agent

This article likely explores how agentic AI systems can proactively monitor patient health data and manage clinical risks at scale, improving healthcare delivery through autonomous AI agents. It proba

agentsvultr
30 Jun 2026
Agents

agents that can write code can solve problems more reliably but you need to make sure you execute that untrusted code in a safe environment …

DGX agent

Code-writing AI agents can solve complex problems more effectively than agents using other approaches, but deploying them requires executing potentially untrusted generated code within sandboxed or is

agentsharrison-chase--x
30 Jun 2026
Agents

Capability Gates Are Not Authorization: Confused-Deputy Failures in LLM Agent Frameworks

DGX agent

arXiv:2606.28679v1 Announce Type: cross Abstract: Tool-using LLM agents increasingly read untrusted content while holding side-effecting tools such as payments, email, CRM, and infrastructure APIs, ye

agentsarxiv-cs-ai
30 Jun 2026
Model Releases

Demonstration-Free Robotic Control via LLM Agents

DGX agent

arXiv:2601.20334v2 Announce Type: replace-cross Abstract: Robotic manipulation has increasingly adopted vision-language-action (VLA) models, which achieve strong performance but typically require task

model-releasesarxiv-cs-ai
30 Jun 2026
Agents

Evaluate agents with Harbor + LangChain ‼️

DGX agent

This post likely discusses how to evaluate AI agents built with LangChain using Harbor, a tool for testing and monitoring. It covers practical approaches for assessing agent performance, reliability,

agentsharrison-chase--x
30 Jun 2026
Agents

Experience Graphs: The Data Foundation for Self-Improving Agents

DGX agent

arXiv:2606.29823v1 Announce Type: cross Abstract: The database community has repeatedly advanced the state of the art by recognizing that new workloads demand new system architectures. We argue that l

agentsarxiv-cs-ai
30 Jun 2026
Model Releases

Multi-Agent Routing as Set-Valued Prediction: A WildChat Benchmark and Cost-Aware Evaluation

DGX agent

arXiv:2606.28925v1 Announce Type: cross Abstract: Tool and agent routing from natural-language prompts is naturally a set-valued prediction problem: a single query may require multiple agents, while o

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

SEATauBench: Adapting Tool-Agent-User Evaluation Into Low-Resource Southeast Asian Languages

DGX agent

arXiv:2606.28715v1 Announce Type: cross Abstract: While AI development and evaluation for Southeast Asia (SEA) has grown rapidly, agent capabilities in regional languages are still poorly understood d

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

TraceLab: Characterizing Coding Agent Workloads for LLM Serving

DGX agent

arXiv:2606.30560v1 Announce Type: cross Abstract: Coding agents are rapidly becoming a major application of agentic LLMs, but serving them efficiently remains challenging. Progress on this challenge r

model-releasesarxiv-cs-ai
30 Jun 2026
Agents

Vercel Agent has updated pricing

DGX agent

Vercel Agent pricing has been updated, likely reflecting changes to the cost structure for using Vercel's AI agent capabilities. The specific details of the new pricing tiers, feature inclusions, and

agentsvercel-blog
30 Jun 2026
Agents

What Capable Agents Must Know: Selection Theorems for Robust Decision-Making under Uncertainty

DGX agent

arXiv:2603.02491v3 Announce Type: replace-cross Abstract: As artificial agents become increasingly capable, what internal structure is necessary for an agent to act competently under uncertainty? Clas

agentsarxiv-cs-ai
30 Jun 2026
Agents

Baz releases Baz Planner, which uses four specialized AI agents to analyze code at the planning stage, and extends its seed funding by 9M to 17M (Mike Wheatley/SiliconANGLE)

DGX agent

Mike Wheatley / SiliconANGLE: Baz releases Baz Planner, which uses four specialized AI agents to analyze code at the planning stage, and extends its seed funding by 9M to 17M — Agentic coding startup

agentstechmeme
29 Jun 2026
Model Releases

Contagion Networks: Evaluator Preference Propagation in Multi-Agent LLM Systems

DGX agent

arXiv:2606.20493v2 Announce Type: replace-cross Abstract: When large language models serve as evaluators in multi-agent systems, their strategy preferences -- whether induced by explicit prompts or by

model-releasesarxiv-cs-ai
29 Jun 2026
Agents

Debugging production agents with Amazon Bedrock AgentCore Observability

DGX agent

In this post, you learn how to debug production agent failures using built-in observability capabilities. We walk through common failure patterns, show how to analyze agent behavior with traces and me

agentsaws-ml-blog
29 Jun 2026
Agents

Hierarchical Control in Multi-Agent Games: LLM-based Planning and RL Execution

DGX agent

arXiv:2606.20014v2 Announce Type: replace-cross Abstract: Reinforcement learning (RL) has achieved strong performance in sequential decision-making, yet scaling to complex multi-agent environments rem

agentsarxiv-cs-ai
29 Jun 2026
Agents

i've been using cursor mobile on the go for the last weeks, and having access to all cloud agents from everywhere is really nice go on a wal…

DGX agent

i've been using cursor mobile on the go for the last weeks, and having access to all cloud agents from everywhere is really nice go on a walk, get an idea, dictate it in the app come back from walk to

agentselon-musk--x
29 Jun 2026
Model Releases

LLawCo: Learning Laws of Cooperation for Modeling Embodied Multi-Agent Behavior

DGX agent

arXiv:2606.28182v1 Announce Type: cross Abstract: Embodied agents operating in decentralized and partially observable environments have attracted growing attention in recent years. However, existing l

model-releasesarxiv-cs-ai
29 Jun 2026
Agents

NEW paper from Google (bookmark it) It's on advancing automated scientific review. Just pay attention to the focus on agentic verification w…

DGX agent

NEW paper from Google (bookmark it) It's on advancing automated scientific review. Just pay attention to the focus on agentic verification which is something I've been writing about recently. AI is ac

agentsdair-ai--x
29 Jun 2026
Agents

Row-Bot runs a LangGraph agent at its core. And supports 2 voice pipelines: - Local STT using whisper + Local TTS using kokoro - A real-time…

DGX agent

Row-Bot runs a LangGraph agent at its core. And supports 2 voice pipelines: - Local STT using whisper + Local TTS using kokoro - A real-time voice pipeline using GPT realtime 2. Both support full agen

agentsharrison-chase--x
29 Jun 2026
Agents

💪deep agents harness

DGX agent

💪deep agents harness The AI SDK Harness API now supports @OpenCode and @LangChain Deep Agents through a single unified interface. Use 𝙷𝚊𝚛𝚗𝚎𝚜𝚜𝙰𝚐𝚎𝚗𝚝 to run any supported runtime without changing your ap

agentsharrison-chase--x
26 Jun 2026
Model Releases

The Red Queen Godel Machine: Co-Evolving Agents and Their Evaluators

DGX agent

arXiv:2606.26294v1 Announce Type: cross Abstract: Self-improving agents are state-of-the-art (SOTA) on agentic coding benchmarks and have recently been extended to general domains. However, their sear

model-releasesarxiv-cs-ai
26 Jun 2026
Agents

Diagnosing and Mitigating Compounding Failures in Agentic Persuasion via Taxonomic Strategy Retrieval

DGX agent

arXiv:2606.24976v1 Announce Type: cross Abstract: Foundation-model agents in multi-step, open-ended environments frequently suffer from compounding errors, where early mistakes contaminate long-horizo

agentsarxiv-cs-cl
25 Jun 2026
Agents

excited to be speaking at @aiDotEngineer World Fair next week on Improving Agents, Continual Learning, and why we think a large part of it i…

DGX agent

excited to be speaking at @aiDotEngineer World Fair next week on Improving Agents, Continual Learning, and why we think a large part of it is...Data Mining! Trace Mining is how we understand agent beh

agentsharrison-chase--x
25 Jun 2026
Agents

Debate2Create: Robot Co-design via Multi-Agent LLM Debate

DGX agent

arXiv:2510.25850v3 Announce Type: replace-cross Abstract: We introduce Debate2Create (D2C), a multi-agent LLM framework that formulates robot co-design as structured, iterative debate grounded in phys

agentsarxiv-cs-lg
24 Jun 2026
Agents

Learn *anything* with our new /learn agent skill.

DGX agent

Learn *anything* with our new /learn agent skill. Obsessed with our new /learn skill. It's my favorite way of learning and researching topics. The agent creates a learning plan and a learning hub (art

agentsdair-ai--x
24 Jun 2026
Model Releases

📣📣 Meet Qwen-AgentWorld — a native language world model that simulates 7 agent environments (MCP, Search, Terminal, SWE, Web, OS, Android)…

DGX agent

📣📣 Meet Qwen-AgentWorld — a native language world model that simulates 7 agent environments (MCP, Search, Terminal, SWE, Web, OS, Android) within a single model. Environment modeling is the training o

model-releasesqwen--x
24 Jun 2026
← Previous
1…6061626364…371
Next →