AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,737 results
7 Jul 2026

DualView: Preventing Indirect Prompt Injection in Personal AI Agents

Model ReleasesDGX agent

arXiv:2607.03821v1 Announce Type: cross Abstract: Personal AI agents that run on the user's local machine, such as OpenClaw, automate daily tasks including web search, email, and file management. Thei

is anyone standardizing on ATIF as a format for agent traces? or a different format? or is it still wild wild west, roll your own

AgentsDGX agent

This post discusses the current state of standardization for agent execution trace formats, questioning whether ATIF (Agent Trace Interchange Format) or alternative formats have achieved adoption, or

Latent Programming Horizons in Coding Agents

AgentsDGX agent

arXiv:2607.05188v1 Announce Type: new Abstract: A coding agent solving a software-engineering task spends dozens of steps reasoning, editing code, and running tests, yet little is known about what the

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Learning Task-Sufficient World Models by Synergizing Agentic Exploration and Structured Modeling

AgentsDGX agent

arXiv:2607.04409v1 Announce Type: new Abstract: Learning and planning in imagination using world models provides an effective paradigm for training agents for decision-making. However, existing approa

Measuring Harness-Induced Belief Divergence in Multi-Step LLM Agents

Model ReleasesDGX agent

arXiv:2607.04528v1 Announce Type: new Abstract: Software-agent benchmarks usually report whether an agent solves a task, but the agent reaches that outcome through a harness that controls what it sees

Organizational Memory for Agentic Business Process Execution

AgentsDGX agent

arXiv:2607.03228v1 Announce Type: new Abstract: LLM-based agents offer new opportunities for automating business process execution beyond the limits of rule-based systems. However, general-purpose LLM

Social Networks of LLM Agents

AgentsDGX agent

arXiv:2607.03695v1 Announce Type: new Abstract: Large language model (LLM) agents are increasingly deployed in interacting populations, raising the question of what such populations come to believe co

We're hosting the LA Agentic AI Meetup this Thursday, July 9, 5 to 7pm at Gulp in Playa Vista. Come talk to engineers, founders, and builder…

AgentsDGX agent

We're hosting the LA Agentic AI Meetup this Thursday, July 9, 5 to 7pm at Gulp in Playa Vista. Come talk to engineers, founders, and builders working across RAG, agentic workflows, and the broader AI

5 Jul 2026

@hwchase17 Yes using deep agents and okf both for my wiki. https://github.com/varunyn/wiki-langGraph

AgentsDGX agent

This post discusses using deep agents and OpenKeyFlow (OKF) for a wiki project built with LangGraph, a framework for constructing agent-based applications. The developer is leveraging LangGraph's capa

3 Jul 2026

Can coding-agents replicate scientific ML papers? We know this is possible because we can already do this @dair_ai. Still a great read. So t…

AgentsDGX agent

Can coding-agents replicate scientific ML papers? We know this is possible because we can already do this @dair_ai. Still a great read. So they try to replicate an ML paper from its materials alone. T

EO-Agents: A Three-Agent LLM Pipeline for Earth Observation Hypothesis Generation

Model ReleasesDGX agent

arXiv:2607.01584v1 Announce Type: new Abstract: Large language models have recently been explored for scientific hypothesis generation, but most prior work relies on unstructured literature and free-f

Safeguarding LLM Agents from Misalignment through Provenance Analysis

SafetyDGX agent

arXiv:2607.01236v1 Announce Type: cross Abstract: As LLM agents gain increasing access to powerful tools, ensuring that their actions are aligned with the user's intent becomes critical. When an agent

The Agentic Garden of Forking Paths

AgentsDGX agent

arXiv:2607.01507v1 Announce Type: new Abstract: Empirical research rarely admits a unique analysis. Different analytical choices can lead to different conclusions from the same data, yet these hidden

UA-ChatDev: Uncertainty-Aware Multi-Agent Collaboration for Reliable Software Development

Model ReleasesDGX agent

arXiv:2607.02186v1 Announce Type: new Abstract: Software development is a complex task that demands cooperation among agents with diverse roles. Large language models (LLMs) have enabled autonomous mu

WorkPods agent memory as wikis. @langchain #openwiki @cognition #deepwiki @karpathy #llmwiki @Factory #autowiki @hwchase17 @BraceSproul

AgentsDGX agent

This post discusses using WorkPods agent memory systems organized as wikis, likely exploring how LLM agents can maintain and access persistent knowledge repositories in wiki format for improved contex

2 Jul 2026

Conversable Complexity: Agentic LLM Collectives as Interpretable Substrates

AgentsDGX agent

arXiv:2607.01047v1 Announce Type: new Abstract: Complexity and interpretability rarely coincide: systems rich enough for complex behaviours to emerge are usually too opaque to question, while transpar

How to evaluate AI agents, avoid reward hacking, and build better specs

AgentsDGX agent

Agent evals are repeatable tests that score whether AI agents completed a task correctly. Learn how to design rubrics, test suites, and trace-based evals that catch failures and prevent reward hacking

LLM Wikis are being slept on. I argue that creating knowledge bases with LLMs or coding agents is one of the most valuable applications of A…

Model ReleasesDGX agent

LLM Wikis are being slept on. I argue that creating knowledge bases with LLMs or coding agents is one of the most valuable applications of AI today. It's about being intentional in building and scalin

Special episode with @bentannyhill on the Max Agency podcast. A month ago, his team shipped LangSmith Engine, our agent that hunts through y…

AgentsDGX agent

Special episode with @bentannyhill on the Max Agency podcast. A month ago, his team shipped LangSmith Engine, our agent that hunts through your agent's failures, prioritizes issues, and drafts the fix

1 Jul 2026

ACE: Pluggable Adaptive Context Elasticizer across Agents

AgentsDGX agent

arXiv:2606.31564v1 Announce Type: new Abstract: The increasing complexity of agentic tasks has led to rapidly growing trajectory lengths, which poses significant challenges for large language model (L

An Agentic AI Framework to Accelerate Scientific Discovery in Plant Phenotyping

AgentsDGX agent

arXiv:2606.31831v1 Announce Type: new Abstract: High-throughput plant phenotyping now generates image derived datasets far faster than scientists can analyze them. At Oak Ridge National Laboratory's A

InfiniteWeb: Scalable Web Environment Synthesis for GUI Agent Training

AgentsDGX agent

arXiv:2601.04126v3 Announce Type: replace-cross Abstract: GUI agents that interact with graphical interfaces on behalf of users represent a promising direction for practical AI assistants. However, tr

ShopX: A Foundation Model for Intent-to-Item Fulfillment in Agentic Shopping

AgentsDGX agent

arXiv:2606.31693v1 Announce Type: cross Abstract: The wave of AI-native applications is moving shopping beyond page- and feed-based browsing toward intent-driven experiences orchestrated by LLM agents

30 Jun 2026

Agentic Coding on Supabase with OpenCode

AgentsDGX agent

The article covers the Supabase plugin for OpenCode, an AI coding agent, which bundles the Supabase MCP server and agent skills for direct database and project interaction . The plugin ships official

AI Trading's Alpha Singularity: Emergent Market Reasoning through Agent-to-Agent Self-Evolution

Local AiDGX agent

arXiv:2606.29194v1 Announce Type: new Abstract: Automated alpha mining holds the scoring function fixed and varies the search algorithm over it. A search that converges against a fixed scorer overfits

Clarus: Coordinating Autonomous Research Agents toward Web-Scale Scientific Collaboration

AgentsDGX agent

arXiv:2606.30246v1 Announce Type: new Abstract: Existing autonomous research agents can support parts of the research process, but most systems still treat research as either an isolated assistant tas

GUICrafter: Weakly-Supervised GUI Agent Leveraging Massive Unannotated Screenshots

AgentsDGX agent

arXiv:2606.29705v1 Announce Type: new Abstract: Data, as the fundamental substrate of modern intelligence, has greatly driven the development of current foundation models. Naturally, researchers aim t

Hybrid Retriever Evolution for Multimodal Document Reasoning Agents

AgentsDGX agent

arXiv:2606.29648v1 Announce Type: cross Abstract: Different retrievers, including lexical, semantic, and multimodal approaches, provide highly complementary strengths for multimodal document understan

It Lied to a Doctor to Buy Poison Ingredients: Quantifying Real-World Misuse of Phone-use Agents

Model ReleasesDGX agent

arXiv:2606.27944v1 Announce Type: cross Abstract: Phone-use Agents can execute complex tasks end to end across real mobile applications. By operating a real device on the user's behalf, they reach far

Mandol: An Agglomerative Agent Memory System for Long-Term Conversations

AgentsDGX agent

arXiv:2606.29778v1 Announce Type: cross Abstract: Long-term conversational agents need to remember and query cross-session, multi-typed information with complex correlations. Existing agent memory sys

Modeling Earth-Scale Human-Like Societies with One Billion Agents

AgentsDGX agent

arXiv:2506.12078v2 Announce Type: replace-cross Abstract: Understanding the dynamic evolution of complex social phenomena requires both high-fidelity modeling of human behavior and large-scale simulat

Neural Procedural Memory: Empowering LLM Agents with Implicit Activation Steering

AgentsDGX agent

arXiv:2606.29824v1 Announce Type: cross Abstract: While Large Language Models (LLMs) excel as static solvers, transforming them into autonomous agents remains challenging. This transition requires con

Self-Evolving World Models for LLM Agent Planning

AgentsDGX agent

arXiv:2606.30639v1 Announce Type: new Abstract: World models offer a principled way to equip long-horizon LLM agents with foresight: predictions of action consequences before execution. However, unrel

SkillOpt: Agent skills as trainable parameters

AgentsDGX agent

AI agents often fail because their instructions, or skills, are manually modified with no guarantee of improvement. Learn how SkillOpt turns skill editing into a training process, making agent behavio

29 Jun 2026

Agentic Publication Protocol: An Attempt to Modernize Scientific Publication

AgentsDGX agent

arXiv:2606.27386v1 Announce Type: cross Abstract: Scientific publication is still organized primarily around static manuscripts, even though much of scientific progress depends on tacit know-how: how

Delayed Verification Destabilizes Multi-Agent LLM Belief: Instability Thresholds and Optimal Corrector Placement

AgentsDGX agent

arXiv:2606.27409v1 Announce Type: cross Abstract: Multi-agent large language model (LLM) systems often rely on verifier and critic agents to suppress hallucinations, but verification is delayed. Durin

Govern the Repository, Not the Agent: Measuring Ecosystem-Level Risk in AI-Native Software

Model ReleasesDGX agent

arXiv:2606.28235v1 Announce Type: cross Abstract: Autonomous coding agents now open and merge pull requests in shared repositories at scale, and the field evaluates them the way it has always evaluate

Straiker, which develops tech for securing enterprise AI agents, raised a 64M Series A, bringing its total funding to 85M (Chris Metinko/Axios)

AgentsDGX agent

Chris Metinko / Axios: Straiker, which develops tech for securing enterprise AI agents, raised a 64M Series A, bringing its total funding to 85M — Straiker, which secures AI agents, raised a $64 milli

26 Jun 2026

A Process Harness for Uplifting Legacy Workflows to Agentic BPM: Design and Realization in CUGA FLO

SafetyDGX agent

arXiv:2606.27188v1 Announce Type: new Abstract: We introduce the process harness, a new mechanism for uplifting legacy workflows into Agentic Business Process Management (Agentic BPM) without replacin

Excited to see @cohere open-source how they use AI coding agents to maintain their vLLM fork. 🙌 Keeping a long-lived fork in sync is tricky…

AgentsDGX agent

Excited to see @cohere open-source how they use AI coding agents to maintain their vLLM fork. 🙌 Keeping a long-lived fork in sync is tricky. Their agent treats it as a control loop: rebase onto each u

What happens when AI agents collaborate on open science? At @aiDotEngineer World’s Fair, @james_y_zou will share work on EinsteinArena and D…

AgentsDGX agent

What happens when AI agents collaborate on open science? At @aiDotEngineer World’s Fair, @james_y_zou will share work on EinsteinArena and DSGym, from multi-agent math discovery to better evaluation f

25 Jun 2026

Agentic System as Compressor: Quantifying System Intelligence in Bits

AgentsDGX agent

arXiv:2606.25960v1 Announce Type: new Abstract: Large language models are turning from isolated predictors into agentic systems: they call tools, retrieve evidence, obey environment constraints, use v

Autodata: An agentic data scientist to create high quality synthetic data

AgentsDGX agent

arXiv:2606.25996v1 Announce Type: cross Abstract: We introduce Autodata, a general method that enables AI agents to act as data scientists who build high quality training and evaluation data. We show

Domain-Specific Agents for Cherenkov Telescope Array Control Software and Gamma-Ray Data Analysis

AgentsDGX agent

arXiv:2510.01299v3 Announce Type: replace-cross Abstract: We present domain-adapted large language model agents designed to support Cherenkov Telescope Array operation and data analysis. The agents co

Evaluating AGENTS.md: Are Repository-Level Context Files Helpful for Coding Agents?

AgentsDGX agent

arXiv:2602.11988v2 Announce Type: replace-cross Abstract: A widespread practice in software development is to tailor coding agents to repositories using context files, such as AGENTS.md. Although this

Low Variance Trust Region Optimization with Independent Actors and Sequential Updates in Cooperative Multi-agent Reinforcement Learning

SafetyDGX agent

arXiv:2606.25526v1 Announce Type: new Abstract: Cooperative multi-agent reinforcement learning assumes each agent shares the same reward function and can be trained effectively using the Trust Region

Salesforce launches Help Agent to simplify AI customer service deployment

Model ReleasesDGX agent

Salesforce Inc. is launching a new prepackaged artificial intelligence agent for customer service, enabling organizations to quickly build and deploy AI agents. Today Salesforce announced Help Agent,

24 Jun 2026

Agent architecture is essentially a solved problem. What's not: memory. Great article by @jakebroekhuizen who's a true expert on this

AgentsDGX agent

This post highlights that while agent architecture design has become well-established and solved, memory systems remain an unsolved challenge in AI agent development. The tweet references an article b

🧠LangSmith Engine as Sleep Time Compute Memory for agents is often described as “sleep time compute” or “dreaming” This involves running a …

AgentsDGX agent

🧠LangSmith Engine as Sleep Time Compute Memory for agents is often described as “sleep time compute” or “dreaming” This involves running a background process to analyze agent trajectories and update a

RIFT-Bench: Dynamic Red-teaming For Agentic AI Systems

AgentsDGX agent

arXiv:2606.23927v1 Announce Type: new Abstract: Agentic AI systems powered by large language models (LLMs) are rapidly evolving into autonomous decision-making systems, exposing attack vectors beyond

23 Jun 2026

AgentMisalignment: Measuring the Propensity for Misaligned Behaviour in LLM-Based Agents

Model ReleasesDGX agent

arXiv:2506.04018v3 Announce Type: replace-cross Abstract: As Large Language Model (LLM) agents become more widespread, associated misalignment risks increase. While prior research has studied agents'

Build real agentic apps using CUGA: two dozen working examples on a lightweight harness

AgentsDGX agent

CUGA is a lightweight framework from IBM Research for building agentic AI applications, featuring two dozen practical working examples. The framework enables developers to create autonomous AI agents

ClickHouse brings real-time analytics to agentic AI

AgentsDGX agent

The growing use of AI agents throughout the enterprise is forcing a thorough reevaluation of the data layer. This shift is driven by the need for millisecond responses that enable agents to make decis

Counsel: A Meta-Evaluation Dataset for Agentic Tasks

SafetyDGX agent

arXiv:2606.21627v1 Announce Type: cross Abstract: As agentic systems tackle increasingly complex multi-step tasks, evaluating their trajectories presents a major bottleneck - human annotation of a sin

Learn to use the new eve agentic framework from Vercel. Go try out the hands-on labs now.

AgentsDGX agent

Learn to use the new eve agentic framework from Vercel. Go try out the hands-on labs now. I'm digging the eve agentic framework from Vercel. I like that everything is files, from the tools to the skil

you need docs built for agents

AgentsDGX agent

you need docs built for agents Docs are the eyes and ears of Agents. But we're moving so fast that they are always outdated. So hence an outdated doc also confuses the models. But it's also hard keep

22 Jun 2026

context engineering docs for agentic engineering - plans, research, etc SHOULD NOT be stored in version control: A good docs management syst…

AgentsDGX agent

context engineering docs for agentic engineering - plans, research, etc SHOULD NOT be stored in version control: A good docs management system keeps them: > outside your repo > accesible to agent via

How does it work? Sakana Fugu is itself an LLM, trained to call various LLMs in an agent pool, including instances of itself recursively. Fu…

AgentsDGX agent

How does it work? Sakana Fugu is itself an LLM, trained to call various LLMs in an agent pool, including instances of itself recursively. Fugu dynamically orchestrates the world's best models to tackl

11 Jun 2026

Agentic Environment Engineering for Large Language Models: A Survey of Environment Modeling, Synthesis, Evaluation, and Application

AgentsDGX agent

arXiv:2606.12191v1 Announce Type: cross Abstract: Environments serve as interactive systems for large language model (LLM) based agents across diverse scenarios and play a crucial role in driving the

AI Coding Agents in Social Science: Methodologically Diverse, Empirically Consistent, Interpretively Vulnerable

Model ReleasesDGX agent

arXiv:2606.11456v1 Announce Type: cross Abstract: The deployment of LLM-based agents in scientific analysis raises opposing concerns: that agents may reduce methodological diversity, or that they may

← Previous
1…4041424344…296
Next →