AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlog
84,433Total entries
1Added by human
84,432Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,911 results
Agents

A Cost-Aware, Paired Protocol for Auditing Dynamic Tool Synthesis in Agentic Video Question Answering

DGX agent

arXiv:2607.01469v2 Announce Type: replace Abstract: Agentic Video Question Answering (VideoQA) systems invoke tools during inference, but their tool libraries are fixed, so recurring procedures are re

agentsarxiv-cs-cv
7 Jul 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Safety

ARLArena: A Unified Framework for Stable Agentic Reinforcement Learning

DGX agent

arXiv:2602.21534v3 Announce Type: replace Abstract: Agentic reinforcement learning (ARL) has rapidly gained attention as a promising paradigm for training agents to solve complex, multi-step interacti

safetyarxiv-cs-ai
7 Jul 2026
Agents

LLMoxie: Exploring Agentic AI for Scientific Software Development

DGX agent

arXiv:2607.02703v1 Announce Type: cross Abstract: In this paper, we describe LLMoxie, an institutional AI platform whose three-tiered architecture supports multi-cloud and on-premise inference, a Lite

agentsarxiv-cs-ai
7 Jul 2026
Safety

MAD-PINN: A Decentralized Physics-Informed Machine Learning Framework for Safe and Optimal Multi-Agent Control

DGX agent

arXiv:2509.23960v2 Announce Type: replace-cross Abstract: Co-optimizing safety and performance in large-scale multi-agent systems remains a fundamental challenge. Existing approaches based on multi-ag

safetyarxiv-cs-ai
7 Jul 2026
Agents

my hot take is that evals are the only part of agent engineering that require real thinking

DGX agent

Harrison Chase argues that evaluations (evals) are the most intellectually demanding aspect of agent engineering, distinguishing them from other engineering tasks that may be more routine or formulaic

agentsharrison-chase--x
7 Jul 2026
Agents

OptiAgent: End-to-End Optimization Modeling via Multi-Agent Iterative Refinement

DGX agent

arXiv:2607.05346v1 Announce Type: new Abstract: We propose OptiAgent, a multi-agent framework that, given a natural language description of an Operations Research problem, is able to output a solver-r

agentsarxiv-cs-ai
7 Jul 2026
Model Releases

Radware adds Claude Code protection and compliance reporting to agent security

DGX agent

Application and network security company Radware Ltd. today expanded its Agentic AI Protection product with compliance reporting, deeper visibility into artificial intelligence agent activity, and new

model-releasessiliconangle
7 Jul 2026
Safety

Securing Multi-Tool AI Agent Chains With Dynamic, Real-Time Compositional Policies

DGX agent

arXiv:2607.03423v1 Announce Type: cross Abstract: Modern AI agent implementations such as frontier coding agents chain multiple tools at runtime that create a security surface that per-tool guardrails

safetyarxiv-cs-ai
7 Jul 2026
Model Releases

SpecEyes: Accelerating Agentic Multimodal LLMs via Speculative Perception and Planning

DGX agent

arXiv:2603.23483v2 Announce Type: replace-cross Abstract: Agentic multimodal large language models (MLLMs) (e.g., OpenAI o3 and Gemini Agentic Vision) achieve remarkable reasoning capabilities through

model-releasesarxiv-cs-cl
7 Jul 2026
Agents

Spinning Straw into Gold: Relabeling LLM Agent Trajectories in Hindsight for Successful Demonstrations

DGX agent

arXiv:2607.04235v1 Announce Type: new Abstract: Large language model agents operate in partially observable, long-horizon settings where obtaining supervision remains a major bottleneck. We address th

agentsarxiv-cs-cl
7 Jul 2026
Agents

Storage gets promoted in the agentic AI era

DGX agent

The year 2026 could be remembered as the moment when storage technology received a massive promotion. The reason is that the current transition from simple chatbots to agentic AI systems has raised th

agentssiliconangle
7 Jul 2026
Safety

Strategic Buying Agents

DGX agent

arXiv:2607.04708v1 Announce Type: cross Abstract: Agentic AI is shifting online shopping from search toward delegated purchasing, where autonomous buying agents monitor markets and decide when to buy

safetyarxiv-cs-ai
7 Jul 2026
Agents

TAMA: A Human-AI Collaborative Thematic Analysis Framework Using Multi-Agent LLMs for Clinical Interviews

DGX agent

arXiv:2503.20666v2 Announce Type: replace-cross Abstract: Thematic analysis (TA) is a widely used qualitative approach for uncovering latent meanings in unstructured text data. TA provides valuable in

agentsarxiv-cs-cl
7 Jul 2026
Agents

The 'I Don't Know' Filter: Enhancing Agentic Reliability in Function Calling

DGX agent

arXiv:2607.04034v1 Announce Type: cross Abstract: The language models that underpin agents have seen a rapid rise in performance on function calling benchmarks. However, the metrics used in the traini

agentsarxiv-cs-ai
7 Jul 2026
Agents

Agent Runs now available in the Vercel MCP and CLI

DGX agent

Vercel has announced the availability of Agent Runs in both the Vercel MCP (Model Context Protocol) and CLI (Command Line Interface), expanding developer capabilities for AI-powered automation and dep

agentsvercel-blog
3 Jul 2026
Model Releases

AgenticRAGTracer: A Hop-Aware Benchmark for Diagnosing Multi-Step Retrieval Reasoning in Agentic RAG

DGX agent

arXiv:2602.19127v2 Announce Type: replace Abstract: With the rapid advancement of agent-based methods in recent years, Agentic RAG has undoubtedly become an important research direction. Multi-hop rea

model-releasesarxiv-cs-cl
3 Jul 2026
Agents

CLAP: Closed-Loop Training, Evaluation, and Release Control for Domain Agent Post-training

DGX agent

arXiv:2607.01846v1 Announce Type: new Abstract: Domain agents often face noisy business data, uncertain post-training gains, offline/application mismatch, and adapter-release risk. This paper presents

agentsarxiv-cs-ai
3 Jul 2026
Agents

ContextNest: Verifiable Context Governance for Autonomous AI Agent

DGX agent

arXiv:2607.02116v1 Announce Type: new Abstract: Autonomous AI agents increasingly depend on external knowledge stores, yet most retrieval pipelines provide relevance without durable guarantees of prov

agentsarxiv-cs-ai
3 Jul 2026
Agents

Leveraging Metamemory Agent for Enhanced Data-Free Code Generation in Large Language Models

DGX agent

arXiv:2501.07892v2 Announce Type: replace-cross Abstract: Large language models (LLMs) have shown strong performance in automated code generation, with few-shot prompting widely used for its simplicit

agentsarxiv-cs-ai
3 Jul 2026
Model Releases

Safety Testing LLM Agents at Scale: From Risk Discovery to Evidence-Grounded Verification

DGX agent

arXiv:2607.01793v1 Announce Type: new Abstract: LLM agents increasingly perform autonomous actions through external tools, leading to complex and evolving safety risks. However, existing safety testin

model-releasesarxiv-cs-ai
3 Jul 2026
Agents

A Task-State Representation for Long-Horizon Mobile GUI Agents

DGX agent

arXiv:2607.00502v1 Announce Type: new Abstract: While long-horizon mobile GUI agents typically rely on thought-action-observation loops, they struggle to separate persistent task states from transient

agentsarxiv-cs-cl
2 Jul 2026
Agents

Behavior-Adaptive Conversational Agents: Toward a Fluid Personality Framework

DGX agent

arXiv:2607.01034v1 Announce Type: cross Abstract: Large language model (LLM)-based conversational agents (CAs) are now ubiquitous, creating new opportunities for AI-mediated behavior change. Their cap

agentsarxiv-cs-ai
2 Jul 2026
Agents

Cheap Code, Costly Judgment: A Case Study on Governable Agentic Software Engineering

DGX agent

arXiv:2607.01087v1 Announce Type: cross Abstract: Generative AI is shifting software engineering from a practice organized around scarce implementation effort toward one organized around abundant, low

agentsarxiv-cs-ai
2 Jul 2026
Model Releases

From Signals to Structure: How Memory Architecture Drives Language Emergence in LLM Agents

DGX agent

arXiv:2607.00233v1 Announce Type: new Abstract: How do two agents invent a shared language from scratch? In a Lewis signaling game, a sender and receiver must coordinate on a code using only their int

model-releasesarxiv-cs-ai
2 Jul 2026
Agents

SkillSelect-Serve: Budget-Controllable and QoS-Aware Skill Service Recommendation and Composition for Small LLM Agents

DGX agent

arXiv:2607.00011v1 Announce Type: cross Abstract: Reusable skill libraries are becoming important infrastructure for large language model (LLM) agents, yet existing selection methods often treat skill

agentsarxiv-cs-ai
2 Jul 2026
Agents

AI-Assisted Discovery of Convex Relaxations via Dual Agents

DGX agent

arXiv:2606.31182v1 Announce Type: new Abstract: Recent work shows that LLM agents can improve sharp-constant inequalities by searching for extremal constructions, which yield upper bounds. We address

agentsarxiv-cs-ai
1 Jul 2026
Safety

GUIDE: Resolving Domain Bias in GUI Agents through Real-Time Web Video Retrieval and Plug-and-Play Annotation

DGX agent

arXiv:2603.26266v3 Announce Type: replace Abstract: Large vision-language models have endowed GUI agents with strong general capabilities for interface understanding and interaction. However, due to i

safetyarxiv-cs-ai
1 Jul 2026
Agents

if i were building an agent from scratch with memory i would use a wiki structure, it's simple and extensible

DGX agent

Harrison Chase suggests using a wiki structure as the foundation for building AI agents with memory systems, citing its simplicity and extensibility as key advantages. This approach would organize age

agentsharrison-chase--x
1 Jul 2026
Agents

if you want agents to do work at scale (like security triage, trace analysis, document parsing), they need structured workflows enforced w/ …

DGX agent

if you want agents to do work at scale (like security triage, trace analysis, document parsing), they need structured workflows enforced w/ code map reduce is a great example -- exactly the kind of pa

agentsharrison-chase--x
1 Jul 2026
Model Releases

SimpleSearch-VL: A Simple Recipe for Multimodal Agentic Deep Search

DGX agent

arXiv:2606.31504v1 Announce Type: new Abstract: We present SimpleSearch-VL, an efficient, reliable, and practical framework for multimodal agentic search. Its core idea is to improve the agent's own s

model-releasesarxiv-cs-cv
1 Jul 2026
Agents

We gave a 2 hr deepdive on how to build inference engines that handle trillion token agentic workloads at @aiDotEngineer. Will drop slides a…

DGX agent

Together AI presented a 2-hour technical deep dive on building inference engines capable of handling trillion-token agentic workloads, covering architecture and optimization strategies for large-scale

agentstogether-ai--x
1 Jul 2026
Agents

Couchbase’s AI Data Plane aims to turn fragmented data into real enterprise agent memory

DGX agent

Couchbase Inc. is trying to solve one of the hardest problems in enterprise artificial intelligence today: turning brittle, chat-style pilots into production-grade agents capable of remembering, reaso

agentssiliconangle
30 Jun 2026
Agents

ManimAgent: Self-Evolving Multimodal Agents for Visual Education

DGX agent

arXiv:2606.30296v1 Announce Type: new Abstract: Multi-round reflection lets agents built on large language models recover from failures within a single task, but each task remains an isolated episode:

agentsarxiv-cs-ai
30 Jun 2026
Safety

Proof-of-Guardrail in AI Agents and What (Not) to Trust from It

DGX agent

arXiv:2603.05786v2 Announce Type: replace-cross Abstract: As AI agents become widely deployed as online services, users often rely on an agent developer's claim about how safety is enforced, which int

safetyarxiv-cs-ai
30 Jun 2026
Agents

Self-Evolving Agentic Image Restoration via Deliberate Planning and Intuitive Execution

DGX agent

arXiv:2606.28971v1 Announce Type: new Abstract: Real-world image restoration (IR) remains challenging due to complex and coupled degradations. While recent agentic IR frameworks leverage Large Languag

agentsarxiv-cs-cv
30 Jun 2026
Agents

Semantic search alone doesn't cut it. Neither does brute-force grep. Agents need both. Today we're shipping the Retrieval Harness in LlamaPa…

DGX agent

Semantic search alone doesn't cut it. Neither does brute-force grep. Agents need both. Today we're shipping the Retrieval Harness in LlamaParse Index: semantic search, server-side grep, and file-level

agentsllamaindex--x
29 Jun 2026
Model Releases

Running the Gauntlet: Re-evaluating the Capabilities of Agents Beyond Familiar Environments

DGX agent

arXiv:2606.14397v2 Announce Type: replace Abstract: As agentic systems continue to evolve and are widely deployed in real-world scenarios, there is a growing demand to faithfully evaluate their capabi

model-releasesarxiv-cs-lg
26 Jun 2026
Safety

When Agents Meet Electric Bus Fleet Operations: Pricing Behavior, Trade-offs, and Policy Implications in an Aggregator Framework

DGX agent

arXiv:2606.26400v1 Announce Type: new Abstract: Agentic systems are changing how complex operational tasks are coordinated, introducing a new paradigm for connecting heterogeneous data sources and aut

safetyarxiv-cs-ai
26 Jun 2026
Applications

deployment cookbook for langchain agents!

DGX agent

deployment cookbook for langchain agents! Agents are easy to demo locally. The hard part is shipping them inside a real app. We published a deployment cookbook for @LangChain agents: full-stack exampl

applicationsharrison-chase--x
25 Jun 2026
Agents

Membox: Weaving Topic Continuity into Long-Range Memory for LLM Agents

DGX agent

arXiv:2601.03785v3 Announce Type: replace Abstract: Long-term human-agent dialogues are organized by topic continuity: adjacent turns often develop the same goal, plan, problem, or event, while relate

agentsarxiv-cs-cl
25 Jun 2026
Model Releases

Multi-Agent Goal Recognition with Team- and Goal-Conditioned Reinforcement Learning and Factorized Branch-and-Bound

DGX agent

arXiv:2606.25978v1 Announce Type: cross Abstract: Multi-agent goal recognition asks an observer to jointly infer which agents act together and what each team is trying to achieve, so the hypothesis sp

model-releasesarxiv-cs-lg
25 Jun 2026
Agents

Plausible but Wrong: A case study on Agentic Failures in Astrophysical Workflows

DGX agent

arXiv:2604.25345v2 Announce Type: replace Abstract: Agentic AI systems are increasingly being integrated into scientific workflows, yet their behavior under realistic conditions remains insufficiently

agentsarxiv-cs-ai
25 Jun 2026
Safety

Escaping the Self-Confirmation Trap: An Execute-Distill-Verify Paradigm for Agentic Experience Learning

DGX agent

arXiv:2606.24428v1 Announce Type: new Abstract: Experience-driven self-evolution is critical for large language model (LLM) agents to improve through open-world interaction. However, existing experien

safetyarxiv-cs-cl
24 Jun 2026
Agents

Grading the Grader: Lessons from Evaluating an Agentic Data Analysis System

DGX agent

arXiv:2606.24839v1 Announce Type: new Abstract: Agentic data analysis systems produce rich outputs, including code, numerical results, and verbal diagnostics. This makes them more challenging to evalu

agentsarxiv-cs-ai
24 Jun 2026
Model Releases

GUI vs. CLI: Execution Bottlenecks in Screen-Only and Skill-Mediated Computer-Use Agents

DGX agent

arXiv:2606.24551v1 Announce Type: new Abstract: Computer-use agents can execute software tasks through either graphical interfaces or programmatic command interfaces, but existing evaluations confound

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

MEMPROBE: Probing Long-Term Agent Memory via Hidden User-State Recovery

DGX agent

arXiv:2606.24595v1 Announce Type: new Abstract: Long-term memory promises LLM agents that grow more capable across sessions, maintaining an accurate, evolving understanding of the user that interactio

model-releasesarxiv-cs-cl
24 Jun 2026
Agents

'Most agents don't learn, they just leave traces.' In 12 minutes, @jakebroekhuizen breaks down how to actually close the loop. Surface issue…

DGX agent

'Most agents don't learn, they just leave traces.' In 12 minutes, @jakebroekhuizen breaks down how to actually close the loop. Surface issues with LangSmith Engine Write memory updates back to Context

agentsharrison-chase--x
24 Jun 2026
Safety

Policy Gradient with Self-Attention for Model-Free Distributed Nonlinear Multi-Agent Games

DGX agent

arXiv:2509.18371v2 Announce Type: replace-cross Abstract: Multi-agent games in dynamic nonlinear settings are challenging due to the time-varying interactions among the agents and the non-stationarity

safetyarxiv-cs-ro
24 Jun 2026
← Previous
1…7172737475…374
Next →