AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,770 results
Model Releases

AgentLens: Production-Assessed Trajectory Reviews for Coding Agent Evaluation

DGX agent

arXiv:2607.06624v1 Announce Type: new Abstract: We present AgentLens, a production-assessed benchmark for interactive code agents. Most code-agent benchmarks reduce a run to a single bit -- did the ta

model-releasesarxiv-cs-ai
9 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Agents

AMD targets system-level AI infrastructure optimization as agentic workloads reshape enterprise compute

DGX agent

Infrastructure design is being redefined by agentic AI, pushing the industry toward system-level AI infrastructure optimization, balancing performance and cost across diverse workloads rather than foc

agentssiliconangle
9 Jul 2026
Safety

Multi-Agent AI Control: Distributed Attacks Hamper Per-Instance Monitors

DGX agent

arXiv:2607.07368v1 Announce Type: cross Abstract: AI control is a family of techniques to prevent an AI with malicious goals from subverting its operator's intent. AI Control usually studies a single

safetyarxiv-cs-ai
9 Jul 2026
Agents

STAGformer: A Spatio-temporal Agent Graph Transformer for Micro Mobility Demand Forecasting

DGX agent

arXiv:2607.06614v1 Announce Type: cross Abstract: Accurate station-level demand forecasting is essential for the efficient operation of bike-sharing systems, yet it remains challenging due to complex

agentsarxiv-cs-ai
9 Jul 2026
Model Releases

An Experimental Design Approach to Evaluating Agentic AI's Autonomous Model Discovery

DGX agent

arXiv:2607.06413v1 Announce Type: cross Abstract: Large language model coding agents increasingly perform open-ended data modeling and analysis. These agents are stochastic and adaptive, and therefore

model-releasesarxiv-cs-ai
8 Jul 2026
Model Releases

dcode <> nemotron 3 ultra <> OpenShell 🔥 Never been easier to own your agent IP!

DGX agent

dcode <> nemotron 3 ultra <> OpenShell 🔥 Never been easier to own your agent IP! Introducing the NemoClaw Deep Agents Blueprint, a reference architecture for building open agent systems developed with

model-releasesharrison-chase--x
8 Jul 2026
Agents

Decision Protocols in Multi-Agent Large Language Model Conversations

DGX agent

arXiv:2607.05477v1 Announce Type: cross Abstract: Improving the task performance of Large Language Models (LLMs) is essential, yet scaling these models faces significant challenges such as diminishing

agentsarxiv-cs-ai
8 Jul 2026
Agents

DeepFabric ships more than 50 AI agents for supply chain operations

DGX agent

Supply chain artificial intelligence startup DeepFabric today announced the general availability of an AI agent platform built for supply chain execution, and a roster of enterprise customers is alrea

agentssiliconangle
8 Jul 2026
Agents

StateFuse: Deterministic Conflict-Preserving Memory for Multi-Agent Systems

DGX agent

arXiv:2607.05844v1 Announce Type: new Abstract: Agent systems accumulate conflicting observations across branches, retries, and replicas, yet many practical memory layers still collapse disagreement b

agentsarxiv-cs-ai
8 Jul 2026
Agents

TypeGo: An OS Runtime for Embodied Agents

DGX agent

arXiv:2607.05482v1 Announce Type: cross Abstract: Large language models (LLMs) can plan behavior for embodied agents from natural language, but treating the LLM as a request/response oracle on the cri

agentsarxiv-cs-ro
8 Jul 2026
Agents

A Cost-Aware, Paired Protocol for Auditing Dynamic Tool Synthesis in Agentic Video Question Answering

DGX agent

arXiv:2607.01469v2 Announce Type: replace Abstract: Agentic Video Question Answering (VideoQA) systems invoke tools during inference, but their tool libraries are fixed, so recurring procedures are re

agentsarxiv-cs-cv
7 Jul 2026
Safety

ARLArena: A Unified Framework for Stable Agentic Reinforcement Learning

DGX agent

arXiv:2602.21534v3 Announce Type: replace Abstract: Agentic reinforcement learning (ARL) has rapidly gained attention as a promising paradigm for training agents to solve complex, multi-step interacti

safetyarxiv-cs-ai
7 Jul 2026
Agents

LLMoxie: Exploring Agentic AI for Scientific Software Development

DGX agent

arXiv:2607.02703v1 Announce Type: cross Abstract: In this paper, we describe LLMoxie, an institutional AI platform whose three-tiered architecture supports multi-cloud and on-premise inference, a Lite

agentsarxiv-cs-ai
7 Jul 2026
Safety

MAD-PINN: A Decentralized Physics-Informed Machine Learning Framework for Safe and Optimal Multi-Agent Control

DGX agent

arXiv:2509.23960v2 Announce Type: replace-cross Abstract: Co-optimizing safety and performance in large-scale multi-agent systems remains a fundamental challenge. Existing approaches based on multi-ag

safetyarxiv-cs-ai
7 Jul 2026
Agents

my hot take is that evals are the only part of agent engineering that require real thinking

DGX agent

Harrison Chase argues that evaluations (evals) are the most intellectually demanding aspect of agent engineering, distinguishing them from other engineering tasks that may be more routine or formulaic

agentsharrison-chase--x
7 Jul 2026
Agents

OptiAgent: End-to-End Optimization Modeling via Multi-Agent Iterative Refinement

DGX agent

arXiv:2607.05346v1 Announce Type: new Abstract: We propose OptiAgent, a multi-agent framework that, given a natural language description of an Operations Research problem, is able to output a solver-r

agentsarxiv-cs-ai
7 Jul 2026
Model Releases

Radware adds Claude Code protection and compliance reporting to agent security

DGX agent

Application and network security company Radware Ltd. today expanded its Agentic AI Protection product with compliance reporting, deeper visibility into artificial intelligence agent activity, and new

model-releasessiliconangle
7 Jul 2026
Safety

Securing Multi-Tool AI Agent Chains With Dynamic, Real-Time Compositional Policies

DGX agent

arXiv:2607.03423v1 Announce Type: cross Abstract: Modern AI agent implementations such as frontier coding agents chain multiple tools at runtime that create a security surface that per-tool guardrails

safetyarxiv-cs-ai
7 Jul 2026
Model Releases

SpecEyes: Accelerating Agentic Multimodal LLMs via Speculative Perception and Planning

DGX agent

arXiv:2603.23483v2 Announce Type: replace-cross Abstract: Agentic multimodal large language models (MLLMs) (e.g., OpenAI o3 and Gemini Agentic Vision) achieve remarkable reasoning capabilities through

model-releasesarxiv-cs-cl
7 Jul 2026
Agents

Spinning Straw into Gold: Relabeling LLM Agent Trajectories in Hindsight for Successful Demonstrations

DGX agent

arXiv:2607.04235v1 Announce Type: new Abstract: Large language model agents operate in partially observable, long-horizon settings where obtaining supervision remains a major bottleneck. We address th

agentsarxiv-cs-cl
7 Jul 2026
Agents

Storage gets promoted in the agentic AI era

DGX agent

The year 2026 could be remembered as the moment when storage technology received a massive promotion. The reason is that the current transition from simple chatbots to agentic AI systems has raised th

agentssiliconangle
7 Jul 2026
Safety

Strategic Buying Agents

DGX agent

arXiv:2607.04708v1 Announce Type: cross Abstract: Agentic AI is shifting online shopping from search toward delegated purchasing, where autonomous buying agents monitor markets and decide when to buy

safetyarxiv-cs-ai
7 Jul 2026
Agents

TAMA: A Human-AI Collaborative Thematic Analysis Framework Using Multi-Agent LLMs for Clinical Interviews

DGX agent

arXiv:2503.20666v2 Announce Type: replace-cross Abstract: Thematic analysis (TA) is a widely used qualitative approach for uncovering latent meanings in unstructured text data. TA provides valuable in

agentsarxiv-cs-cl
7 Jul 2026
Agents

The 'I Don't Know' Filter: Enhancing Agentic Reliability in Function Calling

DGX agent

arXiv:2607.04034v1 Announce Type: cross Abstract: The language models that underpin agents have seen a rapid rise in performance on function calling benchmarks. However, the metrics used in the traini

agentsarxiv-cs-ai
7 Jul 2026
Agents

Agent Runs now available in the Vercel MCP and CLI

DGX agent

Vercel has announced the availability of Agent Runs in both the Vercel MCP (Model Context Protocol) and CLI (Command Line Interface), expanding developer capabilities for AI-powered automation and dep

agentsvercel-blog
3 Jul 2026
Model Releases

AgenticRAGTracer: A Hop-Aware Benchmark for Diagnosing Multi-Step Retrieval Reasoning in Agentic RAG

DGX agent

arXiv:2602.19127v2 Announce Type: replace Abstract: With the rapid advancement of agent-based methods in recent years, Agentic RAG has undoubtedly become an important research direction. Multi-hop rea

model-releasesarxiv-cs-cl
3 Jul 2026
Agents

CLAP: Closed-Loop Training, Evaluation, and Release Control for Domain Agent Post-training

DGX agent

arXiv:2607.01846v1 Announce Type: new Abstract: Domain agents often face noisy business data, uncertain post-training gains, offline/application mismatch, and adapter-release risk. This paper presents

agentsarxiv-cs-ai
3 Jul 2026
Agents

ContextNest: Verifiable Context Governance for Autonomous AI Agent

DGX agent

arXiv:2607.02116v1 Announce Type: new Abstract: Autonomous AI agents increasingly depend on external knowledge stores, yet most retrieval pipelines provide relevance without durable guarantees of prov

agentsarxiv-cs-ai
3 Jul 2026
Agents

Leveraging Metamemory Agent for Enhanced Data-Free Code Generation in Large Language Models

DGX agent

arXiv:2501.07892v2 Announce Type: replace-cross Abstract: Large language models (LLMs) have shown strong performance in automated code generation, with few-shot prompting widely used for its simplicit

agentsarxiv-cs-ai
3 Jul 2026
Model Releases

Safety Testing LLM Agents at Scale: From Risk Discovery to Evidence-Grounded Verification

DGX agent

arXiv:2607.01793v1 Announce Type: new Abstract: LLM agents increasingly perform autonomous actions through external tools, leading to complex and evolving safety risks. However, existing safety testin

model-releasesarxiv-cs-ai
3 Jul 2026
Agents

A Task-State Representation for Long-Horizon Mobile GUI Agents

DGX agent

arXiv:2607.00502v1 Announce Type: new Abstract: While long-horizon mobile GUI agents typically rely on thought-action-observation loops, they struggle to separate persistent task states from transient

agentsarxiv-cs-cl
2 Jul 2026
Agents

Behavior-Adaptive Conversational Agents: Toward a Fluid Personality Framework

DGX agent

arXiv:2607.01034v1 Announce Type: cross Abstract: Large language model (LLM)-based conversational agents (CAs) are now ubiquitous, creating new opportunities for AI-mediated behavior change. Their cap

agentsarxiv-cs-ai
2 Jul 2026
Agents

Cheap Code, Costly Judgment: A Case Study on Governable Agentic Software Engineering

DGX agent

arXiv:2607.01087v1 Announce Type: cross Abstract: Generative AI is shifting software engineering from a practice organized around scarce implementation effort toward one organized around abundant, low

agentsarxiv-cs-ai
2 Jul 2026
Model Releases

From Signals to Structure: How Memory Architecture Drives Language Emergence in LLM Agents

DGX agent

arXiv:2607.00233v1 Announce Type: new Abstract: How do two agents invent a shared language from scratch? In a Lewis signaling game, a sender and receiver must coordinate on a code using only their int

model-releasesarxiv-cs-ai
2 Jul 2026
Agents

SkillSelect-Serve: Budget-Controllable and QoS-Aware Skill Service Recommendation and Composition for Small LLM Agents

DGX agent

arXiv:2607.00011v1 Announce Type: cross Abstract: Reusable skill libraries are becoming important infrastructure for large language model (LLM) agents, yet existing selection methods often treat skill

agentsarxiv-cs-ai
2 Jul 2026
Agents

AI-Assisted Discovery of Convex Relaxations via Dual Agents

DGX agent

arXiv:2606.31182v1 Announce Type: new Abstract: Recent work shows that LLM agents can improve sharp-constant inequalities by searching for extremal constructions, which yield upper bounds. We address

agentsarxiv-cs-ai
1 Jul 2026
Safety

GUIDE: Resolving Domain Bias in GUI Agents through Real-Time Web Video Retrieval and Plug-and-Play Annotation

DGX agent

arXiv:2603.26266v3 Announce Type: replace Abstract: Large vision-language models have endowed GUI agents with strong general capabilities for interface understanding and interaction. However, due to i

safetyarxiv-cs-ai
1 Jul 2026
Agents

if i were building an agent from scratch with memory i would use a wiki structure, it's simple and extensible

DGX agent

Harrison Chase suggests using a wiki structure as the foundation for building AI agents with memory systems, citing its simplicity and extensibility as key advantages. This approach would organize age

agentsharrison-chase--x
1 Jul 2026
Agents

if you want agents to do work at scale (like security triage, trace analysis, document parsing), they need structured workflows enforced w/ …

DGX agent

if you want agents to do work at scale (like security triage, trace analysis, document parsing), they need structured workflows enforced w/ code map reduce is a great example -- exactly the kind of pa

agentsharrison-chase--x
1 Jul 2026
Model Releases

SimpleSearch-VL: A Simple Recipe for Multimodal Agentic Deep Search

DGX agent

arXiv:2606.31504v1 Announce Type: new Abstract: We present SimpleSearch-VL, an efficient, reliable, and practical framework for multimodal agentic search. Its core idea is to improve the agent's own s

model-releasesarxiv-cs-cv
1 Jul 2026
Agents

We gave a 2 hr deepdive on how to build inference engines that handle trillion token agentic workloads at @aiDotEngineer. Will drop slides a…

DGX agent

Together AI presented a 2-hour technical deep dive on building inference engines capable of handling trillion-token agentic workloads, covering architecture and optimization strategies for large-scale

agentstogether-ai--x
1 Jul 2026
Agents

Couchbase’s AI Data Plane aims to turn fragmented data into real enterprise agent memory

DGX agent

Couchbase Inc. is trying to solve one of the hardest problems in enterprise artificial intelligence today: turning brittle, chat-style pilots into production-grade agents capable of remembering, reaso

agentssiliconangle
30 Jun 2026
Agents

ManimAgent: Self-Evolving Multimodal Agents for Visual Education

DGX agent

arXiv:2606.30296v1 Announce Type: new Abstract: Multi-round reflection lets agents built on large language models recover from failures within a single task, but each task remains an isolated episode:

agentsarxiv-cs-ai
30 Jun 2026
Safety

Proof-of-Guardrail in AI Agents and What (Not) to Trust from It

DGX agent

arXiv:2603.05786v2 Announce Type: replace-cross Abstract: As AI agents become widely deployed as online services, users often rely on an agent developer's claim about how safety is enforced, which int

safetyarxiv-cs-ai
30 Jun 2026
Agents

Self-Evolving Agentic Image Restoration via Deliberate Planning and Intuitive Execution

DGX agent

arXiv:2606.28971v1 Announce Type: new Abstract: Real-world image restoration (IR) remains challenging due to complex and coupled degradations. While recent agentic IR frameworks leverage Large Languag

agentsarxiv-cs-cv
30 Jun 2026
Agents

Semantic search alone doesn't cut it. Neither does brute-force grep. Agents need both. Today we're shipping the Retrieval Harness in LlamaPa…

DGX agent

Semantic search alone doesn't cut it. Neither does brute-force grep. Agents need both. Today we're shipping the Retrieval Harness in LlamaParse Index: semantic search, server-side grep, and file-level

agentsllamaindex--x
29 Jun 2026
Model Releases

Running the Gauntlet: Re-evaluating the Capabilities of Agents Beyond Familiar Environments

DGX agent

arXiv:2606.14397v2 Announce Type: replace Abstract: As agentic systems continue to evolve and are widely deployed in real-world scenarios, there is a growing demand to faithfully evaluate their capabi

model-releasesarxiv-cs-lg
26 Jun 2026
Safety

When Agents Meet Electric Bus Fleet Operations: Pricing Behavior, Trade-offs, and Policy Implications in an Aggregator Framework

DGX agent

arXiv:2606.26400v1 Announce Type: new Abstract: Agentic systems are changing how complex operational tasks are coordinated, introducing a new paradigm for connecting heterogeneous data sources and aut

safetyarxiv-cs-ai
26 Jun 2026
← Previous
1…7071727374…371
Next →