AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

agents

GridTimelineEvolution
7,156 results
29 Jul 2026

Sakana is at the frontier of having fun and I respect that

AgentsDGX agent

Sakana is at the frontier of having fun and I respect that We are excited to share our latest work, together with @nyuniversity: 'Dream-Cubed: Controllable Generative Modeling in Minecraft by Training

SnapLogic transforms SnapGPT into a high-powered agentic assistant for the entire integration lifecycle

AgentsDGX agent

Enterprise data integration and automation firm SnapLogic Inc. today announced a significant update to SnapGPT, the company’s artificial intelligence copilot for enterprise data automation, transformi

SourceMinds at CheckThat! 2026: NLI-Grounded Citation Auditing in a Multi-Agent Pipeline for Full Fact-Checking Article Generation

AgentsDGX agent

arXiv:2607.24802v1 Announce Type: cross Abstract: This paper presents our system for Task 3 of the CLEF 2026 CheckThat! Lab, which focuses on generating full fact-checking articles from claims, veraci


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Specula: Scaling formal specifications for autonomous model checking of system code

AgentsDGX agent

arXiv:2607.25333v1 Announce Type: cross Abstract: Specula is a push-button agentic system that generates high-quality formal specifications for large, complex system code and uses the specifications f

Speculate While You Reason: Teaching Agents to Predict Their Next Tool Call via Joint Agent-Speculator RL

AgentsDGX agent

arXiv:2607.25816v1 Announce Type: new Abstract: Large language model agents often spend substantial wall-clock time waiting for tool call results. Tool-call speculation can hide this latency by predic

Steeringless Drifting: Differential-Torque Control of a Four-Wheel Independently Driven Vehicle

AgentsDGX agent

arXiv:2607.24863v1 Announce Type: new Abstract: Control methods for emerging vehicle chassis architectures are important for autonomous driving near handling limits. Unlike conventional drift control,

The Future of Agentic AI Depends on Cloud Cost Optimization

AgentsDGX agent

Capitalizing on agentic AI depends on the right infrastructure investments, yet enterprises struggle to reduce cloud costs. Vultr VX1™ Cloud Compute presents a way to free up budget for CPUs and GPUs

The LAIA Dataset: Labelled Attention for Intelligent Automobiles

AgentsDGX agent

arXiv:2607.25570v1 Announce Type: cross Abstract: The development of autonomous vehicles (AVs) usually relies heavily on data-driven artificial intelligence (AI) models that require large volumes of s

The openai agent sandbox escape actually has very real implications for the diffusion of AI in the enterprise. The incident showed the power…

AgentsDGX agent

The openai agent sandbox escape actually has very real implications for the diffusion of AI in the enterprise. The incident showed the power and capability of agents, and the need to harden systems an

The root cause: request-level engines never see that a series of LLM calls belongs to one longer workflow. ThunderAgent adds that missing vi…

AgentsDGX agent

The root cause: request-level engines never see that a series of LLM calls belongs to one longer workflow. ThunderAgent adds that missing view. It treats each agent workflow as a schedulable program,

The User Asks, Platforms Compete: How Agentic Recommendation Markets Take Shape

AgentsDGX agent

arXiv:2607.25253v1 Announce Type: new Abstract: Online recommendation has traditionally taken place after a user enters a platform, which determines the candidate pool and the ranking shown to the use

ThunderAgent: 2x Faster Agentic Inference for Synthetic Data Generation at Scale

AgentsDGX agent

ThunderAgent is a program-aware scheduler for agentic inference. By treating each agent workflow as a schedulable program, it eliminates KV cache thrashing to deliver more than 2x single-node throughp

Today we’re open-sourcing Numbat, an agent-detection and response layer that is designed to work across agent harnesses. Numbat gives securi…

AgentsDGX agent

Today we’re open-sourcing Numbat, an agent-detection and response layer that is designed to work across agent harnesses. Numbat gives security teams visibility into agent activity, with controls to bl

Tools Are Not Islands: Set-Level Tool Retrieval for LLM Agents via Query-Conditioned Hyperedge Prediction

AgentsDGX agent

arXiv:2607.25718v1 Announce Type: cross Abstract: Large language model (LLM) agents increasingly rely on invoking external tools to complete real-world tasks. Tool retrieval, which selects a small tas

Toward Standardized Cross-Vendor Agent Tool Trust Management in Autonomous Networks

AgentsDGX agent

arXiv:2607.25914v1 Announce Type: new Abstract: Autonomous Network Levels 4-5 require AI agents to invoke tools across vendor boundaries without human oversight, yet existing management standards lack

Towards an Agent Operating System - Lessons from Classical and Cloud OS

AgentsDGX agent

arXiv:2607.25076v1 Announce Type: new Abstract: Every major wave of platform software follows the same arc: an initial period of experimentation with competing frameworks and ad-hoc implementations, f

Understanding User Experiences of Computer Use Agents: Design Space and Opportunities for Building Agent UX Prototypes

AgentsDGX agent

arXiv:2510.04452v3 Announce Type: replace-cross Abstract: Computer use agents (or 'agents') are generative AI that automates actions within user interfaces from user commands. Current research focuses

// Unfolding Sub-Agents for Long-Horizon ML Engineering // Watch a single agent work a machine learning engineering task for six hours you s…

AgentsDGX agent

// Unfolding Sub-Agents for Long-Horizon ML Engineering // Watch a single agent work a machine learning engineering task for six hours you see issues like context fills with stack traces, dead experim

.@UseApolloio will be at Interrupt NYC. Apollo's AI Assistant is one of the earliest production multi-agent systems on LangGraph. Their team…

AgentsDGX agent

.@UseApolloio will be at Interrupt NYC. Apollo's AI Assistant is one of the earliest production multi-agent systems on LangGraph. Their team will share how they migrated it from a hand-rolled supervis

VLD-RAG: Agentic Vision-Language Retrieval-Augmented Generation for Long, Visually-Rich Multi-Page Documents

AgentsDGX agent

arXiv:2607.24748v1 Announce Type: cross Abstract: Visually-rich documents such as reports, slides, and manuals often distribute the evidence needed to answer a question across multiple pages, mixing t

What Gets Lost When Memory Becomes Media? Evaluating AI-Generated Oral History Visualization

AgentsDGX agent

arXiv:2607.24756v1 Announce Type: cross Abstract: What gets lost when memory becomes media? Diaspora oral-history interviews require a double transformation; first-person recollection to third-person

28 Jul 2026

A corrective agentic hybrid RAG and an operations-grounded evaluation for a scientific facility

AgentsDGX agent

arXiv:2607.24663v1 Announce Type: cross Abstract: Scientific user facilities accumulate decades of operational knowledge that no single search index covers: electronic logbooks, technical documents, i

A funny companion: Distinct neural responses to AI- versus human-attributed humor

AgentsDGX agent

arXiv:2509.10847v3 Announce Type: replace-cross Abstract: As artificial intelligence (AI) companions become capable of human-like communication, including telling jokes, understanding how people cogni

A New Role for Relevance: Guiding Corpus Interaction in Agentic Search

AgentsDGX agent

arXiv:2607.24223v1 Announce Type: new Abstract: Relevance is a query-dependent estimate of whether a document or excerpt contains useful evidence. Existing retrieval agents use relevance to select top

A Vocabulary for Multi-Agent Automated Research Systems

AgentsDGX agent

arXiv:2607.22682v1 Announce Type: new Abstract: We introduce a vocabulary for automated research systems built from one or more agents to make their design choices easier to describe and compare. The

Accuracy potential of visual localization exploiting high-end street-level imagery

AgentsDGX agent

arXiv:2607.24409v1 Announce Type: new Abstract: Accurate and reliable pose information with respect to a reference frame is increasingly demanded across applications such as autonomous navigation, sur

ACM: Agentic Context Management for Long Horizon Tasks

AgentsDGX agent

arXiv:2607.23809v1 Announce Type: new Abstract: Agentic tasks are inherently long-horizon and multi-turn, constantly accumulating context through interactions with the environment. Existing context co

Across eight case studies spanning industry and academia, we explore what this shift means for scientific computing—and why human verificati…

AgentsDGX agent

Across eight case studies spanning industry and academia, we explore what this shift means for scientific computing—and why human verification, stewardship, and long-term maintenance matter. https://o

Act Security raises $60M to take action against agentic access sprawl at the infrastructure layer

AgentsDGX agent

Cloud security startup Act Security Ltd. says it’s primed and ready to help enterprises safely deploy artificial intelligence agents at large scale after raising 60 million in funding across two round

Agent-UCT: Upper Confidence Bounds Applied to Trees for Agentic Workflow Optimization with Cost-Awareness

AgentsDGX agent

arXiv:2607.24162v1 Announce Type: new Abstract: Optimizing agentic workflows, such as retrieval-augmented generation (RAG) pipelines, requires navigating a combinatorial space of discrete component ch

Agentic Autoresearch for CT Reconstruction

AgentsDGX agent

arXiv:2607.22824v1 Announce Type: cross Abstract: Comparing CT reconstruction methods fairly is labor-intensive and largely manual, and many benchmarks use idealized data. We ask whether a large langu

Agentic Cloud Decoys: A Deception-Driven Framework for Autonomous Intrusion Investigation

AgentsDGX agent

arXiv:2607.24006v1 Announce Type: cross Abstract: Cloud telemetry arrives at a scale that, paradoxically, makes intrusion understanding harder rather than easier. Attackers operate through legitimate

Agentic Reward Modeling: Verifying GUI Agent via Progressive Trajectory-Grounded Interaction

AgentsDGX agent

arXiv:2602.00575v2 Announce Type: replace Abstract: Reinforcement learning with verifiable rewards (RLVR) provides a promising pathway for continuously advancing GUI agents, yet existing reward modeli

AI agent evaluation: Tips from Anthropic on building evals you can trust

AgentsDGX agent

Learn how to build trustworthy AI agent evals using regression tests, capability evals, production traces, LLM judges, and reproducible environments. The post AI agent evaluation: Tips from Anthropic

An Agentic Orchestration of Atomistic Simulations

AgentsDGX agent

arXiv:2607.22596v1 Announce Type: new Abstract: Atomistic simulations are central to materials design, but their execution involves complex, multi-step workflows that require significant human experti

Are You Still the Agent I Authorized? Earned Authority under a Fixed Ceiling for Evolving Agents

AgentsDGX agent

arXiv:2607.23586v1 Announce Type: new Abstract: Long-lived AI agents increasingly evolve after deployment by retaining experience, acquiring skills and tools, revising workflows, delegating work, and

AutoWorld: Learning Multi-Agent Traffic Simulation with Self-Supervised World Models

AgentsDGX agent

arXiv:2603.28963v2 Announce Type: replace-cross Abstract: Simulation with realistic traffic agents is essential for validating autonomous driving systems. Existing data-driven simulators learn agent b

Beyond Success Rate: Cost-Aware Evaluation of Offensive and Defensive Security Agents

AgentsDGX agent

arXiv:2607.15263v3 Announce Type: replace-cross Abstract: Security-agent evaluations commonly measure peak offensive capability under generous inference budgets, emphasizing vulnerability discovery, e

Building AI That Works: ESnet's Pragmatic Approach to AI-Driven Operational Excellence

AgentsDGX agent

arXiv:2607.22948v1 Announce Type: cross Abstract: The ORBIT (Operations Responses and Business Intelligence Toolkit) project was initiated to assess agentic AI for the upcoming ESnet 7 initiative and

CodeEvo: Interaction-Driven Synthesis of Code-centric Data through Hybrid and Iterative Feedback

AgentsDGX agent

arXiv:2507.22080v2 Announce Type: replace-cross Abstract: Acquiring high-quality instruction-code pairs is essential for training Large Language Models for code generation. While automated synthesis h

CodexGraph: Bridging Large Language Models and Code Repositories via Code Graph Databases

AgentsDGX agent

arXiv:2408.03910v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) excel in stand-alone code tasks like HumanEval and MBPP, but struggle with handling entire code repositories. Thi

Coherent Without Grounding, Grounded Without Success: Observability and Epistemic Failure

AgentsDGX agent

arXiv:2603.28371v2 Announce Type: replace-cross Abstract: When an agent can articulate why something works, we typically take this as evidence of genuine understanding. This presupposes that effective

Commitment To Cooperation With Self-Negotiated Contracts

AgentsDGX agent

arXiv:2607.22750v1 Announce Type: new Abstract: As AI agents operate with increasing autonomy in a multi-agent world, they will need to learn to cooperate with other agents and with humans to generate

Compute Globally, Materialize Locally: The Memory Contract of Sparse Event-KV

AgentsDGX agent

arXiv:2607.23693v1 Announce Type: new Abstract: Long-horizon agents increasingly reuse their KV cache as memory: a serving system keeps a subset of cached entries and drops the rest. Eviction and epis

CORVUS: Context Optimization and Reduction Via Underlying Synchronization for LLM Coding Agents

AgentsDGX agent

arXiv:2607.22711v1 Announce Type: cross Abstract: LLM coding agents operate by constructing trajectories that accumulate reasoning, tool calls, and results to enable multi-step decision-making. Howeve

DeepFaith: Evidence-Grounded LLMs for Faithful Incident Reporting in Multi-Stage APT Defense

AgentsDGX agent

arXiv:2607.24348v1 Announce Type: cross Abstract: Advanced Persistent Threats (APTs) are difficult to detect and interpret due to their multi-stage and stealthy nature. While recent autonomous defense

Delegation Intelligence in Deep Search: A Controllable Framework for Disentangled Capability Diagnosis

AgentsDGX agent

arXiv:2607.23524v1 Announce Type: new Abstract: Deep search is becoming a core capability of modern agent systems, yet it is typically evaluated solely based on end-to-end answer accuracy. This couple

Denial of Deadline: Network-Driven Accuracy Collapse in Distributed Inference Pipelines

AgentsDGX agent

arXiv:2607.24692v1 Announce Type: cross Abstract: Inference systems increasingly combine a fast path that returns predictions within the application's latency deadline together with a higher-accuracy

DeReCo: Decoupling Representation and Coordination Learning for Object-Adaptive Decentralized Multi-Robot Cooperative Transport

AgentsDGX agent

arXiv:2603.08111v2 Announce Type: replace Abstract: Generalizing decentralized multi-robot cooperative transport across objects with diverse shapes and physical properties remains a fundamental challe

DispatchRAG: Grounding Emergency Dispatch Decisions in Real-World Protocols from Traffic Accident Video

AgentsDGX agent

arXiv:2607.23132v1 Announce Type: new Abstract: Assessing the severity of a traffic accident scenario is important to decide which emergency service to dispatch. Missing an ambulance dispatch on a ped

Distributed Coordination for Resilient Multi-UAV Remote Sensing: A Photovoltaic Inspection Case Study

AgentsDGX agent

arXiv:2607.24482v1 Announce Type: new Abstract: Deploying multiple UAVs for remote sensing enables proportional reductions in mission time, but realizing these benefits requires the fleet to coordinat

EviBack: Search-Agent Reinforcement Learning via Evidence-Constrained Teacher Backoff

AgentsDGX agent

arXiv:2607.23955v1 Announce Type: new Abstract: Reinforcement learning enables Agentic RAG systems to learn multi-turn search from verifiable outcome rewards, but all- zero rollout groups provide no c

Failures Reveal What Metrics Miss: An Evidence-Driven Agent for Recursive Refinement of ECG Classifiers

AgentsDGX agent

arXiv:2607.24419v1 Announce Type: new Abstract: Deep models have substantially advanced 12-lead ECG classification, yet their refinement still relies heavily on human experts to inspect failures and i

Fair Division with Strictly Increasing Valuations: A Tight Threshold for Two-Agent EF1 and PO

AgentsDGX agent

arXiv:2607.23367v1 Announce Type: cross Abstract: We study whether strictly positive marginal values restore the compatibility of envy-freeness up to one good (EF1) and Pareto optimality (PO) for indi

Focus Is All You Need: Adaptive Goal-aware Attention Orchestration for Multi-Agent Graph Systems

AgentsDGX agent

arXiv:2607.23678v1 Announce Type: new Abstract: Large language models (LLMs) enable autonomous agents for reasoning, planning, and tool use. Recent systems increasingly organize these agents as graphs

From Cognitive Architectures to Language Agents: A Mechanism-Level Review of Lineage, Convergence, and Migration Gaps

AgentsDGX agent

arXiv:2607.23942v1 Announce Type: new Abstract: Memory, planning, reflection, and tool use are often compared as feature labels, obscuring the control semantics that determine how an agent actually ru

GNN-based Multi-Agent Control of Traffic Shockwaves in Sparse Vehicular Ad-hoc Networks

AgentsDGX agent

arXiv:2607.23792v1 Announce Type: cross Abstract: Traffic shockwaves are stop-and-go waves that propagate upstream through the streams of vehicles and are one of the major causes of traffic congestion

Grading the Narrators: An Isnad-Rijal Framework for Claim-Level Provenance in Multi-Agent Knowledge Systems

AgentsDGX agent

arXiv:2607.24117v1 Announce Type: new Abstract: Modern multi-agent knowledge systems increasingly accumulate knowledge through chains of autonomous transformations rather than direct retrieval. Existi

Greedy dynamical meta-learning

AgentsDGX agent

arXiv:2607.23925v1 Announce Type: new Abstract: Gradient descent scales well to large models, but becomes unstable over long time horizons. Gradient-free optimizers can scale to arbitrary timespans, b

Happy Stateless MCP day!

AgentsDGX agent

Happy Stateless MCP day! I just got taught that you can use Cloudflare's AI Playground: https://playground.ai.cloudflare.com/ to play with MCP servers like: https://keyboardia.dev/mcp or: https://agen

← Previous
1…1011121314…120
Next →