AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,959 results
29 Jun 2026

Multi-tenant LLM analytics with row-level security: How we built a secure agent on AWS

AgentsDGX agent

In this post, we show you how PAR built a production-ready multi-tenant LLM analytics system that enforces row-level security through a three-layer architecture: cryptographic request signing with AWS

Ornith-1.0: Self-Scaffolding LLMs for Agentic Coding

Model ReleasesDGX agent

Ornith-1.0: Self-Scaffolding LLMs for Agentic Coding This is an interesting new open weights (MIT licensed) model, the first model release from DeepReinforce. [...] with variants including 9B Dense, 3

28 Jun 2026

Quoting Jon Udell

AgentsDGX agent

Human Agent in the loop I dislike the phrase “human in the loop” because it cedes authority to the machines. Let’s flip the narrative. It’s our loop, we work the same way we always have, now we recrui

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
27 Jun 2026

Using Local Coding Agents

Model ReleasesDGX agent

This article likely explores how to implement and deploy coding agents that run locally on a user's machine rather than relying on cloud-based services, potentially covering advantages like privacy, l

26 Jun 2026

Autoformalization of Agent Instructions into Policy-as-Code

Model ReleasesDGX agent

arXiv:2606.26649v1 Announce Type: new Abstract: Agent safety in high-stakes domains requires formal policy enforcement, but most existing approaches either rely on probabilistic guardrails (fine-tuned

OPID: On-Policy Skill Distillation for Agentic Reinforcement Learning

SafetyDGX agent

arXiv:2606.26790v1 Announce Type: new Abstract: Outcome-based reinforcement learning provides a stable optimization backbone for language agents, but its sparse trajectory-level rewards provide little

Temporal Validity in Retrieval Memory: Eliminating Stale-Fact Errors for AI Agents over Evolving Knowledge

Local AiDGX agent

arXiv:2606.26511v1 Announce Type: cross Abstract: Retrieval-augmented generation (RAG) gives agents access to accumulated knowledge, but has no model of time. When a fact changes (e.g., a function is

25 Jun 2026

Agentic evolution of physically constrained foundation models

Model ReleasesDGX agent

arXiv:2606.25532v1 Announce Type: cross Abstract: Artificial intelligence increasingly drives automated scientific discovery, yet contemporary generalist agents lack physical grounding, frequently hal

BiPACE: Bisimulation-Guided Policy Optimization with Action Counterfactual Estimation for LLM Agents

Local AiDGX agent

arXiv:2606.25556v1 Announce Type: new Abstract: Stepwise group-based RL is an attractive way to train long-horizon LLM agents without a learned critic: it reuses multiple sampled rollouts to estimate

BrainAgent: A Large Language Model-Driven Multi-Agent Framework for Autonomous Brain Signal Understanding

Model ReleasesDGX agent

arXiv:2606.25400v1 Announce Type: new Abstract: Brain-Computer Interfaces (BCIs) and brain signal understanding are pivotal for clinical health and next-generation interactions. Despite this significa

Neglected Free Lunch from Post-training: Progress Advantage for LLM Agents

SafetyDGX agent

arXiv:2606.26080v1 Announce Type: new Abstract: Process reward models enable fine-grained, step-level evaluation of LLMs, yet building them for agentic settings remains prohibitively difficult: long-h

Staying In Character: Perspective-Bounded Memory For Book-Based Role-Playing Agents

Model ReleasesDGX agent

arXiv:2606.25632v1 Announce Type: new Abstract: Recent LLM role-playing systems build character agents from novels by extracting characters, scenes, and relations. Yet long-narrative role-playing suff

TRUSTMEM: Learning Trustworthy Memory Consolidation for LLM Agents with Long-Term Memory

ResearchDGX agent

arXiv:2606.25161v1 Announce Type: new Abstract: Large language model (LLM) agents rely on long-term memory to support extended interactions and personalized assistance beyond finite context windows. E

While we eagerly await Fable 5's return, our agentic WebGPU kernel optimization framework kept running. Opus 4.8 picked up where Fable left …

Model ReleasesDGX agent

While we eagerly await Fable 5's return, our agentic WebGPU kernel optimization framework kept running. Opus 4.8 picked up where Fable left off, pushing Liquid AI's new LFM2.5 230M to an unbelievable

24 Jun 2026

Google says computer use is now a built-in tool supported in Gemini 3.5 Flash, available via the Gemini API and Gemini Enterprise Agent Platform (Mateo Quiros/The Keyword)

Model ReleasesDGX agent

Mateo Quiros / The Keyword: Google says computer use is now a built-in tool supported in Gemini 3.5 Flash, available via the Gemini API and Gemini Enterprise Agent Platform — Computer use is now a bui

happy karpathy agent day for those who celebrate

AgentsDGX agent

This post by Swyx appears to be a casual, celebratory message referencing Andrej Karpathy (a prominent AI researcher) in the style of a holiday greeting, likely shared on a date significant to the AI

How Loka Built a Natural, Low-Latency Voice Agent with Amazon Nova 2 Sonic

AgentsDGX agent

In this post, we demonstrate the architecture and approach Loka used to solve a common frustration: robotic, slow voice assistants that cause customers to hang up, damaging brand reputation and drivin

LecturaAgents: A Multi-Agent Framework for Adaptive Personalized AI-Assisted Learning and Embodied Teaching

SafetyDGX agent

arXiv:2606.16428v2 Announce Type: replace-cross Abstract: Effective personalized AI-assisted learning demands systems that can not only generate accurate learner-specific educational materials, but al

Multi-agent imitation learning with function approximation: Linear Markov games and beyond

SafetyDGX agent

arXiv:2602.22810v2 Announce Type: replace Abstract: In this work, we present the first theoretical analysis of multi-agent imitation learning (MAIL) in linear Markov games where both the transition dy

NatureBench: Can Coding Agents Match the Published SOTA of Nature-Family Papers?

Model ReleasesDGX agent

arXiv:2606.24530v1 Announce Type: new Abstract: We introduce NatureBench, a cross-discipline benchmark of 90 tasks distilled from peer-reviewed Nature-family publications, designed to evaluate whether

Seltz, which is building a web search engine that can be used by AI agents, raised a $12.5M seed led by Speedinvest and B Capital (Jeremy Kahn/Fortune)

IndustryDGX agent

Jeremy Kahn / Fortune: Seltz, which is building a web search engine that can be used by AI agents, raised a $12.5M seed led by Speedinvest and B Capital — The rise of AI has rekindled the long-dormant

23 Jun 2026

AD-Bench: A Real-World, Trajectory-Aware Advertising Analytics Benchmark for LLM Agents

Model ReleasesDGX agent

arXiv:2602.14257v2 Announce Type: replace-cross Abstract: While Large Language Model (LLM) agents have made remarkable progress on complex reasoning, evaluating them in real-world environments remains

Compressing Observation History into Agent Memory: Distilling Transformers into Recurrent Transformers

AgentsDGX agent

arXiv:2606.21562v1 Announce Type: new Abstract: Transformers are AI's workhorse with strong performance in modeling sequential data, but their computational cost becomes prohibitive when processing lo

ENVS: Environment-Native Verified Search for Long-Horizon GUI Agents

Model ReleasesDGX agent

arXiv:2606.22948v1 Announce Type: cross Abstract: As multimodal agents move from interface understanding to real software control, successful trajectory discovery in live desktop environments becomes

GitOfThoughts: Version-Controlled Reasoning and Agent Memory You Can Replay, Diff, and Merge

Model ReleasesDGX agent

arXiv:2606.14470v2 Announce Type: replace-cross Abstract: Large language model reasoning leaves no trace once it is done. The steps of a chain of thought disappear when the context window closes, a pr

HoloAgent-0: A Unified Embodied Agent Framework with 3D Spatial Memory

SafetyDGX agent

arXiv:2606.23565v1 Announce Type: cross Abstract: LLM agents follow a practical execution loop in digital environments: they reason over structured states, invoke tools, inspect feedback, and revise a

(Look mom - I made it!) Learn more about configuring agent identity and some of the key decisions we made for Claude Tag- https://claude.com…

Model ReleasesDGX agent

This post likely discusses configuration options for setting up Claude agent identity and explains the design decisions behind Claude Tag, Anthropic's tool for customizing Claude's behavior and person

RIZZ: Routing Interactions to Near Zero-Interference Zones for Continual Adaptation of Black-Box Agents

Local AiDGX agent

arXiv:2606.20638v1 Announce Type: cross Abstract: Large language models are increasingly deployed as long-lived agents that must adapt across users, tasks, domains, modalities, and feedback regimes wi

RoboLineage: Agent-Native Data Lifecycle Governance Across Robot Policy Iterations

SafetyDGX agent

arXiv:2606.22142v1 Announce Type: new Abstract: We present RoboLineage, an agent-native data lifecycle governance system for robot policy iteration. Modern robot policies improve through repeated data

Sovereign Execution Broker: Enforcing Certificate-Bound Authority in Agentic Control Planes

SafetyDGX agent

arXiv:2606.20520v2 Announce Type: replace-cross Abstract: Autonomous agents are increasingly connected to cloud, deployment, and data-control workflows, but production mutation authority should not re

When Web Agents Finish but Still Fail: Reproducible Triggers and Trace Diagnostics for Parallel Web Exploration

Model ReleasesDGX agent

arXiv:2606.20724v1 Announce Type: cross Abstract: Long-horizon web agents often fail in ways hidden by final-answer evaluation: they may visit useful pages, produce a well-formed answer, and terminate

22 Jun 2026

Built a local codebase memory for agentic IDEs using Ollama + ChromaDB; zero cloud required

Local AiDGX agent

A developer created a local codebase memory system for agentic integrated development environments (IDEs) using Ollama and ChromaDB, enabling AI-assisted coding without reliance on cloud services. The

GLM-5.2 is the step change for open agents

ResearchDGX agent

GLM-5.2 represents a significant advancement in open-source AI agent capabilities, marking a notable improvement over previous versions in terms of performance and functionality. The article from Inte

Sakana AI launches Fugu, a multi-agent orchestration system accessible through a single model API, claiming Fugu Ultra matches Fable and Mythos on benchmarks (Carl Franzen/VentureBeat)

Model ReleasesDGX agent

Carl Franzen / VentureBeat: Sakana AI launches Fugu, a multi-agent orchestration system accessible through a single model API, claiming Fugu Ultra matches Fable and Mythos on benchmarks — Last night,

That's right, an open, MIT-licensed model beating GPT-5.5 (xhigh) on real-world agentic work! 🔥 Available for free on @huggingface for anyo…

Model ReleasesDGX agent

That's right, an open, MIT-licensed model beating GPT-5.5 (xhigh) on real-world agentic work! 🔥 Available for free on @huggingface for anyone to build on top off GLM-5.2 leads open weights models and

The Starter Tier for Google AI Studio explained

Model ReleasesDGX agent

You've got a working prototype in Google AI Studio. A React frontend, a Node.js backend, maybe a database. Now you want a live URL to share with your team, your users, or a friend who wants to try it.

There are basically two skillsets that matter more in the age of coding agents. The first is the one everyone talks about: product sense and…

ApplicationsDGX agent

There are basically two skillsets that matter more in the age of coding agents. The first is the one everyone talks about: product sense and taste. Code is cheap now, so knowing exactly what to build

11 Jun 2026

A Lightweight Multi-Agent Framework for Automated Concrete Barrier Design

Model ReleasesDGX agent

arXiv:2606.12040v1 Announce Type: new Abstract: The design of reinforced concrete highway barriers is a safety-critical process that requires strict compliance with regulatory provisions such as the A

AerialClaw: An Open-Source Framework for LLM-Driven Autonomous Aerial Agents

SafetyDGX agent

arXiv:2606.12142v1 Announce Type: cross Abstract: Unmanned aerial vehicles (UAVs) are increasingly used in inspection, search and rescue, environmental monitoring, and emergency response. However, mos

FORT-Searcher: Synthesizing Shortcut-Resistant Search Tasks for Training Deep Search Agents

ResearchDGX agent

arXiv:2606.12087v1 Announce Type: new Abstract: Training deep search agents requires verifiable questions whose answers remain unavailable until sufficient evidence has been acquired through search. E

Human-Enhanced Loop Modeling (HELM): Agent-Based Finite Element Modeling of Concrete Bridge Barriers

SafetyDGX agent

arXiv:2606.12025v1 Announce Type: new Abstract: Finite element (FE) modeling of safety-critical infrastructure such as bridge barriers requires high-fidelity nonlinear dynamic analysis, yet the curren

IAPO: Input Attribution-Aware Policy Optimization for Tool Use in Small Multimodal Agents

SafetyDGX agent

arXiv:2606.11652v1 Announce Type: new Abstract: This paper investigates reinforcement learning (RL) methods for improving tool-calling capabilities in multimodal small language model (SLM) agents. Whi

NightFeats @ MMU-RAGent NeurIPS 2025: A Context-Optimized Multi-Agent RAG System for the Text-to-Text Track

Model ReleasesDGX agent

arXiv:2606.11199v1 Announce Type: cross Abstract: We present NightFeats, a structured multi-agent retrieval-augmented generation (RAG) system submitted to the MMU-RAGent competition at NeurIPS 2025, w

Sovereign Assurance Boundary: Certificate-Bound Admission for Agentic Infrastructure

SafetyDGX agent

arXiv:2606.11632v1 Announce Type: cross Abstract: Agentic infrastructure introduces a critical control-plane authorization problem: non-deterministic reasoning systems can propose high-stakes mutation

10 Jun 2026

3SPO: State-Score-Supervised Policy Optimization for LLM Agents

SafetyDGX agent

arXiv:2606.09961v1 Announce Type: cross Abstract: Training large language models (LLMs) as autonomous agents via reinforcement learning (RL) has enabled frontier models to achieve superhuman performan

ASA: Backbone-Training-Free Representation Engineering for Tool-Calling Agents

Model ReleasesDGX agent

arXiv:2602.04935v3 Announce Type: replace-cross Abstract: Adapting LLM agents to domain-specific tool calling remains notably brittle under evolving interfaces. Prompt and schema engineering is easy t

AsyncWebRL: Efficient Multi-Step RL for Visual Web Agents

SafetyDGX agent

arXiv:2606.05597v2 Announce Type: replace Abstract: Training vision-language web agents with multi-step RL is compute-intensive, with two dominant forms of inefficiency: idle GPUs in synchronous RL, a

Can Multi-Agent LLMs Identify Their Peers? Stylometric Fingerprinting in Role-Constrained Political Analysis

Model ReleasesDGX agent

arXiv:2606.09854v1 Announce Type: cross Abstract: Multi-agent large language model (LLM) pipelines for political statement analysis are vulnerable to peer-preservation bias: models tend to protect pee

Decoupling Thought from Speech: Knowledge-Grounded Counterfactual Reasoning for Resilient Multi-Agent Argumentation

SafetyDGX agent

arXiv:2606.10475v1 Announce Type: cross Abstract: Multi-agent debate frameworks have been shown to improve large language model performance in convergent tasks, but they are currently optimized in a w

IntentKV: Cross-Turn Intent-Aware KV Cache Pruning for Agent Inference

Model ReleasesDGX agent

arXiv:2606.09916v1 Announce Type: cross Abstract: Multi-turn LLM agents fan short queries into long trajectories of tool calls, search results, and intermediate reasoning. Both KV memory and KV read b

Jedify raises $24M to give enterprise AI agents the business context they lack

ApplicationsDGX agent

Enterprise artificial intelligence context startup Jedify Inc. today announced that it has raised 24 million in new funding to build what it calls context graphs that give AI agents the business knowl

Learning What to Remember: Observability-Safe Memory Retention via Constrained Optimization for Long-Horizon Language Agents

Local AiDGX agent

arXiv:2606.10616v1 Announce Type: new Abstract: Long-horizon language agents accumulate observations, reasoning traces, and retrieved facts that exceed their finite context windows, making memory rete

New York-based Jedify, whose platform connects to enterprises' knowledge sources via APIs to build a 'context graph' for AI agents, raised a $24M Series A (Ram Iyer/TechCrunch)

ApplicationsDGX agent

Ram Iyer / TechCrunch: New York-based Jedify, whose platform connects to enterprises' knowledge sources via APIs to build a “context graph” for AI agents, raised a $24M Series A — AI vendors promote t

OpenAI and Visa partner to let AI agents make purchases online after users give their permission and to explore enterprise applications for AI-driven payments (Paige Smith/Bloomberg)

ApplicationsDGX agent

Paige Smith / Bloomberg: OpenAI and Visa partner to let AI agents make purchases online after users give their permission and to explore enterprise applications for AI-driven payments — OpenAI and Vis

RedAct: Redacting Agent Capability Traces for Procedural Skill Protection

Model ReleasesDGX agent

arXiv:2606.10813v1 Announce Type: cross Abstract: Users rely on execution traces to observe agent behavior, diagnose failures, and ensure accountability. These traces contain rich procedural detail, i

SkillResolve-Bench: Measuring and Resolving Same-Capability Ambiguity in Agent Skill Retrieval

Model ReleasesDGX agent

arXiv:2606.10388v1 Announce Type: cross Abstract: Agent skill libraries are becoming routable software assets: a retrieved skill can contribute instructions, scripts, resource bindings, and execution

9 Jun 2026

AliyunConsoleAgent: Training Web Agents in Real-World Cloud Environments via Distillation and Reinforcement Learning

Model ReleasesDGX agent

arXiv:2606.09447v1 Announce Type: new Abstract: We present AliyunConsoleAgent, a web agent framework for automated documentation verification in real-world cloud consoles. Major cloud platforms encomp

AWS debuts AWS FinOps Agent to help customers optimize their cloud spending

Model ReleasesDGX agent

Amazon Web Services Inc. today introduced an artificial intelligence tool designed to help organizations lower their cloud bills. AWS FinOps Agent is initially available in public preview. It’s the la

Claude Fable 5 is now supported for use in Hermes Agent via Nous Portal! The first 500 new users get one month free access to the Plus plan …

Model ReleasesDGX agent

Nous Research announced support for Claude Fable 5 integration with Hermes Agent through the Nous Portal, offering new users a promotional one-month free trial of the Plus plan for the first 500 signu

Crayotter: Traceable Multi-Agent Workflows for Long-Form Video Editing

SafetyDGX agent

arXiv:2606.07636v1 Announce Type: new Abstract: Editing a long-form video from heterogeneous footage requires more than selecting clips: an agent must preserve narrative intent across material prepara

← Previous
1…115116117118119…300
Next →