AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlog
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,958 results
Agents

Compass: Navigating Global Marine Lead Data Integration through Expert-Guided LLM Agent

DGX agent

arXiv:2605.29966v1 Announce Type: new Abstract: Marine lead (Pb) and its isotopes are critical tracers for ocean circulation and anthropogenic pollution, yet in-situ observations remain costly and spa

agentsarxiv-cs-ai
29 May 2026
Safety
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

Graph-Enhanced Policy Optimization in LLM Agent Training

DGX agent

arXiv:2510.26270v2 Announce Type: replace Abstract: Multi-step LLM agents in interactive environments represent a crucial step toward long-horizon decision-making. To train such agents, group-based re

safetyarxiv-cs-ai
29 May 2026
Agents

grok-build-0.1 is now available via the xAI API in public beta. This is the same model that powers the Grok Build CLI and excels at agentic …

DGX agent

grok-build-0.1 is now available via the xAI API in public beta. This is the same model that powers the Grok Build CLI and excels at agentic coding. Priced at 1/m input and 2/m output, it’s extremely c

agentselon-musk--x
29 May 2026
Research

Honest Lying: Understanding Memory Confabulation in Reflexive Agents

DGX agent

arXiv:2605.29463v1 Announce Type: cross Abstract: Reflexion-style agents rely on self-generated reflections as memory, implicitly assuming that agents can accurately diagnose their own failures.We sho

researcharxiv-cs-ai
29 May 2026
Agents

i had to pack away my coding agents on monday cuz I knew I’d stay up too late if I played w activegraph during the week (yay, it’s Friday!) …

DGX agent

i had to pack away my coding agents on monday cuz I knew I’d stay up too late if I played w activegraph during the week (yay, it’s Friday!) gautham kept playing with it and just showed me a custom UI

agentsyohei-nakajima--x
29 May 2026
Agents

I literally haven’t typed anything in weeks Since I started using Grok’s speech-to-text in Hermes Agent... I just talk It transcribes everyt…

DGX agent

I literally haven’t typed anything in weeks Since I started using Grok’s speech-to-text in Hermes Agent... I just talk It transcribes everything perfectly Every word. Every single time My thoughts flo

agentsnous-research--x
29 May 2026
Agents

I've been using state-of-the-art models to teach small models running on my computer how I work. The result : a personal agent that runs my …

DGX agent

I've been using state-of-the-art models to teach small models running on my computer how I work. The result : a personal agent that runs my inbox, my deal pipeline, my blog, my calendar, & my research

agentsswyx--x
29 May 2026
Safety

Learning to Choose: An Empowerment-Guided Multi-Agent System with semantic communication for Adaptive Method Selection

DGX agent

arXiv:2605.30042v1 Announce Type: new Abstract: Automating scientific computing workflows requires more than generating executable code: autonomous systems must also select appropriate computational s

safetyarxiv-cs-ai
29 May 2026
Agents

Most people training agentic LLMs with RL right now have a silently broken training loop and have no idea. Here's the trap: single-turn RL w…

DGX agent

Most people training agentic LLMs with RL right now have a silently broken training loop and have no idea. Here's the trap: single-turn RL works beautifully. Clean curves, sane rewards, everything con

agentsclem-delangue--x
29 May 2026
Agents

Open Source Browser Agent That Learns and Repeats Workflows

DGX agent

An open-source browser with built-in AI agents that emphasizes privacy and automation, enabling task automation through natural language without coding. The browser supports multiple AI providers incl

agentsr-ollama
29 May 2026
Safety

PersonaAgent: Bridging Memory and Action for Personalized LLM Agents

DGX agent

arXiv:2506.06254v2 Announce Type: replace Abstract: Large Language Model (LLM) empowered agents have recently emerged as advanced paradigms that exhibit impressive capabilities in a wide range of doma

safetyarxiv-cs-ai
29 May 2026
Agents

Sources: Microsoft is working on an app that will include GitHub Copilot, Copilot chat, Copilot Cowork, and a new agentic workflow tool called Autopilot (Sebastian Herrera/Fortune)

DGX agent

Sebastian Herrera / Fortune: Sources: Microsoft is working on an app that will include GitHub Copilot, Copilot chat, Copilot Cowork, and a new agentic workflow tool called Autopilot — Microsoft needs

agentstechmeme
29 May 2026
Model Releases

STAMP: Training Explicit Memory for Mobile GUI Agents in Controllable and Scalable Virtual Environments

DGX agent

arXiv:2605.29324v1 Announce Type: new Abstract: Mobile GUI agents excel at immediate reactive control but frequently fail in realistic, long-horizon tasks that require memory. This failure stems from

model-releasesarxiv-cs-cl
29 May 2026
Agents

Worked on some code this morning using Opus 4.8 and so far I'm really liking it. Much more cooperative than 4.7 and less 'over agentic'. Sto…

DGX agent

Worked on some code this morning using Opus 4.8 and so far I'm really liking it. Much more cooperative than 4.7 and less 'over agentic'. Stops and asks for my input when needed in places 4.7 (and GPT

agentsjeremy-howard--x
29 May 2026
Agents

Asana acquires StackAI, a no-code platform for building AI agents, for 75M as part of Asana's broader AI pivot; PitchBook: StackAI raised ~20M (Russell Brandom/TechCrunch)

DGX agent

Russell Brandom / TechCrunch: Asana acquires StackAI, a no-code platform for building AI agents, for 75M as part of Asana's broader AI pivot; PitchBook: StackAI raised ~20M — Asana has acquired the wo

agentstechmeme
28 May 2026
Model Releases

Coherence Collapse: Diagnosing Why Code Agents Fail After Reaching the Right Code

DGX agent

arXiv:2603.24631v2 Announce Type: replace-cross Abstract: Code agents resolve 65-70% of SWE-bench Verified issues, but Pass@1 cannot tell us why the rest fail, and, as we show, capable-model failures

model-releasesarxiv-cs-ai
28 May 2026
Agents

Deploying a Hermes Agent with Fly, Modal, OpenRouter, & Cloudflare 02:43 Managed vs VPS 06:08 Architecture 15:14 Setup 22:23 Deployment 28:0…

DGX agent

Deploying a Hermes Agent with Fly, Modal, OpenRouter, & Cloudflare 02:43 Managed vs VPS 06:08 Architecture 15:14 Setup 22:23 Deployment 28:00 Access / OIDC 39:57 Hermes and Open WebUI 52:07 Cloudflare

agentsnous-research--x
28 May 2026
Agents

Discovery Agents for Real-Time Analytics: Toward Proactive Insight Systems

DGX agent

arXiv:2605.27571v1 Announce Type: new Abstract: Modern analytics systems are fundamentally reactive, requiring users to define queries over increasingly complex and continuously evolving data. In real

agentsarxiv-cs-ai
28 May 2026
Agents

Hierarchical Prompt-Domain Control and Learning for Resource-Constrained Agentic Language Models

DGX agent

arXiv:2605.27703v1 Announce Type: new Abstract: Large Language Models are increasingly deployed inside agentic systems, where they must follow structured protocols, adapt to evolving states, and opera

agentsarxiv-cs-ai
28 May 2026
Agents

How we built Cloudflare's data platform and an AI agent on top of it

DGX agent

Cloudflare describes the architecture and development of their unified data platform that integrates data from across their global network, along with an AI agent built on top to enable intelligent qu

agentscloudflare-ai
28 May 2026
Agents

Okta reports Q1 revenue up 11% YoY to 765M, vs. 752M est., says the agentic AI build-out is spiking demand for its identity tools; OKTA jumps 8%+ after hours (Samantha Subin/CNBC)

DGX agent

Samantha Subin / CNBC: Okta reports Q1 revenue up 11% YoY to 765M, vs. 752M est., says the agentic AI build-out is spiking demand for its identity tools; OKTA jumps 8%+ after hours — Okta beat Wall St

agentstechmeme
28 May 2026
Agents

President and Head of AI at Replit, @pirroh is the architect of Replit Agent and former Head of Applied Research at Google X. See him take t…

DGX agent

President and Head of AI at Replit, @pirroh is the architect of Replit Agent and former Head of Applied Research at Google X. See him take the stage with @refikanadol on day two of Vibecon. NYC, June

agentsreplit--x
28 May 2026
Model Releases

📢Qwen3.7-Max just hit #3 on ITbench-AA — a fresh benchmark testing how well models handle real-world enterprise IT tasks, agentic-style. 🔧…

DGX agent

📢Qwen3.7-Max just hit #3 on ITbench-AA — a fresh benchmark testing how well models handle real-world enterprise IT tasks, agentic-style. 🔧Agentic era, go with Qwen.🏃🏃 Artificial Analysis and IBM Resea

model-releasesqwen--x
28 May 2026
Agents

Roles with Rails: Contract-Preserving Role Evolution in Multi-Agent Structured Reasoning

DGX agent

arXiv:2605.28433v1 Announce Type: new Abstract: Role-based LLM multi-agent systems need adaptive role pools, yet adapting such systems is not merely a matter of prompt optimization: roles often carry

agentsarxiv-cs-cl
28 May 2026
Safety

Structured Agent Distillation for Large Language Model

DGX agent

arXiv:2505.13820v5 Announce Type: replace-cross Abstract: Large language models (LLMs) exhibit strong capabilities as decision-making agents by interleaving reasoning and actions, as seen in ReAct-sty

safetyarxiv-cs-ai
28 May 2026
Model Releases

VeriTrip: A Verifiable Benchmark for Travel Planning Agents over Unstructured Web Corpora

DGX agent

arXiv:2605.28683v1 Announce Type: new Abstract: Existing benchmarks have laid the foundation for travel planning agents by establishing API-centric paradigms. However, as the capabilities of Autonomou

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

BeyondSWE: Can Current Code Agent Survive Beyond Single-Repo Bug Fixing?

DGX agent

arXiv:2603.03194v2 Announce Type: replace Abstract: Current code-agent benchmarks primarily evaluate localized issue resolution within a single target repository, leaving under-tested many software en

model-releasesarxiv-cs-cl
27 May 2026
Agents

Building self-improving tax agents with Codex

DGX agent

This article describes how OpenAI's Codex model can be used to build autonomous tax agents capable of self-improvement through code generation and execution. The work demonstrates using large language

agentsopenai
27 May 2026
Agents

Communication Gain and Delay Cost Under Cross-Timestep Delays in Cooperative Multi-Agent Reinforcement Learning

DGX agent

arXiv:2604.03785v2 Announce Type: replace Abstract: Communication is essential for coordination in cooperative multi-agent reinforcement learning under partial observability, yet cross-timestep delays

agentsarxiv-cs-ai
27 May 2026
Safety

Counterfactual Credit Policy Optimization for Multi-Agent Collaboration

DGX agent

arXiv:2603.21563v2 Announce Type: replace Abstract: Collaborative multi-agent large language models (LLMs) can solve complex reasoning tasks by decomposing roles, but reinforcement learning for such s

safetyarxiv-cs-ai
27 May 2026
Agents

Demis Hassabis says he still broadly expects AGI around 2030, though he now sees 2029 as a possibility, and 2026's 'agentic era' is a 'bit like a practice run' (Ina Fried/Axios)

DGX agent

Ina Fried / Axios: Demis Hassabis says he still broadly expects AGI around 2030, though he now sees 2029 as a possibility, and 2026's “agentic era” is a “bit like a practice run” — Google DeepMind CEO

agentstechmeme
27 May 2026
Agents

From data overload to actionable insights: How Verizon Connect scaled agentic AI to 100,000 users

DGX agent

In this post, we show you how Verizon Connect built and scaled an agentic AI solution to transform overwhelming fleet data into clear, actionable insights for 100,000 users daily. We walk you through

agentsaws-ml-blog
27 May 2026
Agents

I really appreciate the lessons and technical ideas @samaysham & team were able to share about their tax agent system, which learns from pro…

DGX agent

I really appreciate the lessons and technical ideas @samaysham & team were able to share about their tax agent system, which learns from production traces to self-improve via detailed tracing tightly

agentslinus-lee--x
27 May 2026
Agents

Interactive Agents: Simulating Counselor-Client Psychological Counseling via Role-Playing LLM-to-LLM Interactions

DGX agent

arXiv:2408.15787v2 Announce Type: replace Abstract: Creating effective dialogue systems for mental health support requires high-quality multi-turn counseling dialogue data, yet collecting real counsel

agentsarxiv-cs-cl
27 May 2026
Model Releases

JobBench: Aligning Agent Work With Human Will

DGX agent

arXiv:2605.26329v1 Announce Type: new Abstract: Current benchmarks for occupational AI agents are scoped primarily by economic values, telling a replacement story. We introduce JobBench, which evaluat

model-releasesarxiv-cs-ai
27 May 2026
Safety

RICE-PO: Turning Retrieval Interactions into Credit Signals for Reasoning Agents

DGX agent

arXiv:2605.26352v1 Announce Type: new Abstract: Retrieval is increasingly moving from one-shot matching toward interactive reasoning, where language agents iteratively inspect evidence, reformulate qu

safetyarxiv-cs-cl
27 May 2026
Agents

SPEAR: Code-Augmented Agentic Prompt Optimization

DGX agent

arXiv:2605.26275v1 Announce Type: new Abstract: Automatic prompt engineering (APE) rewrites prompts to improve downstream task performance, but existing APE loops treat the optimizer itself as a fixed

agentsarxiv-cs-cl
27 May 2026
Model Releases

SWE-Adept: An LLM-Based Agentic Framework for Deep Codebase Analysis and Structured Issue Resolution

DGX agent

arXiv:2603.01327v2 Announce Type: replace-cross Abstract: Large language models (LLMs) exhibit strong performance on self-contained programming tasks. However, they still struggle with repository-leve

model-releasesarxiv-cs-cl
27 May 2026
Safety

Think Twice Before You Act: Enhancing Agent Behavioral Safety with Thought Correction

DGX agent

arXiv:2505.11063v3 Announce Type: replace Abstract: LLM-based agents solve complex tasks through iterative reasoning, tool use, and environment interaction, where each intermediate thought directly sh

safetyarxiv-cs-ai
27 May 2026
Model Releases

Verus-SpecGym: An Agentic Environment for Evaluating Specification Autoformalization

DGX agent

arXiv:2605.26457v1 Announce Type: cross Abstract: AI coding agents are increasingly used to write real-world software, but ensuring that their outputs are correct remains a fundamental challenge. Form

model-releasesarxiv-cs-ai
27 May 2026
Agents

Automated Detection and Classification of Delusion-related Content in Naturalistic Audio Diaries Using Multi-Agent Language Models

DGX agent

arXiv:2605.24755v1 Announce Type: new Abstract: Speech monologues recorded in naturalistic settings provide opportunities to characterize mental illness phenomenology and detect symptom exacerbation.

agentsarxiv-cs-ai
26 May 2026
Agents

Board Mode is here. Instead of one long agent session, you now get a creative workspace where research, slides, websites, apps, and design d…

DGX agent

Board Mode is here. Instead of one long agent session, you now get a creative workspace where research, slides, websites, apps, and design drafts can live side by side. Try it now: https://agent.ii.in

agentsemad-mostaque--x
26 May 2026
Agents

Inference-Time Backdoors via Chat Templates: From LLM Supply Chains to Agentic System Compromise

DGX agent

arXiv:2602.04653v4 Announce Type: replace-cross Abstract: Open-weight language models are increasingly used in production settings, raising new security challenges. One prominent threat is backdoor at

agentsarxiv-cs-lg
26 May 2026
Agents

Just built an insane new agent skill. It can perfectly extract slides from YT videos, then write notes, images, transcripts, and slides into…

DGX agent

Just built an insane new agent skill. It can perfectly extract slides from YT videos, then write notes, images, transcripts, and slides into Obsidian vaults. An HTML artifact allows me to navigate and

agentsdair-ai--x
26 May 2026
Agents

Methods for Formal Verification of Agent Skills: Three Layers Toward a Mechanically Checkable Capability-Containment Proof

DGX agent

arXiv:2605.23951v1 Announce Type: new Abstract: The companion paper introduced a four-level verification lattice on agent-skill manifests (unverified, declared, tested, formal) and left the top level

agentsarxiv-cs-ai
26 May 2026
Model Releases

Novee debuts Agentic Fix, pushing pentest findings into Claude, Copilot and Cursor

DGX agent

Artificial intelligence penetration testing startup Novee Cyber Security Ltd. today launched Agentic Fix, a new capability that pushes validated exploit findings directly into the AI coding agents dev

model-releasessiliconangle
26 May 2026
Agents

VineLM: Trie-Based Fine-Grained Control for Agentic Workflows

DGX agent

arXiv:2605.23914v1 Announce Type: cross Abstract: Agentic workflows interleave configurable LLM stages with tool stages and often include retries or refinement loops. Existing workflow managers profil

agentsarxiv-cs-ai
26 May 2026
Agents

Why Agentic Theorem Prover Works: A Statistical Provability Theory of Mathematical Reasoning Models

DGX agent

arXiv:2602.10538v3 Announce Type: replace-cross Abstract: Agentic theorem provers combine a reasoning model, retrieval, search, and a proof assistant verifier, yet it remains unclear which components

agentsarxiv-cs-lg
26 May 2026
← Previous
1…104105106107108…375
Next →