AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlog
84,630Total entries
1Added by human
84,629Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,980 results
Agents

Better Understanding, Understanding Better

DGX agent

arXiv:2606.31892v1 Announce Type: cross Abstract: 'Any fool can know; the point is to understand.' A well-known remark often attributed to Einstein captures a widely shared intuition: understanding is

agentsarxiv-cs-ai
1 Jul 2026
Agents

Emergent Culture in Minimal LLM Systems

DGX agent

arXiv:2606.30668v1 Announce Type: cross Abstract: What happens when LLM agents operate with no context outside a turn, minimal prompting, and simple tools? Inspired by swarm engineering, we give colle

X Post
Paper
YouTube
Reddit
GitHub
Clear filters
agentsarxiv-cs-ai
1 Jul 2026
Agents

Imagine if, as an engineer, your working memory of a codebase was wiped every time you started a new ticket Sounds ridiculous, but this is f…

DGX agent

Imagine if, as an engineer, your working memory of a codebase was wiped every time you started a new ticket Sounds ridiculous, but this is functionally what happens to agents whenever we start a new t

agentsharrison-chase--x
1 Jul 2026
Agents

Nazrin: An Atomic Neural Proof Automation Tactic in Lean 4

DGX agent

arXiv:2602.18767v3 Announce Type: replace-cross Abstract: In Machine-Assisted Theorem Proving, a theorem proving agent searches for a sequence of expressions and tactics that can prove a statement in

agentsarxiv-cs-lg
1 Jul 2026
Agents

a very cool Harbor x LangSmith flow I love to help you “look at the data”: 1. you do evals or rollouts for RL 2. all reward metrics and trac…

DGX agent

a very cool Harbor x LangSmith flow I love to help you “look at the data”: 1. you do evals or rollouts for RL 2. all reward metrics and traces and rollouts get automatically propulates into Experiment

agentsharrison-chase--x
30 Jun 2026
Agents

Developmental Trajectories of Situation Modeling and Mentalizing in Transformer Language Models

DGX agent

arXiv:2606.28524v1 Announce Type: new Abstract: Recent work suggests that Large Language Models (LLMs) are sensitive to the belief states of agents described by text, as measured by the false belief t

agentsarxiv-cs-cl
30 Jun 2026
Agents

UnfoldArt: Zero-Shot Recovery of Full Articulated 3D Objects from Text or Image

DGX agent

arXiv:2606.30608v1 Announce Type: new Abstract: Articulated 3D objects are essential for interactive environments in embodied AI, robotics, and virtual reality, but reconstructing their structure and

agentsarxiv-cs-cv
30 Jun 2026
Agents

We're all over AI Engineer World's Fair on June 29 to July 2. 🦙 📍Visit us at booth L-G47. LlamaParse demos + Fear of Docs swag 🎤 @jerryjl…

DGX agent

We're all over AI Engineer World's Fair on June 29 to July 2. 🦙 📍Visit us at booth L-G47. LlamaParse demos + Fear of Docs swag 🎤 @jerryjliu0, our Founder & CEO, on agentic document parsing and shippin

agentsjerry-liu--x
29 Jun 2026
Agents

I’m speaking at AIE SF World’s Fair on July 1st in the Sandbox & Platform Engineering track. My talk is called 'From fork() to Fleet: Design…

DGX agent

I’m speaking at AIE SF World’s Fair on July 1st in the Sandbox & Platform Engineering track. My talk is called 'From fork() to Fleet: Designing an Agent Sandbox Cloud'. I’ll cover the OS and Infrastru

agentsswyx--x
27 Jun 2026
Agents

Decoupling Reconnaissance and Exploitation: Measuring the Capability Boundaries of LLM-Based Web Penetration Testing

DGX agent

arXiv:2606.25332v1 Announce Type: cross Abstract: Large Language Models (LLMs) have shown promise for automated penetration testing, yet existing end-to-end black-box evaluations are highly susceptibl

agentsarxiv-cs-ai
25 Jun 2026
Agents

Loops, one of the most symbolic tools ever invented, are rescuing generative AI. Neurosymbolic AI is completely dominating.

DGX agent

Loops, one of the most symbolic tools ever invented, are rescuing generative AI. Neurosymbolic AI is completely dominating. Anthropic engineers just showed how to build agents that can run for days wi

agentsgary-marcus--x
25 Jun 2026
Agents

Tinker Tales: A Tangible Dialogue System for Child-AI Co-Creative Storytelling

DGX agent

arXiv:2602.04109v2 Announce Type: replace-cross Abstract: Conversational AI agents are increasingly explored as creative partners, yet how conversation design shapes child-AI dialogue in co-creative s

agentsarxiv-cs-ai
25 Jun 2026
Agents

From Task-Guided Conversational Graphs to Goal-Oriented Dialogue Runtimes

DGX agent

arXiv:2606.23797v1 Announce Type: cross Abstract: Graph and multi-agent orchestration frameworks make production large language model (LLM) workflows practical, but they do not by themselves solve con

agentsarxiv-cs-ai
24 Jun 2026
Agents

Sleep-time compute is the next scaling axis for intelligence

DGX agent

Sleep-time compute is the next scaling axis for intelligence 🧠LangSmith Engine as Sleep Time Compute Memory for agents is often described as “sleep time compute” or “dreaming” This involves running a

agentsharrison-chase--x
24 Jun 2026
Agents

Reinforcement Learning to Disentangle Multiqubit Quantum States from Partial Observations

DGX agent

arXiv:2406.07884v2 Announce Type: replace-cross Abstract: Using partial knowledge of a quantum state to control multiqubit entanglement is a largely unexplored paradigm in the emerging field of quantu

agentsarxiv-cs-lg
23 Jun 2026
Agents

Sakana Fugu Technical Report

DGX agent

arXiv:2606.21228v1 Announce Type: new Abstract: The capabilities of frontier Large Language Models (LLMs) continue to advance, with different providers increasingly specializing in distinct domains. T

agentsarxiv-cs-lg
23 Jun 2026
Agents

UECP: Uncertainty-Enhanced Collaborative Perception

DGX agent

arXiv:2606.23046v1 Announce Type: new Abstract: Collaborative perception serves as a pivotal solution to enhance the perception capability of individual agents in autonomous driving, where a core chal

agentsarxiv-cs-cv
23 Jun 2026
Agents

Something we’re thinking a bunch about as well Would love thoughts!

DGX agent

Something we’re thinking a bunch about as well Would love thoughts! context engineering docs for agentic engineering - plans, research, etc SHOULD NOT be stored in version control: A good docs managem

agentsharrison-chase--x
22 Jun 2026
Model Releases

Learning Ordinal Response Policies in Rank-Based Stochastic Prize-Collecting Games

DGX agent

arXiv:2510.24515v2 Announce Type: replace Abstract: The Team Orienteering Problem (TOP) generalizes many real-world multi-agent scheduling and routing tasks that occur in autonomous mobility, aerial l

model-releasesarxiv-cs-ro
11 Jun 2026
Agents

paper #1 for context: https://x.com/yoheinakajima/status/2057812713045377055?s=20

DGX agent

paper #1 for context: https://x.com/yoheinakajima/status/2057812713045377055?s=20 babyagi has ~200 citations, but 0 papers... i just published my first paper on arXiv 😆 'The Log is the Agent: Event-So

agentsyohei-nakajima--x
10 Jun 2026
Agents

(Auto)formalization is supposed to be easy: Trellis process semantics for spelling out rigorous proofs

DGX agent

arXiv:2606.09674v1 Announce Type: new Abstract: We present Trellis: an autoformalization system that leverages LLM agents in a deterministically constrained workflow to enforce incremental progress in

agentsarxiv-cs-ai
9 Jun 2026
Agents

FASE: Fast Adaptive Semantic Entropy for Code Quality

DGX agent

arXiv:2606.09800v1 Announce Type: cross Abstract: Multi-agent code generation offers a promising paradigm for autonomous software development by simulating the human software engineering lifecycle. Ho

agentsarxiv-cs-ai
9 Jun 2026
Agents

RepoLaunch: Automating Build and Management of Code Repositories across Languages and Platforms

DGX agent

arXiv:2603.05026v2 Announce Type: replace-cross Abstract: Language model (LM) agents have driven substantial progress in automated software engineering (SWE), yet building and testing software reposit

agentsarxiv-cs-lg
9 Jun 2026
Agents

To Nuke or Not to Nuke: LLMs' (Missing) Ethical Reasoning and Actions in a High-Stakes Decision-Making Simulation

DGX agent

arXiv:2606.08310v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly deployed as long-horizon agents with decision-making capacities. While LLMs can show ethical competence on

agentsarxiv-cs-ai
9 Jun 2026
Agents

What are loops, and how do you build one? A 'loop' is the repeated process where some event or input kicks off an action. For example: 1. CI…

DGX agent

What are loops, and how do you build one? A 'loop' is the repeated process where some event or input kicks off an action. For example: 1. CI fails -> you fix it 2. CI fails again -> you fix it 3. CI p

agentsharrison-chase--x
9 Jun 2026
Agents

IRAF: Interference-Resilient Adaptive Fusion for Noise-Robust End-to-End Full-Duplex Spoken Dialogue Systems

DGX agent

arXiv:2606.06559v1 Announce Type: cross Abstract: Full-duplex spoken dialogue models allow voice agents to listen and speak concurrently, enabling natural interaction with real-time overlap. However,

agentsarxiv-cs-ai
8 Jun 2026
Agents

A Finite Certificate for the Positive n=9 Vasc Inequality

DGX agent

arXiv:2606.06136v1 Announce Type: cross Abstract: We prove the positive-real n=9 case of the Vasc cyclic inequality. The proof was obtained with human-guided assistance from the AI agent MechMath Agen

agentsarxiv-cs-ai
6 Jun 2026
Agents

The Self-Correction Illusion: LLMs Correct Others but Not Themselves

DGX agent

arXiv:2606.05976v1 Announce Type: cross Abstract: Recent work shows that LLM agents struggle to correct errors in their own reasoning traces yet show markedly higher correction rates when identical cl

agentsarxiv-cs-cl
5 Jun 2026
Agents

Beyond Correctness: Rewarding Faithful Reasoning in Retrieval-Augmented Generation

DGX agent

arXiv:2510.13272v3 Announce Type: replace Abstract: Inspired by the success of reinforcement learning (RL) in Large Language Model (LLM) training for domains like math and code, recent work has begun

agentsarxiv-cs-cl
4 Jun 2026
Agents

Trivium: Temporal Regret as a First-Class Objective for Causal-Memory Controllers

DGX agent

arXiv:2606.04421v1 Announce Type: new Abstract: Many current agentic systems and LLM pipelines correct mistakes by optimizing outcome reward. This addresses only the what of failure: when an outcome d

agentsarxiv-cs-ai
4 Jun 2026
Agents

CARVE: Certified Affordable Repair of Vetoed Maneuvers via Envelopes for Interactive Driving

DGX agent

arXiv:2606.02641v1 Announce Type: cross Abstract: Interactive driving exposes a failure mode that is easy to miss in rule-aware autonomous-driving stacks: a hard-rule margin can be negative for an ego

agentsarxiv-cs-ai
3 Jun 2026
Agents

From Control Boundary to Insurance Claim: Reconstructing AI-Mediated Losses Through the CER Framework

DGX agent

arXiv:2606.03777v1 Announce Type: new Abstract: AI losses that arise through an insured organization's generative or agentic AI system require state reconstruction, not merely event reconstruction, be

agentsarxiv-cs-ai
3 Jun 2026
Model Releases

KForge: LLM-Driven Cross-Platform Kernel Generation for AI Accelerators

DGX agent

arXiv:2606.02963v1 Announce Type: new Abstract: Production inference increasingly targets a heterogeneous mix of accelerators. Agentic pipelines interleave reasoning, tool calls, and multi-agent coord

model-releasesarxiv-cs-lg
3 Jun 2026
Agents

MemTrain: Self-Supervised Context Memory Training

DGX agent

arXiv:2606.03197v1 Announce Type: new Abstract: Memory is an indispensable capability for long-horizon LLM agents, enabling them to preserve and utilize information accumulated across extended interac

agentsarxiv-cs-cl
3 Jun 2026
Agents

More info on the Web Dashboard: https://hermes-agent.nousresearch.com/docs/user-guide/features/web-dashboard

DGX agent

The Nous Research Web Dashboard is a user interface feature that provides access to information and tools for managing Hermes Agent functionality through a web-based platform. According to Nous Resear

agentsnous-research--x
3 Jun 2026
Agents

MindClaw: Closed-Loop Embodied Mental-State Reasoning for Precision Intervention

DGX agent

arXiv:2606.01063v1 Announce Type: new Abstract: Theory of Mind (ToM) enables an agent to reason about another actor's beliefs, goals, and intentions, which is essential for human-centered embodied ass

agentsarxiv-cs-ai
2 Jun 2026
Agents

OpenHospital: A Thing-in-itself Arena for Evolving and Benchmarking LLM-based Collective Intelligence

DGX agent

arXiv:2603.14771v3 Announce Type: replace Abstract: Large Language Model (LLM)-based Collective Intelligence (CI) presents a promising approach to overcoming the data wall and continuously boosting th

agentsarxiv-cs-ai
2 Jun 2026
Agents

SWE-rebench V2: Language-Agnostic SWE Task Collection at Scale

DGX agent

arXiv:2602.23866v2 Announce Type: replace-cross Abstract: Software engineering agents (SWE) are improving rapidly, with recent gains largely driven by reinforcement learning (RL). However, RL training

agentsarxiv-cs-cl
2 Jun 2026
Agents

Don't Fool Me Twice: Adapting to Adversity in the Wild with Experience-Driven Reasoning

DGX agent

arXiv:2605.31119v1 Announce Type: cross Abstract: In robotics, dangers and adversity modes are often embodiment-specific and relative to each agent. A frontier of autonomous mobile robotics is to enab

agentsarxiv-cs-lg
1 Jun 2026
Model Releases

PInVerify: An Offline Embodied Benchmark for Active Instance Verification

DGX agent

arXiv:2605.30639v1 Announce Type: cross Abstract: Embodied agents have made strong progress in navigating to target objects, but reaching the goal vicinity does not guarantee that the agent has found

model-releasesarxiv-cs-ai
1 Jun 2026
Agents

Bosses, Kings, and the Commons: Cooperation Under Power Asymmetry in LLM Societies

DGX agent

arXiv:2605.29062v1 Announce Type: new Abstract: Communities can sustainably manage shared resources (commons) through self-governance and cooperative norms, a central finding of Ostrom's theory of sel

agentsarxiv-cs-cl
29 May 2026
Agents

FLIP: Real-Time and Resilient Formation Planning for Large-Scale DIstributed Swarms via Point Cloud Registration

DGX agent

arXiv:2605.29704v1 Announce Type: new Abstract: Traditional large-scale formation planning either oversimplify the formation representation which leads to poor performance, or they employ complete col

agentsarxiv-cs-ro
29 May 2026
Agents

Formalizing Mathematics at Scale

DGX agent

arXiv:2605.29955v1 Announce Type: new Abstract: We present AutoformBot, a multi-agent system for building an Autoformalized Textbook Library At Scale (Atlas) in Lean 4. AutoformBot orchestrates thousa

agentsarxiv-cs-ai
29 May 2026
Agents

FRUC: Feedforward Dynamic Scene Reconstruction from Uncalibrated Collaborative Driving Views

DGX agent

arXiv:2605.29997v1 Announce Type: new Abstract: We present FRUC, a feed-forward 3D Gaussian splatting framework for dynamic scene reconstruction from uncalibrated collaborative driving views. Existing

agentsarxiv-cs-cv
29 May 2026
Model Releases

Indexing the Unreadable: LLM-Native Recursive Construction and Search of Service Taxonomies

DGX agent

arXiv:2605.29270v1 Announce Type: new Abstract: The era of the Internet of Agents (IoA) is taking shape: LLM agents are expected to fulfill user goals by orchestrating fast-growing populations of Mode

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

LogDx-CI: Benchmarking Log Reduction Tools for LLM Root-Cause Diagnosis

DGX agent

arXiv:2605.28876v1 Announce Type: cross Abstract: CI failure logs are large (median 5k lines, max 200k in this corpus) and noisy. Coding agents that try to debug them depend on an upstream tool to red

model-releasesarxiv-cs-ai
29 May 2026
Agents

Our users love @StepFun_ai models and this new release packs a punch at a small size. Looking forward to seeing how well it works with Herme…

DGX agent

Our users love @StepFun_ai models and this new release packs a punch at a small size. Looking forward to seeing how well it works with Hermes Agent! ⚡️ Step 3.7 Flash is here: The new frontier is agen

agentsnous-research--x
29 May 2026
Agents

Production traffic from frontier models is a golden data asset. If you can efficiently mine the traces, filter for quality, and fine-tune sm…

DGX agent

Production traffic from frontier models is a golden data asset. If you can efficiently mine the traces, filter for quality, and fine-tune smaller models on them, you get specialized performance at a f

agentsharrison-chase--x
29 May 2026
← Previous
1…195196197198199…375
Next →