AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

agents

GridTimelineEvolution
7,201 results
11 Jun 2026

SkillJuror: Measuring How Agent Skill Organization Changes Runtime Behavior

AgentsDGX agent

arXiv:2606.11543v1 Announce Type: new Abstract: Agent Skills augment large language model (LLM) agents with procedural knowledge at inference time, but current benchmarks rarely distinguish what a Ski

StatefulDiscovery: Evidence-Calibrated Claim Formation in Open-Ended Scientific Discovery

AgentsDGX agent

arXiv:2606.11851v1 Announce Type: new Abstract: Open-ended scientific discovery asks agents to move beyond executing analyses for predefined questions. Across multiple rounds of exploration, a discove

Sustainability assessment using multimodal AI agents

AgentsDGX agent

arXiv:2507.17012v2 Announce Type: replace Abstract: Reducing the rapidly growing environmental impact of the computing industry requires assessing the emissions of electronics at scale. However, a tra


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

The Impossibility of Eliciting Latent Knowledge

AgentsDGX agent

arXiv:2606.12268v1 Announce Type: new Abstract: Advanced AI systems have extensive knowledge of their environments; in fact, their knowledge may (far) exceed that of their developers or users. Consequ

The Periodic Table of LLM Reasoning: A Structured Survey of Reasoning Paradigms, Methods, and Failure Modes

AgentsDGX agent

arXiv:2606.11470v1 Announce Type: new Abstract: Large Language Models (LLMs) have achieved strong performance across natural language processing tasks, yet reliable reasoning remains an open challenge

Towards Responsibly Non-Compliant Machines

AgentsDGX agent

arXiv:2606.12147v1 Announce Type: new Abstract: We consider the problem of engineering autonomous intelligent agents that are capable to responsibly not comply with user requests. We argue that machin

TreeSeeker: Tree-Structured Trial, Error, and Return in Deep Search

AgentsDGX agent

arXiv:2606.11662v1 Announce Type: new Abstract: Deep search requires agents to answer complex questions through multi-step web search, browsing, evidence comparison, and synthesis. A central challenge

When More Documents Hurt RAG: Mitigating Vector Search Dilution with Domain-Scoped, Model-Agnostic Retrieval

AgentsDGX agent

arXiv:2606.11350v1 Announce Type: new Abstract: Retrieval-augmented generation degrades when scaled to large, heterogeneous document collections, where dense similarity loses discriminative power, and

10 Jun 2026

1. Open the Profiles page from the sidebar 2. Click the Build button in the page header (next to the quick Create button) 3. That takes you …

AgentsDGX agent

1. Open the Profiles page from the sidebar 2. Click the Build button in the page header (next to the quick Create button) 3. That takes you to the /profiles/new builder flow: Identity → Model → Skills

3D-CoS: A New 3D Reconstruction Paradigm Based on VLM Code Synthesis

AgentsDGX agent

arXiv:2606.10478v1 Announce Type: new Abstract: Most recent 3D reconstruction and editing systems operate on implicit and explicit representations such as NeRF, point clouds, or meshes. While these re

A Distributed Multi-UGV Exploration Framework With Loop-Aware Planning and Descriptor-Aided Localization in Resource-Limited Environments

AgentsDGX agent

arXiv:2606.11088v1 Announce Type: new Abstract: Robust and efficient cooperative exploration with multiple unmanned ground vehicles (UGVs) in unknown, GPSdenied, and bandwidth-limited environments wit

A Survey on Semantic Modeling for Building Energy Management

AgentsDGX agent

arXiv:2404.11716v2 Announce Type: replace Abstract: Building Energy Management (BEM) is central to reducing energy use and CO2 emissions in the building sector. Although IoT technologies now provide e

ActiveMem: Distributed Active Memory for Long-Horizon LLM Reasoning

AgentsDGX agent

arXiv:2606.10532v1 Announce Type: new Abstract: Memory is essential for enabling large language model (LLM) agents to handle long-horizon reasoning tasks. Existing memory mechanisms are largely centra

Agentic Social Affordance Framework (ASAF): Agent Identity Design as a Collaboration Interface in Multi-Agent Systems

AgentsDGX agent

arXiv:2606.09832v1 Announce Type: cross Abstract: As AI systems evolve from single conversational agents to complex multi-agent architectures, a critical design dimension has been overlooked: how the

AI agents have become surprisingly good at talking to servers 🤖 The next challenge is getting them to interact with the environment users a…

AgentsDGX agent

AI agents have become surprisingly good at talking to servers 🤖 The next challenge is getting them to interact with the environment users actually live in. 🧑‍💻 Browsers. 🖥️ Apps. 📱 Devices. 🌐 Local st

An Exposure-Time-Aligned Primary-Path Architecture for Autonomous-Driving ECUs

AgentsDGX agent

arXiv:2606.10856v1 Announce Type: new Abstract: While end-to-end (E2E) autonomous driving has become the dominant research direction, production vehicles continue to rely on modular multi-NN pipelines

Are you going to be in Chicago on June 22? Do you want to talk about deepagents? If so - come by our LangChain x @focused_dot_io meetup! htt…

AgentsDGX agent

Harrison Chase announced a LangChain meetup event in Chicago on June 22 featuring discussion of deepagents, in collaboration with focused.io. The event was promoted on X (formerly Twitter) as an oppor

AutoPDE: Reliable Agentic PDE Solving via Explicitly Represented Solver Strategies

AgentsDGX agent

arXiv:2606.10752v1 Announce Type: new Abstract: Numerical solvers for partial differential equations (PDEs) are core computational tools in science and engineering. Building reliable PDE solvers requi

Beyond Static Evaluation: Co-Evolutionary Mechanisms for LLM-Driven Strategy Evolution in Adversarial Games

AgentsDGX agent

arXiv:2606.10389v1 Announce Type: new Abstract: Recent advances in LLM-driven code evolution have enabled automated discovery by iteratively generating and improving programs. However, applying these

Bridging Semantics and Physical Execution: A Neuro-Symbolic Framework for Multi-Pair Robotic Assembly

AgentsDGX agent

arXiv:2606.10808v1 Announce Type: new Abstract: Multi-pair robotic assembly in unstructured environments faces spatial interference and contact uncertainties. Existing paradigms fail to bridge cogniti

Business World Model

AgentsDGX agent

arXiv:2606.10044v1 Announce Type: new Abstract: Businesses are increasingly adopting AI-enabled tools to improve productivity, reduce costs, and enhance products and services. However, the transformat

Catching One in Five: LLM-as-Judge Blind Spots in Production Multi-Turn Transaction Agents

AgentsDGX agent

arXiv:2606.10315v1 Announce Type: cross Abstract: LLM-as-judge is the default instrument for evaluating conversational agents, yet its reliability is almost always reported as agreement with human rat

ChartAgent: A Multimodal Agent for Visually Grounded Reasoning in Complex Chart Question Answering

AgentsDGX agent

arXiv:2510.04514v3 Announce Type: replace Abstract: Recent multimodal LLMs have shown promise in chart-based visual question answering, but their performance declines sharply on unannotated charts-tho

Choosing your surface: Antigravity 2.0, Antigravity CLI, Antigravity IDE, or Antigravity SDK

AgentsDGX agent

TL;DR: Antigravity 2.0: A desktop app to orchestrate multiple autonomous agents working in parallel across independent projects. Antigravity CLI: A terminal interface designed for command-line workflo

Cursor’s code review agent is now over 3x faster, 22% cheaper, and finds 10% more bugs. You can also use /review to run Bugbot locally to ca…

AgentsDGX agent

Cursor has released performance improvements to its code review agent, achieving 3x faster speed, 22% cost reduction, and 10% increased bug detection rates. The update introduces a /review command tha

Data Journalist Agent: Transforming Data into Verifiable Multimodal Stories

AgentsDGX agent

arXiv:2606.11176v1 Announce Type: cross Abstract: Data tells stories that shape society; the data journalist's job is to turn raw information into stories non-experts can trust. A high-quality news fe

Deep dive: How Lightning Engine delivers 4.9x faster Apache Spark performance

AgentsDGX agent

From foundational ETL and analytics to the frontier of generative AI, Apache Spark serves as the architectural backbone for global data processing. However, as data volumes scale, the trade-off betwee

Dmsh: A Multi-Agent Reinforcement Learning Framework for All-Quad Mesh Generation

AgentsDGX agent

arXiv:2606.10601v1 Announce Type: cross Abstract: Generating high-quality meshes for arbitrary geometries remains a fundamental bottleneck in computational engineering, often demanding heuristic tunin

Envision4D: Envisioning Visual Futures via Feed-forward 4D Gaussian Splatting for Autonomous Driving

AgentsDGX agent

arXiv:2606.10656v1 Announce Type: new Abstract: Forecasting the future evolution of dynamic scenes is crucial in autonomous driving. However, existing feed-forward paradigms are primarily designed for

EstRTL: Functional Estimation Guided RTL Code Generation

AgentsDGX agent

arXiv:2606.09867v1 Announce Type: cross Abstract: Optimizing register transfer level (RTL) code is of vital importance in hardware design. Large language models (LLMs) provide new methods for the auto

Exclusive: MotherDuck adds agentic data ingestion to its cloud analytics service

AgentsDGX agent

MotherDuck Corp., the maker of a cloud-native data warehouse based on the open-source DuckDB analytical engine, is betting that artificial intelligence agents will reshape how data pipelines are built

Exclusive: Relai raises $6.9M to enable verifiable and continuous learning for AI agents

AgentsDGX agent

Artificial intelligence infrastructure startup Relai Inc. said today it has closed on 6.9 million in funding as it bids to ensure the reliability of autonomous AI agents for enterprises. The company a

FinOps AI goes beyond token economics as agentic costs emerge

AgentsDGX agent

As FinOps AI strategies continue to emerge,the familiar cloud cost management approach is breaking down — and organizations that fail to adapt risk runaway spending on workloads they barely understand

first in a series of technical blogs of how we build llm infra

AgentsDGX agent

first in a series of technical blogs of how we build llm infra How do you support full-text search JSON filtering over agent traces that span up to hundreds of MBs, while keeping a median (P50) latenc

From Confident Closing to Silent Failure: Characterizing False Success in LLM Agents

AgentsDGX agent

arXiv:2606.09863v1 Announce Type: new Abstract: LLM agents can fail silently by asserting task completion when the environment state shows otherwise. We study this failure mode, false success, across

future of travel!

AgentsDGX agent

future of travel! We raised $6M led by Sequoia to build the future of travel. Watch me plan a perfect trip to Mexico City in 3 minutes. Flights, hotels and a full itinerary that matches my preferences

Goal-oriented Communication for Fast and Robust Robotic Fault Detection and Recovery

AgentsDGX agent

arXiv:2601.18765v2 Announce Type: replace Abstract: Autonomous robotic systems are widely deployed in smart factories and operate in dynamic, uncertain, and human-involved environments that require lo

Harnessing the Collective Intelligence of AI Agents in the Wild for New Discoveries

AgentsDGX agent

arXiv:2606.10402v1 Announce Type: cross Abstract: Scientific discovery is often a collective process: researchers share partial results, inspect failed attempts, and build on each other's ideas over l

HIPIF: Hierarchical Planning and Information Folding for Long-Horizon LLM Agent Learning

AgentsDGX agent

arXiv:2606.10507v1 Announce Type: new Abstract: While Large Language Models (LLMs) have demonstrated strong capabilities as autonomous agents across a wide range of tasks, their performance often degr

How do you support full-text search JSON filtering over agent traces that span up to hundreds of MBs, while keeping a median (P50) latency o…

AgentsDGX agent

How do you support full-text search JSON filtering over agent traces that span up to hundreds of MBs, while keeping a median (P50) latency of 400ms? Here’s an inside look at how we built a custom inve

i showcase 'controlled' self improvement with a novel regime-to-seam approach where failures are categorized and allowed to fix targeted are…

AgentsDGX agent

i showcase 'controlled' self improvement with a novel regime-to-seam approach where failures are categorized and allowed to fix targeted areas of the agent while interesting, it's more to showcase the

Impatient Users Confuse AI Agents: High-fidelity Simulations of Human Traits for Testing Agents

AgentsDGX agent

arXiv:2510.04491v3 Announce Type: replace Abstract: Despite rapid progress in building conversational AI agents, robustness is still largely untested. Small shifts in user behavior, such as being more

in arxiv paper #2, i tackle the last topic from paper #1: @activegraphai as an architectural affordance for self-improving agents 'Regimes: …

AgentsDGX agent

in arxiv paper #2, i tackle the last topic from paper #1: @activegraphai as an architectural affordance for self-improving agents 'Regimes: An Auditable, Held-Out Gated Improvement Loop Demonstrated o

Infini Memory: Maintainable Topic Documents for Long-Term LLM Agent Memory

AgentsDGX agent

arXiv:2606.10677v1 Announce Type: new Abstract: Long-term LLM agents need persistent memory that can track changing facts and provide relevant evidence across sessions. Existing memory systems often s

Introducing @PoeticHQ: a new AI system that executes complex multi-hour tasks with 99%+ accuracy and 10x fewer tokens than agents. We raised…

AgentsDGX agent

Introducing @PoeticHQ: a new AI system that executes complex multi-hour tasks with 99%+ accuracy and 10x fewer tokens than agents. We raised 50M at 500M from Kleiner Perkins, Founders Fund, First Harm

Introducing the Hermes Agent Profile Builder You can now build a complete profile in the dashboard with full control over identity/name/desc…

AgentsDGX agent

Introducing the Hermes Agent Profile Builder You can now build a complete profile in the dashboard with full control over identity/name/description, model/provider, built-in + optional skills, skills-

Introducing Write Gate in Hermes Agent. Now you have the capability to be able to approve/deny memory updates, skill updates, and skill crea…

AgentsDGX agent

Introducing Write Gate in Hermes Agent. Now you have the capability to be able to approve/deny memory updates, skill updates, and skill creation with the same familiar mechanisms as approving dangerou

it’s actually so cool to work at LangChain the…database company (??) yup, the cracked team that built SmithDB is doing a cool blog series on…

AgentsDGX agent

it’s actually so cool to work at LangChain the…database company (??) yup, the cracked team that built SmithDB is doing a cool blog series on “How to build the internals of a database” —> for agent sca

Language-Driven Cost Optimization for Autonomous Driving

AgentsDGX agent

arXiv:2606.10974v1 Announce Type: new Abstract: The driving behavior of autonomous vehicles is typically governed by the cost function of their motion planner, which encodes objectives such as speed t

less novel, but still very interesting impo is the gated approach to self-modification the agent basically forks itself, propose a patch, ru…

AgentsDGX agent

less novel, but still very interesting impo is the gated approach to self-modification the agent basically forks itself, propose a patch, run through multiple tests (static/sandbox/diff), and somethin

Lium raises $5.5M to unlock complex scientific data for AI models

AgentsDGX agent

Lium, a startup formerly known as Astromind, today announced the launch of an “agentic harness” that helps large language models dig into the most complex and messiest datasets. The launch comes after

Lots of people asked how I used Fable to edit its own launch video so I made a video about that! TLDR it wrote a lot of code & tool calls to…

AgentsDGX agent

Lots of people asked how I used Fable to edit its own launch video so I made a video about that! TLDR it wrote a lot of code & tool calls to use transcription services, ffmpeg, do colorgrading, use th

Mobility Anomaly Generation using LLM-Driven Behavior with Kinematic Constraints

AgentsDGX agent

arXiv:2606.10314v1 Announce Type: new Abstract: Although the study of human trajectory anomalies is critical for advancing spatial data mining, empirical research remains severely hindered by a pervas

Most AI agents reset every session. @JenovaAIAgent's don't. Longest session on their platform: 16M tokens. All of it retrievable in <10ms vi…

AgentsDGX agent

Most AI agents reset every session. @JenovaAIAgent's don't. Longest session on their platform: 16M tokens. All of it retrievable in <10ms via Pinecone vector retrieval. Result: Fast ramp to $1M+ ARR,

my weekend hobby: self improvement research

AgentsDGX agent

my weekend hobby: self improvement research in arxiv paper #2, i tackle the last topic from paper #1: @activegraphai as an architectural affordance for self-improving agents 'Regimes: An Auditable, He

ObjSplat: Geometry-Aware Gaussian Surfels for Active Object Reconstruction

AgentsDGX agent

arXiv:2601.06997v2 Announce Type: replace-cross Abstract: Autonomous high-fidelity object reconstruction is fundamental for creating digital assets and bridging the simulation-to-reality gap in roboti

paper #1 for context: https://x.com/yoheinakajima/status/2057812713045377055?s=20

AgentsDGX agent

paper #1 for context: https://x.com/yoheinakajima/status/2057812713045377055?s=20 babyagi has ~200 citations, but 0 papers... i just published my first paper on arXiv 😆 'The Log is the Agent: Event-So

Planar-Sector LOS Guidance for Interception of Agile Targets with Lifting-Wing Quadcopters

AgentsDGX agent

arXiv:2606.10639v1 Announce Type: new Abstract: Autonomous visual interception of agile aerial targets is challenging due to unpredictable target motion, limited sensing, and the strong coupling betwe

Pushing the Limits of LLM Tool Calling via Experiential Knowledge Integration and Activation

AgentsDGX agent

arXiv:2606.10875v1 Announce Type: new Abstract: Large language models (LLMs) rely on tool use to act as autonomous agents, yet often fail in multi-step execution due to insufficient tool-related knowl

Pushing the Performance Limits in Autonomous Racing: Continuous Stability-Aware Adaptive Velocity Planning in Formula Student Driverless

AgentsDGX agent

arXiv:2606.10733v1 Announce Type: new Abstract: In autonomous racing, especially in competitions such as Formula Student Driverless, precise planning of the target velocity of a race car is crucial fo

← Previous
1…4041424344…121
Next →