AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

agents

GridTimelineEvolution
7,215 results
25 May 2026

EVE-Agent: Evidence-Verifiable Self-Evolving Agents

AgentsDGX agent

arXiv:2605.22905v1 Announce Type: new Abstract: Self-evolving agents should not train on examples they cannot justify. Data-free self-evolving search agents offer a scalable route to systems that gene

ExpOS: Explainable Open-Surgery Skills Assessment Using 3D Hand Reconstruction

AgentsDGX agent

arXiv:2605.23653v1 Announce Type: new Abstract: Timely and transparent feedback is essential for effective surgical training, yet current assessment remains dependent on expert observation, limiting s

From Raw Experience to Skill Consumption: A Systematic Study of Model-Generated Agent Skills

AgentsDGX agent

arXiv:2605.23899v1 Announce Type: new Abstract: Language agents increasingly improve by reusing skills -- structured procedural artifacts distilled from past experience. In particular, domain-level an


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

From Residuals to Reasons: LLM-Guided Mechanism Inference from Tabular Data

AgentsDGX agent

arXiv:2605.22897v1 Announce Type: new Abstract: A persistent challenge in machine learning for scientific applications is jointly achieving prediction and understanding. Statistical models excel on st

FusionSense: Tri-Stage Near-Sensor Learning for Runtime-Adaptive Multimodal Edge Intelligence

AgentsDGX agent

arXiv:2605.22868v1 Announce Type: new Abstract: Autonomous systems and smart-industry deployments increasingly split computation across near-sensor, edge, and cloud resources, where tight energy, late

GFSR: Geometric Fidelity and Spatial Refinement for Reliable Lane Detection

AgentsDGX agent

arXiv:2605.23327v1 Announce Type: new Abstract: Lane detection stands as a crucial perception task in autonomous driving and advanced driver assistance systems. However, existing methods still degrade

/goal is really insane! It's how you can get the most out of coding agents today. For efficiency, I find it works best when you do planning …

AgentsDGX agent

/goal is really insane! It's how you can get the most out of coding agents today. For efficiency, I find it works best when you do planning before /goal. This ensures the agent has the right context a

GradingAttack: Exposing Security Vulnerabilities in LLM Based Educational Grading Agents

AgentsDGX agent

arXiv:2602.00979v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly deployed as educational agents for automatic short answer grading (ASAG) in real-world education

Harness, Scaffold, and the AI Agent Terms Worth Getting Right

AgentsDGX agent

This article defines and clarifies key terminology related to AI agents, including the concepts of 'harness' and 'scaffold,' which are important architectural and operational components in building an

(I'm firmly on team red/green TDD for agent code, I like having a test suite that protects against them breaking old features when they make…

AgentsDGX agent

(I'm firmly on team red/green TDD for agent code, I like having a test suite that protects against them breaking old features when they make new changes - https://simonwillison.net/guides/agentic-engi

Join the team on Wednesday for another Hermes Agent Jam!

AgentsDGX agent

Nous Research is inviting developers and enthusiasts to participate in a 'Hermes Agent Jam' event scheduled for Wednesday, likely a hackathon or collaborative workshop focused on building or improving

KPI2KVI: A Multi Agent Workflow for Calculating Key Value Indicators from Service Descriptions

AgentsDGX agent

arXiv:2605.22825v1 Announce Type: cross Abstract: Key Value Indicators (KVIs) provide a decision oriented view of a service by summarizing how operational performance translates into stakeholder value

LACY: A Vision-Language Model-based Language-Action Cycle for Self-Improving Robotic Manipulation

AgentsDGX agent

arXiv:2511.02239v2 Announce Type: replace-cross Abstract: Learning generalizable policies for robotic manipulation increasingly relies on large-scale models that map language instructions to actions (

Latent Cache Flow: Model-to-Model Communication Without Text

AgentsDGX agent

arXiv:2605.22863v1 Announce Type: new Abstract: LLM agents today communicate via text, which incurs considerable latency and information loss due to the need to autoregressively decode the sharer mode

MapGCLR: Geospatial Contrastive Learning of Representations for Online Vectorized HD Map Construction

AgentsDGX agent

arXiv:2603.10688v2 Announce Type: replace-cross Abstract: Autonomous vehicles rely on map information to understand the world around them. However, the creation and maintenance of offline high-definit

MARGIN: Runtime Confidence Calibration for Multi-Agent Foundation Model Coordination

AgentsDGX agent

arXiv:2605.22949v1 Announce Type: new Abstract: Foundation model agents increasingly operate in multi-agent deployments where a coordinator must decide which agent's response to trust. The standard ap

MemAudit: Post-hoc Auditing of Poisoned Agent Memory via Causal Attribution and Structural Anomaly Detection

AgentsDGX agent

arXiv:2605.23723v1 Announce Type: new Abstract: Large language model agents increasingly rely on persistent memory to store past interactions, retrieve relevant demonstrations, and improve long-horizo

Multi-Floor Exploration for Ground Robots via an Incremental Reachable Graph and Structural Priors

AgentsDGX agent

arXiv:2605.23350v1 Announce Type: new Abstract: Autonomous exploration of multi-floor buildings remains challenging for ground robots because conventional 2D and 2.5D maps cannot represent overlapping

NeuroWeaver: An Autonomous Evolutionary Agent for Exploring the Programmatic Space of EEG Analysis Pipelines

AgentsDGX agent

arXiv:2602.13473v2 Announce Type: replace Abstract: Although foundation models have demonstrated remarkable success in general domains, the application of these models to electroencephalography (EEG)

New paper from Microsoft on Self-Evolving Agent Skills

AgentsDGX agent

New paper from Microsoft on Self-Evolving Agent Skills New research from Microsoft Research I see a lot of AI engineers handwriting agent skill docs and hope they generalize. Probably not optimal. Thi

nice write up from the HuggingFace folks aggregating works on defining agents, harnesses, environments, RL, etc. The more we can roughly hav…

AgentsDGX agent

nice write up from the HuggingFace folks aggregating works on defining agents, harnesses, environments, RL, etc. The more we can roughly have a shared vocabulary the better…I still find it confusing (

Plugins can also affect the empty state of the new menu - the latest datasette-agent adds a form for kicking off a new agent conversation - …

AgentsDGX agent

Plugins can also affect the empty state of the new menu - the latest datasette-agent adds a form for kicking off a new agent conversation - live demo (if you sign in with GitHub) on https://agent.data

Query-Adaptive Semantic Chunking for Retrieval-Augmented Generation: A Dynamic Strategy with Contextual Window Expansion

AgentsDGX agent

arXiv:2605.22834v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) systems depend critically on document chunking quality for retrieving relevant context. Fixed chunking segments doc

Ralph Loops are powerful, but wrapping a naive loop in a shell script is a total token burner in production. Pinecone Principal Engineer Jen…

AgentsDGX agent

Ralph Loops are powerful, but wrapping a naive loop in a shell script is a total token burner in production. Pinecone Principal Engineer Jen Hamon breaks down why standard loops collapse: ❌ The Bug: P

Redrawing the AI Map: A Theory of Accountability Boundaries in Agentic Ecosystems

AgentsDGX agent

arXiv:2605.23179v1 Announce Type: new Abstract: Agentic AI orchestrators reduce the interface and assembly costs of composing information systems capabilities across organizational boundaries, seeming

RS2AD-LiDAR: End-to-End Autonomous Driving LiDAR Data Generation from Roadside Sensor Observations

AgentsDGX agent

arXiv:2605.23406v1 Announce Type: new Abstract: End-to-end autonomous driving solutions, which directly process multimodal sensory data and output fine-grained control commands, have gradually become

Scene Reconstruction as Mapping Priors for 3D Detection

AgentsDGX agent

arXiv:2605.22997v1 Announce Type: new Abstract: In autonomous driving, mapping is critical for motion planning but remains an under-utilized resource for perception tasks such as 3D object detection.

SFG-ROS: A Resource-Aware Framework for Dense Multi-Agent Perception

AgentsDGX agent

arXiv:2605.23832v1 Announce Type: new Abstract: Deploying heterogeneous multi-agent robot fleets for collaborative perception requires robust data exchange and scalable software architectures. However

Socially fluent AI decouples conversational signals from source identity in online interaction

AgentsDGX agent

arXiv:2605.23426v1 Announce Type: cross Abstract: Socially fluent agentic AI can now participate in online interaction in ways that resemble ordinary human conversation, potentially weakening people's

TABX: A High-Throughput Sandbox Battle Simulator for Multi-Agent Reinforcement Learning

AgentsDGX agent

arXiv:2602.01665v2 Announce Type: replace-cross Abstract: The design of environments plays a critical role in shaping the development and evaluation of cooperative multi-agent reinforcement learning (

Turning Adaptation into Assets: Cross-Domain Bridging for Online Vision-Language Navigation

AgentsDGX agent

arXiv:2605.23257v1 Announce Type: cross Abstract: Navigating under non-stationary environment shifts poses a critical challenge for a Vision-and-Language Navigation (VLN) agent deployed in the wild. Y

When Is Next-Token Prediction Useful? Marginalization, Ergodicity, Mixture Identifiability, Local Sufficiency, RAG, Tools, and Programming

AgentsDGX agent

arXiv:2605.23278v1 Announce Type: new Abstract: Language models trained on observed sequences are often described as learning the conditional distribution of the next token given previous tokens. This

When Planning Fails Despite Correct Execution: On Epistemic Calibration for LLM-Based Multi-Agent Systems

AgentsDGX agent

arXiv:2605.23414v1 Announce Type: new Abstract: LLM-based multi-agent systems can fail even when planned actions are executed correctly because agents may misjudge their knowledge when evaluating plan

X-TRACK: Physics-Aware xLSTM for Realistic Vehicle Trajectory Prediction

AgentsDGX agent

arXiv:2511.00266v2 Announce Type: replace Abstract: Accurate trajectory prediction is crucial for safe and reliable autonomous driving systems, requiring models that capture long-term temporal depende

24 May 2026

datasette-agent 0.1a4

AgentsDGX agent

Release: datasette-agent 0.1a4 Taking advantage of the new makeJumpSections() JavaScript plugin hook added in Datasette 1.0a30, datasette-agent now presents this 'Start a new agent chat' interface as

Grok Build sub-agent swarm weekend fun. You can reuse the prompt for your projects: Read the proof of `https://cdn.openai.com/pdf/74c24085-1…

AgentsDGX agent

Grok Build sub-agent swarm weekend fun. You can reuse the prompt for your projects: Read the proof of `https://cdn.openai.com/pdf/74c24085-19b0-4534-9c90-465b8e29ad73/unit-distance-proof.pdf` and come

I asked my eng team if I could ship code to prod They told me no 💀

AgentsDGX agent

I asked my eng team if I could ship code to prod They told me no 💀 CEOs are uniquely prone to AI psychosis because they’re sufficiently distant from the last mile of work that still has to happen to g

Please for the love of god don’t take this to heart Go out, have fun, make friends, touch some grass. You can work as hard as you want when …

AgentsDGX agent

Please for the love of god don’t take this to heart Go out, have fun, make friends, touch some grass. You can work as hard as you want when you’re back. kinda fascinating that “going out” became a kin

so many experiments I want to run… 😵‍💫

AgentsDGX agent

Yohei Nakajima expresses the overwhelm of having numerous experimental ideas he wants to pursue, reflecting on the challenge of prioritization and resource constraints in AI research and development.

The Top AI Papers of the Week (May 18 - 24): - AIRA - MetaCogAgent - Memory as a Model - Code as Agent Harness - Weak-Model Critic-Comparato…

AgentsDGX agent

The Top AI Papers of the Week (May 18 - 24): - AIRA - MetaCogAgent - Memory as a Model - Code as Agent Harness - Weak-Model Critic-Comparator - OpenAI Disproves the Unit Distance Conjecture - Producti

23 May 2026

Abstraction for Offline Goal-Conditioned Reinforcement Learning

AgentsDGX agent

arXiv:2605.22711v1 Announce Type: new Abstract: Markov Decision Processes (MDPs) often exhibit significant redundancy due to symmetries and shared structure across state-goal pairs in real-world Goal-

Active Graph is the best, most 'correct' knowledge/context engine I've come across so far (and I've tried or at least researched most of the…

AgentsDGX agent

Active Graph is the best, most 'correct' knowledge/context engine I've come across so far (and I've tried or at least researched most of them.) babyagi has ~200 citations, but 0 papers... i just publi

// Adapt the Interface, Not the Model // I am fascinated by the results across my cheap-model-plus-good-harness builds. This new paper also …

AgentsDGX agent

// Adapt the Interface, Not the Model // I am fascinated by the results across my cheap-model-plus-good-harness builds. This new paper also shows good signs of the code-as-agent-harness thesis. The id

Agents shouldn’t have direct visibility into env vars or credentials that can expose sensitive systems and data. Keeping secrets outside the…

AgentsDGX agent

Agents shouldn’t have direct visibility into env vars or credentials that can expose sensitive systems and data. Keeping secrets outside the agent’s context helps secure the env while still allowing a

[AINews] All Model Labs are now Agent Labs

AgentsDGX agent

Latent Space reports on a rebranding or reorganization where Model Labs have been renamed or converted into Agent Labs, reflecting a shift in focus toward AI agent development and capabilities. This c

Also applies to selling b2b saas

AgentsDGX agent

Also applies to selling b2b saas Spoke with a late-30s girlfriend in Dallas who went on a date last night with a late-50s bachelor from NYC. They were introduced via another matchmaker. She said the v

Am riding in an OG geofrenced autonomous vehicle, otherwise known as an airport monorail. 🚝

AgentsDGX agent

Gary Marcus humorously describes riding an airport monorail as an example of a geofenced autonomous vehicle, highlighting how existing transportation systems already operate with limited autonomy with

Beyond Scalar Objectives: Expert-Feedback-Driven Autonomous Experimentation for Scientific Discovery at the Nanoscale

AgentsDGX agent

arXiv:2605.21820v1 Announce Type: new Abstract: Self-driving laboratories or autonomous experimentation are emerging as transformative platforms for accelerating scientific discovery. Bayesian optimiz

Can frontier models forecast scientific progress? Mostly no, but here is why. This work looks at 4,760 scientific events across disciplines.…

AgentsDGX agent

Can frontier models forecast scientific progress? Mostly no, but here is why. This work looks at 4,760 scientific events across disciplines. Frontier models can identify plausible research directions

Every 'self-evolving agent' paper this year has mutated text: prompts, skill files, workflow graphs, memory schemas. MOSS from USTC & HKUST …

AgentsDGX agent

Every 'self-evolving agent' paper this year has mutated text: prompts, skill files, workflow graphs, memory schemas. MOSS from USTC & HKUST argues this is the wrong layer. The thing that actually brea

If you have massive volumes of documents you're looking to parse, come check us out: https://cloud.llamaindex.ai/?utm_source=xjl&utm_medium=…

AgentsDGX agent

LlamaIndex Cloud is a service designed to handle large-scale document parsing and processing, offering a cloud-based solution for organizations dealing with massive volumes of documents. The platform

If you're on 2nd street facing south and you're curious what we've been up to at @llama_index, this sign might help 🙂

AgentsDGX agent

Jerry Liu shared a photo of a sign on 2nd Street visible from a southward-facing direction that advertises or promotes LlamaIndex, suggesting the company has some physical presence or marketing initia

If you're stopping by the SF Caltrain station over Memorial Day weekend, you might catch a glimpse of our digital ads 📺 We parse (PDFs) (50…

AgentsDGX agent

Jerry Liu announced digital advertisements at the SF Caltrain station during Memorial Day weekend, mentioning the parsing of PDFs and a reference to '50' (likely indicating 50 PDFs or a related metric

“It is built in Rust and leverages the Apache DataFusion query engine” Any new database these days

AgentsDGX agent

“It is built in Rust and leverages the Apache DataFusion query engine” Any new database these days We built SmithDB: the database purpose built for agent observability workloads that now powers many p

it’s the weekend! if you want to try playing w ActiveGraph, just point your AI to http://docs.activegraph.ai and ask it build something :)

AgentsDGX agent

ActiveGraph is an AI tool or framework documented at http://docs.activegraph.ai that enables users to build applications by providing instructions to an AI system. The platform appears designed for we

Last I checked, the @llama_index OSS framework integrates with 78 (!!) vector stores. Investors (and builders starting out, and honestly me …

AgentsDGX agent

Last I checked, the @llama_index OSS framework integrates with 78 (!!) vector stores. Investors (and builders starting out, and honestly me for a while) deemed this space as commoditized. One of the l

last night i got an agent to fork itself, propose a modification to itself on the fork, run through tests (sandbox, etc), and only accept th…

AgentsDGX agent

Yohei Nakajima describes a technical demonstration where an AI agent was configured to create a fork of itself, propose modifications to the forked version, execute tests in a sandboxed environment, a

LCGuard: Latent Communication Guard for Safe KV Sharing in Multi-Agent Systems

AgentsDGX agent

arXiv:2605.22786v1 Announce Type: cross Abstract: Large language model (LLM)-based multi-agent systems increasingly rely on intermediate communication to coordinate complex tasks. While most existing

MOSS: Self-Evolution through Source-Level Rewriting in Autonomous Agent Systems

AgentsDGX agent

arXiv:2605.22794v1 Announce Type: cross Abstract: Autonomous agentic systems are largely static after deployment: they do not learn from user interactions, and recurring failures persist until the nex

optimise for fun when it comes to building for and with agents setting up a hermes agent is fun because it feels like you’re hatching and ra…

AgentsDGX agent

optimise for fun when it comes to building for and with agents setting up a hermes agent is fun because it feels like you’re hatching and raising an incredibly smart digital pet that can do and learn

← Previous
1…6263646566…121
Next →