AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,958 results
26 May 2026

Drift-Resistant Navigation World Model with Anchored Epipolar Guidance

AgentsDGX agent

arXiv:2605.24761v1 Announce Type: cross Abstract: We propose Drift-Resistant Navigation World Model, a generative model that mitigates both perceptual drift and geometric drift in conventional rollout

From Accuracy to Auditability: A Survey of Determinism in Financial AI Systems

AgentsDGX agent

arXiv:2605.23955v1 Announce Type: new Abstract: Deploying machine learning in regulated financial environments -- credit risk, fraud detection, and anti-money laundering -- exposes critical vulnerabil

GRAIL: AI translation for scientists application workflow on satellite data

AgentsDGX agent

arXiv:2605.24784v1 Announce Type: new Abstract: Domain scientists increasingly develop Python scripts to analyze satellite imagery but they lack scalability to large-scale data. This paper demonstrate

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

How we evolved Google’s global and data center networks for the AI era

Model ReleasesDGX agent

Over the last 25 years of building Google’s global network, we’ve navigated major architectural eras — from the Internet, to streaming, and the cloud. Today, we are squarely in the midst of a fourth:

Machine Psychometrics: A Mathematical Psychology of Artificial Intelligence

SafetyDGX agent

arXiv:2605.23952v1 Announce Type: new Abstract: Artificial agents now generate behavior rich enough to invite trust, surprise, and concern, yet our evaluation tools still privilege capability scores o

Multi-market value-stacking: Battery control for combined imbalance participation and non-uniform FCR bidding

AgentsDGX agent

arXiv:2605.23964v1 Announce Type: cross Abstract: The growing share of Renewable Energy Sources (RES) in modern power systems increases both grid imbalances and frequency deviations, reinforcing the n

Persuasion Should be Double-Blind: A Multi-Domain Dialogue Dataset With Faithfulness Based on Causal Theory of Mind

AgentsDGX agent

arXiv:2502.21297v2 Announce Type: replace Abstract: Persuasive dialogue is central to human communication, yet existing datasets often rely on a single language model generating both roles, producing

25 May 2026

IntentionNav: A Benchmark for Intent-Driven Object Navigation from Implicit Human Instruction

Model ReleasesDGX agent

arXiv:2605.23187v1 Announce Type: new Abstract: Existing object navigation benchmarks usually tell an embodied agent which object category to find, such as microwave or chair. Human-facing embodied AI

LACY: A Vision-Language Model-based Language-Action Cycle for Self-Improving Robotic Manipulation

AgentsDGX agent

arXiv:2511.02239v2 Announce Type: replace-cross Abstract: Learning generalizable policies for robotic manipulation increasingly relies on large-scale models that map language instructions to actions (

SciAtlas: A Large-Scale Knowledge Graph for Automated Scientific Research

Model ReleasesDGX agent

arXiv:2605.22878v1 Announce Type: new Abstract: The exponential growth of global academic output has confronted researchers and AI agents with an unprecedented ``information explosion,'' where fragmen

23 May 2026

Abstraction for Offline Goal-Conditioned Reinforcement Learning

AgentsDGX agent

arXiv:2605.22711v1 Announce Type: new Abstract: Markov Decision Processes (MDPs) often exhibit significant redundancy due to symmetries and shared structure across state-goal pairs in real-world Goal-

Harnesses for Inference-Time Alignment over Execution Trajectories

SafetyDGX agent

arXiv:2605.21516v1 Announce Type: new Abstract: Harness engineering has emerged as an important inference-time technique for large language model (LLM) agents, aiming to improve long-term performance

Skill Weaving: Efficient LLM Improvement via Modular Skillpacks

AgentsDGX agent

arXiv:2605.22205v1 Announce Type: cross Abstract: Large language models increasingly require specialization across diverse domains, yet existing approaches struggle to balance multi-domain capacities

22 May 2026

Flying Together: Human-Guided Immersive Shared Control for Aerial Robot Teams in Unknown Environments

AgentsDGX agent

arXiv:2605.21680v1 Announce Type: new Abstract: While autonomous multi-robots can achieve safe and coordinated navigation, they often struggle to adapt to unforeseen conditions and to capture operator

GA-VLN: Geometry-Aware BEV Representation for Efficient Vision-Language Navigation

AgentsDGX agent

arXiv:2605.22036v1 Announce Type: new Abstract: Despite significant progress in Vision-Language Navigation (VLN), existing approaches still rely on dense RGB videos that produce excessive patch tokens

If enterprise data is the oil, Dell wants disaggregated infrastructure to be the pipeline

AgentsDGX agent

Disaggregated infrastructure is displacing the all-in-one data center stack — and the economics of AI are making that reality impossible to ignore. The vast majority of enterprise data still resides o

LIDSA: Cognitive Arbitration for Signal-Free Autonomous Intersection Management

AgentsDGX agent

arXiv:2605.12321v2 Announce Type: replace Abstract: Large language models (LLMs) show strong potential for Intelligent Transportation Systems (ITS), particularly in tasks requiring situational reasoni

Lower Bounds for Advection-Diffusion Equations: An Exploration with AI-Generated Proofs

AgentsDGX agent

arXiv:2605.20623v1 Announce Type: cross Abstract: We establish explicit lower bounds for advection-diffusion equations in three settings: a polynomial ot H^{-1} bound for inviscid shears with uin L^in

NaviAgent: Graph-Driven Bilevel Planning for Scalable Tool Orchestration

SafetyDGX agent

arXiv:2506.19500v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) increasingly act as function-call agents that invoke external tools to tackle tasks beyond their static knowledge

Stdlib or Third-Party? Empirical Performance and Correctness of LLM-Assisted Zero-Dependency Python Libraries

AgentsDGX agent

arXiv:2605.21405v1 Announce Type: cross Abstract: Third-party Python libraries introduce dependency management overhead, supply chain risk, and deployment friction in constrained environments. A natur

21 May 2026

AI-Powered Facial Mask Removal Is Not Suitable For Identification

AgentsDGX agent

arXiv:2603.27747v2 Announce Type: replace Abstract: Recently, crowd-sourced online criminal investigations have used generative-AI to enhance low-quality visual evidence. In one high-profile case, soc

Beyond Words: Multimodal LLM Knows When to Speak

AgentsDGX agent

arXiv:2505.14654v2 Announce Type: replace-cross Abstract: Chatbots via large language models (LLMs) generate fluent responses but often struggle with when to speak, especially for brief, timely listen

Code Generation by Differential Test Time Scaling

AgentsDGX agent

arXiv:2605.20473v1 Announce Type: cross Abstract: Test-time scaling has emerged as a promising approach for improving code generation by exploring large solution spaces at inference time. However, exi

Compositional Transduction with Latent Analogies for Offline Goal-Conditioned Reinforcement Learning

AgentsDGX agent

arXiv:2605.20609v1 Announce Type: new Abstract: Compositional generalization is essential for reaching unseen goals under novel contextual variations in offline goal-conditioned reinforcement learning

Decoupling Communication from Policy: Robust MARL under Bandwidth Constraints

SafetyDGX agent

arXiv:2605.21085v1 Announce Type: cross Abstract: Communication enables coordination in multi-agent reinforcement learning (MARL), but many real-world applications, e.g., search-and-rescue with drone

Draw2Think: Harnessing Geometry Reasoning through Constraint Engine Interaction

AgentsDGX agent

arXiv:2605.20743v1 Announce Type: cross Abstract: Vision-language models solve geometry problems with rising accuracy, yet their intermediate states remain latent and unverifiable: a relation expresse

i feel like there's a general misunderstanding about open source models. most people use a frontier model, switch the api request to open so…

AgentsDGX agent

i feel like there's a general misunderstanding about open source models. most people use a frontier model, switch the api request to open source model, see poor performance, and then churn off. this w

if you work across multiple machines, highly recommend using Grok Build with its subagents to manage SSH tunnels and interact with tmux. We …

AgentsDGX agent

if you work across multiple machines, highly recommend using Grok Build with its subagents to manage SSH tunnels and interact with tmux. We are working on making this experience more native, think of

MC-Risk: Multi-Component Risk Fields for Risk Identification and Motion Planning

AgentsDGX agent

arXiv:2605.21406v1 Announce Type: new Abstract: We present MC-Risk, a planner-aligned, multi-component risk field on a bird's-eye-view grid that yields early, calibrated, and class-aware risk localiza

Our database and data engineering expert @yoniebans made some major improvements to the way sessions are stored and accessed. This will save…

AgentsDGX agent

Our database and data engineering expert @yoniebans made some major improvements to the way sessions are stored and accessed. This will save something like 20-40% of the disk space used by Hermes Agen

Paris-based Pivot, which develops AI tools for procurement and financial workflows, raised a $40M Series B co-led by Forestay Capital and Notion Capital (Tamara Djurickovic/Tech.eu)

AgentsDGX agent

Tamara Djurickovic / Tech.eu: Paris-based Pivot, which develops AI tools for procurement and financial workflows, raised a $40M Series B co-led by Forestay Capital and Notion Capital — With new fundin

Retrieval-Augmented Code Generation: A Survey with Focus on Repository-Level Approaches

AgentsDGX agent

arXiv:2510.04905v3 Announce Type: replace-cross Abstract: Recent advances in large language models (LLMs) have significantly improved automated code generation. While existing approaches have achieved

SubTGraph: Large-Scale Subterranean Environment Synthesis with Controllable Topological Variability for Robotic Autonomy Validation

AgentsDGX agent

arXiv:2605.20917v1 Announce Type: new Abstract: Subterranean (SubT) environments have been a frontier for autonomous robotics, driven by the push for automation of mining operations and the interest i

Three insights you may have missed from theCUBE’s coverage of the DigiCert Trust Summit

AgentsDGX agent

AI is turning digital trust from a security function into an operating model. That shift is putting new pressure on the systems enterprises have long used to verify identity, protect data and keep dig

Why Latent Actions Fail, and How to Prevent It

AgentsDGX agent

arXiv:2605.20223v1 Announce Type: new Abstract: Latent action models (LAMs) aim to learn action-like representations from unlabeled videos by compressing frame-to-frame changes. The frames of in-the-w

20 May 2026

A Geometric Analysis of Small-sized Language Model Hallucinations

AgentsDGX agent

arXiv:2602.14778v3 Announce Type: replace-cross Abstract: Hallucinations -- plausible but factually incorrect responses -- pose a major challenge to the reliability of Large Language Models (LLMs), es

A Logistic Regression Model to Predict Malaria Severity in Children

AgentsDGX agent

arXiv:2605.18900v1 Announce Type: cross Abstract: One of the main causes of death around the globe is malaria. Researchers have sought to develop predictive models for malaria outbreaks based on meteo

Active Graph really feels like the culmination of all of my BabyAGI and graph experiments. [fyi, technical history of babyagi: http://babyag…

AgentsDGX agent

Active Graph really feels like the culmination of all of my BabyAGI and graph experiments. [fyi, technical history of babyagi: http://babyagi.wiki] would love to hear thoughts if you try it out! the e

Causal Evidence that Language Models use Confidence to Drive Behavior

AgentsDGX agent

arXiv:2603.22161v2 Announce Type: replace Abstract: Metacognition -- assessing the quality of one's own cognitive performance -- guides adaptive behavior across species. Substantial research demonstra

CLUE: Adaptively Prioritized Contextual Cues by Leveraging a Unified Semantic Map for Effective Zero-Shot Object-Goal Navigation

AgentsDGX agent

arXiv:2605.19206v1 Announce Type: new Abstract: Zero-shot object-goal navigation (ZSON) is a challenging problem in robotics that requires a comprehensive understanding of both language and visual obs

DECOR: Auditing LLM Deception via Information Manipulation Theory

AgentsDGX agent

arXiv:2605.19270v1 Announce Type: new Abstract: Large language models can deceive by subtly manipulating truthful information -- omitting key facts, shifting focus, or obscuring meaning -- making such

Dual-Gated Epistemic Time-Dilation: Autonomous Compute Modulation in Asynchronous MARL

SafetyDGX agent

arXiv:2603.23722v2 Announce Type: replace-cross Abstract: While Multi-Agent Reinforcement Learning (MARL) algorithms achieve unprecedented successes across complex continuous domains, their standard d

ESLD (External Surrogate Latent Defense): A Latent-Space Architecture for Faster, Stronger Prompt-Injection Defense

SafetyDGX agent

arXiv:2605.18918v1 Announce Type: cross Abstract: Modern AI assistants are agentic. To answer a single user request, the underlying language model pulls in information from many sources, such as web s

FAGER: Factually Grounded Evaluation and Refinement of Text-to-Image Models

AgentsDGX agent

arXiv:2605.19111v1 Announce Type: cross Abstract: Existing text-to-image (T2I) evaluation metrics mainly assess whether generated images align with information explicitly stated in the prompt, but oft

High-quality generation of dynamic game content via small language models: A proof of concept

AgentsDGX agent

arXiv:2601.23206v2 Announce Type: replace Abstract: Large language models (LLMs) offer promise for dynamic game content generation, but they face critical barriers, including narrative incoherence and

Hybrid Training for Vision-Language-Action Models

AgentsDGX agent

arXiv:2510.00600v2 Announce Type: replace-cross Abstract: Using Large Language Models to produce intermediate thoughts, a.k.a. Chain-of-thought (CoT), before providing an answer has been a successful

Library Drift: Diagnosing and Fixing a Silent Failure Mode in Self-Evolving LLM Skill Libraries

AgentsDGX agent

arXiv:2605.19576v1 Announce Type: new Abstract: Self-evolving skill libraries face a silent failure mode we term library drift: unbounded skill accumulation without outcome-driven lifecycle management

LLMs are stateless (every time you reply to an LLM, you re-inject the entire conversation to a fresh inference) the purpose of memory is to …

AgentsDGX agent

LLMs are stateless (every time you reply to an LLM, you re-inject the entire conversation to a fresh inference) the purpose of memory is to provide continuity (in games we'd call it a 'persistent worl

Operationalising Artificial Intelligence Bills of Materials (AIBOMs) for Verifiable AI Provenance and Lifecycle Assurance

AgentsDGX agent

arXiv:2605.19755v1 Announce Type: cross Abstract: Artificial Intelligence (AI) systems are increasingly dependent on complex, multi-layered software supply chains that introduce challenges for reprodu

@Replit narrative walkthrough video of repo for anyone interested: https://x.com/FileCityAI/status/2057164885780226139?s=20

AgentsDGX agent

@Replit narrative walkthrough video of repo for anyone interested: https://x.com/FileCityAI/status/2057164885780226139?s=20 FileCity Tour: activegraph An event-sourced reactive graph runtime for long-

Synthesis and Evaluation of Long-term History-aware Medical Dialogue

Model ReleasesDGX agent

arXiv:2605.19766v1 Announce Type: cross Abstract: An effective healthcare agent must be able to recall and reason over a patient's longitudinal medical history. However, the absence of datasets with r

The 99% Success Paradox: When Near-Perfect Retrieval Equals Random Selection

AgentsDGX agent

arXiv:2605.18857v1 Announce Type: cross Abstract: For most of the history of information retrieval (IR), search results were designed for human consumers who could scan, filter, and discard irrelevant

this is how you add an event, fork and cache a run, and then find the diff between a parent and fork in this example, the fork shares the pa…

AgentsDGX agent

this is how you add an event, fork and cache a run, and then find the diff between a parent and fork in this example, the fork shares the parent's event log up to event 142. from 143 onward it diverge

We're excited to be an official shoutout at the Google I/O Developer Keynote 🔥 @llama_index is building the document infrastructure for AI …

Model ReleasesDGX agent

We're excited to be an official shoutout at the Google I/O Developer Keynote 🔥 @llama_index is building the document infrastructure for AI agents, and we plan to integrate even more heavily with both

We’re hiring for Labs! 🧪 If you’re interested in working with us to push forward Continual Learning, pls DM me with a blurb + link to the b…

AgentsDGX agent

We’re hiring for Labs! 🧪 If you’re interested in working with us to push forward Continual Learning, pls DM me with a blurb + link to the best Applied Research you’ve done (or even better shipped!) yo

YAC: Bridging Natural Language and Interactive Visual Exploration with Generative AI for Biomedical Data Discovery

AgentsDGX agent

arXiv:2509.19182v2 Announce Type: replace-cross Abstract: Incorporating natural language input has the potential to improve the capabilities of biomedical data discovery interfaces. However, user inte

19 May 2026

A Mechanistic Model for Collective Motion from Sensorimotor Regularities

AgentsDGX agent

arXiv:2605.16522v1 Announce Type: new Abstract: Collective behavior in animals has long been modeled through self-propelled particle models, which reproduce striking group-level phenomena through abst

A Pilot Benchmark for NL-to-FOL Translation in Planetary Exploration

Model ReleasesDGX agent

arXiv:2605.17911v1 Announce Type: new Abstract: Future planetary exploration envisions autonomous robotic agents operating under severe communication constraints, without global positioning, and with

Baba in Wonderland: Online Self-Supervised Dynamics Discovery for Executable World Models

AgentsDGX agent

arXiv:2605.16725v1 Announce Type: new Abstract: Executable world models can be read, edited, executed, and reused for planning, but only if the program captures the environment's transition law rather

Causely: A Causal Intelligence Layer for Enterprise AI A Benchmark Study on SRE and Reliability Workflows

Model ReleasesDGX agent

arXiv:2605.18327v1 Announce Type: new Abstract: AI agents deployed into SRE workflows currently derive their understanding of environment state from raw observability telemetry at query time, paying a

← Previous
1…189190191192193…300
Next →