AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,963 results
Agents

Stdlib or Third-Party? Empirical Performance and Correctness of LLM-Assisted Zero-Dependency Python Libraries

DGX agent

arXiv:2605.21405v1 Announce Type: cross Abstract: Third-party Python libraries introduce dependency management overhead, supply chain risk, and deployment friction in constrained environments. A natur

agentsarxiv-cs-ai
22 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Agents

AI-Powered Facial Mask Removal Is Not Suitable For Identification

DGX agent

arXiv:2603.27747v2 Announce Type: replace Abstract: Recently, crowd-sourced online criminal investigations have used generative-AI to enhance low-quality visual evidence. In one high-profile case, soc

agentsarxiv-cs-cv
21 May 2026
Agents

Beyond Words: Multimodal LLM Knows When to Speak

DGX agent

arXiv:2505.14654v2 Announce Type: replace-cross Abstract: Chatbots via large language models (LLMs) generate fluent responses but often struggle with when to speak, especially for brief, timely listen

agentsarxiv-cs-cl
21 May 2026
Agents

Code Generation by Differential Test Time Scaling

DGX agent

arXiv:2605.20473v1 Announce Type: cross Abstract: Test-time scaling has emerged as a promising approach for improving code generation by exploring large solution spaces at inference time. However, exi

agentsarxiv-cs-lg
21 May 2026
Agents

Compositional Transduction with Latent Analogies for Offline Goal-Conditioned Reinforcement Learning

DGX agent

arXiv:2605.20609v1 Announce Type: new Abstract: Compositional generalization is essential for reaching unseen goals under novel contextual variations in offline goal-conditioned reinforcement learning

agentsarxiv-cs-lg
21 May 2026
Safety

Decoupling Communication from Policy: Robust MARL under Bandwidth Constraints

DGX agent

arXiv:2605.21085v1 Announce Type: cross Abstract: Communication enables coordination in multi-agent reinforcement learning (MARL), but many real-world applications, e.g., search-and-rescue with drone

safetyarxiv-cs-lg
21 May 2026
Agents

Draw2Think: Harnessing Geometry Reasoning through Constraint Engine Interaction

DGX agent

arXiv:2605.20743v1 Announce Type: cross Abstract: Vision-language models solve geometry problems with rising accuracy, yet their intermediate states remain latent and unverifiable: a relation expresse

agentsarxiv-cs-cl
21 May 2026
Agents

i feel like there's a general misunderstanding about open source models. most people use a frontier model, switch the api request to open so…

DGX agent

i feel like there's a general misunderstanding about open source models. most people use a frontier model, switch the api request to open source model, see poor performance, and then churn off. this w

agentsharrison-chase--x
21 May 2026
Agents

if you work across multiple machines, highly recommend using Grok Build with its subagents to manage SSH tunnels and interact with tmux. We …

DGX agent

if you work across multiple machines, highly recommend using Grok Build with its subagents to manage SSH tunnels and interact with tmux. We are working on making this experience more native, think of

agentselon-musk--x
21 May 2026
Agents

MC-Risk: Multi-Component Risk Fields for Risk Identification and Motion Planning

DGX agent

arXiv:2605.21406v1 Announce Type: new Abstract: We present MC-Risk, a planner-aligned, multi-component risk field on a bird's-eye-view grid that yields early, calibrated, and class-aware risk localiza

agentsarxiv-cs-ro
21 May 2026
Agents

Our database and data engineering expert @yoniebans made some major improvements to the way sessions are stored and accessed. This will save…

DGX agent

Our database and data engineering expert @yoniebans made some major improvements to the way sessions are stored and accessed. This will save something like 20-40% of the disk space used by Hermes Agen

agentsnous-research--x
21 May 2026
Agents

Paris-based Pivot, which develops AI tools for procurement and financial workflows, raised a $40M Series B co-led by Forestay Capital and Notion Capital (Tamara Djurickovic/Tech.eu)

DGX agent

Tamara Djurickovic / Tech.eu: Paris-based Pivot, which develops AI tools for procurement and financial workflows, raised a $40M Series B co-led by Forestay Capital and Notion Capital — With new fundin

agentstechmeme
21 May 2026
Agents

Retrieval-Augmented Code Generation: A Survey with Focus on Repository-Level Approaches

DGX agent

arXiv:2510.04905v3 Announce Type: replace-cross Abstract: Recent advances in large language models (LLMs) have significantly improved automated code generation. While existing approaches have achieved

agentsarxiv-cs-cl
21 May 2026
Agents

SubTGraph: Large-Scale Subterranean Environment Synthesis with Controllable Topological Variability for Robotic Autonomy Validation

DGX agent

arXiv:2605.20917v1 Announce Type: new Abstract: Subterranean (SubT) environments have been a frontier for autonomous robotics, driven by the push for automation of mining operations and the interest i

agentsarxiv-cs-ro
21 May 2026
Agents

Three insights you may have missed from theCUBE’s coverage of the DigiCert Trust Summit

DGX agent

AI is turning digital trust from a security function into an operating model. That shift is putting new pressure on the systems enterprises have long used to verify identity, protect data and keep dig

agentssiliconangle
21 May 2026
Agents

Why Latent Actions Fail, and How to Prevent It

DGX agent

arXiv:2605.20223v1 Announce Type: new Abstract: Latent action models (LAMs) aim to learn action-like representations from unlabeled videos by compressing frame-to-frame changes. The frames of in-the-w

agentsarxiv-cs-cv
21 May 2026
Agents

A Geometric Analysis of Small-sized Language Model Hallucinations

DGX agent

arXiv:2602.14778v3 Announce Type: replace-cross Abstract: Hallucinations -- plausible but factually incorrect responses -- pose a major challenge to the reliability of Large Language Models (LLMs), es

agentsarxiv-cs-ai
20 May 2026
Agents

A Logistic Regression Model to Predict Malaria Severity in Children

DGX agent

arXiv:2605.18900v1 Announce Type: cross Abstract: One of the main causes of death around the globe is malaria. Researchers have sought to develop predictive models for malaria outbreaks based on meteo

agentsarxiv-cs-lg
20 May 2026
Agents

Active Graph really feels like the culmination of all of my BabyAGI and graph experiments. [fyi, technical history of babyagi: http://babyag…

DGX agent

Active Graph really feels like the culmination of all of my BabyAGI and graph experiments. [fyi, technical history of babyagi: http://babyagi.wiki] would love to hear thoughts if you try it out! the e

agentsyohei-nakajima--x
20 May 2026
Agents

Causal Evidence that Language Models use Confidence to Drive Behavior

DGX agent

arXiv:2603.22161v2 Announce Type: replace Abstract: Metacognition -- assessing the quality of one's own cognitive performance -- guides adaptive behavior across species. Substantial research demonstra

agentsarxiv-cs-lg
20 May 2026
Agents

CLUE: Adaptively Prioritized Contextual Cues by Leveraging a Unified Semantic Map for Effective Zero-Shot Object-Goal Navigation

DGX agent

arXiv:2605.19206v1 Announce Type: new Abstract: Zero-shot object-goal navigation (ZSON) is a challenging problem in robotics that requires a comprehensive understanding of both language and visual obs

agentsarxiv-cs-ro
20 May 2026
Agents

DECOR: Auditing LLM Deception via Information Manipulation Theory

DGX agent

arXiv:2605.19270v1 Announce Type: new Abstract: Large language models can deceive by subtly manipulating truthful information -- omitting key facts, shifting focus, or obscuring meaning -- making such

agentsarxiv-cs-cl
20 May 2026
Safety

Dual-Gated Epistemic Time-Dilation: Autonomous Compute Modulation in Asynchronous MARL

DGX agent

arXiv:2603.23722v2 Announce Type: replace-cross Abstract: While Multi-Agent Reinforcement Learning (MARL) algorithms achieve unprecedented successes across complex continuous domains, their standard d

safetyarxiv-cs-lg
20 May 2026
Safety

ESLD (External Surrogate Latent Defense): A Latent-Space Architecture for Faster, Stronger Prompt-Injection Defense

DGX agent

arXiv:2605.18918v1 Announce Type: cross Abstract: Modern AI assistants are agentic. To answer a single user request, the underlying language model pulls in information from many sources, such as web s

safetyarxiv-cs-ai
20 May 2026
Agents

FAGER: Factually Grounded Evaluation and Refinement of Text-to-Image Models

DGX agent

arXiv:2605.19111v1 Announce Type: cross Abstract: Existing text-to-image (T2I) evaluation metrics mainly assess whether generated images align with information explicitly stated in the prompt, but oft

agentsarxiv-cs-ai
20 May 2026
Agents

High-quality generation of dynamic game content via small language models: A proof of concept

DGX agent

arXiv:2601.23206v2 Announce Type: replace Abstract: Large language models (LLMs) offer promise for dynamic game content generation, but they face critical barriers, including narrative incoherence and

agentsarxiv-cs-ai
20 May 2026
Agents

Hybrid Training for Vision-Language-Action Models

DGX agent

arXiv:2510.00600v2 Announce Type: replace-cross Abstract: Using Large Language Models to produce intermediate thoughts, a.k.a. Chain-of-thought (CoT), before providing an answer has been a successful

agentsarxiv-cs-ai
20 May 2026
Agents

Library Drift: Diagnosing and Fixing a Silent Failure Mode in Self-Evolving LLM Skill Libraries

DGX agent

arXiv:2605.19576v1 Announce Type: new Abstract: Self-evolving skill libraries face a silent failure mode we term library drift: unbounded skill accumulation without outcome-driven lifecycle management

agentsarxiv-cs-ai
20 May 2026
Agents

LLMs are stateless (every time you reply to an LLM, you re-inject the entire conversation to a fresh inference) the purpose of memory is to …

DGX agent

LLMs are stateless (every time you reply to an LLM, you re-inject the entire conversation to a fresh inference) the purpose of memory is to provide continuity (in games we'd call it a 'persistent worl

agentsyohei-nakajima--x
20 May 2026
Agents

Operationalising Artificial Intelligence Bills of Materials (AIBOMs) for Verifiable AI Provenance and Lifecycle Assurance

DGX agent

arXiv:2605.19755v1 Announce Type: cross Abstract: Artificial Intelligence (AI) systems are increasingly dependent on complex, multi-layered software supply chains that introduce challenges for reprodu

agentsarxiv-cs-ai
20 May 2026
Agents

@Replit narrative walkthrough video of repo for anyone interested: https://x.com/FileCityAI/status/2057164885780226139?s=20

DGX agent

@Replit narrative walkthrough video of repo for anyone interested: https://x.com/FileCityAI/status/2057164885780226139?s=20 FileCity Tour: activegraph An event-sourced reactive graph runtime for long-

agentsyohei-nakajima--x
20 May 2026
Model Releases

Synthesis and Evaluation of Long-term History-aware Medical Dialogue

DGX agent

arXiv:2605.19766v1 Announce Type: cross Abstract: An effective healthcare agent must be able to recall and reason over a patient's longitudinal medical history. However, the absence of datasets with r

model-releasesarxiv-cs-ai
20 May 2026
Agents

The 99% Success Paradox: When Near-Perfect Retrieval Equals Random Selection

DGX agent

arXiv:2605.18857v1 Announce Type: cross Abstract: For most of the history of information retrieval (IR), search results were designed for human consumers who could scan, filter, and discard irrelevant

agentsarxiv-cs-ai
20 May 2026
Agents

this is how you add an event, fork and cache a run, and then find the diff between a parent and fork in this example, the fork shares the pa…

DGX agent

this is how you add an event, fork and cache a run, and then find the diff between a parent and fork in this example, the fork shares the parent's event log up to event 142. from 143 onward it diverge

agentsyohei-nakajima--x
20 May 2026
Model Releases

We're excited to be an official shoutout at the Google I/O Developer Keynote 🔥 @llama_index is building the document infrastructure for AI …

DGX agent

We're excited to be an official shoutout at the Google I/O Developer Keynote 🔥 @llama_index is building the document infrastructure for AI agents, and we plan to integrate even more heavily with both

model-releasesjerry-liu--x
20 May 2026
Agents

We’re hiring for Labs! 🧪 If you’re interested in working with us to push forward Continual Learning, pls DM me with a blurb + link to the b…

DGX agent

We’re hiring for Labs! 🧪 If you’re interested in working with us to push forward Continual Learning, pls DM me with a blurb + link to the best Applied Research you’ve done (or even better shipped!) yo

agentsharrison-chase--x
20 May 2026
Agents

YAC: Bridging Natural Language and Interactive Visual Exploration with Generative AI for Biomedical Data Discovery

DGX agent

arXiv:2509.19182v2 Announce Type: replace-cross Abstract: Incorporating natural language input has the potential to improve the capabilities of biomedical data discovery interfaces. However, user inte

agentsarxiv-cs-ai
20 May 2026
Agents

A Mechanistic Model for Collective Motion from Sensorimotor Regularities

DGX agent

arXiv:2605.16522v1 Announce Type: new Abstract: Collective behavior in animals has long been modeled through self-propelled particle models, which reproduce striking group-level phenomena through abst

agentsarxiv-cs-ro
19 May 2026
Model Releases

A Pilot Benchmark for NL-to-FOL Translation in Planetary Exploration

DGX agent

arXiv:2605.17911v1 Announce Type: new Abstract: Future planetary exploration envisions autonomous robotic agents operating under severe communication constraints, without global positioning, and with

model-releasesarxiv-cs-cl
19 May 2026
Agents

Baba in Wonderland: Online Self-Supervised Dynamics Discovery for Executable World Models

DGX agent

arXiv:2605.16725v1 Announce Type: new Abstract: Executable world models can be read, edited, executed, and reused for planning, but only if the program captures the environment's transition law rather

agentsarxiv-cs-ai
19 May 2026
Model Releases

Causely: A Causal Intelligence Layer for Enterprise AI A Benchmark Study on SRE and Reliability Workflows

DGX agent

arXiv:2605.18327v1 Announce Type: new Abstract: AI agents deployed into SRE workflows currently derive their understanding of environment state from raw observability telemetry at query time, paying a

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

CooT: Learning to Coordinate In-Context with Coordination Transformers

DGX agent

arXiv:2506.23549v3 Announce Type: replace Abstract: Effective coordination among unfamiliar partners remains a major challenge in multi-agent systems. Existing approaches, such as population-based met

model-releasesarxiv-cs-ai
19 May 2026
Agents

DeepArrhythmia: Segment-Contextualized ECG Arrhythmia Classification via Selective Evidence Acquisition

DGX agent

arXiv:2605.16441v1 Announce Type: cross Abstract: Beat-level Electrocardiography (ECG) arrhythmia detection aims to assign an arrhythmia class to each beat in a recording, yet many existing systems tr

agentsarxiv-cs-ai
19 May 2026
Agents

Democratizing Large-Scale Re-Optimization with LLM-Guided Model Patches

DGX agent

arXiv:2605.18692v1 Announce Type: new Abstract: Optimization models developed by operations research (OR) experts are often deployed as decision-support systems in industrial settings. However, real-w

agentsarxiv-cs-ai
19 May 2026
Model Releases

DocReward: A Document Reward Model for Structuring and Stylizing

DGX agent

arXiv:2510.11391v3 Announce Type: replace-cross Abstract: Recent agentic workflows automate professional document generation but focus narrowly on textual quality, overlooking structural and stylistic

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

ESI-Bench: Towards Embodied Spatial Intelligence that Closes the Perception-Action Loop

DGX agent

arXiv:2605.18746v1 Announce Type: cross Abstract: Spatial intelligence unfolds through a perception-action loop: agents act to acquire observations, and reason about how observations vary as a functio

model-releasesarxiv-cs-ai
19 May 2026
Agents

Excited to share our new paper: RoPE Distinguishes Neither Positions Nor Tokens in Long Contexts, Provably LLMs often fail on inputs well wi…

DGX agent

Excited to share our new paper: RoPE Distinguishes Neither Positions Nor Tokens in Long Contexts, Provably LLMs often fail on inputs well within their advertised context lengths. We show that these fa

agentsjeremy-howard--x
19 May 2026
Agents

Generative AI Advertising as a Problem of Trustworthy Commercial Intervention

DGX agent

arXiv:2605.18673v1 Announce Type: cross Abstract: Major deployed generative AI advertising systems preserve a visible boundary between commercial content and AI-generated responses. Yet empirical rese

agentsarxiv-cs-cl
19 May 2026
← Previous
1…237238239240241…375
Next →