AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlog
84,433Total entries
1Added by human
84,432Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,911 results
Hardware

Falkor-IRAC: Graph-Constrained Generation for Verified Legal Reasoning in Indian Judicial AI

DGX agent

arXiv:2605.14665v1 Announce Type: new Abstract: Legal reasoning is not semantic similarity search. A court judgment encodes constrained symbolic reasoning: precedent propagation, procedural state tran

hardwarearxiv-cs-ai
15 May 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

MemLens: Benchmarking Multimodal Long-Term Memory in Large Vision-Language Models

DGX agent

arXiv:2605.14906v1 Announce Type: new Abstract: Memory is essential for large vision-language models (LVLMs) to handle long, multimodal interactions, with two method directions providing this capabili

model-releasesarxiv-cs-cv
15 May 2026
Local Ai

Self-Pruned Key-Value Attention: Learning When to Write by Predicting Future Utility

DGX agent

arXiv:2605.14037v1 Announce Type: cross Abstract: Under modern test-time compute and agentic paradigms, language models process ever-longer sequences. Efficient text generation with transformer archit

local-aiarxiv-cs-cl
15 May 2026
Safety

SpeakerLLM: A Speaker-Specialized Audio-LLM for Speaker Understanding and Verification Reasoning

DGX agent

arXiv:2605.15044v1 Announce Type: cross Abstract: As audio-first agents become increasingly common in physical AI, conversational robots, and screenless wearables, audio large language models (audio-L

safetyarxiv-cs-ai
15 May 2026
Model Releases

SPIN: Structural LLM Planning via Iterative Navigation for Industrial Tasks

DGX agent

arXiv:2605.14051v1 Announce Type: new Abstract: Industrial LLM agent systems often separate planning from execution, yet LLM planners frequently produce structurally invalid or unnecessarily long work

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

VerbalValue: A Socially Intelligent Virtual Host for Sales-Driven Live Commerce

DGX agent

arXiv:2605.14542v1 Announce Type: new Abstract: A skilled live-commerce host is not merely a narrator, but a sales agent who converts viewer curiosity into purchase intent through expert product knowl

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

big launches today!

DGX agent

big launches today! 🚀Launching: LangSmith Engine LangSmith Engine is an agent that sits on top of your traces It runs in the background and automatically identifies issues It then proactively suggests

model-releasesharrison-chase--x
14 May 2026
Safety

Causality-Aware End-to-End Autonomous Driving via Ego-Centric Joint Scene Modeling

DGX agent

arXiv:2605.13646v1 Announce Type: cross Abstract: End-to-end autonomous driving, which bypasses traditional modular pipelines by directly predicting future trajectories from sensor inputs, has recentl

safetyarxiv-cs-ai
14 May 2026
Safety

Differentiable Evolutionary Reinforcement Learning

DGX agent

arXiv:2512.13399v2 Announce Type: replace Abstract: Crafting effective reward signals remains a central challenge in Reinforcement Learning (RL), especially for complex reasoning tasks. Existing autom

safetyarxiv-cs-ai
14 May 2026
Applications

EgoForce: Robust Online Egocentric Motion Reconstruction via Diffusion Forcing

DGX agent

arXiv:2605.13041v1 Announce Type: new Abstract: With recent advances in embodied agents and AR devices, egocentric observations are readily available as input for real-world interactive online applica

applicationsarxiv-cs-cv
14 May 2026
Model Releases

Embodied Neurocomputation: A Framework for Interfacing Biological Neural Cultures with Scaled Task-Driven Validation

DGX agent

arXiv:2605.13315v1 Announce Type: cross Abstract: Biological neural networks (BNNs) have been established as a powerful and adaptive substrate that offer the potential for incredibly energy and data e

model-releasesarxiv-cs-lg
14 May 2026
Industry

Grok Build

DGX agent

Grok Build Grok Build is a fully interactive CLI, which means you can actually use your mouse to click. No flickers. Especially useful as I find myself running 5+ agents at a time and jumping between

industryelon-musk--x
14 May 2026
Model Releases

Language Model Goal Selection Differs from Humans' in a Self-Directed Learning Task

DGX agent

arXiv:2603.03295v2 Announce Type: replace-cross Abstract: Whether in agentic workflows, social studies, or chat settings, large language models (LLMs) are increasingly being asked to replace humans in

model-releasesarxiv-cs-ai
14 May 2026
Safety

Learning When to Act: Communication-Efficient Reinforcement Learning via Run-Time Assurance

DGX agent

arXiv:2605.12561v1 Announce Type: new Abstract: Safe reinforcement learning (RL) typically asks extit{what} an agent should do. We ask extit{when} it needs to act, and show that a single policy can jo

safetyarxiv-cs-lg
14 May 2026
Tools

Mitchell's post here reminded me of a similar conversation I had recently about how cheap it can be to port native mobile apps to React Nati…

DGX agent

Mitchell's post here reminded me of a similar conversation I had recently about how cheap it can be to port native mobile apps to React Native using coding agents... and then port them back again late

toolssimon-willison--x
14 May 2026
Tutorials

Probabilistic Prediction Markets with Intermittent Contributions

DGX agent

arXiv:2510.13385v3 Announce Type: replace Abstract: Although both data availability and the demand for accurate forecasts are increasing, collaboration between stakeholders is often constrained by dat

tutorialsarxiv-cs-lg
14 May 2026
Model Releases

the langsmith team cooked on this one @jakebroekhuizen this + claude code + langsmith cli has been 🔥

DGX agent

the langsmith team cooked on this one @jakebroekhuizen this + claude code + langsmith cli has been 🔥 🚀Launching: LangSmith Engine LangSmith Engine is an agent that sits on top of your traces It runs i

model-releasesharrison-chase--x
14 May 2026
Model Releases

ToolWeave: Structured Synthesis of Complex Multi-Turn Tool-Calling Dialogues

DGX agent

arXiv:2605.12521v1 Announce Type: cross Abstract: Multi-turn tool calling is essential for LLMs to function as autonomous agents, yet synthesizing the training data required for these capabilities rem

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

3D-Belief: Embodied Belief Inference via Generative 3D World Modeling

DGX agent

arXiv:2605.11367v1 Announce Type: new Abstract: Recent advances in visual generative models have highlighted the promise of learning generative world models. However, most existing approaches frame wo

model-releasesarxiv-cs-cv
13 May 2026
Model Releases

Covering Human Action Space for Computer Use: Data Synthesis and Benchmark

DGX agent

arXiv:2605.12501v1 Announce Type: new Abstract: Computer-use agents (CUAs) automate on-screen work, as illustrated by GPT-5.4 and Claude. Yet their reliability on complex, low-frequency interactions i

model-releasesarxiv-cs-cv
13 May 2026
Model Releases

DarkQA: Benchmarking Vision-Language Models on Visual-Primitive Question Answering in Low-Light Indoor Scenes

DGX agent

arXiv:2512.24985v4 Announce Type: replace Abstract: Vision Language Models (VLMs) are increasingly adopted as central reasoning modules for embodied agents. Existing benchmarks evaluate their capabili

model-releasesarxiv-cs-cv
13 May 2026
Model Releases

Efficient Algorithms for Logistic Contextual Slate Bandits with Bandit Feedback

DGX agent

arXiv:2506.13163v3 Announce Type: replace Abstract: We study the Logistic Contextual Slate Bandit problem, where, at each round, an agent selects a slate of N items from an exponentially large set (of

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

Fine-Tuning Large Language Models for Cooperative Tactical Deconfliction of Small Unmanned Aerial Systems

DGX agent

arXiv:2603.28561v2 Announce Type: replace Abstract: The growing deployment of small Unmanned Aerial Systems (sUASs) in low-altitude airspaces has increased the need for reliable tactical deconfliction

model-releasesarxiv-cs-ro
13 May 2026
Model Releases

Nautilus: From One Prompt to Plug-and-Play Robot Learning

DGX agent

arXiv:2605.11665v1 Announce Type: new Abstract: Robot learning research is fragmented across policy families, benchmark suites, and real robots; each implementation is entangled with the others in a c

model-releasesarxiv-cs-ro
13 May 2026
Hardware

NVIDIA, Ineffable Intelligence Team Up to Build the Future of Reinforcement Learning Infrastructure

DGX agent

Reinforcement-learning agents — AI systems that learn by trial and error — can convert computation into new knowledge. That’s the focus of a new engineering-level collaboration between NVIDIA and Inef

hardwarenvidia-blog
13 May 2026
Local Ai

p2p ai inference mesh

DGX agent

A peer-to-peer AI inference mesh is a decentralized network architecture where agents connect directly to discover peers and communicate through distributed protocols to share inference workloads . Su

local-air-ollama
13 May 2026
Safety

PriorZero: Bridging Language Priors and World Models for Decision Making

DGX agent

arXiv:2605.12289v1 Announce Type: new Abstract: Leveraging the rich world knowledge of Large Language Models (LLMs) to enhance Reinforcement Learning (RL) agents offers a promising path toward general

safetyarxiv-cs-lg
13 May 2026
Research

Shapley Value Approximation Based on k-Additive Games

DGX agent

arXiv:2502.04763v2 Announce Type: replace-cross Abstract: The Shapley value is the prevalent solution for fair division problems in which a payout is to be divided among multiple agents. By adopting a

researcharxiv-cs-lg
13 May 2026
Safety

The new era of SaMD: Why cloud infrastructure is the foundation for digital health in 2026

DGX agent

In the healthcare and life sciences industries, speed saves lives, but meeting regulatory requirements and other administrative burdens often pumps the brakes for manufacturers of software as a medica

safetygoogle-cloud-ai
13 May 2026
Safety

Transferable Delay-Aware Reinforcement Learning via Implicit Causal Graph Modeling

DGX agent

arXiv:2605.12312v1 Announce Type: new Abstract: Random delays weaken the temporal correspondence between actions and subsequent state feedback, making it difficult for agents to identify the true prop

safetyarxiv-cs-lg
13 May 2026
Hardware

TriBand-BEV: Real-Time LiDAR-Only 3D Pedestrian Detection via Height-Aware BEV and High-Resolution Feature Fusion

DGX agent

arXiv:2605.12220v1 Announce Type: new Abstract: Safe autonomous agents and mobile robots need fast real time 3D perception, especially for vulnerable road users (VRUs) such as pedestrians. We introduc

hardwarearxiv-cs-cv
13 May 2026
Model Releases

UHR-Micro: Diagnosing and Mitigating the Resolution Illusion in Earth Observation VLMs

DGX agent

arXiv:2605.12237v1 Announce Type: new Abstract: Vision-Language Models (VLMs) increasingly operate on ultra-high-resolution (UHR) Earth observation imagery, yet they remain vulnerable to a severe scal

model-releasesarxiv-cs-cv
13 May 2026
Model Releases

Beyond Self-Play and Scale: A Behavior Benchmark for Generalization in Autonomous Driving

DGX agent

arXiv:2605.10034v1 Announce Type: new Abstract: Recent Autonomous Driving (AD) works such as GigaFlow and PufferDrive have unlocked Reinforcement Learning (RL) at scale as a training strategy for driv

model-releasesarxiv-cs-ro
12 May 2026
Research

Beyond Thinking: Imagining in 360^irc for Humanoid Visual Search

DGX agent

arXiv:2605.09146v1 Announce Type: new Abstract: Humanoid Visual Search (HVS) requires agents to actively explore immersive 360^irc environments. While prior methods treat this as a monolithic task rel

researcharxiv-cs-cv
12 May 2026
Model Releases

Do Benchmarks Underestimate LLM Performance? Evaluating Hallucination Detection With LLM-First Human-Adjudicated Assessment

DGX agent

arXiv:2605.08462v1 Announce Type: cross Abstract: Hallucination remains a persistent challenge in Large Language Models (LLMs), particularly in context-grounded settings such as RAG and agentic AI sys

model-releasesarxiv-cs-ai
12 May 2026
Tutorials

Efficient Estimation of Kernel Surrogate Models for Task Attribution

DGX agent

arXiv:2602.03783v2 Announce Type: replace-cross Abstract: Modern AI agents such as large language models are trained on diverse tasks -- translation, code generation, mathematical reasoning, and text

tutorialsarxiv-cs-ai
12 May 2026
Safety

Equivariant Reinforcement Learning for Clifford Quantum Circuit Synthesis

DGX agent

arXiv:2605.10910v1 Announce Type: cross Abstract: We consider the problem of synthesizing Clifford quantum circuits for devices with all-to-all qubit connectivity. We approach this task as a reinforce

safetyarxiv-cs-lg
12 May 2026
Safety

Execution Envelopes: A Shared Admission Contract for Backend AI Execution Requests

DGX agent

arXiv:2605.08267v1 Announce Type: cross Abstract: Enterprise AI backends increasingly admit heterogeneous execution requests across model deployment, inference, evaluation, data movement, and agentic

safetyarxiv-cs-ai
12 May 2026
Safety

Governed Metaprogramming for Intelligent Systems: Reclassifying Eval as a Governed Effect

DGX agent

arXiv:2605.05248v2 Announce Type: replace-cross Abstract: AI systems increasingly synthesize executable structure at runtime: LLMs generate programs, agents construct workflows,self-improving systems

safetyarxiv-cs-ai
12 May 2026
Model Releases

Higher Resolution, Better Generalization: Unlocking Visual Scaling in Deep Reinforcement Learning

DGX agent

arXiv:2605.10546v1 Announce Type: new Abstract: Pixel-based deep reinforcement learning agents are typically trained on heavily downsampled visual observations, a convention inherited from early bench

model-releasesarxiv-cs-lg
12 May 2026
Research

Learning Strategic Value and Cooperation in Multi-Player Stochastic Games through Side Payments

DGX agent

arXiv:2303.05307v2 Announce Type: replace-cross Abstract: We study general-sum, multi-player stochastic games with transferable utility, motivated by settings where agents can use side payments to mak

researcharxiv-cs-ai
12 May 2026
Safety

LLM Advertisement based on Neuron Auctions

DGX agent

arXiv:2605.08326v1 Announce Type: cross Abstract: As Large Language Models (LLMs) transition into conversational agents, generative advertising emerges as a crucial monetization strategy. However, emb

safetyarxiv-cs-ai
12 May 2026
Model Releases

MLS-Bench: A Holistic and Rigorous Assessment of AI Systems on Building Better AI

DGX agent

arXiv:2605.08678v1 Announce Type: new Abstract: Modern AI progress has been driven by ML methods that are generalizable across settings and scalable to larger regimes. As large language models demonst

model-releasesarxiv-cs-lg
12 May 2026
Safety

Neural Co-state Policies: Structuring Hidden States in Recurrent Reinforcement Learning

DGX agent

arXiv:2605.05373v2 Announce Type: replace Abstract: A key capability of intelligent agents is operating under partial observability: reasoning and acting effectively despite missing or incomplete stat

safetyarxiv-cs-lg
12 May 2026
Model Releases

PDEAgent-Bench: A Multi-Metric, Multi-Library Benchmark for PDE Solver Generation

DGX agent

arXiv:2605.09636v1 Announce Type: new Abstract: PDE-to-solver code generation aims to automatically synthesize executable numerical solvers from partial differential equation (PDE) specifications. Thi

model-releasesarxiv-cs-ai
12 May 2026
Local Ai

Playing games with knowledge: AI-Induced delusions need game theoretic interventions

DGX agent

arXiv:2605.08409v1 Announce Type: new Abstract: Conversational AI has a fundamental flaw as a knowledge interface: sycophantic chatbots induce epistemic entrenchment and delusional belief spirals even

local-aiarxiv-cs-ai
12 May 2026
Model Releases

Position: AI Security Policy Should Target Systems, Not Models

DGX agent

arXiv:2605.09504v1 Announce Type: cross Abstract: We present swarm-attack, an open-source adversarial testing framework in which multiple lightweight LLM agents coordinate through shared memory, paral

model-releasesarxiv-cs-ai
12 May 2026
Safety

Reflective Prompted Policy Optimization: Trajectory-Grounded Revision and Salience Bias

DGX agent

arXiv:2605.08315v1 Announce Type: new Abstract: Existing LLM-based policy optimizers see only scalar rewards: that a policy scored 0.45, but not whether the agent got stuck in a loop, fell into a hole

safetyarxiv-cs-lg
12 May 2026
← Previous
1…294295296297298…374
Next →