AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,919 results
Model Releases

Learning Efficient Guardrails for Compliance

DGX agent

arXiv:2510.03485v2 Announce Type: replace Abstract: Autonomous web agents are increasingly deployed for long-horizon tasks, yet their ability to adhere to real-world policies remains critically undere

model-releasesarxiv-cs-ai
20 May 2026
Model Releases
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

// Memory as a Model // The paper augments any LLM with a separate trained memory model that stores, retrieves, and integrates facts on its …

DGX agent

// Memory as a Model // The paper augments any LLM with a separate trained memory model that stores, retrieves, and integrates facts on its behalf. It decouples memory updates from base-model weight u

model-releasesdair-ai--x
20 May 2026
Model Releases

P2DNav: Panorama-to-Downview Reasoning for Zero-shot Vision-and-Language Navigation

DGX agent

arXiv:2605.19634v1 Announce Type: cross Abstract: Vision-and-language navigation (VLN) requires an embodied agent to ground natural-language instructions into executable navigation actions in unseen e

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Reporting from Google I/O 2026 with the four biggest themes from one of the biggest AI labs in the world. 🎤 Voice AI as an interface Google…

DGX agent

Reporting from Google I/O 2026 with the four biggest themes from one of the biggest AI labs in the world. 🎤 Voice AI as an interface Google and Samsung announced new Gemini-powered glasses with Gentle

model-releasesallie-k--miller--x
20 May 2026
Safety

Sampling-Based Safe Reinforcement Learning

DGX agent

arXiv:2605.19469v1 Announce Type: cross Abstract: Safe exploration remains a fundamental challenge in reinforcement learning (RL), limiting the deployment of RL agents in the real world. We propose Sa

safetyarxiv-cs-ai
20 May 2026
Safety

Toward an AI-Powered Computational Testbed for Workforce Policy

DGX agent

arXiv:2605.19064v1 Announce Type: cross Abstract: Workforce transformations are difficult to forecast and costly to mismanage. In particular, the integration of artificial intelligence into knowledge

safetyarxiv-cs-ai
20 May 2026
Applications

We added 600+ new voices on Together AI! Introducing MiniMax Speech 2.8 Turbo on Together AI, an enterprise TTS model for expressive real-ti…

DGX agent

We added 600+ new voices on Together AI! Introducing MiniMax Speech 2.8 Turbo on Together AI, an enterprise TTS model for expressive real-time voice agents. AI natives can now deploy @MiniMax_AI Speec

applicationstogether-ai--x
20 May 2026
Research

A Structural Threshold in Decision Capacity Governs Collapse in Self-Play Reinforcement Learning

DGX agent

arXiv:2605.16315v1 Announce Type: cross Abstract: We show that a threshold in decision capacity determines whether self-play reinforcement learning agents collapse under asymmetric rule perturbations.

researcharxiv-cs-ai
19 May 2026
Safety

ARROW: Augmented Replay for RObust World models

DGX agent

arXiv:2603.11395v2 Announce Type: replace-cross Abstract: Continual reinforcement learning challenges agents to acquire new skills while retaining previously learned ones with the goal of improving pe

safetyarxiv-cs-ai
19 May 2026
Safety

Assured autonomy: How operations research powers and orchestrates generative AI systems

DGX agent

arXiv:2512.23978v2 Announce Type: replace Abstract: Generative artificial intelligence (GenAI) is shifting from conversational assistants toward agentic systems -- autonomous decision-making systems t

safetyarxiv-cs-lg
19 May 2026
Model Releases

Automated Root-Cause Subclassification and No-Code Fix Generation for Invalid Bug Reports

DGX agent

arXiv:2605.17561v1 Announce Type: cross Abstract: Issues faced when using software are reported in the form of bug reports. However, many bug reports are invalid, meaning they do not require code chan

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

CommitDistill: A Lightweight Knowledge-Centric Memory Layer for Software Repositories

DGX agent

arXiv:2605.18284v1 Announce Type: cross Abstract: Software repositories accumulate large amounts of unstructured knowledge in commit messages, pull-request discussions, and issue threads, but develope

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

DecoupleSearch: Decouple Planning and Search via Hierarchical Reward Modeling

DGX agent

arXiv:2510.21712v2 Announce Type: replace-cross Abstract: Retrieval-Augmented Generation (RAG) systems have emerged as a pivotal methodology for enhancing Large Language Models (LLMs) through the dyna

model-releasesarxiv-cs-ai
19 May 2026
Safety

DISA: Offline Importance Sampling for Distribution-Matching LLM-RL

DGX agent

arXiv:2605.17295v1 Announce Type: cross Abstract: Modern reasoning agents are increasingly evaluated on their ability to generate multiple valid solution paths, plans, or tool-use traces for a given i

safetyarxiv-cs-cl
19 May 2026
Model Releases

EPIC-Bench: A Perception-Centric Benchmark for Fine-Grained Embodied Visual Grounding in Vision-Language Models

DGX agent

arXiv:2605.17070v1 Announce Type: new Abstract: While large vision-language models (VLMs) are increasingly adopted as the perceptual backbone for embodied agents, existing benchmarks often rely on que

model-releasesarxiv-cs-cv
19 May 2026
Applications

Excited for this conversation with @hwchase17 , Co-Founder & CEO of @LangChain , on June 2 at the @traversal_ai office in NYC. We’ll be talk…

DGX agent

Excited for this conversation with @hwchase17 , Co-Founder & CEO of @LangChain , on June 2 at the @traversal_ai office in NYC. We’ll be talking about what it actually takes to operate AI agents in pro

applicationsharrison-chase--x
19 May 2026
Model Releases

Firefly: Illuminating Large-Scale Verified Tool-Call Data Generation from Real APIs

DGX agent

arXiv:2605.17558v1 Announce Type: cross Abstract: Training tool-calling agents requires large-scale trajectory data with verifiable labels, yet existing approaches either synthesize environments that

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

Fix the Structural Bottleneck: Context Compression via Explicit Information Transmission

DGX agent

arXiv:2602.03784v2 Announce Type: replace Abstract: Long-context LLM agents often struggle with growing token, memory, and latency costs, making efficient context compression essential for practical d

model-releasesarxiv-cs-cl
19 May 2026
Safety

Generating Realistic Safety-Critical Scenarios for Vehicle-Pedestrian Interactions

DGX agent

arXiv:2605.17229v1 Announce Type: new Abstract: Automated driving system deployment requires rigorous validation across safety-critical vehicle-pedestrian interactions, yet real-world datasets rarely

safetyarxiv-cs-ro
19 May 2026
Research

GRID: Graph Representation of Intelligence Data for Security Text Knowledge Graph Construction

DGX agent

arXiv:2605.16714v1 Announce Type: new Abstract: Security knowledge graphs can provide computable external memory for security agents, but constructing them from long-form cyber threat intelligence (CT

researcharxiv-cs-ai
19 May 2026
Model Releases

HydroAgent: Closing the Gap Between Frontier LLMs and Human Experts in Hydrologic Model Calibration via Simulator-Grounded RL

DGX agent

arXiv:2605.17792v1 Announce Type: new Abstract: Calibrating distributed hydrologic models is a critical bottleneck across operational water resources management - streamflow prediction, reservoir oper

model-releasesarxiv-cs-lg
19 May 2026
Model Releases

Internalizing Tool Knowledge in Small Language Models via QLoRA Fine-Tuning

DGX agent

arXiv:2605.17774v1 Announce Type: new Abstract: Large language models are increasingly used as planning components in agentic systems, but current tool-use pipelines often require full tool schemas to

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

Monocular Open Vocabulary Occupancy Prediction for Indoor Scenes

DGX agent

arXiv:2602.22667v2 Announce Type: replace Abstract: Open-vocabulary 3D occupancy is vital for embodied agents, which need to understand complex indoor environments where semantic categories are abunda

model-releasesarxiv-cs-cv
19 May 2026
Research

MORN: Metacognitive Object-Goal Regulation for Resource-Rational Long-Horizon Navigation

DGX agent

arXiv:2605.16932v1 Announce Type: new Abstract: Robots deployed in unstructured human environments must frequently execute long-horizon missions, such as find the mug, then the chair, then the printer

researcharxiv-cs-ro
19 May 2026
Research

Non-Colliding Biometric Identities for Digital Entities: Geometry, Capacity, and Million-Scale Virtual Identity Provisioning

DGX agent

arXiv:2605.18238v1 Announce Type: new Abstract: Digital entities such as AI agents and humanoid robots increasingly operate alongside real humans, yet their identity infrastructure is based on credent

researcharxiv-cs-cv
19 May 2026
Model Releases

Not What You Asked For: Typographic Attacks in Household Robot Manipulation

DGX agent

arXiv:2605.18593v1 Announce Type: cross Abstract: Open-vocabulary embodied AI agents increasingly rely on vision-language models such as CLIP for object perception and task grounding. However, the sha

model-releasesarxiv-cs-ai
19 May 2026
Local Ai

OxyGen: Unified KV Cache Management for VLA Inference under Multi-Task Parallelism

DGX agent

arXiv:2603.14371v2 Announce Type: replace-cross Abstract: Embodied AI agents increasingly require parallel execution of multiple tasks, such as manipulation, conversation, and memory construction, fro

local-aiarxiv-cs-ai
19 May 2026
Local Ai

Principles of frugal inference and control

DGX agent

arXiv:2406.14427v4 Announce Type: replace Abstract: A central challenge for intelligent agents in an uncertain world is striking the right balance between utility maximization and resource use, not on

local-aiarxiv-cs-ai
19 May 2026
Model Releases

Privacy Preserving Reinforcement Learning with One-Sided Feedback

DGX agent

arXiv:2605.18246v1 Announce Type: cross Abstract: We study reinforcement learning (RL) in multi-dimensional continuous state and action spaces with one-sided feedback, where the agent receives partial

model-releasesarxiv-cs-ai
19 May 2026
Safety

SD-Search: On-Policy Hindsight Self-Distillation for Search-Augmented Reasoning

DGX agent

arXiv:2605.18299v1 Announce Type: new Abstract: Search-augmented reasoning agents interleave internal reasoning with calls to an external retriever, and their performance relies on the quality of each

safetyarxiv-cs-ai
19 May 2026
Tutorials

Starve to Perceive: Taming Lazy Perception in VLMs with Constrained Visual Bandwidth

DGX agent

arXiv:2605.18603v1 Announce Type: new Abstract: Vision-Language Models (VLMs) deployed as situated agents in high-resolution visual environments require active perception -- the ability to dynamically

tutorialsarxiv-cs-cv
19 May 2026
Model Releases

Today, we launched a brand-new intelligent Search box. Here's what that means: An upgrade to the Search experience with our most advanced Ge…

DGX agent

Today, we launched a brand-new intelligent Search box. Here's what that means: An upgrade to the Search experience with our most advanced Gemini 3.5 models, bringing with them our latest agentic capab

model-releasesgoogle-ai--x
19 May 2026
Local Ai

Use your LM Studio models to code locally in @zeddotdev 🚀

DGX agent

Use your LM Studio models to code locally in @zeddotdev 🚀 Local model usage grew 3x in Zed's agent in the last 10 weeks. Cameron Mcloughlin on why he prefers local: 'I worry about over-reliance on pro

local-ailm-studio--x
19 May 2026
Applications

What Drives Success in Physical Planning with Joint-Embedding Predictive World Models?

DGX agent

arXiv:2512.24497v3 Announce Type: replace Abstract: A long-standing challenge in AI is to develop agents capable of solving a wide range of physical tasks and generalizing to new, unseen tasks and env

applicationsarxiv-cs-ai
19 May 2026
Model Releases

When Outcome Looks Right But Discipline Fails: Trace-Based Evaluation Under Hidden Competitor State

DGX agent

arXiv:2605.18580v1 Announce Type: new Abstract: Outcome-only evaluation can certify economically unsafe agents: a policy can hit a business KPI while violating deployable behavioral discipline. In hot

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

WinDeskGround: A Benchmark for Robust GUI Grounding in Complex Multi-Window Desktop Environments

DGX agent

arXiv:2605.16402v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) have revolutionized GUI automation, yet their efficacy is largely established on idealized, single-layer interf

model-releasesarxiv-cs-cv
19 May 2026
Model Releases

WorldArena 2.0: Extending Embodied World Model Benchmarking on Modality, Functionality and Platform

DGX agent

arXiv:2605.17912v1 Announce Type: cross Abstract: World models have emerged as a central paradigm for embodied intelligence, enabling agents to predict action-conditioned future and reason about envir

model-releasesarxiv-cs-cv
19 May 2026
Model Releases

Deterministic Event-Graph Substrates as World Models for Counterfactual Reasoning

DGX agent

arXiv:2605.15967v1 Announce Type: new Abstract: We study event-graph substrates: a class of world models that represent agent state as an append-only log of typed RDF triples and answer counterfactual

model-releasesarxiv-cs-ai
18 May 2026
Research

DRS-GUI: Dynamic Region Search for Training-Free GUI Grounding

DGX agent

arXiv:2605.15542v1 Announce Type: new Abstract: GUI agents powered by Multimodal Large Language Models (MLLMs) have demonstrated impressive capability in understanding and executing user instructions.

researcharxiv-cs-ai
18 May 2026
Model Releases

Hybrid LLM-based Intelligent Framework for Robot Task Scheduling

DGX agent

arXiv:2605.15486v1 Announce Type: cross Abstract: This study introduces intelligent frameworks that use Large Language Models (LLMs) to improve task scheduling for construction robots. The LLM is fed

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

Improved Bounds for Reward-Agnostic and Reward-Free Exploration

DGX agent

arXiv:2602.16363v2 Announce Type: replace Abstract: We study reward-free and reward-agnostic exploration in episodic finite-horizon Markov decision processes (MDPs), where an agent explores an unknown

model-releasesarxiv-cs-lg
18 May 2026
Model Releases

Large Language Models as Optimization Controllers: Adaptive Continuation for SIMP Topology Optimization

DGX agent

arXiv:2603.25099v2 Announce Type: replace-cross Abstract: We present a framework in which a large language model (LLM) acts as an online adaptive controller for SIMP topology optimization, replacing c

model-releasesarxiv-cs-ai
18 May 2026
Safety

Learning Bilevel Policies over Symbolic World Models for Long-Horizon Planning

DGX agent

arXiv:2605.15975v1 Announce Type: new Abstract: We tackle the challenge of building embodied AI agents that can reliably solve long-horizon planning problems. Imitation learning from demonstrations ha

safetyarxiv-cs-ai
18 May 2026
Model Releases

Minerva-Ego: Spatiotemporal Hints for Egocentric Video Understanding

DGX agent

arXiv:2605.15342v1 Announce Type: new Abstract: Video reasoning models are a core component of egocentric and embodied agents. However, standard benchmarks for assessing models provide only evaluation

model-releasesarxiv-cs-cv
18 May 2026
Safety

ScreenSearch: Uncertainty-Aware OS Exploration

DGX agent

arXiv:2605.16024v1 Announce Type: new Abstract: Desktop GUI agents operate under partial observability: visually similar screens can correspond to different underlying workflow states, so locally plau

safetyarxiv-cs-ai
18 May 2026
Model Releases

And, yes, our experiments used a mix of GPT-4 & GPT-4o (publishing takes awhile). I think we would see much larger results with more recent …

DGX agent

And, yes, our experiments used a mix of GPT-4 & GPT-4o (publishing takes awhile). I think we would see much larger results with more recent models, let alone recent agentic tools. 'The Cybernetic Team

model-releasesethan-mollick--x
17 May 2026
Model Releases

AI radio hosts demonstrate why AI can’t be trusted alone

DGX agent

Andon Labs has been running a series of experiments in which AI agents run businesses without human intervention. Its latest is a quartet of radio stations run by some of the most popular AI models ou

model-releasesthe-verge-ai
15 May 2026
Model Releases

Anthropic just went after the 44% of U.S. GDP that enterprise AI has mostly ignored. Claude for Small Business launched this week with 15 pr…

DGX agent

Anthropic just went after the 44% of U.S. GDP that enterprise AI has mostly ignored. Claude for Small Business launched this week with 15 prebuilt agentic workflows and 15 skills connected directly in

model-releasesallie-k--miller--x
15 May 2026
← Previous
1…293294295296297…374
Next →