AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlog
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,920 results
Model Releases

it’s in gemini, just create it in ai studio. oh, that’s for your personal google one account. for workspace you need gemini business. no, no…

DGX agent

it’s in gemini, just create it in ai studio. oh, that’s for your personal google one account. for workspace you need gemini business. no, not gemini advanced, that’s ai pro now. unless you need ai ult

model-releasessimon-willison--x
20 May 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

Learning Efficient Guardrails for Compliance

DGX agent

arXiv:2510.03485v2 Announce Type: replace Abstract: Autonomous web agents are increasingly deployed for long-horizon tasks, yet their ability to adhere to real-world policies remains critically undere

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

// Memory as a Model // The paper augments any LLM with a separate trained memory model that stores, retrieves, and integrates facts on its …

DGX agent

// Memory as a Model // The paper augments any LLM with a separate trained memory model that stores, retrieves, and integrates facts on its behalf. It decouples memory updates from base-model weight u

model-releasesdair-ai--x
20 May 2026
Model Releases

P2DNav: Panorama-to-Downview Reasoning for Zero-shot Vision-and-Language Navigation

DGX agent

arXiv:2605.19634v1 Announce Type: cross Abstract: Vision-and-language navigation (VLN) requires an embodied agent to ground natural-language instructions into executable navigation actions in unseen e

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Reporting from Google I/O 2026 with the four biggest themes from one of the biggest AI labs in the world. 🎤 Voice AI as an interface Google…

DGX agent

Reporting from Google I/O 2026 with the four biggest themes from one of the biggest AI labs in the world. 🎤 Voice AI as an interface Google and Samsung announced new Gemini-powered glasses with Gentle

model-releasesallie-k--miller--x
20 May 2026
Safety

Sampling-Based Safe Reinforcement Learning

DGX agent

arXiv:2605.19469v1 Announce Type: cross Abstract: Safe exploration remains a fundamental challenge in reinforcement learning (RL), limiting the deployment of RL agents in the real world. We propose Sa

safetyarxiv-cs-ai
20 May 2026
Safety

Toward an AI-Powered Computational Testbed for Workforce Policy

DGX agent

arXiv:2605.19064v1 Announce Type: cross Abstract: Workforce transformations are difficult to forecast and costly to mismanage. In particular, the integration of artificial intelligence into knowledge

safetyarxiv-cs-ai
20 May 2026
Applications

We added 600+ new voices on Together AI! Introducing MiniMax Speech 2.8 Turbo on Together AI, an enterprise TTS model for expressive real-ti…

DGX agent

We added 600+ new voices on Together AI! Introducing MiniMax Speech 2.8 Turbo on Together AI, an enterprise TTS model for expressive real-time voice agents. AI natives can now deploy @MiniMax_AI Speec

applicationstogether-ai--x
20 May 2026
Research

A Structural Threshold in Decision Capacity Governs Collapse in Self-Play Reinforcement Learning

DGX agent

arXiv:2605.16315v1 Announce Type: cross Abstract: We show that a threshold in decision capacity determines whether self-play reinforcement learning agents collapse under asymmetric rule perturbations.

researcharxiv-cs-ai
19 May 2026
Safety

ARROW: Augmented Replay for RObust World models

DGX agent

arXiv:2603.11395v2 Announce Type: replace-cross Abstract: Continual reinforcement learning challenges agents to acquire new skills while retaining previously learned ones with the goal of improving pe

safetyarxiv-cs-ai
19 May 2026
Safety

Assured autonomy: How operations research powers and orchestrates generative AI systems

DGX agent

arXiv:2512.23978v2 Announce Type: replace Abstract: Generative artificial intelligence (GenAI) is shifting from conversational assistants toward agentic systems -- autonomous decision-making systems t

safetyarxiv-cs-lg
19 May 2026
Model Releases

Automated Root-Cause Subclassification and No-Code Fix Generation for Invalid Bug Reports

DGX agent

arXiv:2605.17561v1 Announce Type: cross Abstract: Issues faced when using software are reported in the form of bug reports. However, many bug reports are invalid, meaning they do not require code chan

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

CommitDistill: A Lightweight Knowledge-Centric Memory Layer for Software Repositories

DGX agent

arXiv:2605.18284v1 Announce Type: cross Abstract: Software repositories accumulate large amounts of unstructured knowledge in commit messages, pull-request discussions, and issue threads, but develope

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

DecoupleSearch: Decouple Planning and Search via Hierarchical Reward Modeling

DGX agent

arXiv:2510.21712v2 Announce Type: replace-cross Abstract: Retrieval-Augmented Generation (RAG) systems have emerged as a pivotal methodology for enhancing Large Language Models (LLMs) through the dyna

model-releasesarxiv-cs-ai
19 May 2026
Safety

DISA: Offline Importance Sampling for Distribution-Matching LLM-RL

DGX agent

arXiv:2605.17295v1 Announce Type: cross Abstract: Modern reasoning agents are increasingly evaluated on their ability to generate multiple valid solution paths, plans, or tool-use traces for a given i

safetyarxiv-cs-cl
19 May 2026
Model Releases

EPIC-Bench: A Perception-Centric Benchmark for Fine-Grained Embodied Visual Grounding in Vision-Language Models

DGX agent

arXiv:2605.17070v1 Announce Type: new Abstract: While large vision-language models (VLMs) are increasingly adopted as the perceptual backbone for embodied agents, existing benchmarks often rely on que

model-releasesarxiv-cs-cv
19 May 2026
Applications

Excited for this conversation with @hwchase17 , Co-Founder & CEO of @LangChain , on June 2 at the @traversal_ai office in NYC. We’ll be talk…

DGX agent

Excited for this conversation with @hwchase17 , Co-Founder & CEO of @LangChain , on June 2 at the @traversal_ai office in NYC. We’ll be talking about what it actually takes to operate AI agents in pro

applicationsharrison-chase--x
19 May 2026
Model Releases

Firefly: Illuminating Large-Scale Verified Tool-Call Data Generation from Real APIs

DGX agent

arXiv:2605.17558v1 Announce Type: cross Abstract: Training tool-calling agents requires large-scale trajectory data with verifiable labels, yet existing approaches either synthesize environments that

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

Fix the Structural Bottleneck: Context Compression via Explicit Information Transmission

DGX agent

arXiv:2602.03784v2 Announce Type: replace Abstract: Long-context LLM agents often struggle with growing token, memory, and latency costs, making efficient context compression essential for practical d

model-releasesarxiv-cs-cl
19 May 2026
Safety

Generating Realistic Safety-Critical Scenarios for Vehicle-Pedestrian Interactions

DGX agent

arXiv:2605.17229v1 Announce Type: new Abstract: Automated driving system deployment requires rigorous validation across safety-critical vehicle-pedestrian interactions, yet real-world datasets rarely

safetyarxiv-cs-ro
19 May 2026
Research

GRID: Graph Representation of Intelligence Data for Security Text Knowledge Graph Construction

DGX agent

arXiv:2605.16714v1 Announce Type: new Abstract: Security knowledge graphs can provide computable external memory for security agents, but constructing them from long-form cyber threat intelligence (CT

researcharxiv-cs-ai
19 May 2026
Model Releases

HydroAgent: Closing the Gap Between Frontier LLMs and Human Experts in Hydrologic Model Calibration via Simulator-Grounded RL

DGX agent

arXiv:2605.17792v1 Announce Type: new Abstract: Calibrating distributed hydrologic models is a critical bottleneck across operational water resources management - streamflow prediction, reservoir oper

model-releasesarxiv-cs-lg
19 May 2026
Model Releases

Internalizing Tool Knowledge in Small Language Models via QLoRA Fine-Tuning

DGX agent

arXiv:2605.17774v1 Announce Type: new Abstract: Large language models are increasingly used as planning components in agentic systems, but current tool-use pipelines often require full tool schemas to

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

Monocular Open Vocabulary Occupancy Prediction for Indoor Scenes

DGX agent

arXiv:2602.22667v2 Announce Type: replace Abstract: Open-vocabulary 3D occupancy is vital for embodied agents, which need to understand complex indoor environments where semantic categories are abunda

model-releasesarxiv-cs-cv
19 May 2026
Research

MORN: Metacognitive Object-Goal Regulation for Resource-Rational Long-Horizon Navigation

DGX agent

arXiv:2605.16932v1 Announce Type: new Abstract: Robots deployed in unstructured human environments must frequently execute long-horizon missions, such as find the mug, then the chair, then the printer

researcharxiv-cs-ro
19 May 2026
Research

Non-Colliding Biometric Identities for Digital Entities: Geometry, Capacity, and Million-Scale Virtual Identity Provisioning

DGX agent

arXiv:2605.18238v1 Announce Type: new Abstract: Digital entities such as AI agents and humanoid robots increasingly operate alongside real humans, yet their identity infrastructure is based on credent

researcharxiv-cs-cv
19 May 2026
Model Releases

Not What You Asked For: Typographic Attacks in Household Robot Manipulation

DGX agent

arXiv:2605.18593v1 Announce Type: cross Abstract: Open-vocabulary embodied AI agents increasingly rely on vision-language models such as CLIP for object perception and task grounding. However, the sha

model-releasesarxiv-cs-ai
19 May 2026
Local Ai

OxyGen: Unified KV Cache Management for VLA Inference under Multi-Task Parallelism

DGX agent

arXiv:2603.14371v2 Announce Type: replace-cross Abstract: Embodied AI agents increasingly require parallel execution of multiple tasks, such as manipulation, conversation, and memory construction, fro

local-aiarxiv-cs-ai
19 May 2026
Local Ai

Principles of frugal inference and control

DGX agent

arXiv:2406.14427v4 Announce Type: replace Abstract: A central challenge for intelligent agents in an uncertain world is striking the right balance between utility maximization and resource use, not on

local-aiarxiv-cs-ai
19 May 2026
Model Releases

Privacy Preserving Reinforcement Learning with One-Sided Feedback

DGX agent

arXiv:2605.18246v1 Announce Type: cross Abstract: We study reinforcement learning (RL) in multi-dimensional continuous state and action spaces with one-sided feedback, where the agent receives partial

model-releasesarxiv-cs-ai
19 May 2026
Safety

SD-Search: On-Policy Hindsight Self-Distillation for Search-Augmented Reasoning

DGX agent

arXiv:2605.18299v1 Announce Type: new Abstract: Search-augmented reasoning agents interleave internal reasoning with calls to an external retriever, and their performance relies on the quality of each

safetyarxiv-cs-ai
19 May 2026
Tutorials

Starve to Perceive: Taming Lazy Perception in VLMs with Constrained Visual Bandwidth

DGX agent

arXiv:2605.18603v1 Announce Type: new Abstract: Vision-Language Models (VLMs) deployed as situated agents in high-resolution visual environments require active perception -- the ability to dynamically

tutorialsarxiv-cs-cv
19 May 2026
Model Releases

Today, we launched a brand-new intelligent Search box. Here's what that means: An upgrade to the Search experience with our most advanced Ge…

DGX agent

Today, we launched a brand-new intelligent Search box. Here's what that means: An upgrade to the Search experience with our most advanced Gemini 3.5 models, bringing with them our latest agentic capab

model-releasesgoogle-ai--x
19 May 2026
Local Ai

Use your LM Studio models to code locally in @zeddotdev 🚀

DGX agent

Use your LM Studio models to code locally in @zeddotdev 🚀 Local model usage grew 3x in Zed's agent in the last 10 weeks. Cameron Mcloughlin on why he prefers local: 'I worry about over-reliance on pro

local-ailm-studio--x
19 May 2026
Applications

What Drives Success in Physical Planning with Joint-Embedding Predictive World Models?

DGX agent

arXiv:2512.24497v3 Announce Type: replace Abstract: A long-standing challenge in AI is to develop agents capable of solving a wide range of physical tasks and generalizing to new, unseen tasks and env

applicationsarxiv-cs-ai
19 May 2026
Model Releases

When Outcome Looks Right But Discipline Fails: Trace-Based Evaluation Under Hidden Competitor State

DGX agent

arXiv:2605.18580v1 Announce Type: new Abstract: Outcome-only evaluation can certify economically unsafe agents: a policy can hit a business KPI while violating deployable behavioral discipline. In hot

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

WinDeskGround: A Benchmark for Robust GUI Grounding in Complex Multi-Window Desktop Environments

DGX agent

arXiv:2605.16402v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) have revolutionized GUI automation, yet their efficacy is largely established on idealized, single-layer interf

model-releasesarxiv-cs-cv
19 May 2026
Model Releases

WorldArena 2.0: Extending Embodied World Model Benchmarking on Modality, Functionality and Platform

DGX agent

arXiv:2605.17912v1 Announce Type: cross Abstract: World models have emerged as a central paradigm for embodied intelligence, enabling agents to predict action-conditioned future and reason about envir

model-releasesarxiv-cs-cv
19 May 2026
Model Releases

Deterministic Event-Graph Substrates as World Models for Counterfactual Reasoning

DGX agent

arXiv:2605.15967v1 Announce Type: new Abstract: We study event-graph substrates: a class of world models that represent agent state as an append-only log of typed RDF triples and answer counterfactual

model-releasesarxiv-cs-ai
18 May 2026
Research

DRS-GUI: Dynamic Region Search for Training-Free GUI Grounding

DGX agent

arXiv:2605.15542v1 Announce Type: new Abstract: GUI agents powered by Multimodal Large Language Models (MLLMs) have demonstrated impressive capability in understanding and executing user instructions.

researcharxiv-cs-ai
18 May 2026
Model Releases

Hybrid LLM-based Intelligent Framework for Robot Task Scheduling

DGX agent

arXiv:2605.15486v1 Announce Type: cross Abstract: This study introduces intelligent frameworks that use Large Language Models (LLMs) to improve task scheduling for construction robots. The LLM is fed

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

Improved Bounds for Reward-Agnostic and Reward-Free Exploration

DGX agent

arXiv:2602.16363v2 Announce Type: replace Abstract: We study reward-free and reward-agnostic exploration in episodic finite-horizon Markov decision processes (MDPs), where an agent explores an unknown

model-releasesarxiv-cs-lg
18 May 2026
Model Releases

Large Language Models as Optimization Controllers: Adaptive Continuation for SIMP Topology Optimization

DGX agent

arXiv:2603.25099v2 Announce Type: replace-cross Abstract: We present a framework in which a large language model (LLM) acts as an online adaptive controller for SIMP topology optimization, replacing c

model-releasesarxiv-cs-ai
18 May 2026
Safety

Learning Bilevel Policies over Symbolic World Models for Long-Horizon Planning

DGX agent

arXiv:2605.15975v1 Announce Type: new Abstract: We tackle the challenge of building embodied AI agents that can reliably solve long-horizon planning problems. Imitation learning from demonstrations ha

safetyarxiv-cs-ai
18 May 2026
Model Releases

Minerva-Ego: Spatiotemporal Hints for Egocentric Video Understanding

DGX agent

arXiv:2605.15342v1 Announce Type: new Abstract: Video reasoning models are a core component of egocentric and embodied agents. However, standard benchmarks for assessing models provide only evaluation

model-releasesarxiv-cs-cv
18 May 2026
Safety

ScreenSearch: Uncertainty-Aware OS Exploration

DGX agent

arXiv:2605.16024v1 Announce Type: new Abstract: Desktop GUI agents operate under partial observability: visually similar screens can correspond to different underlying workflow states, so locally plau

safetyarxiv-cs-ai
18 May 2026
Model Releases

And, yes, our experiments used a mix of GPT-4 & GPT-4o (publishing takes awhile). I think we would see much larger results with more recent …

DGX agent

And, yes, our experiments used a mix of GPT-4 & GPT-4o (publishing takes awhile). I think we would see much larger results with more recent models, let alone recent agentic tools. 'The Cybernetic Team

model-releasesethan-mollick--x
17 May 2026
Model Releases

AI radio hosts demonstrate why AI can’t be trusted alone

DGX agent

Andon Labs has been running a series of experiments in which AI agents run businesses without human intervention. Its latest is a quartet of radio stations run by some of the most popular AI models ou

model-releasesthe-verge-ai
15 May 2026
← Previous
1…293294295296297…374
Next →