AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,771 results
25 May 2026

SCRIPT: Scalable Diffusion Policy with Multi-stage Training for Language-driven Physics-Based Humanoid Control

SafetyDGX agent

arXiv:2605.22894v1 Announce Type: cross Abstract: Controlling physics-based humanoids from natural-language instructions is a critical step toward general-purpose embodied agents. However, existing me

23 May 2026

Learn anything with our new /lesson-generator skill

Model ReleasesDGX agent

Learn anything with our new /lesson-generator skill Just released my new /lesson-generator skill. Use it with your agent to learn anything: - generate lessons/courses on any topic - include nano-banan

LiteCoOp: Lightweight Multi-LLM Shared-Tree Reasoning for Model-Serving Compiler Optimizations

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Hardware
DGX agent

arXiv:2602.01935v2 Announce Type: replace Abstract: LLM-guided compiler optimization has recently shown promise, but existing approaches rely on a single large LLM throughout search, making them expen

22 May 2026

Catch up on the Dialogues stage at Google I/O 2026.

Model ReleasesDGX agent

The Dialogues stage at Google I/O 2026 brought together Google leaders, scientific minds and creative visionaries to discuss technological breakthroughs. Featured discussions included AI agents and pr

Cursor Composer 2.5's is 3–18x cheaper than Opus 4.7 in Claude Code (medium reasoning), and 5–32x cheaper than GPT-5.5 in Codex (medium) bas…

Model ReleasesDGX agent

Cursor Composer 2.5's is 3–18x cheaper than Opus 4.7 in Claude Code (medium reasoning), and 5–32x cheaper than GPT-5.5 in Codex (medium) based on API pricing This low Cost per Task isn't just driven b

DeepWeb-Bench: A Deep Research Benchmark Demanding Massive Cross-Source Evidence and Long-Horizon Derivation

Model ReleasesDGX agent

arXiv:2605.21482v1 Announce Type: new Abstract: Deep research, in which an agent searches the open web, collects evidence, and derives an answer through extended reasoning, is a prominent use case for

DeferMem: Query-Time Evidence Distillation via Reinforcement Learning for Long-Term Memory QA

ResearchDGX agent

arXiv:2605.22411v1 Announce Type: new Abstract: Large language model (LLM) agents still struggle with long-term memory question answering, where answer-supporting evidence is often scattered across lo

Interpreting and Enhancing Emotional Circuits in Large Vision-Language Models via Cross-Modal Information Flow

ResearchDGX agent

arXiv:2605.21980v1 Announce Type: new Abstract: Large Vision-Language Models (LVLMs) represent a significant leap towards empathetic agents, demonstrating remarkable capabilities in emotion understand

Join @hwchase17 + @traversal_ai founder @_anish_agarwal for a technical fireside chat in NYC on June 2nd. Learn how Traversal builds, ships …

TutorialsDGX agent

Join @hwchase17 + @traversal_ai founder @_anish_agarwal for a technical fireside chat in NYC on June 2nd. Learn how Traversal builds, ships and improves their agents. Enjoy networking, drinks, and a l

LongVT: Incentivizing 'Thinking with Long Videos' via Native Tool Calling

Model ReleasesDGX agent

arXiv:2511.20785v3 Announce Type: replace Abstract: Large multimodal models (LMMs) have shown great potential for video reasoning with textual Chain-of-Thought. However, they remain vulnerable to hall

Runaway token costs and sovereignty concerns are driving enterprises back to the desktop

Local AiDGX agent

The AI PC is being fundamentally redefined as agentic workloads push the boundaries of what local compute can deliver — and as runaway cloud token costs force enterprises to rethink where inference ac

stable-worldmodel: A Platform for Reproducible World Modeling Research and Evaluation

ResearchDGX agent

arXiv:2605.21800v1 Announce Type: cross Abstract: World models are central to building agents that can reason, plan, and generalize beyond their training data. However, research on world models is cur

Steins;Gate Drive: Semantic Safety Arbitration over Structured Futures for Latency-Decoupled LLM Planning

Model ReleasesDGX agent

arXiv:2605.22456v1 Announce Type: new Abstract: Cloud-hosted LLM driver agents provide useful semantic judgments, but their inference latency exceeds stepwise vehicle-control windows. Learned world mo

The $58,000 TV bill: When DirecTV sued O.J. Simpson for piracy

IndustryDGX agent

O.J. Simpson was ordered to pay $25,000 in damages for pirating satellite television signals from DirecTV using illegal devices known as 'bootloaders.' Federal agents seized the illegal devices from S

21 May 2026

AI-Assisted Scientific Assessment: A Case Study on Climate Change

Model ReleasesDGX agent

arXiv:2602.09723v2 Announce Type: replace Abstract: The emerging paradigm of AI co-scientists focuses on tasks characterized by repeatable verification, where agents explore search spaces in 'guess an

API Keys Are Open Secrets

Model ReleasesDGX agent

Today, AI services rely heavily on API keys. To run AI agents, users provide API keys that signify paid tokens, subscriptions, or paid accounts. While API keys are easy to use, it is just as easy to u

Break the context window barrier with Amazon Bedrock AgentCore

Model ReleasesDGX agent

In this post, you will learn how to implement Recursive Language Models (RLM) using Amazon Bedrock AgentCore Code Interpreter and the Strands Agents SDK. By the end, you will know how to process docum

Full interview with Sundar is now live! https://x.com/rowancheung/status/2057491344697012384?s=20

IndustryDGX agent

Full interview with Sundar is now live! https://x.com/rowancheung/status/2057491344697012384?s=20 Google just revealed Omni, personalized cross-device intelligence, and Spark agents at I/O 2025. I sat

GenAI-Driven Threat Detection with Microsoft Security Copilot

Model ReleasesDGX agent

arXiv:2605.20896v1 Announce Type: cross Abstract: Defending against today's increasingly sophisticated cyberattacks requires security analysts to continuously translate evolving attacker tradecraft in

Interaction Locality in Hierarchical Recursive Reasoning

Local AiDGX agent

arXiv:2605.20784v1 Announce Type: cross Abstract: Spatial reasoning requires both location-bound computation and location-invariant structure: agents must make local moves while preserving route, obje

New VIDEO: From LLM Wikis to LLM Artifacts Shared all my thoughts on why LLM wikis and HTML artifacts are a big deal. Plus, new tools to hel…

ResearchDGX agent

New VIDEO: From LLM Wikis to LLM Artifacts Shared all my thoughts on why LLM wikis and HTML artifacts are a big deal. Plus, new tools to help you build wikis and artifacts with agents. Just getting st

20 May 2026

100 things we announced at I/O 2026

Model ReleasesDGX agent

Google I/O 2026 unveiled new models, agents and tools to help users build, search, create, discover, shop and get more done. Key announcements included Gemini Omni, Google Antigravity, and Universal C

Announcing OpenAI-compatible API support for Amazon SageMaker AI endpoints

IndustryDGX agent

Today, Amazon SageMaker AI introduces OpenAI-compatible API support for real-time inference endpoints. If you use the OpenAI SDK, LangChain, or Strands Agents, you can now invoke models on SageMaker A

Anyone understand what Google mean by 'Gemini Spark runs on Gemini 3.5 and uses the Antigravity harness' - is 'Antigravity' a generic term t…

Model ReleasesDGX agent

Anyone understand what Google mean by 'Gemini Spark runs on Gemini 3.5 and uses the Antigravity harness' - is 'Antigravity' a generic term they're using for their agent harnesses now or is their Claw-

AutoResearchClaw: Self-Reinforcing Autonomous Research with Human-AI Collaboration

Model ReleasesDGX agent

arXiv:2605.20025v1 Announce Type: new Abstract: Automating scientific discovery requires more than generating papers from ideas. Real research is iterative: hypotheses are challenged from multiple per

Build real-time voice applications with Amazon SageMaker AI and vLLM

IndustryDGX agent

Voice agents, live captioning, contact center analytics, and accessibility tools all depend on real-time speech-to-text, where your application streams audio in and receives transcription back simulta

CaptchaMind: Training CAPTCHA Solvers via Reinforcement Learning with Explicit Reasoning Supervision

Model ReleasesDGX agent

arXiv:2605.19538v1 Announce Type: cross Abstract: CAPTCHAs are widely deployed as human verification mechanisms and frequently block intelligent agents from completing end-to-end automation in real-wo

Decoding Rewards in Competitive Games: Inverse Game Theory with Entropy Regularization

TutorialsDGX agent

arXiv:2601.12707v2 Announce Type: replace Abstract: Estimating the unknown reward functions driving agents' behaviors is of central interest in inverse reinforcement learning and game theory. To tackl

it’s in gemini, just create it in ai studio. oh, that’s for your personal google one account. for workspace you need gemini business. no, no…

Model ReleasesDGX agent

it’s in gemini, just create it in ai studio. oh, that’s for your personal google one account. for workspace you need gemini business. no, not gemini advanced, that’s ai pro now. unless you need ai ult

Learning Efficient Guardrails for Compliance

Model ReleasesDGX agent

arXiv:2510.03485v2 Announce Type: replace Abstract: Autonomous web agents are increasingly deployed for long-horizon tasks, yet their ability to adhere to real-world policies remains critically undere

// Memory as a Model // The paper augments any LLM with a separate trained memory model that stores, retrieves, and integrates facts on its …

Model ReleasesDGX agent

// Memory as a Model // The paper augments any LLM with a separate trained memory model that stores, retrieves, and integrates facts on its behalf. It decouples memory updates from base-model weight u

P2DNav: Panorama-to-Downview Reasoning for Zero-shot Vision-and-Language Navigation

Model ReleasesDGX agent

arXiv:2605.19634v1 Announce Type: cross Abstract: Vision-and-language navigation (VLN) requires an embodied agent to ground natural-language instructions into executable navigation actions in unseen e

Reporting from Google I/O 2026 with the four biggest themes from one of the biggest AI labs in the world. 🎤 Voice AI as an interface Google…

Model ReleasesDGX agent

Reporting from Google I/O 2026 with the four biggest themes from one of the biggest AI labs in the world. 🎤 Voice AI as an interface Google and Samsung announced new Gemini-powered glasses with Gentle

Sampling-Based Safe Reinforcement Learning

SafetyDGX agent

arXiv:2605.19469v1 Announce Type: cross Abstract: Safe exploration remains a fundamental challenge in reinforcement learning (RL), limiting the deployment of RL agents in the real world. We propose Sa

Toward an AI-Powered Computational Testbed for Workforce Policy

SafetyDGX agent

arXiv:2605.19064v1 Announce Type: cross Abstract: Workforce transformations are difficult to forecast and costly to mismanage. In particular, the integration of artificial intelligence into knowledge

We added 600+ new voices on Together AI! Introducing MiniMax Speech 2.8 Turbo on Together AI, an enterprise TTS model for expressive real-ti…

ApplicationsDGX agent

We added 600+ new voices on Together AI! Introducing MiniMax Speech 2.8 Turbo on Together AI, an enterprise TTS model for expressive real-time voice agents. AI natives can now deploy @MiniMax_AI Speec

19 May 2026

A Structural Threshold in Decision Capacity Governs Collapse in Self-Play Reinforcement Learning

ResearchDGX agent

arXiv:2605.16315v1 Announce Type: cross Abstract: We show that a threshold in decision capacity determines whether self-play reinforcement learning agents collapse under asymmetric rule perturbations.

ARROW: Augmented Replay for RObust World models

SafetyDGX agent

arXiv:2603.11395v2 Announce Type: replace-cross Abstract: Continual reinforcement learning challenges agents to acquire new skills while retaining previously learned ones with the goal of improving pe

Assured autonomy: How operations research powers and orchestrates generative AI systems

SafetyDGX agent

arXiv:2512.23978v2 Announce Type: replace Abstract: Generative artificial intelligence (GenAI) is shifting from conversational assistants toward agentic systems -- autonomous decision-making systems t

Automated Root-Cause Subclassification and No-Code Fix Generation for Invalid Bug Reports

Model ReleasesDGX agent

arXiv:2605.17561v1 Announce Type: cross Abstract: Issues faced when using software are reported in the form of bug reports. However, many bug reports are invalid, meaning they do not require code chan

CommitDistill: A Lightweight Knowledge-Centric Memory Layer for Software Repositories

Model ReleasesDGX agent

arXiv:2605.18284v1 Announce Type: cross Abstract: Software repositories accumulate large amounts of unstructured knowledge in commit messages, pull-request discussions, and issue threads, but develope

DecoupleSearch: Decouple Planning and Search via Hierarchical Reward Modeling

Model ReleasesDGX agent

arXiv:2510.21712v2 Announce Type: replace-cross Abstract: Retrieval-Augmented Generation (RAG) systems have emerged as a pivotal methodology for enhancing Large Language Models (LLMs) through the dyna

DISA: Offline Importance Sampling for Distribution-Matching LLM-RL

SafetyDGX agent

arXiv:2605.17295v1 Announce Type: cross Abstract: Modern reasoning agents are increasingly evaluated on their ability to generate multiple valid solution paths, plans, or tool-use traces for a given i

EPIC-Bench: A Perception-Centric Benchmark for Fine-Grained Embodied Visual Grounding in Vision-Language Models

Model ReleasesDGX agent

arXiv:2605.17070v1 Announce Type: new Abstract: While large vision-language models (VLMs) are increasingly adopted as the perceptual backbone for embodied agents, existing benchmarks often rely on que

Excited for this conversation with @hwchase17 , Co-Founder & CEO of @LangChain , on June 2 at the @traversal_ai office in NYC. We’ll be talk…

ApplicationsDGX agent

Excited for this conversation with @hwchase17 , Co-Founder & CEO of @LangChain , on June 2 at the @traversal_ai office in NYC. We’ll be talking about what it actually takes to operate AI agents in pro

Firefly: Illuminating Large-Scale Verified Tool-Call Data Generation from Real APIs

Model ReleasesDGX agent

arXiv:2605.17558v1 Announce Type: cross Abstract: Training tool-calling agents requires large-scale trajectory data with verifiable labels, yet existing approaches either synthesize environments that

Fix the Structural Bottleneck: Context Compression via Explicit Information Transmission

Model ReleasesDGX agent

arXiv:2602.03784v2 Announce Type: replace Abstract: Long-context LLM agents often struggle with growing token, memory, and latency costs, making efficient context compression essential for practical d

Generating Realistic Safety-Critical Scenarios for Vehicle-Pedestrian Interactions

SafetyDGX agent

arXiv:2605.17229v1 Announce Type: new Abstract: Automated driving system deployment requires rigorous validation across safety-critical vehicle-pedestrian interactions, yet real-world datasets rarely

GRID: Graph Representation of Intelligence Data for Security Text Knowledge Graph Construction

ResearchDGX agent

arXiv:2605.16714v1 Announce Type: new Abstract: Security knowledge graphs can provide computable external memory for security agents, but constructing them from long-form cyber threat intelligence (CT

HydroAgent: Closing the Gap Between Frontier LLMs and Human Experts in Hydrologic Model Calibration via Simulator-Grounded RL

Model ReleasesDGX agent

arXiv:2605.17792v1 Announce Type: new Abstract: Calibrating distributed hydrologic models is a critical bottleneck across operational water resources management - streamflow prediction, reservoir oper

Internalizing Tool Knowledge in Small Language Models via QLoRA Fine-Tuning

Model ReleasesDGX agent

arXiv:2605.17774v1 Announce Type: new Abstract: Large language models are increasingly used as planning components in agentic systems, but current tool-use pipelines often require full tool schemas to

Monocular Open Vocabulary Occupancy Prediction for Indoor Scenes

Model ReleasesDGX agent

arXiv:2602.22667v2 Announce Type: replace Abstract: Open-vocabulary 3D occupancy is vital for embodied agents, which need to understand complex indoor environments where semantic categories are abunda

MORN: Metacognitive Object-Goal Regulation for Resource-Rational Long-Horizon Navigation

ResearchDGX agent

arXiv:2605.16932v1 Announce Type: new Abstract: Robots deployed in unstructured human environments must frequently execute long-horizon missions, such as find the mug, then the chair, then the printer

Non-Colliding Biometric Identities for Digital Entities: Geometry, Capacity, and Million-Scale Virtual Identity Provisioning

ResearchDGX agent

arXiv:2605.18238v1 Announce Type: new Abstract: Digital entities such as AI agents and humanoid robots increasingly operate alongside real humans, yet their identity infrastructure is based on credent

Not What You Asked For: Typographic Attacks in Household Robot Manipulation

Model ReleasesDGX agent

arXiv:2605.18593v1 Announce Type: cross Abstract: Open-vocabulary embodied AI agents increasingly rely on vision-language models such as CLIP for object perception and task grounding. However, the sha

OxyGen: Unified KV Cache Management for VLA Inference under Multi-Task Parallelism

Local AiDGX agent

arXiv:2603.14371v2 Announce Type: replace-cross Abstract: Embodied AI agents increasingly require parallel execution of multiple tasks, such as manipulation, conversation, and memory construction, fro

Principles of frugal inference and control

Local AiDGX agent

arXiv:2406.14427v4 Announce Type: replace Abstract: A central challenge for intelligent agents in an uncertain world is striking the right balance between utility maximization and resource use, not on

Privacy Preserving Reinforcement Learning with One-Sided Feedback

Model ReleasesDGX agent

arXiv:2605.18246v1 Announce Type: cross Abstract: We study reinforcement learning (RL) in multi-dimensional continuous state and action spaces with one-sided feedback, where the agent receives partial

SD-Search: On-Policy Hindsight Self-Distillation for Search-Augmented Reasoning

SafetyDGX agent

arXiv:2605.18299v1 Announce Type: new Abstract: Search-augmented reasoning agents interleave internal reasoning with calls to an external retriever, and their performance relies on the quality of each

Starve to Perceive: Taming Lazy Perception in VLMs with Constrained Visual Bandwidth

TutorialsDGX agent

arXiv:2605.18603v1 Announce Type: new Abstract: Vision-Language Models (VLMs) deployed as situated agents in high-resolution visual environments require active perception -- the ability to dynamically

← Previous
1…232233234235236…297
Next →