AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlog
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,920 results
Model Releases

Hadamard Representation: Scaffolding Performance Across Model-free RL

DGX agent

arXiv:2406.09079v5 Announce Type: replace Abstract: Deep reinforcement learning agents progressively lose representational capacity during training: neurons become dormant, removing active capacity fr

model-releasesarxiv-cs-lg
26 May 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Safety

How to Mitigate the Distribution Shift Problem in Robotics Control: A Robust and Adaptive Approach Based on Offline to Online Imitation Learning

DGX agent

arXiv:2605.25414v1 Announce Type: new Abstract: Distribution shift in imitation learning refers to the problem that the agent cannot plan proper actions for a state that has not been visited during th

safetyarxiv-cs-ro
26 May 2026
Applications

In Search of the Ingredients of Open-Endedness: Replicating Picbreeder with Large Vision-Language Models

DGX agent

arXiv:2605.23908v1 Announce Type: new Abstract: We are in the midst of large-scale industrial and academic efforts to automate the processes of scientific, technological and creative production throug

applicationsarxiv-cs-ai
26 May 2026
Safety

JT-SAFE-V2: Safety-by-Design Foundation Model with World-Context Data

DGX agent

arXiv:2605.24414v1 Announce Type: new Abstract: We introduce JT-Safe-V2, a large language model designed to advance the safety and trustworthiness of foundation models, extending our previous JT-Safe

safetyarxiv-cs-ai
26 May 2026
Safety

Learning in Low-Dimensional Subspaces: Orthogonal Bottlenecks for Reinforcement Learning

DGX agent

arXiv:2605.26012v1 Announce Type: cross Abstract: Deep reinforcement learning (RL) agents commonly rely on high-dimensional neural representations, despite growing evidence that task-relevant value an

safetyarxiv-cs-ai
26 May 2026
Model Releases

NVIDIA Vera CPU Is ‘Packing a Heavy-Hitting Punch’ Against Competition

DGX agent

The shift to agentic AI creates a new CPU requirement for the AI factory: fast cores, massive memory bandwidth and the ability to sustain high performance when all cores are active. Initial benchmark

model-releasesnvidia-blog
26 May 2026
Safety

On the Benefits of Free Exploration for Regret Minimization in Multi-Armed Bandits

DGX agent

arXiv:2605.25789v1 Announce Type: cross Abstract: We study a stochastic multi-armed bandit problem where an agent is granted a free exploration budget before regret accumulates, a setting not captured

safetyarxiv-cs-ai
26 May 2026
Model Releases

ran my first benchmark this weekend (longmemeval) mostly to test activegraph, learned a lot! - this is a stepping stone to show the event ba…

DGX agent

ran my first benchmark this weekend (longmemeval) mostly to test activegraph, learned a lot! - this is a stepping stone to show the event based agent system works. the AI convinced me not to start wit

model-releasesyohei-nakajima--x
26 May 2026
Model Releases

Streaming Reinforcement Learning under Partial Observability with Real-Time Recurrent Learning

DGX agent

arXiv:2605.24709v1 Announce Type: new Abstract: Streaming reinforcement learning has emerged as an online learning paradigm that conforms to the restrictions of natural learning agents that process da

model-releasesarxiv-cs-lg
26 May 2026
Model Releases

Structural Abstraction as an Inductive Bias for Non-Stationary Language Model Training

DGX agent

arXiv:2603.17198v2 Announce Type: replace-cross Abstract: A foundational principle in cognitive science holds that intelligent agents do not learn by storing experiences as isolated instances, but by

model-releasesarxiv-cs-cl
26 May 2026
Model Releases

TS-Skill: A Benchmark for Evaluating Analytical Skills in Time-Series Question Answering

DGX agent

arXiv:2605.24703v1 Announce Type: cross Abstract: Large language models (LLMs) and time-series language models (TSLMs) are increasingly applied to time-series question answering (TSQA). Unlike text-on

model-releasesarxiv-cs-ai
26 May 2026
Research

Turn-Based Structural Triggers: Prompt-Free Backdoors in Multi-Turn LLMs

DGX agent

arXiv:2601.14340v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are widely integrated into interactive systems such as dialogue agents and task-oriented assistants. This growing

researcharxiv-cs-lg
26 May 2026
Model Releases

When Search Becomes Memory: Turning Robot Design Trials into Transferable Skills

DGX agent

arXiv:2605.25832v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used as proposal generators for evolutionary robot design, yet most loops remain memoryless: simulator r

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Design and Report Benchmarks for Knowledge Work

DGX agent

arXiv:2605.23262v1 Announce Type: new Abstract: The development of LLM agents has led to a growing body of work on knowledge-work AI, including coding, research, and healthcare. However, current knowl

model-releasesarxiv-cs-ai
25 May 2026
Model Releases

DRIVESPATIAL: A Benchmark for Spatiotemporal Intelligence in VLMs for Autonomous Driving

DGX agent

arXiv:2605.23176v1 Announce Type: new Abstract: Spatiotemporal intelligence in autonomous driving (AD) requires an agent to integrate multi-view observations into a coherent scene representation, main

model-releasesarxiv-cs-cv
25 May 2026
Model Releases

GENSTRAT: Toward a Science of Strategic Reasoning in Large Language Models

DGX agent

arXiv:2605.23238v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly deployed as economic agents in marketplaces, auctions, and bidding settings. Anticipating their behavior i

model-releasesarxiv-cs-ai
25 May 2026
Local Ai

How Far Will They Go? Red-Teaming Online Influence with Large Language Models

DGX agent

arXiv:2605.22880v1 Announce Type: cross Abstract: As large language model (LLM)-based agents increasingly participate in online discourse, red-teaming their capacity to support political influence cam

local-aiarxiv-cs-ai
25 May 2026
Industry

给小伙买了Hugging Face家的Reachy Mini,目前桌面陪伴机器人里地表最强了。 不光是配件做工好,IDE和开发者生态好。还支持Agentic编程,10岁以下小朋友+codex很容易给它加功能。 手册上说组装需要3小时,小伙今天自己干了两个小时,就剩个头部了,3小时…

DGX agent

The post reviews the Hugging Face Reachy Mini robot, praising its build quality, IDE, and developer ecosystem for desktop companion robotics applications. It highlights the robot's support for agentic

industryclem-delangue--x
25 May 2026
Model Releases

MedExpMem: Adapting Experience Memory for Differential Diagnosis

DGX agent

arXiv:2605.22872v1 Announce Type: cross Abstract: Experienced physicians develop diagnostic expertise through clinical practice, acquiring not only disease knowledge but also the ability to differenti

model-releasesarxiv-cs-ai
25 May 2026
Model Releases

Philosophical Dispositions as Behavioral Constraints for AI-Assisted Code Review: An Empirical Study

DGX agent

arXiv:2605.23108v1 Announce Type: cross Abstract: AI-assisted code review tools typically operate as generic 'expert reviewer' agents, producing homogeneous findings regardless of the analysis type ne

model-releasesarxiv-cs-ai
25 May 2026
Safety

SCRIPT: Scalable Diffusion Policy with Multi-stage Training for Language-driven Physics-Based Humanoid Control

DGX agent

arXiv:2605.22894v1 Announce Type: cross Abstract: Controlling physics-based humanoids from natural-language instructions is a critical step toward general-purpose embodied agents. However, existing me

safetyarxiv-cs-lg
25 May 2026
Model Releases

Learn anything with our new /lesson-generator skill

DGX agent

Learn anything with our new /lesson-generator skill Just released my new /lesson-generator skill. Use it with your agent to learn anything: - generate lessons/courses on any topic - include nano-banan

model-releasesdair-ai--x
23 May 2026
Hardware

LiteCoOp: Lightweight Multi-LLM Shared-Tree Reasoning for Model-Serving Compiler Optimizations

DGX agent

arXiv:2602.01935v2 Announce Type: replace Abstract: LLM-guided compiler optimization has recently shown promise, but existing approaches rely on a single large LLM throughout search, making them expen

hardwarearxiv-cs-lg
23 May 2026
Model Releases

Catch up on the Dialogues stage at Google I/O 2026.

DGX agent

The Dialogues stage at Google I/O 2026 brought together Google leaders, scientific minds and creative visionaries to discuss technological breakthroughs. Featured discussions included AI agents and pr

model-releasesgoogle-ai
22 May 2026
Model Releases

Cursor Composer 2.5's is 3–18x cheaper than Opus 4.7 in Claude Code (medium reasoning), and 5–32x cheaper than GPT-5.5 in Codex (medium) bas…

DGX agent

Cursor Composer 2.5's is 3–18x cheaper than Opus 4.7 in Claude Code (medium reasoning), and 5–32x cheaper than GPT-5.5 in Codex (medium) based on API pricing This low Cost per Task isn't just driven b

model-releaseselon-musk--x
22 May 2026
Model Releases

DeepWeb-Bench: A Deep Research Benchmark Demanding Massive Cross-Source Evidence and Long-Horizon Derivation

DGX agent

arXiv:2605.21482v1 Announce Type: new Abstract: Deep research, in which an agent searches the open web, collects evidence, and derives an answer through extended reasoning, is a prominent use case for

model-releasesarxiv-cs-ai
22 May 2026
Research

DeferMem: Query-Time Evidence Distillation via Reinforcement Learning for Long-Term Memory QA

DGX agent

arXiv:2605.22411v1 Announce Type: new Abstract: Large language model (LLM) agents still struggle with long-term memory question answering, where answer-supporting evidence is often scattered across lo

researcharxiv-cs-cl
22 May 2026
Research

Interpreting and Enhancing Emotional Circuits in Large Vision-Language Models via Cross-Modal Information Flow

DGX agent

arXiv:2605.21980v1 Announce Type: new Abstract: Large Vision-Language Models (LVLMs) represent a significant leap towards empathetic agents, demonstrating remarkable capabilities in emotion understand

researcharxiv-cs-cv
22 May 2026
Tutorials

Join @hwchase17 + @traversal_ai founder @_anish_agarwal for a technical fireside chat in NYC on June 2nd. Learn how Traversal builds, ships …

DGX agent

Join @hwchase17 + @traversal_ai founder @_anish_agarwal for a technical fireside chat in NYC on June 2nd. Learn how Traversal builds, ships and improves their agents. Enjoy networking, drinks, and a l

tutorialsharrison-chase--x
22 May 2026
Model Releases

LongVT: Incentivizing 'Thinking with Long Videos' via Native Tool Calling

DGX agent

arXiv:2511.20785v3 Announce Type: replace Abstract: Large multimodal models (LMMs) have shown great potential for video reasoning with textual Chain-of-Thought. However, they remain vulnerable to hall

model-releasesarxiv-cs-cv
22 May 2026
Local Ai

Runaway token costs and sovereignty concerns are driving enterprises back to the desktop

DGX agent

The AI PC is being fundamentally redefined as agentic workloads push the boundaries of what local compute can deliver — and as runaway cloud token costs force enterprises to rethink where inference ac

local-aisiliconangle
22 May 2026
Research

stable-worldmodel: A Platform for Reproducible World Modeling Research and Evaluation

DGX agent

arXiv:2605.21800v1 Announce Type: cross Abstract: World models are central to building agents that can reason, plan, and generalize beyond their training data. However, research on world models is cur

researcharxiv-cs-ro
22 May 2026
Model Releases

Steins;Gate Drive: Semantic Safety Arbitration over Structured Futures for Latency-Decoupled LLM Planning

DGX agent

arXiv:2605.22456v1 Announce Type: new Abstract: Cloud-hosted LLM driver agents provide useful semantic judgments, but their inference latency exceeds stepwise vehicle-control windows. Learned world mo

model-releasesarxiv-cs-ro
22 May 2026
Industry

The $58,000 TV bill: When DirecTV sued O.J. Simpson for piracy

DGX agent

O.J. Simpson was ordered to pay $25,000 in damages for pirating satellite television signals from DirecTV using illegal devices known as 'bootloaders.' Federal agents seized the illegal devices from S

industryars-technica
22 May 2026
Model Releases

AI-Assisted Scientific Assessment: A Case Study on Climate Change

DGX agent

arXiv:2602.09723v2 Announce Type: replace Abstract: The emerging paradigm of AI co-scientists focuses on tasks characterized by repeatable verification, where agents explore search spaces in 'guess an

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

API Keys Are Open Secrets

DGX agent

Today, AI services rely heavily on API keys. To run AI agents, users provide API keys that signify paid tokens, subscriptions, or paid accounts. While API keys are easy to use, it is just as easy to u

model-releasesgoogle-cloud-ai
21 May 2026
Model Releases

Break the context window barrier with Amazon Bedrock AgentCore

DGX agent

In this post, you will learn how to implement Recursive Language Models (RLM) using Amazon Bedrock AgentCore Code Interpreter and the Strands Agents SDK. By the end, you will know how to process docum

model-releasesaws-ml-blog
21 May 2026
Industry

Full interview with Sundar is now live! https://x.com/rowancheung/status/2057491344697012384?s=20

DGX agent

Full interview with Sundar is now live! https://x.com/rowancheung/status/2057491344697012384?s=20 Google just revealed Omni, personalized cross-device intelligence, and Spark agents at I/O 2025. I sat

industryrowan-cheung--x
21 May 2026
Model Releases

GenAI-Driven Threat Detection with Microsoft Security Copilot

DGX agent

arXiv:2605.20896v1 Announce Type: cross Abstract: Defending against today's increasingly sophisticated cyberattacks requires security analysts to continuously translate evolving attacker tradecraft in

model-releasesarxiv-cs-lg
21 May 2026
Local Ai

Interaction Locality in Hierarchical Recursive Reasoning

DGX agent

arXiv:2605.20784v1 Announce Type: cross Abstract: Spatial reasoning requires both location-bound computation and location-invariant structure: agents must make local moves while preserving route, obje

local-aiarxiv-cs-lg
21 May 2026
Research

New VIDEO: From LLM Wikis to LLM Artifacts Shared all my thoughts on why LLM wikis and HTML artifacts are a big deal. Plus, new tools to hel…

DGX agent

New VIDEO: From LLM Wikis to LLM Artifacts Shared all my thoughts on why LLM wikis and HTML artifacts are a big deal. Plus, new tools to help you build wikis and artifacts with agents. Just getting st

researchdair-ai--x
21 May 2026
Model Releases

100 things we announced at I/O 2026

DGX agent

Google I/O 2026 unveiled new models, agents and tools to help users build, search, create, discover, shop and get more done. Key announcements included Gemini Omni, Google Antigravity, and Universal C

model-releasesgoogle-ai
20 May 2026
Industry

Announcing OpenAI-compatible API support for Amazon SageMaker AI endpoints

DGX agent

Today, Amazon SageMaker AI introduces OpenAI-compatible API support for real-time inference endpoints. If you use the OpenAI SDK, LangChain, or Strands Agents, you can now invoke models on SageMaker A

industryaws-ml-blog
20 May 2026
Model Releases

Anyone understand what Google mean by 'Gemini Spark runs on Gemini 3.5 and uses the Antigravity harness' - is 'Antigravity' a generic term t…

DGX agent

Anyone understand what Google mean by 'Gemini Spark runs on Gemini 3.5 and uses the Antigravity harness' - is 'Antigravity' a generic term they're using for their agent harnesses now or is their Claw-

model-releasessimon-willison--x
20 May 2026
Model Releases

AutoResearchClaw: Self-Reinforcing Autonomous Research with Human-AI Collaboration

DGX agent

arXiv:2605.20025v1 Announce Type: new Abstract: Automating scientific discovery requires more than generating papers from ideas. Real research is iterative: hypotheses are challenged from multiple per

model-releasesarxiv-cs-ai
20 May 2026
Industry

Build real-time voice applications with Amazon SageMaker AI and vLLM

DGX agent

Voice agents, live captioning, contact center analytics, and accessibility tools all depend on real-time speech-to-text, where your application streams audio in and receives transcription back simulta

industryaws-ml-blog
20 May 2026
Model Releases

CaptchaMind: Training CAPTCHA Solvers via Reinforcement Learning with Explicit Reasoning Supervision

DGX agent

arXiv:2605.19538v1 Announce Type: cross Abstract: CAPTCHAs are widely deployed as human verification mechanisms and frequently block intelligent agents from completing end-to-end automation in real-wo

model-releasesarxiv-cs-ai
20 May 2026
Tutorials

Decoding Rewards in Competitive Games: Inverse Game Theory with Entropy Regularization

DGX agent

arXiv:2601.12707v2 Announce Type: replace Abstract: Estimating the unknown reward functions driving agents' behaviors is of central interest in inverse reinforcement learning and game theory. To tackl

tutorialsarxiv-cs-lg
20 May 2026
← Previous
1…292293294295296…374
Next →