AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlog
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,920 results
Model Releases

Datadog launches more than 100 features at DASH to push autonomous AI ops

DGX agent

Observability and security platform company Datadog Inc. today unveiled more than 100 new capabilities at its annual DASH 2026 conference, headlined by a major expansion of its Bits AI agents that the

model-releasessiliconangle
9 Jun 2026
Research
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

DIJIT: A Robotic Head for an Active Observer

DGX agent

arXiv:2512.07998v2 Announce Type: replace-cross Abstract: We present DIJIT, a novel binocular robotic head expressly designed for mobile agents that behave as active observers. DIJIT's unique breadth

researcharxiv-cs-cv
9 Jun 2026
Model Releases

If you thought AI progress was slowing down, well here's the immediate answer to that. Huge jump in capability across the board. This is goi…

DGX agent

If you thought AI progress was slowing down, well here's the immediate answer to that. Huge jump in capability across the board. This is going to deliver major improvement in agents across almost all

model-releasesboris-cherny--x
9 Jun 2026
Model Releases

Language-based Trial and Error Falls Behind in the Era of Experience

DGX agent

arXiv:2601.21754v3 Announce Type: replace Abstract: While Large Language Models (LLMs) excel in language-based agentic tasks, their applicability to unseen, nonlinguistic environments (e.g., symbolic

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

LargeMonitor: Monitoring Online Task-Free Continual Learning via Large Pretrained Models

DGX agent

arXiv:2606.09430v1 Announce Type: cross Abstract: Online task-free continual learning (TFCL) requires intelligent agents to sequentially accumulate knowledge from an unbounded, non-stationary data str

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

LLM Inference at the Edge: Mobile, NPU, and GPU Performance Efficiency Trade-offs Under Sustained Load

DGX agent

arXiv:2603.23640v2 Announce Type: replace-cross Abstract: Deploying large language models on-device for always-on personal agents demands sustained inference from hardware tightly constrained in power

model-releasesarxiv-cs-lg
9 Jun 2026
Hardware

MuJoCo-Drones-Gym: A GPU-Accelerated Multi-Drone Simulator for Control and Reinforcement Learning

DGX agent

arXiv:2606.08039v1 Announce Type: new Abstract: Robotic simulators are a cornerstone of modern research in aerial robotics, serving both as a vehicle for the development of new control algorithms and

hardwarearxiv-cs-ro
9 Jun 2026
Model Releases

Setting a custom price for a model in AgentsView

DGX agent

TIL: Setting a custom price for a model in AgentsView I've been really enjoying AgentsView by Wes McKinney as a tool for exploring my token usage across different coding agents running on my laptop. C

model-releasessimon-willison
9 Jun 2026
Model Releases

Storage Insights datasets: Enabling org-wide operational discovery with activity insights

DGX agent

As enterprise storage footprints scale to billions of objects, AI applications and agentic workloads are fundamentally shifting the role of storage from a passive repository to the foundation of the d

model-releasesgoogle-cloud-ai
9 Jun 2026
Hardware

Towards Automated Kernel Generation in the Era of LLMs

DGX agent

arXiv:2601.15727v3 Announce Type: replace Abstract: The performance of modern AI systems is fundamentally constrained by the quality of their underlying GPU kernels, which translate high-level algorit

hardwarearxiv-cs-lg
9 Jun 2026
Model Releases

Where Instruction Hierarchy Breaks: Diagnosing and Repairing Failures in Reasoning Language Models

DGX agent

arXiv:2606.07808v1 Announce Type: new Abstract: Reasoning language models deployed in agentic workflows must follow an instruction hierarchy: when instructions from different sources conflict, the mod

model-releasesarxiv-cs-ai
9 Jun 2026
Safety

WhiFlash: Accelerating Speculative Decoding with Token-Level Cross-Paradigm Routing

DGX agent

arXiv:2606.07710v1 Announce Type: cross Abstract: The autoregressive nature of large language models (LLMs) remains a significant bottleneck for inference, particularly in complex agentic workloads. W

safetyarxiv-cs-ai
9 Jun 2026
Research

Accelerated Decentralized Stochastic Gradient Descent for Strongly Convex Optimization

DGX agent

arXiv:2606.07496v1 Announce Type: new Abstract: Decentralized stochastic optimization is a fundamental paradigm for large-scale learning over networks, where agents communicate only with their neighbo

researcharxiv-cs-lg
8 Jun 2026
Model Releases

Beyond Waypoints: A Trajectory-Centric Waypointing Paradigm for Vision-Language Navigation

DGX agent

arXiv:2606.07244v1 Announce Type: cross Abstract: Vision-Language Navigation in Continuous Environments (VLN-CE) requires agents to follow natural-language instructions while navigating in real-world-

model-releasesarxiv-cs-ai
8 Jun 2026
Tutorials

CAPE: Contrastive Action-conditioned Parallel Encoding for Embodied Planning

DGX agent

arXiv:2606.07304v1 Announce Type: new Abstract: Embodied agents need to predict the future consequences of candidate actions in order to plan effectively before execution. Existing visual dynamics mod

tutorialsarxiv-cs-ro
8 Jun 2026
Local Ai

Learning Explicit Behavioral Models with Adaptive Questions and World-Model Probes

DGX agent

arXiv:2606.07127v1 Announce Type: new Abstract: Interactive agents trained only against task return can achieve high scores while failing to represent the mechanisms that make their actions succeed. T

local-aiarxiv-cs-lg
8 Jun 2026
Model Releases

NTILC: Neural Tool Invocation via Learned Compression

DGX agent

arXiv:2606.06566v1 Announce Type: cross Abstract: Agentic tool-calling language models depend on large registries of callable APIs, functions, and local actions. Placing full tool specifications direc

model-releasesarxiv-cs-ai
8 Jun 2026
Model Releases

RECAP: Regression Evaluation for Continual Adaptation of Prompts

DGX agent

arXiv:2606.06698v1 Announce Type: cross Abstract: Production agentic systems routinely face evolving constraints and must comply from the very next interaction. Scenarios like a tool-call notification

model-releasesarxiv-cs-cl
8 Jun 2026
Safety

Robust Driving Control for Autonomous Vehicles: An Intelligent General-sum Constrained Adversarial Reinforcement Learning Approach

DGX agent

arXiv:2510.09041v3 Announce Type: replace-cross Abstract: Deep reinforcement learning (DRL) has demonstrated remarkable success in developing autonomous driving policies. However, its vulnerability to

safetyarxiv-cs-ai
8 Jun 2026
Model Releases

Rubrics are even more flexible than /goal You can define a custom subagent for the grading, including custom tools, prompt, and iteration li…

DGX agent

Rubrics are even more flexible than /goal You can define a custom subagent for the grading, including custom tools, prompt, and iteration limits. Try it out and let us know what you think! we just shi

model-releasesharrison-chase--x
8 Jun 2026
Hardware

NVIDIA, KRAFTON, NC and Reigning ‘League of Legends’ Champions T1 Celebrate RTX Spark at Korea’s PC Bangs

DGX agent

At GTC Taipei at COMPUTEX last week, NVIDIA unveiled RTX Spark, the superchip that reinvents Windows PCs for the era of personal AI agents. On the heels of this announcement, NVIDIA founder and CEO Je

hardwarenvidia-blog
7 Jun 2026
Industry

OpenAI plots biggest ChatGPT overhaul since launch

DGX agent

OpenAI is planning its biggest ChatGPT overhaul yet, aiming to turn it into a 'superapp' with coding tools and AI agents to boost revenue ahead of a potential stock market listing. The redesigned Chat

industryr-chatgpt
7 Jun 2026
Model Releases

Brick-Composer: Using MLLMs for Assembly with Diverse Bricks

DGX agent

arXiv:2606.05445v1 Announce Type: new Abstract: We dream of AI agents that can read arbitrary designs and construct real-world objects from reusable building blocks. As a first step toward this vision

model-releasesarxiv-cs-ai
6 Jun 2026
Tutorials

// Continual Learning Bench // One of the research areas with lots of investments is continual learning. While there are many efforts, there…

DGX agent

// Continual Learning Bench // One of the research areas with lots of investments is continual learning. While there are many efforts, there is very little progress in measuring it. So the big questio

tutorialsdair-ai--x
6 Jun 2026
Model Releases

DragOn: A Benchmark and Dataset for Drag-Based GUI Interactions

DGX agent

arXiv:2606.06322v1 Announce Type: new Abstract: GUI agents - vision-based models that control desktops, web browsers, and mobile devices through graphical user interfaces - promise to automate a wide

model-releasesarxiv-cs-ai
6 Jun 2026
Applications

Five labs, five minds: building a multi-model finance drama on small models

DGX agent

This article describes a collaborative hackathon project involving five research labs that developed a financial simulation drama using small language models, focusing on building complex multi-agent

applicationshugging-face
6 Jun 2026
Safety

GIPO: Gaussian Importance Sampling Policy Optimization

DGX agent

arXiv:2603.03955v2 Announce Type: replace-cross Abstract: Post-training with reinforcement learning (RL) has recently shown strong promise for advancing multimodal agents beyond supervised imitation.

safetyarxiv-cs-ai
6 Jun 2026
Model Releases

Goedel-Architect: Streamlining Formal Theorem Proving with Blueprint Generation and Refinement

DGX agent

arXiv:2606.06468v1 Announce Type: new Abstract: We introduce Goedel-Architect, an agentic framework for formal theorem proving in Lean 4 centered on blueprint generation and refinement. A blueprint is

model-releasesarxiv-cs-ai
6 Jun 2026
Safety

Flow-based Policy Adaptation without Policy Updates

DGX agent

arXiv:2606.06461v1 Announce Type: new Abstract: Leveraging prior knowledge from pretrained policies, foundation models, or human operators offers an efficient alternative to learning robot skills from

safetyarxiv-cs-ro
5 Jun 2026
Industry

Grok supports worktrees

DGX agent

Grok supports worktrees Grok Build tip of the day: worktrees! If you're unfamiliar with worktrees, they're essentially lightweight copies of your repo, allowing you to run parallel agents within their

industryelon-musk--x
5 Jun 2026
Model Releases

KV-Control: Parameter-Efficient K/V Injection for Trajectory-Controlled Text-to-Motion

DGX agent

arXiv:2606.05624v1 Announce Type: new Abstract: Text-conditioned 3D human motion models now synthesize plausible motions from prompts, but practical animation and embodied-agent workflows rarely stop

model-releasesarxiv-cs-cv
5 Jun 2026
Local Ai

LeanMarathon: Toward Reliable AI Co-Mathematicians through Long-Horizon Lean Autoformalization

DGX agent

arXiv:2606.05400v1 Announce Type: cross Abstract: Long-horizon autoformalization of research mathematics fails not only at hard lemmas, but at scale: statements drift, dependencies tangle, context dec

local-aiarxiv-cs-cl
5 Jun 2026
Model Releases

LLM-Guided ANN Index Optimization for Human-Object Interaction Retrieval

DGX agent

arXiv:2606.05489v1 Announce Type: new Abstract: Retrieval systems underpin modern AI applications -- spanning visual search, recommendation engines, and multi-modal question answering. Modern multi-st

model-releasesarxiv-cs-cv
5 Jun 2026
Research

PHUMA: Physically Reliable Humanoid Locomotion Dataset

DGX agent

arXiv:2510.26236v2 Announce Type: replace Abstract: Motion imitation is a promising approach for humanoid locomotion, enabling agents to acquire humanlike behaviors. Existing methods typically rely on

researcharxiv-cs-ro
5 Jun 2026
Safety

QueryAgent-R1: Bridging Query Generation and Product Retrieval for E-Commerce Query Recommendation

DGX agent

arXiv:2606.05671v1 Announce Type: new Abstract: Query recommendation in e-commerce search aims to proactively suggest queries that match users' potential interests. However, existing methods mainly op

safetyarxiv-cs-cl
5 Jun 2026
Safety

RiskFlow: Fast and Faithful Safety-Critical Traffic Scenario Generation

DGX agent

arXiv:2606.06423v1 Announce Type: new Abstract: Safety-critical traffic scenario generation is essential for evaluating autonomous driving systems under rare but high-risk interactions. Existing diffu

safetyarxiv-cs-ro
5 Jun 2026
Model Releases

The latest AI news we announced in May 2026

DGX agent

Google's May 2026 AI updates center on the new 'agentic' era, featuring the Gemini 3.5 model and Gemini Omni for advanced reasoning and creation. Gemini Omni is a new model that can create anything fr

model-releasesgoogle-ai
5 Jun 2026
Safety

Adaptive Information Control for Search-Augmented LLM Reasoning

DGX agent

arXiv:2602.01672v2 Announce Type: replace Abstract: Search-augmented reasoning agents interleave multi-step reasoning with external retrieval, but uncontrolled retrieval can introduce redundant eviden

safetyarxiv-cs-cl
4 Jun 2026
Model Releases

BioBlue: Systematic runaway-optimiser-like LLM failure modes on biologically and economically aligned AI safety benchmarks for LLMs with simplified observation format

DGX agent

arXiv:2509.02655v3 Announce Type: replace-cross Abstract: Many AI alignment discussions of 'runaway optimisation' focus on RL agents: unbounded utility maximisers that over-optimise a proxy objective

model-releasesarxiv-cs-ai
4 Jun 2026
Applications

CYGNET: Cypher Gate for Neural Execution Triage and Cost Containment

DGX agent

arXiv:2606.04645v1 Announce Type: new Abstract: Language models acting as agents over knowledge graphs generate Cypher queries that fail structurally (crashing at the database) or semantically (execut

applicationsarxiv-cs-cl
4 Jun 2026
Model Releases

Nemotron 3 Ultra (550B-A55B) is here - our strongest open-weight model and full training recipe to date. Heavy emphasis on real-world infere…

DGX agent

Nemotron 3 Ultra (550B-A55B) is here - our strongest open-weight model and full training recipe to date. Heavy emphasis on real-world inference efficiency for long-context agentic workloads. Everythin

model-releasesclem-delangue--x
4 Jun 2026
Model Releases

ollama run gemma4:12b Gemma 4 12B is updated on Ollama, and available across all platforms! Try it on: Claude Code ollama launch claude --mo…

DGX agent

ollama run gemma4:12b Gemma 4 12B is updated on Ollama, and available across all platforms! Try it on: Claude Code ollama launch claude --model gemma4:12b Hermes Agent ollama launch hermes --model gem

model-releasesollama--x
4 Jun 2026
Safety

Policy Gradient for Continuous-Time Robust Markov Decision Processes

DGX agent

arXiv:2606.04335v1 Announce Type: new Abstract: The framework of robust Markov decision processes (RMDPs) allows the design of reinforcement learning agents that satisfy performance guarantees under w

safetyarxiv-cs-lg
4 Jun 2026
Local Ai

R-APS: Compositional Reasoning and In-Context Meta-Learning for Constrained Design via Reflective Adversarial Pareto Search

DGX agent

arXiv:2606.04823v1 Announce Type: new Abstract: Large language models (LLMs) are fluent on open-ended tasks, yet in agentic settings, where a system must plan, use tools, and act over extended horizon

local-aiarxiv-cs-ai
4 Jun 2026
Safety

Unlocking Proactivity in Task-Oriented Dialogue

DGX agent

arXiv:2605.22240v2 Announce Type: replace Abstract: Proactive task-oriented dialogue (TOD), such as outbound sales, demands a persuasive agent that actively probes the user's concerns and steers the c

safetyarxiv-cs-ai
4 Jun 2026
Model Releases

As AI gets better, it reveals an empty promise

DGX agent

This week we've got tandem hands-ons with Google's new Gemini AI agent - Spark - from my colleagues David Pierce and Jay Peters. Their takeaways are similar: It's so effective that it's scary. Spark k

model-releasesthe-verge-ai
3 Jun 2026
Model Releases

Benchmarking Visual State Tracking in Multimodal Video Understanding

DGX agent

arXiv:2606.03920v1 Announce Type: new Abstract: Understanding a video requires more than recognizing isolated moments, as humans continuously track entities, states, and events over time. This capacit

model-releasesarxiv-cs-cv
3 Jun 2026
Model Releases

Chatbots Output Meaningful (but Problematic) Language

DGX agent

arXiv:2606.02973v1 Announce Type: new Abstract: Are utterances by AI chatbots meaningful? Concretely, if a user asks, say, Anthropic's agent Claude, 'What is the capital of Spain?' and Claude answers,

model-releasesarxiv-cs-cl
3 Jun 2026
← Previous
1…289290291292293…374
Next →