AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlog
84,619Total entries
1Added by human
84,618Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,976 results
Safety

When Evidence is Sparse: Weakly Supervised Early Failure Alerting in Dialogs and LLM-Agent Trajectories

DGX agent

arXiv:2606.05414v1 Announce Type: new Abstract: Early failure alerting requires deciding, while a dialog or agent trajectory is still unfolding, whether to flag it as likely to fail. This is challengi

safetyarxiv-cs-cl
5 Jun 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Industry

Working with agents should feel like working with a colleague. You should be able “speak to” them not just with text chats, but by gesturing…

DGX agent

Working with agents should feel like working with a colleague. You should be able “speak to” them not just with text chats, but by gesturing at a screen together, talking live, etc. With Design Mode,

industryelon-musk--x
5 Jun 2026
Model Releases

4/5 Still came in ~$0.30 under Claude Code’s spend at a similar score. So we added a lightweight Test Agent that writes repo tests and filte…

DGX agent

4/5 Still came in ~$0.30 under Claude Code’s spend at a similar score. So we added a lightweight Test Agent that writes repo tests and filters failing patches, pushing our final result to 60.9% - surp

model-releasesai21-labs--x
4 Jun 2026
Model Releases

Adaptive Minds: Empowering Agents with LoRA-as-Tools

DGX agent

arXiv:2510.15416v2 Announce Type: replace Abstract: We investigate a framework in which LoRA adapters are treated as callable tools that a base language model can dynamically select and invoke. We hyp

model-releasesarxiv-cs-ai
4 Jun 2026
Industry

An interview with Naomi Gleit, Meta's head of product who joined the company 20 years ago, on Zuckerberg's 'unfair' reputation, AI agents' capabilities, more (Zoe Kleinman/BBC)

DGX agent

Zoe Kleinman / BBC: An interview with Naomi Gleit, Meta's head of product who joined the company 20 years ago, on Zuckerberg's “unfair” reputation, AI agents' capabilities, more — When Naomi Gleit joi

industrytechmeme
4 Jun 2026
Model Releases

Caught in the Act(ivation): Toward Pre-Output and Multi-Turn Detection of Credential Exfiltration by LLM Agents

DGX agent

arXiv:2606.04141v1 Announce Type: cross Abstract: LLM agents often place sensitive credentials in the same context window as untrusted retrieved content, creating a direct path for indirect prompt inj

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

Nemotron 3.5 ASR is built for streaming multilingual speech recognition and voice agents. One 0.6B checkpoint. 40 language-locales. Sub-100m…

DGX agent

Nemotron 3.5 ASR is built for streaming multilingual speech recognition and voice agents. One 0.6B checkpoint. 40 language-locales. Sub-100ms latency. Cache-aware FastConformer carries context forward

model-releasestogether-ai--x
4 Jun 2026
Industry

OpenAI banned my account the day after I paid. 3 years of work, 30-40 Codex agents, all my client income — locked. No reason given. I'm the sole provider for my family. Has this happened to anyone?

DGX agent

This post describes a user's account suspension by OpenAI shortly after making a payment, resulting in loss of access to multiple Codex agents and client-dependent income, with no explanation provided

industryr-chatgpt
4 Jun 2026
Industry

Shared my first trace from @NanoClaw_AI to @huggingface yesterday. Very cool! By default, all agents should store their traces on HF (in pri…

DGX agent

Shared my first trace from @NanoClaw_AI to @huggingface yesterday. Very cool! By default, all agents should store their traces on HF (in private) so that you can keep a history of them, analyze them,.

industryclem-delangue--x
4 Jun 2026
Hardware

Together AI provides the inference stack behind both: high-throughput serving on the latest NVIDIA Blackwell GPUs for agentic workloads, and…

DGX agent

Together AI provides the inference stack behind both: high-throughput serving on the latest NVIDIA Blackwell GPUs for agentic workloads, and TensorRT engines plus event-driven streaming I/O for low-la

hardwaretogether-ai--x
4 Jun 2026
Local Ai

ARBOR: Online Process Rewards via a Reusable Rubric Buffer for Search Agents

DGX agent

arXiv:2606.03239v1 Announce Type: new Abstract: LLM-based search agents are trained predominantly with outcome-only reward, leaving the search process itself unsupervised. This signal degenerates on o

local-aiarxiv-cs-cl
3 Jun 2026
Model Releases

Assistax: A Multi-Agent Hardware-Accelerated Reinforcement Learning Benchmark for Assistive Robotics

DGX agent

arXiv:2507.21638v2 Announce Type: replace Abstract: The development of reinforcement learning (RL) algorithms has been largely driven by ambitious challenge tasks and benchmarks. Games have dominated

model-releasesarxiv-cs-ai
3 Jun 2026
Model Releases

EvoDrive: Pareto Evolution for Safety-Critical Autonomous Driving via Self-Improving LLM Agents

DGX agent

arXiv:2606.03678v1 Announce Type: new Abstract: Generating safety-critical scenarios is essential for validating and improving autonomous driving systems, yet it inherently requires maximizing adversa

model-releasesarxiv-cs-ai
3 Jun 2026
Model Releases

Frontier models are powerful advisors. On @harvey's Legal Agent Benchmark, a GLM 5.1 worker using Claude Opus 4.7 as a sparse advisor reache…

DGX agent

Frontier models are powerful advisors. On @harvey's Legal Agent Benchmark, a GLM 5.1 worker using Claude Opus 4.7 as a sparse advisor reached 18/100 all-pass versus 14/100 for Opus alone, at 39% of th

model-releasesfireworks-ai--x
3 Jun 2026
Model Releases

.@GoogleDeepMind's Gemma 4 - 12B is available on Ollama! Chat: ollama run gemma4:12b-mlx Hermes Agent: ollama launch hermes --model gemma4:1…

DGX agent

.@GoogleDeepMind's Gemma 4 - 12B is available on Ollama! Chat: ollama run gemma4:12b-mlx Hermes Agent: ollama launch hermes --model gemma4:12b-mlx Claude Code: ollama launch claude --model gemma4:12b-

model-releasesollama--x
3 Jun 2026
Safety

Margin Play: A Multi-Agent System For Public Policy Analysis In The Brazilian Equatorial Margin

DGX agent

arXiv:2606.02614v1 Announce Type: cross Abstract: The Brazilian Equatorial Margin (BEM) is Brazil's next offshore oil frontier, with operations expected to begin in 2026 in the Foz do Amazonas basin.

safetyarxiv-cs-ai
3 Jun 2026
Hardware

MOSAIC: Efficient Mixture-of-Agent Scheduling via Adaptive Aggregation and Inference Concurrency

DGX agent

arXiv:2606.03014v1 Announce Type: new Abstract: Mixture-of-Agents (MoA) systems improve reasoning accuracy by routing each query to multiple expert LLMs and aggregating their outputs. Efficiently exec

hardwarearxiv-cs-lg
3 Jun 2026
Hardware

NVIDIA Enables the Next Era Of Physical AI Research With Agent Skills For Autonomous Vehicles, Robotics And Vision AI

DGX agent

At CVPR, NVIDIA is unveiling new physical AI agent skills that help researchers and developers speed the development of autonomous vehicles, robots and vision AI systems. The core challenge in physica

hardwarenvidia-blog
3 Jun 2026
Model Releases

OpenAI ran a hiring challenge, but the top candidate was one they couldn’t hire: our autonomous research agent, Aiden. In Parameter Golf, Ai…

DGX agent

OpenAI ran a hiring challenge, but the top candidate was one they couldn’t hire: our autonomous research agent, Aiden. In Parameter Golf, Aiden ran for 22 days, and out-outperformed all 1,016 other re

model-releasesclem-delangue--x
3 Jun 2026
Applications

Pinecone Nexus Now Integrates with Microsoft OneLake, Bringing AI Agents Directly to Enterprise Data

DGX agent

Pinecone Nexus has integrated with Microsoft OneLake, enabling AI agents to access and work directly with enterprise data stored in OneLake without requiring data movement or copying. This integration

applicationspinecone
3 Jun 2026
Research

RGMem: Renormalization Group-inspired Memory Evolution for Language Agents

DGX agent

arXiv:2510.16392v3 Announce Type: replace Abstract: Personalized and continuous interactions are critical for LLM-based conversational agents, yet finite context windows and static parametric memory h

researcharxiv-cs-ai
3 Jun 2026
Tutorials

the best agents aren't just built with the best models: they're built with harnesses purpose-built for the task at hand here's a guide on ho…

DGX agent

the best agents aren't just built with the best models: they're built with harnesses purpose-built for the task at hand here's a guide on how to build a harness that's really good at feeding the model

tutorialsharrison-chase--x
3 Jun 2026
Safety

Tool-Aware Optimization with Entropy Guidance for Efficient Agentic Reinforcement Learning

DGX agent

arXiv:2606.03762v1 Announce Type: cross Abstract: Agentic reinforcement learning (RL) equips large language models (LLMs) with tool-use capabilities that substantially improve reasoning on complex tas

safetyarxiv-cs-ai
3 Jun 2026
Model Releases

TSQAgent: Rating Time Series Data Quality via Dedicated Agentic Reasoning

DGX agent

arXiv:2606.03629v1 Announce Type: new Abstract: Assessing the quality of time series (TS) data is fundamental yet inherently challenging due to the multifaceted nature of quality dimensions. Recently,

model-releasesarxiv-cs-ai
3 Jun 2026
Tools

Uber reportedly now caps coding agents at $1,500/month per employee per tool - seems sensible to me, but it's also an interesting hint at th…

DGX agent

Uber reportedly now caps coding agents at $1,500/month per employee per tool - seems sensible to me, but it's also an interesting hint at the value Uber thinks these tools are providing https://simonw

toolssimon-willison--x
3 Jun 2026
Local Ai

Build on-device personal AI agents on Windows PCs with new tools from NVIDIA and Microsoft, including secure sandboxing, faster local infere…

DGX agent

Build on-device personal AI agents on Windows PCs with new tools from NVIDIA and Microsoft, including secure sandboxing, faster local inference, multi-GPU support, and RTX acceleration for Windows AI

local-aicomfyui--x
2 Jun 2026
Industry

Cisco bets on AI agents to redefine workplace collaboration

DGX agent

Cisco Systems Inc. today unveiled a set of collaboration and customer experience products that transform its Webex videoconferencing platform into what executives describe as an intelligent operating

industrysiliconangle
2 Jun 2026
Tools

Conductor's parallel coding agents were local-only, but now they can run remotely on Vercel. 'Our customers can't tell the difference becaus…

DGX agent

Conductor's parallel coding agents were local-only, but now they can run remotely on Vercel. 'Our customers can't tell the difference because Vercel's Sandboxes are so fast.' https://vercel.com/blog/h

toolsvercel--x
2 Jun 2026
Model Releases

Cross-Environment Neural Reranking for Sample-Efficient Action Selection in Text-Based Agents

DGX agent

arXiv:2606.02204v1 Announce Type: new Abstract: Large language model agents achieve strong performance on text-based benchmarks but incur prohibitive inference costs, motivating the use of compact neu

model-releasesarxiv-cs-cl
2 Jun 2026
Hardware

Deploy Agentic-Ready AI at the Edge with Memory Efficiency in NVIDIA JetPack 7.2

DGX agent

NVIDIA JetPack 7.2 enables deployment of AI agents to edge devices with optimized memory and performance for real-world applications. The release directly supports one-command deployment of NVIDIA Nem

hardwarenvidia-developer
2 Jun 2026
Research

Doing What They Say, Not What They Reason: Locating the Faithfulness Gap in LLM Agents

DGX agent

arXiv:2606.00476v1 Announce Type: new Abstract: Do LLM agents act on the reasoning they state? This question of process fidelity is central to using LLMs in social simulation, yet it is hard to measur

researcharxiv-cs-ai
2 Jun 2026
Tools

Holo3.1: Fast & Local Computer Use Agents

DGX agent

Holo3.1 is a computer use agent developed by Hugging Face that enables fast, local execution of tasks on computing systems without requiring cloud infrastructure. The model is designed to interpret an

toolshugging-face
2 Jun 2026
Safety

Leyline: KV Cache Directives for Agentic Inference

DGX agent

arXiv:2606.01065v1 Announce Type: cross Abstract: Modern KV cache management assumes the chatbot workload: prompts arrive once and the cache grows append-only, so prefix caching and forward-only evict

safetyarxiv-cs-ai
2 Jun 2026
Research

LLM agents patch security bugs, pass all tests, but still leave the vulnerability open [R]

DGX agent

Research demonstrates that LLM-based agents can generate functionally correct patches that pass all tests while still containing security vulnerabilities, challenging the assumption that test-passing

researchr-machinelearning
2 Jun 2026
Model Releases

LLM Consortium for Software Design Refinement: A Controlled Experiment on Multi-Agent Collaboration Topologies

DGX agent

arXiv:2606.01490v1 Announce Type: cross Abstract: We present a controlled experiment evaluating 12 multi-agent LLM collaboration topologies for software architecture design. Using a 2imes2imes2 factor

model-releasesarxiv-cs-ai
2 Jun 2026
Safety

MobEvolve: An Agentic Self-Evolving Heuristic System for Interpretable Human Mobility Generation

DGX agent

arXiv:2606.01640v1 Announce Type: new Abstract: Human mobility generation aims to synthesize realistic trip chains for target populations based on individual features. Existing paradigms, including de

safetyarxiv-cs-ai
2 Jun 2026
Local Ai

NVIDIA Partners With Microsoft on Unified Stack for Agentic AI Deployment, From Windows Devices to Cloud to Local

DGX agent

The agentic AI moment has arrived, but delivering on its promise requires more than good models. It also takes fast hardware, secure runtimes, a responsive data layer and models tuned for long-running

local-ainvidia-blog
2 Jun 2026
Hardware

Observation, Not Prediction: Conversation-Level Disaggregated Scheduling for Agentic Serving

DGX agent

arXiv:2606.01839v1 Announce Type: cross Abstract: LLM-based agents resolve a user task through many turns of dependent inference and tool calls, producing a workload whose total cost is unknown when t

hardwarearxiv-cs-lg
2 Jun 2026
Safety

On Effectiveness and Efficiency of Agentic Tool-calling and RL Training

DGX agent

arXiv:2606.00135v1 Announce Type: cross Abstract: Tool-calling is a central component of modern large language model (LLM) agents, equipping them with skills beyond their parametric knowledge. This pa

safetyarxiv-cs-ai
2 Jun 2026
Model Releases

RescueBench: Can Embodied Agents Save Lives in the Wild ?

DGX agent

arXiv:2606.01848v1 Announce Type: new Abstract: Search-and-rescue (SAR) requires embodied agents to explore unfamiliar environments under multimodal uncertainty, perform multi-stage interactions, and

model-releasesarxiv-cs-cv
2 Jun 2026
Local Ai

Today we're announcing that hybrid agentic inference is coming to Perplexity Computer. Computer can split tasks between a local model runnin…

DGX agent

Today we're announcing that hybrid agentic inference is coming to Perplexity Computer. Computer can split tasks between a local model running on your machine and frontier models in the cloud. This kee

local-aiperplexity--x
2 Jun 2026
Model Releases

Uber says it has limited all employees to $1,500 in monthly token spending per AI coding tool 'to responsibly encourage agentic AI adoption' (Natalie Lung/Bloomberg)

DGX agent

Natalie Lung / Bloomberg: Uber says it has limited all employees to $1,500 in monthly token spending per AI coding tool “to responsibly encourage agentic AI adoption” — Uber Technologies Inc. has set

model-releasestechmeme
2 Jun 2026
Tools

Using Parallel Agents to Move Faster in Replit https://x.com/i/broadcasts/1NxarrEMVOnKj

DGX agent

This broadcast discusses techniques for improving performance and speed in Replit by utilizing parallel agents or concurrent processing methods. The content likely covers how developers can leverage p

toolsreplit--x
2 Jun 2026
Hardware

A new beginning of PC starts with @NVIDIARTXSpark, supercharging what's possible in Hermes Agent.

DGX agent

A new beginning of PC starts with @NVIDIARTXSpark, supercharging what's possible in Hermes Agent. This is the NVIDIA RTX Spark Superchip. A new beginning for personal computers. Designed for creators,

hardwarenous-research--x
1 Jun 2026
Safety

Counterfactual Evaluation Reveals Hidden Capability Profiles in Clinical LLMs and Agents

DGX agent

arXiv:2605.30590v1 Announce Type: cross Abstract: Two clinical AI systems can score nearly identically on coverage-based rubrics yet behave radically differently when their patient inputs change: one

safetyarxiv-cs-ai
1 Jun 2026
Model Releases

DeepImageSearch: Benchmarking Multimodal Agents for Context-Aware Image Retrieval in Visual Histories

DGX agent

arXiv:2602.10809v2 Announce Type: replace Abstract: Existing multimodal retrieval systems excel at semantic matching but implicitly assume that query-image relevance can be measured in isolation. This

model-releasesarxiv-cs-cv
1 Jun 2026
Safety

Detect in Any Scene: An Agentic Framework for Object Detection with Experience-Aware Reasoning

DGX agent

arXiv:2605.31174v1 Announce Type: new Abstract: Object detection in real-world scenarios remains challenging due to diverse image degradations and heterogeneous object distributions, which significant

safetyarxiv-cs-cv
1 Jun 2026
Model Releases

Eywa: Provenance-Grounded Long-Term Memory for AI Agents

DGX agent

arXiv:2605.30771v1 Announce Type: new Abstract: AI agents that persist across sessions need memory they can retrieve, audit, update, and erase. Existing memory systems often collapse source evidence,

model-releasesarxiv-cs-cl
1 Jun 2026
← Previous
1…163164165166167…375
Next →