AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,964 results
28 Apr 2026

Introducing NVIDIA Nemotron 3 Nano Omni: Long-Context Multimodal Intelligence for Documents, Audio and Video Agents

Model ReleasesDGX agent

NVIDIA's Nemotron 3 Nano Omni is a lightweight multimodal AI model capable of processing documents, audio, and video inputs for building intelligent agents. The model supports long-context understandi

LLM-Assisted Op-Amp Behavioral-Level Design via Agentic Human-Mimicking Reasoning

Model ReleasesDGX agent

arXiv:2601.21321v2 Announce Type: replace Abstract: This paper proposes White-Op, an operational amplifier (op-amp) behavioral-level parameter design framework assisted by the human-mimicking reasonin

LLM-Guided Agentic Floor Plan Parsing for Accessible Indoor Navigation of Blind and Low-Vision People

Model Releases
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2604.23970v1 Announce Type: new Abstract: Indoor navigation remains a critical accessibility challenge for the blind and low-vision (BLV) individuals, as existing solutions rely on costly per-bu

NVIDIA Nemotron 3 Nano Omni Powers Multimodal Agent Reasoning in a Single Efficient Open Model

Model ReleasesDGX agent

NVIDIA Nemotron 3 Nano Omni is a 30B hybrid mixture-of-experts model that brings multimodal perception and reasoning into a single system, natively supporting text, image, video, and audio inputs whil

OpenAI models, Codex, and Managed Agents come to AWS

Model ReleasesDGX agent

OpenAI announced the availability of its models, including Codex, and managed agent capabilities on Amazon Web Services (AWS) infrastructure. This integration enables AWS customers to access OpenAI's

PExA: Parallel Exploration Agent for Complex Text-to-SQL

Model ReleasesDGX agent

arXiv:2604.22934v1 Announce Type: new Abstract: LLM-based agents for text-to-SQL often struggle with latency-performance trade-off, where performance improvements come at the cost of latency or vice v

Q&A with Sam Altman and AWS CEO Matt Garman about OpenAI's new partnership with AWS, Bedrock Managed Agents, Trainium chips, and more (Ben Thompson/Stratechery)

IndustryDGX agent

Ben Thompson / Stratechery: Q&A with Sam Altman and AWS CEO Matt Garman about OpenAI's new partnership with AWS, Bedrock Managed Agents, Trainium chips, and more — As I noted yesterday, today's Strate

Today we’re releasing Laguna XS.2, Poolside’s first open-weight model. It’s a 33B total / 3B active MoE model built for agentic coding and l…

Model ReleasesDGX agent

Today we’re releasing Laguna XS.2, Poolside’s first open-weight model. It’s a 33B total / 3B active MoE model built for agentic coding and long-horizon tasks. Trained fully in-house on our own stack.

27 Apr 2026

A completely local agent that lives right inside your browser. Powered by Gemma 4 E2B and WebGPU, it uses native tool calling to: 🔍 Search …

Model ReleasesDGX agent

A completely local agent that lives right inside your browser. Powered by Gemma 4 E2B and WebGPU, it uses native tool calling to: 🔍 Search browsing history 📄 Read and summarize pages 🔗 Manage tabs 100

MolClaw: An Autonomous Agent with Hierarchical Skills for Drug Molecule Evaluation, Screening, and Optimization

Model ReleasesDGX agent

arXiv:2604.21937v1 Announce Type: new Abstract: Computational drug discovery, particularly the complex workflows of drug molecule screening and optimization, requires orchestrating dozens of specializ

ParseBench: A benchmark for document parsing agents @llama_index just shipped a benchmark with 2k verified pages for real enterprise documen…

Model ReleasesDGX agent

ParseBench: A benchmark for document parsing agents @llama_index just shipped a benchmark with 2k verified pages for real enterprise documents. Benchmarks are the major underrated component in the ML

RTX 5090 users: TensorRT-LLM vs llama.cpp (GGUF) for Coding Agents (Cline/RooCode) – Is the speed worth the VRAM limit?

Model ReleasesDGX agent

This post compares TensorRT-LLM and llama.cpp (GGUF) as inference frameworks for running coding agents like Cline and RooCode on RTX 5090 GPUs, examining the tradeoff between inference speed and VRAM

This is totally wrong. Blaming the user is missing the point that (a) coding agents have been overhyped and (b) can’t reliably obey the rule…

SafetyDGX agent

This is totally wrong. Blaming the user is missing the point that (a) coding agents have been overhyped and (b) can’t reliably obey the rules given to them in system prompts and other guardrails. Sorr

TuneForge: an MCP server that lets your coding agent (Claude, Cursor, etc.) handle dataset generation, LoRA fine-tuning, RL, and evaluation directly in chat

Model ReleasesDGX agent

TuneForge is an MCP (Model Context Protocol) server that enables coding agents like Claude and Cursor to perform machine learning operations directly within chat interfaces, including dataset generati

We power 1M+ developers globally. Our researchers developed FlashAttention, Mixture of Agents, EinsteinArena, and more. Our platform is buil…

ToolsDGX agent

We power 1M+ developers globally. Our researchers developed FlashAttention, Mixture of Agents, EinsteinArena, and more. Our platform is built for large-scale, latency-sensitive workloads on open-sourc

26 Apr 2026

Collov Labs, whose visual interface lets users feed images and camera input into a model that AI agents can reason over and act on, raised a $23M Series A (Chris Metinko/Axios)

ApplicationsDGX agent

Chris Metinko / Axios: Collov Labs, whose visual interface lets users feed images and camera input into a model that AI agents can reason over and act on, raised a 23M Series A — Collov Labs, which tu

25 Apr 2026

🚀Meet Carnice-V2-27b🚀 → Carnice is a 27 billion parameter model capable of beating models 10x the size in Hermes-agent, fully open-source …

Model ReleasesDGX agent

🚀Meet Carnice-V2-27b🚀 → Carnice is a 27 billion parameter model capable of beating models 10x the size in Hermes-agent, fully open-source and built on top of Qwen3.6-27B →Build to fit on Consumer GPU

24 Apr 2026

Cross-Session Threats in AI Agents: Benchmark, Evaluation, and Algorithms

Model ReleasesDGX agent

arXiv:2604.21131v1 Announce Type: cross Abstract: AI-agent guardrails are memoryless: each message is judged in isolation, so an adversary who spreads a single attack across dozens of sessions slips p

GPT-5.5 and GPT-5.5 Pro are now available in Hermes Agent through the Nous Portal and OpenRouter providers! (alongside the direct openai oau…

Model ReleasesDGX agent

GPT-5.5 and GPT-5.5 Pro are now available in Hermes Agent through the Nous Portal and OpenRouter providers! (alongside the direct openai oauth provider from yesterday) GPT-5.5 and GPT-5.5 Pro are now

GPT-5.5 is a giant leap forward for handling ambiguity compared to previous GPT models. As Windsurf 2.0 focuses more on parallel agents, thi…

Model ReleasesDGX agent

GPT-5.5 is a giant leap forward for handling ambiguity compared to previous GPT models. As Windsurf 2.0 focuses more on parallel agents, this model is key for long-horizon tasks — it excels at underst

GPT-5.5 now available in Deep Agents!

Model ReleasesDGX agent

GPT-5.5 now available in Deep Agents! GPT-5.5 is now available in the API. The model brings higher intelligence and stronger token efficiency to complex work, helping tasks get done with fewer retries

Omission Constraints Decay While Commission Constraints Persist in Long-Context LLM Agents

Model ReleasesDGX agent

arXiv:2604.20911v1 Announce Type: cross Abstract: LLM agents deployed in production operate under operator-defined behavioral policies (system-prompt instructions such as prohibitions on credential di

Reinforcement Learning with Foundation Priors: Let the Embodied Agent Efficiently Learn on Its Own

SafetyDGX agent

arXiv:2310.02635v5 Announce Type: replace-cross Abstract: Reinforcement learning (RL) is a promising approach for solving robotic manipulation tasks. However, it is challenging to apply the RL algorit

This is where we are right now. And i’m not gonna lie it feels pretty magical 🧚‍♀️ Qwen3.6 27B running inside of Pi coding agent via Llama.…

Model ReleasesDGX agent

This is where we are right now. And i’m not gonna lie it feels pretty magical 🧚‍♀️ Qwen3.6 27B running inside of Pi coding agent via Llama.cpp on the MacBook Pro For non-trivial tasks on the @huggingf

23 Apr 2026

Accelerating PayPal's Commerce Agent with Speculative Decoding: An Empirical Study on EAGLE3 with Fine-Tuned Nemotron Models

Model ReleasesDGX agent

arXiv:2604.19767v1 Announce Type: cross Abstract: We evaluate speculative decoding with EAGLE3 as an inference-time optimization for PayPal's Commerce Agent, powered by a fine-tuned llama3.1-nemotron-

Earth Day at #GoogleCloudNext, I’m demoing a Sustainability Agent at the @nvidia booth. Built with @Google ADK, @googlegemma, #nemotron, Clo…

Model ReleasesDGX agent

Earth Day at #GoogleCloudNext, I’m demoing a Sustainability Agent at the @nvidia booth. Built with @Google ADK, @googlegemma, #nemotron, Cloud Run, @milvusio , @LangChain & @ollama to reason across im

Interval POMDP Shielding for Imperfect-Perception Agents

SafetyDGX agent

arXiv:2604.20728v1 Announce Type: new Abstract: Autonomous systems that rely on learned perception can make unsafe decisions when sensor readings are misclassified. We study shielding for this setting

Introducing GPT-5.5 A new class of intelligence for real work and powering agents, built to understand complex goals, use tools, check its w…

Model ReleasesDGX agent

Introducing GPT-5.5 A new class of intelligence for real work and powering agents, built to understand complex goals, use tools, check its work, and carry more tasks through to completion. It marks a

OpenAI says GPT-5.5's improvements are strongest in agentic coding, computer use, and early scientific research, which require reasoning across longer contexts (Madison Mills/Axios)

Model ReleasesDGX agent

Madison Mills / Axios: OpenAI says GPT-5.5's improvements are strongest in agentic coding, computer use, and early scientific research, which require reasoning across longer contexts — OpenAI on Thurs

SceneOrchestra: Efficient Agentic 3D Scene Synthesis via Full Tool-Call Trajectory Generation

Model ReleasesDGX agent

arXiv:2604.19907v1 Announce Type: new Abstract: Recent agentic frameworks for 3D scene synthesis have advanced realism and diversity by integrating heterogeneous generation and editing tools. These to

Semantic Prompting: Agentic Incremental Narrative Refinement through Spatial Semantic Interaction

SafetyDGX agent

arXiv:2604.19971v1 Announce Type: cross Abstract: Interactive spatial layouts empower users to synthesize information and organize findings for sensemaking. While Large Language Models (LLMs) can auto

22 Apr 2026

CentaurTA Studio: A Self-Improving Human-Agent Collaboration System for Thematic Analysis

SafetyDGX agent

arXiv:2604.18589v1 Announce Type: cross Abstract: Thematic analysis is difficult to scale: manual workflows are labor-intensive, while fully automated pipelines often lack controllability and transpar

Michele Catasta (@pirroh) at Google Cloud Next talking about what's new with Gemini, agentic applications, and what's next for builders. Fre…

Model ReleasesDGX agent

Michele Catasta (@pirroh) at Google Cloud Next talking about what's new with Gemini, agentic applications, and what's next for builders. Fresh off Replit being named Google Cloud Partner of the Year.

New in Claude Code: /ultrareview (research preview) runs a fleet of bug-hunting agents in the cloud. Findings land in the CLI or Desktop aut…

Model ReleasesDGX agent

New in Claude Code: /ultrareview (research preview) runs a fleet of bug-hunting agents in the cloud. Findings land in the CLI or Desktop automatically. Run it before merging critical changes—auth, dat

Qwen3.6 27B is now in LM Studio! Vision, reasoning, and agentic tool calling - running locally on your computer. Surpasses previous Qwen mod…

Model ReleasesDGX agent

Qwen3.6 27B is now in LM Studio! Vision, reasoning, and agentic tool calling - running locally on your computer. Surpasses previous Qwen models many times its size 🚀👾🔥 https://lmstudio.ai/models/qwen/

Tencent launches an international beta for QClaw, its OpenClaw-based AI agent, and says the Chinese version, launched in March, reached over 1M users in 10 days (T. K. Lin/KrASIA)

Model ReleasesDGX agent

T. K. Lin / KrASIA: Tencent launches an international beta for QClaw, its OpenClaw-based AI agent, and says the Chinese version, launched in March, reached over 1M users in 10 days — OpenClaw's founde

The Essence of Balance for Self-Improving Agents in Vision-and-Language Navigation

SafetyDGX agent

arXiv:2604.19064v1 Announce Type: new Abstract: In vision-and-language navigation (VLN), self-improvement from policy-induced experience, using only standard VLN action supervision, critically depends

Visual Reasoning Agent: Robust Vision Systems in Remote Sensing via Inference-Time Scaling

Model ReleasesDGX agent

arXiv:2509.16343v2 Announce Type: replace-cross Abstract: Building robust vision systems for high-stakes domains such as remote sensing requires stronger visual reasoning than what single-pass inferen

We need open traces so that everyone can train open agent models! cc @steipete @badlogicgames @thdxr @matanSF @hwchase17

Model ReleasesDGX agent

We need open traces so that everyone can train open agent models! cc @steipete @badlogicgames @thdxr @matanSF @hwchase17 People are misreading the SpaceX/Cursor deal as an M&A story. It’s actually a b

We're launching two specialized TPUs for the agentic era.

HardwareDGX agent

Google announced Ironwood, its seventh-generation TPU that is twice as power efficient as the previous generation, alongside specialized hardware designed to support the emerging agentic AI era. Ironw

21 Apr 2026

Beyond Verifiable Rewards: Rubric-Based GRM for Reinforced Fine-Tuning SWE Agents

ResearchDGX agent

arXiv:2604.16335v1 Announce Type: new Abstract: Despite recent progress in Large Language Model (LLM) Agents for Software Engineering (SWE) tasks, end-to-end fine-tuning typically relies on verifiable

BridgeEQA: Virtual Embodied Agents for Real Bridge Inspections

Model ReleasesDGX agent

arXiv:2511.12676v2 Announce Type: replace Abstract: Deploying embodied agents that can answer questions about their surroundings in realistic real-world settings remains difficult, partly due to the s

Evaluating Tool-Using Language Agents: Judge Reliability, Propagation Cascades, and Runtime Mitigation in AgentProp-Bench

Model ReleasesDGX agent

arXiv:2604.16706v1 Announce Type: cross Abstract: Automated evaluation of tool-using large language model (LLM) agents is widely assumed to be reliable, but this assumption has rarely been validated a

Friendly Machines is happening next Tuesday! We’re exploring how Hermes by @NousResearch and Claude are used to build useful agents: – live …

Model ReleasesDGX agent

Friendly Machines is happening next Tuesday! We’re exploring how Hermes by @NousResearch and Claude are used to build useful agents: – live demos – Q&A with @NousResearch team – discuss your workflows

HiGMem: A Hierarchical and LLM-Guided Memory System for Long-Term Conversational Agents

Model ReleasesDGX agent

arXiv:2604.18349v1 Announce Type: new Abstract: Long-term conversational large language model (LLM) agents require memory systems that can recover relevant evidence from historical interactions withou

HUGGING FACE JUST AUTOMATED THEIR ENTIRE POST-TRAINING TEAM WITH AN AGENT. It reads papers, runs GPU experiments, iterates, and builds resea…

Model ReleasesDGX agent

HUGGING FACE JUST AUTOMATED THEIR ENTIRE POST-TRAINING TEAM WITH AN AGENT. It reads papers, runs GPU experiments, iterates, and builds research-backed models autonomously. Pushed a benchmark from 10%

Privacy-R1: Privacy-Aware Multi-LLM Agent Collaboration via Reinforcement Learning

Local AiDGX agent

arXiv:2510.16054v2 Announce Type: replace-cross Abstract: When users submit queries to Large Language Models (LLMs), their prompts can often contain sensitive data, forcing a difficult choice: Send th

Report: Meta will train AI agents by tracking employees' mouse, keyboard use

IndustryDGX agent

Meta is installing tracking software on U.S.-based employees' computers to capture mouse movements, clicks, and keystrokes for training AI models to build autonomous AI agents . The tool focuses on ar

20 Apr 2026

life when you discover an open-source model that runs 300 parallel agents, executes for 12+ hours straight, beats GPT-5.4 and opus 4.6 on mu…

Model ReleasesDGX agent

life when you discover an open-source model that runs 300 parallel agents, executes for 12+ hours straight, beats GPT-5.4 and opus 4.6 on multiple benchmarks... and the weights are on huggingface Medi

🔥 NEW: I just dropped my free Claude Cowork in 5 Minutes quick start for business professionals who want to use AI agents without the termi…

Model ReleasesDGX agent

🔥 NEW: I just dropped my free Claude Cowork in 5 Minutes quick start for business professionals who want to use AI agents without the terminal. Cowork is made for less technical, everyday business wor

Subliminal Transfer of Unsafe Behaviors in AI Agent Distillation

SafetyDGX agent

arXiv:2604.15559v1 Announce Type: new Abstract: Recent work on subliminal learning demonstrates that language models can transmit semantic traits through data that is semantically unrelated to those t

19 Apr 2026

ollama launch copilot Ollama now supports GitHub's Copilot CLI, the terminal agent that works directly with repositories on GitHub. You can …

Local AiDGX agent

ollama launch copilot Ollama now supports GitHub's Copilot CLI, the terminal agent that works directly with repositories on GitHub. You can use it to: Explore issues and PRs. Search across repos by la

18 Apr 2026

I prefer my design tool to be more closely integrated with where my agents work. I spent a few hours building my own design tool (inspired b…

Model ReleasesDGX agent

I prefer my design tool to be more closely integrated with where my agents work. I spent a few hours building my own design tool (inspired by Claude Design) inside my orchestrator. I can use this with

17 Apr 2026

AccelOpt: A Self-Improving LLM Agentic System for AI Accelerator Kernel Optimization

Model ReleasesDGX agent

arXiv:2511.15915v2 Announce Type: replace-cross Abstract: We present AccelOpt, a self-improving large language model (LLM) agentic system that autonomously optimizes kernels for emerging AI acclerator

AgentIAD: Agentic Industrial Anomaly Detection via Adaptive Memory Augmentation

Model ReleasesDGX agent

arXiv:2512.13671v2 Announce Type: replace Abstract: Industrial anomaly detection (IAD) is challenging due to the subtle and highly localized nature of many defects, which single-pass vision--language

ContractSkill: Repairable Contract-Based Skills for Multimodal Web Agents

Model ReleasesDGX agent

arXiv:2603.20340v3 Announce Type: replace-cross Abstract: Self-generated skills for web agents are often unstable and can even hurt performance relative to direct acting. We argue that the key bottlen

ECM Contracts: Contract-Aware, Versioned, and Governable Capability Interfaces for Embodied Agents

Model ReleasesDGX agent

arXiv:2604.13097v1 Announce Type: cross Abstract: Embodied agents increasingly rely on modular capabilities that can be installed, upgraded, composed, and governed at runtime. Prior work has introduce

GeoAgentBench: A Dynamic Execution Benchmark for Tool-Augmented Agents in Spatial Analysis

Model ReleasesDGX agent

arXiv:2604.13888v1 Announce Type: new Abstract: The integration of Large Language Models (LLMs) into Geographic Information Systems (GIS) marks a paradigm shift toward autonomous spatial analysis. How

LLMOrbit: A Circular Taxonomy of Large Language Models -From Scaling Walls to Agentic AI Systems

Model ReleasesDGX agent

arXiv:2601.14053v2 Announce Type: replace-cross Abstract: The field of artificial intelligence has undergone a revolution from foundational Transformer architectures to reasoning-capable systems appro

we just shipped support for subagents with `deepagents deploy`! add an agents/ dir to your project with an AGENTS.md per specialized subagen…

ApplicationsDGX agent

we just shipped support for subagents with `deepagents deploy`! add an agents/ dir to your project with an AGENTS.md per specialized subagent. subagents are great for task delegation with isolated/opt

← Previous
1…135136137138139…300
Next →