AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,964 results
24 Jun 2026

Multimedia and Visual Analytics in the Agentic Era

Model ReleasesDGX agent

arXiv:2504.06138v3 Announce Type: replace-cross Abstract: Professional users need tools to help them gain actionable insights from large multimedia collections. Foundation models and AI agents have ra

Runlayer, which provides an infrastructure and control layer for enterprise AI agents, raised a 30M Series A led by Felicis, bringing its total funding to 42M (Lily Mae Lazarus/Fortune)

ApplicationsDGX agent

Lily Mae Lazarus / Fortune: Runlayer, which provides an infrastructure and control layer for enterprise AI agents, raised a 30M Series A led by Felicis, bringing its total funding to 42M — When longti

Trase, which is building an operating system and infrastructure layer for AI agents in industries like health care and defense, raised a $107M seed led by Arch (Brock E.W. Turner/Axios)

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
IndustryDGX agent

Brock E.W. Turner / Axios: Trase, which is building an operating system and infrastructure layer for AI agents in industries like health care and defense, raised a 107M seed led by Arch — Trase, an AI

23 Jun 2026

An agentic loop (compile, test, profile, revise) helps. Gemini 3 Pro went from 24 to 35/87 correct, then plateaued after ~20 steps. Feedback…

Model ReleasesDGX agent

An agentic loop (compile, test, profile, revise) helps. Gemini 3 Pro went from 24 to 35/87 correct, then plateaued after ~20 steps. Feedback fixes syntax, not rank coordination, collective ordering, o

Anthropic launches Claude Tag, an agentic AI coworker for Slack that can learn context, give suggestions, and more, in beta for Claude Team and Enterprise tiers (David Gewirtz/ZDNET)

Model ReleasesDGX agent

David Gewirtz / ZDNET: Anthropic launches Claude Tag, an agentic AI coworker for Slack that can learn context, give suggestions, and more, in beta for Claude Team and Enterprise tiers — ZDNET's key ta

Claude Tag is an incredible new form factor for agents, so I think it's going to take some time to figure out the best practices, but these …

Model ReleasesDGX agent

Claude Tag is an incredible new form factor for agents, so I think it's going to take some time to figure out the best practices, but these are some of my favorites 🧵 Introducing Claude Tag, a new way

Evo-RAD: Navigating Rare Retinal Disease Diagnosis via Self-Evolving Agentic Retrieval

Model ReleasesDGX agent

arXiv:2606.22955v1 Announce Type: new Abstract: Large-scale pretrained foundation models have revolutionized general medical screening, but often falter on rare diseases because such conditions are un

Fara-1.5: Scalable Learning Environments for Computer Use Agents

Model ReleasesDGX agent

arXiv:2606.20785v1 Announce Type: cross Abstract: Collecting computer use data from human demonstrations is expensive and slow, motivating the need for scalable generation strategies. This requires tw

HEAS: Hierarchical Evolutionary Agent-Based Simulation Framework for Multi-Objective Policy Search

SafetyDGX agent

arXiv:2508.15555v4 Announce Type: replace-cross Abstract: HEAS is a Python framework that connects agent-based simulation, evolutionary search, and scenario-based evaluation in a single reproducible p

Introducing Claude code for 3D modeling You can now create a production-level 3D models from any agent you are using. Ready to plug into you…

Model ReleasesDGX agent

Introducing Claude code for 3D modeling You can now create a production-level 3D models from any agent you are using. Ready to plug into your own apps or games Made possible by @trysuzanne and @monid_

Nous: A Predictive World Model for Long-Term Agent Memory

Model ReleasesDGX agent

arXiv:2606.22030v1 Announce Type: cross Abstract: We present Nous, a novel agent memory architecture grounded in the principle that knowledge is prediction, not storage. Rather than persisting facts a

Self-Evolution for Multi-Turn Tool-Calling Agents via Divergence-Point Preference Learning

Model ReleasesDGX agent

arXiv:2606.23112v1 Announce Type: new Abstract: Multi-turn tool-using agents must coordinate long-horizon tool sequences while tracking dialogue state and policy constraints. Existing approaches often

SkillHarness: Harnessing Safe Skills for Computer-Use Agents

SafetyDGX agent

arXiv:2606.20636v1 Announce Type: cross Abstract: Computer-Use Agents (CUAs) are increasingly deployed in dynamic interactive environments, creating a growing need for continual skill learning during

With agentic coding, complexity compounds in a mechanical way: unnecessary code ends up in the codebase, moves to the context window, degrad…

Model ReleasesDGX agent

With agentic coding, complexity compounds in a mechanical way: unnecessary code ends up in the codebase, moves to the context window, degrades the model's reasoning abilities, leads to more unnecessar

22 Jun 2026

An interesting new paper by my recent PhD graduate on how AI agents' greed for visible incentives can lead them to abandon their safety alig…

SafetyDGX agent

An interesting new paper by my recent PhD graduate on how AI agents' greed for visible incentives can lead them to abandon their safety alignment. You can read it here: https://arxiv.org/abs/2606.1691

Building pay-per-intelligence for AI agents: How Ampersend uses Amazon Bedrock AgentCore Payments

TutorialsDGX agent

In this post, you will learn how Ampersend built a pay-per-intelligence routing layer on top of Amazon Bedrock AgentCore Payments. AI agents autonomously route tasks to the most effective model, pay p

My parallel agent side-project today was having Claude Code port the new Moebius image pinpointing model to ONNX in order to run it entirely…

Model ReleasesDGX agent

My parallel agent side-project today was having Claude Code port the new Moebius image pinpointing model to ONNX in order to run it entirely in the browser https://simonwillison.net/2026/Jun/22/portin

What’s the best place to rent on demand B200s? Ideally CLI so agents can spin them up

IndustryDGX agent

This post discusses options for renting on-demand B200 GPUs with command-line interface (CLI) capabilities that would allow automated agent-based provisioning. The query seeks recommendations for clou

21 Jun 2026

Voice agents get a lot more interesting when they can use the screen 🔥 This demo runs the full loop on Together AI: STT, voice, and reasoni…

ToolsDGX agent

Voice agents get a lot more interesting when they can use the screen 🔥 This demo runs the full loop on Together AI: STT, voice, and reasoning across Parakeet, MiniMax Speech 2.8, and MiniMax M3. Real-

20 Jun 2026

Yeah I agree. The big winner is going to be Ollama: I've offloaded all my supervisory, code review, and ontology learning agents to my $20 a…

Local AiDGX agent

Yeah I agree. The big winner is going to be Ollama: I've offloaded all my supervisory, code review, and ontology learning agents to my $20 a month Ollama subscription, because I keep getting weight li

11 Jun 2026

HERO: Hindsight-Enhanced Reflection from Environment Observations for Agentic Self-Distillation

SafetyDGX agent

arXiv:2606.11559v1 Announce Type: new Abstract: Reinforcement learning typically improves multi-turn agent capabilities through the terminal outcome of the trajectories, which makes it difficult to de

MASK: Multi-Agent Semantic K-Scheduling for Risk-Sensitive 6G Robotics

SafetyDGX agent

arXiv:2606.11249v1 Announce Type: cross Abstract: Realizing the vision of 6G connected robotics requires reconciling high-performance collaborative control with the rigid spectral limitations of physi

MODF-SIR: A Multi-agent Omni-modal Distilled Framework for Social Intelligence Reasoning

Local AiDGX agent

arXiv:2606.12018v1 Announce Type: new Abstract: We propose a multi-agent collaborative framework built upon a lightweight Multimodal Large Language Model (MLLM), specifically designed for social intel

The model subsidies will eventually end and this workflow of “creating loops that will prompt your agents” will result in massive amounts of…

Model ReleasesDGX agent

The model subsidies will eventually end and this workflow of “creating loops that will prompt your agents” will result in massive amounts of code that’s not well understood that you will have to pay l

10 Jun 2026

A History-Aware Visually Grounded Critic for Computer Use Agents

Model ReleasesDGX agent

arXiv:2606.11078v1 Announce Type: new Abstract: Various test-time interventions for Computer Use Agents (CUAs), including critic models, have been developed to improve performance through pre-executio

Coding agents break when models are 'almost' bug-free. But almost valid JSON is just not the same valid JSON. Fun piece here from @akshay_pa…

ToolsDGX agent

Coding agents break when models are 'almost' bug-free. But almost valid JSON is just not the same valid JSON. Fun piece here from @akshay_pachaar shows why SFT can't fix this, and how GRPO trains agai

EEVEE: Towards Test-time Prompt Learning in the Real World for Self-Improving Agents

Model ReleasesDGX agent

arXiv:2606.11182v1 Announce Type: cross Abstract: In this paper, we propose EEVEE, the first multi-dataset test-time prompt learning framework for LLM agents, enabling test-time prompt learning under

Effective Reinforcement Learning for Agentic Search by Recycling Zero-Variance Queries During Training

Model ReleasesDGX agent

arXiv:2606.10709v1 Announce Type: cross Abstract: The use of GRPO-style algorithms has become the standard strategy for training LLM search agents under outcome-only rewards. With these algorithms, a

Human-AI Coordination Zones: A Framework for Designing Human-in-the-Loop Experiences with Agentic AI

SafetyDGX agent

arXiv:2606.09848v1 Announce Type: cross Abstract: As generative and agentic AI becomes embedded in everyday products, practitioners face a persistent challenge: how to design human-AI coordination --

Less Context, More Accuracy: A Bi-Temporal Memory Engine for LLM Agents Where a Lean Retrieved Context Beats the Full History

Model ReleasesDGX agent

arXiv:2606.09900v1 Announce Type: cross Abstract: Long-term memory is the missing layer for LLM agents: across sessions they forget, and the common workaround -- replaying the whole history into the p

Moonshine: An Autonomous Mathematical Research Agent Centered on Conjecture Generation

Model ReleasesDGX agent

arXiv:2606.10806v1 Announce Type: new Abstract: Moonshine is an autonomous agent whose central objective is to generate mathematical conjectures. Its core capability is to extract structure from class

The FBI seizes 13 domains allegedly tied to fake consulting firms that sought information from US government and military employees for suspected Chinese agents (A.J. Vicens/Reuters)

ApplicationsDGX agent

A.J. Vicens / Reuters: The FBI seizes 13 domains allegedly tied to fake consulting firms that sought information from US government and military employees for suspected Chinese agents — Federal author

What Fits (Into Few Tokens) Doesn't Overfit: Compression and Generalization in ML Research Agents

Model ReleasesDGX agent

arXiv:2606.11045v1 Announce Type: new Abstract: Reusing a held-out benchmark adaptively should, in principle, invite overfitting. Yet benchmark-driven machine learning (ML) has produced surprisingly l

9 Jun 2026

Accelerating Federated Learning Research with AI Agents and NVIDIA FLARE Auto-FL

HardwareDGX agent

NVIDIA FLARE Auto-FL automates federated learning research by constraining agent actions through a control plane, enforcing fixed benchmark contracts, and using an experiment ledger to ensure reproduc

Can the Environment Speak for Itself? T^{2}-GRPO: A Turn-Trajectory Group Relative Policy Optimization for Caregiver Agents

SafetyDGX agent

arXiv:2606.08875v1 Announce Type: new Abstract: Optimizing large language models (LLMs) for long-horizon caregiver agents requires balancing delayed task objectives with immediate environment dynamics

I built a mobile app, promo video, and pitch deck for my travel app at the same time using Replit's parallel agents 👇

ToolsDGX agent

A developer used Replit's parallel agents feature to simultaneously create multiple deliverables for a travel app startup: a functional mobile application, promotional video, and investor pitch deck.

Introducing the Fast Gemma Challenge with Hugging Face Over the next few days, dozens of agents will collaborate to make Gemma 4 E4B even fa…

Model ReleasesDGX agent

Google and Hugging Face are launching the Fast Gemma Challenge, where multiple agents will collaborate to optimize the performance and speed of Gemma 4 E4B model. The initiative aims to improve the ef

Major update: my Ollama coding assistant now has a full autonomous agent + biggest release yet + qwen2.5-coder:7b

Local AiDGX agent

A user announced a major update to their Ollama coding assistant featuring Cline, an autonomous coding agent that operates within IDEs and can create, edit files, execute commands, and interact with t

Rosetta Memory: Adaptive Memory for Cross-LLM Agents

Model ReleasesDGX agent

arXiv:2606.07711v1 Announce Type: cross Abstract: Memory is the key component for transforming a stateless LLM into a persistent, evolving agent through experience accumulation, long-horizon planning,

Three MLX videos dropped at WWDC: Running agents locally by @angeloskath https://www.youtube.com/watch?v=wykPErJ8M-8 Distributed inference a…

Local AiDGX agent

Three MLX videos dropped at WWDC: Running agents locally by @angeloskath https://www.youtube.com/watch?v=wykPErJ8M-8 Distributed inference and training by Tatiana Likhomanenko https://www.youtube.com/

8 Jun 2026

Attack Selection in Agentic AI Control Evaluations Meaningfully Decreases Safety

SafetyDGX agent

arXiv:2606.06529v1 Announce Type: new Abstract: An attacker that strategically chooses when to attack is much harder to catch than one that attacks indiscriminately. AI control is a safety framework f

Autonomous heterogeneous catalyst discovery with a self-evolving multi-agent digital twin

Model ReleasesDGX agent

arXiv:2606.05050v1 Announce Type: cross Abstract: Theoretical heterogeneous catalysis promises rapid catalyst discovery, yet computational and machine-learning predictions often deviate from experimen

Google upgrades NotebookLM, which now runs on Gemini 3.5 and Antigravity, to deliver new agentic capabilities and more advanced reasoning for AI Ultra users (Ivan Mehta/TechCrunch)

Model ReleasesDGX agent

Ivan Mehta / TechCrunch: Google upgrades NotebookLM, which now runs on Gemini 3.5 and Antigravity, to deliver new agentic capabilities and more advanced reasoning for AI Ultra users — Google on Monday

Great tips. In practice, this is how it roughly looks to run agents autonomously for hours or days. /goal or /loop to keep it going. Verific…

Model ReleasesDGX agent

Great tips. In practice, this is how it roughly looks to run agents autonomously for hours or days. /goal or /loop to keep it going. Verification is crucial here. Seeing a number of benchmarks showing

Hi everyone, I'm building a local AI agent with Ollama and exploring dynamic PDF extraction. Since Ollama can't directly process PDFs, I'm extracting text and passing it via prompts. Should I use PDFPlumber, a vector database (RAG), or another approach for accurate document understanding? Guide me !

Local AiDGX agent

User seeks guidance on PDF processing methods for local AI agents built with Ollama, comparing approaches like PDFPlumber extraction, vector database RAG systems, and alternative techniques for accura

Interesting model here 35b a3b trained for agentic use It gets 60.7 on Terminal Bench2 qwen 3.6 27b gets 59.3 Essentially the same Going to …

Model ReleasesDGX agent

Interesting model here 35b a3b trained for agentic use It gets 60.7 on Terminal Bench2 qwen 3.6 27b gets 59.3 Essentially the same Going to have to try it out https://huggingface.co/nex-agi/Nex-N2-min

It’s safe to close your laptop now: Hosting coding agents on Amazon Bedrock AgentCore

Model ReleasesDGX agent

Amazon Bedrock AgentCore Runtime gives each agent session its own isolated microVM with a persistent workspace, secure tool access through Gateway, and built-in observability—so you can run Claude Cod

MADRAG: Multi-Agent Debate with Retrieval-Augmented Generation for Training-Free Analytic Essay Scoring

SafetyDGX agent

arXiv:2606.06754v1 Announce Type: cross Abstract: We present MADRAG, a training-free framework for analytic essay scoring that combines multi-agent reasoning with retrieval-augmented grounding. Unlike

MemDreamer: Decoupling Perception and Reasoning for Long Video Understanding via Hierarchical Graph Memory and Agentic Retrieval Mechanism

Model ReleasesDGX agent

arXiv:2606.07512v1 Announce Type: cross Abstract: Current Vision-Language Models struggle with hours-long videos because processing full-length visual sequences induces prohibitive token explosion and

SlimSearcher: Training Efficiency-Aware Web Agents via Adaptive Reward Gating

SafetyDGX agent

arXiv:2606.07074v1 Announce Type: cross Abstract: Deep research agents have demonstrated remarkable capabilities in complex information-seeking tasks, yet this power comes at a steep computational cos

Transforming solar and wind maintenance reports with Genie and AI agents

IndustryDGX agent

This article explores how Databricks' Genie product and AI agents can automate and improve the processing of maintenance reports for renewable energy installations, specifically solar and wind facilit

7 Jun 2026

Need guidance setting up Local AI, Agents, MCP & RAG on an all-AMD Linux rig (7900 XTX / CachyOS)

Local AiDGX agent

This post seeks guidance on configuring local AI infrastructure on an AMD-based Linux system (7900 XTX GPU with CachyOS), specifically covering the setup of large language models via Ollama, AI agents

OpenAI plans to overhaul ChatGPT in the coming weeks, turning it into a superapp with coding tools and AI agents to serve as a gateway to higher-margin products (Cristina Criddle/Financial Times)

IndustryDGX agent

Cristina Criddle / Financial Times: OpenAI plans to overhaul ChatGPT in the coming weeks, turning it into a superapp with coding tools and AI agents to serve as a gateway to higher-margin products — $

6 Jun 2026

CangLing-KnowFlow: A Unified Knowledge-and-Flow-fused Agent for Comprehensive Remote Sensing Applications

Model ReleasesDGX agent

arXiv:2512.15231v3 Announce Type: replace Abstract: The automated and intelligent processing of massive remote sensing (RS) datasets is critical in Earth observation (EO). Existing automated systems a

Critic-Guided Heterogeneous Multi-Agent Reasoning for Reliable Mathematical Problem Solving

Model ReleasesDGX agent

arXiv:2606.05704v1 Announce Type: new Abstract: Recent Large Language Models (LLMs) have shown impressive reasoning abilities; but they are still susceptible to hallucinations, intermediate reasoning

Data Flow Control: Data Safety Policies for AI Agents

Model ReleasesDGX agent

arXiv:2606.05679v1 Announce Type: cross Abstract: Agents increasingly generate SQL, orchestrate pipelines, and automate data analysis on behalf of users. While recent work improves query correctness,

Introducing Harness-1, a 20B search agent trained with a state-externalizing harness. > frontier-level long-horizon search, rivaling Opus-4.…

Model ReleasesDGX agent

Introducing Harness-1, a 20B search agent trained with a state-externalizing harness. > frontier-level long-horizon search, rivaling Opus-4.6 and outperforming GPT-5.4 > Context-1-level cost and laten

TAPO: Tool-Aware Policy Optimization via Credit Transfer for Multimodal Search Agents

Model ReleasesDGX agent

arXiv:2606.05784v1 Announce Type: new Abstract: We identify and formally characterize credit misassignment as a systematic failure mode of GRPO in tool-augmented multimodal search agents: its uniform

TOKI: A Bitemporal Operator Algebra for Contradiction Resolution in LLM-Agent Persistent Memory

Local AiDGX agent

arXiv:2606.06240v1 Announce Type: cross Abstract: Persistent memory for an LLM agent is a write-heavy substrate: every belief update is a versioned write, and a new claim may contradict a stored one.

ToolChoiceConfusion: Causal Minimal Tool Filtering for Reliable LLM Agents

Model ReleasesDGX agent

arXiv:2606.06284v1 Announce Type: new Abstract: Large language model agents increasingly rely on external tools, but larger tool menus can reduce reliability and efficiency by increasing wrong-tool ca

← Previous
1…129130131132133…300
Next →