AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlog
84,619Total entries
1Added by human
84,618Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,976 results
Local Ai

Yeah I agree. The big winner is going to be Ollama: I've offloaded all my supervisory, code review, and ontology learning agents to my $20 a…

DGX agent

Yeah I agree. The big winner is going to be Ollama: I've offloaded all my supervisory, code review, and ontology learning agents to my $20 a month Ollama subscription, because I keep getting weight li

local-aiollama--x
20 Jun 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Safety

HERO: Hindsight-Enhanced Reflection from Environment Observations for Agentic Self-Distillation

DGX agent

arXiv:2606.11559v1 Announce Type: new Abstract: Reinforcement learning typically improves multi-turn agent capabilities through the terminal outcome of the trajectories, which makes it difficult to de

safetyarxiv-cs-ai
11 Jun 2026
Safety

MASK: Multi-Agent Semantic K-Scheduling for Risk-Sensitive 6G Robotics

DGX agent

arXiv:2606.11249v1 Announce Type: cross Abstract: Realizing the vision of 6G connected robotics requires reconciling high-performance collaborative control with the rigid spectral limitations of physi

safetyarxiv-cs-lg
11 Jun 2026
Local Ai

MODF-SIR: A Multi-agent Omni-modal Distilled Framework for Social Intelligence Reasoning

DGX agent

arXiv:2606.12018v1 Announce Type: new Abstract: We propose a multi-agent collaborative framework built upon a lightweight Multimodal Large Language Model (MLLM), specifically designed for social intel

local-aiarxiv-cs-ai
11 Jun 2026
Model Releases

The model subsidies will eventually end and this workflow of “creating loops that will prompt your agents” will result in massive amounts of…

DGX agent

The model subsidies will eventually end and this workflow of “creating loops that will prompt your agents” will result in massive amounts of code that’s not well understood that you will have to pay l

model-releasesjerry-liu--x
11 Jun 2026
Model Releases

A History-Aware Visually Grounded Critic for Computer Use Agents

DGX agent

arXiv:2606.11078v1 Announce Type: new Abstract: Various test-time interventions for Computer Use Agents (CUAs), including critic models, have been developed to improve performance through pre-executio

model-releasesarxiv-cs-ai
10 Jun 2026
Tools

Coding agents break when models are 'almost' bug-free. But almost valid JSON is just not the same valid JSON. Fun piece here from @akshay_pa…

DGX agent

Coding agents break when models are 'almost' bug-free. But almost valid JSON is just not the same valid JSON. Fun piece here from @akshay_pachaar shows why SFT can't fix this, and how GRPO trains agai

toolsfireworks-ai--x
10 Jun 2026
Model Releases

EEVEE: Towards Test-time Prompt Learning in the Real World for Self-Improving Agents

DGX agent

arXiv:2606.11182v1 Announce Type: cross Abstract: In this paper, we propose EEVEE, the first multi-dataset test-time prompt learning framework for LLM agents, enabling test-time prompt learning under

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

Effective Reinforcement Learning for Agentic Search by Recycling Zero-Variance Queries During Training

DGX agent

arXiv:2606.10709v1 Announce Type: cross Abstract: The use of GRPO-style algorithms has become the standard strategy for training LLM search agents under outcome-only rewards. With these algorithms, a

model-releasesarxiv-cs-ai
10 Jun 2026
Safety

Human-AI Coordination Zones: A Framework for Designing Human-in-the-Loop Experiences with Agentic AI

DGX agent

arXiv:2606.09848v1 Announce Type: cross Abstract: As generative and agentic AI becomes embedded in everyday products, practitioners face a persistent challenge: how to design human-AI coordination --

safetyarxiv-cs-ai
10 Jun 2026
Model Releases

Less Context, More Accuracy: A Bi-Temporal Memory Engine for LLM Agents Where a Lean Retrieved Context Beats the Full History

DGX agent

arXiv:2606.09900v1 Announce Type: cross Abstract: Long-term memory is the missing layer for LLM agents: across sessions they forget, and the common workaround -- replaying the whole history into the p

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

Moonshine: An Autonomous Mathematical Research Agent Centered on Conjecture Generation

DGX agent

arXiv:2606.10806v1 Announce Type: new Abstract: Moonshine is an autonomous agent whose central objective is to generate mathematical conjectures. Its core capability is to extract structure from class

model-releasesarxiv-cs-ai
10 Jun 2026
Applications

The FBI seizes 13 domains allegedly tied to fake consulting firms that sought information from US government and military employees for suspected Chinese agents (A.J. Vicens/Reuters)

DGX agent

A.J. Vicens / Reuters: The FBI seizes 13 domains allegedly tied to fake consulting firms that sought information from US government and military employees for suspected Chinese agents — Federal author

applicationstechmeme
10 Jun 2026
Model Releases

What Fits (Into Few Tokens) Doesn't Overfit: Compression and Generalization in ML Research Agents

DGX agent

arXiv:2606.11045v1 Announce Type: new Abstract: Reusing a held-out benchmark adaptively should, in principle, invite overfitting. Yet benchmark-driven machine learning (ML) has produced surprisingly l

model-releasesarxiv-cs-ai
10 Jun 2026
Hardware

Accelerating Federated Learning Research with AI Agents and NVIDIA FLARE Auto-FL

DGX agent

NVIDIA FLARE Auto-FL automates federated learning research by constraining agent actions through a control plane, enforcing fixed benchmark contracts, and using an experiment ledger to ensure reproduc

hardwarenvidia-developer
9 Jun 2026
Safety

Can the Environment Speak for Itself? T^{2}-GRPO: A Turn-Trajectory Group Relative Policy Optimization for Caregiver Agents

DGX agent

arXiv:2606.08875v1 Announce Type: new Abstract: Optimizing large language models (LLMs) for long-horizon caregiver agents requires balancing delayed task objectives with immediate environment dynamics

safetyarxiv-cs-ai
9 Jun 2026
Tools

I built a mobile app, promo video, and pitch deck for my travel app at the same time using Replit's parallel agents 👇

DGX agent

A developer used Replit's parallel agents feature to simultaneously create multiple deliverables for a travel app startup: a functional mobile application, promotional video, and investor pitch deck.

toolsreplit--x
9 Jun 2026
Model Releases

Introducing the Fast Gemma Challenge with Hugging Face Over the next few days, dozens of agents will collaborate to make Gemma 4 E4B even fa…

DGX agent

Google and Hugging Face are launching the Fast Gemma Challenge, where multiple agents will collaborate to optimize the performance and speed of Gemma 4 E4B model. The initiative aims to improve the ef

model-releasesclem-delangue--x
9 Jun 2026
Local Ai

Major update: my Ollama coding assistant now has a full autonomous agent + biggest release yet + qwen2.5-coder:7b

DGX agent

A user announced a major update to their Ollama coding assistant featuring Cline, an autonomous coding agent that operates within IDEs and can create, edit files, execute commands, and interact with t

local-air-ollama
9 Jun 2026
Model Releases

Rosetta Memory: Adaptive Memory for Cross-LLM Agents

DGX agent

arXiv:2606.07711v1 Announce Type: cross Abstract: Memory is the key component for transforming a stateless LLM into a persistent, evolving agent through experience accumulation, long-horizon planning,

model-releasesarxiv-cs-ai
9 Jun 2026
Local Ai

Three MLX videos dropped at WWDC: Running agents locally by @angeloskath https://www.youtube.com/watch?v=wykPErJ8M-8 Distributed inference a…

DGX agent

Three MLX videos dropped at WWDC: Running agents locally by @angeloskath https://www.youtube.com/watch?v=wykPErJ8M-8 Distributed inference and training by Tatiana Likhomanenko https://www.youtube.com/

local-aiollama--x
9 Jun 2026
Safety

Attack Selection in Agentic AI Control Evaluations Meaningfully Decreases Safety

DGX agent

arXiv:2606.06529v1 Announce Type: new Abstract: An attacker that strategically chooses when to attack is much harder to catch than one that attacks indiscriminately. AI control is a safety framework f

safetyarxiv-cs-ai
8 Jun 2026
Model Releases

Autonomous heterogeneous catalyst discovery with a self-evolving multi-agent digital twin

DGX agent

arXiv:2606.05050v1 Announce Type: cross Abstract: Theoretical heterogeneous catalysis promises rapid catalyst discovery, yet computational and machine-learning predictions often deviate from experimen

model-releasesarxiv-cs-ai
8 Jun 2026
Model Releases

Google upgrades NotebookLM, which now runs on Gemini 3.5 and Antigravity, to deliver new agentic capabilities and more advanced reasoning for AI Ultra users (Ivan Mehta/TechCrunch)

DGX agent

Ivan Mehta / TechCrunch: Google upgrades NotebookLM, which now runs on Gemini 3.5 and Antigravity, to deliver new agentic capabilities and more advanced reasoning for AI Ultra users — Google on Monday

model-releasestechmeme
8 Jun 2026
Model Releases

Great tips. In practice, this is how it roughly looks to run agents autonomously for hours or days. /goal or /loop to keep it going. Verific…

DGX agent

Great tips. In practice, this is how it roughly looks to run agents autonomously for hours or days. /goal or /loop to keep it going. Verification is crucial here. Seeing a number of benchmarks showing

model-releasesdair-ai--x
8 Jun 2026
Local Ai

Hi everyone, I'm building a local AI agent with Ollama and exploring dynamic PDF extraction. Since Ollama can't directly process PDFs, I'm extracting text and passing it via prompts. Should I use PDFPlumber, a vector database (RAG), or another approach for accurate document understanding? Guide me !

DGX agent

User seeks guidance on PDF processing methods for local AI agents built with Ollama, comparing approaches like PDFPlumber extraction, vector database RAG systems, and alternative techniques for accura

local-air-ollama
8 Jun 2026
Model Releases

Interesting model here 35b a3b trained for agentic use It gets 60.7 on Terminal Bench2 qwen 3.6 27b gets 59.3 Essentially the same Going to …

DGX agent

Interesting model here 35b a3b trained for agentic use It gets 60.7 on Terminal Bench2 qwen 3.6 27b gets 59.3 Essentially the same Going to have to try it out https://huggingface.co/nex-agi/Nex-N2-min

model-releasesclem-delangue--x
8 Jun 2026
Model Releases

It’s safe to close your laptop now: Hosting coding agents on Amazon Bedrock AgentCore

DGX agent

Amazon Bedrock AgentCore Runtime gives each agent session its own isolated microVM with a persistent workspace, secure tool access through Gateway, and built-in observability—so you can run Claude Cod

model-releasesaws-ml-blog
8 Jun 2026
Safety

MADRAG: Multi-Agent Debate with Retrieval-Augmented Generation for Training-Free Analytic Essay Scoring

DGX agent

arXiv:2606.06754v1 Announce Type: cross Abstract: We present MADRAG, a training-free framework for analytic essay scoring that combines multi-agent reasoning with retrieval-augmented grounding. Unlike

safetyarxiv-cs-cl
8 Jun 2026
Model Releases

MemDreamer: Decoupling Perception and Reasoning for Long Video Understanding via Hierarchical Graph Memory and Agentic Retrieval Mechanism

DGX agent

arXiv:2606.07512v1 Announce Type: cross Abstract: Current Vision-Language Models struggle with hours-long videos because processing full-length visual sequences induces prohibitive token explosion and

model-releasesarxiv-cs-ai
8 Jun 2026
Safety

SlimSearcher: Training Efficiency-Aware Web Agents via Adaptive Reward Gating

DGX agent

arXiv:2606.07074v1 Announce Type: cross Abstract: Deep research agents have demonstrated remarkable capabilities in complex information-seeking tasks, yet this power comes at a steep computational cos

safetyarxiv-cs-ai
8 Jun 2026
Industry

Transforming solar and wind maintenance reports with Genie and AI agents

DGX agent

This article explores how Databricks' Genie product and AI agents can automate and improve the processing of maintenance reports for renewable energy installations, specifically solar and wind facilit

industrydatabricks
8 Jun 2026
Local Ai

Need guidance setting up Local AI, Agents, MCP & RAG on an all-AMD Linux rig (7900 XTX / CachyOS)

DGX agent

This post seeks guidance on configuring local AI infrastructure on an AMD-based Linux system (7900 XTX GPU with CachyOS), specifically covering the setup of large language models via Ollama, AI agents

local-air-ollama
7 Jun 2026
Industry

OpenAI plans to overhaul ChatGPT in the coming weeks, turning it into a superapp with coding tools and AI agents to serve as a gateway to higher-margin products (Cristina Criddle/Financial Times)

DGX agent

Cristina Criddle / Financial Times: OpenAI plans to overhaul ChatGPT in the coming weeks, turning it into a superapp with coding tools and AI agents to serve as a gateway to higher-margin products — $

industrytechmeme
7 Jun 2026
Model Releases

CangLing-KnowFlow: A Unified Knowledge-and-Flow-fused Agent for Comprehensive Remote Sensing Applications

DGX agent

arXiv:2512.15231v3 Announce Type: replace Abstract: The automated and intelligent processing of massive remote sensing (RS) datasets is critical in Earth observation (EO). Existing automated systems a

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

Critic-Guided Heterogeneous Multi-Agent Reasoning for Reliable Mathematical Problem Solving

DGX agent

arXiv:2606.05704v1 Announce Type: new Abstract: Recent Large Language Models (LLMs) have shown impressive reasoning abilities; but they are still susceptible to hallucinations, intermediate reasoning

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

Data Flow Control: Data Safety Policies for AI Agents

DGX agent

arXiv:2606.05679v1 Announce Type: cross Abstract: Agents increasingly generate SQL, orchestrate pipelines, and automate data analysis on behalf of users. While recent work improves query correctness,

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

Introducing Harness-1, a 20B search agent trained with a state-externalizing harness. > frontier-level long-horizon search, rivaling Opus-4.…

DGX agent

Introducing Harness-1, a 20B search agent trained with a state-externalizing harness. > frontier-level long-horizon search, rivaling Opus-4.6 and outperforming GPT-5.4 > Context-1-level cost and laten

model-releasesclem-delangue--x
6 Jun 2026
Model Releases

TAPO: Tool-Aware Policy Optimization via Credit Transfer for Multimodal Search Agents

DGX agent

arXiv:2606.05784v1 Announce Type: new Abstract: We identify and formally characterize credit misassignment as a systematic failure mode of GRPO in tool-augmented multimodal search agents: its uniform

model-releasesarxiv-cs-ai
6 Jun 2026
Local Ai

TOKI: A Bitemporal Operator Algebra for Contradiction Resolution in LLM-Agent Persistent Memory

DGX agent

arXiv:2606.06240v1 Announce Type: cross Abstract: Persistent memory for an LLM agent is a write-heavy substrate: every belief update is a versioned write, and a new claim may contradict a stored one.

local-aiarxiv-cs-ai
6 Jun 2026
Model Releases

ToolChoiceConfusion: Causal Minimal Tool Filtering for Reliable LLM Agents

DGX agent

arXiv:2606.06284v1 Announce Type: new Abstract: Large language model agents increasingly rely on external tools, but larger tool menus can reduce reliability and efficiency by increasing wrong-tool ca

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

When Should Memory Stay Silent: Measuring Memory-Use Boundaries in Memory-Augmented Conversational Agents

DGX agent

arXiv:2606.06055v1 Announce Type: new Abstract: Long-term memory enables language model agents to support personalized interactions, but it remains unclear when available memories warrant integration

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

When Tools Fail: Benchmarking Dynamic Replanning and Anomaly Recovery in LLM Agents

DGX agent

arXiv:2606.05806v1 Announce Type: new Abstract: Existing benchmarks evaluate Tool-Integrated Reasoning (TIR) in LLMs on idealized ''happy paths'', largely overlooking real-world tool failures. We intr

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

ArcANE: Do Role-Playing Language Agents Stay in Character at the Right Time?

DGX agent

arXiv:2606.05553v1 Announce Type: new Abstract: Role-playing language agents (RPLAs) should play characters whose values and behavior evolve as the story progresses, not maintain a fixed persona. Exis

model-releasesarxiv-cs-cl
5 Jun 2026
Model Releases

Asuka-Bench: Benchmarking Code Agents on Underspecified User Intent and Multi-Round Refinement

DGX agent

arXiv:2606.05920v1 Announce Type: cross Abstract: Existing code-generation benchmarks score a single mapping from a complete prompt to a one-shot output. However, real web development is different. Us

model-releasesarxiv-cs-cl
5 Jun 2026
Safety

EMBER: Efficient Memory via Budgeted Evidence Retention for Long-Horizon Agents

DGX agent

arXiv:2606.05894v1 Announce Type: new Abstract: Long-horizon agents can archive large histories, but future answers still incur retrieval, rereading, and context costs. When retained memory misses ans

safetyarxiv-cs-cl
5 Jun 2026
Industry

Storyboarding in Grok @imagine was pretty fun. I really like how the agents help visualize some of the important historical events. I was tr…

DGX agent

Storyboarding in Grok @imagine was pretty fun. I really like how the agents help visualize some of the important historical events. I was trying to iterate through the clothing and voices from the anc

industryelon-musk--x
5 Jun 2026
Safety

Toward Culturally Aligned LLMs through Ontology-Guided Multi-Agent Reasoning

DGX agent

arXiv:2601.21700v3 Announce Type: replace Abstract: Large Language Models (LLMs) increasingly support culturally sensitive decision making, yet often exhibit misalignment due to skewed pretraining dat

safetyarxiv-cs-cl
5 Jun 2026
← Previous
1…162163164165166…375
Next →