AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,963 results
Agents

“We believe in intelligence as a public good before everything else.” Here’s my new episode with @karan4d, who co-founded @NousResearch and …

DGX agent

“We believe in intelligence as a public good before everything else.” Here’s my new episode with @karan4d, who co-founded @NousResearch and helped build Hermes, the #1 personal agent and AI app on Ope

agentsnous-research--x
2 Aug 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Functional Cache Grafting: Robust and Rapid Code-Policy Synthesis for Embodied Agents

DGX agent

arXiv:2606.13097v2 Announce Type: replace-cross Abstract: Code-writing large language models (CodeLLMs) generate executable code policies for embodied agents by translating natural language goals and

safetyarxiv-cs-ai
31 Jul 2026
Local Ai

I built a hybrid Transformer–SSM LLM agent with a local CLI, active control, and run receipts

DGX agent

I’m one of the builders of LOLM, a hybrid Transformer–SSM model and agent system from Qira. The model separates surface token processing from persistent latent-state tracking. An NFET controller can s

local-air-ollama
31 Jul 2026
Model Releases

'Intelligence too cheap to meter' battle is on! Given that DeepSeek-V4-Flash-Preview is already great for agentic tasks, there is no doubt t…

DGX agent

'Intelligence too cheap to meter' battle is on! Given that DeepSeek-V4-Flash-Preview is already great for agentic tasks, there is no doubt this new checkpoint must be an absolute beast. 20+ point jump

model-releasesdair-ai--x
31 Jul 2026
Agents

SkillSmith: Learning to Compose Parametric Skills and Textual Knowledge

DGX agent

arXiv:2607.27497v1 Announce Type: new Abstract: Agentic systems driven by large language models (LLMs) regularly feature two key mechanisms to autonomously solve complex problems: synthesizing text-ba

agentsarxiv-cs-cl
31 Jul 2026
Safety

AgentGFM: A Graph Foundation Model with Node-Agent Information-Flow Control

DGX agent

arXiv:2607.26533v1 Announce Type: new Abstract: Graph Foundation Models (GFMs) aim to learn transferable knowledge from multi-domain graphs and adapt to unseen scenarios. As a fundamental source of re

safetyarxiv-cs-lg
30 Jul 2026
Model Releases

MiniIO debuts AIStor Memory, the long-term memory AI agents need to scale safely

DGX agent

Object storage software company MiniIO Inc. says it has cracked the persistent memory problem for artificial intelligence agents with the launch of a new offering called AIStor Memory. Whereas convent

model-releasessiliconangle
30 Jul 2026
Safety

The narrative: Blame the Agent, instead of the Agency that told him to “apply all your powers and told to achieve this win.” Ironically, man…

DGX agent

The narrative: Blame the Agent, instead of the Agency that told him to “apply all your powers and told to achieve this win.” Ironically, many forefront members of the AI-Safety community, in their fer

safetyyann-lecun--x
30 Jul 2026
Local Ai

A Control System, a Dataset, and a Recipe for Making Frozen LLM Agents Learn a Domain

DGX agent

arXiv:2607.25415v1 Announce Type: new Abstract: Production LLM agents are increasingly assembled from a frozen model wrapped in a harness: a prompt template, a tool set, a memory/retrieval layer, a pl

local-aiarxiv-cs-ai
29 Jul 2026
Model Releases

Addressable Recall Compaction for Long Context-Window Control in AI Agents

DGX agent

arXiv:2607.25066v1 Announce Type: new Abstract: Long-horizon LLM agents accumulate reasoning traces, actions, and tool observations that can eventually exceed a model's fixed context window. Existing

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

Evaluating VLMs for Autonomous Agent-Driven Geometry Clipping Detection in Video Game QA

DGX agent

arXiv:2607.25921v1 Announce Type: cross Abstract: In this work, we study the use of Vision-Language Models (VLMs) for anomaly detection in an agent-driven game Quality Assurance (QA) pipeline focusing

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

GAUGE: Grading Agent-Built Financial Models Without a Golden Answer

DGX agent

arXiv:2607.24889v1 Announce Type: cross Abstract: Financial models combine public disclosures with analyst assumptions to produce forecasts and valuations. While some components can be checked mechani

model-releasesarxiv-cs-ai
29 Jul 2026
Agents

Generate Autonomous Business Insights with AI Agent and MCP Servers

DGX agent

Learn how Amazon Bedrock AgentCore delivers autonomous, cross-system business intelligence through configuration rather than custom code. Using pre-built MCP server connectors, fine-grained access con

agentsaws-ml-blog
29 Jul 2026
Research

MemLens: A Value-Aware Memory Management System with Interactive Analytics for LLM-based Agents

DGX agent

arXiv:2607.25992v1 Announce Type: cross Abstract: Recently, memory management has become a key infrastructure for LLM-based agents, as it directly affects long-horizon reasoning, personalized response

researcharxiv-cs-ai
29 Jul 2026
Safety

Shared Voxel-Map-Based Cooperative Indoor UAV Guidance with a Multi-Agent Soft Actor-Critic Controller

DGX agent

arXiv:2607.25728v1 Announce Type: cross Abstract: This paper presents a cooperative indoor UAV guidance framework that combines a shared voxel-map world model with a multi-agent Soft Actor-Critic (MAS

safetyarxiv-cs-ai
29 Jul 2026
Safety

VetClaw: An Edge-Cloud Multimodal Agentic System for Veterinary Disease Screening

DGX agent

arXiv:2607.26042v1 Announce Type: new Abstract: We present VetClaw, an edge-cloud multimodal agentic system for early veterinary disease screening. VetClaw uses a camera module as an edge sensing devi

safetyarxiv-cs-cv
29 Jul 2026
Agents

Delegation Intelligence in Deep Search: A Controllable Framework for Disentangled Capability Diagnosis

DGX agent

arXiv:2607.23524v1 Announce Type: new Abstract: Deep search is becoming a core capability of modern agent systems, yet it is typically evaluated solely based on end-to-end answer accuracy. This couple

agentsarxiv-cs-ai
28 Jul 2026
Agents

Lexical discovery in unknown environments orchestrated by Large Language Models

DGX agent

arXiv:2607.22591v1 Announce Type: new Abstract: Populations of autonomous agents deployed in unknown environments (e.g. planetary or deep-sea exploration) must develop shared vocabularies to refer to

agentsarxiv-cs-ai
28 Jul 2026
Local Ai

ReflexTrack: A Feedback-Driven Agent for Training-Free Referring Video Object Segmentation

DGX agent

arXiv:2607.24098v1 Announce Type: new Abstract: Referring video object segmentation (RVOS) requires segmenting a target specified by natural language throughout a video. Recent agentic approaches comb

local-aiarxiv-cs-cv
28 Jul 2026
Tutorials

AI customer service: strategy, agents, and solutions guide

DGX agent

This guide explains how AI customer service operates, focusing on the use of AI agents and sentiment analysis to interpret and respond to customer interactions. It provides practical instructions for

tutorialsdatabricks
27 Jul 2026
Research

Really interesting concept -- instead of optimizing against static benchmarks or internally generated reward models, agents are evaluated th…

DGX agent

Really interesting concept -- instead of optimizing against static benchmarks or internally generated reward models, agents are evaluated through real economic interactions. Using an external market a

researchfrancois-chollet--x
27 Jul 2026
Model Releases

90 agentic bakeoff runs: ThinkingCap vs Fable Fusion vs stock Qwen3.6-27B

DGX agent

Last week someone here said ThinkingCap and Fable Fusion 'really do beat the OG' for agentic work, so I ran it: 6 self-grading tasks, 5 reps, 3 models, 90 isolated runs. Tooling, since that's half the

model-releasesr-localllama
26 Jul 2026
Model Releases

Agentic coding without the cloud: evaluating open-weight large language models on longitudinal data preparation tasks

DGX agent

arXiv:2607.21482v1 Announce Type: new Abstract: Large language models (LLMs) and agents are now widely used tools in code development, with data typically sent to third-party cloud-based models. Their

model-releasesarxiv-cs-ai
24 Jul 2026
Safety

AttriMem: Attribution-Guided Process Feedback for Agent Memory Learning

DGX agent

arXiv:2607.21106v1 Announce Type: new Abstract: Effective memory is crucial for LLM agents, yet constructing it effectively remains challenging. A memory-construction policy decides what information t

safetyarxiv-cs-ai
24 Jul 2026
Model Releases

CAMeR: Keyword-Gated Hybrid Activation for Adaptive Memory Retention in LLM Agents

DGX agent

arXiv:2607.20458v1 Announce Type: cross Abstract: Large language model (LLM) agents operating over extended dialogues accumulate vast amounts of information, yet existing memory systems either retain

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

CANN Bench: Benchmarking Agent Generated Kernels against Real NPU and Algorithmic Limits

DGX agent

arXiv:2607.20518v1 Announce Type: new Abstract: AI agents are now capable of writing, compiling, and iteratively optimizing low-level operator kernels on different hardware platforms. Existing benchma

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

DFAH-Bench: Benchmarking Observable Agent Instability in Financial Decision-Making

DGX agent

arXiv:2607.20491v1 Announce Type: new Abstract: Standard evaluation benchmarks measure what a tool-using agent decides, not whether it arrives at that decision through the same process each time. We i

model-releasesarxiv-cs-ai
24 Jul 2026
Safety

From Agent Failures to Text Policies: What Works and What Breaks

DGX agent

arXiv:2607.20668v1 Announce Type: cross Abstract: TextGrad improves language-model systems by revising text from feedback. Its core thesis is that natural-language feedback can act as a gradient for o

safetyarxiv-cs-ai
24 Jul 2026
Model Releases

GuardianAgentBench: Where Agents Fail and How to Guard Them

DGX agent

arXiv:2607.20982v1 Announce Type: new Abstract: As large language model agents increasingly operate autonomously with access to tools and external environments, ensuring their safe and reliable behavi

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

MemTools: A Unified Research Framework for Interoperable Agent Memory

DGX agent

arXiv:2607.21404v1 Announce Type: new Abstract: While memory systems are essential for agent architectures, pervasive architectural fragmentation restricts systematic research. Existing implementation

model-releasesarxiv-cs-cl
24 Jul 2026
Agents

MKEvolve: A Modular Multi-Agent Framework for Kernel Code Generation

DGX agent

arXiv:2607.20501v1 Announce Type: new Abstract: Despite rapid progress in LLM-based code generation, writing correct and performant kernels for hardware accelerators remains a key bottleneck in scalin

agentsarxiv-cs-ai
24 Jul 2026
Model Releases

People are using Minecraft farms as AI agent benchmarks

DGX agent

Someone modelled sugarcane farming as an integer program. See, sugarcane only grows next to water. Water costs one tile and can feed at most four cane tiles. The layout therefore becomes a coverage pr

model-releasesr-chatgpt
24 Jul 2026
Safety

Regulating autonomous and agentic AI

DGX agent

arXiv:2607.21345v1 Announce Type: new Abstract: Regulating activities where regulatees use autonomous and agentic AI is challenging. Regulatory assumptions about regulatee knowledge and control no lon

safetyarxiv-cs-ai
24 Jul 2026
Model Releases

Student-Centered Distillation Narrows the Agentic Gap Between Small and Large LLMs

DGX agent

arXiv:2509.14257v3 Announce Type: replace-cross Abstract: Large Language Model agents achieve strong performance on multi-step reasoning and tool-use tasks, but their impressive capabilities typically

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

ETPDesigner: Multi-Agent Orchestration for Interactive Multimodal Electronic Theater Program

DGX agent

arXiv:2607.19947v1 Announce Type: new Abstract: Electronic Theater Programs (ETPs) serve as critical promotional media in the performing arts, comprising a multi-page collection of heterogeneous visua

model-releasesarxiv-cs-cv
23 Jul 2026
Model Releases

NEXUS: Structured Runtime Safety for Tool-Using LLM Agents

DGX agent

arXiv:2607.19356v1 Announce Type: new Abstract: Tool-using LLM agents increasingly execute high-impact actions, making runtime safety monitoring essential. We present NEXUS (Neural EXecution Utility a

model-releasesarxiv-cs-ai
23 Jul 2026
Industry

OpenAI introduces Presence to help enterprises build AI agents

DGX agent

OpenAI Group PBC today introduced a product called Presence that enterprises can use to build artificial intelligence agents. The company is using the software to power its call center. According to O

industrysiliconangle
23 Jul 2026
Agents

OpenAI should release a detailed transcript from the Hugging Face hacking incident -- it would be helpful for the field learn from. Did the …

DGX agent

OpenAI should release a detailed transcript from the Hugging Face hacking incident -- it would be helpful for the field learn from. Did the top-level agent know about the hacking, or was there some 'v

agentsclem-delangue--x
23 Jul 2026
Agents

// Programmatic Memory Enables Long-Horizon Reasoning // Keep the entire interaction log and search it. It works great and beats bespoke mem…

DGX agent

// Programmatic Memory Enables Long-Horizon Reasoning // Keep the entire interaction log and search it. It works great and beats bespoke memory harnesses on long-horizon tasks. New research introduces

agentsdair-ai--x
23 Jul 2026
Safety

Stress Testing Concept Erasure with Large Language Model Agents

DGX agent

arXiv:2607.17890v2 Announce Type: replace Abstract: Concept erasure aims to remove semantic concepts from a trained generative model and is increasingly important for responsible AI deployment. Howeve

safetyarxiv-cs-ai
23 Jul 2026
Model Releases

The first known runaway AI agent - or a very bad marketing stunt?

DGX agent

The first known runaway AI agent - or a very bad marketing stunt? Martin Alderson's commentary on the OpenAI accidental cyberattack against Hugging Face includes a couple of details I hadn't considere

model-releasessimon-willison
23 Jul 2026
Safety

Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review

DGX agent

arXiv:2507.10142v2 Announce Type: replace Abstract: Multi-Agent Reinforcement Learning (MARL) has achieved strong performance in simulated benchmarks, yet real deployments often violate the assumption

safetyarxiv-cs-ai
23 Jul 2026
Tutorials

Are structured outputs in agents always good? This paper suggests that you might have to take a closer look. Your product's structured outpu…

DGX agent

Are structured outputs in agents always good? This paper suggests that you might have to take a closer look. Your product's structured output surface is measurably more homogeneous than the chat surfa

tutorialsdair-ai--x
22 Jul 2026
Model Releases

microsoft/Fara1.5-27B · Hugging Face

DGX agent

Fara1.5-27B is a multimodal computer use agent (CUA) for web browsers, from Microsoft Research AI Frontiers. It observes the browser through screenshots and acts on the user's behalf by emitting struc

model-releasesr-localllama
22 Jul 2026
Model Releases

OpenAI’s accidental cyberattack against Hugging Face is science fiction that happened

DGX agent

This story is wild. The short version: OpenAI were running a cybersecurity test against an unreleased model, with the model's guardrail features turned off. Rather than solve the test, the model broke

model-releasessimon-willison
22 Jul 2026
Model Releases

This is OpenSWE! It's an OSS coding agent that runs in the cloud and lives in Slack. Use it for coding, general Q&A, planning, etc etc etc. …

DGX agent

This is OpenSWE! It's an OSS coding agent that runs in the cloud and lives in Slack. Use it for coding, general Q&A, planning, etc etc etc. It's truly a jack of all trades (and very widely used at Lan

model-releasesharrison-chase--x
22 Jul 2026
Hardware

NVIDIA Vera CPU: Olympus Cores Built for Maximum Single-Thread Performance in Agentic AI

DGX agent

NVIDIA’s Vera CPU, built around the Olympus core, is engineered for agentic‑AI workloads that rely heavily on single‑thread performance, deep memory‑level parallelism, and efficient handling of irregu

hardwarenvidia-developer
21 Jul 2026
Hardware

At SIGGRAPH, NVIDIA Advances Graphics and Simulation With Agentic and Physical AI

DGX agent

At SIGGRAPH 2026, NVIDIA unveiled a suite of AI‑driven graphics and simulation advances, highlighting neural rendering, agentic and physical AI world models, and real‑time simulation methods. Key rele

hardwarenvidia-blog
20 Jul 2026
← Previous
1…141142143144145…375
Next →