AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,958 results
Model Releases

MMShopBench: A Real-Log Benchmark for Multimodal, Multi-Turn Shopping Agents

DGX agent

arXiv:2607.29002v1 Announce Type: new Abstract: Online shoppers increasingly turn to AI shopping assistants, using images and multi-turn dialogue to express and refine product needs that are difficult

model-releasesarxiv-cs-ai
3 Aug 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Real-world mainframe modernization with AI: A safe, scalable path from mainframe to cloud

DGX agent

For too long, enterprises with legacy mainframe estates have been faced with a high-stakes dilemma: continue maintaining their mainframes, essentially kicking the modernization can down the road (they

model-releasesgoogle-cloud-ai
3 Aug 2026
Model Releases

SeekBrain: An Autonomous Multi-Agent System for Accelerating Neuroscience Discovery

DGX agent

arXiv:2607.29347v1 Announce Type: cross Abstract: Modern neuroscience relies on integrating multi-scale, multimodal datasets to uncover the neural principles underlying intelligence. However, analytic

model-releasesarxiv-cs-ai
3 Aug 2026
Local Ai

Zero-Mem: Zero-Token Memory Operations for LLM Agents

DGX agent

arXiv:2607.29377v1 Announce Type: new Abstract: LLM agents need memory to act consistently over long interactions, yet many systems use additional LLM calls to operate that memory. Generating intermed

local-aiarxiv-cs-cl
3 Aug 2026
Model Releases

Real-world reality check on Qwen for autonomous coding agents

DGX agent

TLDR below 👇🏼 I’ve seen a lot of hype around Qwen 3.6 35B and 3.5 120B lately, especially regarding coding and tool-use capabilities. On this subreddit it is the defacto recommended model for everyone

model-releasesr-localllama
2 Aug 2026
Model Releases

DeepSeek V4 Flash 0731 in Hermes Agent and one prompt, took 32 minutes and cost 0.07$, this model is so cheap to the point where 2 dollars c…

DGX agent

**DeepSeek V4 Flash 0731 Performance Test** On July 31 2026, a single prompt executed via the Hermes Agent on DeepSeek V4 Flash 0731 completed in 32 minutes and incurred an estimated cost of 0.07 USD.

model-releasesnous-research--x
31 Jul 2026
Safety

Harness-G: A Graph-Structured Harness for Search Agents

DGX agent

arXiv:2607.27652v1 Announce Type: new Abstract: Reinforcement learning (RL) search agents commonly model retrieval as free-form natural-language query generation and optimize multi-turn interactions u

safetyarxiv-cs-cl
31 Jul 2026
Model Releases

OK, GPT-5.6 Luna is a bit of a beast. Given the 80% price drop today I decided to try it in Datasette Agent, and it's furiously quick and ge…

DGX agent

OK, GPT-5.6 Luna is a bit of a beast. Given the 80% price drop today I decided to try it in Datasette Agent, and it's furiously quick and generates all the SQL, HTML and JavaScript (for Datasette Apps

model-releasessimon-willison--x
31 Jul 2026
Agents

RadHarmony: Radiological Data Handling in the Era of Agentic AI

DGX agent

arXiv:2607.27235v1 Announce Type: cross Abstract: Training deep learning models on radiological images requires integrating heterogeneous datasets across different sources, file formats, directory lay

agentsarxiv-cs-cv
31 Jul 2026
Safety

TAPO: Transition-Aware Policy Optimization for LLM Agents

DGX agent

arXiv:2607.27973v1 Announce Type: new Abstract: Recently, Reinforcement Learning (RL) has emerged as a crucial paradigm for the post-training of Large Language Model (LLM) agents. However, existing me

safetyarxiv-cs-lg
31 Jul 2026
Model Releases

The new DeepSeek v4 Flash is now available in Hermes Agent through Nous Portal and OpenRouter

DGX agent

The new DeepSeek v4 Flash is now available in Hermes Agent through Nous Portal and OpenRouter 🚀 DeepSeek-V4-Flash Official API is now LIVE in public beta! 🔷 We’ve massively upgraded its Agent capabili

model-releasesnous-research--x
31 Jul 2026
Model Releases

“AI Meltdown” is the scenario where the agent goes off the rails and forgets/ignores all previous instructions and soft guardrails. No promp…

DGX agent

“AI Meltdown” is the scenario where the agent goes off the rails and forgets/ignores all previous instructions and soft guardrails. No prompt injection or malicious actor is needed for this to happen.

model-releasesperplexity--x
30 Jul 2026
Safety

Agentic AI in medicine: architectures, applications, evaluation, and challenges for clinical translation

DGX agent

arXiv:2607.25489v1 Announce Type: new Abstract: Large language models and multimodal foundation models are enabling medical artificial intelligence (AI) systems to move beyond isolated prediction and

safetyarxiv-cs-cv
29 Jul 2026
Model Releases

Cooperative Multi-UAV Navigation in Complex Environments via Systematic Multi-Agent Deep Reinforcement Learning

DGX agent

arXiv:2607.25754v1 Announce Type: new Abstract: Cooperative navigation of multi-agent UAVs in complex environments faces key challenges including local optima traps, sparse rewards, learning imbalance

model-releasesarxiv-cs-ro
29 Jul 2026
Model Releases

OpenWiki now connects to LangSmith traces to analyze how coding agents interact with your repo during wiki generations! We added a LangSmith…

DGX agent

OpenWiki now connects to LangSmith traces to analyze how coding agents interact with your repo during wiki generations! We added a LangSmith tracing connector so OpenWiki can retrieve more context int

model-releasesharrison-chase--x
29 Jul 2026
Model Releases

PATHFinder Agent for Tailored Prenatal Care

DGX agent

arXiv:2607.24768v1 Announce Type: new Abstract: Prenatal care is an important preventive service designed to improve outcomes for pregnant individuals. The American College of Obstetricians and Gyneco

model-releasesarxiv-cs-ai
29 Jul 2026
Local Ai

ProcAgent: An Agentic Framework for Procedural Task Guidance on Edge with Human-in-the-Loop

DGX agent

arXiv:2607.24770v1 Announce Type: new Abstract: Procedural tasks such as furniture assembly and home repair impose substantial cognitive demands because users must interpret instructions, track task p

local-aiarxiv-cs-ai
29 Jul 2026
Model Releases

RSMeM: Knowledge-Enhanced Memory Evolution for Remote Sensing Agents with Systematic Evaluation

DGX agent

arXiv:2607.24772v1 Announce Type: new Abstract: Geoscience research requires complex analysis and domain expertise, with remote sensing (RS) observations as a key foundation. However, existing RS agen

model-releasesarxiv-cs-ai
29 Jul 2026
Agents

SnapLogic transforms SnapGPT into a high-powered agentic assistant for the entire integration lifecycle

DGX agent

Enterprise data integration and automation firm SnapLogic Inc. today announced a significant update to SnapGPT, the company’s artificial intelligence copilot for enterprise data automation, transformi

agentssiliconangle
29 Jul 2026
Model Releases

Evaluating Fuzz Testing for Reinforcement Learning Agents

DGX agent

arXiv:2607.24577v1 Announce Type: new Abstract: Reinforcement Learning (RL) agents are increasingly deployed in safety-critical domains such as robotics, autonomous driving, and drone control, where u

model-releasesarxiv-cs-lg
28 Jul 2026
Model Releases

Gubernaut: A Deterministic Homeostatic Controller for Affect-Regulated LLM Agents, Validated Across Independent Model Families

DGX agent

arXiv:2607.24339v1 Announce Type: new Abstract: Large language model (LLM) agents inherit reactive failure modes: escalation under provocation, sycophantic drift under flattery, perseveration when stu

model-releasesarxiv-cs-ai
28 Jul 2026
Agents

HELIOS: An LLM-Driven Autonomous Indirect Trajectory Optimization Agent

DGX agent

arXiv:2607.24051v1 Announce Type: cross Abstract: Low-thrust trajectory optimization is a core technology in deep-space mission design. Indirect methods based on Pontryagin's Minimum Principle (PMP) o

agentsarxiv-cs-ai
28 Jul 2026
Model Releases

LazyMem: Retrieve Broadly, Construct Selectively for Efficient Long-Term Agent Memory

DGX agent

arXiv:2607.22690v1 Announce Type: new Abstract: Long-term memory lets LLM agents reuse past interactions, but raw dialogue histories are verbose and information-sparse. Retrieving broadly improves evi

model-releasesarxiv-cs-ai
28 Jul 2026
Safety

Looping Is Not Reliability: State-Bound Evidence and Typed Revision Contracts for Agentic Code Repair

DGX agent

arXiv:2607.24604v1 Announce Type: cross Abstract: Generate--test--revise loops are common in coding agents, but repetition alone provides no reliability guarantee. We study the gap between finding a c

safetyarxiv-cs-ai
28 Jul 2026
Safety

MedDDC-Eval: Diagnosis-Decoupled Evaluation of Multi-Turn Medical Consultation Agents

DGX agent

arXiv:2607.18999v2 Announce Type: replace-cross Abstract: Evaluating multi-turn medical consultation agents requires judging the diagnostic support provided by the histories they elicit through intera

safetyarxiv-cs-ai
28 Jul 2026
Model Releases

QFoldAgent: An Autonomous Quantum Optimization Multi-Agent System for Protein Structure Prediction

DGX agent

arXiv:2607.22549v1 Announce Type: new Abstract: Hybrid quantum-classical protein structure prediction depends strongly on Hamiltonian penalty weights, yet existing lattice-based workflows typically fi

model-releasesarxiv-cs-ai
28 Jul 2026
Safety

The Agentic AI Foundation updates MCP with a fully stateless architecture, a hardened authentication model, a formal 12-month deprecation policy, and more (Michael Nuñez/VentureBeat)

DGX agent

Michael Nuñez / VentureBeat: The Agentic AI Foundation updates MCP with a fully stateless architecture, a hardened authentication model, a formal 12-month deprecation policy, and more — The Model Cont

safetytechmeme
28 Jul 2026
Agents

Agentic Root Cause Analysis through Evidence-Grounded Reasoning

DGX agent

arXiv:2607.22385v1 Announce Type: cross Abstract: Diagnosing the root cause of anomalies is essential for safe industrial operation. Despite extensive sensor instrumentation, formulating hypotheses an

agentsarxiv-cs-lg
27 Jul 2026
Model Releases

Great technical paper from Google. Great read on why context beats scale for agents working against unfamiliar APIs. (bookmark it) GPU kerne…

DGX agent

Great technical paper from Google. Great read on why context beats scale for agents working against unfamiliar APIs. (bookmark it) GPU kernel optimization has KernelBench to hillclimb on. TPUs had not

model-releasesdair-ai--x
27 Jul 2026
Agents

Looking forward to chatting with @hugobowne today (July 27) at 4 pm PT on the Vanishing Gradient livestream on YouTube. Will cover open sour…

DGX agent

Looking forward to chatting with @hugobowne today (July 27) at 4 pm PT on the Vanishing Gradient livestream on YouTube. Will cover open source, the newest LLMs & trends, agent frameworks, and whatever

agentssebastian-raschka--x
27 Jul 2026
Safety

Molt: A Scalable PyTorch-Native Training Framework for Agentic Reinforcement Learning

DGX agent

arXiv:2607.21653v1 Announce Type: cross Abstract: Agentic reinforcement learning research is constant algorithm modification, new estimators, new pipeline stages, new rollout schemes, and in mainstrea

safetyarxiv-cs-cl
27 Jul 2026
Local Ai

My Ollama box picks the music now: an agentic DJ running on a 9B model

DGX agent

I got tired of my Ollama server sitting idle between chat experiments, so I pointed it at my Navidrome library and made it run a radio station. The DJ is an agent, not a shuffler. Each turn it gets to

local-air-localllama
27 Jul 2026
Agents

the paper:

DGX agent

the paper: babyagi has ~200 citations, but 0 papers... i just published my first paper on arXiv 😆 'The Log is the Agent: Event-Sourced Reactive Graphs for Auditable, Forkable Agentic Systems' https://

agentsyohei-nakajima--x
27 Jul 2026
Agents

Agent-Guided Relational Concept Discovery: Toward Interpretable Surgical Margin Assessment

DGX agent

arXiv:2607.21437v1 Announce Type: new Abstract: Deep learning models can effectively use Rapid Evaporative Ionization Mass Spectrometry (REIMS) data for surgical margin assessment. However, their clin

agentsarxiv-cs-ai
24 Jul 2026
Safety

ARCO: Adaptive Rubrics with Co-Evolution for Multi-Step LLM-Based Agents

DGX agent

arXiv:2606.21262v2 Announce Type: replace Abstract: Reinforcement learning for multi-step LLM agents often relies on scalar rewards that indicate success but cannot explain why a trajectory is good or

safetyarxiv-cs-ai
24 Jul 2026
Model Releases

AREX: Towards a Recursively Self-Improving Agent for Deep Research

DGX agent

arXiv:2607.21461v1 Announce Type: new Abstract: Deep research requires agents to find answers that jointly satisfy multiple constraints. Discovering such answers is costly, whereas verifying a candida

model-releasesarxiv-cs-ai
24 Jul 2026
Agents

Bayesian uncertainty estimation improves clinical decision making in medical AI agents

DGX agent

arXiv:2607.20582v1 Announce Type: cross Abstract: Machine learning models for medical image analysis typically lack a reliable measure of confidence, limiting their use in ambiguous or atypical cases.

agentsarxiv-cs-ai
24 Jul 2026
Hardware

Congrats to @lmstudio on Bionic! Put open models to work with a local-first agent for docs, coding, voice, and more. Run it today on NVIDIA …

DGX agent

Congrats to @lmstudio on Bionic! Put open models to work with a local-first agent for docs, coding, voice, and more. Run it today on NVIDIA RTX GPUs. 🚀 Meet LM Studio Bionic. The Agent made for Open M

hardwarelm-studio--x
24 Jul 2026
Agents

OPTScientist: Multi-Agent Discovery of Typed Optimizer Programs for Transformer Pretraining

DGX agent

arXiv:2607.20486v1 Announce Type: new Abstract: Designing optimizers for modern deep learning remains a challenging scientific problem, requiring the joint consideration of optimization geometry, stat

agentsarxiv-cs-ai
24 Jul 2026
Safety

The Dark Room in the Reward Channel: Dense Prediction Rewards Collapse GRPO-Trained LLM Agents -- and What Actually Works

DGX agent

arXiv:2607.21273v1 Announce Type: new Abstract: Dense per-step supervision is an appealing remedy for sparse-reward, long-horizon LLM agents: reward the agent for predicting its next observation, and

safetyarxiv-cs-lg
24 Jul 2026
Model Releases

Workflow-Localized Mechanism Learning: Attribution-Guided Repair and Knowledge Reuse for Structured Agent Skills

DGX agent

arXiv:2607.20999v1 Announce Type: new Abstract: Agent Skills package reusable procedural knowledge as external artifacts for frozen language-model agents, yet existing optimizers do not jointly resolv

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

AgentCgroup: Understanding and Controlling OS Resources of AI Agents

DGX agent

arXiv:2602.09345v3 Announce Type: replace-cross Abstract: AI agents are increasingly deployed in multi-tenant cloud environments, where they execute diverse tool calls within sandboxed containers, eac

model-releasesarxiv-cs-ai
23 Jul 2026
Agents

An LLM-powered Agentic Recommendation System for Connected TV Content Discovery

DGX agent

arXiv:2607.09988v3 Announce Type: replace-cross Abstract: Recommendation systems, from traditional multi-stage to recent unified generative architectures, face challenges in incorporating diverse cont

agentsarxiv-cs-ai
23 Jul 2026
Model Releases

CEO-Bench: Can Agents Play the Long Game?

DGX agent

arXiv:2606.18543v2 Announce Type: replace Abstract: Language model agents are becoming proficient executors at isolated, short-horizon tasks such as software engineering and customer service. Yet real

model-releasesarxiv-cs-ai
23 Jul 2026
Safety

Coordinating from Memory: Graph-Structured Experience Reuse for Multi-Agent Adaptation in Dynamic Manufacturing

DGX agent

arXiv:2607.19985v1 Announce Type: new Abstract: Dynamic manufacturing environments require multi-agent systems to coordinate effectively under frequent operational disturbances such as machine failure

safetyarxiv-cs-ai
23 Jul 2026
Safety

Guardrails as Scapegoats: Auditing Unfaithful Safety Refusals in Tool-Augmented LLM Agents

DGX agent

arXiv:2607.19449v1 Announce Type: cross Abstract: Evaluation frameworks for tool-augmented LLM agents focus overwhelmingly on capability metrics or explicit tool crashes, leaving silent infrastructure

safetyarxiv-cs-ai
23 Jul 2026
Model Releases

The Chronos Vulnerability: A Taxonomy of Temporal Persistence and Memory-Based Deception in Agentic AI

DGX agent

arXiv:2607.19433v1 Announce Type: new Abstract: The transition from stateless generative models in artificial intelligence to stateful, autonomous agents represents an architectural evolution that, wh

model-releasesarxiv-cs-ai
23 Jul 2026
Model Releases

Kimi K3 is second only to Fable 5 on AA-Briefcase, our agentic knowledge work benchmark, but costs more than Opus 4.8 to run while averaging…

DGX agent

Kimi K3 is second only to Fable 5 on AA-Briefcase, our agentic knowledge work benchmark, but costs more than Opus 4.8 to run while averaging nearly an hour per task Last week @Kimi_Moonshot released K

model-releaseskimi-moonshot--x
21 Jul 2026
← Previous
1…114115116117118…375
Next →