AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,951 results
3 Aug 2026

MMShopBench: A Real-Log Benchmark for Multimodal, Multi-Turn Shopping Agents

Model ReleasesDGX agent

arXiv:2607.29002v1 Announce Type: new Abstract: Online shoppers increasingly turn to AI shopping assistants, using images and multi-turn dialogue to express and refine product needs that are difficult

Real-world mainframe modernization with AI: A safe, scalable path from mainframe to cloud

Model ReleasesDGX agent

For too long, enterprises with legacy mainframe estates have been faced with a high-stakes dilemma: continue maintaining their mainframes, essentially kicking the modernization can down the road (they

SeekBrain: An Autonomous Multi-Agent System for Accelerating Neuroscience Discovery

Model ReleasesDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2607.29347v1 Announce Type: cross Abstract: Modern neuroscience relies on integrating multi-scale, multimodal datasets to uncover the neural principles underlying intelligence. However, analytic

Zero-Mem: Zero-Token Memory Operations for LLM Agents

Local AiDGX agent

arXiv:2607.29377v1 Announce Type: new Abstract: LLM agents need memory to act consistently over long interactions, yet many systems use additional LLM calls to operate that memory. Generating intermed

2 Aug 2026

Real-world reality check on Qwen for autonomous coding agents

Model ReleasesDGX agent

TLDR below 👇🏼 I’ve seen a lot of hype around Qwen 3.6 35B and 3.5 120B lately, especially regarding coding and tool-use capabilities. On this subreddit it is the defacto recommended model for everyone

31 Jul 2026

DeepSeek V4 Flash 0731 in Hermes Agent and one prompt, took 32 minutes and cost 0.07$, this model is so cheap to the point where 2 dollars c…

Model ReleasesDGX agent

**DeepSeek V4 Flash 0731 Performance Test** On July 31 2026, a single prompt executed via the Hermes Agent on DeepSeek V4 Flash 0731 completed in 32 minutes and incurred an estimated cost of 0.07 USD.

Harness-G: A Graph-Structured Harness for Search Agents

SafetyDGX agent

arXiv:2607.27652v1 Announce Type: new Abstract: Reinforcement learning (RL) search agents commonly model retrieval as free-form natural-language query generation and optimize multi-turn interactions u

OK, GPT-5.6 Luna is a bit of a beast. Given the 80% price drop today I decided to try it in Datasette Agent, and it's furiously quick and ge…

Model ReleasesDGX agent

OK, GPT-5.6 Luna is a bit of a beast. Given the 80% price drop today I decided to try it in Datasette Agent, and it's furiously quick and generates all the SQL, HTML and JavaScript (for Datasette Apps

RadHarmony: Radiological Data Handling in the Era of Agentic AI

AgentsDGX agent

arXiv:2607.27235v1 Announce Type: cross Abstract: Training deep learning models on radiological images requires integrating heterogeneous datasets across different sources, file formats, directory lay

TAPO: Transition-Aware Policy Optimization for LLM Agents

SafetyDGX agent

arXiv:2607.27973v1 Announce Type: new Abstract: Recently, Reinforcement Learning (RL) has emerged as a crucial paradigm for the post-training of Large Language Model (LLM) agents. However, existing me

The new DeepSeek v4 Flash is now available in Hermes Agent through Nous Portal and OpenRouter

Model ReleasesDGX agent

The new DeepSeek v4 Flash is now available in Hermes Agent through Nous Portal and OpenRouter 🚀 DeepSeek-V4-Flash Official API is now LIVE in public beta! 🔷 We’ve massively upgraded its Agent capabili

30 Jul 2026

“AI Meltdown” is the scenario where the agent goes off the rails and forgets/ignores all previous instructions and soft guardrails. No promp…

Model ReleasesDGX agent

“AI Meltdown” is the scenario where the agent goes off the rails and forgets/ignores all previous instructions and soft guardrails. No prompt injection or malicious actor is needed for this to happen.

29 Jul 2026

Agentic AI in medicine: architectures, applications, evaluation, and challenges for clinical translation

SafetyDGX agent

arXiv:2607.25489v1 Announce Type: new Abstract: Large language models and multimodal foundation models are enabling medical artificial intelligence (AI) systems to move beyond isolated prediction and

Cooperative Multi-UAV Navigation in Complex Environments via Systematic Multi-Agent Deep Reinforcement Learning

Model ReleasesDGX agent

arXiv:2607.25754v1 Announce Type: new Abstract: Cooperative navigation of multi-agent UAVs in complex environments faces key challenges including local optima traps, sparse rewards, learning imbalance

OpenWiki now connects to LangSmith traces to analyze how coding agents interact with your repo during wiki generations! We added a LangSmith…

Model ReleasesDGX agent

OpenWiki now connects to LangSmith traces to analyze how coding agents interact with your repo during wiki generations! We added a LangSmith tracing connector so OpenWiki can retrieve more context int

PATHFinder Agent for Tailored Prenatal Care

Model ReleasesDGX agent

arXiv:2607.24768v1 Announce Type: new Abstract: Prenatal care is an important preventive service designed to improve outcomes for pregnant individuals. The American College of Obstetricians and Gyneco

ProcAgent: An Agentic Framework for Procedural Task Guidance on Edge with Human-in-the-Loop

Local AiDGX agent

arXiv:2607.24770v1 Announce Type: new Abstract: Procedural tasks such as furniture assembly and home repair impose substantial cognitive demands because users must interpret instructions, track task p

RSMeM: Knowledge-Enhanced Memory Evolution for Remote Sensing Agents with Systematic Evaluation

Model ReleasesDGX agent

arXiv:2607.24772v1 Announce Type: new Abstract: Geoscience research requires complex analysis and domain expertise, with remote sensing (RS) observations as a key foundation. However, existing RS agen

SnapLogic transforms SnapGPT into a high-powered agentic assistant for the entire integration lifecycle

AgentsDGX agent

Enterprise data integration and automation firm SnapLogic Inc. today announced a significant update to SnapGPT, the company’s artificial intelligence copilot for enterprise data automation, transformi

28 Jul 2026

Evaluating Fuzz Testing for Reinforcement Learning Agents

Model ReleasesDGX agent

arXiv:2607.24577v1 Announce Type: new Abstract: Reinforcement Learning (RL) agents are increasingly deployed in safety-critical domains such as robotics, autonomous driving, and drone control, where u

Gubernaut: A Deterministic Homeostatic Controller for Affect-Regulated LLM Agents, Validated Across Independent Model Families

Model ReleasesDGX agent

arXiv:2607.24339v1 Announce Type: new Abstract: Large language model (LLM) agents inherit reactive failure modes: escalation under provocation, sycophantic drift under flattery, perseveration when stu

HELIOS: An LLM-Driven Autonomous Indirect Trajectory Optimization Agent

AgentsDGX agent

arXiv:2607.24051v1 Announce Type: cross Abstract: Low-thrust trajectory optimization is a core technology in deep-space mission design. Indirect methods based on Pontryagin's Minimum Principle (PMP) o

LazyMem: Retrieve Broadly, Construct Selectively for Efficient Long-Term Agent Memory

Model ReleasesDGX agent

arXiv:2607.22690v1 Announce Type: new Abstract: Long-term memory lets LLM agents reuse past interactions, but raw dialogue histories are verbose and information-sparse. Retrieving broadly improves evi

Looping Is Not Reliability: State-Bound Evidence and Typed Revision Contracts for Agentic Code Repair

SafetyDGX agent

arXiv:2607.24604v1 Announce Type: cross Abstract: Generate--test--revise loops are common in coding agents, but repetition alone provides no reliability guarantee. We study the gap between finding a c

MedDDC-Eval: Diagnosis-Decoupled Evaluation of Multi-Turn Medical Consultation Agents

SafetyDGX agent

arXiv:2607.18999v2 Announce Type: replace-cross Abstract: Evaluating multi-turn medical consultation agents requires judging the diagnostic support provided by the histories they elicit through intera

QFoldAgent: An Autonomous Quantum Optimization Multi-Agent System for Protein Structure Prediction

Model ReleasesDGX agent

arXiv:2607.22549v1 Announce Type: new Abstract: Hybrid quantum-classical protein structure prediction depends strongly on Hamiltonian penalty weights, yet existing lattice-based workflows typically fi

The Agentic AI Foundation updates MCP with a fully stateless architecture, a hardened authentication model, a formal 12-month deprecation policy, and more (Michael Nuñez/VentureBeat)

SafetyDGX agent

Michael Nuñez / VentureBeat: The Agentic AI Foundation updates MCP with a fully stateless architecture, a hardened authentication model, a formal 12-month deprecation policy, and more — The Model Cont

27 Jul 2026

Agentic Root Cause Analysis through Evidence-Grounded Reasoning

AgentsDGX agent

arXiv:2607.22385v1 Announce Type: cross Abstract: Diagnosing the root cause of anomalies is essential for safe industrial operation. Despite extensive sensor instrumentation, formulating hypotheses an

Great technical paper from Google. Great read on why context beats scale for agents working against unfamiliar APIs. (bookmark it) GPU kerne…

Model ReleasesDGX agent

Great technical paper from Google. Great read on why context beats scale for agents working against unfamiliar APIs. (bookmark it) GPU kernel optimization has KernelBench to hillclimb on. TPUs had not

Looking forward to chatting with @hugobowne today (July 27) at 4 pm PT on the Vanishing Gradient livestream on YouTube. Will cover open sour…

AgentsDGX agent

Looking forward to chatting with @hugobowne today (July 27) at 4 pm PT on the Vanishing Gradient livestream on YouTube. Will cover open source, the newest LLMs & trends, agent frameworks, and whatever

Molt: A Scalable PyTorch-Native Training Framework for Agentic Reinforcement Learning

SafetyDGX agent

arXiv:2607.21653v1 Announce Type: cross Abstract: Agentic reinforcement learning research is constant algorithm modification, new estimators, new pipeline stages, new rollout schemes, and in mainstrea

My Ollama box picks the music now: an agentic DJ running on a 9B model

Local AiDGX agent

I got tired of my Ollama server sitting idle between chat experiments, so I pointed it at my Navidrome library and made it run a radio station. The DJ is an agent, not a shuffler. Each turn it gets to

the paper:

AgentsDGX agent

the paper: babyagi has ~200 citations, but 0 papers... i just published my first paper on arXiv 😆 'The Log is the Agent: Event-Sourced Reactive Graphs for Auditable, Forkable Agentic Systems' https://

24 Jul 2026

Agent-Guided Relational Concept Discovery: Toward Interpretable Surgical Margin Assessment

AgentsDGX agent

arXiv:2607.21437v1 Announce Type: new Abstract: Deep learning models can effectively use Rapid Evaporative Ionization Mass Spectrometry (REIMS) data for surgical margin assessment. However, their clin

ARCO: Adaptive Rubrics with Co-Evolution for Multi-Step LLM-Based Agents

SafetyDGX agent

arXiv:2606.21262v2 Announce Type: replace Abstract: Reinforcement learning for multi-step LLM agents often relies on scalar rewards that indicate success but cannot explain why a trajectory is good or

AREX: Towards a Recursively Self-Improving Agent for Deep Research

Model ReleasesDGX agent

arXiv:2607.21461v1 Announce Type: new Abstract: Deep research requires agents to find answers that jointly satisfy multiple constraints. Discovering such answers is costly, whereas verifying a candida

Bayesian uncertainty estimation improves clinical decision making in medical AI agents

AgentsDGX agent

arXiv:2607.20582v1 Announce Type: cross Abstract: Machine learning models for medical image analysis typically lack a reliable measure of confidence, limiting their use in ambiguous or atypical cases.

Congrats to @lmstudio on Bionic! Put open models to work with a local-first agent for docs, coding, voice, and more. Run it today on NVIDIA …

HardwareDGX agent

Congrats to @lmstudio on Bionic! Put open models to work with a local-first agent for docs, coding, voice, and more. Run it today on NVIDIA RTX GPUs. 🚀 Meet LM Studio Bionic. The Agent made for Open M

OPTScientist: Multi-Agent Discovery of Typed Optimizer Programs for Transformer Pretraining

AgentsDGX agent

arXiv:2607.20486v1 Announce Type: new Abstract: Designing optimizers for modern deep learning remains a challenging scientific problem, requiring the joint consideration of optimization geometry, stat

The Dark Room in the Reward Channel: Dense Prediction Rewards Collapse GRPO-Trained LLM Agents -- and What Actually Works

SafetyDGX agent

arXiv:2607.21273v1 Announce Type: new Abstract: Dense per-step supervision is an appealing remedy for sparse-reward, long-horizon LLM agents: reward the agent for predicting its next observation, and

Workflow-Localized Mechanism Learning: Attribution-Guided Repair and Knowledge Reuse for Structured Agent Skills

Model ReleasesDGX agent

arXiv:2607.20999v1 Announce Type: new Abstract: Agent Skills package reusable procedural knowledge as external artifacts for frozen language-model agents, yet existing optimizers do not jointly resolv

23 Jul 2026

AgentCgroup: Understanding and Controlling OS Resources of AI Agents

Model ReleasesDGX agent

arXiv:2602.09345v3 Announce Type: replace-cross Abstract: AI agents are increasingly deployed in multi-tenant cloud environments, where they execute diverse tool calls within sandboxed containers, eac

An LLM-powered Agentic Recommendation System for Connected TV Content Discovery

AgentsDGX agent

arXiv:2607.09988v3 Announce Type: replace-cross Abstract: Recommendation systems, from traditional multi-stage to recent unified generative architectures, face challenges in incorporating diverse cont

CEO-Bench: Can Agents Play the Long Game?

Model ReleasesDGX agent

arXiv:2606.18543v2 Announce Type: replace Abstract: Language model agents are becoming proficient executors at isolated, short-horizon tasks such as software engineering and customer service. Yet real

Coordinating from Memory: Graph-Structured Experience Reuse for Multi-Agent Adaptation in Dynamic Manufacturing

SafetyDGX agent

arXiv:2607.19985v1 Announce Type: new Abstract: Dynamic manufacturing environments require multi-agent systems to coordinate effectively under frequent operational disturbances such as machine failure

Guardrails as Scapegoats: Auditing Unfaithful Safety Refusals in Tool-Augmented LLM Agents

SafetyDGX agent

arXiv:2607.19449v1 Announce Type: cross Abstract: Evaluation frameworks for tool-augmented LLM agents focus overwhelmingly on capability metrics or explicit tool crashes, leaving silent infrastructure

The Chronos Vulnerability: A Taxonomy of Temporal Persistence and Memory-Based Deception in Agentic AI

Model ReleasesDGX agent

arXiv:2607.19433v1 Announce Type: new Abstract: The transition from stateless generative models in artificial intelligence to stateful, autonomous agents represents an architectural evolution that, wh

21 Jul 2026

Kimi K3 is second only to Fable 5 on AA-Briefcase, our agentic knowledge work benchmark, but costs more than Opus 4.8 to run while averaging…

Model ReleasesDGX agent

Kimi K3 is second only to Fable 5 on AA-Briefcase, our agentic knowledge work benchmark, but costs more than Opus 4.8 to run while averaging nearly an hour per task Last week @Kimi_Moonshot released K

This is what strong inference economics unlock. @rox_ai built its own search agent, ran it in production for 6+ months, and reached 91.3% ac…

Model ReleasesDGX agent

This is what strong inference economics unlock. @rox_ai built its own search agent, ran it in production for 6+ months, and reached 91.3% accuracy at 1.03¢ per query. Together AI is proud to help powe

16 Jul 2026

COLMAR: Cooperative View Policy Learning for Multi-Agent Active 3D Reconstruction

Model ReleasesDGX agent

arXiv:2607.13524v1 Announce Type: new Abstract: Active 3D reconstruction requires selecting informative viewpoints under limited sensing budgets. In multi-agent settings, coordination inefficiencies s

Exploratory, Communicative, and Deployable: Vision-Driven Embodied Agents for Open-World Mobile Manipulation

Model ReleasesDGX agent

arXiv:2607.13653v1 Announce Type: new Abstract: Real-world deployment of embodied agents requires active exploration, visual grounding, and interactive intent disambiguation. However, existing framewo

Inference Economics of Enterprise Coding Agents: A Case Study of Cloud vs. On-Premise LLMs

Model ReleasesDGX agent

arXiv:2607.13080v1 Announce Type: cross Abstract: Autonomous coding agents force engineering organizations to choose between API-based frontier models -- strong reasoning at high token cost -- and on-

Learning Engagement Assistant (LEA): Cross-Course Scalability and Classroom Evaluation of an Agentic AI Tutoring System

AgentsDGX agent

arXiv:2607.13370v1 Announce Type: cross Abstract: This paper is an extension of a paper presented at the ICAART 2026 conference, which introduced LEA (Learning Engagement Assistant), an adaptive AI tu

Memory as a Controlled Process: Learned Adaptive Memory Management for LLM Agents

SafetyDGX agent

arXiv:2607.13591v1 Announce Type: cross Abstract: Large Language Model (LLM) agents increasingly rely on external memory systems to accumulate experience across tasks. Yet nearly all existing approach

TRACE: Turn-level Reward Assignment via Credit Estimation for Long-Horizon Agents

Model ReleasesDGX agent

arXiv:2607.13988v1 Announce Type: new Abstract: Multi-turn agents solve complex tasks through extended sequences of tool interactions before producing a final answer, making credit assignment a fundam

15 Jul 2026

Agentic vision: Building visual intelligence with Amazon Bedrock and MCP servers

AgentsDGX agent

In this post, we walk you through the Computer Vision MCP Server, which illustrates this approach, representing how AI systems can process visual information and make intelligent decisions through a s

Cadence extends its AI agents beyond chips with AuraStack for circuit boards and packaging

HardwareDGX agent

Integrated circuit and electronic hardware design company Cadence Design Systems Inc. today announced a new artificial intelligence agent that assists with packaging and system design, the step after

The Blender MCP is now part of the Hermes Agent MCP Catalog! Easily activate and install the Blender MCP by running `hermes mcp install blen…

Model ReleasesDGX agent

The Blender MCP is now part of the Hermes Agent MCP Catalog! Easily activate and install the Blender MCP by running `hermes mcp install blender` and ask your agent to start using blender. Hermes will

14 Jul 2026

continued working on @activegraphai reference packs last weekend, resulted in needing to harden the runtime: it already could... - keep a co…

AgentsDGX agent

continued working on @activegraphai reference packs last weekend, resulted in needing to harden the runtime: it already could... - keep a complete history of everything an agent did - replay that hist

10 Jul 2026

Game Theory Driven Multi-Agent Framework Mitigates Language Model Hallucination

AgentsDGX agent

arXiv:2607.08403v1 Announce Type: new Abstract: The application of lightweight Large Language Models in rule-based scientific domains remains severely limited by their tendency to mimic linguistic pat

← Previous
1…9192939495…300
Next →