AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlog
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “tools”

GridTimelineEvolution
10,103 results
Agents

Living-Harness Is an Interactive-Agent Evolver

DGX agent

arXiv:2607.26598v1 Announce Type: cross Abstract: Large language model (LLM) agents may recover from a failure within an episode or after a retry, yet the same execution failure can recur in later tas

agentsarxiv-cs-cl
30 Jul 2026
Model Releases

llm 0.32rc2

DGX agent

Release: llm 0.32rc2 Hot on the heels of RC1, this fixes a dependency issue and also adds two neat new features: The default model for users who have not set their own default is now GPT-5.6 Luna. It

X Post
Paper
YouTube
Reddit
GitHub
Clear filters
model-releasessimon-willison
30 Jul 2026
Model Releases

Nanbeige4.2-3B: I'm not impressed

DGX agent

I've tested Nanbeige-4.2-3B. On paper, the benchmarks promise it blows away Qwen3.5-9B and Gemma4-12B. My goal was to have something very light and fast to replace Qwen3.6-35B (or finetunes thereof) f

model-releasesr-localllama
30 Jul 2026
Model Releases

Reeling It In: Flexible Needle Pick Up via Thread Manipulation for Autonomous Suturing

DGX agent

arXiv:2607.26337v1 Announce Type: new Abstract: Suture-needle pickup is necessary for autonomous suturing, as a needle can be unexpectedly dropped or strategically released to adjust the grasping conf

model-releasesarxiv-cs-ro
30 Jul 2026
Model Releases

This is absolutely wild... Anthropic reviewed their logs and found out that their own supposedly-sandboxed cyber evals had hacked three sepa…

DGX agent

This is absolutely wild... Anthropic reviewed their logs and found out that their own supposedly-sandboxed cyber evals had hacked three separate companies back in April without them noticing! In a rev

model-releasessimon-willison--x
30 Jul 2026
Hardware

What to expect during Black Hat USA: Join theCUBE Aug. 5-6

DGX agent

Artificial intelligence has propelled the cybersecurity world into a new phase, driven by startlingly advanced autonomous attacks and an urgent need to adopt technology to defend against them. This we

hardwaresiliconangle
30 Jul 2026
Model Releases

After a few more hours, I think I've figured out Opus 5. Opus 5 is trained to be more agentic than anything I've used. All Claude 5 models a…

DGX agent

After a few more hours, I think I've figured out Opus 5. Opus 5 is trained to be more agentic than anything I've used. All Claude 5 models are like that. So what changes? The way to interact with Opus

model-releasesdair-ai--x
29 Jul 2026
Model Releases

Agentic AI for Scientific Reasoning in Autonomous Quantum Sensing Experiments

DGX agent

arXiv:2607.25145v1 Announce Type: cross Abstract: We implement an agentic AI workflow built around a large language model (LLM) agent for autonomous experiments with nitrogen-vacancy (NV) centers in d

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

AnnoBench: A Benchmark for Visualization Annotation Generation

DGX agent

arXiv:2607.25911v1 Announce Type: cross Abstract: Annotation is among the most demanding visualization tasks to automate, as it simultaneously requires correctly navigating visual, semantic, and styli

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

Authoring Agent Skills: A Software-Engineering Approach

DGX agent

arXiv:2607.25032v1 Announce Type: cross Abstract: Agent Skills are an emerging way to extend large language model agents with reusable procedural knowledge that the agent loads on demand. Anthropic in

model-releasesarxiv-cs-ai
29 Jul 2026
Research

Automated Numerical Stability Analysis of Deep Learning Operators

DGX agent

arXiv:2607.25494v1 Announce Type: cross Abstract: Finite-precision arithmetic unavoidably introduces numerical approximation errors. Numerical computations may use insufficient precision or an imprope

researcharxiv-cs-ai
29 Jul 2026
Research

dtControl2+arepsilon: Trading Optimality for Explainability in MDPs via Decision Trees

DGX agent

arXiv:2607.25925v1 Announce Type: new Abstract: Over the past decade, decision trees have been used to represent controllers (a.k.a. policies) in an explainable way, with dtControl2 as a current state

researcharxiv-cs-ai
29 Jul 2026
Research

Exploring Line Bundle Standard Models with Transformers

DGX agent

arXiv:2607.00078v2 Announce Type: cross Abstract: We propose a Transformer-based Reinforcement Learning architecture, 'LB-Explorer', to search for heterotic line bundle standard models arising from co

researcharxiv-cs-lg
29 Jul 2026
Research

FIDAC: An Easy-to-use Pipeline to Extract and Interpret Interpersonal Distance From Video

DGX agent

arXiv:2607.25146v1 Announce Type: new Abstract: The distance between persons reveals significant information about their perception of each other. However, such information is not easily extractable a

researcharxiv-cs-cv
29 Jul 2026
Local Ai

i'll say it plainly, hermes agent desktop is the best agentic app i've used, and i'm a little mad i didn't find it sooner. it auto see the m…

DGX agent

i'll say it plainly, hermes agent desktop is the best agentic app i've used, and i'm a little mad i didn't find it sooner. it auto see the models i'm serving, laguna s 2.1 sitting on my dgx spark and

local-ainous-research--x
29 Jul 2026
Research

Long-Term PM2.5 Forecasting Using a DTW-Enhanced CNN-GRU Model

DGX agent

arXiv:2510.22863v2 Announce Type: replace-cross Abstract: Reliable long-term forecasting of PM2.5 concentrations is critical for public health early-warning systems, yet existing deep learning approac

researcharxiv-cs-ai
29 Jul 2026
Hardware

Super interesting new work from NVIDIA. (bookmark it) They suggest building agents as Python objects. Very cool idea and I think it could a …

DGX agent

Super interesting new work from NVIDIA. (bookmark it) They suggest building agents as Python objects. Very cool idea and I think it could a lot with agent reliability. More below: Agent development to

hardwaredair-ai--x
29 Jul 2026
Model Releases

The borderless Lakehouse: Bring AWS, Databricks and Snowflake data to your AI agents

DGX agent

Today’s data lakehouse is no longer mere data repository, but increasingly a system of action, actively executing tasks via always-on, autonomous AI agents. Rather than waiting for static reports, the

model-releasesgoogle-cloud-ai
29 Jul 2026
Safety

VetClaw: An Edge-Cloud Multimodal Agentic System for Veterinary Disease Screening

DGX agent

arXiv:2607.26042v1 Announce Type: new Abstract: We present VetClaw, an edge-cloud multimodal agentic system for early veterinary disease screening. VetClaw uses a camera module as an edge sensing devi

safetyarxiv-cs-cv
29 Jul 2026
Model Releases

What’s new in Gemini Enterprise Agent Platform

DGX agent

Since we launched Gemini Enterprise Agent Platform a few months ago, we’ve seen inspiring progress from businesses and builders alike. To stir up development, we’ve also shared 13 demos that can walk

model-releasesgoogle-cloud-ai
29 Jul 2026
Agents

A corrective agentic hybrid RAG and an operations-grounded evaluation for a scientific facility

DGX agent

arXiv:2607.24663v1 Announce Type: cross Abstract: Scientific user facilities accumulate decades of operational knowledge that no single search index covers: electronic logbooks, technical documents, i

agentsarxiv-cs-ai
28 Jul 2026
Model Releases

Agent Team Work Zone: An Automated, Persistent Workspace for Long-Lived Coding Agent Teams

DGX agent

arXiv:2607.22917v1 Announce Type: new Abstract: Large Language Model (LLM) agents have significantly improved coding and programming workflows. Claude Code, in particular, is one of the most powerful

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

AgentOmnia: Scaling Agentic Models for Full-Scenario Applications

DGX agent

arXiv:2607.23124v1 Announce Type: new Abstract: Large language model agents have advanced rapidly, yet progress remains fragmented across domains, capabilities, task difficulty, and interaction settin

model-releasesarxiv-cs-ai
28 Jul 2026
Agents

Building AI That Works: ESnet's Pragmatic Approach to AI-Driven Operational Excellence

DGX agent

arXiv:2607.22948v1 Announce Type: cross Abstract: The ORBIT (Operations Responses and Business Intelligence Toolkit) project was initiated to assess agentic AI for the upcoming ESnet 7 initiative and

agentsarxiv-cs-ai
28 Jul 2026
Model Releases

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding

DGX agent

arXiv:2607.24743v1 Announce Type: cross Abstract: Multimodal large language models (MLLMs) hold immense potential to revolutionize clinical practice, yet deploying them in the medical domain is fundam

model-releasesarxiv-cs-ai
28 Jul 2026
Applications

Codex from 0 to 10M Users: Building ChatGPT Work — Akshay Nathan, OpenAI

DGX agent

OpenAI’s Codex, an AI‑powered coding engine, has surged beyond the developer community to reach over 10 million users—combining Codex with its new “ChatGPT Work” platform that turns the model into a g

applicationslatent-space
28 Jul 2026
Agents

CodexGraph: Bridging Large Language Models and Code Repositories via Code Graph Databases

DGX agent

arXiv:2408.03910v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) excel in stand-alone code tasks like HumanEval and MBPP, but struggle with handling entire code repositories. Thi

agentsarxiv-cs-ai
28 Jul 2026
Model Releases

Constraint-Bound Agnostic Bayesian Optimization: One Model for All Thresholds

DGX agent

arXiv:2607.23448v1 Announce Type: cross Abstract: Expensive constrained optimization problems in real-world industry design often involve constraint thresholds that are difficult to determine in advan

model-releasesarxiv-cs-ai
28 Jul 2026
Agents

Delegation Intelligence in Deep Search: A Controllable Framework for Disentangled Capability Diagnosis

DGX agent

arXiv:2607.23524v1 Announce Type: new Abstract: Deep search is becoming a core capability of modern agent systems, yet it is typically evaluated solely based on end-to-end answer accuracy. This couple

agentsarxiv-cs-ai
28 Jul 2026
Model Releases

Diagrid Catalyst 2.0 adds durable execution to more than 10 agent frameworks

DGX agent

Agent infrastructure startup Diagrid Inc. today released Catalyst 2.0, an update to its managed workflow engine that adds automatic failure recovery and cryptographic verification to artificial intell

model-releasessiliconangle
28 Jul 2026
Local Ai

DRC-Aid: Design-Rule Correction via Agentic Framework utilizing Inference-Time Large Language Models

DGX agent

arXiv:2607.22761v1 Announce Type: cross Abstract: Resolving Design Rule Violations (DRVs) in layouts entails an iterative loop of geometric edits and verification. We present DRC-Aid, a closed-loop ag

local-aiarxiv-cs-lg
28 Jul 2026
Model Releases

Embodied GPT-5.1: Evidence of a World Model?

DGX agent

arXiv:2607.23899v1 Announce Type: cross Abstract: This exploratory study examines whether a large multimodal language model, GPT-5.1, can serve as the high-level controller of a physical mobile robot

model-releasesarxiv-cs-ai
28 Jul 2026
Agents

Failures Reveal What Metrics Miss: An Evidence-Driven Agent for Recursive Refinement of ECG Classifiers

DGX agent

arXiv:2607.24419v1 Announce Type: new Abstract: Deep models have substantially advanced 12-lead ECG classification, yet their refinement still relies heavily on human experts to inspect failures and i

agentsarxiv-cs-ai
28 Jul 2026
Safety

From Proprietary to Open-Source: Bridging the Distribution Gap via Multi-Agent Protocol Distillation in Agentic Search

DGX agent

arXiv:2607.24280v1 Announce Type: new Abstract: Agentic search enables large language models to solve knowledge-intensive tasks by interleaving multi-step reasoning with retrieval, yet optimizing this

safetyarxiv-cs-ai
28 Jul 2026
Model Releases

It’s been five years since ChatGPT was released, AGI is not coming, not even the most modest AI predictions will ever become a reality, and …

DGX agent

It’s been five years since ChatGPT was released, AGI is not coming, not even the most modest AI predictions will ever become a reality, and there needs to be some of severe criminal sanction applied t

model-releasesgary-marcus--x
28 Jul 2026
Model Releases

KAYROS: An Anytime and Exact Open-Source Solver for Duration-Minimization Time-Dependent Vehicle Routing. A Technical Report and a Case Study in Human-AI Engineering

DGX agent

arXiv:2607.23116v1 Announce Type: cross Abstract: KAYROS is an open-source solver for duration-minimization time-dependent vehicle routing problems, with or without time windows (TDVRPTW, TDVRP). In t

model-releasesarxiv-cs-ai
28 Jul 2026
Safety

Neuro-Symbolic Meta-Policies for Temporal Knowledge-Graph Memory under Partial Observability

DGX agent

arXiv:2607.18368v2 Announce Type: replace Abstract: Partially observable reinforcement learning requires deciding what to retain, retrieve, and forget over time. We introduce a neuro-symbolic meta-pol

safetyarxiv-cs-ai
28 Jul 2026
Agents

PlanCraft: Sketch, Refine, and Furnish for Architect-Inspired Progressive 3D Residential Scene Generation

DGX agent

arXiv:2607.23491v1 Announce Type: cross Abstract: Two structural insights have been overlooked in automated residential floor plan generation. First, design is inherently progressive. Architects begin

agentsarxiv-cs-cl
28 Jul 2026
Model Releases

Spatial Reasoning in LLM Game Agents: Impact of Causal Context and Multi-Step Planning

DGX agent

arXiv:2607.22732v1 Announce Type: new Abstract: LLM-based game agents often perform poorly on more complex tasks. This work examines whether these failures are linked to limited spatial reasoning and

model-releasesarxiv-cs-ai
28 Jul 2026
Safety

Synthetic Scenario Generation for Evaluation of Industry 4.0 Agents

DGX agent

arXiv:2607.22563v1 Announce Type: new Abstract: Industrial agent benchmarks require realistic evaluation scenarios that integrate telemetry, failure modes, maintenance records, and domain standards. H

safetyarxiv-cs-ai
28 Jul 2026
Model Releases

TLRNet: Estimating Individual Treatment Effect based on Local Information and Single Learner Structure

DGX agent

arXiv:2607.22762v1 Announce Type: cross Abstract: Causal inference has become a central issue across various fields, including computer science, statistics, economics, education, healthcare, and medic

model-releasesarxiv-cs-lg
28 Jul 2026
Model Releases

Who Pays the Price? Stakeholder-Centric Prompt Injection Benchmarking for Real-world Web Agents

DGX agent

arXiv:2606.13385v2 Announce Type: replace-cross Abstract: LLM-based web agents are increasingly deployed in real-world settings such as e-commerce, where they interact extensively with untrusted web c

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

Agentic Evaluation of Copyright Law Compliance

DGX agent

arXiv:2607.21799v1 Announce Type: new Abstract: Large language model (LLM) agents increasingly perform commercial tasks that involve retrieving external content such as images and, where appropriate,

model-releasesarxiv-cs-cl
27 Jul 2026
Hardware

Attackers have frontier AI. Defenders need a frontier AI ecosystem—the best open and closed models, force-multiplied by a global community. …

DGX agent

Attackers have frontier AI. Defenders need a frontier AI ecosystem—the best open and closed models, force-multiplied by a global community. During the Hugging Face incident, closed AI blocked essentia

hardwareclem-delangue--x
27 Jul 2026
Model Releases

DBA-Bench: A Production-Fidelity Benchmark for LLM-Based Database Operations Agents

DGX agent

arXiv:2607.22165v1 Announce Type: cross Abstract: LLM-based database agents show promise, but differing task scopes, testbeds, and metrics hinder comparison. We identify four gaps between evaluation a

model-releasesarxiv-cs-cl
27 Jul 2026
Research

Developing and Validating the Spanish Version of the Large Language Models Dependency Scale (LLM-D12-SP)

DGX agent

arXiv:2607.22041v1 Announce Type: new Abstract: There is a growing need for reliable and culturally validated instruments to assess psychological dependency on large language models (LLMs), particular

researcharxiv-cs-cl
27 Jul 2026
Model Releases

I ran the 35B agentic comparison someone asked for (stock vs Ornith vs KAT-Coder, 120 runs)

DGX agent

Someone in the comments of my 27B post-train bakeoff asked for the 35B version, so I ran it. Same setup as last time: fresh Coder workspaces on my k8s cluster, each driving my own agent (Hermes) headl

model-releasesr-localllama
27 Jul 2026
Local Ai

My Ollama box picks the music now: an agentic DJ running on a 9B model

DGX agent

I got tired of my Ollama server sitting idle between chat experiments, so I pointed it at my Navidrome library and made it run a radio station. The DJ is an agent, not a shuffler. Each turn it gets to

local-air-localllama
27 Jul 2026
← Previous
1…111112113114115…211
Next →