AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “tools”

GridTimelineEvolution
10,030 results
10 Aug 2026

WebRider: Persona-Conditioned Intent Controllers for Live-Web Assistance

Model ReleasesDGX agent

arXiv:2608.06704v1 Announce Type: new Abstract: Delegating a web task involves more than asking a question; it requires transferring a policy: what to verify, how to handle uncertainty, which preferen

9 Aug 2026

Stripe just published how their company-wide AI agent works. The bar for building one just dropped to one engineer and one week It is called…

AgentsDGX agent

Stripe just published how their company-wide AI agent works. The bar for building one just dropped to one engineer and one week It is called Kai. Their own words: a coding agent for non-engineers. You

The Gemma team will host a special event on August 20

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases
DGX agent

Tweet by u/hackerllama Could be copium, but I would love to see Gemma 4.1 there with unified audio input for all model sizes perhaps even up to 120B, much improved tool calling (even with the latest t

8 Aug 2026

b10331

Model ReleasesDGX agent

server: report the isolate working directory from get_info (#26773) server: report the isolate working directory from get_info Without an explicit cwd, get_info fell back to the server process working

7 Aug 2026

Agentic self-driving microscopy benchmarks support qualification but do not necessarily generalize to unseen tasks

Model ReleasesDGX agent

arXiv:2608.05266v1 Announce Type: new Abstract: Large language model agents are increasingly being developed to control a wide range of scientific characterization tools including microscopes and sync

ChainClaw: A Layered Agent Framework for Reliable On-Chain Execution

Model ReleasesDGX agent

arXiv:2608.05790v1 Announce Type: new Abstract: General-purpose large language model agents have achieved strong performance on tool-augmented tasks, yet they rely on assumptions break down in blockch

CohortHijack: Robustness of Single Cell Annotation to Companion Cell Removal

ResearchDGX agent

arXiv:2608.05900v1 Announce Type: new Abstract: Many single-cell annotation tools refine an initial cell label using nearby cells or cluster-level voting. We study whether this refinement can be manip

Confidence matters: Leveraging Multi-view Geometric Priors for GS-based Reconstruction

ResearchDGX agent

arXiv:2608.06117v1 Announce Type: new Abstract: 3D Gaussian splatting (3DGS) has emerged as a widely-used tool for novel view synthesis, offering real-time rendering in a sparse representation. Howeve

DreamGuard: Efficient Runtime Guardrail for LLM Agents via Risk-Aware World Model

SafetyDGX agent

arXiv:2608.05695v1 Announce Type: new Abstract: As large language model (LLM) agents increasingly invoke external tools and interact with real-world systems, unsafe actions may cause irreversible cons

ECG-LENS: Lead-Aware Clinical Context Enriched ECG Report Generation and Evaluation

Local AiDGX agent

arXiv:2608.05893v1 Announce Type: new Abstract: Electrocardiography (ECG) is one of the most widely used non-invasive tools for diagnosing cardiovascular disease, but transforming multi-lead ECG recor

EcoAgent-Bench: Evaluating Economic Decision-Making in Budget-Constrained LLM Agents

Model ReleasesDGX agent

arXiv:2608.05519v1 Announce Type: new Abstract: Agent benchmarks usually measure task completion and treat resource use as an auxiliary statistic. In deployment, however, the choice among a local look

From Passive Mirrors to Active Agents: Holonic Digital Twins for Physical AI over Networks

Local AiDGX agent

arXiv:2608.06227v1 Announce Type: cross Abstract: Despite advances in artificial intelligence (AI) across multiple sectors, today's AI tools, including deep learning and generative AI, still fail when

HERALD: Counterfactual Audits and Minimal Repairs for Proof-of-Retrieval Rewards

Model ReleasesDGX agent

arXiv:2608.06012v1 Announce Type: new Abstract: Search-agent rewards mix answer quality, citation grounding, tool cost, and anti-hacking terms; a high score therefore need not imply that cited evidenc

How Cohere Health digitizes clinical policies using Amazon Bedrock AgentCore

SafetyDGX agent

In this post, you learn how Cohere Health built a multi-tenant agentic architecture on AgentCore using AgentCore Runtime’s secure MicroVM isolation, unified tool access through AgentCore Gateway, Agen

Managing AI Coding Costs at Scale

AgentsDGX agent

**Managing AI Coding Costs at Scale** This article addresses the economic challenges of deploying AI coding tools across large teams or organizations. It explores strategies for tracking resource usag

OmniMech: All-in-one Multimodal Mechanical Benchmark for 3D Reconstruction

Model ReleasesDGX agent

arXiv:2608.05539v1 Announce Type: new Abstract: Recent vision-language models (VLMs) can generate executable CAD programs from images, but existing methods mainly target coarse, general-purpose 3D obj

SkillTrace: Multi-Trace Provenance Auditing for LLM-Agent Skill Reuse

AgentsDGX agent

arXiv:2608.05204v1 Announce Type: new Abstract: LLM-agent ecosystems are rapidly growing around reusable skills: mixed-modality packages of metadata, natural-language instructions, code, tools, refere

SkillTV-Bench: Benchmarking How Well Judges Perform on Skill-Augmented Agentic Execution

Model ReleasesDGX agent

arXiv:2608.05573v1 Announce Type: new Abstract: LLM agents increasingly execute long-horizon tasks through tool use and environment interaction, shifting evaluation from final-response scoring to veri

Sparse Mixture-of-Experts for Non-Uniform Noise Reduction in MRI Images

ApplicationsDGX agent

arXiv:2501.14198v3 Announce Type: replace-cross Abstract: Magnetic Resonance Imaging (MRI) is an essential diagnostic tool in clinical settings, but its utility is often hindered by noise artifacts in

Tytan: Interactive Neurosymbolic Construction of Analytic Semantic Schemas from Relational Data

Model ReleasesDGX agent

arXiv:2608.06331v1 Announce Type: cross Abstract: From natural-language query interfaces to automated report generation, data analysis tools need a description of the data: the real-world entities it

When Do Prompt-Side Agent Playbooks Transfer? Accuracy, Cost, and Runtime Shift in Agent Deployment

AgentsDGX agent

arXiv:2608.05778v1 Announce Type: new Abstract: Prompt-side playbooks can improve tool-using language agents without retraining, but their portability beyond the source setting is unclear. We study fr

6 Aug 2026

Broadcom updates VMware security for AI-era threats

IndustryDGX agent

Broadcom Inc. today introduced new security, performance and automation capabilities for VMware Cloud Foundation that are intended to help enterprises defend private clouds against artificial intellig

Building an agentic app deployer with Amazon Bedrock and AWS Lambda

AgentsDGX agent

PDI Technologies built PDI Brew, an agentic platform on AWS where non-technical employees describe a tool in plain English and receive a fully provisioned, multi-tenant web application in seconds. See

Capability-Gated Planning: Cost-to-Goal Discovery and the Limits of Myopic Experiment Selection

ResearchDGX agent

arXiv:2608.05085v1 Announce Type: cross Abstract: Systems that automate scientific discovery must repeatedly decide which experiment to run, which hypothesis to test, which tool to build, and when to

DeepAmbigQA: Ambiguous Multi-hop Questions for Benchmarking LLM Answer Completeness

ResearchDGX agent

Large language models (LLMs) with integrated search tools show strong promise in open-domain question answering (QA), yet they often struggle to produce complete answer set to complex questions such a

EvolveNet: Collaborative Harness Evolution for Agent Self-Improvement

Local AiDGX agent

arXiv:2608.04968v1 Announce Type: new Abstract: The capabilities of an LLM agent depend not only on its model but on the harness: the executable program that constructs context, invokes tools, verifie

Introducing Kitesurf: The agent-first browser that runs in V8 isolates on Cloudflare Workers

AgentsDGX agent

We should be giving all agents tools that excel at what’s important for an AI model. Kitesurf is Cloudflare’s new stateless, highly scalable, and cost-effective web browser that runs entirely on top o

nvidias nemotron omni only loads its text half on a mac, so i wrote the vision and audio towers in mlx

Model ReleasesDGX agent

nvidias nemotron omni is open weights and it sees, hears and reasons. theres already a 4bit mlx quant on hugging face but only the text backbone loads with standard mlx tooling. the model card says it

OmniRouting: A Semantic-Coupled Multimodal Benchmark for Constraint-Aware Spatial Reasoning in PCB Routing

Model ReleasesDGX agent

arXiv:2608.04434v1 Announce Type: new Abstract: Recent large language models (LLMs) have demonstrated remarkable progress in constraint-aware navigation, maze reasoning, and graph reasoning. However,

Open models give teams more room to run agent loops, with API prices at a fraction of GPT-5.6 Sol and Claude Fable 5. That matters as planni…

Model ReleasesDGX agent

Open models give teams more room to run agent loops, with API prices at a fraction of GPT-5.6 Sol and Claude Fable 5. That matters as planning, tool calls, retries, and long contexts compound token us

Q-CueGraph: Query-Conditioned Visual Evidence Graphs for Multimodal Reasoning

SafetyDGX agent

arXiv:2608.04452v1 Announce Type: cross Abstract: High-resolution pixels and crop or zoom tools give multimodal large language models the ability to inspect an image, but they do not provide a reliabl

Skill-Use: Can LLMs Actually Use Skills in Agentic Harnesses?

Model ReleasesDGX agent

arXiv:2608.04828v1 Announce Type: new Abstract: Large language model (LLM) agents increasingly rely on skills, structured documents that specify when to act, which procedure to follow, and which tools

The Neural Echo: A Signal Processing Perspective for Understanding Neural Networks

Local AiDGX agent

arXiv:2608.04864v1 Announce Type: cross Abstract: We introduce the neural echo as a tool for understanding the behavior of neural networks. It generalizes the model-based concepts of impulse responses

5 Aug 2026

Agentic AI forces a reckoning on governance as autonomous actors enter production

AgentsDGX agent

As AI agents move from experimental chatbots into production systems, enterprises must rethink agent governance as autonomous actors gain access to sensitive data, tools and business processes that tr

AI-Assisted Peer Review Across Research Communities: From Reviewer AI Policies to LLM Review Quality

SafetyDGX agent

arXiv:2608.03581v1 Announce Type: cross Abstract: AI-assisted peer review is increasingly discussed and adopted as a tool to support the scientific publishing process, yet there is little systematic u

Conformal risk control for model-form uncertainty in parametric non-intrusive reduced-order models

ApplicationsDGX agent

arXiv:2608.03360v1 Announce Type: cross Abstract: Non-intrusive reduced-order models (NIROMs) have become a standard tool for approximating parametric partial differential equations from computer desi

Formal Verification of Agentic Systems over Operational Data

AgentsDGX agent

arXiv:2608.03609v1 Announce Type: new Abstract: Agentic systems driven by large language models (LLMs) are increasingly deployed in real-world workflows where they act on persistent operational data.

Hybrid LLM-Augmented Reinforcement Learning Agents for Complex Sequential Decision Tasks

AgentsDGX agent

arXiv:2608.03502v1 Announce Type: new Abstract: Large Language Models (LLMs) have recently shown strong capabilities in reasoning, planning, and tool-use, enabling new forms of autonomous agents. Howe

Meta is offering a cheaper Muse Spark 1.2 'contributor' tier priced at 0.10/1M input and 0.20/1M output tokens in exchange for using user prompts for training (Wall Street Journal)

AgentsDGX agent

Wall Street Journal: Meta is offering a cheaper Muse Spark 1.2 “contributor” tier priced at 0.10/1M input and 0.20/1M output tokens in exchange for using user prompts for training — The company, press

NOMADD: Numerical Optimization of Models Adapting to Data Drift

Model ReleasesDGX agent

arXiv:2608.02845v1 Announce Type: new Abstract: Tabular model performance degrades when feature distributions change over time or the relationship between features and outcome variables change over ti

Robust Biharmonic Skinning Using Geometric Fields

ResearchDGX agent

arXiv:2406.00238v3 Announce Type: replace-cross Abstract: Bounded bihramonic weights are a popular tool used to rig and deform characters for animation, to compute reduced-order simulations, and to de

S^3: Improving Agent Safety through Multi-Stage Defense

Model ReleasesDGX agent

arXiv:2608.02683v1 Announce Type: cross Abstract: Large Language Model (LLM) agents rely on multi-stage agentic workflows, with stages such as memory, planning, and tool execution, to accomplish compl

SAT-Edge-Agent: Hardware-in-the-Loop Edge-Agent Orchestration for Onboard Satellite Intelligence

Local AiDGX agent

arXiv:2608.03728v1 Announce Type: new Abstract: Onboard satellite intelligence requires a task layer that translates mission intent into local tool calls, exposes execution state, and returns machine-

Scaling agentic AI: How UiPath built its high-performance GPU platform on AI Hypercomputer

Model ReleasesDGX agent

As a market leader in enterprise agentic automation and business orchestration, UiPath is helping to pioneer an industry shift toward agentic AI. With it, the company is deploying autonomous agents to

Suffix-Constrained Greedy Search Algorithms for Causal Language Models

ResearchDGX agent

arXiv:2603.01243v2 Announce Type: replace Abstract: Large language models (LLMs) are powerful tools that have found applications beyond human-machine interfaces and chatbots. Beside free-form generati

Surrogate Substitution Preserves PHI Detectability: A Multi-Detector Equivalence Study

ResearchDGX agent

arXiv:2608.03172v1 Announce Type: new Abstract: Structure-preserving de-identification replaces protected health information (PHI) with realistic same-type surrogates -- 'Anna S.' becomes 'Maria S.',

The Agent Operating System (AOS): A Reference Operating Architecture for Distributed Agentic Systems

SafetyDGX agent

arXiv:2608.03214v1 Announce Type: new Abstract: Large language models have transformed artificial intelligence from isolated prediction services into components of long-running, distributed systems th

Video-DeepResearch: Towards the Next-Generation Multimodal Deepresearch Agent

Model ReleasesDGX agent

arXiv:2608.03979v1 Announce Type: cross Abstract: We introduce Video-DeepResearch (Video-DR), extending multimodal agents from static images to continuous video streams, a setting that demands dense s

Vulnerabilities, Secrets and Misconfiguration in the Highest-Exposure Docker Hub Images

Model ReleasesDGX agent

arXiv:2608.02669v1 Announce Type: cross Abstract: Docker Hub is the registry underneath most container deployments, and a flaw in a widely reused base image is inherited by every image built on it. Pr

Your Agentic LLMs Secretly Encode Latent Signals of Indirect Prompt-Injection Exposure

Model ReleasesDGX agent

arXiv:2608.02657v1 Announce Type: cross Abstract: Agentic LLMs are vulnerable to indirect prompt injection (IPI) attacks, e.g., malicious side-tasks hidden in external tool results. While many efforts

4 Aug 2026

Agentic Bayesian Optimization through Surrogate-Augmented Autoresearch

AgentsDGX agent

arXiv:2608.00316v1 Announce Type: new Abstract: Bayesian optimization (BO) has become the standard tool for sample-efficient optimization and owes its efficiency to uncertainty-aware search driven by

b10254

Model ReleasesDGX agent

chat : add new template for DeepSeek V4 Flash 0731 (#26398) common/chat: update DeepSeek V4 templates Align the DeepSeek V4 templates with the official encoders while keeping parser behavior out of th

b10275

Model ReleasesDGX agent

server: decode Windows OEM output to UTF-8 in built-in tools (#26597) a child process writes in the OEM code page, which is not UTF-8 on a western Windows install, so accented output reaches the JSON

Capability Provenance in Language Models: A Case Study in Social Reasoning

Model ReleasesDGX agent

arXiv:2606.19625v2 Announce Type: replace Abstract: We use training-data attribution as an interpretable tool for capability discovery, mapping which regions of the pretraining corpus support social-r

Constructing Parallel Multidimensional Chromatic Lexicons for Corpus-Assisted Analysis of Russian and English Texts

ResearchDGX agent

arXiv:2608.01752v1 Announce Type: new Abstract: This article addresses the relative scarcity of research tools for the corpus-assisted linguistic analysis of colour terms in literary texts. It describ

DeepSeek V4 Flash is now live on Together AI. Frontier agent performance is getting dramatically cheaper. DSV4 on Together AI brings a major…

Model ReleasesDGX agent

DeepSeek V4 Flash is now live on Together AI. Frontier agent performance is getting dramatically cheaper. DSV4 on Together AI brings a major jump in coding, tool use, and long-running agent performanc

Extended KAFR: A kinematic-adaptive paradigm for the efficient analysis of surgical video

Model ReleasesDGX agent

arXiv:2608.01058v1 Announce Type: new Abstract: Artificial Intelligence is increasingly applied to surgical video analysis for phase segmentation, skill assessment, and workflow optimization. A key ch

How the GitHub legal team used Copilot CLI to streamline their workflows

TutorialsDGX agent

Learn how to build tools to simplify how you work—without writing a single line of code. The post How the GitHub legal team used Copilot CLI to streamline their workflows appeared first on The GitHub

Instruction-Conditioned Exploration with Asymmetric Reinforcement Learning and Self-Distillation

SafetyDGX agent

arXiv:2608.02087v1 Announce Type: cross Abstract: Post-training Large Language Models (LLMs) with Reinforcement Learning (RL) has become an important tool for improving model capabilities, but the LLM

LongHorizon-Harness: Advancing Long-Horizon Agents for Real-World Tasks

Model ReleasesDGX agent

arXiv:2608.01964v1 Announce Type: new Abstract: Large language model (LLM) agents increasingly undertake long-horizon tasks that require sustained reasoning, tool use, and revision across many interde

← Previous
1…6061626364…168
Next →