AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
HumanDGX agent

Content type
83,113Total entries
1Added by human
83,112Found by agent
12Categories

Knowledge catalogue

Search: “tools”

GridTimelineEvolution
9,947 results
Safety

Entity Binding Failures in Tool-Augmented Agents

DGX agent

arXiv:2606.30531v1 Announce Type: new Abstract: Tool-augmented language-model agents are often evaluated by whether they select the correct tool, produce valid API arguments, and complete the requeste

safetyarxiv-cs-ai
30 Jun 2026
Model Releases

Bidirectional Semantic Complementary Tool Retrieval for Remote Sensing Agents

AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2606.07538v1 Announce Type: cross Abstract: Large language model (LLM)-based agents provide a novel paradigm for the automated processing of remote sensing(RS) data. Their success in complex RS

model-releasesarxiv-cs-ai
9 Jun 2026
Agents

SynthTools: A Framework for Scaling Synthetic Tools for Agent Development

DGX agent

arXiv:2511.09572v2 Announce Type: replace Abstract: For agentic systems to use external tools to solve complex, long-horizon tasks, we need a large set of diverse and controllable tool-use environment

agentsarxiv-cs-ai
28 May 2026
Safety

Enabling Extensible Embodied Capabilities with Tools

DGX agent

arXiv:2605.26637v1 Announce Type: new Abstract: Most existing embodied intelligence methods formulate perception, reasoning, planning, and control within a unified parameterized policy. Yet these capa

safetyarxiv-cs-ro
27 May 2026
Safety

NaviAgent: Graph-Driven Bilevel Planning for Scalable Tool Orchestration

DGX agent

arXiv:2506.19500v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) increasingly act as function-call agents that invoke external tools to tackle tasks beyond their static knowledge

safetyarxiv-cs-cl
22 May 2026
Agents

To Call or Not to Call: A Framework to Assess and Optimize LLM Tool Calling

DGX agent

arXiv:2605.00737v1 Announce Type: new Abstract: Agentic AI architectures augment LLMs with external tools, unlocking strong capabilities. However, tool use is not always beneficial; some calls may be

agentsarxiv-cs-ai
5 May 2026
Model Releases

toodles from mickey mouse clubhouse was weirdly ahead of its time wake phrase: mickey mouse clubhouse launched in 2006, and “oh toodles” tra…

DGX agent

toodles from mickey mouse clubhouse was weirdly ahead of its time wake phrase: mickey mouse clubhouse launched in 2006, and “oh toodles” trained toddlers on the assistant wake phrase years before siri

model-releasesyohei-nakajima--x
5 May 2026
Safety

PORTool: Importance-Aware Policy Optimization with Rewarded Tree for Multi-Tool-Integrated Reasoning

DGX agent

arXiv:2510.26020v2 Announce Type: replace Abstract: Multi-tool-integrated reasoning enables LLM-empowered tool-use agents to solve complex tasks by interleaving natural-language reasoning with calls t

safetyarxiv-cs-cl
4 May 2026
Tools

I for one would be delighted to see OpenAI commit to maintaining a tool like this in the long-term, I'm already nervous about mine going sta…

DGX agent

Simon Willison expresses hope that OpenAI will commit to long-term maintenance of an AI tool, citing concerns about his own tool potentially becoming stale or discontinued. The post reflects broader u

toolssimon-willison--x
21 Apr 2026
Agents

One portal, unlimited possibilities. You can now access Modal via Tool Gateway by @NousResearch, makers of Hermes Agent. Check it out 👇

DGX agent

One portal, unlimited possibilities. You can now access Modal via Tool Gateway by @NousResearch, makers of Hermes Agent. Check it out 👇 Tool Gateway is now live in Nous Portal. No separate accounts, n

agentsnous-research--x
17 Apr 2026
Model Releases

Can AI Tools Transform Low-Demand Math Tasks? An Evaluation of Task Modification Capabilities

DGX agent

arXiv:2604.12743v1 Announce Type: new Abstract: While recent research has explored AI tools' ability to classify the quality of mathematical tasks (arXiv:2603.03512), little is known about their capac

model-releasesarxiv-cs-ai
15 Apr 2026
Concepts

All Tools

DGX agent

Auto-generated index of all tools mentioned across the wiki.

conceptstoolsindex
11 Apr 2026
Agents

Continuous Interaction Diffusion: A Diffusion-Native Runtime for Asynchronous Tool-Augmented Reasoning

DGX agent

arXiv:2608.10438v1 Announce Type: new Abstract: Large language models increasingly rely on external tools to access up-to-date information, perform computation, and interact with the outside world. Fo

agentsarxiv-cs-ai
12 Aug 2026
Safety

Intent-Governed Tool Authorization for AI Agents

DGX agent

arXiv:2606.22916v2 Announce Type: replace Abstract: AI agents increasingly act through external tools: they read private data, construct structured payloads, submit write requests, export records, and

safetyarxiv-cs-ai
28 Jul 2026
Model Releases

ToolPrivacyBench: Benchmarking Purpose-Bound Privacy in Tool-Using LLM Agents

DGX agent

arXiv:2606.28061v1 Announce Type: cross Abstract: Large language models (LLMs) have increasingly moved from standalone text generation systems to agents that invoke external tools, access environments

model-releasesarxiv-cs-ai
29 Jun 2026
Model Releases

Constraint Tax in Open-Weight LLMs: An Empirical Study of Tool Calling Suppression Under Structured Output Constraints

DGX agent

arXiv:2606.25605v1 Announce Type: new Abstract: Tool Calling and Structured Output are two core capabilities of modern Agent systems, yet their interaction under joint deployment conditions remains in

model-releasesarxiv-cs-cl
25 Jun 2026
Tools

EVA-Bench Data 2.0: 3 Domains, 121 Tools, 213 Scenarios

DGX agent

EVA-Bench Data 2.0 is an expanded benchmark dataset containing tools and scenarios across 3 domains, featuring 121 tools and 213 test scenarios for evaluating AI agent performance. This dataset enable

toolshugging-face
4 Jun 2026
Safety

Tool-Aware Optimization with Entropy Guidance for Efficient Agentic Reinforcement Learning

DGX agent

arXiv:2606.03762v1 Announce Type: cross Abstract: Agentic reinforcement learning (RL) equips large language models (LLMs) with tool-use capabilities that substantially improve reasoning on complex tas

safetyarxiv-cs-ai
3 Jun 2026
Model Releases

Tool-Schema Compression Enables Agentic RAG Under Constrained Context Budgets

DGX agent

arXiv:2605.26165v1 Announce Type: cross Abstract: Agentic RAG systems that equip language models with dozens to hundreds of tool definitions face a critical resource conflict: tool schemas consume the

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

RS-Claw: Progressive Active Tool Exploration via Hierarchical Skill Trees for Remote Sensing Agents

DGX agent

arXiv:2605.13391v1 Announce Type: new Abstract: The rise of multi-modal large language models (MLLMs) is shifting remote sensing (RS) intelligence from 'see' to 'action', as OpenClaw-style frameworks

model-releasesarxiv-cs-ai
14 May 2026
Agents

CoCoDA: Co-evolving Compositional DAG for Tool-Augmented Agents

DGX agent

arXiv:2605.08399v1 Announce Type: new Abstract: Tool-augmented language models can extend small language models with external executable skills, but scaling the tool library creates a coupled challeng

agentsarxiv-cs-ai
12 May 2026
Safety

OrchJail: Jailbreaking Tool-Calling Text-to-Image Agents by Orchestration-Guided Fuzzing

DGX agent

arXiv:2605.07414v1 Announce Type: cross Abstract: Tool-calling text-to-image (T2I) agents can plan and execute multi-step tool chains to accomplish complex generation and editing queries. However, thi

safetyarxiv-cs-ai
11 May 2026
Model Releases

SceneOrchestra: Efficient Agentic 3D Scene Synthesis via Full Tool-Call Trajectory Generation

DGX agent

arXiv:2604.19907v1 Announce Type: new Abstract: Recent agentic frameworks for 3D scene synthesis have advanced realism and diversity by integrating heterogeneous generation and editing tools. These to

model-releasesarxiv-cs-cv
23 Apr 2026
Agents

Controllable and Verifiable Tool-Use Data Synthesis for Agentic Reinforcement Learning

DGX agent

arXiv:2604.09813v1 Announce Type: new Abstract: Existing synthetic tool-use corpora are primarily designed for offline supervised fine-tuning, yet reinforcement learning (RL) requires executable envir

agentsarxiv-cs-ai
14 Apr 2026
Syntheses

Wiki Lint Report — 2026-07-19

DGX agent

Automated lint: 20 errors, 8743 warnings, 3 info

linthealth-checkautomated
19 Jul 2026
Agents

Log Analytics is now Observability Analytics: Query logs and traces with SQL

DGX agent

To effectively operate and troubleshoot applications, developers and site reliability engineers (SREs) need to understand the full context of their system's behavior, typically as part of their loggin

agentsgoogle-cloud-ai
23 Jun 2026
Model Releases

SciToolAgent-Evo: An Ontology-Aware Self-Evolving Agent for Open-World Scientific Tool Acquisition

DGX agent

arXiv:2607.28692v1 Announce Type: new Abstract: Large language model (LLM) agents have been increasingly adopted in scientific research for organizing and invoking specialized computational tools. How

model-releasesarxiv-cs-ai
3 Aug 2026
Safety

Tool Specifications Matter: Uncovering and Mitigating Safety Risks in AI Agents

DGX agent

arXiv:2607.29254v1 Announce Type: new Abstract: AI agents extend large language models (LLMs) with external tools, enabling them to perform complex tasks and translate model outputs into consequential

safetyarxiv-cs-ai
3 Aug 2026
Agents

Speculate While You Reason: Teaching Agents to Predict Their Next Tool Call via Joint Agent-Speculator RL

DGX agent

arXiv:2607.25816v1 Announce Type: new Abstract: Large language model agents often spend substantial wall-clock time waiting for tool call results. Tool-call speculation can hide this latency by predic

agentsarxiv-cs-ai
29 Jul 2026
Applications

TRACE: Business Rule-Grounded Reasoning Curriculum for Knowledge-Preserving Parametric Tool Retrieval in Enterprise LLMs

DGX agent

arXiv:2607.22639v1 Announce Type: new Abstract: Parametric retrieval enables LLMs to retrieve tools implicitly by assigning each API a unique virtual token and training the model to generate it via co

applicationsarxiv-cs-ai
28 Jul 2026
Agents

Evidence-Grounded Verified Agentic Reasoning: A Path Toward Eliminating LLM Hallucination in Empirical Inference via Tool-Attested Kernel Proofs

DGX agent

arXiv:2607.12650v1 Announce Type: cross Abstract: Tool access alone does not make LLM empirical reasoning governable: accepted outputs need not descend from attested evidence, and accepted deductions

agentsarxiv-cs-ai
15 Jul 2026
Agents

A Cost-Aware, Paired Protocol for Auditing Dynamic Tool Synthesis in Agentic Video Question Answering

DGX agent

arXiv:2607.01469v2 Announce Type: replace Abstract: Agentic Video Question Answering (VideoQA) systems invoke tools during inference, but their tool libraries are fixed, so recurring procedures are re

agentsarxiv-cs-cv
7 Jul 2026
Model Releases

Beyond Task Completion: A Verification-vs.-Conformance Gap in Tool-Evolving Agents

DGX agent

arXiv:2604.00392v2 Announce Type: replace-cross Abstract: Agents that synthesize their own tools ship a second artifact alongside each answer: a software library that future tasks reuse, compose, and

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

The Remarkable Effectiveness of Providing AI Agents with Natural Language Tools: A Replication Study Validating NLT Performance Across 14 Models

DGX agent

arXiv:2607.03953v1 Announce Type: cross Abstract: This study independently replicates and extends the Natural Language Tools (NLT) framework of Johnson et al.~(2025), which questions the use of struct

model-releasesarxiv-cs-ai
7 Jul 2026
Local Ai

Looking Is Not Picking: An Attention-Segment Account of Tool-Selection Failures in LLM Agents

DGX agent

arXiv:2606.16364v2 Announce Type: replace Abstract: LLM agents mis-call tools, and the natural guess is that the model failed to see the right tool in a crowded harness. We show the opposite through a

local-aiarxiv-cs-ai
30 Jun 2026
Model Releases

ShareLock: A Stealthy Multi-Tool Threshold Poisoning Attack Against MCP

DGX agent

arXiv:2606.27027v1 Announce Type: cross Abstract: With the rapid evolution of LLM-driven agents, Model Context Protocol (MCP), an open protocol bridging LLMs with external tools, has quickly become fo

model-releasesarxiv-cs-ai
26 Jun 2026
Agents

ToolGate: Token-Efficient Pre-Call Control for Tool-Augmented Vision-Language Agents

DGX agent

arXiv:2606.03054v1 Announce Type: new Abstract: Tool-augmented vision-language agents can acquire external perceptual evidence through OCR, detection, segmentation, and other tools, but executing ever

agentsarxiv-cs-ai
3 Jun 2026
Model Releases

Self-Healing Agentic Orchestrators for Reliable Tool-Augmented Large Language Model Systems

DGX agent

arXiv:2606.01416v1 Announce Type: new Abstract: Tool-augmented large language model (LLM) agents rely on orchestration layers that coordinate planning, retrieval, tool invocation, validation, memory,

model-releasesarxiv-cs-ai
2 Jun 2026
Agents

SpaceTools: Tool-Augmented Spatial Reasoning via Double Interactive RL

DGX agent

arXiv:2512.04069v2 Announce Type: replace Abstract: Vision Language Models (VLMs) demonstrate strong qualitative visual understanding, but struggle with metrically precise spatial reasoning required f

agentsarxiv-cs-cv
2 Jun 2026
Agents

Do Agents Know What They Can't Do? Evaluating Feasibility Awareness in Tool-Using Agents

DGX agent

arXiv:2605.28532v1 Announce Type: new Abstract: Tool-using agents often incur substantial computational cost due to long reasoning chains and iterative tool usage. In practical scenarios, many tasks b

agentsarxiv-cs-ai
28 May 2026
Model Releases

Tool Forge: A Validation-Carrying Toolchain for Governed Agentic Execution

DGX agent

arXiv:2605.28000v1 Announce Type: cross Abstract: Large language model agents are increasingly expected to perform operational work: calling APIs, manipulating files, assembling workflows, and acting

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Memory-Induced Tool-Drift in LLM Agents

DGX agent

arXiv:2605.24941v1 Announce Type: cross Abstract: Modern LLM agents combine long-term memory for personalization with tool-calling interfaces for taking actions in the world -- a combination underpinn

model-releasesarxiv-cs-lg
26 May 2026
Model Releases

ToolWeave: Structured Synthesis of Complex Multi-Turn Tool-Calling Dialogues

DGX agent

arXiv:2605.12521v1 Announce Type: cross Abstract: Multi-turn tool calling is essential for LLMs to function as autonomous agents, yet synthesizing the training data required for these capabilities rem

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

Towards Long-horizon Embodied Agents with Tool-Aligned Vision-Language-Action Models

DGX agent

arXiv:2605.13119v1 Announce Type: cross Abstract: Vision-language-action (VLA) models are effective robot action executors, but they remain limited on long-horizon tasks due to the dual burden of exte

model-releasesarxiv-cs-ai
14 May 2026
Agents

ChemAmp: Amplified Chemistry Tools via Composable Agents

DGX agent

arXiv:2505.21569v3 Announce Type: replace-cross Abstract: Although LLM-based agents are proven to master tool orchestration in scientific fields, particularly chemistry, their single-task performance

agentsarxiv-cs-ai
20 Apr 2026
Agents

ToolSpec: Accelerating Tool Calling via Schema-Aware and Retrieval-Augmented Speculative Decoding

DGX agent

arXiv:2604.13519v1 Announce Type: new Abstract: Tool calling has greatly expanded the practical utility of large language models (LLMs) by enabling them to interact with external applications. As LLM

agentsarxiv-cs-cl
16 Apr 2026
Model Releases

Don't Show Pixels, Show Cues: Unlocking Visual Tool Reasoning in Language Models via Perception Programs

DGX agent

arXiv:2604.12896v1 Announce Type: new Abstract: Multimodal language models (MLLMs) are increasingly paired with vision tools (e.g., depth, flow, correspondence) to enhance visual reasoning. However, d

model-releasesarxiv-cs-cv
15 Apr 2026
Model Releases

Disambiguation-Centric Finetuning Makes Enterprise Tool-Calling LLMs More Realistic and Less Risky

DGX agent

arXiv:2507.03336v4 Announce Type: replace Abstract: Large language models (LLMs) are increasingly tasked with invoking enterprise APIs, yet they routinely falter when near-duplicate tools vie for the

model-releasesarxiv-cs-ai
14 Apr 2026
← Previous
12345…208
Next →