AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “tools”

GridTimelineEvolution
5,141 results
Agents

When Does Memory Help Multi-Trajectory Inference for Tool-Use LLM Agents?

DGX agent

arXiv:2605.28224v1 Announce Type: new Abstract: Multi-trajectory inference for tool-use LLM agents - generating multiple reasoning attempts and selecting among them - benefits from transferring knowle

agentsarxiv-cs-ai
28 May 2026
Model Releases
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

LiveMCP-101: Stress Testing and Diagnosing MCP-enabled Agents on Challenging Queries

DGX agent

arXiv:2508.15760v2 Announce Type: replace-cross Abstract: Tool calling has emerged as a critical capability for AI agents. In contrast to conventional tool calling frameworks that rely on static, prov

model-releasesarxiv-cs-ai
26 May 2026
Local Ai

IndusAgent: Reinforcing Open-Vocabulary Industrial Anomaly Detection with Agentic Tools

DGX agent

arXiv:2605.20682v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have shown remarkable capability in bridging visual perception and textual reasoning, enabling zero-shot unders

local-aiarxiv-cs-cv
21 May 2026
Research

A Cloud-Based Tool for Meteorite Recovery Using Drones and Machine Learning

DGX agent

arXiv:2605.19179v1 Announce Type: cross Abstract: We present a cloud-based tool that uses drones and machine learning to help recover instrumentally observed meteorite falls. We showcase a collection

researcharxiv-cs-lg
20 May 2026
Agents

EnvFactory: Scaling Tool-Use Agents via Executable Environments Synthesis and Robust RL

DGX agent

arXiv:2605.18703v1 Announce Type: new Abstract: Equipping LLMs with tool-use capabilities via Agentic Reinforcement Learning (Agentic RL) is bottlenecked by two challenges: the lack of scalable, robus

agentsarxiv-cs-cl
19 May 2026
Research

ClickRemoval: An Interactive Open-Source Tool for Object Removal in Diffusion Models

DGX agent

arXiv:2605.14461v1 Announce Type: new Abstract: Existing object removal tools often rely on manual masks or text prompts, making precise removal difficult for non-expert users in complex scenes and of

researcharxiv-cs-cv
15 May 2026
Model Releases

Building Interactive Real-Time Agents with Asynchronous I/O and Speculative Tool Calling

DGX agent

arXiv:2605.13360v1 Announce Type: new Abstract: There is a growing demand for agentic AI technologies for a range of downstream applications like customer service and personal assistants. For applicat

model-releasesarxiv-cs-lg
14 May 2026
Safety

When to Ask a Question: Understanding Communication Strategies in Generative AI Tools

DGX agent

arXiv:2605.11240v1 Announce Type: cross Abstract: Generative AI models differ from traditional machine learning tools in that they allow users to provide as much or as little information as they choos

safetyarxiv-cs-lg
13 May 2026
Model Releases

Trajectory Supervision for Continual Tool-Use Learning in LLMs

DGX agent

arXiv:2605.09734v1 Announce Type: cross Abstract: Most language-model training data shows final artifacts, not the process that produced them. We study a tractable version of this question in tool use

model-releasesarxiv-cs-ai
12 May 2026
Research

Exploring the Adoption Intention in Using AI-Enabled Educational Tools Among Preservice Teachers in the Philippines: A Partial-Least Square Modeling

DGX agent

arXiv:2604.27346v1 Announce Type: cross Abstract: This study examines the factors influencing pre-service teachers' behavioral intention to use AI-enabled educational tools during their practicum, usi

researcharxiv-cs-ai
1 May 2026
Safety

ATLAS: An Annotation Tool for Long-horizon Robotic Action Segmentation

DGX agent

arXiv:2604.26637v1 Announce Type: cross Abstract: Annotating long-horizon robotic demonstrations with precise temporal action boundaries is crucial for training and evaluating action segmentation and

safetyarxiv-cs-ai
30 Apr 2026
Local Ai

A Digital Pathology Resource for Liver Cancer Quantification with Datasets, Benchmarks, and Tools

DGX agent

arXiv:2604.22858v1 Announce Type: new Abstract: Liver cancer, especially hepatocellular carcinoma (HCC), imposes a substantial global disease burden. Accurate diagnosis and prognostic assessment direc

local-aiarxiv-cs-cv
28 Apr 2026
Model Releases

Complete Cyclic Subtask Graphs for Tool-Using LLM Agents: Flexibility, Cost, and Bottlenecks in Multi-Agent Workflows

DGX agent

arXiv:2604.22820v1 Announce Type: cross Abstract: Long-horizon tool-using tasks sometimes benefit from revisiting earlier subtasks for recovery and exploration, but added multi-agent workflow flexibil

model-releasesarxiv-cs-ai
28 Apr 2026
Research

Video-STAR: Reinforcing Open-Vocabulary Action Recognition with Tools

DGX agent

arXiv:2510.08480v2 Announce Type: replace Abstract: Multimodal large language models (MLLMs) have demonstrated remarkable potential in bridging visual and textual reasoning, yet their reliance on text

researcharxiv-cs-cv
20 Apr 2026
Model Releases

ClawVM: Harness-Managed Virtual Memory for Stateful Tool-Using LLM Agents

DGX agent

arXiv:2604.10352v1 Announce Type: new Abstract: Stateful tool-using LLM agents treat the context window as working memory, yet today's agent harnesses manage residency and durability as best-effort, c

model-releasesarxiv-cs-ai
14 Apr 2026
Research

DynamicsLLM: a Dynamic Analysis-based Tool for Generating Intelligent Execution Traces Using LLMs to Detect Android Behavioural Code Smells

DGX agent

arXiv:2604.10661v1 Announce Type: cross Abstract: Mobile apps have become essential of our daily lives, making code quality a critical concern for developers. Behavioural code smells are characteristi

researcharxiv-cs-ai
14 Apr 2026
Agents

HTAA: Enhancing LLM Planning via Hybrid Toolset Agentization & Adaptation

DGX agent

arXiv:2604.10917v1 Announce Type: new Abstract: Enabling large language models to scale and reliably use hundreds of tools is critical for real-world applications, yet challenging due to the inefficie

agentsarxiv-cs-cl
14 Apr 2026
Safety

E3-TIR: Enhanced Experience Exploitation for Tool-Integrated Reasoning

DGX agent

arXiv:2604.09455v1 Announce Type: new Abstract: While Large Language Models (LLMs) have demonstrated significant potential in Tool-Integrated Reasoning (TIR), existing training paradigms face signific

safetyarxiv-cs-ai
13 Apr 2026
Model Releases

TiAb Review Plugin: A Browser-Based Tool for AI-Assisted Title and Abstract Screening

DGX agent

arXiv:2604.08602v1 Announce Type: cross Abstract: Background: Server-based screening tools impose subscription costs, while open-source alternatives require coding skills. Objectives: We developed a b

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

TraceSafe: A Systematic Assessment of LLM Guardrails on Multi-Step Tool-Calling Trajectories

DGX agent

arXiv:2604.07223v1 Announce Type: cross Abstract: As large language models (LLMs) evolve from static chatbots into autonomous agents, the primary vulnerability surface shifts from final outputs to int

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

Set-shifting Behavioral Test for Harnessed Agents

DGX agent

arXiv:2607.13396v1 Announce Type: new Abstract: What happens to an LLM agent's tool choice when the reliable tool silently changes within an ongoing session? We borrow set-shifting from cognitive psyc

model-releasesarxiv-cs-ai
16 Jul 2026
Safety

Behavior Leverage Imbalance in Multi-Teacher On-Policy Distillation

DGX agent

arXiv:2607.07050v1 Announce Type: new Abstract: Agentic language models must learn when to call tools, when to consume tool responses, and when to answer directly. This makes multi-teacher on-policy d

safetyarxiv-cs-cl
9 Jul 2026
Model Releases

AsyncTool: Evaluating the Asynchronous Function Calling Capability under Multi-Task Scenarios

DGX agent

arXiv:2605.27995v1 Announce Type: new Abstract: Large language model (LLM)-based agents have shown strong capabilities in using external tools to solve complex tasks. However, existing evaluations oft

model-releasesarxiv-cs-ai
28 May 2026
Safety

Efficient Agentic Reinforcement Learning with On-Policy Intrinsic Knowledge Boundary Enhancement

DGX agent

arXiv:2605.26952v1 Announce Type: new Abstract: Agentic reinforcement learning (RL) has proven effective for training LLM-based agents with external tool-use capabilities. However, we identify that ag

safetyarxiv-cs-cl
27 May 2026
Model Releases

Persistent Semantic Entities in Tool-Augmented LLM Systems

DGX agent

arXiv:2608.07952v1 Announce Type: cross Abstract: Tool-augmented LLM agents can harbor implicit state that persists across sessions, activates through events, and propagates across agent boundaries---

model-releasesarxiv-cs-ai
11 Aug 2026
Research

Preview-Based Relative-Motion Control of an Insertion Tool for Neural-Thread Placement in Pulsating Tissue

DGX agent

arXiv:2608.08860v1 Announce Type: cross Abstract: Robotic neural-thread placement requires regulating the insertion-tool tip relative to tissue that moves with cardiac and respiratory pulsation. This

researcharxiv-cs-ro
11 Aug 2026
Model Releases

Hallucinations on the Board: Tool-Augmented Evaluation of LLM Chess Commentary

DGX agent

arXiv:2608.04240v1 Announce Type: cross Abstract: Superhuman game engines in domains like chess have made expert-level evaluations easily accessible, yet they communicate what is true without the natu

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

AMTFV: Agentic Mathematical Tool-Flow Verification for LLM Self-Correction

DGX agent

arXiv:2607.29549v1 Announce Type: new Abstract: Large language models have demonstrated strong mathematical problem-solving capabilities, yet reliably verifying their candidate answers remains challen

model-releasesarxiv-cs-ai
3 Aug 2026
Local Ai

MOT-SR: Multi-Objective Tool-Augmented Scientific Equation Discovery with Large Language Models

DGX agent

arXiv:2607.29561v1 Announce Type: cross Abstract: Symbolic Regression (SR) aims to discover analytical equations from observational data and plays a central role in scientific modeling. While recent L

local-aiarxiv-cs-ai
3 Aug 2026
Safety

What Can Be Enforced? A Theory of Certified Runtime Safety for Tool-Using Agents

DGX agent

arXiv:2607.22868v1 Announce Type: new Abstract: Runtime guardrails act before irreversible tool calls, but their guarantees depend on what policy state is representable, what a judge observes, and whe

safetyarxiv-cs-ai
28 Jul 2026
Safety

DynaMark: A Reinforcement Learning Framework for Dynamic Watermarking in Industrial Machine Tool Controllers

DGX agent

arXiv:2508.21797v2 Announce Type: replace-cross Abstract: Industry 4.0's highly networked Machine Tool Controllers (MTCs) are prime targets for replay attacks that use outdated sensor data to manipula

safetyarxiv-cs-ai
24 Jul 2026
Research

emb-diversity: A Tool for Embedding-Based Measurement of Data Diversity

DGX agent

arXiv:2607.19848v1 Announce Type: new Abstract: There is growing evidence that data diversity is crucial for developing fair and robust NLP models. However, current approaches to measure diversity rem

researcharxiv-cs-cl
23 Jul 2026
Model Releases

NEXUS: Structured Runtime Safety for Tool-Using LLM Agents

DGX agent

arXiv:2607.19356v1 Announce Type: new Abstract: Tool-using LLM agents increasingly execute high-impact actions, making runtime safety monitoring essential. We present NEXUS (Neural EXecution Utility a

model-releasesarxiv-cs-ai
23 Jul 2026
Agents

Multi-Agent Collaborative Reasoning with Tool-Augmented Evidence for Urban Region Profiling

DGX agent

arXiv:2607.13558v1 Announce Type: new Abstract: Urban region profiling constitutes a core problem in urban computing, supporting applications such as population estimation, economic assessment, and en

agentsarxiv-cs-ai
16 Jul 2026
Model Releases

NKI-Agent: Domain-Specific Fine-Tuning and Agentic Tool Use for Neuron Kernel Generation

DGX agent

arXiv:2607.04395v1 Announce Type: new Abstract: Recent agentic approaches to LLM-based kernel generation have achieved impressive results on CUDA. For emerging AI accelerators such as AWS Trainium and

model-releasesarxiv-cs-lg
7 Jul 2026
Model Releases

VideoSearcher: Empowering Video Deep Research with Multi-Tool Agentic Reasoning via Reinforcement Learning

DGX agent

arXiv:2607.02927v1 Announce Type: cross Abstract: Video understanding is moving beyond closed-context perception toward open-world evidence exploration, a paradigm formalized as Video Deep Research (V

model-releasesarxiv-cs-ai
7 Jul 2026
Research

KnowledgeDebugger -- an Exploration Tool for Knowledge Localization and Editing in Transformers

DGX agent

arXiv:2607.01000v1 Announce Type: new Abstract: Recent research has increasingly focused on understanding how Transformers store and process knowledge, as well as how this knowledge can be edited. Res

researcharxiv-cs-cl
2 Jul 2026
Model Releases

An Executable Benchmarking Suite for Tool-Using Agents

DGX agent

arXiv:2605.11030v2 Announce Type: replace-cross Abstract: Closed-loop tool-using agents are increasingly evaluated in executable web, code, and micro-task environments, but benchmark reports often con

model-releasesarxiv-cs-ai
1 Jul 2026
Model Releases

Think in English, Answer in Korean: Efficient Adaptation of Multilingual Tool-Using Agents

DGX agent

arXiv:2606.31648v1 Announce Type: new Abstract: We present LuckyStar 111B, a 111B-parameter hybrid reasoning model developed through a collaboration between Cohere and LG CNS for Korean-English enterp

model-releasesarxiv-cs-ai
1 Jul 2026
Model Releases

From Tool Connection to Execution Control: Benchmarking Security Invariants in MCP-Style Agent Runtimes

DGX agent

arXiv:2606.29073v1 Announce Type: cross Abstract: Model Context Protocol (MCP)-style ecosystems give language-model applications a practical connection layer for tools, resources, prompts, and transpo

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

How Do Tool-Augmented LLM Agents Perform on Real-World Energy Analytics Tasks?

DGX agent

arXiv:2606.26346v1 Announce Type: new Abstract: Agentic benchmarks have emerged across general-purpose and domain-specific settings, including finance, coding, law, and drug discovery, yet energy-doma

model-releasesarxiv-cs-ai
26 Jun 2026
Safety

JailbreakOPT: Tool-Assisted Iterative Jailbreak Prompt Optimization

DGX agent

arXiv:2606.11425v1 Announce Type: cross Abstract: Jailbreak attacks expose persistent safety weaknesses in large language models (LLMs), but existing stateless single-turn methods face a trade-off: ha

safetyarxiv-cs-ai
11 Jun 2026
Local Ai

AnomaMind: Agentic Time Series Anomaly Detection with Tool-Augmented Reasoning

DGX agent

arXiv:2602.13807v2 Announce Type: replace Abstract: Time series anomaly detection is critical in many real-world applications, where effective solutions must localize anomalous regions and support rel

local-aiarxiv-cs-lg
10 Jun 2026
Model Releases

Decision-Aware Memory Cards: Counterfactual-Inspired Context Selection and Compression for Tool-Using LLM Agents

DGX agent

arXiv:2606.08151v1 Announce Type: new Abstract: Tool-using LLM agents often fail not because relevant text is absent, but because decisive evidence is not selected, compressed, or surfaced at action t

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Declarative Skills for AI Agents in Knowledge-Grounded Tool-Use Workflows

DGX agent

arXiv:2606.06923v1 Announce Type: new Abstract: We study orchestration mechanisms for tool-using AI agents in realistic customer-service workflows over an unstructured knowledge base. We argue that de

model-releasesarxiv-cs-ai
8 Jun 2026
Research

GlossAssist -- A Tool to Simplify Corpus Creation and Study the Effect of NLP Models in Low-Resource Documentation Settings

DGX agent

arXiv:2606.04367v1 Announce Type: new Abstract: Interlinear glossed text (IGT) is the standard format for linguistic annotation in language documentation. Producing it manually, however, is often slow

researcharxiv-cs-cl
4 Jun 2026
Model Releases

How AI Fails: An Interactive Pedagogical Tool for Demonstrating Dialectal Bias in Automated Toxicity Models

DGX agent

arXiv:2511.06676v3 Announce Type: replace Abstract: Now that AI-driven moderation has become pervasive in everyday life, we often hear claims that 'the AI is biased'. While this is often said jokingly

model-releasesarxiv-cs-cl
2 Jun 2026
Agents

RadAgent: A tool-using AI agent for stepwise interpretation of chest computed tomography

DGX agent

arXiv:2604.15231v2 Announce Type: replace Abstract: Vision-language models (VLM) have markedly advanced AI-driven interpretation and reporting of complex medical imaging, such as computed tomography (

agentsarxiv-cs-ai
2 Jun 2026
← Previous
1…678910…108
Next →