AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “tools”

GridTimelineEvolution
5,189 results
Agents

MATRA: Modeling the Attack Surface of Agentic AI Systems -- OpenClaw Case Study

DGX agent

arXiv:2605.10763v1 Announce Type: new Abstract: LLMs are increasingly deployed as autonomous agents with access to tools, databases, and external services, yet practitioners (across different sectors)

agentsarxiv-cs-ai
12 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

PINS: Proximal Iterations with Sparse Newton and Sinkhorn for Optimal Transport

DGX agent

arXiv:2502.03749v2 Announce Type: replace Abstract: Optimal transport (OT) is a widely used tool in machine learning, but computing high-accuracy solutions for large instances remains costly. Entropic

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

PrepBench: How Far Are We from Natural-Language-Driven Data Preparation?

DGX agent

arXiv:2605.08687v1 Announce Type: cross Abstract: Data preparation is a central and time-consuming stage in data analysis workflows. Traditionally, commercial tools have relied on graphical user inter

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Quantitative Sobolev Approximation Bounds for Neural Operators with Empirical Validation on Burgers Equation

DGX agent

arXiv:2605.08170v1 Announce Type: new Abstract: Neural operators have emerged as a powerful tool for learning mappings between infinite-dimensional function spaces. However, their approximation proper

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Re^2Math: Benchmarking Theorem Retrieval in Research-Level Mathematics

DGX agent

arXiv:2605.09012v1 Announce Type: new Abstract: Large language models are increasingly capable at closed-world mathematical reasoning, but research assistance also requires source-grounded use of the

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

RewardHarness: Self-Evolving Agentic Post-Training

DGX agent

arXiv:2605.08703v1 Announce Type: new Abstract: Evaluating instruction-guided image edits requires rewards that reflect subtle human preferences, yet current reward models typically depend on large-sc

model-releasesarxiv-cs-ai
12 May 2026
Safety

Route by State, Recover from Trace: STAR with Failure-Aware Markov Routing for Multi-Agent Spatiotemporal Reasoning

DGX agent

arXiv:2605.10057v1 Announce Type: new Abstract: Compositional spatiotemporal reasoning often requires a system to invoke multiple heterogeneous specialists, such as geometric, temporal, topological, a

safetyarxiv-cs-ai
12 May 2026
Safety

RUBEN: Rule-Based Explanations for Retrieval-Augmented LLM Systems

DGX agent

arXiv:2605.10862v1 Announce Type: new Abstract: This paper demonstrates RUBEN, an interactive tool for discovering minimal rules to explain the outputs of retrieval-augmented large language models (LL

safetyarxiv-cs-cl
12 May 2026
Safety

Skill-R1: Agent Skill Evolution via Reinforcement Learning

DGX agent

arXiv:2605.09359v1 Announce Type: cross Abstract: Agentic large language models often rely on skills, reusable natural language procedures that guide planning, action, and tool use. In practice, skill

safetyarxiv-cs-ai
12 May 2026
Model Releases

The Last Word Often Wins: A Format Confound in Chain-of-Thought Corruption Studies

DGX agent

arXiv:2605.10799v1 Announce Type: cross Abstract: Corruption studies, the primary tool for evaluating chain-of-thought (CoT) faithfulness, identify which chain positions are 'computationally important

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Towards a Large Language-Vision Question Answering Model for MSTAR Automatic Target Recognition

DGX agent

arXiv:2605.10772v1 Announce Type: cross Abstract: Large language-vision models (LLVM), such as OpenAI's ChatGPT and GPT-4, have gained prominence as powerful tools for analyzing text and imagery. The

model-releasesarxiv-cs-ai
12 May 2026
Research

Trapping Attacker in Dilemma: Examining Internal Correlations and External Influences of Trigger for Defending GNN Backdoors

DGX agent

arXiv:2605.08278v1 Announce Type: cross Abstract: GNNs have become a standard tool for learning on relational data, yet they remain highly vulnerable to backdoor attacks. Prior defenses often depend o

researcharxiv-cs-ai
12 May 2026
Model Releases

WildClawBench: A Benchmark for Real-World, Long-Horizon Agent Evaluation

DGX agent

arXiv:2605.10912v1 Announce Type: new Abstract: Large language and vision-language models increasingly power agents that act on a user's behalf through command-line interface (CLI) harnesses. However,

model-releasesarxiv-cs-cl
12 May 2026
Research

AdaHOP: Fast and Accurate Low-Precision Training via Outlier-Pattern-Aware Rotation

DGX agent

arXiv:2604.02525v2 Announce Type: replace Abstract: Hadamard transforms have become a key tool for stabilizing low-precision training, but existing methods apply them uniformly across tensors and comp

researcharxiv-cs-lg
11 May 2026
Agents

Agentic AI and the Industrialization of Cyber Offense: Forecast, Consequences, and Defensive Priorities for Enterprises and the Mittelstand

DGX agent

arXiv:2605.06713v1 Announce Type: cross Abstract: Agentic AI systems can plan, call tools, inspect code, interact with web applications, and coordinate multi-step workflows. These same capabilities ch

agentsarxiv-cs-ai
11 May 2026
Safety

Bias and Uncertainty in LLM-as-a-Judge Estimation

DGX agent

arXiv:2605.06939v1 Announce Type: new Abstract: LLM-as-a-Judge evaluation has become a standard tool for assessing base model performance. However, characterizing performance via the naive estimator,

safetyarxiv-cs-lg
11 May 2026
Safety

Can David Beat Goliath? On Multi-Hop Reasoning with Resource-Constrained Agents

DGX agent

arXiv:2601.21699v2 Announce Type: replace Abstract: Multi-turn reasoning agents solve complex questions by decomposing them into intermediate retrieval or tool-use steps, for accumulating supporting e

safetyarxiv-cs-cl
11 May 2026
Safety

EvolveR: Self-Evolving LLM Agents through an Experience-Driven Lifecycle

DGX agent

arXiv:2510.16079v2 Announce Type: replace-cross Abstract: Current Large Language Model (LLM) agents show strong performance in tool use, but lack the crucial capability to systematically learn from th

safetyarxiv-cs-ai
11 May 2026
Agents

From Storage to Experience: A Survey on the Evolution of LLM Agent Memory Mechanisms

DGX agent

arXiv:2605.06716v1 Announce Type: new Abstract: Large Language Model (LLM)-based agents have fundamentally reshaped artificial intelligence by integrating external tools and planning capabilities. Whi

agentsarxiv-cs-ai
11 May 2026
Research

It Just Takes Two: Scaling Amortized Inference to Large Sets

DGX agent

arXiv:2605.07972v1 Announce Type: cross Abstract: Neural posterior estimation has emerged as a powerful tool for amortized inference, with growing adoption across scientific and applied domains. In ma

researcharxiv-cs-ai
11 May 2026
Model Releases

MIPIAD: Multilingual Indirect Prompt Injection Attack Defense with Qwen -- TF-IDF Hybrid and Meta-Ensemble Learning

DGX agent

arXiv:2605.07269v1 Announce Type: new Abstract: Indirect prompt injection remains a persistent weakness in retrieval-augmented and tool-using LLM systems, and the problem becomes harder to characteris

model-releasesarxiv-cs-cl
11 May 2026
Applications

Multi-Stage Prototype Learning for Interpretable Time Series Classification

DGX agent

arXiv:2106.09636v2 Announce Type: replace Abstract: Deep learning methods are powerful tools in classifying multivariate time series data. Despite their high performance, these methods are hard to int

applicationsarxiv-cs-lg
11 May 2026
Research

Pre-training Enables Extraordinary All-optical Image Denoising

DGX agent

arXiv:2605.07810v1 Announce Type: cross Abstract: Optical neural networks are emerging as powerful machine learning and information processing tools because of their potential advantages in speed and

researcharxiv-cs-cv
11 May 2026
Safety

Safactory: A Scalable Agentic Infrastructure for Training Trustworthy Autonomous Intelligence

DGX agent

arXiv:2605.06230v2 Announce Type: replace Abstract: As large models evolve from conversational assistants into autonomous agents, challenges increasingly arise from long-horizon decision making, tool

safetyarxiv-cs-ai
11 May 2026
Model Releases

ScrapeGraphAI-100k: Dataset for Schema-Constrained LLM Generation

DGX agent

arXiv:2602.15189v2 Announce Type: replace-cross Abstract: Producing output that conforms to a specified JSON schema underlies tool use, structured extraction, and knowledge base construction in modern

model-releasesarxiv-cs-ai
11 May 2026
Agents

Securing Computer-Use Agents: A Unified Architecture-Lifecycle Framework for Deployment-Grounded Reliability

DGX agent

arXiv:2605.07110v1 Announce Type: new Abstract: Computer-use agents(CUAs)are moving frombounded benchmarks toward real software environments, wherethey operate browsers, desktops, mobile applications,

agentsarxiv-cs-cl
11 May 2026
Safety

SimCT: Recovering Lost Supervision for Cross-Tokenizer On-Policy Distillation

DGX agent

arXiv:2605.07711v1 Announce Type: new Abstract: On-policy distillation (OPD) is a standard tool for transferring teacher behavior to a smaller student, but it implicitly assumes that teacher and stude

safetyarxiv-cs-cl
11 May 2026
Local Ai

SoftSAE: Dynamic Top-K Selection for Adaptive Sparse Autoencoders

DGX agent

arXiv:2605.06610v2 Announce Type: replace-cross Abstract: Sparse Autoencoders (SAEs) have become an important tool in mechanistic interpretability, helping to analyze internal representations in both

local-aiarxiv-cs-cv
11 May 2026
Research

Towards an Inferentialist Account of Information Through Proof-theoretic Semantics

DGX agent

arXiv:2605.05368v2 Announce Type: replace-cross Abstract: Information is one of the most widely-discussed concepts of the current era. However, a great deal of insightful work notwithstanding, it is y

researcharxiv-cs-ai
11 May 2026
Research

VNN-LIB 2.0: Rigorous Foundations for Neural Network Verification

DGX agent

arXiv:2605.07451v1 Announce Type: new Abstract: Neural network verification is an active and rapidly maturing research area, with a growing ecosystem of solvers and tools. The VNN-LIB standard was int

researcharxiv-cs-lg
11 May 2026
Agents

WebClipper: Efficient Evolution of Web Agents with Graph-based Trajectory Pruning

DGX agent

arXiv:2602.12852v2 Announce Type: replace Abstract: Deep Research systems based on web agents have shown strong potential in solving complex information-seeking tasks, yet their search efficiency rema

agentsarxiv-cs-ai
11 May 2026
Research

Active Contact Sensing for Robust Robot-to-Human Object Handover

DGX agent

arXiv:2605.04610v1 Announce Type: new Abstract: Robot-to-human object handover is an essential skill for robot assistants, from serving drinks at home to passing surgical tools in the operating room.

researcharxiv-cs-ro
7 May 2026
Safety

Agent-Based Modeling of Low-Emission Fertilizer Adoption for Dairy Farm Decarbonisation using Empirical Farm Data

DGX agent

arXiv:2605.03648v1 Announce Type: new Abstract: To understand complex system dynamics in dairy farming, it is essential to use modeling tools that capture farm heterogeneity, social interactions, and

safetyarxiv-cs-ai
7 May 2026
Local Ai

FaSTA^*: Fast-Slow Toolpath Agent with Subroutine Mining for Efficient Multi-turn Image Editing

DGX agent

arXiv:2506.20911v2 Announce Type: replace Abstract: We develop a cost-efficient neurosymbolic agent to address challenging multi-turn image editing tasks such as ``Detect the bench in the image while

local-aiarxiv-cs-cv
7 May 2026
Agents

GeoDecider: A Coarse-to-Fine Agentic Workflow for Explainable Lithology Classification

DGX agent

arXiv:2605.03383v1 Announce Type: new Abstract: Lithology classification aims to infer subsurface rock types from well-logging signals, supporting downstream applications like reservoir characterizati

agentsarxiv-cs-ai
7 May 2026
Model Releases

MEMTIER: Tiered Memory Architecture and Retrieval Bottleneck Analysis for Long-Running Autonomous AI Agents

DGX agent

arXiv:2605.03675v1 Announce Type: new Abstract: Long-running autonomous AI agents suffer from a well-documented memory coherence problem: tool-execution success rates degrade 14 percentage points over

model-releasesarxiv-cs-ai
7 May 2026
Hardware

Quadrature-TreeSHAP: Depth-Independent TreeSHAP and Shapley Interactions

DGX agent

arXiv:2605.04497v1 Announce Type: new Abstract: Shapley values are a standard tool for explaining predictions of tree ensembles, with Path-Dependent SHAP being the most widely used variant. Despite su

hardwarearxiv-cs-lg
7 May 2026
Safety

Time series causal discovery with variable lags

DGX agent

arXiv:2605.04081v1 Announce Type: new Abstract: Causal Bayesian Networks (CBNs) are a powerful tool for reasoning under uncertainty about complex real-world problems. Such problems evolve over time, r

safetyarxiv-cs-lg
7 May 2026
Hardware

When Agents Handle Secrets: A Survey of Confidential Computing for Agentic AI

DGX agent

arXiv:2605.03213v1 Announce Type: cross Abstract: Agentic AI systems, specifically LLM-driven agents that plan, invoke tools, maintain persistent memory, and delegate tasks to peer agents via protocol

hardwarearxiv-cs-ai
7 May 2026
Research

A Sentence Relation-Based Approach to Sanitizing Malicious Instructions

DGX agent

arXiv:2605.01078v1 Announce Type: cross Abstract: Retrieval-augmented generation and tool-integrated LLM agents increasingly depend on external textual sources. This reliance broadens the available at

researcharxiv-cs-ai
6 May 2026
Research

Adaptive Estimation and Optimal Control in Offline Contextual MDPs without Stationarity

DGX agent

arXiv:2605.03393v1 Announce Type: cross Abstract: Contextual MDPs are powerful tools with wide applicability in areas from biostatistics to machine learning. However, specializing them to offline data

researcharxiv-cs-lg
6 May 2026
Research

AMSnet-q: Unsupervised Circuit Identification and Performance Labeling for AMS Circuits

DGX agent

arXiv:2605.01404v1 Announce Type: cross Abstract: Analog and mixed-signal (AMS) circuit design remains heavily reliant on expert knowledge. While recent AI-driven automation tools can generate candida

researcharxiv-cs-ai
6 May 2026
Safety

Architectural Obsolescence of Unhardened Agentic-AI Runtimes

DGX agent

arXiv:2605.01740v1 Announce Type: cross Abstract: An agentic-AI runtime issues tool calls, sends messages, and actuates devices on behalf of an LLM. Catching the four ways an action can diverge from i

safetyarxiv-cs-ai
6 May 2026
Applications

Free Decompression with Algebraic Spectral Curves

DGX agent

arXiv:2605.03634v1 Announce Type: cross Abstract: Tools from random matrix theory have become central to deep learning theory, using spectral information to provide mechanisms for modeling generalizat

applicationsarxiv-cs-lg
6 May 2026
Agents

HeavySkill: Heavy Thinking as the Inner Skill in Agentic Harness

DGX agent

arXiv:2605.02396v1 Announce Type: new Abstract: Recent advances in agentic harness with orchestration frameworks that coordinate multiple agents with memory, skills, and tool use have achieved remarka

agentsarxiv-cs-ai
6 May 2026
Safety

Model Routing as a Trust Problem: Route Receipts for Adaptive AI Systems

DGX agent

arXiv:2605.01710v1 Announce Type: new Abstract: AI products often route requests through version aliases, service tiers, tool choices, regional endpoints, fallback rules, or safety handling before res

safetyarxiv-cs-ai
6 May 2026
Tutorials

Neural Decision-Propagation for Answer Set Programming

DGX agent

arXiv:2605.01797v1 Announce Type: new Abstract: Integration of Answer Set Programming (ASP) with neural networks has emerged as a promising tool in Neuro-symbolic AI. While existing approaches extend

tutorialsarxiv-cs-ai
6 May 2026
Model Releases

The Oracle's Fingerprint: Correlated AI Forecasting Errors and the Limits of Bias Transmission

DGX agent

arXiv:2605.00844v1 Announce Type: cross Abstract: When large language models (LLMs) are consulted as forecasting tools, the independence of individual errors -- the foundation of collective intelligen

model-releasesarxiv-cs-ai
6 May 2026
← Previous
1…3637383940…109
Next →