AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “tools”

GridTimelineEvolution
5,141 results
Model Releases

LayerRAG-Bench: A Cross-Layer Reliability Benchmark for Agentic Retrieval-Augmented Generation

DGX agent

arXiv:2607.27353v1 Announce Type: new Abstract: Agentic retrieval-augmented generation systems can produce answers that appear grounded while failing at the evidence, tool-contract, authorization, or

model-releasesarxiv-cs-cl
31 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Agents

LEDGERMIND: Provenance-Constrained Multimodal Agentic Reasoning with a Structured Evidence Ledger

DGX agent

arXiv:2607.28374v1 Announce Type: new Abstract: Multimodal agents for visual question answering increasingly operate as multi-step trajectories that interleave perception, retrieval, and reasoning, ye

agentsarxiv-cs-lg
31 Jul 2026
Agents

Partner Capability Estimation for Task-Agnostic Adaptation in Ad-Hoc Teamwork

DGX agent

arXiv:2607.27177v1 Announce Type: new Abstract: Effective collaboration with novel and diverse partners is a crucial skill for autonomous agents. Most current ad-hoc teamwork (AHT) approaches assume t

agentsarxiv-cs-ai
31 Jul 2026
Research

AIGen: Automating AI Bill of Materials Generation Through Hybrid MLOps Integration

DGX agent

arXiv:2607.26652v1 Announce Type: new Abstract: The responsible development and deployment of artificial intelligence (AI) systems requires rigorous documentation of their constituent artifacts, e.g.,

researcharxiv-cs-lg
30 Jul 2026
Local Ai

A Control System, a Dataset, and a Recipe for Making Frozen LLM Agents Learn a Domain

DGX agent

arXiv:2607.25415v1 Announce Type: new Abstract: Production LLM agents are increasingly assembled from a frozen model wrapped in a harness: a prompt template, a tool set, a memory/retrieval layer, a pl

local-aiarxiv-cs-ai
29 Jul 2026
Research

Semantic Space Search Trajectory Networks

DGX agent

arXiv:2607.25122v1 Announce Type: new Abstract: Search Trajectory Networks (STNs) are a graph-based tool for visualizing and characterizing the behavior of optimization algorithms. STNs' reliance on d

researcharxiv-cs-lg
29 Jul 2026
Model Releases

Decentralized Granular Access Control for Agentic AI Systems in Critical Infrastructure

DGX agent

arXiv:2607.22611v1 Announce Type: new Abstract: The deployment of autonomous AI agents in production infrastructure introduces fundamental security challenges that traditional role-based access contro

model-releasesarxiv-cs-ai
28 Jul 2026
Research

Fashion-3DLR: A Controllable 3D Garment Generation Using Pairwise Fashion Elements for Intelligent Design

DGX agent

arXiv:2607.23189v1 Announce Type: cross Abstract: AI-generated content (AIGC) has made significant progress, with 2D generative models becoming ready-to-use tools for the digital fashion industry. How

researcharxiv-cs-ai
28 Jul 2026
Safety

RMS@CC-MMD 2026: Multimodal Misogyny Detection via Geometric Interaction and Multi-View Consensus

DGX agent

arXiv:2607.22709v1 Announce Type: cross Abstract: The proliferation of internet memes has introduced new complexities to automated content moderation, particularly in detecting misogyny. Memes often r

safetyarxiv-cs-ai
28 Jul 2026
Model Releases

SAGE: Safety-First Defense-in-Depth Guardrails for Verified Lifecycle Control of High-Impact Generative AI

DGX agent

arXiv:2607.22926v1 Announce Type: new Abstract: High-impact generative AI makes catastrophic misuse a lifecycle-control problem, not merely a prompt-filtering problem. SAGE is a safety-first, authoriz

model-releasesarxiv-cs-ai
28 Jul 2026
Agents

SetGo: Metadata Readiness for Scientific AI Datasets

DGX agent

arXiv:2607.22677v1 Announce Type: cross Abstract: Scientific datasets intended for AI use require both computational readiness for model training and metadata readiness for discovery, sharing, and reu

agentsarxiv-cs-ai
28 Jul 2026
Agents

SpecBox: Speculative Sandbox Scheduling for Efficient LLM Agent Serving

DGX agent

arXiv:2607.23933v1 Announce Type: cross Abstract: As LLM agents increasingly rely on the Model Context Protocol (MCP) to invoke isolated external sandboxes, disaggregated sandbox deployment introduces

agentsarxiv-cs-ai
28 Jul 2026
Research

Computer Vision Based Neurology Brain Activity Rejection Architecture and Implementation

DGX agent

arXiv:2607.21654v1 Announce Type: cross Abstract: The electroencephalogram (EEG) is a valuable and widely applied tool for investigating brain disorders and behavioral changes. It offers a minimally r

researcharxiv-cs-lg
27 Jul 2026
Model Releases

AISE-Bench: A Full-Cycle Curated Benchmark for Information Seeking on Academic Knowledge Graphs

DGX agent

arXiv:2607.20498v1 Announce Type: new Abstract: Large language models (LLMs) augmented with tools are emerging as autonomous agents capable of using Web engine, APIs, and code to solve complex, long-h

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

ArbiGraph: Arbitrarily Scalable Verifiable Task Graphs for Evaluating Context Management

DGX agent

arXiv:2607.20764v1 Announce Type: new Abstract: We introduce ARBIGRAPH, a benchmark generator for evaluating whether tool-assisted language agents can retain, update, compose, and discard task-relevan

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

Autonomous Topology Mutation: Safe Runtime Restructuring for Multi-Agent LLM Systems with Capability, State, and Shadow Invariants

DGX agent

arXiv:2607.20488v1 Announce Type: new Abstract: Multi-agent LLM frameworks typically fix their team topology at boot time. When an individual agent becomes overloaded at runtime, for example by mixing

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

OpenForgeRL: Train Harness-native Agents in Any Environment

DGX agent

arXiv:2607.21557v1 Announce Type: new Abstract: Modern AI agents rely on elaborate inference harnesses such as Claude Code, Codex, and OpenClaw to drive multi-turn reasoning, tool use, and access to e

model-releasesarxiv-cs-ai
24 Jul 2026
Agents

Silent Failures in Multimodal Agentic Search:A Diagnostic Taxonomy and Cross-Judge Evaluation

DGX agent

arXiv:2607.19793v1 Announce Type: new Abstract: Multimodal agentic search systems increasingly rely on external tools to answer knowledge-intensive visual questions. However, existing evaluations main

agentsarxiv-cs-ai
23 Jul 2026
Model Releases

Evaluating Vision Foundation Models for Pixel and Object Classification in Microscopy

DGX agent

arXiv:2603.19802v2 Announce Type: replace Abstract: Deep learning underlies most modern approaches and tools in computer vision, including biomedical imaging. However, for interactive semantic segment

model-releasesarxiv-cs-cv
16 Jul 2026
Research

Lag Operator SSMs: A Geometric Framework for Structured State Space Modeling

DGX agent

arXiv:2512.18965v2 Announce Type: replace Abstract: Structured State Space Models (SSMs), which are at the heart of the recently popular Mamba architecture, are powerful tools for sequence modeling. H

researcharxiv-cs-lg
16 Jul 2026
Safety

Operational Evidence Gaps for LLMs in Fraud Detection and Trust-and-Safety Workflows

DGX agent

arXiv:2607.13078v1 Announce Type: cross Abstract: LLMs are now proposed for fraud detection, scam investigation, content moderation, and other trust-and-safety workflows. Much of the public literature

safetyarxiv-cs-ai
16 Jul 2026
Safety

AAAI-26 Dual Submissions: Novel Challenges

DGX agent

arXiv:2607.11918v1 Announce Type: cross Abstract: Dual submissions, in which identical or substantially similar papers are simultaneously submitted to one or more archival venues, without cross-citati

safetyarxiv-cs-ai
15 Jul 2026
Safety

Function-Aware Fill-in-the-Middle as Mid-Training for Coding Agent Foundation Models

DGX agent

arXiv:2607.12463v1 Announce Type: new Abstract: Coding agents must integrate external tool returns into ongoing reasoning - a capability that standard left-to-right pretraining on code exposes only in

safetyarxiv-cs-ai
15 Jul 2026
Agents

IdeaTrail: Full-Process Agent Trajectories for Scientific Ideation

DGX agent

arXiv:2607.10144v2 Announce Type: replace Abstract: Scientific ideation unfolds over multiple stages, including literature search, paper reading, tool use, claim checking, cross-paper synthesis, brain

agentsarxiv-cs-ai
15 Jul 2026
Model Releases

TerraLogic: A Benchmark for Hierarchical Geospatial Reasoning in Earth Observation

DGX agent

arXiv:2607.12497v1 Announce Type: new Abstract: Beyond perception, reasoning is essential in remote sensing for advanced interpretation, inference, and decision-making. Recent advances in large langua

model-releasesarxiv-cs-cv
15 Jul 2026
Research

The Model Knows Your Project, Not You: Measuring Recognition in LLMs with NameRank

DGX agent

arXiv:2607.12520v1 Announce Type: new Abstract: What a frontier model recalls about a person or tool from its own weights -- before any retrieval step -- often shapes the first description a human see

researcharxiv-cs-ai
15 Jul 2026
Research

Data Alchemy: Mitigating Cross-Site Model Variability Through Test Time Data Calibration

DGX agent

arXiv:2407.13632v2 Announce Type: replace Abstract: Deploying deep learning-based imaging tools across various clinical sites poses significant challenges due to inherent domain shifts and regulatory

researcharxiv-cs-cv
10 Jul 2026
Agents

DeepSearch-World: Self-Distillation for Deep Search Agents in a Verifiable Environment

DGX agent

arXiv:2607.07820v1 Announce Type: new Abstract: Training tool-use agents to improve from their own experience remains challenging, as supervised fine-tuning relies on fixed teacher-distilled trajector

agentsarxiv-cs-cl
10 Jul 2026
Local Ai

Segmenting Low-Contrast XCTs of Concrete: An Unsupervised Approach

DGX agent

arXiv:2603.00127v2 Announce Type: replace Abstract: X-Ray Computed Tomography (XCT) is a compelling tool in experimental mechanics, capable of non-destructively extracting information pertaining to th

local-aiarxiv-cs-cv
9 Jul 2026
Model Releases

The Harness Effect: How Orchestration Design Sets the Token Economics of Enterprise Agentic AI

DGX agent

arXiv:2607.06906v1 Announce Type: new Abstract: Agentic AI development today runs on token maxing: buying capability with tokens -- longer reasoning traces, more turns, wider tool payloads, bigger rep

model-releasesarxiv-cs-ai
9 Jul 2026
Safety

When Agents Remember Too Much: Memory Poisoning Attacks on Large Language Model Agents

DGX agent

arXiv:2607.06595v1 Announce Type: cross Abstract: Personal AI agents powered by large language models can reason and act using available tools to access emails, manage calendars, and push code to remo

safetyarxiv-cs-ai
9 Jul 2026
Safety

From Passive Retrieval to Active Memory Navigation: Learning to Use Memory as a Structured Action Space

DGX agent

arXiv:2607.05794v1 Announce Type: new Abstract: Long-term user memory is essential for personalized conversational agents, yet many memory systems still expose memory through passive retrieval interfa

safetyarxiv-cs-ai
8 Jul 2026
Safety

Lingering Authority: Revocable Resource-and-Effect Capabilities for Coding Agents

DGX agent

arXiv:2606.22504v1 Announce Type: cross Abstract: Coding agents often receive broad tool access for an entire task, even when a resource is needed only for one subgoal. We call this gap lingering auth

safetyarxiv-cs-ai
8 Jul 2026
Model Releases

Scientific Code Search at Scale: A Multi-Domain Dataset and Benchmark

DGX agent

arXiv:2607.05443v1 Announce Type: cross Abstract: Scientists increasingly rely on open-source tools to support their research workflows, yet discovering relevant software among over 600 million GitHub

model-releasesarxiv-cs-ai
8 Jul 2026
Model Releases

AgentGym2: Benchmarking Large Language Model Agents in De-Idealized Real-World Environments

DGX agent

arXiv:2607.05174v1 Announce Type: new Abstract: Language agents, i.e., LLM agents, progress rapidly and are increasingly deployed in production environments. This trend underscores the urgent need for

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

Forethought: Verifiable Reasoning from Neurosymbolic Primitive Programming

DGX agent

arXiv:2607.04096v1 Announce Type: new Abstract: Current agentic workflows usually involve decomposing user requests into sequences of tool calls with correctly resolved parameters, the results of whic

model-releasesarxiv-cs-ai
7 Jul 2026
Applications

P^3: Toward Versatile Embodied Agents

DGX agent

arXiv:2508.07033v2 Announce Type: replace Abstract: Embodied agents have shown promising generalization capabilities across diverse physical environments, making them essential for a wide range of rea

applicationsarxiv-cs-ro
7 Jul 2026
Agents

SelfMem: Self-Optimizing Memory for AI Agents

DGX agent

arXiv:2607.03726v1 Announce Type: new Abstract: While current AI agents support increasingly long context windows, tool use, and skill execution for long-horizon tasks, they still require memory syste

agentsarxiv-cs-cl
7 Jul 2026
Model Releases

SovereignPA-Bench: Evaluating User-Owned Personal Agents under Evolving Intent, Platform Mediation, and Consent Constraints

DGX agent

arXiv:2607.05363v1 Announce Type: new Abstract: Personal agents are becoming persistent user-owned intermediaries: they remember preferences, filter platform-mediated information, use tools, and negot

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

SpecEyes: Accelerating Agentic Multimodal LLMs via Speculative Perception and Planning

DGX agent

arXiv:2603.23483v2 Announce Type: replace-cross Abstract: Agentic multimodal large language models (MLLMs) (e.g., OpenAI o3 and Gemini Agentic Vision) achieve remarkable reasoning capabilities through

model-releasesarxiv-cs-cl
7 Jul 2026
Safety

Teaming Up with AI: Coordination and Cooperation

DGX agent

arXiv:2607.03181v1 Announce Type: cross Abstract: Successful diffusion of AI in the workforce hinges on the economic value that AI brings to human endeavors. Bringing AI into the workforce is more tha

safetyarxiv-cs-ai
7 Jul 2026
Model Releases

Adoption and Impact of Command-Line AI Coding Agents: A Study of Microsoft's Early 2026 Rollout of Claude Code and GitHub Copilot CLI

DGX agent

arXiv:2607.01418v1 Announce Type: cross Abstract: Organizations rolling out agentic command line tools like Anthropic's Claude Code and GitHub's Copilot CLI need to know who will try them, who will ke

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Bringing Agentic Search to Earth Observation Data Discovery

DGX agent

arXiv:2607.02387v1 Announce Type: cross Abstract: NASA and its data centers hold thousands of geoscience datasets and tools like Worldview, Giovanni, the Science Discovery Engine, and Harmony. Finding

model-releasesarxiv-cs-lg
3 Jul 2026
Local Ai

Self-explainable Operator Learning for Discovering Spatial Patterns in Functional Data

DGX agent

arXiv:2607.02203v1 Announce Type: new Abstract: Operator learning has emerged as a powerful tool for modeling complex physical systems in functional spaces. However, their neural network-based archite

local-aiarxiv-cs-lg
3 Jul 2026
Safety

OpenReward: Learning to Reward Long-form Agentic Tasks via Reinforcement Learning

DGX agent

arXiv:2510.24636v3 Announce Type: replace Abstract: Reward models (RMs) have become essential for aligning large language models (LLMs), serving as scalable proxies for human evaluation in both traini

safetyarxiv-cs-cl
2 Jul 2026
Model Releases

Retrieved Images as Visual Thought: Training-Free Multimodal In-Context Learning for the Open-vs-Closed Gap

DGX agent

arXiv:2607.00606v1 Announce Type: new Abstract: Recent work on Thinking with Images makes vision a dynamic part of reasoning, but does so through generation: the model invokes external tools, synthesi

model-releasesarxiv-cs-cv
2 Jul 2026
Agents

Agentic AI Enhances Physician Trust in Clinical Decision Making

DGX agent

arXiv:2606.30658v1 Announce Type: cross Abstract: Medical AI has shifted from reasoning to agentic AI, a new paradigm that autonomously invokes external tools during reasoning, rendering intermediate

agentsarxiv-cs-ai
1 Jul 2026
Hardware

AgRefactor: Self-Evolving Agentic Workflow for HLS Compatibility and Performance

DGX agent

arXiv:2606.30949v1 Announce Type: new Abstract: High-Level Synthesis (HLS) provides a fast path from concepts to silicon, but converting real-world software into synthesizable HLS code remains challen

hardwarearxiv-cs-ai
1 Jul 2026
← Previous
1…1617181920…108
Next →