AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “tools”

GridTimelineEvolution
5,189 results
Model Releases

Student-Centered Distillation Narrows the Agentic Gap Between Small and Large LLMs

DGX agent

arXiv:2509.14257v3 Announce Type: replace-cross Abstract: Large Language Model agents achieve strong performance on multi-step reasoning and tool-use tasks, but their impressive capabilities typically

model-releasesarxiv-cs-ai
24 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Telco-GAIA: Bilingual Benchmark for Agents in Telecom Domain

DGX agent

arXiv:2607.20510v1 Announce Type: new Abstract: We introduce Telco-GAIA, a bilingual, multi-modal benchmark for evaluating tool-using agents on the data of a real-world telecommunications operator. Te

model-releasesarxiv-cs-ai
24 Jul 2026
Safety

Toward cryptographically verifiable authorization for autonomous AI agents: A security hypothesis, preliminary formal model, and proof-of-concept implementation

DGX agent

arXiv:2607.21325v1 Announce Type: cross Abstract: Autonomous AI agents increasingly execute actions, invoke tools, and operate on protected resources with limited human oversight. Existing authenticat

safetyarxiv-cs-ai
24 Jul 2026
Safety

Understanding Critical Thinking in Generative Artificial Intelligence Use: Development, Validation, and Correlates of the Critical Thinking in AI Use Scale

DGX agent

arXiv:2512.12413v2 Announce Type: replace Abstract: Generative AI tools are increasingly embedded in everyday work and learning, yet their fluency, opacity, and propensity to hallucinate mean that use

safetyarxiv-cs-ai
24 Jul 2026
Model Releases

Unlearning Under Imbalance: Benchmarking Fairness in Multimodal LLM Unlearning

DGX agent

arXiv:2607.21300v1 Announce Type: cross Abstract: Machine unlearning has emerged as a tool for removing personal data from trained models to comply with recent AI regulations. To evaluate unlearning e

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

A Multiclass Quantum Aligned Centroid Kernel

DGX agent

arXiv:2607.19782v1 Announce Type: cross Abstract: Kernel methods are powerful tools in machine learning but commonly used full-Gram kernels face three key limitations: (1) quadratic scaling with train

model-releasesarxiv-cs-lg
23 Jul 2026
Model Releases

ChannelGuard: Safe Models Do Not Compose into Safe Multi-Agent Systems

DGX agent

arXiv:2607.19430v1 Announce Type: cross Abstract: Multi-agent LLM applications chain a planner, worker agents, a verifier, and a synthesizer, and every hop between agents is an unmonitored channel thr

model-releasesarxiv-cs-ai
23 Jul 2026
Safety

JANUS: Foreseeing Latent Risk for Long-Horizon Agent Safety

DGX agent

arXiv:2607.19913v1 Announce Type: new Abstract: Agent safety is moving from content moderation toward preventing operational failures before tool-using agents act. We propose Janus, a foresight-orient

safetyarxiv-cs-ai
23 Jul 2026
Model Releases

MissingBench-Verified: Probing Vision-Language Models' Inability to Detect Missing Object Parts

DGX agent

arXiv:2607.18673v2 Announce Type: new Abstract: Vision Language Models (VLMs) are well known for hallucinating non-existent objects in images. Objects with missing parts present a unique challenge for

model-releasesarxiv-cs-cv
23 Jul 2026
Model Releases

Multi-modal transformer for signal classification in nanopore blockade experiments

DGX agent

arXiv:2607.20323v1 Announce Type: new Abstract: Nanopore devices have emerged as powerful tools for single-molecule sensing, with potential for rapid, portable diagnostics. They detect changes in ioni

model-releasesarxiv-cs-lg
23 Jul 2026
Safety

REGEN: Replay-recycling for Expert-to-Generalist distillation with Offline Reinforcement Learning

DGX agent

arXiv:2607.19450v1 Announce Type: cross Abstract: Large-scale online reinforcement learning (RL) is the predominant means of eliciting advanced abilities including long-term reasoning and agentic tool

safetyarxiv-cs-ai
23 Jul 2026
Research

SIINR: Structurally Informed Implicit Neural Representations for super-resolution with uncertainty quantification of clinical quality diffusion MRI datasets

DGX agent

arXiv:2607.19943v1 Announce Type: new Abstract: Diffusion Magnetic Resonance Imaging (dMRI) is a powerful tool for probing brain microstructure, but clinical acquisitions are often limited by low out-

researcharxiv-cs-cv
23 Jul 2026
Local Ai

Active Learning for Efficient Annotation of Surgical Videos with Weak Supervision

DGX agent

arXiv:2607.13237v1 Announce Type: new Abstract: Precise spatial-temporal annotation of laparoscopic videos is time-consuming and requires expert knowledge. We propose a human-in-the-loop knowledge acq

local-aiarxiv-cs-cv
16 Jul 2026
Safety

AI-Native Insurance for Agentic AI: Pricing, Underwriting, and End-to-End Automation

DGX agent

arXiv:2607.13230v1 Announce Type: new Abstract: Agentic AI introduces new insurance challenges because autonomous AI systems can make decisions, invoke tools, modify external environments, and interac

safetyarxiv-cs-ai
16 Jul 2026
Research

An Efficient Newton Algorithm for Nonnegative Matrix Factorization with the Kullback-Leibler Divergence

DGX agent

arXiv:2607.13919v1 Announce Type: new Abstract: Nonnegative Matrix Factorization (NMF) is a fundamental tool in unsupervised learning, which approximates a nonnegative matrix by the product of two low

researcharxiv-cs-lg
16 Jul 2026
Model Releases

CAVA: Canonical Action Verification and Attestation for Runtime Governance of Agentic AI Systems

DGX agent

arXiv:2607.13716v1 Announce Type: new Abstract: Agentic AI systems increasingly act through heterogeneous runtimes: local coding hooks, SDK tools, browser automation, managed-agent traces, API gateway

model-releasesarxiv-cs-ai
16 Jul 2026
Local Ai

Differentiable Polarized Path Tracing

DGX agent

arXiv:2607.13265v1 Announce Type: new Abstract: Physically based differentiable rendering has proven to be a powerful tool for inverse rendering problems (e.g., 3D reconstruction, reflectance estimati

local-aiarxiv-cs-cv
16 Jul 2026
Model Releases

Discriminative Barrier Functions for Safe Adversarial Imitation Learning from Observation

DGX agent

arXiv:2607.13938v1 Announce Type: new Abstract: Inverse Reinforcement Learning (IRL) algorithms are powerful tools for learning from and generalizing expert demonstrations, but they often rely on unco

model-releasesarxiv-cs-ro
16 Jul 2026
Research

OriginBlame: Record- and Token-Level Data Provenance for AI Training Datasets

DGX agent

arXiv:2607.13037v1 Announce Type: new Abstract: When a data contributor requests removal, model trainers face a practical gap: unlearning algorithms require a forget set, yet no tool can locate which

researcharxiv-cs-ai
16 Jul 2026
Safety

SAFETY SENTRY: Context-Aware Human Intervention via EXECUTE-ASK-REFUSE Routing

DGX agent

arXiv:2607.13594v1 Announce Type: new Abstract: LLM agents act on real-world environments through tool calls, and a single misjudged action can cause irreversible harm. The standard safeguard is a gua

safetyarxiv-cs-ai
16 Jul 2026
Agents

Self-Improvements in Modern Agentic Systems: A Survey

DGX agent

arXiv:2607.13104v1 Announce Type: new Abstract: Self-improving autonomous agents are moving from research prototypes to deployed systems. The primary goal is controllable evolution, or adaptation, fro

agentsarxiv-cs-ai
16 Jul 2026
Model Releases

SingGuard-NSFA: Extensible Guardrails for Agentic AI via Generative Reasoning and Real-Time Classification

DGX agent

arXiv:2607.13081v1 Announce Type: cross Abstract: We present nsfaguard, a guardrail framework for securing agentic AI systems against operational threats, such as prompt injection, sensitive informati

model-releasesarxiv-cs-ai
16 Jul 2026
Model Releases

STOCKTAKE: Measuring the Gap Between Perception and Action in LLM Agents with a Fair Oracle

DGX agent

arXiv:2607.13618v1 Announce Type: new Abstract: LLM agents are increasingly evaluated on multi-week decision tasks in which the state that drives cost is never directly observed. On such tasks the fin

model-releasesarxiv-cs-ai
16 Jul 2026
Agents

When Bots Join the Team: Bot Adoption and the Institutional Fabric of Open-Source Software Projects

DGX agent

arXiv:2607.13679v1 Announce Type: new Abstract: AI agents are joining human teams, raising a basic question: when an automated agent becomes a regular participant, does group organization strengthen o

agentsarxiv-cs-ai
16 Jul 2026
Safety

Git-Assistant: Planning-Based Support for Updating Git Repositories

DGX agent

arXiv:2607.09224v2 Announce Type: replace-cross Abstract: Version control systems are essential for collaborative software development, yet tools like git remain challenging for many practitioners. Re

safetyarxiv-cs-ai
15 Jul 2026
Safety

Isolation as a First-Class Principle for LLM-Agent System Safety: Concepts, Taxonomy, Challenges and Future Directions

DGX agent

arXiv:2607.12406v1 Announce Type: new Abstract: The capability of LLM agents to function as the ``brain'' of a system fundamentally expands the scope of analysis beyond a standalone model. Consequentl

safetyarxiv-cs-ai
15 Jul 2026
Model Releases

RCWT: Measuring Task-Budget Displacement from Coordination Content in LLM Calls

DGX agent

arXiv:2607.12216v1 Announce Type: cross Abstract: Multi-agent and memory-augmented LLM systems often place coordination content, shared state, prior discussion, tool outputs, summaries, and role instr

model-releasesarxiv-cs-ai
15 Jul 2026
Agents

The Emerging Paradigm of Geospatial Foundation Models: From Pre-Training to Agentic Reasoning

DGX agent

arXiv:2607.12177v1 Announce Type: new Abstract: The analysis of satellite and aerial imagery has entered a new era with the advent of foundation models. This paper describes the concept of Geospatial

agentsarxiv-cs-ai
15 Jul 2026
Model Releases

Token Reduction Is Not Cost Reduction

DGX agent

arXiv:2607.12161v1 Announce Type: new Abstract: Context-reduction layers for API-based coding agents, including command-output compressors, retrieval rankers, and payload-optimizing proxies, are usual

model-releasesarxiv-cs-cl
15 Jul 2026
Model Releases

Track, Rank, Crack: Epistemic Working Memory Scales Multi-Hop Reasoning in Language Agents

DGX agent

arXiv:2607.12267v1 Announce Type: cross Abstract: Language agents that interleave reasoning and tool use degrade sharply as reasoning chains lengthen, even when each individual step is easy. We trace

model-releasesarxiv-cs-ai
15 Jul 2026
Research

AI-integrated models for assessing agricultural resilience

DGX agent

arXiv:2607.07759v1 Announce Type: new Abstract: Agricultural supply chains are vulnerable to disruptions through linked biophysical and economic systems. We develop an AI-powered tool that integrates

researcharxiv-cs-ai
10 Jul 2026
Model Releases

Aligning Clinical Needs and AI Capabilities: A Survey on LLMs for Medical Reasoning

DGX agent

arXiv:2607.07761v1 Announce Type: new Abstract: Large language models (LLMs) have emerged as important tools in healthcare, showing growing potential for clinical reasoning and patient care. This surv

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

HumanForge: A Human-Centric Deepfake Video Benchmark with Multi-Agent Forgery Rationales

DGX agent

arXiv:2607.08705v1 Announce Type: new Abstract: Rapid advancements in video diffusion models and temporal editing tools have enabled the generation of highly realistic human-centric videos, posing unp

model-releasesarxiv-cs-cv
10 Jul 2026
Agents

Multi-Agent Firewall Architecture for Privacy Protection of Sensitive Data in Interactions with Language Models

DGX agent

arXiv:2607.08282v1 Announce Type: cross Abstract: While Large Language Models (LLMs) have become essential productivity tools, their integration into workflows without adequate safeguards creates sign

agentsarxiv-cs-ai
10 Jul 2026
Research

Overthinking: Amplifying Reasoning Weights to Extract Learned Secrets

DGX agent

arXiv:2607.08173v1 Announce Type: new Abstract: Black box auditing of language models is an essential pre-deployment tool, but it may miss subtle forms of misalignment and hidden information. To bette

researcharxiv-cs-ai
10 Jul 2026
Safety

Persona Cartography: Charting Language Model Personality Traits in Weight Space

DGX agent

arXiv:2607.07916v1 Announce Type: new Abstract: Large language models exhibit recurring behavioural patterns -- personas -- that shape generalisation and safety, but we lack reliable tools for decompo

safetyarxiv-cs-ai
10 Jul 2026
Research

Stop Guessing When to Stop Testing: Efficient Model Evaluation with Just Enough Data

DGX agent

arXiv:2607.08522v1 Announce Type: new Abstract: The inherent rigidity of fixed-size benchmarks makes them an inefficient tool for model evaluation. Diverse evaluation objectives, including model ranki

researcharxiv-cs-lg
10 Jul 2026
Agents

TTHE: Test-Time Harness Evolution

DGX agent

arXiv:2607.08124v1 Announce Type: cross Abstract: The behavior of an LLM agent is determined not only by the underlying model, but also by its harness: the executable program that constructs context,

agentsarxiv-cs-lg
10 Jul 2026
Safety

Workflow as Knowledge: Semantic Persistence for LLM-Mediated Workflows

DGX agent

arXiv:2607.08740v1 Announce Type: new Abstract: Large language model (LLM) applications increasingly use explicit workflows for tool use, retrieval, branching, checkpointing, and human approval. Exist

safetyarxiv-cs-ai
10 Jul 2026
Research

Amortized Inference for Correlated Discrete Choice Models via Equivariant Neural Networks

DGX agent

arXiv:2603.24705v3 Announce Type: replace-cross Abstract: Discrete choice models are fundamental tools in management science, economics, and marketing for understanding and predicting decision-making.

researcharxiv-cs-lg
9 Jul 2026
Tutorials

Creating Power Distribution Network Layouts Using Generative Adversarial Networks and Image-Based Representations

DGX agent

arXiv:2607.06622v1 Announce Type: cross Abstract: Utilities increasingly rely on planning and operational tools to cope with the increased penetrations of distributed energy resources, yet the lack of

tutorialsarxiv-cs-lg
9 Jul 2026
Model Releases

Explain Before You Answer: A Survey on Compositional Visual Reasoning

DGX agent

arXiv:2508.17298v3 Announce Type: replace-cross Abstract: Compositional visual reasoning has emerged as a key research frontier in multimodal AI, aiming to endow machines with the human-like ability t

model-releasesarxiv-cs-ai
9 Jul 2026
Safety

Mathematical methods of reinforcement learning

DGX agent

arXiv:2607.06935v1 Announce Type: cross Abstract: Reinforcement learning (RL) is increasingly grounded in tools from probability, optimization, and operator theory. This survey organizes the mathemati

safetyarxiv-cs-lg
9 Jul 2026
Model Releases

MIRA-Math: A Benchmark for Minimal Information Requesting and Mathematical Reasoning

DGX agent

arXiv:2607.07391v1 Announce Type: new Abstract: Mathematical reasoning benchmarks typically provide all facts needed to solve each problem, while interactive benchmarks often mix reasoning with tools,

model-releasesarxiv-cs-ai
9 Jul 2026
Research

SPEAR: A Simulator for Photorealistic Embodied AI Research

DGX agent

arXiv:2607.06701v1 Announce Type: cross Abstract: Interactive simulators have become powerful tools for training embodied agents and generating synthetic visual data, but existing photorealistic simul

researcharxiv-cs-ai
9 Jul 2026
Model Releases

A VLM-Enhanced Framework for Comprehensive Traffic Sign Condition Assessment Integrating Daytime Visual Performance and Nighttime Retroreflectivity Evaluation

DGX agent

arXiv:2607.06478v1 Announce Type: new Abstract: Traffic signs are crucial components of road safety, serving as visual tools under all lighting conditions. The Manual on Uniform Traffic Control Device

model-releasesarxiv-cs-cv
8 Jul 2026
Agents

Akashic: A Low-Overhead LLM Inference Service with MemAttention

DGX agent

arXiv:2607.05708v1 Announce Type: new Abstract: Recent LLM-based agent systems continuously accumulate context across multi-turn interactions, tool invocations, and cross-session workflows. Replaying

agentsarxiv-cs-ai
8 Jul 2026
Research

Decision-Focused Scenario Generation and Selection for Efficient and Robust Grid Dispatch

DGX agent

arXiv:2607.05830v1 Announce Type: cross Abstract: The increasing uncertainty from flexible demand and renewable generation has made distributionally robust optimization (DRO) an important tool for rob

researcharxiv-cs-ai
8 Jul 2026
← Previous
1…2526272829…109
Next →