AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “tools”

GridTimelineEvolution
5,141 results
Agents

A Gateway Architecture for Enterprise MCP Authentication: Unifying Heterogeneous Auth, Identity Delegation, and the User / Non-User Persona Problem

DGX agent

arXiv:2608.10760v1 Announce Type: cross Abstract: The Model Context Protocol (MCP) has become the de-facto interface for connecting LLM agents to enterprise tools, and adoption has been explosive: wit

agentsarxiv-cs-ai
12 Aug 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Bayesian-Agent: Posterior-Guided Skill Evolution Across LLM Agent Harnesses

DGX agent

arXiv:2606.08348v2 Announce Type: replace Abstract: LLM agents increasingly rely on prompts, tools, memory, SOPs, skills, and harness feedback, yet current self-evolution pipelines often update these

model-releasesarxiv-cs-cl
12 Aug 2026
Applications

Bayesian Federated Cause-of-Death Classification and Quantification Under Distribution Shift

DGX agent

arXiv:2505.02257v2 Announce Type: replace-cross Abstract: In regions lacking medically certified causes of death, verbal autopsy (VA) is a widely used tool to ascertain the cause of death through inte

applicationsarxiv-cs-lg
12 Aug 2026
Agents

Do Personalized Skills Help Coding Agents? An Empirical Study of Developer Interaction Histories

DGX agent

arXiv:2608.10319v1 Announce Type: cross Abstract: Large language model (LLM)-powered agents have rapidly evolved from code-completion tools into solvers of complex software engineering tasks. As devel

agentsarxiv-cs-ai
12 Aug 2026
Safety

Enhancing Automated Essay Scoring With Three Techniques: Two-Stage Fine-Tuning, Score Alignment, and Self-Training

DGX agent

arXiv:2602.01747v2 Announce Type: replace Abstract: Automated Essay Scoring (AES) plays a crucial role in education by providing scalable and efficient assessment tools. However, in real-world setting

safetyarxiv-cs-cl
12 Aug 2026
Agents

How to Dogfood Your AI Chat Agent: A Three-Layer Evaluation Framework with Goal-Directed NPC Simulation

DGX agent

arXiv:2608.09939v1 Announce Type: cross Abstract: Production teams deploying LLM chat agents face a specific quality assurance gap: existing evaluation tools test individual responses or simulate soci

agentsarxiv-cs-ai
12 Aug 2026
Safety

Most biomedical publications show signs of LLM-assisted writing

DGX agent

arXiv:2608.10715v1 Announce Type: cross Abstract: Over the past several years, LLM-powered chatbots and agents have become widely used as a tool for academic writing. LLM-assisted writing can be valua

safetyarxiv-cs-ai
12 Aug 2026
Agents

The Signal Rail: A Deterministic Motion Grammar for Communicating Conversational Agent State in Terminal Interfaces

DGX agent

arXiv:2608.10689v1 Announce Type: cross Abstract: Terminal interfaces to conversational agents report rich internal state (listening, thinking, executing tools, awaiting input, failing) almost entirel

agentsarxiv-cs-cl
12 Aug 2026
Research

Uncertainty-Aware Deep Learning for Genomics Applications: Insights from an Empirical Study

DGX agent

arXiv:2608.11054v1 Announce Type: new Abstract: Deep learning models have emerged as the standard computational tool for a wide range of applications in genomics. Yet, uncertainty quantification (UQ)

researcharxiv-cs-lg
12 Aug 2026
Research

A Content-Aware Pure Permutation with Intrinsic Avalanche Effect: Breaking the Diffusion-Permutation Dichotomy

DGX agent

arXiv:2608.09452v1 Announce Type: new Abstract: Pixel permutation is a fundamental tool in image processing, image encryption, and data hiding (including watermarking and steganography) that rearrange

researcharxiv-cs-cv
11 Aug 2026
Safety

Agentic Harnesses: LLM-Driven Verification Layers for Robot Autonomy

DGX agent

arXiv:2608.09857v1 Announce Type: cross Abstract: Advances in advanced artificial intelligence tools have sparked research in robot autonomy, but the development of such systems has largely focused on

safetyarxiv-cs-ai
11 Aug 2026
Safety

Beyond cognacy

DGX agent

arXiv:2507.03005v3 Announce Type: replace Abstract: Computational phylogenetics has become an established tool in historical linguistics, with many language families now analyzed using likelihood-base

safetyarxiv-cs-cl
11 Aug 2026
Applications

Carnot: Interpretable, Interactive, and Optimized Execution of Deep Research Queries

DGX agent

arXiv:2608.09532v1 Announce Type: cross Abstract: Enterprises increasingly seek to query data lakes using natural language via AI-driven tools like semantic operators or deep research agents. However,

applicationsarxiv-cs-ai
11 Aug 2026
Research

CuteTTS: Efficient and High-Quality Speech Synthesis via Autoregressive Modeling of Continuous Latents

DGX agent

arXiv:2608.08638v1 Announce Type: cross Abstract: Zero-shot text-to-speech (TTS) now supports interactive assistants, personalized media, and accessibility tools. All TTS systems require faithful ling

researcharxiv-cs-ai
11 Aug 2026
Model Releases

DarwinX: Evolving Agent Harnesses Through Natural Selection

DGX agent

arXiv:2608.07545v1 Announce Type: cross Abstract: An LLM agent's capability depends not only on model weights but on its harness: prompts, tools, skills, and control flow. Self-improvement loops alrea

model-releasesarxiv-cs-ai
11 Aug 2026
Hardware

Design Space of Self--Consistent Electrostatic Machine Learning Interatomic Potentials

DGX agent

arXiv:2603.14700v2 Announce Type: replace-cross Abstract: Machine learning interatomic potentials (MLIPs) have become widely used tools in atomistic simulations. For much of the history of this field,

hardwarearxiv-cs-lg
11 Aug 2026
Safety

Distributed Optimization with Streaming Data: A Temporal Weighting Perspective

DGX agent

arXiv:2608.09565v1 Announce Type: cross Abstract: Optimization theory is a widely used tool for intelligent decision-making. While classical optimization deals with fixed, time-invariant objective fun

safetyarxiv-cs-ai
11 Aug 2026
Research

Efficient identification of critical regions via Flow Matching-based Monte Carlo initialization

DGX agent

arXiv:2508.15318v5 Announce Type: replace-cross Abstract: Markov chain Monte Carlo (MCMC) is a standard tool for studying many-body systems, but its practical cost can become substantial, especially w

researcharxiv-cs-lg
11 Aug 2026
Research

EvalConvoLearn: An Open-Source Framework for Evaluating Grounded Learner Simulations in Tutoring Conversations

DGX agent

arXiv:2608.07497v1 Announce Type: cross Abstract: Conversational learner simulations are valuable tools for testing learning theories, evaluating instructional materials and automated tutors, or power

researcharxiv-cs-cl
11 Aug 2026
Model Releases

Harmful Content Is Not Enough: Continuation Framing Moderates In-Context Emergent Misalignment

DGX agent

arXiv:2608.08212v1 Announce Type: new Abstract: In-context learning (ICL) can induce emergent misalignment (EM), where narrow misaligned examples alter answers to unrelated questions. Existing prompts

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Hierarchical Self-Improvement: A Framework for Task-Specific Evolvable Agent Harnesses

DGX agent

arXiv:2608.08466v1 Announce Type: new Abstract: Modern LLM agents are often improved by modifying prompts, tools, or workflows manually, while the executable scaffold surrounding the model---the harne

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

How to Ask the AI: A User Perspective Survey for Large Language Model Prompting

DGX agent

arXiv:2608.07494v1 Announce Type: cross Abstract: AI tools like ChatGPT and DeepSeek, powered by Large Language Models (LLMs), allow users to obtain instant and effective content responses simply by t

model-releasesarxiv-cs-ai
11 Aug 2026
Tutorials

Idea Search: Guiding Tree Search with Ideas to Explore Diverse Scientific Methods

DGX agent

arXiv:2608.08958v1 Announce Type: cross Abstract: Tree Search-based test-time scaling of LLMs is a powerful tool for automated scientific coding. However, pure Tree Search sometimes struggles with sys

tutorialsarxiv-cs-ai
11 Aug 2026
Model Releases

Knowing You Is Everything: LLM Agents Achieve Near-Perfect Profile-Consistent Reaction Prediction in Social Media Simulation

DGX agent

arXiv:2608.07498v1 Announce Type: cross Abstract: Autonomous AI agents in social media present concrete risks to democratic discourse and platform governance, while also offering tools for pre-deploym

model-releasesarxiv-cs-ai
11 Aug 2026
Research

Logarithmic-Free Moment and Generalization Bounds for Uniformly Stable Algorithms

DGX agent

arXiv:2608.09870v1 Announce Type: cross Abstract: Uniform stability is a classical tool for controlling the generalization error of a learning algorithm. Bousquet, Klochkov, and Zhivotovskiy (2020) sh

researcharxiv-cs-lg
11 Aug 2026
Model Releases

Open Evaluation Agent: Efficient and Promptable Evaluation of Visual Generative Models

DGX agent

arXiv:2608.09666v1 Announce Type: new Abstract: Recent advances in visual generative models have enabled high-quality image and video generation, but evaluating these models often demands sampling hun

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

OpenLoopEvolve: A Verifiable Self-Evolution Framework for Loop Policies in Long-Horizon Complex Tasks

DGX agent

arXiv:2608.09380v1 Announce Type: new Abstract: Long-horizon complex tasks require agents to repeatedly observe states, formulate plans, invoke tools, verify results, and recover from failures in cont

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Ouroboros: A Self-Developing Frontier Coding Agent with Reviewed Core Evolution

DGX agent

arXiv:2608.08311v1 Announce Type: cross Abstract: We present Ouroboros, a self-developing agent harness whose tools, prompts, context assembly, and core implementation improve through reviewed commits

model-releasesarxiv-cs-ai
11 Aug 2026
Research

RoboSeg: Online Part-Level Semantic Reconstruction for Robotic Manipulation via a Single Eye-in-Hand Camera

DGX agent

arXiv:2608.09778v1 Announce Type: new Abstract: Robotic manipulation requires perception systemsthat identify actionable parts such as handles, rims, triggers,and tool tips, not merely object categori

researcharxiv-cs-ro
11 Aug 2026
Safety

Safety Cost of Steering Vectors Is Separable and Reducible

DGX agent

arXiv:2608.08383v1 Announce Type: new Abstract: Steering vectors are a lightweight tool for controlling LLM behavior. However, emerging evidence shows that steering vectors can unintentionally comprom

safetyarxiv-cs-cl
11 Aug 2026
Model Releases

SkillSentry: Reliable Skill Execution for LLM Agents via Runtime Assurance

DGX agent

arXiv:2608.09253v1 Announce Type: new Abstract: LLM agents are increasingly equipped with skills to perform complex tasks through multi-step reasoning and tool use. Although skills provide reusable pr

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Sparse corruption in low-rank matrix inference: the PCA benchmark

DGX agent

arXiv:2511.11927v2 Announce Type: replace-cross Abstract: Principal Component Analysis (PCA) is a standard tool for extracting a low-rank signal from noisy observations. It is known that applying PCA

model-releasesarxiv-cs-lg
11 Aug 2026
Model Releases

Stateful CARS: Exact Cross-History Reuse for Policy-Constrained LLM Agents

DGX agent

arXiv:2608.08282v1 Announce Type: new Abstract: Tool-using language-model agents face constraints whose meaning changes with observations and prior actions. We study exact sampling from the model dist

model-releasesarxiv-cs-lg
11 Aug 2026
Tutorials

Transformer Explainer: Learning LLM Transformers with Interactive Visual Explanation and Experimentation

DGX agent

arXiv:2408.04619v2 Announce Type: replace-cross Abstract: The Transformer architecture underpins modern large language models powering state-of-the-art text generation and AI applications. However, it

tutorialsarxiv-cs-ai
11 Aug 2026
Model Releases

TREAT: Evaluating Access to Formal Knowledge across Equivalent Mathematical Representations

DGX agent

arXiv:2608.07540v1 Announce Type: new Abstract: AI systems increasingly operate between flexible input representations and formal objects used by downstream tools. A key challenge is recognizing when

model-releasesarxiv-cs-ai
11 Aug 2026
Research

TSPORec: Token Selection via Preference Optimization for LLM-Based Sequential Recommendation

DGX agent

arXiv:2608.09605v1 Announce Type: cross Abstract: Large Language Models (LLMs) have emerged as powerful tools for improving recommendation systems. The effectiveness of LLMs arises from their ability

researcharxiv-cs-ai
11 Aug 2026
Model Releases

Two-Layer Linear Auto-Regressive Models Estimate Latent States

DGX agent

arXiv:2606.12691v2 Announce Type: replace-cross Abstract: Auto-regressive models have emerged as powerful tools for sequential data, from language to video. Understanding how and why these models lear

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

When Counterbalancing Hides the Bias: Access-Conditioned Position Lock in Forced-Choice LLM Evaluation

DGX agent

arXiv:2607.10202v2 Announce Type: replace Abstract: Forced-choice probes with counterbalanced orientations are a standard tool for measuring language-model 'value dispositions,' and a concentration/ex

model-releasesarxiv-cs-lg
11 Aug 2026
Safety

Blind to the Pivotal Vote: Aggregate Independence Metrics Miss Where Verification Actually Helps

DGX agent

arXiv:2608.06940v1 Announce Type: new Abstract: LLM judge panels are a standard evaluation tool, but prior work reports highly correlated panel errors: nine judges provide roughly the effective inform

safetyarxiv-cs-ai
10 Aug 2026
Agents

CAi Copilot: Reducing Operational Workload in Molecular Design through Intent-Driven Agentic Workflows

DGX agent

arXiv:2608.06961v1 Announce Type: new Abstract: Early-stage molecular design is an iterative process, not just a task of generating molecules. Researchers turn broad goals into design strategies, refi

agentsarxiv-cs-ai
10 Aug 2026
Research

CAS2UML: A Handwritten Sketch-to-PlantUML Dataset for Class and Activity Diagrams

DGX agent

arXiv:2608.07036v1 Announce Type: cross Abstract: Automated UML generation from sketches and images is gaining renewed attention with the rise of large language models and multimodal AI. However, repr

researcharxiv-cs-cv
10 Aug 2026
Research

EpiFlow: A framework for improving the utility of wastewater signals for disease forecasting

DGX agent

arXiv:2608.06671v1 Announce Type: new Abstract: Wastewater-based surveillance is an effective tool for disease monitoring and can provide early warning of outbreaks. Although wastewater viral loads (W

researcharxiv-cs-lg
10 Aug 2026
Research

From Points to Edges: Edge-Conditioned Spectral Operators for Physics-Sensitive PDE Learning

DGX agent

arXiv:2608.06894v1 Announce Type: new Abstract: Neural operators have become a central tool for solving partial differential equations (PDEs), with spectral operators offering efficient global mixing

researcharxiv-cs-ai
10 Aug 2026
Model Releases

HarnessSafe: Evaluating Safety Across Persistent Carriers in Agent Harnesses

DGX agent

arXiv:2608.06984v1 Announce Type: cross Abstract: Modern agent harnesses persist state across tasks and sessions through persistent carriers like memory, skills, tools, and shared artifacts. However,

model-releasesarxiv-cs-ai
10 Aug 2026
Safety

KnifeHunter: Structured Local Representation Learning for Fine-Grained Knife Image Retrieval in Law Enforcement

DGX agent

arXiv:2608.07057v1 Announce Type: new Abstract: Knife-enabled violence presents a major public safety challenge, and law enforcement agencies require scalable tools for catalogue-level knife identific

safetyarxiv-cs-cv
10 Aug 2026
Model Releases

Long-Horizon Agent Trajectory Attribution: A Unified Benchmark and Fine-Grained Annotation Framework

DGX agent

arXiv:2608.06909v1 Announce Type: new Abstract: Large language model (LLM) agents increasingly operate through long-horizon trajectories involving user instructions, tool use, external observations, a

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

Science Edge Evaluation: SEE the Missing Step Toward Real Scientific Discovery

DGX agent

arXiv:2608.06931v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly involved in scientific discovery, yet it remains unclear whether they can support complex real laboratory

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

WebRider: Persona-Conditioned Intent Controllers for Live-Web Assistance

DGX agent

arXiv:2608.06704v1 Announce Type: new Abstract: Delegating a web task involves more than asking a question; it requires transferring a policy: what to verify, how to handle uncertainty, which preferen

model-releasesarxiv-cs-ai
10 Aug 2026
← Previous
1…2122232425…108
Next →