AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “tools”

GridTimelineEvolution
5,189 results
Model Releases

Provably Data-driven Multiple Hyper-parameter Tuning with Structured Loss Function

DGX agent

arXiv:2602.02406v2 Announce Type: replace-cross Abstract: Data-driven algorithm design automates hyperparameter tuning, but its statistical foundations remain limited because model performance can dep

model-releasesarxiv-cs-lg
13 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Applications

Quantifying the Reconstructability of Astrophysical Methods with Large Language Models and Information Theory: A Case Study in Spectral Reconstruction

DGX agent

arXiv:2605.11154v1 Announce Type: cross Abstract: Modern astrophysical studies rely heavily on complex data analysis pipelines; however, published descriptions often lack the detail required for compu

applicationsarxiv-cs-lg
13 May 2026
Safety

Red-Teaming Text-to-Image Models via In-Context Experience Replay and Semantic-Preserving Prompt Rewriting

DGX agent

arXiv:2411.16769v3 Announce Type: replace-cross Abstract: Understanding the capabilities of text-to-image (T2I) models in harmful content generation is essential to safety and compliance. However, hum

safetyarxiv-cs-cl
13 May 2026
Safety

Sequential Off-Policy Learning with Logarithmic Smoothing

DGX agent

arXiv:2506.10664v2 Announce Type: replace-cross Abstract: Off-policy learning enables training policies from logged interaction data. Most prior work considers the batch setting, where a policy is lea

safetyarxiv-cs-lg
13 May 2026
Applications

Sign Language Recognition and Translation for Low-Resource Languages: Challenges and Pathways Forward

DGX agent

arXiv:2605.12096v1 Announce Type: new Abstract: Sign languages are natural, visual-gestural languages used by Deaf communities worldwide. Over 300 distinct sign languages remain severely low-resource

applicationsarxiv-cs-cl
13 May 2026
Safety

Space Syntax-guided Post-training for Residential Floor Plan Generation

DGX agent

arXiv:2602.22507v2 Announce Type: replace-cross Abstract: Residential floor plan generation requires not only geometric fidelity but also spatial configurational logic: shared living spaces should be

safetyarxiv-cs-cv
13 May 2026
Tutorials

The Confusion is Real: GRAPHIC -- A Network Science Approach to Confusion Matrices in Deep Learning

DGX agent

arXiv:2602.19770v2 Announce Type: replace Abstract: Explainable artificial intelligence has emerged as a promising field of research to address reliability concerns in artificial intelligence. Despite

tutorialsarxiv-cs-lg
13 May 2026
Applications

Towards Affordable Energy: A Gymnasium Environment for Electric Utility Demand-Response Programs

DGX agent

arXiv:2605.12462v1 Announce Type: cross Abstract: Extreme weather and volatile wholesale electricity markets expose residential consumers to catastrophic financial risks, yet demand response at the di

applicationsarxiv-cs-lg
13 May 2026
Applications

UniVLR: Unifying Text and Vision in Visual Latent Reasoning for Multimodal LLMs

DGX agent

arXiv:2605.11856v1 Announce Type: cross Abstract: Multimodal large language models are increasingly expected to perform thinking with images, yet existing visual latent reasoning methods still rely on

applicationsarxiv-cs-cl
13 May 2026
Safety

When Does ell_2-Boosting Overfit Benignly? High-Dimensional Risk Asymptotics and the ell_1 Implicit Bias

DGX agent

arXiv:2605.06314v2 Announce Type: replace Abstract: Benign overfitting is well-characterized in ell_2 geometries, but its behavior under the ell_1 implicit bias of greedy ensembles remains challenging

safetyarxiv-cs-lg
13 May 2026
Model Releases

A Geometric Perspective on Next-Token Prediction in Large Language Models: Three Emerging Phases

DGX agent

arXiv:2605.09011v1 Announce Type: cross Abstract: We investigate the geometry of predictive information across the layers of large language models (LLMs). We repurpose representation lenses-learned af

model-releasesarxiv-cs-ai
12 May 2026
Local Ai

A Physical Theory of Backpropagation: Exact Gradients from the Least-Action Principle

DGX agent

arXiv:2602.02281v2 Announce Type: replace-cross Abstract: Backpropagation is typically presented as a symbolic procedure: a backward pass topologically distinct from inference, with non-local error si

local-aiarxiv-cs-ai
12 May 2026
Research

A Unified Lyapunov-IQC Framework for Uniform Stability of Smooth Quadratic First-Order Accelerated Optimizers

DGX agent

arXiv:2605.08488v1 Announce Type: cross Abstract: We develop a unified Lyapunov-integral quadratic constraint (IQC) framework for establishing uniform stability of first-order accelerated optimization

researcharxiv-cs-lg
12 May 2026
Agents

A Versatile AI Agent for Rare Disease Diagnosis and Risk Gene Prioritization

DGX agent

arXiv:2605.06226v2 Announce Type: replace Abstract: Accurate and timely diagnosis is essential for effective treatment, particularly in the context of rare diseases. However, current diagnostic workfl

agentsarxiv-cs-ai
12 May 2026
Applications

Absurd World: A Simple Yet Powerful Method to Absurdify the Real-world for Probing LLM Reasoning Capabilities

DGX agent

arXiv:2605.09678v1 Announce Type: new Abstract: While extremely powerful and versatile at various tasks, the thinking capabilities of large language models (LLMs) are often put under scrutiny as they

applicationsarxiv-cs-ai
12 May 2026
Agents

Ace-Skill: Bootstrapping Multimodal Agents with Prioritized and Clustered Evolution

DGX agent

arXiv:2605.08887v1 Announce Type: new Abstract: Self-evolving agents present a promising path toward continual adaptation by distilling task interactions into reusable knowledge artifacts. In practice

agentsarxiv-cs-ai
12 May 2026
Safety

Agent-Sentry: Bounding LLM Agents via Execution Provenance

DGX agent

arXiv:2603.22868v2 Announce Type: replace-cross Abstract: Agentic computing systems, while immensely capable, raise serious security, privacy, and safety concerns. A key issue is that the full set of

safetyarxiv-cs-ai
12 May 2026
Agents

Agentic AI Scientists Are Not Built For Autonomous Scientific Discovery

DGX agent

arXiv:2605.08956v1 Announce Type: new Abstract: A growing body of work pursues AI scientists capable of end-to-end autonomous scientific discovery. This position paper argues that although they alread

agentsarxiv-cs-ai
12 May 2026
Safety

Alignment as Jurisprudence

DGX agent

arXiv:2605.08416v1 Announce Type: new Abstract: Jurisprudence, the study of how judges should properly decide cases, and alignment, the science of getting AI models to conform to human values, share a

safetyarxiv-cs-ai
12 May 2026
Agents

An agentic framework for gravitational-wave counterpart association in the multi-messenger era

DGX agent

arXiv:2605.10584v1 Announce Type: cross Abstract: With the detection of gravitational waves (GWs), multi-messenger astronomy has opened a new window for advancing our understanding of astrophysics, de

agentsarxiv-cs-ai
12 May 2026
Research

AU-Harness: An Open-Source Toolkit for Holistic Evaluation of Audio LLMs

DGX agent

arXiv:2509.08031v3 Announce Type: replace-cross Abstract: Large Audio Language Models (LALMs) are rapidly advancing, but evaluating them remains challenging due to inefficient and non-standardized too

researcharxiv-cs-ai
12 May 2026
Safety

Auditing Data Membership in Reinforcement Learning With Verifiable Rewards

DGX agent

arXiv:2511.14045v2 Announce Type: replace-cross Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has become a core training stage in recent large language models (LLMs). Its reliance on

safetyarxiv-cs-ai
12 May 2026
Model Releases

Beyond the All-in-One Agent: Benchmarking Role-Specialized Multi-Agent Collaboration in Enterprise Workflows

DGX agent

arXiv:2605.08761v1 Announce Type: cross Abstract: Large language model (LLM) agents are increasingly expected to operate in enterprise environments, where work is distributed across specialized roles,

model-releasesarxiv-cs-lg
12 May 2026
Research

Causal Parametric Drift Simulation: A Digital Twin Framework for Classifier Robustness Evaluation

DGX agent

arXiv:2605.09663v1 Announce Type: cross Abstract: Machine learning classifiers in dynamic environments face concept drift -- changes in the data-generating process that degrade performance. Convention

researcharxiv-cs-ai
12 May 2026
Safety

Conformity Generates Collective Misalignment in AI Agents Societies

DGX agent

arXiv:2605.10721v1 Announce Type: cross Abstract: Artificial intelligence safety research focuses on aligning individual language models with human values, yet deployed AI systems increasingly operate

safetyarxiv-cs-cl
12 May 2026
Agents

Consistency as a Testable Property: Statistical Methods to Evaluate AI Agent Reliability

DGX agent

arXiv:2605.10516v1 Announce Type: new Abstract: This paper establishes a rigorous measurement science for AI agent reliability, providing a foundational framework for quantifying consistency under sem

agentsarxiv-cs-ai
12 May 2026
Research

CONTRA: Conformal Prediction Region via Normalizing Flow Transformation

DGX agent

arXiv:2605.08561v1 Announce Type: cross Abstract: Density estimation and reliable prediction regions for outputs are crucial in supervised and unsupervised learning. While conformal prediction effecti

researcharxiv-cs-lg
12 May 2026
Applications

Counterfactual Stress Testing for Image Classification Models

DGX agent

arXiv:2605.10894v1 Announce Type: new Abstract: Deep learning models in medical imaging often fail when deployed in new clinical environments due to distribution shifts in demographics, scanner hardwa

applicationsarxiv-cs-cv
12 May 2026
Research

Cplus2ASP: Computing Action Language C+ in Answer Set Programming

DGX agent

arXiv:2605.09528v1 Announce Type: new Abstract: We present Version 2 of system Cplus2ASP, which implements the definite fragment of action language C+. Its input language is fully compatible with the

researcharxiv-cs-ai
12 May 2026
Applications

DataArc-SynData-Toolkit: A Unified Closed-Loop Framework for Multi-Path, Multimodal, and Multilingual Data Synthesis

DGX agent

arXiv:2605.08138v1 Announce Type: new Abstract: Synthetic data has emerged as a crucial solution to the data scarcity bottleneck in large language models (LLMs), particularly for specialized domains a

applicationsarxiv-cs-lg
12 May 2026
Research

Defense effectiveness across architectural layers: a mechanistic evaluation of persistent memory attacks on stateful LLM agents

DGX agent

arXiv:2605.08442v1 Announce Type: cross Abstract: Persistent memory attacks against LLM agents achieve high attack success rates against open-source models. In these attacks, malicious instructions in

researcharxiv-cs-ai
12 May 2026
Agents

Delivering Science as a Service: Sci-Orchestra's Cloud-Native Approach to HPC

DGX agent

arXiv:2605.08396v1 Announce Type: new Abstract: The increasing complexity of modern computational environments often burdens researchers with infrastructure management, authentication protocols, and c

agentsarxiv-cs-cv
12 May 2026
Applications

Discriminative Span as a Predictor of Synthetic Data Utility via Classifier Reconstruction

DGX agent

arXiv:2605.09697v1 Announce Type: new Abstract: In many real-world computer vision applications, including medical imaging and industrial inspection, binary classification tasks are characterized by a

applicationsarxiv-cs-cv
12 May 2026
Tutorials

Do LLMs Experience an Internal Polylogue? Investigating Reasoning through the Lens of Personas

DGX agent

arXiv:2605.09159v1 Announce Type: new Abstract: Recent work shows that large language models (LLMs) encode behavioural traits ('personas') as linear directions in activation space, often called 'perso

tutorialsarxiv-cs-ai
12 May 2026
Research

Efficient Statistics With Unknown Truncation, Polynomial Time Algorithms, Beyond Gaussians

DGX agent

arXiv:2410.01656v2 Announce Type: replace-cross Abstract: We study the estimation of distributional parameters when samples are shown only if they fall in some unknown set S subseteq R^d. Kontonis, Tz

researcharxiv-cs-lg
12 May 2026
Research

Enabling Structure-Only Initialization and Out-of-Distribution Generalization in GNN-based Molecular Dynamics Simulators

DGX agent

arXiv:2605.09495v1 Announce Type: cross Abstract: Machine learning-based simulators offer the potential to model the dynamics of complex systems more efficiently than classical approaches, while retai

researcharxiv-cs-lg
12 May 2026
Agents

Engineering Robustness into Personal Agents with the AI Workflow Store

DGX agent

arXiv:2605.10907v1 Announce Type: cross Abstract: The dominant paradigm for AI agents is an 'on-the-fly' loop in which agents synthesize plans and execute actions within seconds or minutes in response

agentsarxiv-cs-ai
12 May 2026
Model Releases

ERIS: Enhancing Privacy and Scalability in Federated Learning via Federated Shard Aggregation

DGX agent

arXiv:2602.08617v2 Announce Type: replace Abstract: Scaling Federated Learning (FL) to billion-parameter models forces a challenging trade-off between privacy, scalability, and model utility. Existing

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Exploring the AI Obedience: Why is Generating a Pure Color Image Harder than CyberPunk?

DGX agent

arXiv:2603.00166v2 Announce Type: replace-cross Abstract: Recent advances in generative AI have shown human-level performance in complex content creation. However, we identify a 'Paradox of Simplicity

model-releasesarxiv-cs-ai
12 May 2026
Safety

Extended Wasserstein-GAN Approach to Causal Distribution Learning: Density-Free Estimation and Minimax Optimality

DGX agent

arXiv:2605.10206v1 Announce Type: cross Abstract: Distributional causal inference requires estimating not only average treatment effects but also interventional outcome distributions, including quanti

safetyarxiv-cs-lg
12 May 2026
Research

Factual recall in linear associative memories: sharp asymptotics and mechanistic insights

DGX agent

arXiv:2605.10795v1 Announce Type: cross Abstract: Large language models demonstrate remarkable ability in factual recall, yet the fundamental limits of storing and retrieving input--output association

researcharxiv-cs-lg
12 May 2026
Agents

FocuSFT: Bilevel Optimization for Dilution-Aware Long-Context Fine-Tuning

DGX agent

arXiv:2605.09932v1 Announce Type: new Abstract: Large language models can now process increasingly long inputs, yet their ability to effectively use information spread across long contexts remains lim

agentsarxiv-cs-cl
12 May 2026
Applications

From Controlled to the Wild: Evaluation of Pentesting Agents for the Real-World

DGX agent

arXiv:2605.10834v1 Announce Type: new Abstract: AI pentesting agents are increasingly credible as offensive security systems, but current benchmarks still provide limited guidance on which will perfor

applicationsarxiv-cs-ai
12 May 2026
Research

From pre-training to downstream performance: Does domain-specific pre-training make sense?

DGX agent

arXiv:2605.08819v1 Announce Type: new Abstract: Deep learning techniques have revolutionised medical imaging, improving diagnostic accuracy and enabling both more accurate and earlier disease detectio

researcharxiv-cs-cv
12 May 2026
Model Releases

General Agent Evaluation

DGX agent

arXiv:2602.22953v2 Announce Type: replace Abstract: General-purpose agents perform tasks in unfamiliar environments without domain-specific manual customization. Yet no study has systematically measur

model-releasesarxiv-cs-ai
12 May 2026
Hardware

GPU-Accelerated Synthesis of Mixed-Boolean Arithmetic: Beyond Caching

DGX agent

arXiv:2605.08243v1 Announce Type: cross Abstract: Synthesizing Mixed-Boolean Arithmetic (MBA) expressions from input-output examples is central to program deobfuscation and also useful for compiler op

hardwarearxiv-cs-lg
12 May 2026
Model Releases

Grounding the Score: Explicit Visual Premise Verification for Reliable Vision-Language Process Reward Models

DGX agent

arXiv:2603.16253v2 Announce Type: replace-cross Abstract: Vision-language process reward models (VL-PRMs) are increasingly used to score intermediate reasoning steps and rerank candidates under test-t

model-releasesarxiv-cs-ai
12 May 2026
Research

Harmonized Feature Conditioning and Frequency-Prompt Personalization for Multi-Rater Medical Segmentation

DGX agent

arXiv:2605.08210v1 Announce Type: new Abstract: Multi-rater medical image segmentation captures the inherent ambiguity of clinical interpretation, where diagnostic boundaries vary across experts and i

researcharxiv-cs-cv
12 May 2026
← Previous
1…9091929394…109
Next →