AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,557 results
26 May 2026

The Normalized Maximum Likelihood for Regular Non-Smooth Models: Measure-Theoretic Foundations and Geometric Sampling

ResearchDGX agent

arXiv:2605.24477v1 Announce Type: new Abstract: The Normalized Maximum Likelihood (NML) codelength, or stochastic complexity, represents a principled criterion for universal coding. While recent coare

The Perception-Physics Paradox: Probing Scientific Alignment with TC-Bench

Model ReleasesDGX agent

arXiv:2605.24782v1 Announce Type: new Abstract: While Vision Foundation Models (VFMs) excel at predictive tasks on satellite imagery, their performance can arise from visual correlations rather than u

The Quantization Benefits of Residual-Free Transformers

ResearchDGX agent

arXiv:2605.25880v1 Announce Type: new Abstract: Large-scale transformer training and deployment are increasingly constrained by the transfer of activations, gradients, and optimizer states across acce


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

The Timing Dependencies of Trust: Speed, Accuracy, and cBCI Neuro-Decoupling in Human-AI Teams

ResearchDGX agent

arXiv:2605.25868v1 Announce Type: cross Abstract: The speed and accuracy of an artificial teammate fundamentally alter the failure states of Human-AI integration. While high-speed AI interventions ris

TorchLean: Formalizing Neural Networks in Lean

SafetyDGX agent

arXiv:2602.22631v2 Announce Type: replace-cross Abstract: Neural networks are increasingly deployed in scientific, safety critical, and mission critical pipelines, yet verification and analysis are of

Towards Cognitively-Faithful Decision-Making Models to Improve AI Alignment

SafetyDGX agent

arXiv:2509.04445v2 Announce Type: replace Abstract: Recent AI trends seek to align AI models to learned human-centric objectives, such as personal preferences, utility, or societal values. Using stand

Towards Large Model Feature Coding

Model ReleasesDGX agent

arXiv:2605.24025v1 Announce Type: cross Abstract: Large models have delivered remarkable performance across a wide range of perception and generation tasks, yet practical deployment is increasingly co

Towards Long-Horizon Interpretability: Efficient and Faithful Multi-Token Attribution for Reasoning LLMs

ResearchDGX agent

arXiv:2602.01914v2 Announce Type: replace Abstract: Token attribution methods provide intuitive explanations for language model outputs by identifying causally important input tokens. However, as mode

Towards Understanding Adam Convergence on Highly Degenerate Polynomials

SafetyDGX agent

arXiv:2603.09581v2 Announce Type: replace Abstract: Adam is a widely used optimization algorithm in deep learning, yet the specific class of objective functions where it exhibits inherent advantages r

Towards Verifiable Transformers: Solver-Checkable Circuit Explanations

Local AiDGX agent

arXiv:2605.24033v1 Announce Type: new Abstract: Mechanistic interpretability often identifies circuits inside Transformer models, but explanations of those circuits are usually validated through examp

Trade-off Functions for DP-SGD with Subsampling based on Random Shuffling: Tight Upper and Lower Bounds

Model ReleasesDGX agent

arXiv:2605.06259v2 Announce Type: replace Abstract: We derive a tight analysis of the trade-off function for Differentially Private Stochastic Gradient Descent (DP-SGD) with subsampling based on rando

Train-Free Segmentation in MRI with Cubical Persistent Homology

ResearchDGX agent

arXiv:2401.01160v3 Announce Type: replace-cross Abstract: We investigate a framework for train-free MRI segmentation based on Topological Data Analysis. The pipeline proceeds in three steps, first ide

Trained quantum neural networks are Gaussian processes

ResearchDGX agent

arXiv:2402.08726v2 Announce Type: replace-cross Abstract: We study quantum neural networks made by parametric one-qubit gates and fixed two-qubit gates in the limit of infinite width, where the genera

Trajectory-Based Difficulty Scoring for Reliable Learning on Tabular Data

ResearchDGX agent

arXiv:2605.24680v1 Announce Type: new Abstract: Gradient-boosted trees achieve strong performance on tabular data, yet often leave a long tail of poorly predicted instances. We introduce a Trajectory-

Transformer-based few-shot learning for modeling Electricity Consumption Profiles with minimal data across thousands of domains

ResearchDGX agent

arXiv:2408.08399v3 Announce Type: replace Abstract: Electricity Consumption Profiles (ECPs) are crucial for operating and planning power distribution systems, especially with the increasing number of

TSFLora: Token-Compressed Split Fine-Tuning for Wireless Edge Networks

ResearchDGX agent

arXiv:2605.23988v1 Announce Type: cross Abstract: Adapting large AI models (LAMs) to personalized edge data is challenging because wireless devices have limited memory, computation, and uplink capacit

TUBE: Tangent Upper Bound on Evidence for Discrete Diffusion Language Models

ResearchDGX agent

arXiv:2605.24292v1 Announce Type: new Abstract: Log-likelihood is a standard metric for evaluating generative models. Unfortunately, in contrast to autoregressive models (ARMs), discrete diffusion mod

Turn-Based Structural Triggers: Prompt-Free Backdoors in Multi-Turn LLMs

ResearchDGX agent

arXiv:2601.14340v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are widely integrated into interactive systems such as dialogue agents and task-oriented assistants. This growing

UNATE: UNsupervised ATomic Embedding for crystal structures property prediction

TutorialsDGX agent

arXiv:2605.25866v1 Announce Type: new Abstract: Accurately predicting crystal properties is critical for accelerating materials discovery, but it is often limited by scarce labeled data and costly the

Unifying Value Alignment and Assignment in Cross-Domain Offline Reinforcement Learning with Heterogeneous Datasets

SafetyDGX agent

arXiv:2605.24862v1 Announce Type: new Abstract: Cross-domain offline reinforcement learning (RL) aims to learn a policy in the target domain with a limited target domain dataset and a source domain da

URS: A Unified Neural Routing Solver for Cross-Problem Zero-Shot Generalization

Model ReleasesDGX agent

arXiv:2509.23413v2 Announce Type: replace Abstract: Multi-task neural routing solvers have emerged as a promising paradigm for their ability to solve multiple vehicle routing problems (VRPs) using a s

V3H: View Variation and View Heredity for Incomplete Multi-view Clustering

Model ReleasesDGX agent

arXiv:2011.11194v4 Announce Type: replace Abstract: Real data often appear in the form of multiple incomplete views. Incomplete multi-view clustering is an effective method to integrate these incomple

Variable Clustering via Distributionally Robust Nodewise Regression

ResearchDGX agent

arXiv:2212.07944v3 Announce Type: replace Abstract: We study a multi-factor block model for variable clustering and connect it to regularized subspace clustering through a distributionally robust vers

ViroBench: Benchmarking Nucleotide Foundation Models on Viral Genomics Tasks

Model ReleasesDGX agent

arXiv:2605.25388v1 Announce Type: new Abstract: Nucleotide sequences constitute the fundamental genetic basis of biological systems, rendering viral genomic analysis critical for biomedical advancemen

Vision-Guided Outdoor Flight and Obstacle Evasion via Reinforcement Learning

SafetyDGX agent

arXiv:2605.24449v1 Announce Type: cross Abstract: Although quadcopters boast impressive traversal capabilities enabled by their omnidirectional maneuverability, the need for continuous pilot control i

Visual-Redundancy-Controlled Parallel Decoding for Diffusion-Based Multimodal Large Language Models

Model ReleasesDGX agent

arXiv:2605.25820v1 Announce Type: new Abstract: Diffusion-based multimodal large language models (dMLLMs) decode by iteratively predicting tokens at multiple masked positions in parallel. This turns e

Volatility Surface Reconstruction using Deep Learning under No-Arbitrage Constraints

ResearchDGX agent

arXiv:2605.24031v1 Announce Type: cross Abstract: We study the reconstruction of implied volatility surfaces from sparse and noisy option quotes using deep learning models under no-arbitrage constrain

When Interpretability Becomes a Liability: Adversarial Attacks on CBM Concept Layers

ResearchDGX agent

arXiv:2605.25304v1 Announce Type: new Abstract: Concept Bottleneck Models (CBMs) have emerged as a cornerstone approach for interpretable machine learning, providing human-understandable intermediate

Which Is Better For Reducing Outdated and Vulnerable Dependencies: Pinning or Floating?

ResearchDGX agent

arXiv:2510.08609v3 Announce Type: replace-cross Abstract: Developers consistently use version constraints to specify acceptable versions of the dependencies for their project. Pinning dependencies can

Why Agentic Theorem Prover Works: A Statistical Provability Theory of Mathematical Reasoning Models

AgentsDGX agent

arXiv:2602.10538v3 Announce Type: replace-cross Abstract: Agentic theorem provers combine a reasoning model, retrieval, search, and a proof assistant verifier, yet it remains unclear which components

WINO: A Weak-Form Physics Informed Neural Operator for Hyperelasticity on Variable Domains

ResearchDGX agent

arXiv:2605.24651v1 Announce Type: cross Abstract: We propose a Weak-form Physics-Informed Neural Operator (WINO), a data-free framework that combines the efficiency of neural operators with the geomet

WLNO: Wavelet-Laplace Neural Operator for Solving Partial Differential Equations

Model ReleasesDGX agent

arXiv:2605.24658v1 Announce Type: new Abstract: This work introduces the Wavelet-Laplace Neural Operator (WLNO), a novel neural operator that fuses Haar wavelet multi-scale spatial decomposition with

XRPO: Pushing the limits of GRPO with Targeted Exploration and Exploitation

SafetyDGX agent

arXiv:2510.06672v3 Announce Type: replace Abstract: Reinforcement learning algorithms such as GRPO have driven recent advances in large language model (LLM) reasoning. While scaling the number of roll

Zeroth-Order Nonconvex Nonsmooth Optimization with Heavy-Tailed Noise

ResearchDGX agent

arXiv:2605.24513v1 Announce Type: new Abstract: This paper considers the nonconvex nonsmooth problem in which the objective function is Lipschitz continuous. We focus on the stochastic setting where t

25 May 2026

A comprehensive evaluation of pretraining strategies for channel-agnostic contrastive self-supervision of biosignals

ResearchDGX agent

arXiv:2410.19842v2 Announce Type: replace-cross Abstract: Contrastive learning yields impressive results for self-supervision in computer vision. The approach relies on the creation of positive pairs,

A Simple Plug-in for Improving Eviction-Based KV Cache Compression

ResearchDGX agent

arXiv:2605.23258v1 Announce Type: new Abstract: KV cache growth is a major bottleneck for long-context inference in large language models. Existing methods are often dominated by binary eviction or re

Accelerating Divisible Load Processing Through Machine Learning: A Practical Framework for Large-Scale Workloads

TutorialsDGX agent

arXiv:2605.23247v1 Announce Type: new Abstract: In this paper, we introduce the first machine learning framework for predicting optimal processing times in Single-Level Tree Network (SLTN) architectur

Accelerating ground state search of spatial photonic Ising machines with genetic-simulated annealing hybrid algorithm

ResearchDGX agent

arXiv:2605.23295v1 Announce Type: cross Abstract: Spatial photonic Ising machines (SPIMs) based on spatial light modulators (SLMs) have emerged as highly effective solvers for many tasks, including co

Active Sensing Subserves Task-Level Control

ResearchDGX agent

arXiv:2605.22988v1 Announce Type: cross Abstract: Active sensing is traditionally defined as the expenditure of energy, typically in the form of movement, for obtaining information. Here, we propose t

Advanced AI Service Provisioning in O-RAN through LLM Engine Integration

ResearchDGX agent

arXiv:2605.23809v1 Announce Type: cross Abstract: The Open Radio Access Network (O-RAN) architecture allows AI to be embedded directly into the RAN through modular xApps and rApps, yet creating these

AGZO: Activation-Guided Zeroth-Order Optimization for LLM Fine-Tuning

ResearchDGX agent

arXiv:2601.17261v4 Announce Type: replace Abstract: Zeroth-Order (ZO) optimization has emerged as a promising solution for fine-tuning LLMs under strict memory constraints, as it avoids the prohibitiv

Amortized Simulation-Based Inference in Generalized Bayes via Neural Posterior Estimation

ResearchDGX agent

arXiv:2601.22367v2 Announce Type: replace-cross Abstract: Generalized Bayesian Inference (GBI) tempers a loss with a temperature eta > 0 to mitigate overconfidence and improve robustness under model m

An Open-Source Training Dataset for Foundation Models for Black-box Optimization

TutorialsDGX agent

arXiv:2605.23417v1 Announce Type: new Abstract: Most black-box optimization methods require extensive hyperparameter tuning, often limiting their ability to generalize across different optimization do

Any-Dimensional Invariant Universality

ResearchDGX agent

arXiv:2605.23156v1 Announce Type: new Abstract: Several machine learning models are defined for inputs of any size, such as graphs with different numbers of nodes and point clouds containing varying n

Approaching I/O-optimality for Approximate Attention

Model ReleasesDGX agent

arXiv:2605.23751v1 Announce Type: new Abstract: We revisit the I/O complexity of attention in large language models. Given query-key-value matrices Q,K,VinR^{nimes d}, and a machine with fast memory s

Archimedean Copula Inference via Taylor-Mode AD

Model ReleasesDGX agent

arXiv:2605.23134v1 Announce Type: new Abstract: No existing nested Archimedean copula tool handles all three of (a) arbitrary per-variable (right-)censoring in survival analysis, (b) arbitrary nesting

Are Targeted Data Poisoning Attacks as Effective as We Think?

ResearchDGX agent

arXiv:2509.06896v2 Announce Type: replace Abstract: Targeted data poisoning attacks manipulate model predictions on specific test samples by injecting malicious data into training. Yet existing evalua

Assessing Predictive Models for Fairness Based on Movement Patterns

SafetyDGX agent

arXiv:2605.23234v1 Announce Type: new Abstract: Assessing the spatial fairness of predictive models involves establishing whether they are statistically penalizing (favoring) individuals associated wi

Asymmetric Scaling Laws from Sparse Features

ResearchDGX agent

arXiv:2605.23591v1 Announce Type: cross Abstract: We introduce a model for neural scaling laws under sparse activations. In the model, test loss is often dominated by rare coordinates that are never o

Automatic Construction of Clinical Scoring Systems with LLM Agents

ResearchDGX agent

arXiv:2601.22324v2 Announce Type: replace Abstract: Modern clinical practice relies on evidence-based guidelines implemented as compact scoring systems composed of a small number of interpretable deci

Building a privacy-preserving Federated Recommender system for mobile devices

Local AiDGX agent

arXiv:2605.22924v1 Announce Type: new Abstract: Serving personalized content on mobile devices has traditionally required pooling sensitive user data on centralized servers, a practice increasingly at

CapTrack: Multifaceted Evaluation of Forgetting in LLM Post-Training

SafetyDGX agent

arXiv:2603.06610v2 Announce Type: replace Abstract: Large language model (LLM) post-training enhances latent skills, unlocks value alignment, improves performance, and enables domain adaptation. Unfor

Cascaded Transfer: Learning Many Tasks under Budget Constraints

ResearchDGX agent

arXiv:2601.21513v2 Announce Type: replace Abstract: In distributed applications, such as energy demand forecasting at the substation level or federated learning, a large number of related tasks must b

Causal Additive Models with Unobserved Causal Paths and Backdoor Paths

ResearchDGX agent

arXiv:2502.07646v3 Announce Type: replace Abstract: Causal additive models provide a tractable yet expressive framework for causal discovery in the presence of hidden variables. When unobserved backdo

Certification from Examples is Hard for Circuits and Transformers under Minimal Overparametrization

ResearchDGX agent

arXiv:2605.22964v1 Announce Type: new Abstract: As state-of-the-art neural networks are deployed on reasoning and algorithmic tasks, exactness guarantees become increasingly important. However, high a

Certified Per-Instance Unlearning Using Individual Sensitivity Bounds

ResearchDGX agent

arXiv:2602.15602v2 Announce Type: replace Abstract: Certified machine unlearning can be achieved via noise injection leading to differential privacy guarantees, where noise is calibrated to worst-case

Class-Dependent Hybrid Data Augmentation for Multiclass Migraine Classification under Severe Class Imbalance

SafetyDGX agent

arXiv:2605.23453v1 Announce Type: new Abstract: We conducted a reproducibility-oriented re-evaluation of prior migraine classification studies, correcting for data leakage and metric bias. We then int

Classification of IED-free EEG Responses for Assisted Epilepsy Diagnosis

ResearchDGX agent

arXiv:2605.22858v1 Announce Type: cross Abstract: Diagnosing epilepsy is challenging when routine EEGs lack interictal epileptiform discharges (IEDs). Intermittent photic stimulation (IPS) and hyperve

Complete-muE: Optimal Hyperparameter Transfer and Scaling for MoE Models

Model ReleasesDGX agent

arXiv:2605.23893v1 Announce Type: new Abstract: We propose Complete-muE, a framework which targets hyperparameter transfer across dense FFN and any Mixture-of-Experts (MoE) setups in transformer block

Contrast to Detect: Dynamic Graph Contrastive Regularization for Unsupervised Anomaly Detection in Multivariate Time Series

ApplicationsDGX agent

arXiv:2605.23744v1 Announce Type: new Abstract: Anomaly detection in multivariate time series (MTS) is hindered by dynamic inter-variable dependencies and feature entanglement under spectral noise, an

← Previous
1…129130131132133…243
Next →