AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,557 results
15 May 2026

Separating Intrinsic Ambiguity from Estimation Uncertainty in Deep Generative Models for Linear Inverse Problems

ResearchDGX agent

arXiv:2605.15050v1 Announce Type: new Abstract: Recently, deep generative models have been used for posterior inference in inverse problems, including high-stakes applications in medical imaging and s

Silent Collapse in Recursive Learning Systems

AgentsDGX agent

arXiv:2605.14588v1 Announce Type: new Abstract: Recursive learning -- where models are trained on data generated by previous versions of themselves -- is increasingly common in large language models,

Slower Generalization, Faster Memorization: A Sweet Spot in Algorithmic Learning

ResearchDGX agent

arXiv:2605.14659v1 Announce Type: new Abstract: Critical-data-size accounts of grokking suggest a natural post-threshold intuition: once training data is sufficient to identify the underlying rule, ad


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Smooth Multi-Policy Causal Effect Estimation in Longitudinal Settings

SafetyDGX agent

arXiv:2605.14284v1 Announce Type: new Abstract: Comparative evaluation of multiple dynamic treatment policies is essential for healthcare and policy decisions, yet conventional longitudinal causal inf

SplineFlow: Flow Matching for Dynamical Systems with B-Spline Interpolants

TutorialsDGX agent

arXiv:2601.23072v2 Announce Type: replace Abstract: Flow matching is a scalable generative framework for characterizing continuous normalizing flows with wide-range applications. However, current stat

Stochastic Attention via Langevin Dynamics on the Modern Hopfield Energy

Model ReleasesDGX agent

arXiv:2603.06875v3 Announce Type: replace Abstract: Attention heads retrieve: given a query, they return a weighted average of stored values. We showed that this computation is one step of gradient de

Stochastic dynamics learning with state-space systems

AgentsDGX agent

arXiv:2508.07876v2 Announce Type: replace-cross Abstract: This work advances the theoretical foundations of reservoir computing (RC) by providing a unified treatment of fading memory and the echo stat

Stochastic Matching via Local Sparsification

ResearchDGX agent

arXiv:2605.14195v1 Announce Type: cross Abstract: The classic online stochastic matching problem typically requires immediate and irrevocable matching decisions. However, in many modern decentralized

Support Before Frequency in Discrete Diffusion

TutorialsDGX agent

arXiv:2605.13999v1 Announce Type: new Abstract: Discrete diffusion models are increasingly competitive for language modeling, yet it remains unclear how their denoising objectives organize learning. A

SurF: A Generative Model for Multivariate Irregular Time Series Forecasting

ApplicationsDGX agent

arXiv:2605.14069v1 Announce Type: new Abstract: Irregularly sampled multivariate event streams remain a stubbornly difficult modality for generative modeling: tokenization-based approaches break down

Synthetic American Option Pricing via Jump-HMM-Driven Heston Implied Volatility

Model ReleasesDGX agent

arXiv:2605.13998v1 Announce Type: cross Abstract: Generating realistic synthetic option prices requires implied volatility as an input, yet implied volatility is itself derived from observed option pr

Synthetic Sociality: How Generative Models Privatize the Social Fabric

HardwareDGX agent

arXiv:2605.14090v1 Announce Type: cross Abstract: We put forth a critical theoretical framework for analyzing generative models both descriptively and normatively. Our thesis is that generative models

TabClustPFN: A Prior-Fitted Network for Tabular Data Clustering

ApplicationsDGX agent

arXiv:2601.21656v3 Announce Type: replace Abstract: Clustering tabular data is a fundamental yet challenging problem due to heterogeneous feature types, diverse data-generating mechanisms, and the abs

TabPFN-3: Technical Report

Model ReleasesDGX agent

arXiv:2605.13986v1 Announce Type: new Abstract: Tabular data underpins most high-value prediction problems in science and industry, and TabPFN has driven the foundation model revolution for this modal

Temporal Fair Division in Multi-Agent Systems: From Precise Alternation Metrics to Scalable Coordination Proxies

SafetyDGX agent

arXiv:2605.14879v1 Announce Type: cross Abstract: A plethora real-world environments require agents to compete repeatedly for the same limited resource, calling for a temporal notion of fairness judge

Test-Time Learning with an Evolving Library

Model ReleasesDGX agent

arXiv:2605.14477v1 Announce Type: new Abstract: We introduce EvoLib, a test-time learning framework that enables large language models to accumulate, reuse, and evolve knowledge across problem instanc

Text-Dependent Speaker Verification (TdSV) Challenge 2024: Team Naive System Report

ResearchDGX agent

arXiv:2605.14896v1 Announce Type: cross Abstract: This paper presents a system for the 2024 Text-Dependent Speaker Verification (TdSV) Challenge. The system achieved a Minimum Detection Cost Function

The Rate-Distortion-Polysemanticity Tradeoff in SAEs

Model ReleasesDGX agent

arXiv:2605.14694v1 Announce Type: new Abstract: Sparse Autoencoders (SAEs) that can accurately reconstruct their input (minimizing distortion) by making efficient use of few features (minimizing the r

The Spheres Dataset: Multitrack Orchestral Recordings for Music Source Separation and Information Retrieval

ResearchDGX agent

arXiv:2511.21247v2 Announce Type: replace-cross Abstract: This paper introduces The Spheres dataset, multitrack orchestral recordings designed to advance machine learning research in music source sepa

TILBench: A Systematic Benchmark for Tabular Imbalanced Learning Across Data Regimes

Model ReleasesDGX agent

arXiv:2605.14915v1 Announce Type: new Abstract: Imbalanced learning remains a fundamental challenge in tabular data applications. Despite decades of research and numerous proposed algorithms, a system

TILT: Target-induced loss tilting under covariate shift

Model ReleasesDGX agent

arXiv:2605.14280v1 Announce Type: new Abstract: We introduce and analyze Target-Induced Loss Tilting (TILT) for unsupervised domain adaptation under covariate shift. It is based on a novel objective f

To discretize continually: Mean shift interacting particle systems for Bayesian inference

Model ReleasesDGX agent

arXiv:2605.14142v1 Announce Type: cross Abstract: Integration against a probability distribution given its unnormalized density is a central task in Bayesian inference and other fields. We introduce n

ToMAToMP: Robust and Multi-Parameter Topological Clustering

Model ReleasesDGX agent

arXiv:2605.14824v1 Announce Type: new Abstract: Topological clustering, and its main algorithm ToMATo, is a clustering method from Topological Data Analysis (TDA) which has been applied successfully i

TopoPrimer: The Missing Topological Context in Forecasting Models

ResearchDGX agent

arXiv:2605.15035v1 Announce Type: new Abstract: We introduce TopoPrimer, a framework that makes the global topological structure of the series population an explicit input to any forecasting model. To

Training-Free Generative Sampling via Moment-Matched Score Smoothing

SafetyDGX agent

arXiv:2605.14276v1 Announce Type: cross Abstract: Diffusion models generate samples by denoising along the score of a perturbed target distribution. In practice, one trains a neural diffusion model, w

Training ML Models with Predictable Failures

SafetyDGX agent

arXiv:2605.15134v1 Announce Type: new Abstract: Estimating how often an ML model will fail at deployment scale is central to pre-deployment safety assessment, but a feasible evaluation set is rarely l

Unbiased and Second-Order-Free Training for High-Dimensional PDEs

SafetyDGX agent

arXiv:2605.14643v1 Announce Type: new Abstract: Deep learning methods based on backward stochastic differential equations (BSDEs) have emerged as competitive alternatives to physics-informed neural ne

Uncovering Trajectory and Topological Signatures in Multimodal Pediatric Sleep Embeddings

ResearchDGX agent

arXiv:2605.14156v1 Announce Type: new Abstract: While generative models have shown promise in pediatric sleep analysis, the latent structure of their multimodal embeddings remains poorly understood. T

Unsupervised simulation of incompressible flows with physics- and equality- constrained artificial neural networks

ResearchDGX agent

arXiv:2511.18820v2 Announce Type: replace-cross Abstract: Physics-informed neural networks (PINNs) have shown promise for solving partial differential equations, yet their success in simulating incomp

Vendor-Conditioned Contrastive Learning for Predicting Organizational Cyber Threat Targets

Model ReleasesDGX agent

arXiv:2012.14425v2 Announce Type: replace-cross Abstract: Cyberattacks cause billions of dollars in damage annually, with malicious hackers often sharing exploit code and techniques on underground for

Vision-LLMs for Spatiotemporal Traffic Forecasting

SafetyDGX agent

arXiv:2510.11282v2 Announce Type: replace Abstract: Accurate spatiotemporal traffic forecasting is a critical prerequisite for proactive resource management in dense urban mobile networks. While large

Wahkon: A Statistically Principled Deep RKHS Superposition Network

ResearchDGX agent

arXiv:2605.14041v1 Announce Type: cross Abstract: Deep learning excels at prediction but often lacks finite-sample guarantees and calibrated uncertainty; RKHS (Reproducing Kernel Hilbert Space)-based

Watch your neighbors: Training statistically accurate chaotic systems with local phase space information

Local AiDGX agent

arXiv:2605.14405v1 Announce Type: new Abstract: Chaotic systems pose fundamental challenges for data-driven dynamics discovery, as small modeling errors lead to exponentially growing trajectory discre

What if Tomorrow is the World Cup Final? Counterfactual Time Series Forecasting with Textual Conditions

ApplicationsDGX agent

arXiv:2605.14422v1 Announce Type: new Abstract: Time series forecasting has become increasingly critical in real-world scenarios, where future sequences are influenced not only by historical patterns

When Are Two Networks the Same? Tensor Similarity for Mechanistic Interpretability

ResearchDGX agent

arXiv:2605.15183v1 Announce Type: new Abstract: Mechanistic interpretability aims to break models into meaningful parts; verifying that two such parts implement the same computation is a prerequisite.

Winning Lottery Tickets in Neural Networks via a Quantum-Inspired Classical Algorithm

ResearchDGX agent

arXiv:2605.13979v1 Announce Type: cross Abstract: Quantum machine learning (QML) aims to accelerate machine learning tasks by exploiting quantum computation. Previous work studied a QML algorithm for

Woodelf++: A Fast and Unified Partial Dependence Plot Algorithm for Decision Tree Ensembles

HardwareDGX agent

arXiv:2605.14578v1 Announce Type: new Abstract: Partial Dependence Plots (PDPs) visualize how changes in a single feature affect the average model prediction. They are widely used in practice to inter

XAI and Statistical Analysis for Reliable Intrusion Detection in the UAVIDS-2025 Dataset: From Tree to Hybrid and Tabular DNN Ensembles

ResearchDGX agent

arXiv:2605.13922v1 Announce Type: cross Abstract: During the last few years, the term Mechanistic Interpretability, a specific area, under the umbrella of explainable artificial intelligence (XAI), ha

14 May 2026

A Faster Generalized Two-Stage Approximate Top-K

ResearchDGX agent

arXiv:2506.04165v3 Announce Type: replace Abstract: We consider the Top-K selection problem, which aims to identify the largest K elements in an array. Top-K selection arises in many machine learning

A Five-Layer MLOps Architecture for Connected Automated Driving

SafetyDGX agent

arXiv:2605.12719v1 Announce Type: cross Abstract: The continual assurance of safety and performance of automated driving systems (ADSs) poses significant challenges. ADSs operate in complex, dynamic,

A Hybrid Tucker-LSTM Tensor Network Model for SOC Prediction in Electric Vehicles

ApplicationsDGX agent

arXiv:2605.13200v1 Announce Type: new Abstract: Accurate state of charge estimation is critical for the success of electric vehicle battery management strategies, but it is well known that conventiona

A Resampling-Based Framework for Network Structure Learning in High-Dimensional Data

ResearchDGX agent

arXiv:2605.12706v1 Announce Type: new Abstract: RSNet is an open-source R package that provides a resampling-based framework for robust and interpretable network inference, designed to address the lim

A Unified Framework for Critical Scaling of Inverse Temperature in Self-Attention

ResearchDGX agent

arXiv:2605.12697v1 Announce Type: cross Abstract: Length-dependent logit rescaling is widely used to stabilize long-context self-attention, but existing analyses and methods suggest conflicting invers

A Unified Three-Stage Machine Learning Framework for Diabetes Detection, Subtype Discrimination, and Cognitive-Metabolic Hypothesis Testing

ApplicationsDGX agent

arXiv:2605.13464v1 Announce Type: new Abstract: Diabetes mellitus affects over 537 million adults worldwide and remains a major challenge in preventive healthcare. Existing machine-learning studies pr

Accelerating Particle-based Energetic Variational Inference

ResearchDGX agent

arXiv:2504.03158v2 Announce Type: replace-cross Abstract: In this work, we propose a new particle-based variational inference (ParVI) method for accelerating the Energetic Variational Inference with I

Achieving epsilon^{-2} Sample Complexity for Single-Loop Actor-Critic under Minimal Assumptions

SafetyDGX agent

arXiv:2605.13639v1 Announce Type: new Abstract: In this paper, we establish last-iterate convergence rates for off-policy actor--critic methods in reinforcement learning. In particular, under a single

Adam-SHANG: A Convergent Adam-Type Method for Stochastic Smooth Convex Optimization

SafetyDGX agent

arXiv:2605.12878v1 Announce Type: cross Abstract: We propose Adam-SHANG, a Lyapunov-guided Adam-type method that couples momentum, adaptive preconditioning, and a curvature-aware correction through a

Adaptive Kernel Density Estimation with Pre-training

Local AiDGX agent

arXiv:2605.13092v1 Announce Type: cross Abstract: Density estimation in high-dimensional settings is an important and challenging statistical problem.Traditional methods based on kernel smoothing are

AdaptNC: Adaptive Nonconformity Scores for Conformal Prediction under Distribution Shift

SafetyDGX agent

arXiv:2602.01629v2 Announce Type: replace Abstract: Rigorous uncertainty quantification is essential for the safe deployment of autonomous systems in unconstrained environments. Conformal Prediction (

Addressing Finite-Horizon MDPs via Low-Rank Tensor Value Approximation

SafetyDGX agent

arXiv:2501.10598v3 Announce Type: replace Abstract: We study the problem of learning optimal policies in finite-horizon Markov Decision Processes (MDPs) using low-rank reinforcement learning (RL) meth

AGOP as Explanation: From Feature Learning to Per-Sample Attribution in Image Classifiers

Model ReleasesDGX agent

arXiv:2605.12816v1 Announce Type: new Abstract: The Average Gradient Outer Product (AGOP) governs feature learning in neural networks: the Neural Feature Ansatz states that weight Gram matrices at eac

Amortized Neural Clustering of Time Series based on Statistical Features

ResearchDGX agent

arXiv:2605.13128v1 Announce Type: cross Abstract: This paper introduces an algorithm-agnostic approach to feature-based time series clustering via amortized neural inference. By training neural networ

ASAP: Amortized Doubly-Stochastic Attention via Sliced Dual Projection

Model ReleasesDGX agent

arXiv:2605.12879v1 Announce Type: new Abstract: Doubly-stochastic attention has emerged as a transport-based alternative to row-softmax attention, with recent Transformer variants using it to reduce a

Asynchronous Reasoning: Training-Free Interactive Thinking LLMs

SafetyDGX agent

arXiv:2512.10931v3 Announce Type: replace Abstract: Many state-of-the-art LLMs are trained to think before giving their answer. Reasoning can greatly improve language model capabilities, but it also m

Attention Once Is All You Need: Efficient Streaming Inference with Stateful Transformers

Model ReleasesDGX agent

arXiv:2605.13784v1 Announce Type: new Abstract: Conventional transformer inference engines are request-driven, paying an O(n) prefill cost on every query. In streaming workloads, where data arrives co

Backdoor Channels Hidden in Latent Space: Cryptographic Undetectability in Modern Neural Networks

ResearchDGX agent

arXiv:2605.13214v1 Announce Type: cross Abstract: Recent cryptographic results establish that neural networks can be backdoored such that no efficient algorithm can distinguish them from a clean model

Bayesian Nonparametric Mixed-Effect ODEs with Gaussian Processes

ResearchDGX agent

arXiv:2605.13088v1 Announce Type: new Abstract: Dynamical modelling is central to many scientific domains, including pharmacometrics, systems biology, physiology, and epidemiology. In these settings,

Before the Last Token: Diagnosing Final-Token Safety Probe Failures

SafetyDGX agent

arXiv:2605.12726v1 Announce Type: new Abstract: Final-token safety probes monitor a single hidden state after prompt prefill, but jailbreak prompts can contain probe-visible unsafe evidence distribute

Benchmarking Attribute Discrimination in Infant-Scale Vision-Language Models

Model ReleasesDGX agent

arXiv:2512.18951v3 Announce Type: replace Abstract: Infants learn not only object categories but also fine-grained visual attributes such as color, size, and texture from limited experience. Prior inf

Beyond Explained Variance: A Cautionary Tale of PCA

ResearchDGX agent

arXiv:2605.13520v1 Announce Type: cross Abstract: We address shortcomings of principal component analysis (PCA) for visualizing high-dimensional data lying on a nonlinear low-dimensional manifold via

← Previous
1…156157158159160…243
Next →