AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,432 results
15 Jul 2026

When Directional Accuracy Lies: A Base-Rate-Honest Benchmark for LoRA-Adapted TimesFM on Equity Forecasting

Model ReleasesDGX agent

arXiv:2607.12248v1 Announce Type: cross Abstract: Large pretrained time-series models such as TimesFM are attractive for financial forecasting, but raw directional accuracy is a misleading scoreboard

When Does Reward Teach State? A Hidden-Automaton Instrument and the Group-Language Boundary

SafetyDGX agent

arXiv:2607.11953v1 Announce Type: new Abstract: Does a reinforcement-learning agent that earns high reward represent its task's latent state, or only a reward-correlated shortcut? The question is usua

10 Jul 2026

A law of robustness for two-layer neural networks with arbitrary weights


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model ReleasesDGX agent

arXiv:2607.07778v1 Announce Type: new Abstract: Bubeck, Li and Nagaraj conjectured that, for generic data, any two-layer neural network with m neurons that fits n noisy labels must have Lipschitz cons

A Quantum Reservoir Architecture for Chaotic Forecasting and a Test of Whether Its High Dimension Helps

ResearchDGX agent

arXiv:2607.07978v1 Announce Type: cross Abstract: Quantum reservoir computing uses a fixed quantum circuit as a feature generator and trains only a simple linear readout on top of it. This makes it ch

A Self-Supervised Approach for Minimal-Annotation Hydroacoustic Data Exploration

ResearchDGX agent

arXiv:2607.07733v1 Announce Type: cross Abstract: Passive hydroacoustic monitoring often generates large volumes of continuous recordings that are only partially exploited due to the cost of manual an

A Sliced-Wasserstein Framework on Correlation Matrices for EEG Decoding

ApplicationsDGX agent

arXiv:2606.06104v2 Announce Type: replace Abstract: Electroencephalography (EEG) offers noninvasive, millisecond resolution recordings of neuronal activity and is widely used in neuroscience and healt

An exact information theory of generalization phase transitions in Bayesian diffusion models

Local AiDGX agent

arXiv:2607.08041v1 Announce Type: new Abstract: How diffusion models circumvent the curse of dimensionality to learn complex distributions over high dimensional spaces from a finite training set, inst

An interpretable Good--Turing restart criterion for k-means++

ResearchDGX agent

arXiv:2607.08243v1 Announce Type: new Abstract: The k-means++ algorithm is commonly restarted multiple times to avoid poor local optima, yet the number of restarts is almost always chosen arbitrarily

AutoAnchor: Stable Diffusion Unlearning Using Cross-Attention as a Manifold Surrogate

ResearchDGX agent

arXiv:2607.08337v1 Announce Type: new Abstract: Diffusion unlearning is essential for mitigating the generation of harmful or copyrighted content in text-to-image models. Current diffusion unlearning

BACH: A Bayesian Admixture of Contrastive Heads for Multi-Interest Two-Tower Retrieval

ResearchDGX agent

arXiv:2607.08107v1 Announce Type: cross Abstract: Two-tower retrievers compress each user into a single embedding, limiting their ability to serve diverse interests. Multi-interest models give each us

Bayesian Deep Learning for Discrete Choice

Model ReleasesDGX agent

arXiv:2505.18077v3 Announce Type: replace-cross Abstract: Discrete choice models (DCMs) are used to analyze individual decision-making in contexts such as transportation choices, political elections,

Bayesian Experimental Design via Score Matching

SafetyDGX agent

arXiv:2607.08335v1 Announce Type: cross Abstract: Policy-based approaches to Bayesian experimental design (BED) allow the learning of deep policy networks that adaptively make intelligent design decis

Beyond Backpropagation: Monte Carlo Method Can Train Deep Neural Networks

Model ReleasesDGX agent

arXiv:2607.08406v1 Announce Type: new Abstract: Backpropagation (BP) dominates deep learning training, but its reliance on gradients brings inherent troubles -- vanishing and exploding gradients. The

Beyond Success Rates: Trainability and Extractability for Offline GCRL

SafetyDGX agent

arXiv:2602.05459v2 Announce Type: replace Abstract: Offline goal-conditioned reinforcement learning (GCRL) is typically benchmarked by the best tuned success rate of each method. This score measures a

BiSCo-LLM: Lookup-Free Binary Spherical Coding for Extreme Low-Bit Large Language Model Compression

Local AiDGX agent

arXiv:2607.08643v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly constrained by memory capacity, weight bandwidth, and checkpoint storage during deployment. Existing low-b

CAAD: Causality-Aware Multivariate Time Series Anomaly Detection via Multi-Scale Alignment and Structural Causal Consistency

SafetyDGX agent

arXiv:2607.08555v1 Announce Type: new Abstract: The operational integrity of complex industrial systems relies on precise anomaly detection and diagnosis. The vast majority of existing methods narrowl

Calibrated Stackelberg Games: Learning Optimal Commitments Against Calibrated Agents

AgentsDGX agent

arXiv:2306.02704v2 Announce Type: replace-cross Abstract: We introduce Calibrated Stackelberg Games (CSGs), a generalization of the standard Stackelberg Games (SGs) framework. In CSGs, a principal rep

CASL-VAE: Learning Structured Latent Variables from Unpaired Data for Semi-supervised Clustering and Paired Sample Generation

ResearchDGX agent

arXiv:2607.08254v1 Announce Type: new Abstract: Quantifying variability in a target population relative to a reference population is central to many scientific and clinical problems (e.g., diseased vs

Certified Interventional Fidelity: Anytime-Valid, Adaptive Evaluation of Causal Claims in Mechanistic Interpretability

ResearchDGX agent

arXiv:2607.08349v1 Announce Type: new Abstract: Mechanistic interpretability often evaluates explanations by intervening on a model: swapping hidden states, patching activations, ablating components,

Classifier Chain-based Pathological Test Recommendation

ResearchDGX agent

arXiv:2607.08299v1 Announce Type: new Abstract: Accurate and timely diagnoses are essential for quality patient care. However, delayed recommendation of diagnostic tests and physicians' subjective int

Collate: Collaborative Neural Network Learning for Latency-Critical Edge Systems

Local AiDGX agent

arXiv:2607.08013v1 Announce Type: new Abstract: Federated Learning (FL) empowers multiple clients to collaboratively learn a model, enlarging the training data of each client for high accuracy while p

Communication-Efficient Byzantine-Robust Federated Conformal Prediction via Partial Model Sharing

ResearchDGX agent

arXiv:2602.18396v2 Announce Type: replace Abstract: We propose PRISM-FCP (Partial shaRing and robust calIbration with Statistical Margins for Federated Conformal Prediction), a communication-efficient

Conformal Predictive Programming for Chance Constrained Optimization

ResearchDGX agent

arXiv:2402.07407v3 Announce Type: replace-cross Abstract: We propose conformal predictive programming (CPP), a framework to solve chance constrained optimization problems, i.e., optimization problems

Contrastive Order Learning: A General Framework for Ordinal Regression

ResearchDGX agent

arXiv:2607.08109v1 Announce Type: new Abstract: We propose contrastive order learning (ConOrd), a contrastive learning framework for ordinal regression that integrates the strengths of contrastive lea

Contravariance Theory: Strong Alignment for Minimal Solutions to Hard Tasks

SafetyDGX agent

arXiv:2607.08561v1 Announce Type: new Abstract: A series of results from the NeuroAI over the past fifteen years have raised core questions both about how to compare Deep Neural Network (DNN) models t

Cross-Modal Generative Framework for Signal Translation from Fetal-Maternal Electrocardiograms to Fetal Doppler Waveforms

ResearchDGX agent

arXiv:2607.08073v1 Announce Type: new Abstract: Fetal electrocardiogram (fECG) and Doppler ultrasound provide complementary views of fetal cardiovascular function: fECG captures electrical activity wh

CTA-Pipelining: A Latency-Oriented Spatial Scaling Method for Multi-GPU Systems

HardwareDGX agent

arXiv:2607.07862v1 Announce Type: cross Abstract: The evolution of compute infrastructure has transformed multi-GPU systems into tightly integrated shared-memory structures. However, current software

Deep Learning for Joint Narrowband Interference Cancellation and Soft Demodulation in OFDM Systems

ResearchDGX agent

arXiv:2607.08717v1 Announce Type: new Abstract: Narrowband interference (NBI) severely degrades orthogonal frequency-division multiplexing (OFDM) systems by corrupting subcarriers and rendering classi

DeepPySR -- A Symbolic Regression Framework with Dynamic Pruning, Pareto Selection, and Hierarchical Composition for Real-World Scientific Discovery

ApplicationsDGX agent

arXiv:2607.08150v1 Announce Type: new Abstract: Symbolic regression (SR) discovers analytical equations from data, yielding glass-box models with directly interpretable formulas, unlike black-box meth

DeepSWE: Measuring Frontier Coding Agents on Original, Long-Horizon Engineering Tasks

Model ReleasesDGX agent

arXiv:2607.07946v1 Announce Type: cross Abstract: DeepSWE is a benchmark of 113 original, long-horizon software engineering tasks for evaluating coding agents. Most public agentic coding benchmarks fo

Diagnosing Shape-Prior Shortcuts in Long-Range Single-Shot Fringe Projection Profilometry

Model ReleasesDGX agent

arXiv:2606.17093v2 Announce Type: replace Abstract: Learning-based single-shot fringe projection profilometry (FPP) has been studied almost entirely at close range, and the networks used are evaluated

Distributed Sketching on Data Partitions for OLS Regression

ResearchDGX agent

arXiv:2607.07888v1 Announce Type: new Abstract: This paper studies distributed sketching for ordinary least squares (OLS) regression, an approach that distributes small sketches of a large data set ov

Distributionally Faithful Imputation via Positive Semi-Definite Kernel Density Estimation

ApplicationsDGX agent

arXiv:2607.07767v1 Announce Type: cross Abstract: Missing values undermine statistical inference and machine learning pipelines, yet most imputation methods rely on heuristics or restrictive parametri

Douglas-Rachford Splitting for Group-Sparse Feedback Linear-Quadratic Control

ResearchDGX agent

arXiv:2507.19895v4 Announce Type: replace-cross Abstract: In this paper, we study the distributed linear quadratic problem with fixed communication topology (DFT-LQ) and the sparse feedback linear qua

Dropping Just a Handful of Preferences Can Change Top Large Language Model Rankings

ResearchDGX agent

arXiv:2508.11847v4 Announce Type: replace-cross Abstract: We propose a method for evaluating the robustness of widely used LLM ranking systems -- variants of a Bradley--Terry model -- to dropping a wo

Dynamics of Gradient Descent with Large Step Size Near a Manifold of Flat Minima

ResearchDGX agent

arXiv:2607.08380v1 Announce Type: new Abstract: An important quantity in the theory of gradient descent (GD) is the sharpness, defined as the largest eigenvalue of the objective Hessian. Classical ana

EdgeRefine: Privacy-Utility Balance for Graphs via Jaccard Sampling under Edge Differential Privacy

ResearchDGX agent

arXiv:2607.08659v1 Announce Type: new Abstract: Graph Neural Networks (GNNs) have shown considerable success in learning from graph-structured data, but their use in privacy-sensitive areas remains di

Efficient Partitioning Method of Large-Scale Public Safety Spatio-Temporal Data based on Information Loss Constraints

SafetyDGX agent

arXiv:2306.12857v3 Announce Type: replace Abstract: The storage, management, and application of massive spatio-temporal data are widely used in practical scenarios, including public safety. However, d

Eigenvalue Calibration for Semantic Embeddings of Large Language Models

ApplicationsDGX agent

arXiv:2607.08377v1 Announce Type: new Abstract: Uncertainty quantification is central to the reliable deployment of large language models (LLMs), and eigenvalues of semantic embeddings have recently e

Evaluating the Generalizability of Foundation Models for Extreme Environmental Events: Case Study of California Wildfire PM2.5

Model ReleasesDGX agent

arXiv:2607.07951v1 Announce Type: new Abstract: Wildfire smoke events produce extreme PM_{2.5} concentrations that pose severe public health risks, yet forecasting rare, hazardous-level spikes remains

Explaining Near-Zero Hessian Eigenvalues Through Approximate Symmetries in Neural Networks

Local AiDGX agent

arXiv:2607.07845v1 Announce Type: new Abstract: The Hessian of the training loss governs the local geometry of the loss landscape, yet despite existing explanations for its largest eigenvalues, the or

Expressivity and Statistical Trade-offs in Diffusion Policy Learning

SafetyDGX agent

arXiv:2607.07967v1 Announce Type: cross Abstract: Diffusion-based policies have recently emerged as powerful policy parameterizations for reinforcement learning, representing state-conditioned action

Federated Deep Learning for Privacy-Preserving Cardiovascular Disease Risk Prediction

ApplicationsDGX agent

arXiv:2607.08595v1 Announce Type: new Abstract: Cardiovascular disease risk prediction models often rely on data from a single institution or centrally pooled datasets. Extending these models across i

FPGN: Redefining Ultra-Fast Programmable Gate-based Neural Acceleration with Differentiable LUTs

Local AiDGX agent

arXiv:2607.08427v1 Announce Type: cross Abstract: Achieving nanosecond-scale inference latency for deep neural networks (DNNs) has become a primary architectural concern for latency-critical applicati

Frequency-Domain Multi-Modality Transportation Modeling

TutorialsDGX agent

arXiv:2607.08475v1 Announce Type: new Abstract: Multi-modality transportation refers to urban systems composed of multiple transportation modes, such as traffic flow and public transit, whose dynamics

From Theory to Application: A Practical Introduction to Neural Operators in Scientific Computing

ResearchDGX agent

arXiv:2503.05598v3 Announce Type: replace-cross Abstract: This review examines neural operator architectures for learning solution operators of parametric partial differential equations (PDEs), with a

Functional and Secure Code Generation with Task Vectors

Model ReleasesDGX agent

arXiv:2607.07881v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used for code generation, but they struggle to generate functional code free of security vulnerabilities

Generalization Theory for Through-the-Wall Radar Human Activity Recognition

Model ReleasesDGX agent

arXiv:2607.08144v1 Announce Type: cross Abstract: Through-the-wall radar (TWR) human activity recognition (HAR) is important for non-line-of-sight indoor sensing, security monitoring, and emergency re

Geometry-Aware Deep Congruence Networks for Manifold Learning in Cross-Subject Motor Imagery

SafetyDGX agent

arXiv:2511.18940v3 Announce Type: replace Abstract: Cross-subject motor imagery decoding remains a fundamental challenge in EEG-based brain-computer interfaces due to substantial inter-subject variabi

GradInf: Gradient Estimation as Probabilistic Inference

ApplicationsDGX agent

arXiv:2607.07840v1 Announce Type: cross Abstract: Gradient estimation -- the task of computing the gradient of the expected value of a probabilistic program -- has diverse applications in scientific c

(heta_l, heta_u)-Parametric Multi-Task Optimization: Joint Search in Solution and Infinite Task Spaces

ApplicationsDGX agent

arXiv:2503.08394v5 Announce Type: replace-cross Abstract: Multi-task optimization is typically characterized by a fixed and finite set of tasks. The present paper relaxes this condition by considering

High-Dimensional Procrustes Matching via Tree Counts

ResearchDGX agent

arXiv:2607.08538v1 Announce Type: cross Abstract: Suppose we observe two sets of n Gaussian vectors in R^d, with the promise that, after applying a permutation of [n] and a rotation of R^d, the two se

Image classification via a quantum-inspired strategy involving a mixture of experts

HardwareDGX agent

arXiv:2607.07754v1 Announce Type: new Abstract: Pattern recognition problems arise in a variety of physical image processing situations, and convolutional neural networks are a popular scheme for the

Improving RCT-Based Treatment Effect Estimation Under Covariate Mismatch via Calibrated Alignment

SafetyDGX agent

arXiv:2603.19186v3 Announce Type: replace Abstract: Randomized controlled trials (RCTs) are the gold standard for estimating treatment effects, yet they are often underpowered for detecting effect het

ImputeViz: A Visual Analytics Dashboard for Diagnosing Missing Data and Comparing Imputation Methods

ResearchDGX agent

arXiv:2607.08579v1 Announce Type: cross Abstract: Missing data is a persistent obstacle in scientific, social science, and public health research, often biasing analyses and placing accountability on

Instance Generation for Patient-to-room Assignment and Admission Scheduling Based on Real Hospital Data

ResearchDGX agent

arXiv:2507.03423v2 Announce Type: replace-cross Abstract: Developing algorithms for real-life problems that perform well in practice depends on the availability of realistic data for testing. Obtainin

Joint Bayesian Parameter and Model Order Estimation for Low-Rank Probability Mass Tensors

Model ReleasesDGX agent

arXiv:2410.06329v4 Announce Type: replace-cross Abstract: Obtaining a reliable estimate of the joint probability mass function (PMF) of a set of random variables from observed data is a significant ob

Joint Discrete-Continuous Flow Matching for Open-Vocabulary Inverse Design of Multilayer Optical Coatings

Model ReleasesDGX agent

arXiv:2607.08392v1 Announce Type: cross Abstract: Amortized neural inverse design typically remains closed-world: component choices are fixed vocabulary tokens, coordinate grids are frozen at training

Koopman-informed recurrent neural networks

ApplicationsDGX agent

arXiv:2410.23467v3 Announce Type: replace Abstract: Recurrent neural networks are a successful neural architecture for many time-dependent problems, including time series analysis, forecasting, and mo

KronQ: LLM Quantization via Kronecker-Factored Hessian

Model ReleasesDGX agent

arXiv:2607.07964v1 Announce Type: new Abstract: Post-training quantization (PTQ) is a widely adopted technique for compressing large language models (LLMs) without retraining. Existing second-order PT

← Previous
1…4243444546…241
Next →