AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,557 results
14 May 2026

MaskPro: Linear-Space Probabilistic Learning for Strict (N:M)-Sparsity on LLMs

SafetyDGX agent

arXiv:2506.12876v2 Announce Type: replace Abstract: The rapid scaling of large language models~(LLMs) has made inference efficiency a primary bottleneck in the practical deployment. To address this, s

Mechanistic Evidence for Spectral Structures in Prior-Data Fitted Networks

Model ReleasesDGX agent

arXiv:2601.21731v2 Announce Type: replace Abstract: Prior-Data Fitted Networks (PFNs) enable amortized Bayesian inference in a single forward pass, yet their internal representations remain opaque. It

MIDST Challenge at SaTML 2025: Membership Inference over Diffusion-models-based Synthetic Tabular data

ResearchDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2603.19185v2 Announce Type: replace Abstract: Synthetic data is often perceived as a silver-bullet solution to data anonymization and privacy-preserving data publishing. Drawn from generative mo

MILM: Large Language Models for Multimodal Irregular Time Series with Informative Sampling

ApplicationsDGX agent

arXiv:2605.13711v1 Announce Type: new Abstract: Multimodal irregular time series (MITS) consist of asynchronous and irregularly sampled observations from heterogeneous numerical and textual channels.

Min-Max Optimization Requires Exponentially Many Queries

ResearchDGX agent

arXiv:2605.13806v1 Announce Type: cross Abstract: We study the query complexity of min-max optimization of a nonconvex-nonconcave function f over [0,1]^d imes [0,1]^d. We show that, given oracle acces

Mix, Don't Tune: Bilingual Pre-Training Outperforms Hyperparameter Search in Data-Constrained Settings

ResearchDGX agent

arXiv:2605.13225v1 Announce Type: new Abstract: For most languages of the world, language model pre-training operates in a data-constrained regime where models must repeat their training data many tim

Mixed neural posterior estimation for simulators with discrete and continuous parameters

Model ReleasesDGX agent

arXiv:2605.13551v1 Announce Type: new Abstract: Neural Posterior Estimation (NPE) enables rapid parameter inference for complex simulators with intractable likelihoods. NPE trains an inference network

MPINeuralODE: Multiple-Initial-Condition Physics-Informed Neural ODEs for Globally Consistent Dynamical System Learning

ResearchDGX agent

arXiv:2605.13305v1 Announce Type: new Abstract: Neural ordinary differential equations (Neural ODEs) often fit training trajectories while generalizing poorly to unseen initial conditions and long hor

Multi-Armed Sampling Problem and the End of Exploration

Model ReleasesDGX agent

arXiv:2507.10797v2 Announce Type: replace Abstract: This paper introduces the framework of multi-armed sampling, which serves as the sampling counterpart to the optimization problem of multi-armed ban

Multi-Objective and Mixed-Reward Reinforcement Learning via Reward-Decorrelated Policy Optimization

SafetyDGX agent

arXiv:2605.13641v1 Announce Type: new Abstract: Complex reinforcement learning environments frequently employ multi-task and mixed-reward formulations. In these settings, heterogeneous reward distribu

Multimodal Graph-based Classification of Esophageal Motility Disorders

TutorialsDGX agent

arXiv:2605.13623v1 Announce Type: new Abstract: Diagnosing esophageal motility disorders pose significant challenges due to the complexity of high-resolution impedance manometry (HRIM) data and variab

Multitask Multimodal Fusion with Tabular Foundation Models for Peak and Durability Prediction of Pertussis Booster Response

ResearchDGX agent

arXiv:2605.12852v1 Announce Type: new Abstract: Pertussis booster vaccination produces immune responses that vary widely across individuals in both peak magnitude and long-term durability. These two p

Neurodata Without Boredom: Benchmarking Agentic AI for Data Reuse

AgentsDGX agent

arXiv:2605.12808v1 Announce Type: new Abstract: Neuroscience data are highly fragmented across labs, formats, and experimental paradigms, and reuse often requires substantial manual effort. A persiste

NeuroRisk: Physics-Informed Neural Optimization for Risk-Aware Traffic Engineering

SafetyDGX agent

arXiv:2605.12862v1 Announce Type: cross Abstract: In production Wide-Area Networks (WANs), correlated failures dominate availability losses, forcing operators to reserve large safety margins that leav

OceanCBM: A Concept Bottleneck Model for Mechanistic Interpretability in Ocean Forecasting

TutorialsDGX agent

arXiv:2605.12639v1 Announce Type: new Abstract: Extreme ocean phenomena are challenging not only to predict but to diagnose, as accurate forecasts alone do not reveal the underlying physical drivers.

Offline Two-Player Zero-Sum Markov Games with KL Regularization

ResearchDGX agent

arXiv:2605.13025v1 Announce Type: new Abstract: We study the problem of learning Nash equilibria in offline two-player zero-sum Markov games. While existing approaches often rely on explicit pessimism

On Privacy-Preserving Image Transmission in Low-Altitude Networks: A Swin Transformer-Based Framework with Federated Learning

ResearchDGX agent

arXiv:2605.12566v1 Announce Type: cross Abstract: The rapid development of low-altitude economy has driven the proliferation of Unmanned Aerial Vehicle (UAV) applications, including logistics, inspect

On the Generalization of Knowledge Distillation: An Information-Theoretic View

SafetyDGX agent

arXiv:2605.13143v1 Announce Type: cross Abstract: Knowledge distillation is widely used to improve generalization in practice, yet its theoretical understanding remains elusive. In the standard distil

On the Limits of Latent Reuse in Diffusion Models

ResearchDGX agent

arXiv:2605.13448v1 Announce Type: cross Abstract: Diffusion models are often trained in low-dimensional latent spaces, which are then reused for related but shifted datasets. In this work, we study wh

Online Conformal Prediction: Enforcing monotonicity via Online Optimization

ApplicationsDGX agent

arXiv:2605.12668v1 Announce Type: cross Abstract: Conformal prediction provides a principled framework for uncertainty quantification with finite-sample coverage guarantees. While recent work has exte

OSDN: Improving Delta Rule with Provable Online Preconditioning in Linear Attention

Model ReleasesDGX agent

arXiv:2605.13473v1 Announce Type: new Abstract: Linear attention and state-space models offer constant-memory alternatives to softmax attention, but often struggle with in-context associative recall.

PaMM: Periodic Motif Memory for Atomistic Models with an Explicit Local-Structure Interface

Model ReleasesDGX agent

arXiv:2605.13297v1 Announce Type: new Abstract: Periodic crystals repeatedly instantiate similar local coordination motifs across translated cells and chemically related structures, but current equiva

Parallel Scan Recurrent Neural Quantum States for Scalable Variational Monte Carlo

Model ReleasesDGX agent

arXiv:2605.13807v1 Announce Type: cross Abstract: Neural-network quantum states have emerged as a powerful variational framework for quantum many-body systems, with recent progress often driven by mas

Partial Optimality in the Preordering Problem

ResearchDGX agent

arXiv:2602.17346v2 Announce Type: replace-cross Abstract: Preordering is a generalization of clustering and partial ordering with applications in bioinformatics and social network analysis. Given a fi

Path-independent Flow Matching for Multi-parameter Generative Dynamics

Model ReleasesDGX agent

arXiv:2605.13487v1 Announce Type: new Abstract: Flow Matching is a powerful framework for learning transport maps between probability distributions. Yet its standard single-parameter formulation is no

Perceptrons and localization of attention's mean-field landscape

ResearchDGX agent

arXiv:2601.21366v2 Announce Type: replace Abstract: The forward pass of a Transformer can be seen as an interacting particle system on the unit sphere: time plays the role of layers, particles that of

Phasor Memory Networks: Stable Backpropagation Through Time for Scalable Explicit Memory

Model ReleasesDGX agent

arXiv:2605.13370v1 Announce Type: new Abstract: For over a decade, explicit memory architectures like the Neural Turing Machine have remained theoretically appealing yet practically intractable for la

Physics Guided Generative Optimization for Trotter Suzuki Decomposition

ResearchDGX agent

arXiv:2605.13268v1 Announce Type: cross Abstract: Product formulas for Trotter Suzuki simulation remain a practical route to Hamiltonian evolution on noisy intermediate scale quantum (NISQ) hardware,

Physics-informed neural particle flow for the Bayesian update step

ResearchDGX agent

arXiv:2602.23089v2 Announce Type: replace Abstract: The Bayesian update step poses significant computational challenges in high-dimensional nonlinear estimation. While log-homotopy particle flow filte

Pitfalls of Unlabeled Disagreement-Based Drift Detection in Streaming Tree Ensembles

Model ReleasesDGX agent

arXiv:2605.12803v1 Announce Type: new Abstract: Detecting concept drift in high-speed data streams remains challenging, particularly when models must operate on unlabeled data and avoid false alarms c

Plug-In Classification of Drift Functions in Diffusion Processes Using Neural Networks

ResearchDGX agent

arXiv:2602.02791v2 Announce Type: replace-cross Abstract: We study supervised multiclass classification for diffusion processes, where each class is characterized by a distinct drift function and traj

Polyhedral Instability Governs Regret in Online Learning

ResearchDGX agent

arXiv:2605.13692v1 Announce Type: new Abstract: Many online decision problems over combinatorial actions are addressed via convex relaxations, leading to online convex optimization with piecewise line

Population Risk Bounds for Kolmogorov-Arnold Networks Trained by DP-SGD with Correlated Noise

ResearchDGX agent

arXiv:2605.12648v1 Announce Type: new Abstract: We establish the first population risk bounds for Kolmogorov-Arnold Networks (KANs) trained by mini-batch SGD with gradient clipping, covering non-priva

Posterior Bayesian Neural Networks with Dependent Weights

ResearchDGX agent

arXiv:2507.22095v5 Announce Type: replace-cross Abstract: We consider fully connected and feedforward deep neural networks with dependent and possibly heavy-tailed weights, as introduced in [26], to a

Potential and challenges of generative adversarial networks for super-resolution in 4D Flow MRI

ResearchDGX agent

arXiv:2508.14950v2 Announce Type: replace-cross Abstract: 4D Flow Magnetic Resonance Imaging (4D Flow MRI) enables non-invasive quantification of blood flow and hemodynamic parameters. However, its cl

Pragmatic Curiosity: A Unified Framework for Hybrid Learning and Optimization via Active Inference

ResearchDGX agent

arXiv:2602.06104v2 Announce Type: replace Abstract: Many engineering and scientific workflows rely on expensive black-box evaluations, requiring sequential decisions that must both improve task perfor

Predict-Project-Renoise: Sampling Diffusion Models under Hard Constraints

ResearchDGX agent

arXiv:2601.21033v2 Announce Type: replace Abstract: Diffusion models cannot enforce hard constraints, yet applications in the physical sciences demand exact satisfaction of conservation laws, boundary

Predicting Channel Closures in the Lightning Network with Machine Learning

Model ReleasesDGX agent

arXiv:2605.12759v1 Announce Type: new Abstract: The Lightning Network (LN) is a second-layer protocol for Bitcoin designed to enable fast and cost-efficient off-chain transactions. Channels in the LN

Pretraining Language Models with Subword Regularization: An Empirical Study of BPE Dropout in Low-Resource NLP

SafetyDGX agent

arXiv:2605.13436v1 Announce Type: cross Abstract: Subword regularization methods such as BPE dropout are typically applied only during fine-tuning, while pretraining is usually done with deterministic

Probabilistic Prediction Markets with Intermittent Contributions

TutorialsDGX agent

arXiv:2510.13385v3 Announce Type: replace Abstract: Although both data availability and the demand for accurate forecasts are increasing, collaboration between stakeholders is often constrained by dat

Profit Maximization in Bilateral Trade against a Smooth Adversary

ResearchDGX agent

arXiv:2605.12664v1 Announce Type: cross Abstract: Bilateral trade models the task of intermediating between two strategic agents, a seller and a buyer, who wish to trade a good. We study this problem

Progressively Sampled Equality-Constrained Optimization

ApplicationsDGX agent

arXiv:2510.00417v2 Announce Type: replace-cross Abstract: An algorithm is proposed, analyzed, and tested for solving continuous nonlinear-equality-constrained optimization problems where the objective

Protein Circuit Tracing via Cross-layer Transcoders

TutorialsDGX agent

arXiv:2602.12026v2 Announce Type: replace Abstract: Protein language models (pLMs) have emerged as powerful predictors of protein structure and function. However, the computational circuits underlying

Provable Quantization with Randomized Hadamard Transform

ResearchDGX agent

arXiv:2605.13810v1 Announce Type: new Abstract: Vector quantization via random projection followed by scalar quantization is a fundamental primitive in machine learning, with applications ranging from

Provably avoiding over-optimization in Direct Preference Optimization without knowing the data distribution

SafetyDGX agent

arXiv:2602.06239v2 Announce Type: replace Abstract: We introduce PEPO (Pessimistic Ensemble based Preference Optimization), a single-step Direct Preference Optimization (DPO)-like algorithm to mitigat

Proximal-Based Generative Modeling for Bayesian Inverse Problems

SafetyDGX agent

arXiv:2605.13278v1 Announce Type: cross Abstract: Score-based diffusion models demonstrate superior performance in generative tasks but encounter fundamental bottlenecks in inverse problems due to the

QSMOTE-PGM/kPGM: QSMOTE Based PGM and kPGM for Imbalanced Dataset Classification

ResearchDGX agent

arXiv:2512.16960v2 Announce Type: replace Abstract: Quantum-inspired machine learning (QiML) employs mathematical principles from quantum theory, such as Hilbert-space representations and quantum stat

Quantifying Potential Observation Missingness in Inverse Reinforcement Learning

ApplicationsDGX agent

arXiv:2605.12831v1 Announce Type: new Abstract: Inverse reinforcement learning (IRL), which infers reward functions from demonstrations, is a valuable tool for modeling and understanding decision-maki

Re-evaluating Minimum Bayes Risk Decoding for Automatic Speech Recognition

ResearchDGX agent

arXiv:2510.19471v2 Announce Type: replace-cross Abstract: Recent work has shown that sample-based Minimum Bayes Risk (MBR) decoding outperforms beam search in text-to-text generation tasks, such as ma

Real-World Challenges in Fake News Detection: Dealing with Posts by Cold Users

ApplicationsDGX agent

arXiv:2605.12511v1 Announce Type: cross Abstract: Social media serves as a primary source of information in the current digital era. Many people consume a vast range of information in a very short spa

Recurrent Transformer-Based Near- and Far-Field THz Wideband Channel Estimation for UM-MIMO

ResearchDGX agent

arXiv:2605.12578v1 Announce Type: cross Abstract: The integration of terahertz communications and ultra-massive multiple-input multiple-output (UM-MIMO) systems in 6G networks is motivated by their ab

Reducing cross-sample prediction churn in scientific machine learning

Model ReleasesDGX agent

arXiv:2605.13826v1 Announce Type: new Abstract: Scientific machine learning reports predictive performance. It does not report whether the same prediction would survive a different draw of training da

Reframing preprocessing selection as model-internal calibration in near-infrared spectroscopy: A large-scale benchmark of operator-adaptive PLS and Ridge models

Model ReleasesDGX agent

arXiv:2605.13587v1 Announce Type: cross Abstract: Near-infrared spectroscopy (NIRS) is rapid and non-destructive, but reliable calibration still depends heavily on spectral preprocessing. In routine p

Reinforced Collaboration in Multi-Agent Flow Networks

AgentsDGX agent

arXiv:2605.12943v1 Announce Type: new Abstract: Multi-agent systems provide a powerful way to extend large language models (LLMs) by decomposing a complex task into specialized subtasks handled by dif

Rescaled Asynchronous SGD: Optimal Distributed Optimization under Data and System Heterogeneity

Local AiDGX agent

arXiv:2605.13434v1 Announce Type: new Abstract: Asynchronous stochastic gradient descent (ASGD) is a standard way to exploit heterogeneous compute resources in distributed learning: instead of forcing

Rethinking Generalization in Graph Neural Networks: A Structural Complexity Perspective

Model ReleasesDGX agent

arXiv:2605.13597v1 Announce Type: new Abstract: Graph neural networks (GNNs) have emerged as a fundamental tool for learning from graph-structured data, achieving strong performance across a wide rang

Revisiting DAgger in the Era of LLM-Agents

SafetyDGX agent

arXiv:2605.12913v1 Announce Type: new Abstract: Long-horizon LM agents learn from multi-turn interaction, where a single early mistake can alter the subsequent state distribution and derail the whole

Reward-Weighted On-Policy Distillation with an Open Property-Equivalence Verifier for NL-to-SVA Generation

SafetyDGX agent

arXiv:2605.13501v1 Announce Type: cross Abstract: LLM-based generation of SystemVerilog Assertions (SVA) is often reported as nearing saturation, with the strongest specialized model reaching {sim}76%

RMNP: Row-Momentum Normalized Preconditioning for Scalable Matrix-Based Optimization

ResearchDGX agent

arXiv:2603.20527v3 Announce Type: replace Abstract: Preconditioned adaptive methods have gained significant attention for training deep neural networks, as they capture rich curvature information of t

Robust Sequential Experimental Design for A/B Testing

ApplicationsDGX agent

arXiv:2605.12899v1 Announce Type: cross Abstract: Experimental design has emerged as a powerful approach for improving the sample efficiency of A/B testing, yet existing designs rely critically on cor

← Previous
1…159160161162163…243
Next →