AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,432 results
11 May 2026

Quotient Semivalues for False-Name-Resistant Data Attribution

Model ReleasesDGX agent

arXiv:2605.07663v1 Announce Type: cross Abstract: Data valuation methods allocate payments and audit training data's contribution to machine-learning pipelines; however, they often assume passive cont

Regret-Oracle Complexity Tradeoffs in Agnostic Online Learning

ResearchDGX agent

arXiv:2605.07155v1 Announce Type: new Abstract: Agnostic online learning is classically solved via a reduction to the realizable setting, utilizing Littlestone's Standard Optimal Algorithm (SOA) as a

Reinforcement Learning for Exponential Utility: Algorithms and Convergence in Discounted MDPs

SafetyDGX agent

arXiv:2605.08053v1 Announce Type: new Abstract: Reinforcement learning (RL) for exponential-utility optimization in discounted Markov decision processes (MDPs) lacks principled value-based algorithms.


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

RelAgent: LLM Agents as Data Scientists for Relational Learning

AgentsDGX agent

arXiv:2605.07840v1 Announce Type: new Abstract: Relational learning is a challenging problem that has motivated a wide range of approaches, including graph-based models (e.g., graph neural networks, g

Relay Buffer Independent Communication over Pooled HBM for Efficient MoE Inference on Ascend

ResearchDGX agent

arXiv:2605.06055v2 Announce Type: replace-cross Abstract: Mixture-of-Experts (MoE) inference requires large-scale token exchange across devices, making dispatch and combine major bottlenecks in both p

Response Time Enhances Alignment with Heterogeneous Preferences

SafetyDGX agent

arXiv:2605.06987v1 Announce Type: new Abstract: Aligning large language models (LLMs) to human preferences typically relies on aggregating pooled feedback into a single reward model. However, this sta

Retrieval from Within: An Intrinsic Capability of Attention-Based Models

ResearchDGX agent

arXiv:2605.05806v2 Announce Type: replace Abstract: Retrieval-augmented generation (RAG) typically treats retrieval and generation as separate systems. We ask whether an attention-based encoder-decode

RIDER: 3D RNA Inverse Design with Reinforcement Learning-Guided Diffusion

SafetyDGX agent

arXiv:2602.16548v2 Announce Type: replace Abstract: The inverse design of RNA three-dimensional (3D) structures is crucial for engineering functional RNAs in synthetic biology and therapeutics. While

Risk-Consistent Multiclass Learning from Random Label-Subset Membership Queries

SafetyDGX agent

arXiv:2605.07413v1 Announce Type: new Abstract: Obtaining accurate class labels is often costly or unreliable, and may also be limited by privacy or other practical conditions. Compared with asking an

RNAGenScape: Property-Guided, Optimized Generation of mRNA Sequences with Manifold Langevin Dynamics

Local AiDGX agent

arXiv:2510.24736v3 Announce Type: replace-cross Abstract: Generating property-optimized mRNA sequences is central to applications such as vaccine design and protein replacement therapy, but remains ch

Robust and Reliable AI for Predictive Quality in Semiconductor Materials Manufacturing with MLOps and Uncertainty Quantification

ApplicationsDGX agent

arXiv:2605.07752v1 Announce Type: new Abstract: Semiconductor materials manufacturing presents unique challenges for machine learning deployment due to evolving process conditions, equipment degradati

Robust stochastic first order methods in heavy-tailed noise via medoid mini-batch gradient sampling

ResearchDGX agent

arXiv:2605.07634v1 Announce Type: cross Abstract: We consider a first order stochastic optimization framework where, at each iteration, K independent identically distributed (i.i.d.) data point sample

Robust Sublinear Convergence Rates for Iterative Bregman Projections

Model ReleasesDGX agent

arXiv:2602.01372v2 Announce Type: replace-cross Abstract: Entropic regularization provides a simple way to approximate linear programs whose constraints split into two or more tractable blocks. The re

Robustness of Refugee-Matching Gains to Off-Policy Evaluation Choices

SafetyDGX agent

arXiv:2605.06686v1 Announce Type: new Abstract: Previous research has investigated the potential of refugee matching for boosting refugee outcomes, first considered by Bansak et al. (2018). This paper

Rollback-Free Stable Brick Structures Generation

SafetyDGX agent

arXiv:2605.06947v1 Announce Type: new Abstract: While autoregressive models have advanced 3D generation, creating physically stable brick structures remains a challenge due to the strict requirements

Sample Complexity of Stochastic Optimization with Integer Variables

ResearchDGX agent

arXiv:2605.07239v1 Announce Type: new Abstract: We establish sample complexity results for stochastic optimization over the integers, especially with a view to understand the complexity with respect t

Scalable Equilibrium Propagation via Intermediate Error Signals for Deep Convolutional CRNNs

ResearchDGX agent

arXiv:2508.15989v2 Announce Type: replace Abstract: Equilibrium Propagation (EP) is a biologically inspired local learning rule first proposed for convergent recurrent neural networks (CRNNs), in whic

Scaling Categorical Flow Maps

Model ReleasesDGX agent

arXiv:2605.07820v1 Announce Type: new Abstract: Continuous diffusion and flow matching models could represent a powerful alternative to autoregressive approaches for language modelling (LM), as they u

Self Driving Datasets: From 20 Million Papers to Nuanced Biomedical Knowledge at Scale

Local AiDGX agent

arXiv:2605.07022v1 Announce Type: new Abstract: Manually curated biomedical repositories -- spanning bioactivity, genomics, and chemistry -- are expensive to maintain, lag behind primary literature, a

Self-Play Enhancement via Advantage-Weighted Refinement in Online Federated LLM Fine-Tuning with Real-Time Feedback

Model ReleasesDGX agent

arXiv:2605.07977v1 Announce Type: new Abstract: Recent works have advanced feedback-based learning systems, whereby a foundation model is able to intake incoming feedback (e.g., a user) to self-improv

Semantic State Abstraction Interfaces for LLM-Augmented Portfolio Decisions: Multi-Axis News Decomposition and RL Diagnostics

ResearchDGX agent

arXiv:2605.06730v1 Announce Type: new Abstract: We introduce Semantic State Abstraction Interfaces (SSAI): a methodological template for mapping sparse unstructured text into K auditable, named coordi

Semiparametric Efficient Test for Interpretable Distributional Treatment Effects

ResearchDGX agent

arXiv:2605.08034v1 Announce Type: cross Abstract: Distributional treatment effects can be invisible to means: a treatment may preserve average outcomes while changing tails, modes, dispersion, or rare

SGD for Variational Inference: Tackling Unbounded Variance via Preconditioning and Dynamic Batching

ResearchDGX agent

arXiv:2605.07531v1 Announce Type: new Abstract: Black-Box Variational Inference (BBVI) typically relies on Stochastic Gradient Descent (SGD) to optimize the Evidence Lower Bound (ELBO). However, the s

SHARP: A Self-Evolving Human-Auditable Rubric Policy for Financial Trading Agents

SafetyDGX agent

arXiv:2605.06822v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly deployed for autonomous financial trading, a domain requiring continuous adaptation to noisy, non-stationa

Simple KNN-Based Outlier Detection Achieves Robust Clustering

ApplicationsDGX agent

arXiv:2605.07130v1 Announce Type: new Abstract: Being robust to the presence of outliers is crucial for applying clustering algorithms in practice. In the extit{robust k-Means} problem (i.e., k-Means

SLOPE: Optimistic Potential Landscape Shaping for Model-based Reinforcement Learning

TutorialsDGX agent

arXiv:2602.03201v3 Announce Type: replace Abstract: Model-based reinforcement learning (MBRL) is sample-efficient but struggles in sparse reward settings. A critical bottleneck arises from the lack of

Slowly Annealed Langevin Dynamics: Theory and Applications to Training-Free Guided Generation

SafetyDGX agent

arXiv:2605.07950v1 Announce Type: new Abstract: We study Slowly Annealed Langevin Dynamics (SALD), a sampler for tracking a path of moving target distributions and approximating the terminal target th

SMT-Based Active Learning of Weighted Automata

ResearchDGX agent

arXiv:2605.07758v1 Announce Type: cross Abstract: We present an SMT-based active learning algorithm for nondeterministic weighted automata (WFAs) as a practical and robust alternative to Hankel/L*-sty

SOCKET: SOft Collision Kernel EsTimator for Sparse Attention

HardwareDGX agent

arXiv:2602.06283v2 Announce Type: replace Abstract: Exploiting sparsity during long-context inference is key to scaling large language models, as attention dominates the cost of autoregressive decodin

Solving Max-Cut to Global Optimality via Feasibility-Preserving Graph Neural Networks

ResearchDGX agent

arXiv:2605.07113v1 Announce Type: new Abstract: Exact solution of hard combinatorial optimization problems often relies on strong convex relaxations, but solving these relaxations repeatedly inside a

Sparse Attention as a Range Searching Problem: Towards an Inference-Efficient Index for KV Cache

HardwareDGX agent

arXiv:2605.06763v1 Announce Type: new Abstract: Sparse attention improves LLM inference efficiency by selecting a subset of key-value entries, but at the cost of potential accuracy degradation. In par

Sparse Attention as Compact Kernel Regression

ResearchDGX agent

arXiv:2601.22766v3 Announce Type: replace Abstract: Recent work has revealed a link between self-attention mechanisms in transformers and test-time kernel regression via the Nadaraya-Watson estimator,

Sparse Random-Feature Neural Networks with Krylov-Based SVD for Singularly Perturbed ODE

Model ReleasesDGX agent

arXiv:2605.07286v1 Announce Type: cross Abstract: Random-feature neural networks (RFNNs), including architectures with fixed hidden layers and analytically determined output weights, offer fast traini

Spectrum-Adaptive Generalization Bounds for Trained Deep Transformers

ResearchDGX agent

arXiv:2605.07297v1 Announce Type: cross Abstract: Understanding why trained Transformers generalize well is a fundamental problem in modern machine learning theory, and complexity-based generalization

Star Elastic: Many-in-One Reasoning LLMs with Efficient Budget Control

Model ReleasesDGX agent

arXiv:2605.07182v1 Announce Type: new Abstract: Training a family of large language models (LLMs), either from scratch or via iterative compression, is prohibitively expensive and inefficient, requiri

Stationary Reweighting Yields Local Convergence of Soft Fitted Q-Iteration

SafetyDGX agent

arXiv:2512.23927v2 Announce Type: replace-cross Abstract: Fitted Q-iteration (FQI) and soft FQI are widely used value-based methods for offline reinforcement learning, but their standard stability gua

STEPS: A Temporal Smooth Error Propagation Solver on the Manifolds for Test-Time Adaptation in Time Series Forecasting

ResearchDGX agent

arXiv:2605.08005v1 Announce Type: new Abstract: Test-Time Adaptation (TTA) aims to improve time series forecasting under distribution shifts by using limited observations revealed during inference. Ho

Streaming Adversarial Robustness in Fuzzy ARTMAP: Mechanism-Aligned Evaluation, Progressive Training, and Interpretable Diagnostics

ResearchDGX agent

arXiv:2605.06902v1 Announce Type: new Abstract: Adversarial robustness has been studied extensively for offline deep networks, but less is known about strict single-pass streaming neural learners. Thi

StreamPhy: Streaming Inference of High-Dimensional Physical Dynamics via State Space Models

ResearchDGX agent

arXiv:2605.07384v1 Announce Type: new Abstract: Inferring the evolution of high-dimensional and multi-modal (e.g., spatio-temporal) physical fields from irregular sparse measurements in real time is a

Structure learning of Hamiltonians from real-time evolution

TutorialsDGX agent

arXiv:2405.00082v4 Announce Type: replace-cross Abstract: We study the problem of Hamiltonian structure learning from real-time evolution: given the ability to apply e^{-i Ht} for an unknown local Ham

Structured Coupling for Flow Matching

TutorialsDGX agent

arXiv:2605.07676v1 Announce Type: new Abstract: Standard flow matching scales well but typically relies on an unstructured source distribution, limiting its ability to learn interpretable latent struc

Structured Prototype-Guided Adaptation for EEG Foundation Models

Model ReleasesDGX agent

arXiv:2602.17251v2 Announce Type: replace Abstract: Electroencephalography (EEG) foundation models (EFMs) have shown strong potential for transferable representation learning, yet their adaptation in

Susceptibilities and Patterning: A Primer on Linear Response in Bayesian Learning

Local AiDGX agent

arXiv:2605.07980v1 Announce Type: new Abstract: These notes introduce the theory of susceptibilities as developed in [arXiv:2504.18274, arXiv:2601.12703] for interpreting neural networks. The suscepti

SWaRL: Safeguard Code Watermarking via Reinforcement Learning

ResearchDGX agent

arXiv:2601.02602v2 Announce Type: replace-cross Abstract: We present SWaRL, a robust and fidelity-preserving watermarking framework designed to protect the intellectual property of code LLMs by embedd

Synergistic Benefits of Joint Molecule Generation and Property Prediction

ResearchDGX agent

arXiv:2504.16559v3 Announce Type: replace Abstract: Modeling the joint distribution of data samples and their properties allows to construct a single model for both data generation and property predic

Target-Aware Data Augmentation for SAT Prediction

Model ReleasesDGX agent

arXiv:2605.06931v1 Announce Type: new Abstract: Learning-based approaches to NP-hard problems have shown increasing promise, but their progress is fundamentally constrained by the high cost of generat

Temporal Attention for Adaptive Control of Euler-Lagrange Systems with Unobservable Memory

SafetyDGX agent

arXiv:2605.06877v1 Announce Type: new Abstract: Adaptive control of Euler-Lagrange systems is challenging when friction is governed by a finite-horizon internal state that is not directly observable f

Tessellations of Semi-Discrete Flow Matching

ResearchDGX agent

arXiv:2605.07513v1 Announce Type: new Abstract: We study Flow Matching in a semi-discrete setting where a Gaussian source is transported toward a discrete target supported on finitely many points. Thi

Test-Time Compositional Generalization in Diffusion Models via Concept Discovery

Local AiDGX agent

arXiv:2605.07078v1 Announce Type: new Abstract: Compositional generalization requires models to produce novel configurations from familiar parts. In diffusion models, prior compositional generation me

Testing Noise Assumptions of Learning Algorithms

ResearchDGX agent

arXiv:2501.09189v3 Announce Type: replace Abstract: We pose a fundamental question in computational learning theory: can we efficiently test whether a training set satisfies the assumptions of a given

The Convergence Gap: Instruction-Tuned Language Models Stabilize Later in the Forward Pass

Model ReleasesDGX agent

arXiv:2605.07282v1 Announce Type: new Abstract: Final outputs hide when a checkpoint commits to its next-token prediction. We introduce the convergence gap, a model-diffing diagnostic that decodes eac

The Coupling Tax: How Shared Token Budgets Undermine Visible Chain-of-Thought Under Fixed Output Limits

Model ReleasesDGX agent

arXiv:2605.07686v1 Announce Type: new Abstract: Chain-of-thought reasoning is often treated as a monotone way to improve language-model accuracy by letting a model think longer. We identify a counterv

The Minimax Rate of Second-Order Calibration

ResearchDGX agent

arXiv:2605.07808v1 Announce Type: new Abstract: We characterize the minimax rate of estimating the second-order calibration error for binary classification, which quantifies whether a higher-order pre

Theory of Optimal Learning Rate Schedules and Scaling Laws for a Random Feature Model

Model ReleasesDGX agent

arXiv:2602.04774v2 Announce Type: replace-cross Abstract: Setting the learning rate (LR) for a deep learning model is a critical part of successful training. Choosing LRs is often done empirically wit

ThinKV: Thought-Adaptive KV Cache Compression for Efficient Reasoning Models

Model ReleasesDGX agent

arXiv:2510.01290v2 Announce Type: replace Abstract: The long-output context generation of large reasoning models enables extended chain of thought (CoT) but also drives rapid growth of the key-value (

Toward Better Geometric Representations for Molecule Generative Models

SafetyDGX agent

arXiv:2605.07693v1 Announce Type: new Abstract: Geometric representation-conditioned molecule generation provides an effective paradigm that decouples molecule representation modeling from structure g

TRACE: Transport Alignment Conformal Prediction via Diffusion and Flow Matching Models

SafetyDGX agent

arXiv:2605.07100v1 Announce Type: cross Abstract: Constructing valid and informative conformal prediction regions for multi-dimensional outputs remains a fundamental challenge. While conformal predict

Training-Induced Escape from Token Clustering in a Mean-Field Formulation of Transformers

Model ReleasesDGX agent

arXiv:2605.07772v1 Announce Type: new Abstract: Transformers perform inference by iteratively transforming token representations across layers. This layerwise computation has been studied empirically,

Transfer Learning Across Fast- and Full-Simulation Domains in High-Energy Physics

TutorialsDGX agent

arXiv:2605.07471v1 Announce Type: new Abstract: Machine-learning models in high-energy physics are often trained on simulated data, where fully simulated samples are computationally expensive while fa

Transformer-Based Wildlife Species Classification from Daily Movement Trajectories

ResearchDGX agent

arXiv:2605.06726v1 Announce Type: new Abstract: Inferring the identity of wildlife species from daily movement data alone is a challenging task. We train sequence models on large-scale, 7-species GPS

← Previous
1…177178179180181…241
Next →