AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent
84,433Total entries
1Added by human
84,432Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,557 results
12 May 2026

Kernel-Gradient Drifting Models

ResearchDGX agent

arXiv:2605.10727v1 Announce Type: new Abstract: We propose kernel-gradient drifting, a one-step generative modeling framework that replaces the fixed Euclidean displacement direction in drifting model

Kinetic theory for Transformers and the lost-in-the-middle phenomenon

ResearchDGX agent

arXiv:2605.09213v1 Announce Type: cross Abstract: We study causal self-attention dynamics -- a toy model for decoder Transformers -- which we interpret as a non-exchangeable interacting particle syste

Kintsugi: Learning Policies by Repairing Executable Knowledge Bases

Local AiDGX agent

arXiv:2605.09487v1 Announce Type: new Abstract: Modern embodied agents achieve impressive performance, but their task knowledge is often stored in neural weights, latent state, or prompt-bound memory,


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

LagrangianSplats: Divergence-Free Transport of Gaussian Primitives for Fluid Reconstruction

ApplicationsDGX agent

arXiv:2605.09299v1 Announce Type: cross Abstract: Reconstructing 3D fluid velocity fields from sparse 2D video observations is a highly ill-posed inverse problem, demanding both transport consistency

Lakestream: A Consistent and Brokerless Data Plane for Large Foundation Model Training

ResearchDGX agent

arXiv:2605.09994v1 Announce Type: cross Abstract: Modern Large Foundation Model (LFM) training has transformed the data pipeline from a static ingestion layer into a dynamic component that must co-evo

Laplacian Heads Improve Transformers by Smoothing Token Representations

ResearchDGX agent

arXiv:2602.09297v2 Announce Type: replace Abstract: Transformers update token representations through multi-head attention and residual connections as X leftarrow X + sum_{i} P^{(i)}XW_{V_i}W_{o_i}, w

LAQuant: A Simple Overhead-free Large Reasoning Model Quantization by Layer-wise Lookahead Loss

SafetyDGX agent

arXiv:2605.08755v1 Announce Type: new Abstract: Large reasoning models (LRMs) reach competition-level math and coding accuracy via long autoregressive decoding, making per-token decoding cost a primar

Large Language Models over Networks: Collaborative Intelligence under Resource Constraints

Local AiDGX agent

arXiv:2605.08626v1 Announce Type: cross Abstract: Large language models (LLMs) are transforming society, powering applications from smartphone assistants to autonomous driving. Yet cloud-based LLM ser

Latent Geometry Beyond Search: Amortizing Planning in World Models

Model ReleasesDGX agent

arXiv:2605.08732v1 Announce Type: cross Abstract: Modern vision-based world models can represent observations as compact yet expressive latent manifolds, but fast goal-oriented planning in these space

Layer Collapse in Diffusion Language Models

Model ReleasesDGX agent

arXiv:2605.06366v2 Announce Type: replace Abstract: Diffusion language models (DLMs) have recently emerged as competitive alternatives to autoregressive (AR) language models, yet differences in their

LBI: Parallel Scan Backpropagation via Latent Bounded Interfaces

ResearchDGX agent

arXiv:2605.09204v1 Announce Type: new Abstract: Backpropagation is inherently sequential across depth, creating an O(K)-deep dependency chain that bottlenecks parallel training. While parallel-scan fo

Learnability and Competition in High-Dimensional Multi-Component ICA

ResearchDGX agent

arXiv:2605.08552v1 Announce Type: cross Abstract: Independent Component Analysis (ICA) is a foundational tool for unsupervised representation learning, yet its high-dimensional theory remains largely

Learngene Search Across Multiple Datasets for Building Variable-Sized Models

ResearchDGX agent

arXiv:2605.08209v1 Announce Type: new Abstract: Deep learning methods are widely used under diverse resource constraints, resulting in models of varying sizes, such as the Vision Transformer (ViT) ser

Learning Confidence Ellipsoids and Applications to Robust Subspace Recovery

Model ReleasesDGX agent

arXiv:2512.16875v4 Announce Type: replace-cross Abstract: We study the problem of finding confidence ellipsoids for an arbitrary distribution in high dimensions. Given samples from a distribution D an

Learning from Acceptance: Cumulative Regret in the Game of Coding

TutorialsDGX agent

arXiv:2605.09754v1 Announce Type: cross Abstract: Classical coding-theoretic guarantees often rely on trust assumptions, such as requiring sufficiently many honest nodes compared with adversarial ones

Learning Graph Foundation Models on Riemannian Graph-of-Graphs

ResearchDGX agent

arXiv:2605.09993v1 Announce Type: new Abstract: Graph foundation models (GFMs), pretrained on massive graph data, have transformed graph machine learning by supporting general-purpose reasoning across

Learning Polyhedral Conformal Sets for Robust Optimization

ResearchDGX agent

arXiv:2605.08506v1 Announce Type: new Abstract: Robust optimization (RO) provides a principled framework for decision-making under uncertainty, but its performance critically depends on the choice of

Learning predictive models for combinations of heterogeneous proteomic data sources

ResearchDGX agent

arXiv:2605.08958v1 Announce Type: new Abstract: Multiple technologies that measure expression levels of protein mixtures in the human body offer a potential for detection and understanding the disease

Learning Pure Quantum States in Any Dimension (Almost) Without Regret

TutorialsDGX agent

arXiv:2605.09019v1 Announce Type: cross Abstract: We extend quantum state tomography with minimal cumulative disturbance, first investigated in [arXiv:2406.18370], to arbitrary finite-dimensional pure

Learning Rate Scheduling with Matrix Factorization for Private Training

ResearchDGX agent

arXiv:2511.17994v2 Announce Type: replace Abstract: We study differentially private model training with stochastic gradient descent under learning rate scheduling and correlated noise. Although correl

Learning stochastic multiscale models through normalizing flows

ResearchDGX agent

arXiv:2605.09718v1 Announce Type: cross Abstract: Many systems in physics, engineering, and biology exhibit multiscale stochastic dynamics, where low-dimensional slow variables evolve under the influe

Learning the Channel Gain from Anywhere to Anywhere via Cross-environment Transformer Estimators

AgentsDGX agent

arXiv:2605.08211v1 Announce Type: cross Abstract: Channel-gain maps provide the channel gain between any two locations in a geographical region. They find numerous applications, from resource allocati

Learning Theory of Transformers: Local-to-Global Approximation via Softmax Partition of Unity

ResearchDGX agent

arXiv:2605.08811v1 Announce Type: cross Abstract: This paper investigates the learning theory of Transformer networks for regression tasks on the compact Euclidean domain [0,1]^d and d-dimensional com

Learning to Bid with Unknown Private Values in Budget-Constrained First-Price Auctions

ResearchDGX agent

arXiv:2605.09448v1 Announce Type: new Abstract: The transition to First-Price Auctions (FPA) in digital advertising has spurred significant research, yet existing work typically assumes access to a va

Learning to Compress Time-to-Control: A Reinforcement Learning Framework for Chronic Disease Management

SafetyDGX agent

arXiv:2605.09818v1 Announce Type: new Abstract: Reinforcement learning (RL) in healthcare has had mixed results, with reward sparsity, unreliable off-policy evaluation, and deployment-simulation gap a

Learning to Learn the Macroscopic Fundamental Diagram using Physics-Informed and meta Machine Learning techniques

TutorialsDGX agent

arXiv:2508.14137v2 Announce Type: replace Abstract: The Macroscopic Fundamental Diagram is a popular tool used to describe traffic dynamics in an aggregated way, with applications ranging from traffic

Learning to Sparsify Stochastic Linear Bandits

Model ReleasesDGX agent

arXiv:2605.10151v1 Announce Type: new Abstract: This paper addresses the problem of learning to sparsify stochastic linear bandits, where a decision-maker sequentially selects actions from a high-dime

Learning When to Stop: Selective Imitation Learning Under Arbitrary Dynamics Shift

SafetyDGX agent

arXiv:2605.09183v1 Announce Type: new Abstract: Behavior cloning provides strong imitation learning guarantees when training and test environments share the same dynamics. However, in many deployment

Learning When to Trust LLM Priors: A Validated Framework for Semantic Prior Integration

ResearchDGX agent

arXiv:2601.21410v3 Announce Type: replace-cross Abstract: Large language models (LLMs) encode rich semantic knowledge that can be useful for supervised learning, but their outputs are unreliable as st

Lecture Notes on Statistical Physics and Neural Networks

ResearchDGX agent

arXiv:2605.06394v1 Announce Type: cross Abstract: These lecture notes introduce some topics of classical statistical physics, particularly those that are relevant for neural networks and deep learning

Less is More: Towards Simple Graph Contrastive Learning

ResearchDGX agent

arXiv:2509.25742v4 Announce Type: replace Abstract: Graph Contrastive Learning (GCL) has shown strong promise for unsupervised graph representation learning, yet its effectiveness on heterophilic grap

Likelihood scoring for continuations of mathematical text: a self-supervised benchmark with tests for shortcut vulnerabilities

Model ReleasesDGX agent

arXiv:2605.10810v1 Announce Type: new Abstract: We introduce an automatically generated benchmark for predicting hidden text in technical papers. A paper supplies visible context X and a hidden contin

LiLAW: Lightweight Learnable Adaptive Weighting to Learn Sample Difficulty & Improve Noisy Training

SafetyDGX agent

arXiv:2509.20786v3 Announce Type: replace Abstract: Training deep neural networks with noise and data heterogeneity is a major challenge. We introduce Lightweight Learnable Adaptive Weighting (LiLAW),

Liouville PDE-based sliced-Wasserstein flow

SafetyDGX agent

arXiv:2505.17204v3 Announce Type: replace-cross Abstract: The sliced Wasserstein flow (SWF), a nonparametric and implicit generative gradient flow, is transformed into a Liouville partial differential

LLM-Driven Performance-Space Augmentation for Meta-Learning-Based Algorithm Selection

SafetyDGX agent

arXiv:2605.09518v1 Announce Type: new Abstract: Meta-learning for algorithm selection relies on a meta-dataset in which each row corresponds to a supervised learning dataset described by meta-features

LLMs for Secure Hardware Design and Related Problems: Opportunities and Challenges

HardwareDGX agent

arXiv:2605.10807v1 Announce Type: cross Abstract: The integration of Large Language Models (LLMs) into Electronic Design Automation (EDA) and hardware security is rapidly reshaping the semiconductor i

Local LMO: Constrained Gradient Optimization via a Local Linear Minimization Oracle

ResearchDGX agent

arXiv:2605.08850v1 Announce Type: cross Abstract: We design Local LMO - a new projection-free gradient-type method for constrained optimization. The key algorithmic idea is to replace the global linea

Locking Pretrained Weights via Deep Low-Rank Residual Distillation

Model ReleasesDGX agent

arXiv:2605.10777v1 Announce Type: new Abstract: The quality of open-weight language models has dramatically improved in recent years. Sharing weights greatly facilitates model adoption by enabling the

Machine Learning-Based Graph Simplification for Symbolic Accelerators

ResearchDGX agent

arXiv:2605.08996v1 Announce Type: new Abstract: Graph-based accelerators have been widely adopted in symbolic data processing applications such as genomics, cybersecurity, and artificial intelligence.

Make Each Token Count: Towards Improving Long-Context Performance with KV Cache Eviction

SafetyDGX agent

arXiv:2605.09649v1 Announce Type: new Abstract: The key-value (KV) cache is a major bottleneck in long-context inference, where memory and computation grow with sequence length. Existing KV eviction m

Many Needles in a Haystack: Active Hit Discovery for Perturbation Experiments

ResearchDGX agent

arXiv:2605.10196v1 Announce Type: new Abstract: High-throughput gene perturbation experiments can test several genetic interventions in parallel, yet experimental budgets remain limited. A central goa

MARGIN: Margin-Aware Regularized Geometry for Imbalanced Vulnerability Detection

ApplicationsDGX agent

arXiv:2605.10240v1 Announce Type: cross Abstract: Software vulnerability detection is critical for ensuring software security and reliability. Despite recent advances in deep learning, real-world vuln

MASS-DPO: Multi-negative Active Sample Selection for Direct Policy Optimization

SafetyDGX agent

arXiv:2605.10784v1 Announce Type: new Abstract: Multi-negative preference optimization under the Plackett--Luce (PL) model extends Direct Preference Optimization (DPO) by leveraging comparative signal

Matrix Factorization for Practical Continual Mean Estimation Under User-Level Differential Privacy

ResearchDGX agent

arXiv:2601.22320v2 Announce Type: replace Abstract: We study continual mean estimation, where data vectors arrive sequentially and the goal is to maintain accurate estimates of the running mean. We ad

MDL-GBG: A Non-parametric and Interpretable Granular-Ball Generation Method for Clustering

Local AiDGX agent

arXiv:2605.08759v1 Announce Type: new Abstract: Existing granular-ball generation methods are still mainly driven by handcrafted quality measures and heuristic splitting or stopping criteria, which we

Measuring and Decomposing Mode Separation via the Canonical Diffusion

ResearchDGX agent

arXiv:2605.08777v1 Announce Type: cross Abstract: Mode separation, namely how sharply a distribution fragments into barrier-separated clusters, is a fundamental geometric property of densities, diffic

Mechanistic Independence: A Principle for Identifiable Disentangled Representations

ResearchDGX agent

arXiv:2509.22196v2 Announce Type: replace Abstract: Disentangled representations seek to recover latent factors of variation underlying observed data, yet their identifiability is still not fully unde

MEG-XL: Data-Efficient Brain-to-Text via Long-Context Pre-Training

TutorialsDGX agent

arXiv:2602.02494v2 Announce Type: replace Abstract: Clinical brain-to-text interfaces are designed for paralysed patients who cannot provide extensive training recordings. Pre-training improves data-e

Meta-reinforcement learning with minimum attention

SafetyDGX agent

arXiv:2505.16741v4 Announce Type: replace Abstract: Minimum attention applies the least action principle to changes of control concerning state and time, first proposed by Brockett. The involved regul

METBRA25Y: Brazil Surface Meteorology Archive with Harmonized Variables and Quality Control

ResearchDGX agent

arXiv:2605.08701v1 Announce Type: new Abstract: This data paper describes METBRA25Y, a harmonized archive of hourly surface meteorological observations from Brazil derived from public historical recor

Metropolis-Adjusted Diffusion Models

SafetyDGX agent

arXiv:2605.09654v1 Announce Type: cross Abstract: Sampling from score-based diffusion models incurs bias due to both time discretisation and the approximation of the score function. A common strategy

MicroFuse: Protein-to-Genome Expert Fusion for Microbial Operon Reasoning

Model ReleasesDGX agent

arXiv:2605.08815v1 Announce Type: new Abstract: Predicting microbial operon co-membership requires integrating two complementary biological signals: protein-scale molecular identity and genome-context

Minimal Filling Architectures of Polynomial Neural Networks: Counterexamples, Frontier Search, and Defects

ResearchDGX agent

arXiv:2605.09609v1 Announce Type: new Abstract: We provide a counterexample to the minimal unimodal conjecture for polynomial neural networks (PNNs) with power activation functions. Fixing the input a

Mistake-Bounded Language Generation

ResearchDGX agent

arXiv:2605.10809v1 Announce Type: new Abstract: We investigate the learning task of language generation in the limit, but shift focus from the traditional time-of-last-mistake metric of a generator's

Mitigating Data Scarcity in Spaceflight Applications for Offline Reinforcement Learning Using Physics-Informed Deep Generative Models

SafetyDGX agent

arXiv:2604.02438v2 Announce Type: replace Abstract: The deployment of reinforcement learning (RL)-based controllers on physical systems is often limited by poor generalization to real-world scenarios,

Mitigating Membership Inference in Intermediate Representations with Differentially Private Training

ResearchDGX agent

arXiv:2602.22611v2 Announce Type: replace Abstract: In Embedding-as-an-Interface (EaaI) settings, pre-trained models are queried for Intermediate Representations (IRs). The distributional properties o

Mixture of Experts for Recognizing Depression from Interview and Reading Tasks

ResearchDGX agent

arXiv:2502.20213v2 Announce Type: replace Abstract: Depression is a mental disorder and can cause a variety of symptoms, including psychological, physical, and social. Speech has been proved an object

MLS-Bench: A Holistic and Rigorous Assessment of AI Systems on Building Better AI

Model ReleasesDGX agent

arXiv:2605.08678v1 Announce Type: new Abstract: Modern AI progress has been driven by ML methods that are generalizable across settings and scalable to larger regimes. As large language models demonst

Model Capacity Determines Grokking through Competing Memorisation and Generalisation Speeds

Model ReleasesDGX agent

arXiv:2605.09724v1 Announce Type: new Abstract: Existing accounts of grokking explain the phenomena in terms of mechanistic frameworks such as circuit efficiency or lazy-to-rich transitions. However,

Model-Free Neural Filtering: A Comparison with Classical Filters in Nonlinear Systems

Model ReleasesDGX agent

arXiv:2601.21266v3 Announce Type: replace Abstract: Neural network models are increasingly used for state estimation in control and decision-making, yet it remains unclear to what extent they behave a

← Previous
1…171172173174175…243
Next →