AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,557 results
1 Jun 2026

Approximation and learning of anisotropic and mixed smooth functions by deep ReLU neural networks

TutorialsDGX agent

arXiv:2605.31152v1 Announce Type: cross Abstract: This paper studies how efficiently deep ReLU neural networks can approximate and learn smooth functions. When the error is measured in L^p([0,1]^d) no

Assign and Add: A Mechanistic Study of Compositional Arithmetic

ResearchDGX agent

arXiv:2605.31497v1 Announce Type: new Abstract: Large language models are able to compose skills in order to perform complex tasks, many of which might not have been seen during training. The details

Asymptotically Optimal Sequential Testing with Markovian Data

ResearchDGX agent

arXiv:2602.17587v2 Announce Type: replace-cross Abstract: We study one-sided and alpha-correct sequential hypothesis testing for data generated by an ergodic, finite-state Markov chain. The null hypot


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Attention-based optimizer for symmetry finding

ResearchDGX agent

arXiv:2605.30429v1 Announce Type: cross Abstract: Finding symmetries is crucial for understanding physical models. In this work, we present an optimization framework that searches Pauli symmetries of

Augmented Lagrangian Predictive Coding

Local AiDGX agent

arXiv:2605.31022v1 Announce Type: new Abstract: Predictive coding (PC) is a local-learning alternative to backpropagation (BP), training deep networks via local energy-minimization dynamics rather tha

Automating Formal Verification with Reinforcement Learning and Recursive Inference

Model ReleasesDGX agent

arXiv:2605.30914v1 Announce Type: new Abstract: Automated formal verification remains challenging for large language models because data for proof assistants and verification-aware languages is scarce

Balanced LoRA: Removing Parameter Invariance to Accelerate Convergence

Model ReleasesDGX agent

arXiv:2605.31484v1 Announce Type: new Abstract: Low-Rank Adaptation (LoRA) is the most widely adopted method for fine-tuning large language models. Notably, LoRA is inherently overparameterized: multi

Bandwidth Allocation with Device Partitioning for Federated Learning over Industrial IoT networks

Model ReleasesDGX agent

arXiv:2605.30892v1 Announce Type: new Abstract: We consider a federated learning (FL) system in which Industrial Internet-of-Things (IIoT) devices collaboratively train a global model over wireless ch

BAT: Better Audio Transformer Guided by Convex Gated Probing

TutorialsDGX agent

arXiv:2602.16305v2 Announce Type: replace-cross Abstract: Probing is widely adopted in computer vision to faithfully evaluate self-supervised learning (SSL) embeddings, as finetuning may misrepresent

Batched Stochastic Linear Bandits with 1-Bit Communication Constraints

AgentsDGX agent

arXiv:2605.30976v1 Announce Type: cross Abstract: We study stochastic linear bandits under a natural combination of batching and communication constraints: the time horizon is partitioned into batches

Bayesian Inference with Shaped Deep Non-linear MLPs

ResearchDGX agent

arXiv:2605.30860v1 Announce Type: cross Abstract: A central aim of deep learning theory is to characterize how neural networks make predictions in the regime of simultaneously large model and training

Bayesian Rain Field Reconstruction using Commercial Microwave Links and Diffusion Model Priors

ApplicationsDGX agent

arXiv:2605.05520v2 Announce Type: replace Abstract: Commercial Microwave Links (CMLs) offer dense spatial coverage for rainfall sensing but produce path-integrated measurements that make accurate grou

Benchmarking Uncertainty and its Disentanglement in multi-label Chest X-Ray Classification

Model ReleasesDGX agent

arXiv:2508.04457v2 Announce Type: replace-cross Abstract: Reliable uncertainty quantification is crucial for trustworthy decision-making and the deployment of AI models in medical imaging. While prior

Best-Arm Identification-Based Trust Region Selection for Bayesian Optimization on Multimodal Functions

Local AiDGX agent

arXiv:2605.31050v1 Announce Type: new Abstract: Gaussian process-based Bayesian optimization (BO) is a popular approach for expensive black-box optimization, but its performance often degrades on comp

Beyond Additive Decompositions: Interpretability Through Separability

ResearchDGX agent

arXiv:2605.31200v1 Announce Type: new Abstract: Interpretable machine learning requires models that are accurate and structurally faithful to the data.Existing explainability methods rely heavily on a

Beyond ReLU: Bifurcation, Oversmoothing, and Topological Priors

Model ReleasesDGX agent

arXiv:2602.15634v2 Announce Type: replace Abstract: Graph Neural Networks (GNNs) learn node representations through iterative network-based message-passing. While powerful, deep GNNs suffer from overs

Beyond Test-Time Memory: State-Space Optimal Control for LLM Reasoning

HardwareDGX agent

arXiv:2603.09221v2 Announce Type: replace Abstract: Associative memory has long underpinned the design of sequential models. Beyond recall, humans reason by projecting future states and selecting goal

Beyond Tokens: Enhancing RTL Quality Estimation via Structural Graph Learning

ResearchDGX agent

arXiv:2508.18730v2 Announce Type: replace Abstract: Estimating the quality of register transfer level (RTL) designs is crucial in the electronic design automation (EDA) workflow, as it enables instant

Bifurcated Remaining Useful Life Prediction: A Hybrid Approach for Realistic Uncertainty Characterization

ResearchDGX agent

arXiv:2605.31241v1 Announce Type: new Abstract: This study presents a novel hybrid prognostic framework for uncertainty-aware Remaining Useful Life (RUL) estimation in turbofan engines using the NASA

BOKBO (Best of K Bad Options): Calibrated Abstention for VLA Policies

Model ReleasesDGX agent

arXiv:2605.30660v1 Announce Type: new Abstract: Test-time scaling for vision-language-action (VLA) policies, methods such as RoboMonkey, SEAL, MG-Select, and V-GPS, samples K candidate action chunks a

Bridging the Gap Between Natural Language and Market Dynamics via High-Dimensional Representation Learning

ResearchDGX agent

arXiv:2605.30652v1 Announce Type: new Abstract: Traditional multi-modal financial forecasting often relies on scalar sentiment scores, which fail to capture the nuances of financial news. To address t

CacheProbe: Auditing Prompt Cache Isolation in Gateway APIs

ResearchDGX agent

arXiv:2605.30613v1 Announce Type: cross Abstract: Over the past year, prompt caching in Large Language Models (LLMs) has become increasingly more popular across inference APIs. Prompt caching helps sa

Can Subgraph Explanations Be Weaponized to Steal Graph Neural Networks?

Model ReleasesDGX agent

arXiv:2605.30470v1 Announce Type: new Abstract: Graph Machine Learning as a Service (GMLaaS) platforms increasingly implement explainability interfaces to meet regulatory transparency requirements. Ho

Causal Evaluation of Membership Inference Attacks

SafetyDGX agent

arXiv:2602.02819v3 Announce Type: replace Abstract: Membership Inference Attacks (MIAs) aim to distinguish training points (members) from unseen data (non-members), and are widely used to quantify mem

CellBRIDGE: Learning Cellular Trajectories via Interaction-Aware Alignment

SafetyDGX agent

arXiv:2605.30635v1 Announce Type: new Abstract: Inferring dynamics from population snapshots is a fundamental challenge in machine learning and biology. In scRNA-sequencing (scRNA-seq), destructive me

Chain-of-Thought and Compressed Looped Transformers: A Memory-Budget Separation

ResearchDGX agent

arXiv:2605.30757v1 Announce Type: new Abstract: Chain-of-thought prompting and looped Transformers both give a fixed model more test-time computation, but they differ in what they remember. Chain-of-t

Chem-PerturBridge: a harmonized compendium of small molecule perturbation transcriptomic effects

ResearchDGX agent

arXiv:2605.31522v1 Announce Type: new Abstract: Large perturbation models require training data encompassing chemical, cellular, and assay diversity. Current transcriptomic resources for small-molecul

CoMem: Context Management with A Decoupled Long-Context Model

AgentsDGX agent

arXiv:2605.30842v1 Announce Type: new Abstract: Context management enables agentic models to solve long-horizon tasks through iterative summarization of previous interaction histories. However, this p

Conformal C2ST: Turning weak classifiers into strong two-sample tests

ResearchDGX agent

arXiv:2507.17026v2 Announce Type: replace-cross Abstract: The two-sample testing problem, a fundamental task in statistics and machine learning, seeks to determine whether two sets of samples, drawn f

Conformal Reliability: A New Evaluation Metric for Conditional Generation

ResearchDGX agent

arXiv:2605.30807v1 Announce Type: new Abstract: Conditional generative models have recently achieved remarkable success in various applications. However, a suitable metric for evaluating the reliabili

Constrained Flow Optimization via Sequential Fine Tuning for Molecular Design

TutorialsDGX agent

arXiv:2605.30610v1 Announce Type: new Abstract: Adapting generative foundation models, in particular diffusion and flow models, to optimize given reward functions (e.g., binding affinity) while satisf

Constrained Multi-Objective Reinforcement Learning with Max-Min Criterion

SafetyDGX agent

arXiv:2605.31388v1 Announce Type: new Abstract: Multi-Objective Reinforcement Learning (MORL) extends standard RL by optimizing policies with respect to multiple, often conflicting, objectives. While

Contextual Scalarisation Thompson Sampling for multi-objective decisions in public media

SafetyDGX agent

arXiv:2605.31291v1 Announce Type: cross Abstract: Recommender systems may operate under multiple, competing objectives. For example, audience reach, cultural values, public service mandate, and operat

Convergence of Steepest Descent and Adam under Non-Uniform Smoothness

Model ReleasesDGX agent

arXiv:2605.30648v1 Announce Type: new Abstract: Recent work has analyzed the convergence of first-order methods under non-uniform smoothness assumptions that better model the loss landscape in machine

Convergence of Two-Timescale Markovian Stochastic Approximations with Applications in Reinforcement Learning

Model ReleasesDGX agent

arXiv:2605.31172v1 Announce Type: new Abstract: This work studies the convergence of two-timescale stochastic approximations (SA), a class of iterative algorithms that update two sets of parameters in

Cost-aware Stopping for Bayesian Optimization

SafetyDGX agent

arXiv:2507.12453v5 Announce Type: replace Abstract: In automated machine learning, scientific discovery, and other applications of Bayesian optimization, deciding when to stop evaluating expensive bla

Cross-Layer Subspace Coupling for LLM Compression: A Unifying Framework and Its Empirical Limits

ResearchDGX agent

arXiv:2605.30836v1 Announce Type: new Abstract: Recent SVD based compression methods for large language models like SVD LLM and Basis Sharing can be unified under one optimization problem. While mathe

Data-driven Progressive Discovery of Physical Laws

ResearchDGX agent

arXiv:2603.13727v2 Announce Type: replace Abstract: Symbolic regression is a powerful tool for knowledge discovery, enabling the extraction of interpretable mathematical expressions directly from data

Deep-learning-based low-energy trigger algorithms for the Hyper-Kamiokande experiment

HardwareDGX agent

arXiv:2605.31391v1 Announce Type: cross Abstract: Modern machine learning techniques have become increasingly important in particle physics because of their powerful pattern-recognition capabilities,

Delayed Momentum Aggregation: Communication-efficient Byzantine-robust Federated Learning with Partial Participation

ResearchDGX agent

arXiv:2509.02970v3 Announce Type: replace Abstract: Partial participation is essential for communication-efficient federated learning at scale, yet existing Byzantine-robust methods typically assume f

Density-Guided Robust Counterfactual Explanations on Tabular Data under Model Multiplicity

Local AiDGX agent

arXiv:2605.30901v1 Announce Type: new Abstract: Counterfactual explanations (CEs) are essential for actionable recourse, yet their reliability is often compromised in low-density regions, where classi

Destruction is a General Strategy to Learn Generation; Diffusion's Strength is to Take it Seriously; Exploration is the Future

TutorialsDGX agent

arXiv:2605.30553v1 Announce Type: new Abstract: I present diffusion models as part of a family of machine learning techniques that withhold information from a model's input and train it to guess the w

DG-CoLearn: An Efficient Collaborative Learning Framework for Dynamic Graphs

ResearchDGX agent

arXiv:2605.31427v1 Announce Type: new Abstract: Dynamic graph learning (DGL) is essential for modelling evolving graph data, but existing methods suffer from significant computational overhead due to

dgMARK: Decoding-Guided Watermarking for Diffusion Language Models

ResearchDGX agent

arXiv:2601.22985v2 Announce Type: replace Abstract: We propose dgMARK, a decoding-guided watermarking method for discrete diffusion language models (dLLMs). Unlike autoregressive models, dLLMs can gen

Diffusion Models Preferentially Memorize Prototypical Examples or: Why Does My Diffusion Model Love Slop?

ApplicationsDGX agent

arXiv:2605.30642v1 Announce Type: new Abstract: Generative models have a persistent limitation: their tendency to memorize training data can create legal liabilities and erode creative diversity. Unde

DisasterLex: An Expert Concept-to-Schema Knowledge Graph for Geospatial Reasoning in Disaster Analytics

ResearchDGX agent

arXiv:2605.30538v1 Announce Type: new Abstract: Disasters are inevitable and increasingly costly, and effective response depends on querying structured tabular data: precise, information-dense records

Discovering a Zeta Map Algorithm on Dyck Paths via Mechanistic Interpretability

ResearchDGX agent

arXiv:2605.30482v1 Announce Type: new Abstract: Machine learning is increasingly used in mathematical discovery, but in mathematics the desired output is often not a prediction itself, but an explicit

Discovering Thermodynamically Admissible Dissipation Potentials via Grammar-Based Symbolic Regression

ResearchDGX agent

arXiv:2605.31532v1 Announce Type: cross Abstract: Constitutive laws for inelastic materials must satisfy strict thermodynamic admissibility requirements, yet current data-driven approaches sacrifice i

DisjunctiveNet: Neural Symbolic Learning via Differentiable Convexified Optimization Layers

ApplicationsDGX agent

arXiv:2605.30456v1 Announce Type: new Abstract: Many learning tasks in science and engineering are characterized by sparse datasets, which limits the effectiveness of purely data-driven approaches. At

Diving into Kronecker Adapters: Component Design Matters

Model ReleasesDGX agent

arXiv:2602.01267v2 Announce Type: replace Abstract: Kronecker adapters have emerged as a promising approach for fine-tuning large-scale models, enabling high-rank updates through tunable component str

Do covariates explain why these groups differ? The choice of reference group can reverse conclusions in the Oaxaca-Blinder decomposition

Model ReleasesDGX agent

arXiv:2603.29972v2 Announce Type: replace-cross Abstract: Scientists often want to explain why an outcome is different in two groups. For instance, differences in patient mortality rates across two ho

Don't be so Stief! Learning KV Cache low-rank approximation over the Stiefel manifold

ResearchDGX agent

arXiv:2601.21686v2 Announce Type: replace Abstract: Key-value (KV) caching enables fast autoregressive decoding but at long contexts becomes a dominant bottleneck in High Bandwidth Memory (HBM) capaci

Don't Fool Me Twice: Adapting to Adversity in the Wild with Experience-Driven Reasoning

AgentsDGX agent

arXiv:2605.31119v1 Announce Type: cross Abstract: In robotics, dangers and adversity modes are often embodiment-specific and relative to each agent. A frontier of autonomous mobile robotics is to enab

Dynamical local Frechet curve regression in manifolds

ResearchDGX agent

arXiv:2505.05168v3 Announce Type: replace-cross Abstract: Under mild conditions, this paper derives a least-squares local linear Frechet curve predictor for response and regressor evaluated in a separ

Early Prediction of Future Behavioral Strategy from Process Traces

ResearchDGX agent

arXiv:2605.30550v1 Announce Type: new Abstract: Adaptive systems often need to make task-specific decisions about people from limited evidence: a tutor may need to anticipate how a learner will approa

Effective Biological Representation Learning by Masking Gene Expression

ResearchDGX agent

arXiv:2605.31562v1 Announce Type: new Abstract: RNA sequencing produces rich and diverse datasets of gene expression, offering compelling insights into cellular state and function that have many appli

Efficient and Uncertainty-Aware Diffusion Framework for Offline-to-Online Reinforcement Learning

SafetyDGX agent

arXiv:2605.30776v1 Announce Type: new Abstract: Offline-to-Online Reinforcement Learning (O2O-RL) leverages an offline, pre-trained policy to minimize costly online interactions. Although data-efficie

Eigenvectors of Experts are Training-free Non-collapsing Routers

TutorialsDGX agent

arXiv:2605.30992v1 Announce Type: new Abstract: Sparse Mixture of Experts (SMoE) architectures improve the training efficiency of Large Language Models (LLMs) by routing input tokens to a selected sub

End-to-End Compression for Tabular Foundation Models

Model ReleasesDGX agent

arXiv:2602.05649v2 Announce Type: replace Abstract: The long-standing dominance of gradient-boosted decision trees for tabular data has recently been challenged by in-context learning tabular foundati

Error Amplification Limits ANN-to-SNN Conversion in Continuous Control

Model ReleasesDGX agent

arXiv:2601.21778v2 Announce Type: replace-cross Abstract: Spiking Neural Networks (SNNs) can achieve competitive performance by converting already existing well-trained Artificial Neural Networks (ANN

← Previous
1…110111112113114…243
Next →