AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,557 results
13 May 2026

Online Continual Learning with Dynamic Label Hierarchies

TutorialsDGX agent

arXiv:2605.11742v1 Announce Type: new Abstract: Online Continual Learning (OCL) aims to learn from endless nonext{-}stationary data streams, yet most existing methods assume a flat label space and ove

Online Learning-to-Defer with Varying Experts

ApplicationsDGX agent

arXiv:2605.12340v1 Announce Type: cross Abstract: Learning-to-Defer (L2D) methods route each query either to a predictive model or to external experts. While existing work studies this problem in batc

Operator Spectroscopy of Trained Lattice Samplers

ResearchDGX agent

arXiv:2605.11199v1 Announce Type: cross Abstract: Trained lattice samplers are usually judged by the ensembles they generate. Here we instead analyze the trained field-space function itself: a flow-ma


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Optimal Policy Learning under Budget and Coverage Constraints

SafetyDGX agent

arXiv:2605.12235v1 Announce Type: cross Abstract: We study optimal policy learning under combined budget and minimum coverage constraints. We show that the problem admits a knapsack-type structure and

Optimal Representations for Generalized Contrastive Learning with Imbalanced Datasets

ResearchDGX agent

arXiv:2605.11291v1 Announce Type: new Abstract: In this paper, we provide a computable characterization of the geometry of optimal representations in Contrastive Learning (CL) when the classes are imb

Optimistic Dual Averaging Unifies Modern Optimizers

ResearchDGX agent

arXiv:2605.11172v1 Announce Type: new Abstract: We introduce SODA, a generalization of Optimistic Dual Averaging, which provides a common perspective on state-of-the-art optimizers like Muon, Lion, Ad

Oscillators Are All You Need: Irregular Time Series Modelling via Damped Harmonic Oscillators with Closed-Form Solutions

ResearchDGX agent

arXiv:2602.12139v2 Announce Type: replace Abstract: Transformers excel at time series modelling through attention mechanisms that capture long-term temporal patterns. However, they assume uniform time

OUI as a Structural Observable: Towards an Activation-Centric View of Neural Network Training

ResearchDGX agent

arXiv:2605.11570v1 Announce Type: new Abstract: Activation functions are what make deep networks expressive: without them, the model collapses to a linear map. Yet we still evaluate training mostly fr

OverNaN: NaN-Aware Oversampling for Imbalanced Learning with Meaningful Missingness

SafetyDGX agent

arXiv:2605.11525v1 Announce Type: new Abstract: Missing values are routinely treated as defects to be eliminated through deletion or imputation prior to machine learning. In many applied domains, howe

Overparametrized models with posterior drift

ResearchDGX agent

arXiv:2506.23619v2 Announce Type: replace-cross Abstract: This paper investigates the impact of posterior drift on out-of-sample forecasting accuracy in overparametrized machine learning models. We do

Oversmoothing as Representation Degeneracy in Neural Sheaf Diffusion

SafetyDGX agent

arXiv:2605.11178v1 Announce Type: new Abstract: Neural Sheaf Diffusion (NSD) generalizes diffusion-based Graph Neural Networks by replacing scalar graph Laplacians with sheaf Laplacians whose learned

Overtrained, Not Misaligned

Model ReleasesDGX agent

arXiv:2605.12199v1 Announce Type: new Abstract: Emergent misalignment (EM), where fine-tuning on a narrow task (like insecure code) causes broad misalignment across unrelated domains, was first demons

Partial Model Sharing Improves Byzantine Resilience in Federated Conformal Prediction

ResearchDGX agent

arXiv:2605.11684v1 Announce Type: new Abstract: We propose a Byzantine-resilient federated conformal prediction (FCP) method that leverages partial model sharing, where only a subset of model paramete

Partition Tree: Conditional Density Estimation over General Outcome Spaces

ResearchDGX agent

arXiv:2602.04042v2 Announce Type: replace Abstract: We propose Partition Tree, a novel tree-based framework for conditional density estimation over general outcome spaces that supports both continuous

Patterns behind Chaos: Forecasting Data Movement for Efficient Large-Scale MoE LLM Inference

Model ReleasesDGX agent

arXiv:2510.05497v5 Announce Type: replace-cross Abstract: Large-scale Mixture of Experts (MoE) Large Language Models (LLMs) have recently become the frontier open-weight models, achieving remarkable m

PENEX: AdaBoost-Inspired Neural Network Regularization

ResearchDGX agent

arXiv:2510.02107v4 Announce Type: replace Abstract: AdaBoost sequentially fits so-called weak learners to minimize an exponential loss, which penalizes misclassified data points more severely than oth

Persona-Conditioned Adversarial Prompting: Multi-Identity Red-Teaming for Adversarial Discovery and Mitigation

SafetyDGX agent

arXiv:2605.11730v1 Announce Type: new Abstract: Automated red-teaming for LLMs often discovers narrow attack slices, missing diverse real-world threats, and yielding insufficient data for safety fine-

Physics Aware Neural Networks: Denoising for Magnetic Navigation

ResearchDGX agent

arXiv:2602.13690v2 Announce Type: replace Abstract: Magnetic-anomaly navigation, leveraging small-scale variations in the Earth's magnetic field, is a promising alternative when GPS is unavailable or

Physics-Informed Teacher-Student Ensemble Learning for Traffic State Estimation with a Varying Speed Limit Scenario

Local AiDGX agent

arXiv:2605.11346v1 Announce Type: new Abstract: Physics-informed deep learning (PIDL) neural networks have shown their capability as a useful instrument for transportation practitioners in utilizing t

Pion: A Spectrum-Preserving Optimizer via Orthogonal Equivalence Transformation

ResearchDGX agent

arXiv:2605.12492v1 Announce Type: new Abstract: We introduce Pion, a spectrum-preserving optimizer for large language model (LLM) training based on orthogonal equivalence transformation. Unlike additi

PIVOT: Bridging Planning and Execution in LLM Agents via Trajectory Refinement

AgentsDGX agent

arXiv:2605.11225v1 Announce Type: cross Abstract: Large language model (LLM)-based agents frequently generate seemingly coherent plans that fail upon execution due to infeasible actions, constraint vi

POP: Prior-Fitted First-Order Optimization Policies

Model ReleasesDGX agent

arXiv:2602.15473v2 Announce Type: replace Abstract: Gradient-based optimizers are highly sensitive to design choices in their adaptive learning rate mechanisms. To address this limitation, we introduc

Post-ADC Inference: Valid Inference After Active Data Collection

SafetyDGX agent

arXiv:2605.11511v1 Announce Type: cross Abstract: The validity of statistical inference depends critically on how data are collected. When data gathered through active data collection (ADC) are reused

Posterior Contraction Rates for Sparse Kolmogorov-Arnold Networks in Anisotropic Besov Spaces

Model ReleasesDGX agent

arXiv:2605.11652v1 Announce Type: cross Abstract: We study posterior contraction rates for sparse Bayesian Kolmogorov-Arnold networks (KANs) over anisotropic Besov spaces, providing a statistical foun

Predictive Maps of Multi-Agent Reasoning: A Successor-Representation Spectrum for LLM Communication Topologies

SafetyDGX agent

arXiv:2605.11453v1 Announce Type: cross Abstract: Practitioners deploying multi-agent large language model (LLM) systems must currently choose between communication topologies such as chain, star, mes

Pretraining Strategies and Scaling for ECG Foundation Models: A Systematic Study

ResearchDGX agent

arXiv:2605.12241v1 Announce Type: cross Abstract: Specialized foundation models are beginning to emerge in various medical subdomains, but pretraining methodologies and parametric scaling with the siz

Primal-Dual Policy Optimization for Linear CMDPs with Adversarial Losses

SafetyDGX agent

arXiv:2605.11535v1 Announce Type: new Abstract: Existing work on linear constrained Markov decision processes (CMDPs) has primarily focused on stochastic settings, where the losses and costs are eithe

Principled Latent Diffusion for Graphs via Laplacian Autoencoders

ResearchDGX agent

arXiv:2601.13780v3 Announce Type: replace Abstract: Graph diffusion models achieve state-of-the-art performance in graph generation but suffer from quadratic complexity in the number of nodes -- and m

PriorZero: Bridging Language Priors and World Models for Decision Making

SafetyDGX agent

arXiv:2605.12289v1 Announce Type: new Abstract: Leveraging the rich world knowledge of Large Language Models (LLMs) to enhance Reinforcement Learning (RL) agents offers a promising path toward general

PrivacySIM: Evaluating LLM Simulation of User Privacy Behavior

ApplicationsDGX agent

arXiv:2605.12147v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used to simulate human behavior, but their ability to simulate individual privacy decisions is not well

Probabilistic Computers for Neural Quantum States

Model ReleasesDGX agent

arXiv:2512.24558v2 Announce Type: replace-cross Abstract: Neural quantum states efficiently represent many-body wavefunctions with neural networks, but the cost of Monte Carlo sampling limits their sc

Probabilistic Modeling of Latent Agentic Substructures in Deep Neural Networks

SafetyDGX agent

arXiv:2509.06701v2 Announce Type: replace Abstract: We develop a theory of intelligent agency grounded in probabilistic modeling for neural models. Agents are represented as outcome distributions with

Probing Non-Equilibrium Grain Boundary Dynamics with XPCS and Domain-Adaptive Machine Learning

Model ReleasesDGX agent

arXiv:2605.12194v1 Announce Type: cross Abstract: Grain-boundary (GB) dynamics control the stability, mechanical, and functional response of nanocrystalline materials, but direct experimental access t

Procedural-skill SFT across capacity tiers: A W-Shaped pre-SFT Trajectory and Regime-Asymmetric Mechanism on 0.8B-4B Qwen3.5 Models

Model ReleasesDGX agent

arXiv:2605.11907v1 Announce Type: new Abstract: We measure procedural-skill SFT contribution across three Qwen3.5 dense scales (0.8B, 2B, 4B) on a 200-task / 40-skill holdout, with Claude Haiku 4.5 as

Provably Data-driven Multiple Hyper-parameter Tuning with Structured Loss Function

Model ReleasesDGX agent

arXiv:2602.02406v2 Announce Type: replace-cross Abstract: Data-driven algorithm design automates hyperparameter tuning, but its statistical foundations remain limited because model performance can dep

Pure Exploration Beyond Reward Feedback: The Role of Post-Action Context

ResearchDGX agent

arXiv:2502.03061v2 Announce Type: replace Abstract: We introduce the problem of best arm identification (BAI) with post-action context, a new BAI problem in a stochastic multi-armed bandit environment

QDSB: Quantized Diffusion Schrodinger Bridges

ApplicationsDGX agent

arXiv:2605.11983v1 Announce Type: new Abstract: Learning generative models in settings where the source and target distributions are only specified through unpaired samples is gaining in importance. H

Quantifying the Reconstructability of Astrophysical Methods with Large Language Models and Information Theory: A Case Study in Spectral Reconstruction

ApplicationsDGX agent

arXiv:2605.11154v1 Announce Type: cross Abstract: Modern astrophysical studies rely heavily on complex data analysis pipelines; however, published descriptions often lack the detail required for compu

QuIDE: Mastering the Quantized Intelligence Trade-off via Active Optimization

Model ReleasesDGX agent

arXiv:2605.10959v1 Announce Type: new Abstract: There is currently no unified metric for evaluating the efficiency of quantized neural networks. We propose QuIDE, built around the Intelligence Index I

Quotient-Categorical Representations for Bellman-Compatible Average-Reward Distributional Reinforcement Learning

SafetyDGX agent

arXiv:2605.11289v1 Announce Type: new Abstract: Average-reward reinforcement learning requires estimating the gain and the bias, which is defined only up to an additive constant. This makes direct dis

Random-Set Graph Neural Networks

AgentsDGX agent

arXiv:2605.11987v1 Announce Type: cross Abstract: Uncertainty quantification has become an important factor in understanding the data representations produced by Graph Neural Networks (GNNs). Despite

Rank Is Not Capacity: Spectral Occupancy for Latent Graph Models

Local AiDGX agent

arXiv:2605.11142v1 Announce Type: new Abstract: Graph representation learning has become a standard approach for analyzing networked data, with latent embeddings widely used for link prediction, commu

Read, Extract, Classify: A Tool for Smarter Requirements Engineering

ResearchDGX agent

arXiv:2605.11045v1 Announce Type: cross Abstract: This paper presents the ReXCL tool, which automates the extraction and classification processes in requirements engineering, enhancing the software de

Reconsidering the energy efficiency of spiking neural networks

Model ReleasesDGX agent

arXiv:2409.08290v4 Announce Type: replace-cross Abstract: Spiking Neural Networks (SNNs) promise higher energy efficiency over conventional Quantized Artificial Neural Networks (QNNs) due to their eve

Regret minimization in Linear Bandits with offline data via extended D-optimal exploration

ResearchDGX agent

arXiv:2508.08420v3 Announce Type: replace Abstract: We consider the problem of online regret minimization in linear bandits with access to prior observations (offline data) from the underlying bandit

Rethink the Role of Neural Decoders in Quantum Error Correction

SafetyDGX agent

arXiv:2605.12046v1 Announce Type: cross Abstract: Quantum error correction (QEC) is essential for enabling quantum advantages, with decoding as a central algorithmic primitive. Owing to its importance

Rethinking external validation for the target population: Capturing patient-level similarity with a generative model

SafetyDGX agent

arXiv:2605.11284v1 Announce Type: cross Abstract: Background: External validation is essential for assessing the transportability of predictive models. However, its interpretation is often confounded

Rethinking LLMOps for Fraud and AML: Building a Compliance-Grade LLM Serving Stack

Model ReleasesDGX agent

arXiv:2605.11232v1 Announce Type: cross Abstract: Fraud detection and anti-money-laundering (AML) compliance are high-value domains for large language models (LLMs), but their serving requirements dif

Revisiting GAN with Bayes-Optimal Discrimination

ResearchDGX agent

arXiv:2510.25609v3 Announce Type: replace Abstract: We propose an alternative to the standard GAN training approach, in which the discriminator is a binary classifier trained by cross-entropy to disti

Robust Multi-Agent Path Finding under Observation Attacks: A Principled Adversarial-Plus-Smoothing Training Recipe

SafetyDGX agent

arXiv:2605.11469v1 Announce Type: new Abstract: Decentralized multi-agent path finding (MAPF) routes a team of agents on a shared grid, each acting from its own local view. The standard solution train

Robust Policy Optimization to Prevent Catastrophic Forgetting

SafetyDGX agent

arXiv:2602.08813v2 Announce Type: replace Abstract: Large language models are commonly trained through multi-stage post-training: first via RLHF, then fine-tuned for other downstream objectives. Yet e

Robustness Certificates for Neural Networks against Adversarial Attacks

SafetyDGX agent

arXiv:2512.20865v2 Announce Type: replace Abstract: The increasing use of machine learning in safety-critical domains amplifies the risk of adversarial threats, especially data poisoning attacks that

Rotary Masked Autoencoders are Versatile Learners

ResearchDGX agent

arXiv:2505.20535v3 Announce Type: replace Abstract: Applying Transformers to irregular time-series typically requires specializations to their baseline architecture, which can result in additional com

Rotation-Preserving Supervised Fine-Tuning

ResearchDGX agent

arXiv:2605.10973v1 Announce Type: new Abstract: Supervised fine-tuning (SFT) improves in-domain performance but can degrade out-of-domain (OOD) generalization. Prior work suggests that this degradatio

RT-Transformer: The Transformer Block as a Spherical State Estimator

ResearchDGX agent

arXiv:2605.11007v1 Announce Type: new Abstract: We show that the core components of the Transformer block -- attention, residual connections, and normalization -- arise naturally from a single geometr

Scaling Laws and Tradeoffs in Recurrent Networks of Expressive Neurons

Model ReleasesDGX agent

arXiv:2605.12049v1 Announce Type: new Abstract: Cortical neurons are complex, multi-timescale processors wired into recurrent circuits, shaped by long evolutionary pressure under stringent biological

SCOPE: Siamese Contrastive Operon Pair Embeddings for Functional Sequence Representation and Classification

Model ReleasesDGX agent

arXiv:2605.11022v1 Announce Type: cross Abstract: Identifying operons is a fundamental step in understanding prokaryotic gene regulation, as classifying genes into operons supports the reconstruction

Search Your Block Floating Point Scales!

Model ReleasesDGX agent

arXiv:2605.12464v1 Announce Type: new Abstract: Quantization has emerged as a standard technique for accelerating inference for generative models by enabling faster low-precision computations and redu

Seeing the Needle in the Haystack: Towards Weakly-Supervised Log Instance Anomaly Localization via Counterfactual Perturbation

Local AiDGX agent

arXiv:2605.10988v1 Announce Type: new Abstract: Log anomaly detection is a critical task for system operations and security assurance. However, in networked systems at scale, log data are generated at

Self-Supervised Laplace Approximation for Bayesian Uncertainty Quantification

Model ReleasesDGX agent

arXiv:2605.12208v1 Announce Type: cross Abstract: Approximate Bayesian inference typically revolves around computing the posterior parameter distribution. In practice, however, the main object of inte

← Previous
1…165166167168169…243
Next →