AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,557 results
13 May 2026

Understanding Sample Efficiency in Predictive Coding

SafetyDGX agent

arXiv:2605.11911v1 Announce Type: new Abstract: Predictive Coding (PC) is an influential account of cortical learning. Much of recent work has focused on comparing PC to Backpropagation (BP) to find w

Uniform Scaling Limits in AdamW-Trained Transformers

ResearchDGX agent

arXiv:2605.11059v1 Announce Type: cross Abstract: We study the large-depth limit of transformers trained with AdamW, by modelling the hidden-state dynamics as an interacting particle system (IPS) coup

Unlearning with Asymmetric Sources: Improved Unlearning-Utility Trade-off with Public Data

ResearchDGX agent

arXiv:2605.11170v1 Announce Type: new Abstract: Noise-based certified machine unlearning currently faces a hard ceiling: the noise magnitude required to certify unlearning typically destroys model uti


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Variance-aware Reward Modeling with Anchor Guidance

ApplicationsDGX agent

arXiv:2605.11865v1 Announce Type: cross Abstract: Standard Bradley--Terry (BT) reward models are limited when human preferences are pluralistic. Although soft preference labels preserve disagreement i

Variational Linear Attention: Stable Associative Memory for Long-Context Transformers

ResearchDGX agent

arXiv:2605.11196v1 Announce Type: new Abstract: Linear attention reduces the quadratic cost of softmax attention to O(T), but its memory state grows as O(T) in Frobenius norm, causing progressive inte

Vertex-Softmax: Tight Transformer Verification via Exact Softmax Optimization

ResearchDGX agent

arXiv:2605.10974v1 Announce Type: new Abstract: Certified verification of transformer attention requires bounding the softmax function over interval constraints on the pre-softmax scores. Existing ver

Welfare as a Guiding Principle for Machine Learning -- From Compass, to Lens, to Roadmap

ResearchDGX agent

arXiv:2502.11981v3 Announce Type: replace Abstract: Decades of research in machine learning have given us powerful tools for making accurate predictions. But when used in social settings and on human

When and How to Canonize: A Generalization Perspective

TutorialsDGX agent

arXiv:2605.11008v1 Announce Type: new Abstract: While invariant architectures are standard for processing symmetric data, there is growing interest in achieving invariance by applying group averaging

When Does ell_2-Boosting Overfit Benignly? High-Dimensional Risk Asymptotics and the ell_1 Implicit Bias

SafetyDGX agent

arXiv:2605.06314v2 Announce Type: replace Abstract: Benign overfitting is well-characterized in ell_2 geometries, but its behavior under the ell_1 implicit bias of greedy ensembles remains challenging

When to Ask a Question: Understanding Communication Strategies in Generative AI Tools

SafetyDGX agent

arXiv:2605.11240v1 Announce Type: cross Abstract: Generative AI models differ from traditional machine learning tools in that they allow users to provide as much or as little information as they choos

Why Conclusions Diverge from the Same Observations: Formalizing World-Model Non-Identifiability via an Inference

ApplicationsDGX agent

arXiv:2605.12255v1 Announce Type: cross Abstract: When people share the same documents and observations yet reach different conclusions, the disagreement often shifts into a judgment that the other pa

xi-DPO: Direct Preference Optimization via Ratio Reward Margin

ResearchDGX agent

arXiv:2605.10981v1 Announce Type: new Abstract: Reference-free preference optimization has emerged as an efficient alternative to reinforcement learning from human feedback, with Simple Preference Opt

12 May 2026

A Call to Lagrangian Action: Learning Population Mechanics from Temporal Snapshots

ResearchDGX agent

arXiv:2605.08550v1 Announce Type: new Abstract: The population dynamics of molecules, cells, and organisms are governed by a number of unknown forces. In the last decade, population dynamics have pred

A Controlled Diagnostic Study of Hardware-Induced Distortions in Hardware-Aware Training

ResearchDGX agent

arXiv:2605.09416v1 Announce Type: new Abstract: Hardware-aware training (HAT) is widely used to improve the robustness of neural networks on non-ideal AI accelerators, such as analog in-memory computi

A Cross-Layered Multi-Drone Coordination for Medical Supply Delivery during Disaster Response Management

SafetyDGX agent

arXiv:2605.09342v1 Announce Type: cross Abstract: Autonomous drone fleets have immense potential in medical supply delivery during disaster incident response. However, coordinating multiple drones in

A Market-Rule-Informed Neural Network for Efficient Imbalance Electricity Price Forecasting

ResearchDGX agent

arXiv:2605.09061v1 Announce Type: cross Abstract: Accurate and efficient imbalance electricity price forecasting is critical for industrial energy trading systems, especially as battery assets and aut

A new initialisation to Control Gradients in Sinusoidal Neural network

Model ReleasesDGX agent

arXiv:2512.06427v2 Announce Type: replace Abstract: Proper initialisation strategy is of primary importance to mitigate gradient explosion or vanishing when training neural networks. Yet, the impact o

A PyTorch Library of Turing-Complete Neural Networks

ResearchDGX agent

arXiv:2605.08150v1 Announce Type: new Abstract: We present a PyTorch package that compiles neural networks and their weights from Turing machine descriptions, producing models that exactly simulate th

A Random-Matrix Criterion for Initializing Gated Recurrent Neural Networks

ResearchDGX agent

arXiv:2605.10650v1 Announce Type: new Abstract: Proper weight initialization prior to training has historically been one of the key factors that helped kick off the deep learning revolution. Initializ

A Simulated Federated Analysis of MS-Induced Brain Lesions

ApplicationsDGX agent

arXiv:2605.08223v1 Announce Type: new Abstract: Federated techniques such as federated learning and federated analysis have emerged as a powerful paradigm for enabling multi-center research on sensiti

A Single Deep Preference-Conditioned Policy for Learning Pareto Coverage Sets

SafetyDGX agent

arXiv:2605.08946v1 Announce Type: new Abstract: Preference-conditioned multi-objective reinforcement learning aims to learn a single policy that captures trade-offs across preferences, but under nonli

A Spectral Framework for Closed-Form Relative Density Estimation

ResearchDGX agent

arXiv:2605.10668v1 Announce Type: new Abstract: We propose a closed-form spectral framework for relative log-density estimation in linearly parameterized probabilistic models, including unnormalized a

A Stability Benchmark of Generative Regularizers for Inverse Problems

Model ReleasesDGX agent

arXiv:2605.10076v1 Announce Type: cross Abstract: Generative (diffusion) priors demonstrate remarkable performance in addressing inverse problems in imaging. Yet, for scientific and medical imaging, i

A Tale of Two Problems: Multi-Task Bilevel Learning Meets Equality Constrained Multi-Objective Optimization

ResearchDGX agent

arXiv:2605.09094v1 Announce Type: new Abstract: In recent years, bilevel optimization (BLO) has attracted significant attention for its broad applications in machine learning. However, most existing w

A Unified Lyapunov-IQC Framework for Uniform Stability of Smooth Quadratic First-Order Accelerated Optimizers

ResearchDGX agent

arXiv:2605.08488v1 Announce Type: cross Abstract: We develop a unified Lyapunov-integral quadratic constraint (IQC) framework for establishing uniform stability of first-order accelerated optimization

A Unified Representation of Neural Networks Architectures

Model ReleasesDGX agent

arXiv:2512.17593v3 Announce Type: replace Abstract: In this paper we consider the limiting case of neural networks (NNs) architectures when the number of neurons in each hidden layer and the number of

Accelerating Power Method with Fast Sketching for Stronger Low-Rank Approximation

Model ReleasesDGX agent

arXiv:2605.09755v1 Announce Type: cross Abstract: The power method is one of the most fundamental tools for extracting top principal components from data through low-rank matrix approximation. Yet, wh

Accelerating Zeroth-Order Spectral Optimization with Partial Orthogonalization from Power Iteration

Local AiDGX agent

arXiv:2605.09034v1 Announce Type: new Abstract: Zeroth-order (ZO) optimization has become increasingly popular and important in fine-tuning large language models (LLMs), especially on edge devices due

Achieving Better Local Regret Bound for Online Non-Convex Bilevel Optimization

ResearchDGX agent

arXiv:2602.06457v2 Announce Type: replace Abstract: Online bilevel optimization (OBO) has emerged as a powerful framework for many machine learning problems. Prior works have developed several algorit

Active Multiple-Prediction-Powered Inference

ApplicationsDGX agent

arXiv:2605.08429v1 Announce Type: cross Abstract: Post-deployment monitoring of healthcare AI requires statistically valid, label-efficient methods, but gold-standard labels from clinician chart revie

AdamFLIP: Adaptive Momentum Feedback Linearization Optimization for Hard Constrained PINN Training

Model ReleasesDGX agent

arXiv:2605.08408v1 Announce Type: new Abstract: Physics-informed neural networks (PINNs) provide a flexible framework for solving forward and inverse problems governed by partial differential equation

AdaPaD: Adaptive Parallel Deflation for PEFT with Self-Correcting Rank Discovery

Model ReleasesDGX agent

arXiv:2605.10741v1 Announce Type: new Abstract: Fine-tuning large language models with LoRA requires choosing a rank r before training starts. Existing approaches either extract rank-1 components sequ

Adaptive digital twins for predictive decision-making: Online Bayesian learning of transition dynamics

ApplicationsDGX agent

arXiv:2512.13919v2 Announce Type: replace Abstract: This work shows how adaptivity can enhance value realization of digital twins in civil engineering. We focus on adapting the state transition models

Adaptive Memory Momentum via a Model-Based Framework for Deep Learning Optimization

ResearchDGX agent

arXiv:2510.04988v3 Announce Type: replace Abstract: The vast majority of modern deep learning models are trained with momentum-based first-order optimizers. The momentum term governs the optimizer's m

Adaptive Multi-view Graph Contrastive Learning via Fractional-order Neural Diffusion Networks

Model ReleasesDGX agent

arXiv:2511.06216v4 Announce Type: replace Abstract: Graph contrastive learning (GCL) learns node and graph representations by contrasting multiple views of the same graph. Existing methods typically r

Additive Atomic Forests for Symbolic Function and Antiderivative Discovery

ResearchDGX agent

arXiv:2605.08130v1 Announce Type: new Abstract: We present a framework for the simultaneous symbolic recovery of a function and its antiderivative from data. The framework rests on three ideas. First,

Adversary-Robust Learning from Fully Asynchronous Directional Derivative Estimates

Model ReleasesDGX agent

arXiv:2605.09337v1 Announce Type: new Abstract: We propose FAR-SIGN (Fully Asynchronous Robust optimization via SIGNed directional projections) for adversary-resilient learning in parameter-server--wo

Affine Tracing: A New Paradigm for Probabilistic Linear Solvers

ResearchDGX agent

arXiv:2605.10566v1 Announce Type: cross Abstract: Probabilistic linear solvers (PLSs) return probability distributions that quantify uncertainty due to limited computation in the solution of linear sy

AgentSlimming: Towards Efficient and Cost-Aware Multi-Agent Systems

AgentsDGX agent

arXiv:2605.08813v1 Announce Type: new Abstract: Large Language Model-based Multi-Agent Systems (MAS) have demonstrated remarkable capabilities in complex tasks. However, manually designing optimal com

Aligning Validation with Deployment: Target-Weighted Cross-Validation for Spatial Prediction

SafetyDGX agent

arXiv:2603.29981v2 Announce Type: replace Abstract: Reliable estimation of predictive performance is essential for spatial environmental modeling, where machine-learning models are used to generate ma

Alignment-Sensitive Minimax Rates for Spectral Algorithms with Learned Kernels

SafetyDGX agent

arXiv:2509.20294v4 Announce Type: replace Abstract: We study spectral algorithms in the setting where kernels are learned from data. We introduce the effective span dimension (ESD), an alignment-sensi

(alpha,eta)-Stability for Boosting Vector-Valued Prediction

ResearchDGX agent

arXiv:2602.18866v2 Announce Type: replace Abstract: Despite the widespread use of boosting in structured prediction, a general theoretical understanding of aggregation beyond scalar prediction remains

AlphaExploitem: Going Beyond the Nash Equilibrium in Poker by Learning to Exploit Suboptimal Play

Model ReleasesDGX agent

arXiv:2605.09150v1 Announce Type: new Abstract: Poker is an imperfect information game that has served as a long-standing benchmark for decision-making under uncertainty. To maximize utility beyond th

Amortizing Causal Sensitivity Analysis via Prior Data-Fitted Networks

ResearchDGX agent

arXiv:2605.10590v1 Announce Type: cross Abstract: Causal sensitivity analysis aims to provide bounds for causal effect estimates in the presence of unobserved confounding. However, existing methods fo

An Empirical Analysis of Calibration and Selective Prediction in Multimodal Clinical Condition Classification

SafetyDGX agent

arXiv:2603.02719v2 Announce Type: replace Abstract: As artificial intelligence systems move toward clinical deployment, ensuring reliable prediction behavior is fundamental for safety-critical decisio

An Integrative Genome-Scale Metabolic Modeling and Machine Learning Framework for Predicting and Optimizing Single-Cell Protein Production in Saccharomyces cerevisiae

ApplicationsDGX agent

arXiv:2603.25561v2 Announce Type: replace Abstract: Saccharomyces cerevisiae is increasingly recognised as a key source for single-cell protein (SCP) production, a rising solution to global protein-su

Anchor-guided Hypergraph Condensation with Dual-level Discrimination

ResearchDGX agent

arXiv:2605.10001v1 Announce Type: new Abstract: The increasing prevalence of large-scale hypergraphs poses significant computational challenges for hypergraph neural network (HNN) training. To address

APEX: Audio Prototype EXplanations for Classification Tasks

ResearchDGX agent

arXiv:2605.10153v1 Announce Type: cross Abstract: Explainable AI (XAI) has achieved remarkable success in image classification, yet the audio domain lacks equally mature solutions. Current methods app

Applying Graph Analysis for Unsupervised Fast Malware Fingerprinting

ResearchDGX agent

arXiv:2510.12811v2 Announce Type: replace-cross Abstract: Malware proliferation is increasing at a tremendous rate, with hundreds of thousands of new samples identified daily. Manual investigation of

Assessing the robustness of heterogeneous treatment effects in survival analysis under informative censoring

SafetyDGX agent

arXiv:2510.13397v3 Announce Type: replace Abstract: Dropout is common in clinical studies, with up to half of patients leaving early due to side effects or other reasons. When dropout is informative (

Auction-Based Online Policy Adaptation for Evolving Objectives

SafetyDGX agent

arXiv:2604.02151v2 Announce Type: replace Abstract: We consider multi-objective reinforcement learning problems where objectives come from an identical family -- such as the class of reachability obje

AxiomOcean: Forecasting the Three-Dimensional Structure of the Upper Ocean

ResearchDGX agent

arXiv:2605.10455v1 Announce Type: new Abstract: Short-term ocean forecast skill depends strongly on the three-dimensional ocean structure of the upper ocean, which governs stratification, subsurface h

Balancing Efficiency and Fairness in Traffic Light Control through Deep Reinforcement Learning

SafetyDGX agent

arXiv:2605.10170v1 Announce Type: new Abstract: Urban traffic congestion presents a significant challenge for modern cities, which impacts mobility and sustainability. Traditional traffic light contro

Bayesian Optimization with Structured Measurements: A Vector-Valued RKHS Framework

ResearchDGX agent

arXiv:2605.09775v1 Announce Type: new Abstract: Bayesian optimization (BO) is an efficient framework for optimizing expensive black-box functions. However, it is typically formulated as learning an en

Bayesian Reasoning for Physics Informed Neural Networks

ResearchDGX agent

arXiv:2308.13222v3 Announce Type: replace-cross Abstract: We introduce an evidence-driven Bayesian formulation of physics-informed neural networks that enables automatic optimization of loss weights b

BCJR-QAT: A Differentiable Relaxation of Trellis-Coded Weight Quantization

Model ReleasesDGX agent

arXiv:2605.10655v1 Announce Type: new Abstract: Trellis-coded quantization sets the current 2-bit post-training frontier for LLMs (QTIP), but pushing below the PTQ ceiling requires quantization-aware

Benchmarking Sensor-Fault Robustness in Forecasting

Model ReleasesDGX agent

arXiv:2605.10822v1 Announce Type: new Abstract: Cyber-physical system (CPS) forecasting models depend on sensor streams with noisy, biased, missing, or temporally misaligned readings, yet standard for

Benchmarking Transformer and xLSTM for Time-Series Forecasting of Heat Consumption

Model ReleasesDGX agent

arXiv:2605.09722v1 Announce Type: new Abstract: Obtaining an accurate short-term forecasting for heat demand is an essential part of operating district heating networks cost-efficient and reliable. He

Beyond Accuracy: Evaluating Posterior Fidelity of Diffusion Inverse Solvers

ApplicationsDGX agent

arXiv:2602.04189v2 Announce Type: replace Abstract: Uncertainty evaluation is critical in scientific and engineering inverse problems. However, existing benchmarks on Diffusion Inverse Solvers (DIS) p

Beyond Hard Writes and Rigid Preservation: Soft Recursive Least-Squares for Lifelong LLM Editing

ResearchDGX agent

arXiv:2601.15686v2 Announce Type: replace Abstract: Model editing updates a pre-trained LLM with new facts or rules without retraining while preserving unrelated behavior. In real deployment, edits ar

← Previous
1…167168169170171…243
Next →