AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent
84,433Total entries
1Added by human
84,432Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,557 results
12 May 2026

Modeling Atomic Conformational Ensembles of Proteins via Test-Time Supervision of Boltz-2 on Cryo-EM Density Maps

ResearchDGX agent

arXiv:2605.09832v1 Announce Type: new Abstract: Knowledge of a protein's atomic conformational ensemble is critical to determining its function, yet state-of-the-art ensemble prediction models are lim

MoMo: Conditioned Contrastive Representation Learning for Preference-Modulated Planning

SafetyDGX agent

arXiv:2605.08512v1 Announce Type: new Abstract: Temporally contrastive representation learning induces a latent structure capable of reducing long-horizon planning to inference in a low-dimensional li

Multi-scale Predictive Representations for Goal-conditioned Reinforcement Learning

SafetyDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2605.09364v1 Announce Type: new Abstract: This paper investigates robust representation learning in offline goal-conditioned reinforcement learning (GCRL). Particularly in sparse reward scenario

Multifidelity Gaussian process regression for solving nonlinear partial differential equations

ResearchDGX agent

arXiv:2605.10383v1 Announce Type: cross Abstract: Solving nonlinear partial differential equations (PDEs) using kernel methods offers a compelling alternative to traditional numerical solvers. However

Muon Does Not Converge on Convex Lipschitz Functions

ResearchDGX agent

arXiv:2605.08980v1 Announce Type: new Abstract: Muon and its variants have shown strong empirical performance in a variety of deep learning tasks. Existing convergence analyses of Muon rely on smoothn

Muon-OGD: Muon-based Spectral Orthogonal Gradient Projection for LLM Continual Learning

Model ReleasesDGX agent

arXiv:2605.08949v1 Announce Type: new Abstract: A central challenge in continual learning for large language models (LLMs) is catastrophic forgetting, where adapting to new tasks can substantially deg

MuonEq: Balancing Before Orthogonalization with Lightweight Equilibration

ResearchDGX agent

arXiv:2603.28254v2 Announce Type: replace Abstract: Orthogonalized-update optimizers such as Muon improve training of matrix-valued parameters, but existing extensions typically either rescale updates

Muown: Row-Norm Control for Muon Optimization

ResearchDGX agent

arXiv:2605.10797v1 Announce Type: new Abstract: Muon has emerged as a strong competitor to AdamW for language model pre-training, yet its behavior at scale is sensitive to weight decay. Recent work ha

Mutual Information Optimal Density Control of Linear Systems and Generalized Schrodinger Bridges with Reference Refinement

SafetyDGX agent

arXiv:2605.09349v1 Announce Type: cross Abstract: We consider a mutual information (MI) regularized version of optimal density control of a discrete-time linear system. MI optimal control has been pro

Natural Policy Gradient as Doubly Smoothed Policy Iteration: A Bellman-Operator Framework

SafetyDGX agent

arXiv:2605.10671v1 Announce Type: new Abstract: In this work, we show that natural policy gradient, a core algorithm in reinforcement learning, admits an exact formulation as a smoothed and averaged f

Near-Optimal Last-Iterate Convergence for Zero-Sum Games with Bandit Feedback and Opponent Actions

ResearchDGX agent

arXiv:2605.09363v1 Announce Type: new Abstract: Last-iterate convergence of learning dynamics in games has attracted significant recent attention. In two-player zero-sum games with bandit feedback, wh

Nearly-Optimal Algorithm for Adversarial Kernelized Bandits

ResearchDGX agent

arXiv:2605.10299v1 Announce Type: new Abstract: This paper studies kernelized bandits (also known as Gaussian process bandits) in an adversarial environment, where the reward functions in a known repr

Nested Slice Sampling: Vectorized Nested Sampling for GPU-Accelerated Inference

Model ReleasesDGX agent

arXiv:2601.23252v2 Announce Type: replace-cross Abstract: Model comparison and calibrated uncertainty quantification often require integrating over parameters, but scalable inference can be challengin

Neural Co-state Policies: Structuring Hidden States in Recurrent Reinforcement Learning

SafetyDGX agent

arXiv:2605.05373v2 Announce Type: replace Abstract: A key capability of intelligent agents is operating under partial observability: reasoning and acting effectively despite missing or incomplete stat

Neural Posterior Estimation of Terrain Parameters from Radar Sounder Data

Model ReleasesDGX agent

arXiv:2605.08179v1 Announce Type: cross Abstract: Radar sounders are electromagnetic instruments that can probe deep into the subsurface of Earth and other planetary bodies by processing the echo of t

Neural Variance-aware Dueling Bandits with Deep Representation and Shallow Exploration

ApplicationsDGX agent

arXiv:2506.01250v2 Announce Type: replace Abstract: In this paper, we address the contextual dueling bandit problem by proposing variance-aware algorithms that leverage neural networks to approximate

Neural Weight Norm = Kolmogorov Complexity

Model ReleasesDGX agent

arXiv:2605.10878v1 Announce Type: new Abstract: Why does weight decay work? We prove that, in any fixed-precision regime, the smallest weight norm of a looped neural network outputting a binary string

NeuralBench: A Unifying Framework to Benchmark NeuroAI Models

Model ReleasesDGX agent

arXiv:2605.08495v1 Announce Type: new Abstract: Deep learning and large public datasets have recently catalyzed the proliferation of AI models for processing brain recordings. However, systematically

Neurally-plausible radial basis kernels using distributed Fourier embeddings

ResearchDGX agent

arXiv:2605.08458v1 Announce Type: new Abstract: Coherent, continuous spatial representations are critical for synthesizing physical and perceptual phenomena into a single representational space. Radia

Neuroprobe: Evaluating Intracranial Brain Responses to Naturalistic Stimuli

ResearchDGX agent

arXiv:2509.21671v2 Announce Type: replace Abstract: High-resolution neural datasets enable foundation models for the next generation of brain-computer interfaces and neurological treatments. The commu

Non-intrusive Body Composition Assessment from Full-body mmWave Scans

ResearchDGX agent

arXiv:2605.08306v1 Announce Type: cross Abstract: Body composition assessment (BCA) provides detailed information about the distribution of different tissue types in the body, enabling more precise ch

Non-Parametric Rehearsal Learning via Conditional Mean Embeddings

ResearchDGX agent

arXiv:2605.08999v1 Announce Type: new Abstract: In machine learning, a critical class of decision-related problems concerns preventing predicted undesirable outcomes, referred to as the extit{avoiding

Nonlinear GENERIC Informed Neural Networks (N-GINNs): learning GENERIC dynamics with non-quadratic dissipation potentials

ResearchDGX agent

arXiv:2605.09058v1 Announce Type: cross Abstract: We introduce Nonlinear GENERIC Informed Neural Networks (N-GINNs), a deep learning framework for discovering evolution equations of systems governed b

NoRIN: Backbone-Adaptive Reversible Normalization for Time-Series Forecasting

ApplicationsDGX agent

arXiv:2605.10823v1 Announce Type: new Abstract: Reversible instance normalization (RevIN) and its successors (Dish-TS, SAN, FAN) have become the de facto plug-in for time-series forecasting, yet the m

Objective-Specific Privileged Bases via Full-Prefix Matryoshka Learning

ResearchDGX agent

arXiv:2605.09160v1 Announce Type: new Abstract: Learned representations are often invariant to rotational transformations, leaving individual dimensions non-identifiable and interchangeable. We study

Omni-DeepSearch: A Benchmark for Audio-Driven Omni-Modal Deep Search

Model ReleasesDGX agent

arXiv:2605.08762v1 Announce Type: cross Abstract: Current omni-modal benchmarks mainly evaluate models under settings where multiple modalities are provided simultaneously, while the ability to start

On Characterizing Learnability for Adversarial Noisy Bandits

ResearchDGX agent

arXiv:2605.09200v1 Announce Type: new Abstract: We study adversarial noisy bandits given a known function class F. In each round, the adversary selects a function f in F, the learner chooses an arm, a

On Improving Graph Neural Networks for QSAR by Pre-training on Extended-Connectivity Fingerprints

ResearchDGX agent

arXiv:2605.10722v1 Announce Type: new Abstract: Molecular Graph Neural Networks (GNNs) are increasingly common in drug discovery, particularly for Quantitative Structure-Activity Relationship (QSAR) s

On Observation Time for Recovering Latent Hawkes Networks

ApplicationsDGX agent

arXiv:2605.08400v1 Announce Type: cross Abstract: Dynamics of interacting systems in engineering, society, and nature often evolve over latent networks that govern which entities can interact. We stud

On periodic distributed representations using Fourier embeddings

ResearchDGX agent

arXiv:2605.10818v1 Announce Type: new Abstract: Periodic signals are critical for representing physical and perceptual phenomena. Scalar, real angular measures, e.g., radians and degrees, result in di

On the Convergence of Muon and Beyond

ResearchDGX agent

arXiv:2509.15816v5 Announce Type: replace Abstract: The Muon optimizer has demonstrated remarkable empirical success in handling matrix-structured parameters for training neural networks. However, a s

On the Convergence Rate of LoRA Gradient Descent

ResearchDGX agent

arXiv:2512.18248v3 Announce Type: replace Abstract: The low-rank adaptation (LoRA) algorithm for fine-tuning large models has grown popular in recent years due to its remarkable performance and low co

On the global convergence of gradient descent for wide shallow models with bounded nonlinearities

ResearchDGX agent

arXiv:2605.10775v1 Announce Type: cross Abstract: A surprising phenomenon in the training of neural networks is the ability of gradient descent to find global minimizers of the training loss despite i

On Uniform Error Bounds for Kernel Regression under Non-Gaussian Noise

SafetyDGX agent

arXiv:2605.09757v1 Announce Type: new Abstract: Providing non-conservative uncertainty quantification for function estimates derived from noisy observations remains a fundamental challenge in statisti

Online Set Learning from Precision and Recall Feedback

ResearchDGX agent

arXiv:2605.09565v1 Announce Type: new Abstract: We consider the problem of learning an unknown subset N_ext{target} of a domain in an online setting. In each round t, the learner predicts a set of ite

Online Sharp-Calibrated Bayesian Optimization

ApplicationsDGX agent

arXiv:2605.10572v1 Announce Type: new Abstract: Bayesian optimization (BO) is a widely used framework for optimizing expensive black-box functions, commonly based on Gaussian process (GP) surrogate mo

Optimal and Scalable MAPF via Multi-Marginal Optimal Transport and Schrodinger Bridges

AgentsDGX agent

arXiv:2605.10917v1 Announce Type: new Abstract: We consider anonymous multi-agent path finding (MAPF) where a set of robots is tasked to travel to a set of targets on a finite, connected graph. We sho

Optimal Attention Temperature Improves the Robustness of In-Context Learning under Distribution Shift in High Dimensions

ResearchDGX agent

arXiv:2511.01292v2 Announce Type: replace-cross Abstract: Pretrained Transformers can perform in-context learning (ICL) from a few demonstrations, but this ability can fail sharply when the test distr

Optimal Regret for Single Index Bandits

ResearchDGX agent

arXiv:2605.09454v1 Announce Type: cross Abstract: We study the extit{single-index bandit} problem, where rewards depend on an unknown one-dimensional projection of high-dimensional contexts through an

Optimality of Sub-network Laplace Approximations: New Results and Methods

Model ReleasesDGX agent

arXiv:2605.09075v1 Announce Type: cross Abstract: Although the Laplace approximation offers a simple route to uncertainty quantification in deep neural networks, its reliance on inverting large Hessia

Optimised Support Vector Regression for California Housing Price Prediction: The Critical Role of Feature Engineering and Hyperparameter Tuning

Model ReleasesDGX agent

arXiv:2605.08660v1 Announce Type: new Abstract: In the recent literature, Support Vector Regression (SVR) has been cited as one of the weakest performers on the California Housing benchmark dataset, w

Optimizing Server Placement for Vertical Federated Learning in Dynamic Edge/Fog Networks

ResearchDGX agent

arXiv:2605.09813v1 Announce Type: cross Abstract: We investigate the control and optimization of vertical federated learning (VFL), a class of distributed machine learning (ML) methods in which edge/f

OTora: A Unified Red Teaming Framework for Reasoning-Level Denial-of-Service in LLM Agents

Model ReleasesDGX agent

arXiv:2605.08876v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly deployed as autonomous agents that execute tool-augmented, multi-step tasks, where latency is a critical f

OUIDecay: Adaptive Layer-wise Weight Decay for CNNs Using Online Activation Patterns

ResearchDGX agent

arXiv:2605.10161v1 Announce Type: new Abstract: Weight decay remains one of the most widely used regularization mechanisms for training convolutional neural networks, yet it is still commonly applied

Outlier detection for patient monitoring and alerting

ResearchDGX agent

arXiv:2605.08955v1 Announce Type: new Abstract: We develop and evaluate a data-driven approach for detecting unusual (anomalous) patient-management decisions using past patient cases stored in electro

Outlier-robust Diffusion Posterior Sampling for Bayesian Inverse Problems

TutorialsDGX agent

arXiv:2602.02045v2 Announce Type: replace Abstract: Diffusion models have emerged as powerful learned priors for Bayesian inverse problems (BIPs). Diffusion-based solvers rely on a presumed likelihood

PA-RNet: Perturbation-Aware Residual Network for Robust Multimodal Time Series Forecasting

ApplicationsDGX agent

arXiv:2508.04750v2 Announce Type: replace Abstract: In real-world applications, multimodal time-series forecasting faces a key challenge: textual information is often useful but unreliable. Auxiliary

PACT: Peak-Aware Cross-Attention Graph Transformers for Efficient Storm-Surge Emulation

ResearchDGX agent

arXiv:2605.09036v1 Announce Type: new Abstract: Accurate and efficient storm-surge emulation is essential for coastal hazard assessment, yet high-fidelity hydrodynamic models remain too expensive for

Parameterized Complexity of Stationarity Testing for Piecewise-Affine Functions and Shallow CNN Losses

Model ReleasesDGX agent

arXiv:2605.10219v1 Announce Type: cross Abstract: We study the parameterized complexity of testing approximate first-order stationarity at a prescribed point for continuous piecewise-affine (PA) funct

Path-Based Gradient Boosting for Graph-Level Prediction

Model ReleasesDGX agent

arXiv:2605.08102v1 Announce Type: new Abstract: We propose PathBoost, a gradient tree boosting method for graph-level classification and regression that learns discriminative path-based features direc

Path-Dependent Denoising: A Non-Conservative Field Perspective on Order Collapse in Diffusion Language Models

Local AiDGX agent

arXiv:2605.09303v1 Announce Type: new Abstract: Diffusion language models (DLMs) offer a structural alternative to autoregressive generation: denoising can update tokens in arbitrary orders or in para

PC3D: Zero-Shot Cooperation Across Variable Rosters via Personalized Context Distillation

AgentsDGX agent

arXiv:2605.10377v1 Announce Type: new Abstract: Cooperative multi-agent reinforcement learning often assumes a fixed execution team, yet many decentralized systems must operate with varying numbers of

Per-Loss Adapters for Gradient Conflict in Physics-Informed Neural Networks

Model ReleasesDGX agent

arXiv:2605.10136v1 Announce Type: new Abstract: Physics-informed neural networks (PINNs) train a single neural approximation by minimizing multiple physics- and data-derived losses, but the gradients

Performance and Energy Trade-Off Analysis of Hierarchical Federated Learning for Plant Disease Classification

ResearchDGX agent

arXiv:2605.08121v1 Announce Type: cross Abstract: Early detection of plant diseases is critical for improving crop productivity, while it also facilitates the foundations of precision agriculture. Rec

PFN-TS: Thompson Sampling for Contextual Bandits via Prior-Data Fitted Networks

SafetyDGX agent

arXiv:2605.10137v1 Announce Type: cross Abstract: Thompson sampling is a widely used strategy for contextual bandits: at each round, it samples a reward function from a Bayesian posterior and acts gre

Phases of Muon: When Muon Eclipses SignSGD

ResearchDGX agent

arXiv:2605.09552v1 Announce Type: cross Abstract: Recently, Muon and related spectral optimizers have demonstrated strong empirical performance as scalable stochastic methods, often outperforming Adam

PHIDA: Persistence-Guided Node-to-Cluster Mapping for Online Clustering

Model ReleasesDGX agent

arXiv:2605.08673v1 Announce Type: new Abstract: Online clustering methods that adaptively create and update nodes as data arrive often make node learning explicit, whereas the mapping from the learned

PhysEDA: Physics-Aware Learning Framework for Efficient EDA With Manhattan Distance Decay

SafetyDGX agent

arXiv:2605.10547v1 Announce Type: new Abstract: Electronic design automation (EDA) addresses placement, routing, timing analysis, and power-integrity verification for integrated circuits. Learning met

Physics-Informed Neural PDE Solvers via Spatio-Temporal MeanFlow

ResearchDGX agent

arXiv:2605.08915v1 Announce Type: new Abstract: Deep learning paradigms, such as PINNs and neural operators, have significantly advanced the solving of PDEs. However, they often struggle to capture th

Physics-Modeled Neural Networks

ResearchDGX agent

arXiv:2605.08176v1 Announce Type: new Abstract: We introduce Dynamical Physics-Modeled Neural Networks (DynPMNNs), a continuous-time deep learning architecture in which each hidden layer is defined as

← Previous
1…172173174175176…243
Next →