AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,557 results
2 Jun 2026

Efficient Hamiltonian, structure and trace distance learning of Gaussian states

TutorialsDGX agent

arXiv:2411.03163v4 Announce Type: replace-cross Abstract: In this work, we initiate the study of Hamiltonian learning for positive temperature bosonic Gaussian states, the quantum generalization of th

Efficient Synthetic Network Generation via Latent Embedding Reconstruction

ResearchDGX agent

arXiv:2606.00934v1 Announce Type: cross Abstract: Network data are ubiquitous across the social sciences, biology, and information systems. Generating realistic synthetic network data has broad applic

Embedding-Space Diffusion for Zero-Shot Environmental Sound Classification

Model ReleasesDGX agent

arXiv:2412.03771v3 Announce Type: replace-cross Abstract: Zero-shot learning enables models to generalise to unseen classes by leveraging semantic information, bridging the gap between training and te


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Enhancing LLM Metacognition via Cognitive Pairwise Training

Model ReleasesDGX agent

arXiv:2606.00869v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) has become central to LLM reasoning, but its outcome-level rewards can make models more willing to

ERICA: Quantifying Replicability of Cluster Analysis

ResearchDGX agent

arXiv:2606.00302v1 Announce Type: cross Abstract: Despite being ubiquitous in science, clustering remains a technique whose results are not quantitatively scrutinized via a framework. We present an an

Error Bounds for a Diffusion Model-Based Drift Estimator

Model ReleasesDGX agent

arXiv:2606.02115v1 Announce Type: cross Abstract: Parameter estimation in stochastic differential equations is a classical statistical problem of much importance in many scientific fields. Recent work

EST-PRM: Stress-Testing Process Reward Models Before They Become Load-Bearing

ResearchDGX agent

arXiv:2606.00437v1 Announce Type: new Abstract: Process reward models (PRMs) are widely used in language-model training with dense step-level supervision. They assume PRM scores are stable proxies for

Evaluating and Learning Robust Bandit Policies Under Uncertain Causal Mechanisms

SafetyDGX agent

arXiv:2508.02812v3 Announce Type: replace Abstract: Causal graphical models can encode large amounts structural knowledge, both from the background knowledge of domain experts and the structural knowl

Evaluating Real-World Generalizability of Algorithm Selection Models

Model ReleasesDGX agent

arXiv:2606.02016v1 Announce Type: new Abstract: Algorithm Selection (AS) aims to automatically identify the most suitable optimization algorithm for a given problem instance by leveraging measurable p

Everywhere Learning: Artificial Intelligence with Pointwise Constraints

AgentsDGX agent

arXiv:2606.01557v1 Announce Type: new Abstract: Everywhere learning is a new paradigm whereby Artificial Intelligence (AI) systems are trained to satisfy loss constraints with probability one over the

Exploiting Similarities in A/B Testing with Off-Policy Estimation

SafetyDGX agent

arXiv:2506.10677v3 Announce Type: replace-cross Abstract: We study A/B testing, the standard protocol for measuring the performance gain of a new decision system relative to a baseline. Traditional A/

Exploiting weight-space symmetries for approximating curvature

ApplicationsDGX agent

arXiv:2606.00442v1 Announce Type: new Abstract: Many machine learning techniques rely on approximating a loss function's curvature, but this is notoriously hard to do at the scale of modern deep netwo

Exposing Vulnerabilities in Explanation for Time Series Classifiers via Dual-Target Attacks

ResearchDGX agent

arXiv:2602.02763v3 Announce Type: replace Abstract: Interpretable time series deep learning systems are often assessed by checking temporal consistency on explanations, implicitly treating this as evi

Expressivity of congruence-based architectures for DNNs on positive-definite matrices

ResearchDGX agent

arXiv:2606.02490v1 Announce Type: new Abstract: This work studies neural architectures for classifying symmetric positive-definite matrices, focusing on congruence-like layers, in which the input matr

Fairness in two-player zero-sum games with bandit feedback

SafetyDGX agent

arXiv:2606.01159v1 Announce Type: new Abstract: We study two-player zero-sum games (TPZSGs) with bandit feedback under fairness constraints requiring every action to be played with probability at leas

FAiT: Frequency-Aware Inverted Transformer for Multivariate Time Series Forecasting

SafetyDGX agent

arXiv:2606.01306v1 Announce Type: new Abstract: While Transformer-based architectures have established themselves as a dominant paradigm in Multivariate Time Series Forecasting (MTSF), their core self

Fast Generalization after Interpolation via Critically Damped Momentum Optimization

Local AiDGX agent

arXiv:2606.01521v1 Announce Type: new Abstract: A central problem in machine learning is that models can achieve near-perfect training performance while generalizing substantially less well to unseen

Feature to Dynamics: Feature-space to Autoregression strategy for Zero-shot Time Series Forecasting

Model ReleasesDGX agent

arXiv:2606.01289v1 Announce Type: new Abstract: Zero-shot time series forecasting aims to predict future values for previously unseen series, requiring models to generalize temporal dynamics beyond th

FedCF: Fair Federated Conformal Prediction

SafetyDGX agent

arXiv:2509.22907v2 Announce Type: replace Abstract: Conformal Prediction (CP) is a widely used technique for quantifying uncertainty in machine learning models. In its standard form, CP offers probabi

Federated Learning via Variational Bayesian Inference: Personalization, Sparsity and Clustering

ResearchDGX agent

arXiv:2303.04345v2 Announce Type: replace Abstract: Federated learning (FL) is a promising framework that models distributed machine learning while protecting the privacy of clients. However, FL suffe

FIRM: Federated In-client Regularized Multi-objective Alignment for Large Language Models

SafetyDGX agent

arXiv:2511.16992v3 Announce Type: replace Abstract: Aligning Large Language Models (LLMs) with human values often involves balancing multiple, conflicting objectives such as helpfulness and harmlessne

Fixed-Mean Gaussian Processes for Post-hoc Bayesian Deep Learning

ResearchDGX agent

arXiv:2412.04177v2 Announce Type: replace Abstract: Recently, there has been an increasing interest in performing post-hoc uncertainty estimation about the predictions of pre-trained deep neural netwo

FLaG: Fine-Grained Latent Grouping for Hallucination Detection

ResearchDGX agent

arXiv:2606.00301v1 Announce Type: new Abstract: Hallucinations in large language models (LLMs) arise from heterogeneous failure mechanisms, making reliable detection difficult for any single global un

Flexible Online Representation Learning Based on Similarity Matching

ResearchDGX agent

arXiv:2606.01546v1 Announce Type: new Abstract: Sparse high-dimensional representations are conducive to uncovering nontrivial structures in unsupervised exploration of data. Such a representation can

Flow-Based Density Ratio Estimation for Intractable Distributions with Applications in Genomics

ResearchDGX agent

arXiv:2602.24201v2 Announce Type: replace Abstract: Estimating density ratios between pairs of intractable data distributions is a core problem in probabilistic modeling, enabling principled compariso

Flow Matching for Convective-Scale Precipitation Downscaling

Model ReleasesDGX agent

arXiv:2606.00281v1 Announce Type: cross Abstract: Generative machine learning is an increasingly important complement to dynamical downscaling for producing high-resolution precipitation projections,

Flow-Transformed Implicit Processes for Function-Space Variational Inference

ResearchDGX agent

arXiv:2606.01954v1 Announce Type: new Abstract: Implicit-process priors define distributions over functions through flexible generative mechanisms, making them attractive for Bayesian function-space m

Flowers: A Warp Drive for Neural PDE Solvers

Model ReleasesDGX agent

arXiv:2603.04430v2 Announce Type: replace Abstract: We introduce Flowers, a neural architecture for learning PDE solution operators built entirely from multihead warps. Aside from pointwise channel mi

FM-IRL: Flow-Matching for Reward Modeling and Policy Regularization in Reinforcement Learning

SafetyDGX agent

arXiv:2510.09222v3 Announce Type: replace Abstract: Flow Matching (FM) has shown remarkable ability in modeling complex distributions and achieves strong performance in offline imitation learning for

Frequentist Consistency of Prior-Data Fitted Networks for Causal Inference

SafetyDGX agent

arXiv:2603.12037v2 Announce Type: replace Abstract: Foundation models based on prior-data fitted networks (PFNs) have shown strong empirical performance in causal inference by framing the task as an i

From Moments to Models: Graphon-Mixture Learning for Mixup and Contrastive Learning

Model ReleasesDGX agent

arXiv:2510.03690v4 Announce Type: replace Abstract: Real-world graph datasets often arise from mixtures of populations, where graphs are generated by multiple distinct underlying distributions. In thi

From Performance to Viability: A Bootstrap Framework for Latent-Space Representation Learning in Adaptive Biological Systems

ResearchDGX agent

arXiv:2606.01374v1 Announce Type: new Abstract: Observable performance is commonly used to characterize biological systems. In adaptive systems, however, similar performances may arise from distinct o

From Reward-Free Representations to Preferences: Rethinking Offline Preference-Based Reinforcement Learning

ResearchDGX agent

arXiv:2606.01123v1 Announce Type: new Abstract: Preference-based reinforcement learning (PbRL) avoids explicit reward engineering by learning from pairwise human preference feedback. Existing offline

From Scaling to Structured Expressivity: Rethinking Transformers for CTR Prediction

Model ReleasesDGX agent

arXiv:2511.12081v2 Announce Type: replace-cross Abstract: Despite massive investments in scale, deep models for click-through rate (CTR) prediction often exhibit rapidly diminishing returns -- a stark

From Zero to Hero: Advancing Zero-Shot Foundation Models for Tabular Outlier Detection

ResearchDGX agent

arXiv:2602.03018v2 Announce Type: replace Abstract: Outlier detection (OD) is widely used in practice; but its effective deployment on new tasks is hindered by lack of labeled outliers, which makes al

Fundamental bounds on efficiency-confidence trade-off for transductive conformal prediction

ResearchDGX agent

arXiv:2509.04631v2 Announce Type: replace Abstract: Transductive conformal prediction addresses the simultaneous prediction for multiple data points. Given a desired confidence level, the objective is

G2LoRA: Gradient Orthogonal Low-Rank Adaptation Framework for Graph Continual Learning on Text-Attributed Graphs

Model ReleasesDGX agent

arXiv:2606.01873v1 Announce Type: new Abstract: LLM-as-Aligner has emerged as a prevalent pre-training paradigm for Text-Attributed Graphs(TAGS), aligning graph and text modalities into a shared embed

Gate the Filter, Not the Message: Node-Channel Mixtures for Pre-Propagation GNNs

TutorialsDGX agent

arXiv:2606.01660v1 Announce Type: new Abstract: Pre-propagation graph neural networks (PPGNNs) push all graph-dependent computation into a preprocessing step and train only on the resulting dense hop

Generalization of Gibbs and Langevin Monte Carlo Algorithms in the Interpolation Regime

ResearchDGX agent

arXiv:2510.06028v3 Announce Type: replace Abstract: This paper provides data-dependent bounds on the expected error of the Gibbs algorithm in the overparameterized interpolation regime, where low trai

Generalized Guarantees for Variational Inference in the Presence of Even and Elliptical Symmetry

ResearchDGX agent

arXiv:2511.01064v3 Announce Type: replace-cross Abstract: Variational inference (VI) approximates a target density p by the best match q in a family of tractable distributions. The best variational ap

Genotype-Conditioned Molecular Generation via Evidence-Grounded Multi-Objective Latent Perturbation in Diffusion Models

AgentsDGX agent

arXiv:2606.01461v1 Announce Type: new Abstract: Developing effective anticancer therapeutics remains challenging due to tumor heterogeneity and the absence of well-defined molecular targets across can

GIFT: Geometry-Induced Functional Transfer for Category-level Object Manipulation

ApplicationsDGX agent

arXiv:2503.15371v2 Announce Type: replace-cross Abstract: Robotic manipulation of unfamiliar objects in new environments is challenging due to limited generalisation capabilities. We propose a new ski

GLENS: Global Search via Learning from Solver Iterates with Diffusion Models

Model ReleasesDGX agent

arXiv:2606.00366v1 Announce Type: new Abstract: We consider the problem of generating a large collection of initial guesses for local minima of multimodal non-convex continuous optimization problems.

GLIDE: Graph-guided Leap Inference for Diffusion Estimation of Spatio-Temporal Point Processes

Local AiDGX agent

arXiv:2606.01273v1 Announce Type: new Abstract: Spatio-temporal point processes (STPPs) provide a principled framework for modeling asynchronous events in continuous time and space. Recent diffusion-b

Global Convergence of Adaptive Sensing for Principal Eigenvector Estimation

ResearchDGX agent

arXiv:2505.10882v2 Announce Type: replace Abstract: Principal component analysis classically requires full d-dimensional samples, yet in various applications hardware limits acquisition to a few scala

GNMR: Runtime Stability Control for Low-Precision Large Language Model Training

Model ReleasesDGX agent

arXiv:2606.00539v1 Announce Type: new Abstract: Training stability is a key bottleneck in low-precision language model training: efficient low-cost paths can still produce short-lived numerical risks

GNN-Enabled Robust Hybrid Beamforming with Score-Based CSI Generation and Denoising

TutorialsDGX agent

arXiv:2511.06663v2 Announce Type: replace-cross Abstract: Accurate Channel State Information (CSI) is critical for Hybrid Beamforming (HBF) tasks. However, obtaining high-resolution CSI remains challe

GPTQ-intrinsic LoRA: A Near-optimal Algorithm for Low-precision Quantization with Low-rank Adaptation

ResearchDGX agent

arXiv:2606.01412v1 Announce Type: new Abstract: Post-training quantization is widely used for compressing large neural networks, but aggressive low-bit quantization can significantly degrade model qua

Graph Edit Distance Formulation for the Vehicle Routing Problem: Theory and Analysis

Model ReleasesDGX agent

arXiv:2606.01987v1 Announce Type: cross Abstract: We show that the Vehicle Routing Problem (VRP) can be reformulated as a Graph Edit Distance (GED) maximization problem. Under a simple edge-deletion c

Graph Transfer Learning via Shared Latent Geometry: Theory and Applications

ResearchDGX agent

arXiv:2606.00716v1 Announce Type: new Abstract: Inference and control in engineered physical systems pay a heavy physics cost at deployment: state estimators, inverse-problem solvers, model-predictive

GREAT: Generalizable Backdoor Attacks in RLHF via Emotion-Aware Trigger Synthesis

Model ReleasesDGX agent

arXiv:2510.09260v2 Announce Type: replace-cross Abstract: Recent work has shown that RLHF is highly susceptible to backdoor attacks. However, existing methods often rely on rare tokens or fixed trigge

Grounded Decoding: Retrieval-Anchored Probability Fusion for Faithful RAG

ResearchDGX agent

arXiv:2606.00432v1 Announce Type: new Abstract: As retrieval-augmented generation (RAG) systems scale, it becomes increasingly challenging to ensure faithful grounding in external evidence. Large lang

HOIST: Humanoid Optimization with Imitation and Sample-efficient Tuning for Manipulating Suspended Loads

SafetyDGX agent

arXiv:2606.00252v1 Announce Type: cross Abstract: Manipulating suspended payloads with humanoid robots is challenging because the robot can only influence an underactuated, oscillatory load through wh

How Accurately Can a Gaussian Approximate Stochastic Approximation Iterates?

ResearchDGX agent

arXiv:2602.13906v2 Announce Type: replace-cross Abstract: Stochastic approximation (SA) is a method for finding the root of an operator perturbed by noise. The focus of this paper is studying the dist

How (and when) can you fit examples to logic-based hypothesis classes over infinite structures?

ResearchDGX agent

arXiv:2606.01107v1 Announce Type: cross Abstract: We study fitting problems, sometimes called ``training problems'', where we have a finite sample consisting of inputs and outputs, and we want to know

How Many Domains Suffice for Domain Generalization? A Tight Characterization via the Domain Shattering Dimension

TutorialsDGX agent

arXiv:2506.16704v3 Announce Type: replace Abstract: We study a fundamental question of domain generalization: given a family of domains (i.e., data distributions), how many randomly sampled domains do

How Much Orthogonalization Does Muon Need?

ResearchDGX agent

arXiv:2606.00371v1 Announce Type: new Abstract: Muon optimizers improve neural-network training by replacing ill-conditioned momentum updates with approximately semi-orthogonal updates. This motivates

How Neural Losses Shape VAE Latents

ResearchDGX agent

arXiv:2606.00635v1 Announce Type: new Abstract: Modern VAEs are rarely trained with the pointwise likelihood implied by the standard eta-VAE objective. In practice, pointwise reconstruction is often c

How Optimality Structures Sparse Dictionaries: A Theory for Understanding SAE Representations

Local AiDGX agent

arXiv:2606.02385v1 Announce Type: cross Abstract: Sparse Autoencoders (SAEs) have found success parsing neural representations into interpretable concepts, providing a basis for understanding and cont

HR-VILAGE-3K3M: A Human Respiratory Viral Immunization Longitudinal Gene Expression Dataset for Systems Immunity

Model ReleasesDGX agent

arXiv:2505.14725v2 Announce Type: replace-cross Abstract: Respiratory viral infections pose a global health burden, yet the cellular immune mechanisms underlying protection and pathology remain unclea

← Previous
1…105106107108109…243
Next →