AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,329 results
4 Aug 2026

Retrieval-Augmented Interpretable Learning: Towards Task-Specific Zero-Shot Models in Healthcare

ApplicationsDGX agent

arXiv:2607.17508v2 Announce Type: replace Abstract: We introduce Retrieval-Augmented Interpretable Learning (RAIL), a probabilistic meta-learning framework for zero-shot generation of task-specific in

Reusing Rollouts under Policy Lag: Prefix-Normalized Policy Optimization for LLM Reinforcement Learning

Model ReleasesDGX agent

arXiv:2608.01418v1 Announce Type: cross Abstract: Autoregressive rollout generation is a major computational cost in reinforcement learning for large language models. Reusing each rollout batch for ad

RHEA: Reliability-Harmonized Reconstruction and Assignment for Robust Multimodal-Attributed Graph Clustering

Research

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2608.00621v1 Announce Type: new Abstract: Multimodal-attributed graphs (MAGs), whose nodes carry heterogeneous attributes such as text and images over a relational structure, have become a funda

Riemannian Attention Mechanisms for Transformers: A Theoretical Framework and Architecture Design

Model ReleasesDGX agent

arXiv:2608.01283v1 Announce Type: new Abstract: All Transformer-based large language models compute attention via the Euclidean inner product, an architectural choice that Dong et al. (2021) proved ca

Robust Bayesian Optimization via Tempered Posteriors

ResearchDGX agent

arXiv:2601.07094v2 Announce Type: replace-cross Abstract: Bayesian optimization (BO) iteratively fits a Gaussian process (GP) surrogate to accumulated evaluations and selects new queries via an acquis

Round-Trip Consistency: Bidirectional Diffusion Models Can Predict Their Own Rollout Errors

Model ReleasesDGX agent

arXiv:2608.00675v1 Announce Type: cross Abstract: Autoregressive models accumulate error over long rollouts, yet at deployment there is no ground truth to measure it against. We train a single conditi

SAFE-Merge: Data-Free Continual Model Merging with General Knowledge Preservation

Model ReleasesDGX agent

arXiv:2608.01184v1 Announce Type: new Abstract: Data-free continual model merging must incorporate a stream of specialized models while retaining both pretrained general knowledge and previously acqui

Scikit-fingerprints: Python library for scikit-learn compatible molecular fingerprints and chemoinformatics

TutorialsDGX agent

arXiv:2608.02027v1 Announce Type: new Abstract: We present scikit-fingerprints, a comprehensive, fully scikit-learn compatible library for molecular machine learning in Python, based on RDKit. Molecul

SCOPE: Entanglement Frontier Escape for Source-Free Class Unlearning

Model ReleasesDGX agent

arXiv:2608.02058v1 Announce Type: new Abstract: Source-free class unlearning erases whole classes using only the forget data, judged at the representation level, where features can leak a class the he

Scoring Rules! Statistical and Strategic Alignment for Text Evaluation Metrics

SafetyDGX agent

arXiv:2608.01423v1 Announce Type: cross Abstract: Reference-based text evaluation metrics, which are widely used to assess natural language generation systems, score a candidate response by comparing

Searching for Quantum Effects in the Brain: A Bell-Type Test for Nonclassical Latent Representations in Autoencoders

ResearchDGX agent

arXiv:2601.10588v2 Announce Type: replace-cross Abstract: Whether neural information processing is entirely classical or involves quantum-mechanical elements remains an open question. Here we propose

Secrets Everywhere: Auditing Memorization in Mobility Prediction Models

ResearchDGX agent

arXiv:2608.02052v1 Announce Type: new Abstract: Human mobility prediction models, which forecast the next location in a user's trajectory, are increasingly deployed in urban analytics, navigation, and

Self-Certification of Representation Adequacy: Sequential Certification at Minimum Task Loss

SafetyDGX agent

arXiv:2608.02267v1 Announce Type: cross Abstract: Agents that act on a compressed representation of their history face a structural risk: if the representation aliases histories with different optimal

Self-composing neural operators for high-frequency and multiscale PDE surrogates

Model ReleasesDGX agent

arXiv:2508.20650v2 Announce Type: replace Abstract: Addressing the computational challenges of high-frequency and multiscale partial differential equations (PDEs), this work introduces a self-composin

Self-Supervised Representations for Binary Program Clustering: From Empirical Study to Retrieval-Augmented Learning

ResearchDGX agent

arXiv:2608.02348v1 Announce Type: cross Abstract: Malware clustering is a critical task in cybersecurity that helps discover threats and analyze evolving malware families. While self-supervised learni

Sharp Root Anti-Concentration via Projective Incidence and Ordered Root Laws

Local AiDGX agent

arXiv:2608.01670v1 Announce Type: new Abstract: This paper answers the one-dimensional local root anti-concentration questions posed by Balcan, Pegden, and Sharma in the context of online optimization

Sheaf-theoretic Signal Processing on Graphs: Spectral Theory, Filtering, and Sampling

ResearchDGX agent

arXiv:2608.01318v1 Announce Type: cross Abstract: Modern sensing, communication, and learning systems generate heterogeneous network signals, with local data differing in dimension, modality, and geom

SIGMA: Semantic Identifier Grouping for Molecular Autoregression

SafetyDGX agent

arXiv:2603.25062v2 Announce Type: replace Abstract: Autoregressive molecular models assign probability to molecular serializations even though chemical identity is invariant to serialization. Equivale

Similarity-Aware Machine Unlearning

Model ReleasesDGX agent

arXiv:2608.00246v1 Announce Type: new Abstract: Machine unlearning removes the influence of user-specified training examples from a trained model, avoiding the need to retrain it from scratch. Localiz

Simplified Quadratic Gradient: A Unified Framework Bridging Gradient Descent and Newton-Type Methods by Synthesizing Hessians and Gradients

ResearchDGX agent

arXiv:2209.03282v5 Announce Type: replace-cross Abstract: Accelerating the convergence of second-order optimization, particularly Newton-type methods, remains a pivotal challenge in algorithmic resear

Simulation-Based Plate-Reverb Parameter Estimation from a Single Impulse Response

Model ReleasesDGX agent

arXiv:2608.00656v1 Announce Type: cross Abstract: We present a simulation-trained, non-iterative estimator for Task A of the 1st DAFx Parameter Estimation Challenge. Each unnormalized plate-reverb imp

SingLEM: Single-Channel Large EEG Model

Local AiDGX agent

arXiv:2509.17920v2 Announce Type: replace Abstract: Current deep learning models for electroencephalography (EEG) are often task-specific and depend on large labeled datasets, limiting their adaptabil

Smooth Reparameterizations of Functions on Simplicial Product Spaces: Applications to Probabilistic Tensor Decomposition and Functional Data Registration

ResearchDGX agent

arXiv:2608.02576v1 Announce Type: new Abstract: We consider optimization problems defined on product spaces of simplices. Examples of this class of problems include learning low-rank discrete multivar

SoniSpeech: A Large-Scale Open-Vocabulary Tri-Modal Dataset for Wearable Silent Speech Interfaces

Model ReleasesDGX agent

arXiv:2608.00803v1 Announce Type: cross Abstract: Wearable silent speech interfaces (SSIs) are limited to small, closed vocabularies. Approaches achieving larger vocabularies require obtrusive hardwar

Sparse Covariance Neural Networks

ResearchDGX agent

arXiv:2410.01669v3 Announce Type: replace Abstract: Covariance Neural Networks (VNNs) perform graph convolutions on the covariance matrix of input data to leverage correlation information as pairwise

SparseKAN: Compressing Kolmogorov--Arnold Networks Across Basis Functions, Neurons, and Bits

HardwareDGX agent

arXiv:2608.00859v1 Announce Type: new Abstract: Kolmogorov--Arnold Networks (KANs) replace scalar edge weights with learnable univariate functions parameterized by multiple basis coefficients. This in

Spatiotemporal Proximal Causal Inference under Hidden Confounding and Interference

Local AiDGX agent

arXiv:2608.01352v1 Announce Type: new Abstract: Estimating causal effects from real-world spatiotemporal data is challenging due to hidden confounders and interference. Standard causal identification

Stabilized Best-of-K Training for Neural Combinatorial Optimization

ResearchDGX agent

arXiv:2608.00296v1 Announce Type: new Abstract: Leader Reward modifies POMO training to emphasize the best trajectory produced by repeated inference. We test a narrow extension: replace its binary lea

Staged Multi-Agent Training (SMAT) for Hip Exoskeletons: Metabolic and Biomechanical Validation of a Simulation-Trained Co-Adaptive Controller

Model ReleasesDGX agent

arXiv:2608.00715v1 Announce Type: cross Abstract: Learning-based controllers can deliver exoskeleton assistance after training entirely in physics-based simulation, yet few controllers that address hu

Start Classifying: Categorical Critics for LLM Reinforcement Learning

SafetyDGX agent

arXiv:2608.02181v1 Announce Type: new Abstract: Proximal Policy Optimization (PPO) for large language models typically trains its critic by mean-squared-error (MSE) regression on scalar value targets.

Statistical comparisons of time-series feature sets on classification tasks

ResearchDGX agent

arXiv:2608.01586v1 Announce Type: cross Abstract: In recent years, numerous open-source software libraries have been developed for computing sets of features from univariate time series. The type and

Statistical Mechanics of Learning on Product Wasserstein Manifolds

ResearchDGX agent

arXiv:2608.01434v1 Announce Type: new Abstract: Normally the statistical mechanics of learning treats constraints on weight distributions as restrictions that shrink the space of possible solutions. T

Stop When Memory Suffices: Evidence-Conditioned Progressive Execution for LLM Agents

AgentsDGX agent

arXiv:2608.01285v1 Announce Type: new Abstract: The continued development of LLMs toward persistent and adaptive intelligence increasingly requires long-term memory mechanisms that preserve and reuse

Structured Memory for Edge Language Models: Persistent Context and Corpus Retrieval via O(1) SSM State Injection

Model ReleasesDGX agent

arXiv:2608.02560v1 Announce Type: new Abstract: Retrieval-augmented generation (RAG) imposes a prefill cost proportional to retrieved context length, and -- with Transformer backbones -- a KV-cache th

Subtype Robustness Is Not Just Accuracy: Calibration Under Unseen Subtype Shift

ResearchDGX agent

arXiv:2608.00928v1 Announce Type: new Abstract: Subtype robustness asks whether a model keeps the correct coarse prediction when test examples come from fine-grained subtypes absent from training but

Surrogate Modeling for the Design of Optimal Lattice Structures using Tensor Completion

ResearchDGX agent

arXiv:2510.07474v2 Announce Type: replace Abstract: When designing new materials, it is often necessary to design a material with specific desired properties. Unfortunately, as new design variables ar

T-TAMER: Provably Taming Trade-offs in ML Serving

ResearchDGX agent

arXiv:2509.22992v2 Announce Type: replace Abstract: As machine learning models continue to grow in size and complexity, efficient serving faces increasingly broad trade-offs spanning accuracy, latency

TabDPT-Turbo: Efficient In-Context Learning for Tabular Prediction

Model ReleasesDGX agent

arXiv:2608.01400v1 Announce Type: new Abstract: Tabular foundation models, driven by in-context learning, have rapidly grown in quality and popularity. However, recent approaches with either cell-base

Tevatron Meets Megatron: Expert-Parallel LLM Reranker Training on an Academic Budget

Model ReleasesDGX agent

arXiv:2608.00916v1 Announce Type: cross Abstract: Modern reranking recipes---billion-scale cross-encoders, mixture-of-experts (MoE) backbones, and distillation against strong teachers---have outpaced

tFUSOperator: Operator Learning for Transcranial Focused Ultrasound Digital Twins

ResearchDGX agent

arXiv:2608.01839v1 Announce Type: new Abstract: Transcranial focused ultrasound (tFUS) requires accurate estimation of the intracranial acoustic field, which is distorted by skull-induced aberrations.

The Bayesian Reflex: A Predictive Coding Engine for Artificial Intelligence

ResearchDGX agent

arXiv:2608.00492v1 Announce Type: cross Abstract: Predictive coding offers a powerful theory of cortical computation, but corresponding scalable algorithmic implementations for artificial intelligence

The Condition-Number Barrier in Sparse Least Squares

Model ReleasesDGX agent

arXiv:2608.02588v1 Announce Type: cross Abstract: In [AS21], Axiotis and Sviridenko conjectured that the linear dependence on the restricted condition number in sparse convex optimization cannot be im

The Elements of Differentiable Programming

ResearchDGX agent

arXiv:2403.14606v4 Announce Type: replace Abstract: Artificial intelligence has recently experienced remarkable advances, fueled by large models, vast datasets, accelerated hardware, and, last but not

The Fourth Quadrant: A Stylized View of Benign Misfitting

ResearchDGX agent

arXiv:2608.01032v1 Announce Type: new Abstract: Training error is what we can observe on a training set; test error is the quantity we actually care about. We study linear regression with squared-erro

The Label Defines the Timescale: Trait-State Limits of Temporal-Aggregate Learning

ResearchDGX agent

arXiv:2608.01587v1 Announce Type: cross Abstract: Machine-learning benchmarks often pair a label that aggregates a long temporal horizon with input observed through one or a few short windows. Their a

The No-Clash Teaching Dimension is Bounded by VC Dimension

ResearchDGX agent

arXiv:2603.23561v4 Announce Type: replace-cross Abstract: In the realm of machine learning theory, to prevent unnatural coding schemes between teacher and learner, No-Clash Teaching Dimension was intr

Thermalizing Stochastic Programs

ResearchDGX agent

arXiv:2608.01615v1 Announce Type: cross Abstract: We present a set of tools for mapping general stochastic programs to thermodynamic hardware designed for energy-efficient stochastic sampling. Given a

Towards Anomaly Detection on Relational Data

Model ReleasesDGX agent

arXiv:2606.18621v2 Announce Type: replace Abstract: Relational databases are widely used for managing structured data in real-world systems. Detecting anomalies from such relational data is crucial fo

Towards Effective Federated Multimodal Graph Learning via Navigating Multifaceted Heterogeneity

ResearchDGX agent

arXiv:2608.00623v1 Announce Type: new Abstract: Multimodal-attributed graphs (MAGs), where nodes carry heterogeneous semantic content across multiple modalities while edges encode relational dependenc

Towards General Language-Conditioned Latent Safety Filters

SafetyDGX agent

arXiv:2608.00315v1 Announce Type: cross Abstract: Robot policies are becoming increasingly general, with vision-language-action (VLA) models enabling a single policy to execute diverse tasks specified

Training Deep Morphological Neural Networks as Universal Approximators

Model ReleasesDGX agent

arXiv:2505.09710v4 Announce Type: replace Abstract: We investigate deep morphological neural networks (DMNNs), studying how changes in algebraic structure affect the expressivity and trainability of d

Training nGPT

Model ReleasesDGX agent

arXiv:2608.01284v1 Announce Type: new Abstract: The normalized Transformer (nGPT) realizes hyperspherical representation learning by constraining model parameter vectors and activation vectors to the

Training Small LLMs as Spatial Multi-Agent Policies

SafetyDGX agent

arXiv:2608.01425v1 Announce Type: cross Abstract: Training LLM-based multi-agent systems with multi-agent reinforcement learning is rapidly gaining traction, and a parallel line of work argues that su

Trajectories That Segment Themselves: Agent-Declared Boundaries as a Training Unit

AgentsDGX agent

arXiv:2608.02302v1 Announce Type: cross Abstract: Long-horizon coding-agent trajectories are poorly matched to the credit units available to train on: a single action has no stable value, an episode l

Transfer Learning of CATE with Kernel Ridge Regression

ApplicationsDGX agent

arXiv:2502.11331v4 Announce Type: replace-cross Abstract: The proliferation of data has sparked significant interest in leveraging findings from one study to estimate treatment effects in a different

Trust or Check? Understanding the (Evolutionary) Dynamics of User Trust in AI Systems

SafetyDGX agent

arXiv:2603.24742v2 Announce Type: replace-cross Abstract: As the capabilities and adoption of Artificial Intelligence (AI) systems grow, trust in these AI systems is an increasingly urgent concern. Mu

Trustworthy AI in Digital Health: A Comprehensive Review of Robustness and Explainability

SafetyDGX agent

arXiv:2608.02238v1 Announce Type: cross Abstract: Ensuring trust in AI systems is essential for the safe and ethical integration of machine learning systems into high-stakes domains such as digital he

Tunneling the Loss Landscape: Bypassing Memorization with Monte Carlo Parameter Swapping

Model ReleasesDGX agent

arXiv:2608.01833v1 Announce Type: cross Abstract: Grokking is a striking phenomenon in neural network training, where a model can undergo a prolonged period of pure memorization before abrupt generali

Uncertainty-Aware Simulation-Based Inference for Operations Research with Large Language Models

Local AiDGX agent

arXiv:2608.00019v1 Announce Type: new Abstract: Deploying large language models (LLMs) for operations research (OR) tasks remains challenging because correctness depends on a coherent modeling process

Uncertainty-guided active learning for surrogate prediction of stream-finishing wear fields

ResearchDGX agent

arXiv:2608.00593v1 Announce Type: cross Abstract: In stream finishing, the wear experienced by a workpiece depends strongly on its orientation within the rotating abrasive media. Determining suitable

← Previous
1…1617181920…239
Next →