AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,557 results
29 May 2026

Echoes within the Reasoning: Stealthy and Effective Watermarking via Chain of Thought

ResearchDGX agent

arXiv:2605.28890v1 Announce Type: cross Abstract: Large Language Models with Chain-of-Thought reasoning capabilities represent valuable intellectual property, yet existing black-box watermarking metho

Efficient Test-Time Finetuning of LLMs via Convex Reconstruction and Gradient Caching

ResearchDGX agent

arXiv:2605.30337v1 Announce Type: new Abstract: Test-time finetuning (TTFT) is a rapidly evolving paradigm that adapts a language model to each prompt by retrieving related sequences, updating the mod

Eigen-Spike Emergence and Quadratic Equivalents for Conjugate Kernels on Nonlinearly Separable Data

ResearchDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2605.29669v1 Announce Type: cross Abstract: Recent work in random matrix theory (RMT) has developed the notion of deterministic equivalents: typically linear surrogate models that approximate th

EMAG: Differentiable 4D Gaussian Mixture Splatting for EEG Spatial Super-Resolution

ResearchDGX agent

arXiv:2605.29731v1 Announce Type: new Abstract: High-density electroencephalography (HD-EEG) enables fine-grained measurement of cortical activity but requires expensive hardware and lengthy setup tim

Enhancing LLM Training via Spectral Clipping

ResearchDGX agent

arXiv:2603.14315v2 Announce Type: replace Abstract: While spectral-based optimizers like Muon operate directly on the spectrum of updates, standard adaptive methods such as AdamW do not account for th

Enhancing Membership Inference Attacks on Diffusion Models from a Frequency-Domain Perspective

ResearchDGX agent

arXiv:2505.20955v4 Announce Type: replace-cross Abstract: Diffusion models have achieved tremendous success in image generation, but they also raise significant concerns regarding privacy and copyrigh

Ensemble Score Filtering for Real-Data Energy Consumption Forecast Correction

ResearchDGX agent

arXiv:2605.29072v1 Announce Type: new Abstract: Accurate estimation and forecasting of energy consumption are important for power-system operation, planning, and demand-side management. In practice, h

Envy-Free Allocation of Indivisible Goods via Noisy Queries

AgentsDGX agent

arXiv:2602.06361v2 Announce Type: replace-cross Abstract: We introduce a problem of fairly allocating indivisible goods (items) in which the agents' valuations cannot be observed directly, but instead

ExDBSCAN: Explaining DBSCAN with Counterfactual Reasoning -- Additional Material

ResearchDGX agent

arXiv:2605.30225v1 Announce Type: new Abstract: Clustering is an unsupervised technique for grouping data points by similarity. While explainability methods exist for supervised machine learning, they

Explaining Concept Shift with Interpretable Feature Attribution

TutorialsDGX agent

arXiv:2505.20634v2 Announce Type: replace Abstract: Concept shift occurs when the distribution of labels conditioned on the features changes between domains, which can make even a well-tuned ML model

Fairness-Aware Federated Learning with Trajectory Shapley Value

Model ReleasesDGX agent

arXiv:2605.30336v1 Announce Type: new Abstract: Federated learning is an emerging distributed paradigm that addresses the challenges posed by heterogeneous, privacy-sensitive data. It enables multiple

Faithful Embeddings of Irregular and Asynchronous Data for Online Log-NCDEs

TutorialsDGX agent

arXiv:2605.30213v1 Announce Type: new Abstract: Continuous-time models are a natural choice for irregular and asynchronous data. A central design choice is how to embed discrete observations into cont

FarSkip-Collective: Unhobbling Blocking Communication in Mixture of Experts Models

Model ReleasesDGX agent

arXiv:2511.11505v3 Announce Type: replace Abstract: Blocking communication presents a major hurdle in running MoEs efficiently in distributed settings. To address this, we present FarSkip-Collective w

Faster Molecular Dynamics with Neural Network Potentials via Distilled Multiple Time-Stepping and Non-Conservative Forces

ApplicationsDGX agent

arXiv:2602.14975v3 Announce Type: replace-cross Abstract: Following our previous work (J. Phys. Chem. Lett., 2026, 17, 5, 1288-1295), we propose the DMTS-NC approach, a distilled multi-time-step (DMTS

Feature Geometry of LoRA Adapters: A Sparse Autoencoder Analysis of Representational Divergence in Fine-Tuned Language Models

Model ReleasesDGX agent

arXiv:2605.28896v1 Announce Type: new Abstract: Low-Rank Adaptation (LoRA) has emerged as a widely adopted approach for adapting large language models, yet the internal representational changes induce

FedQHD: Closed-Form Function-Space Federated Reinforcement Learning

Model ReleasesDGX agent

arXiv:2605.29002v1 Announce Type: new Abstract: Federated reinforcement learning enables decentralized agents to collaboratively improve policies or value estimates without exchanging raw trajectories

Feedback-to-Rubrics: Can We Learn Expert Criteria from Inline Comments?

TutorialsDGX agent

arXiv:2605.29857v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used for writing and review support, but their usefulness depends on context-dependent criteria, such as e

Financially Guided Deep Portfolio Optimization

TutorialsDGX agent

arXiv:2605.28853v1 Announce Type: cross Abstract: Portfolio optimization in real-world financial markets is notoriously difficult due to non-stationarity, noisy data, and high transaction costs. Stand

Fingerprinting Inference Systems of Large Language Models

ResearchDGX agent

arXiv:2605.29979v1 Announce Type: cross Abstract: The behavior of LLMs does not depend solely on the model itself. Components of the inference system, such as the inference engine, attention backend,

Fisher-Preserving Guidance: Training-Free Manifold Constraints for Safe Diffusion Control

SafetyDGX agent

arXiv:2605.29937v1 Announce Type: cross Abstract: Diffusion models are effective for waypoint prediction in visual navigation, but standard sampling and test time guidance can produce unreliable or in

Follow-the-Perturbed-Leader for Decoupled Bandits: Best-of-Both-Worlds and Practicality

SafetyDGX agent

arXiv:2510.12152v2 Announce Type: replace-cross Abstract: We study the decoupled multi-armed bandit problem, where the learner separately selects one arm for exploration and one, possibly different, a

FPLIER: Federated Pathway-Level Information Extractor

Model ReleasesDGX agent

arXiv:2605.29587v1 Announce Type: cross Abstract: In transcriptomics, gene-set-aware factorization methods such as the Pathway Level Information Extractor (PLIER) are most effective when trained on la

From Short Histories to Long Futures: Horizon-Aware Graph Neural Networks for Long Horizon Forecasting

Local AiDGX agent

arXiv:2605.29952v1 Announce Type: new Abstract: Accurate long-range prediction of geophysical systems is difficult due to strongly nonlinear dynamics, the high computational cost of full-physics simul

From Sublinear to Linear: Local Convergence in Finite-Width Networks via Locally Polyak-Lojasiewicz Regions

Model ReleasesDGX agent

arXiv:2507.21429v3 Announce Type: replace-cross Abstract: We study local linear convergence of gradient descent for finite-width feedforward networks under the squared empirical loss. Prior work shows

Gated Graph Attention Networks with Learnable Temperature

TutorialsDGX agent

arXiv:2605.29803v1 Announce Type: new Abstract: Graph attention networks learn neighbor importance through data-dependent coefficients, but standard layers lack explicit control over unreliable featur

Generative Spatiotemporal Intent Sequence Recommendation via Implicit Reasoning in Amap

ApplicationsDGX agent

arXiv:2605.28888v1 Announce Type: cross Abstract: Real-world user behavior rarely consists of isolated actions; instead, it often forms intent flows governed by spatiotemporal dependencies. To provide

Gesture-Aware Indoor THz ISAC Systems for Adaptive Resource Allocation

ResearchDGX agent

arXiv:2605.29913v1 Announce Type: cross Abstract: This paper investigates a multi-user indoor integrated sensing and communication (ISAC) system operating in the terahertz (THz) band, designed for ada

Gradient Perturbation: Learning to Perturb Gradients for Adaptive Training

Model ReleasesDGX agent

arXiv:2605.29494v1 Announce Type: new Abstract: Deep neural network training involves both forward propagation (from features through logits to loss) and backward propagation (from loss through gradie

Gradient Preconditioning for Efficient and Reliable Reward-Guided Generation

ResearchDGX agent

arXiv:2602.08646v2 Announce Type: replace Abstract: We propose a gradient preconditioning method that makes reward-guided generation with one-step generative models both efficient and reliable. Test-t

Harmless Yet Harmful: Neutral Prompting Attacks for Stealthy Hallucination Steering in Agent Skills

AgentsDGX agent

arXiv:2605.29354v1 Announce Type: cross Abstract: LLM-powered coding agents increasingly participate in software development workflows by generating code, selecting dependencies, and producing package

Horizon Activation Mapping for Neural Networks in Time Series Forecasting

ResearchDGX agent

arXiv:2601.02094v4 Announce Type: replace Abstract: Neural networks for time series forecasting have relied on error metrics and architecture-specific interpretability approaches for model selection t

Improving Full Waveform Inversion in Large Model Era

Model ReleasesDGX agent

arXiv:2603.00377v2 Announce Type: replace Abstract: Full Waveform Inversion (FWI) is a highly nonlinear and ill-posed problem that aims to recover subsurface velocity maps from surface-recorded seismi

In-Place Feedback: Reliable Refinement for Multi-Turn Expert-LLM Collaboration

ResearchDGX agent

arXiv:2510.00777v2 Announce Type: replace Abstract: LLM-generated drafts often contain subtle factual or logical errors, yet prior work shows that models struggle to reliably integrate multi-turn feed

Inferring the Size of Large Language Models From Popular Text Memorization

Model ReleasesDGX agent

arXiv:2605.29223v1 Announce Type: new Abstract: The parameter counts of the most widely used large language models (LLMs) are often withheld by their developers, leaving model size -- a primary refere

Information-Directed Offline-to-Online Reinforcement Learning

SafetyDGX agent

arXiv:2605.29405v1 Announce Type: new Abstract: Decision-making from offline datasets typically warm-starts a policy or score model from fixed offline data and then refines it with limited online inte

Instance-dependent Stochastic Lipschitz bandit

ResearchDGX agent

arXiv:2605.29748v1 Announce Type: cross Abstract: We study the Lipschitz bandit problem, where a learner sequentially maximizes an unknown Lipschitz function f over a domain X subset [0,1]^d using noi

Is Your Diffusion Sampler Actually Correct? A Sampler-Centric Evaluation of Discrete Diffusion Language Models

ResearchDGX agent

arXiv:2602.19619v2 Announce Type: replace Abstract: Discrete diffusion language models (dLLMs) provide a fast and flexible alternative to autoregressive models (ARMs) via iterative denoising with para

Joint Model and Data Sparsification via the Marginal Likelihood

ResearchDGX agent

arXiv:2605.29908v1 Announce Type: cross Abstract: Sparse recovery in linear systems underpins applications from signal processing to high-dimensional regression. Sparse Bayesian Learning, grounded in

K-FinHallu: A Hallucination Detection Benchmark for Multi-Turn RAG in Korean Finance

Model ReleasesDGX agent

arXiv:2605.29523v1 Announce Type: new Abstract: Large Language Models (LLMs) have advanced financial automation through Retrieval-Augmented Generation (RAG), yet hallucinations remain a critical barri

KAN-AD: Time Series Anomaly Detection with Kolmogorov-Arnold Networks

Local AiDGX agent

arXiv:2411.00278v4 Announce Type: replace Abstract: Time series anomaly detection (TSAD) underpins real-time monitoring in cloud services and web systems, allowing rapid identification of anomalies to

Kernel-based potential mean-field games with unbiased random Fourier U-statistics

Model ReleasesDGX agent

arXiv:2605.29371v1 Announce Type: cross Abstract: We study the subclass of potential mean-field games in which the running interaction cost and the terminal target cost are both expressed through repr

Kernel Renormalization in Bayesian Deep Neural Networks: the Equivalent Wishart Ansatz in the Proportional Regime

Model ReleasesDGX agent

arXiv:2605.29684v1 Announce Type: new Abstract: The scaling limit where both the size of the training set P and the width N of a deep neural network grow at the same rate, the so-called proportional-w

Knowledge Offloading: Decomposing LLMs into Sparse Backbones and Memory Modules

Model ReleasesDGX agent

arXiv:2605.29075v1 Announce Type: new Abstract: LLMs encode both general capabilities and domain-specific knowledge in a single set of parameters. We ask whether this capacity can be reorganized: keep

Leak@k: Unlearning Does Not Make LLMs Forget Under Probabilistic Decoding

Model ReleasesDGX agent

arXiv:2511.04934v3 Announce Type: replace Abstract: Unlearning in large language models (LLMs) is critical for regulatory compliance and for building ethical generative AI systems that avoid producing

Learning Robust and Task-Invariant Functional Representation from fMRI through Siamese Self-Supervised Learning

TutorialsDGX agent

arXiv:2605.28990v1 Announce Type: new Abstract: Functional magnetic resonance imaging (fMRI) is a powerful tool for investigating human brain function. However, the high cost of data acquisition and t

Learning to Extrapolate to New Tasks: A Relational Approach to Task Extrapolation

Model ReleasesDGX agent

arXiv:2605.30132v1 Announce Type: new Abstract: Modern learning systems excel at interpolation but struggle to generalize to unseen tasks outside the training distribution's support. This failure occu

Learning to Perturb Hidden Representations for Generalizable Deep Learning

ResearchDGX agent

arXiv:2605.29525v1 Announce Type: new Abstract: Deep neural networks process data through a cascade of representations: input features, hidden activations, logits, and loss. While perturbations at the

Learning to Solve PDEs on Neural Shape Representations

Local AiDGX agent

arXiv:2512.21311v2 Announce Type: replace Abstract: Solving partial differential equations (PDEs) on shapes underpins many shape analysis and engineering tasks; yet, prevailing PDE solvers operate on

Leave a Window Out: Modifying the Jackknife for Predictive Inference in Time Series

ResearchDGX agent

arXiv:2605.30292v1 Announce Type: cross Abstract: Conformal prediction methods enjoy strong theoretical and empirical predictive inference performance, provided the data is exchangeable, and predictor

libhmm: A Modern C++20 Library for Hidden Markov Models with Correct MLE Emission M-Steps

Model ReleasesDGX agent

arXiv:2605.29208v1 Announce Type: cross Abstract: We describe libhmm, a C++20 library for Hidden Markov Model parameter estimation, sequence decoding, and model selection. libhmm addresses two gaps in

Looking around you: external information enhances representations for event sequences

ApplicationsDGX agent

arXiv:2502.10205v3 Announce Type: replace Abstract: Representation learning produces models in different domains, such as store purchases, client transactions, and general people's behavior. However,

Manifold-based Algorithms for the Hadamard Decomposition

ResearchDGX agent

arXiv:2605.28980v1 Announce Type: cross Abstract: Given a matrix X, and two ranks r_1 and r_2, the Hadamard decomposition (HD) looks for two low-rank matrices, X_1 of rank r_1 and X_2 of rank r_2, bot

MarginGate: Sparse Margin-Triggered Verification for Batch-Invariant LLM Inference

Model ReleasesDGX agent

arXiv:2605.30218v1 Announce Type: new Abstract: Temperature-zero BF16 LLM inference is often treated as reproducible, yet the same request can emit different tokens when decoded alone or inside a larg

Matching Rates and Optimal Allocation for Federated Probe-Logit Distillation under Heterogeneous Bandwidth Budgets

ResearchDGX agent

arXiv:2605.29642v1 Announce Type: cross Abstract: In federated language modeling, K nodes each hold n samples but cannot pool data or exchange full-precision gradients or weights. We study the minimax

Matrix Completion with Hypergraphs:Sharp Thresholds and Efficient Algorithms

ResearchDGX agent

arXiv:2401.08197v3 Announce Type: replace Abstract: This paper considers the problem of completing a rating matrix based on sub-sampled matrix entries as well as observed social graphs and hypergraphs

Mean-Field Diffuser: Scaling Offline MARL to Thousands of Agents

SafetyDGX agent

arXiv:2605.30190v1 Announce Type: new Abstract: Diffusion-based planning has achieved strong results in single-agent offline reinforcement learning, yet scaling to many-agent systems remains intractab

Measure flow path recovery in Bayes Hilbert spaces

ResearchDGX agent

arXiv:2603.20329v2 Announce Type: replace-cross Abstract: We study the ill-posed problem of recovering a probability measure flow from finitely many moving localized sensors using a Bayes Hilbert fram

MEC: Machine-Learning-Assisted Generalized Entropy Calibration for Semi-Supervised Mean Estimation

ResearchDGX agent

arXiv:2604.05446v2 Announce Type: replace-cross Abstract: Obtaining high-quality labels is costly, whereas unlabeled covariates are often abundant, motivating semi-supervised inference methods with re

Midpoint Generative Models

ResearchDGX agent

arXiv:2605.29920v1 Announce Type: new Abstract: We introduce Midpoint Generative Models (MGM), a principled framework for training one-step generative models. MGM is based on a simple symmetry of Flow

MIRAGE: Adaptive Multimodal Gating for Whole-Brain fMRI Encoding

ResearchDGX agent

arXiv:2605.29850v1 Announce Type: new Abstract: Recent progress in task-optimized neural networks has established encoding models as a powerful tool for predicting brain responses to naturalistic stim

← Previous
1…115116117118119…243
Next →