AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlog
84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,557 results
Research

On Statistical Estimation of Edge-Reinforced Random Walks

DGX agent

arXiv:2503.06115v2 Announce Type: replace-cross Abstract: Reinforced random walks (RRWs), including vertex-reinforced random walks (VRRWs) and edge-reinforced random walks (ERRWs), model random walks

researcharxiv-cs-lg
23 May 2026
Safety
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

On the Sample Complexity of Discounted Reinforcement Learning with Optimized Certainty Equivalents

DGX agent

arXiv:2605.21763v1 Announce Type: new Abstract: We study risk-sensitive reinforcement learning in finite discounted MDPs, where a generative model of the MDP is assumed to be available. We consider a

safetyarxiv-cs-lg
23 May 2026
Model Releases

One LR Doesn't Fit All: Heavy-Tail Guided Layerwise Learning Rates for LLMs

DGX agent

arXiv:2605.22297v1 Announce Type: new Abstract: Learning rate configuration is a fundamental aspect of modern deep learning. The prevailing practice of applying a uniform learning rate across all laye

model-releasesarxiv-cs-lg
23 May 2026
Safety

One-Way Policy Optimization for Self-Evolving LLMs

DGX agent

arXiv:2605.22156v1 Announce Type: new Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has become a promising paradigm for scaling reasoning capabilities of Large Language Models (LLMs)

safetyarxiv-cs-lg
23 May 2026
Local Ai

OPPO: Bayesian Value Recursion for Token-Level Credit Assignment in LLM Reasoning

DGX agent

arXiv:2605.21851v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards has become the standard recipe for improving LLM reasoning, but the dominant algorithm GRPO assigns a sin

local-aiarxiv-cs-lg
23 May 2026
Research

Optimal Guarantees for Auditing Renyi Differentially Private Machine Learning

DGX agent

arXiv:2605.21938v1 Announce Type: new Abstract: We study black-box auditing for machine learning algorithms that claim R 'enyi differential privacy (RDP) guarantees. We introduce an auditing framework

researcharxiv-cs-lg
23 May 2026
Research

Optimization over the intersection of manifolds

DGX agent

arXiv:2605.22736v1 Announce Type: cross Abstract: Optimization over the intersection of two manifolds arises in a broad range of applications, but is hindered by the coupled geometry of the feasible r

researcharxiv-cs-lg
23 May 2026
Research

Partial Fusion of Neural Networks: Efficient Tradeoffs Between Ensembles and Weight Aggregation

DGX agent

arXiv:2605.22350v1 Announce Type: new Abstract: Ensembles of neural networks typically outperform individual networks but incur large computational costs, whereas weight aggregation produces less cost

researcharxiv-cs-lg
23 May 2026
Research

PeakFocus: Bridging Peak Localization and Intensity Regression via a Unified Multi-Scale Framework for Electricity Load Forecasting

DGX agent

arXiv:2605.21550v1 Announce Type: new Abstract: Electricity load peak forecasting (ELPF), simultaneously predicting peak timing and intensity, is a prerequisite for effective grid scheduling and risk

researcharxiv-cs-lg
23 May 2026
Safety

PEARL: Unbiased Percentile Estimation via Contrastive Learning for Industrial-Scale Livestream Recommendation

DGX agent

arXiv:2605.21752v1 Announce Type: new Abstract: Recommender systems trained on user interaction data are susceptible to behavioral intensity imbalance--a systematic distortion arising from heterogeneo

safetyarxiv-cs-lg
23 May 2026
Tutorials

PhylaFlow: Hybrid Flow Matching in Billera-Holmes-Vogtmann Tree Space for Phylogenetic Inference

DGX agent

arXiv:2605.21859v1 Announce Type: cross Abstract: Phylogenetic trees are hybrid objects: branch lengths vary continuously, while topologies change discretely through edge contractions and expansions.

tutorialsarxiv-cs-lg
23 May 2026
Applications

Physics-Informed Generative Solver: Bridging Data-Driven Priors and Conservation Laws for Stable Spatiotemporal Field Reconstruction

DGX agent

arXiv:2605.22338v1 Announce Type: new Abstract: Reconstructing continuous physical fields from sparse measurements is a central inverse problem, but data-driven generative models can produce states th

applicationsarxiv-cs-lg
23 May 2026
Research

Physics Priors Offer Useful Accuracy-Carbon Trade-Offs in Spatio-Temporal Forecasting

DGX agent

arXiv:2509.24517v2 Announce Type: replace Abstract: Development of modern deep learning methods has been driven primarily by the push for improving model efficacy (accuracy metrics). This sole focus o

researcharxiv-cs-lg
23 May 2026
Applications

Plug-in Losses for Evidential Deep Learning: A Simplified Framework for Uncertainty Estimation that Includes the Softmax Classifier

DGX agent

arXiv:2605.22746v1 Announce Type: new Abstract: Real-world sensor-based learning systems require uncertainty estimation that is both reliable and computationally efficient. Evidential Deep Learning (E

applicationsarxiv-cs-lg
23 May 2026
Tutorials

Position: The Time for Sampling Is Now! Charting a New Course for Bayesian Deep Learning

DGX agent

arXiv:2605.21765v1 Announce Type: new Abstract: The practical adoption of sampling-based inference (SAI) in Bayesian neural networks (BNNs) remains limited, partly due to persistent misconceptions abo

tutorialsarxiv-cs-lg
23 May 2026
Safety

Post-Training is About States, Not Tokens: A State Distribution View of SFT, RL, and On-Policy Distillation

DGX agent

arXiv:2605.22731v1 Announce Type: new Abstract: Large language model post-training methods such as supervised fine-tuning (SFT), reinforcement learning (RL), and distillation are often analyzed throug

safetyarxiv-cs-lg
23 May 2026
Model Releases

Posterior Collapse as Automatic Spectral Pruning

DGX agent

arXiv:2605.22691v1 Announce Type: new Abstract: We show that posterior collapse in eta-VAEs implements automatic spectral pruning. A latent mode collapses if its contribution to reconstruction is belo

model-releasesarxiv-cs-lg
23 May 2026
Research

Predicting Performance of Symbolic and Prompt Programs with Examples

DGX agent

arXiv:2605.21515v1 Announce Type: new Abstract: LLM prompting is widely used for naturally stated tasks, yet it is unreliable it may succeed on a few test cases but fail at deployment time. We study p

researcharxiv-cs-lg
23 May 2026
Model Releases

Prior Knowledge-enhanced Spatio-temporal Epidemic Forecasting

DGX agent

arXiv:2602.22270v2 Announce Type: replace Abstract: Spatio-temporal epidemic forecasting is critical for public health management, yet existing methods often struggle with insensitivity to weak epidem

model-releasesarxiv-cs-lg
23 May 2026
Research

Prior shift estimation for positive unlabeled data through the lens of kernel embedding

DGX agent

arXiv:2502.21194v3 Announce Type: replace-cross Abstract: We study estimation of a class prior for unlabeled target samples which possibly differs from that of source population. Moreover, it is assum

researcharxiv-cs-lg
23 May 2026
Model Releases

Protein Thoughts: Interpretable Reasoning with Tree of Thoughts and Embedding-Space Flow Matching for Protein-Protein Interaction Discovery

DGX agent

arXiv:2605.21522v1 Announce Type: cross Abstract: Protein-protein interactions (PPIs) govern nearly all cellular processes, yet computational methods for identifying binding partners typically produce

model-releasesarxiv-cs-lg
23 May 2026
Research

Prototype-Guided Classification Sub-Task Decoupling Framework: Enhancing Generalization and Interpretability for Multivariate Time Series

DGX agent

arXiv:2605.22055v1 Announce Type: new Abstract: Time Series Classification (TSC) is a long-standing research problem that has gained increasing attention in recent years with the rapid growth of large

researcharxiv-cs-lg
23 May 2026
Model Releases

Provable Joint Decontamination for Benchmarking Multiple Large Language Models

DGX agent

arXiv:2605.21543v1 Announce Type: new Abstract: Benchmark data contamination has become a central challenge in LLM evaluation: when evaluation examples appear in the training data of one or more audit

model-releasesarxiv-cs-lg
23 May 2026
Applications

Provable Robustness against Backdoor Attacks via the Primal-Dual Perspective on Differential Privacy

DGX agent

arXiv:2605.21780v1 Announce Type: new Abstract: Randomized smoothing is a powerful tool for certifying robustness to adversarial perturbations, including poisoning attacks via randomized training and

applicationsarxiv-cs-lg
23 May 2026
Research

Provably Protecting Fine-Tuned LLMs from Training Data Extraction while Preserving Utility

DGX agent

arXiv:2602.00688v2 Announce Type: replace Abstract: Fine-tuning large language models (LLMs) on sensitive datasets raises privacy concerns, as training data extraction (TDE) attacks can expose highly

researcharxiv-cs-lg
23 May 2026
Safety

Proxy-Based Approximation of Shapley and Banzhaf Interactions

DGX agent

arXiv:2605.22738v1 Announce Type: new Abstract: Shapley and Banzhaf interactions capture the complex dynamics inherent in modern machine learning applications. However, current estimators for these hi

safetyarxiv-cs-lg
23 May 2026
Research

Q-PhotoNAS: Hybrid Quantum Neural Architecture Search Framework on Photonic Devices

DGX agent

arXiv:2605.22097v1 Announce Type: cross Abstract: Photonic quantum computing is a promising platform for scalable quantum machine learning, but designing effective hybrid architectures remains challen

researcharxiv-cs-lg
23 May 2026
Research

Quantitative coronary calcification analysis for prediction of myocardial ischemia using non-contrast CT calcium scoring

DGX agent

arXiv:2605.21745v1 Announce Type: new Abstract: Non-contrast computed tomography calcium scoring (CTCS) is widely recognized as an effective tool for cardiovascular risk stratification. This study aim

researcharxiv-cs-lg
23 May 2026
Research

RADAR: Defending RAG Dynamically against Retrieval Corruption

DGX agent

arXiv:2605.22041v1 Announce Type: cross Abstract: While RAG systems are increasingly deployed in dynamic web search, temporal volatility amplifies their vulnerability to adversarial attacks. Existing

researcharxiv-cs-lg
23 May 2026
Safety

[Re] FairDICE: A Fair Tradeoff in Multi-objective Offline RL

DGX agent

arXiv:2603.03454v2 Announce Type: replace Abstract: Offline Reinforcement Learning (RL) is an emerging field of RL in which policies are learned solely from demonstrations. Within offline RL, some env

safetyarxiv-cs-lg
23 May 2026
Hardware

Reading Task Failure Off the Activations: A Sparse-Feature Audit of GPT-2 Small on Indirect Object Identification

DGX agent

arXiv:2605.22719v1 Announce Type: new Abstract: We report a small, reproducible audit of which sparse-autoencoder (SAE) features of GPT-2 small fire differently on failed versus successful trials of t

hardwarearxiv-cs-lg
23 May 2026
Model Releases

Reasoning through Verifiable Forecast Actions: Consistency-Grounded RL for Financial LLMs

DGX agent

arXiv:2605.21975v1 Announce Type: new Abstract: Financial markets are characterized by extreme non-stationarity, low signal-to-noise ratios, and strong dependence on external information such as news,

model-releasesarxiv-cs-lg
23 May 2026
Research

Regret-Based (epsilon,elta)-optimal Stopping Criteria for Bayesian Optimization

DGX agent

arXiv:2605.22561v1 Announce Type: new Abstract: Bayesian optimization (BO) is a widely used iterative black-box optimization method that utilizes Gaussian process (GP) surrogate models. In practice, B

researcharxiv-cs-lg
23 May 2026
Research

Reinforced Graph of Thoughts: RL-Driven Adaptive Prompting for LLMs

DGX agent

arXiv:2605.22195v1 Announce Type: new Abstract: Graph of Thoughts (GoT), a generalized form of recent prompting paradigms for large language models (LLMs), has been shown to be useful for elaborate pr

researcharxiv-cs-lg
23 May 2026
Research

Reinforcement learning for ion shuttling on trapped-ion quantum computers

DGX agent

arXiv:2605.22463v1 Announce Type: cross Abstract: Scalable trapped-ion quantum computing is commonly realized with modular chips that feature distinct zones with specific functionalities, such as stor

researcharxiv-cs-lg
23 May 2026
Research

Relational Linear Properties in Language Models: An Empirical Investigation

DGX agent

arXiv:2605.22532v1 Announce Type: new Abstract: Linear properties are ubiquitous in the representations of language models; however, testing them experimentally remains a challenging task. This work f

researcharxiv-cs-lg
23 May 2026
Local Ai

Reliable Wireless Indoor Localization via Cross-Validated Prediction-Powered Calibration

DGX agent

arXiv:2507.20268v3 Announce Type: replace Abstract: Wireless indoor localization using predictive models with received signal strength information (RSSI) requires proper calibration for reliable posit

local-aiarxiv-cs-lg
23 May 2026
Local Ai

Remember to be Curious: Episodic Context and Persistent Worlds for 3D Exploration

DGX agent

arXiv:2605.22814v1 Announce Type: new Abstract: Exploration is a prerequisite for learning useful behaviors in sparse-reward, long-horizon tasks, particularly within 3D environments. Curiosity-driven

local-aiarxiv-cs-lg
23 May 2026
Model Releases

Representation Gap: Explaining the Unreasonable Effectiveness of Neural Networks from a Geometric Perspective

DGX agent

arXiv:2605.21692v1 Announce Type: new Abstract: Characterizing precisely the asymptotic generalization error of neural networks using parameters that can be estimated efficiently is a crucial problem

model-releasesarxiv-cs-lg
23 May 2026
Local Ai

Represented Is Not Computed: A Causal Test of Candidate Algorithmic Intermediates in a Transformer

DGX agent

arXiv:2605.22488v1 Announce Type: new Abstract: Structured prompts require integrating components according to task-relevant relations. How a network implements this integration is often hard to judge

local-aiarxiv-cs-lg
23 May 2026
Model Releases

Rethinking Forward Processes for Score-Based Nonlinear Data Assimilation in High Dimensions

DGX agent

arXiv:2604.02889v2 Announce Type: replace-cross Abstract: Data assimilation is the process of estimating the state of a dynamical system over time by combining model predictions with measurements. Thi

model-releasesarxiv-cs-lg
23 May 2026
Safety

Revisiting Regularized Policy Optimization for Stable and Efficient Reinforcement Learning in Two-Player Games

DGX agent

arXiv:2602.10894v2 Announce Type: replace Abstract: Two-player games such as board games have long been used as traditional benchmarks for reinforcement learning. This work revisits a policy optimizat

safetyarxiv-cs-lg
23 May 2026
Model Releases

Revisiting Robustness for LLM Safety Alignment via Selective Geometry Control

DGX agent

arXiv:2602.07340v2 Announce Type: replace Abstract: Safety alignment of large language models remains brittle under domain shift and noisy preference supervision. Most existing robust alignment method

model-releasesarxiv-cs-lg
23 May 2026
Research

Richer Bayesian Last Layers with Subsampled NTK Features

DGX agent

arXiv:2602.01279v2 Announce Type: replace Abstract: Bayesian Last Layers (BLLs) provide a convenient and computationally efficient way to estimate uncertainty in neural networks. However, they underes

researcharxiv-cs-lg
23 May 2026
Applications

Riemannian geometry meets fMRI: the advantages of modeling correlation manifolds and eigenvector subspaces

DGX agent

arXiv:2605.22334v1 Announce Type: new Abstract: Correlation matrices are fundamental summaries of functional brain networks, yet standard analyses often treat entries independently, ignoring the curve

applicationsarxiv-cs-lg
23 May 2026
Model Releases

RobustSpeechFlow: Learning Robust Text-to-Speech Trajectories via Augmentation-based Contrastive Flow Matching

DGX agent

arXiv:2605.22083v1 Announce Type: cross Abstract: While flow-matching text-to-speech (TTS) achieves strong zero-shot speaker similarity and naturalness, it remains susceptible to content fidelity issu

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

Rule-State Inference (RSI): A Bayesian Framework for Compliance Monitoring in Rule-Governed Domains

DGX agent

arXiv:2603.21610v2 Announce Type: replace Abstract: Compliance monitoring in rule-governed domains (tax administration, clinical protocol adherence, environmental regulation) faces three structural ob

model-releasesarxiv-cs-lg
23 May 2026
Research

Same Architecture, Different Capacity: Optimizer-Induced Spectral Scaling Laws

DGX agent

arXiv:2605.21803v1 Announce Type: new Abstract: Scaling laws have made language-model performance predictable from model size, data, and compute, but they typically treat the optimizer as a fixed trai

researcharxiv-cs-lg
23 May 2026
← Previous
1…169170171172173…304
Next →