AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,557 results
23 May 2026

On Statistical Estimation of Edge-Reinforced Random Walks

ResearchDGX agent

arXiv:2503.06115v2 Announce Type: replace-cross Abstract: Reinforced random walks (RRWs), including vertex-reinforced random walks (VRRWs) and edge-reinforced random walks (ERRWs), model random walks

On the Sample Complexity of Discounted Reinforcement Learning with Optimized Certainty Equivalents

SafetyDGX agent

arXiv:2605.21763v1 Announce Type: new Abstract: We study risk-sensitive reinforcement learning in finite discounted MDPs, where a generative model of the MDP is assumed to be available. We consider a

One LR Doesn't Fit All: Heavy-Tail Guided Layerwise Learning Rates for LLMs

Model ReleasesDGX agent

arXiv:2605.22297v1 Announce Type: new Abstract: Learning rate configuration is a fundamental aspect of modern deep learning. The prevailing practice of applying a uniform learning rate across all laye


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

One-Way Policy Optimization for Self-Evolving LLMs

SafetyDGX agent

arXiv:2605.22156v1 Announce Type: new Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has become a promising paradigm for scaling reasoning capabilities of Large Language Models (LLMs)

OPPO: Bayesian Value Recursion for Token-Level Credit Assignment in LLM Reasoning

Local AiDGX agent

arXiv:2605.21851v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards has become the standard recipe for improving LLM reasoning, but the dominant algorithm GRPO assigns a sin

Optimal Guarantees for Auditing Renyi Differentially Private Machine Learning

ResearchDGX agent

arXiv:2605.21938v1 Announce Type: new Abstract: We study black-box auditing for machine learning algorithms that claim R 'enyi differential privacy (RDP) guarantees. We introduce an auditing framework

Optimization over the intersection of manifolds

ResearchDGX agent

arXiv:2605.22736v1 Announce Type: cross Abstract: Optimization over the intersection of two manifolds arises in a broad range of applications, but is hindered by the coupled geometry of the feasible r

Partial Fusion of Neural Networks: Efficient Tradeoffs Between Ensembles and Weight Aggregation

ResearchDGX agent

arXiv:2605.22350v1 Announce Type: new Abstract: Ensembles of neural networks typically outperform individual networks but incur large computational costs, whereas weight aggregation produces less cost

PeakFocus: Bridging Peak Localization and Intensity Regression via a Unified Multi-Scale Framework for Electricity Load Forecasting

ResearchDGX agent

arXiv:2605.21550v1 Announce Type: new Abstract: Electricity load peak forecasting (ELPF), simultaneously predicting peak timing and intensity, is a prerequisite for effective grid scheduling and risk

PEARL: Unbiased Percentile Estimation via Contrastive Learning for Industrial-Scale Livestream Recommendation

SafetyDGX agent

arXiv:2605.21752v1 Announce Type: new Abstract: Recommender systems trained on user interaction data are susceptible to behavioral intensity imbalance--a systematic distortion arising from heterogeneo

PhylaFlow: Hybrid Flow Matching in Billera-Holmes-Vogtmann Tree Space for Phylogenetic Inference

TutorialsDGX agent

arXiv:2605.21859v1 Announce Type: cross Abstract: Phylogenetic trees are hybrid objects: branch lengths vary continuously, while topologies change discretely through edge contractions and expansions.

Physics-Informed Generative Solver: Bridging Data-Driven Priors and Conservation Laws for Stable Spatiotemporal Field Reconstruction

ApplicationsDGX agent

arXiv:2605.22338v1 Announce Type: new Abstract: Reconstructing continuous physical fields from sparse measurements is a central inverse problem, but data-driven generative models can produce states th

Physics Priors Offer Useful Accuracy-Carbon Trade-Offs in Spatio-Temporal Forecasting

ResearchDGX agent

arXiv:2509.24517v2 Announce Type: replace Abstract: Development of modern deep learning methods has been driven primarily by the push for improving model efficacy (accuracy metrics). This sole focus o

Plug-in Losses for Evidential Deep Learning: A Simplified Framework for Uncertainty Estimation that Includes the Softmax Classifier

ApplicationsDGX agent

arXiv:2605.22746v1 Announce Type: new Abstract: Real-world sensor-based learning systems require uncertainty estimation that is both reliable and computationally efficient. Evidential Deep Learning (E

Position: The Time for Sampling Is Now! Charting a New Course for Bayesian Deep Learning

TutorialsDGX agent

arXiv:2605.21765v1 Announce Type: new Abstract: The practical adoption of sampling-based inference (SAI) in Bayesian neural networks (BNNs) remains limited, partly due to persistent misconceptions abo

Post-Training is About States, Not Tokens: A State Distribution View of SFT, RL, and On-Policy Distillation

SafetyDGX agent

arXiv:2605.22731v1 Announce Type: new Abstract: Large language model post-training methods such as supervised fine-tuning (SFT), reinforcement learning (RL), and distillation are often analyzed throug

Posterior Collapse as Automatic Spectral Pruning

Model ReleasesDGX agent

arXiv:2605.22691v1 Announce Type: new Abstract: We show that posterior collapse in eta-VAEs implements automatic spectral pruning. A latent mode collapses if its contribution to reconstruction is belo

Predicting Performance of Symbolic and Prompt Programs with Examples

ResearchDGX agent

arXiv:2605.21515v1 Announce Type: new Abstract: LLM prompting is widely used for naturally stated tasks, yet it is unreliable it may succeed on a few test cases but fail at deployment time. We study p

Prior Knowledge-enhanced Spatio-temporal Epidemic Forecasting

Model ReleasesDGX agent

arXiv:2602.22270v2 Announce Type: replace Abstract: Spatio-temporal epidemic forecasting is critical for public health management, yet existing methods often struggle with insensitivity to weak epidem

Prior shift estimation for positive unlabeled data through the lens of kernel embedding

ResearchDGX agent

arXiv:2502.21194v3 Announce Type: replace-cross Abstract: We study estimation of a class prior for unlabeled target samples which possibly differs from that of source population. Moreover, it is assum

Protein Thoughts: Interpretable Reasoning with Tree of Thoughts and Embedding-Space Flow Matching for Protein-Protein Interaction Discovery

Model ReleasesDGX agent

arXiv:2605.21522v1 Announce Type: cross Abstract: Protein-protein interactions (PPIs) govern nearly all cellular processes, yet computational methods for identifying binding partners typically produce

Prototype-Guided Classification Sub-Task Decoupling Framework: Enhancing Generalization and Interpretability for Multivariate Time Series

ResearchDGX agent

arXiv:2605.22055v1 Announce Type: new Abstract: Time Series Classification (TSC) is a long-standing research problem that has gained increasing attention in recent years with the rapid growth of large

Provable Joint Decontamination for Benchmarking Multiple Large Language Models

Model ReleasesDGX agent

arXiv:2605.21543v1 Announce Type: new Abstract: Benchmark data contamination has become a central challenge in LLM evaluation: when evaluation examples appear in the training data of one or more audit

Provable Robustness against Backdoor Attacks via the Primal-Dual Perspective on Differential Privacy

ApplicationsDGX agent

arXiv:2605.21780v1 Announce Type: new Abstract: Randomized smoothing is a powerful tool for certifying robustness to adversarial perturbations, including poisoning attacks via randomized training and

Provably Protecting Fine-Tuned LLMs from Training Data Extraction while Preserving Utility

ResearchDGX agent

arXiv:2602.00688v2 Announce Type: replace Abstract: Fine-tuning large language models (LLMs) on sensitive datasets raises privacy concerns, as training data extraction (TDE) attacks can expose highly

Proxy-Based Approximation of Shapley and Banzhaf Interactions

SafetyDGX agent

arXiv:2605.22738v1 Announce Type: new Abstract: Shapley and Banzhaf interactions capture the complex dynamics inherent in modern machine learning applications. However, current estimators for these hi

Q-PhotoNAS: Hybrid Quantum Neural Architecture Search Framework on Photonic Devices

ResearchDGX agent

arXiv:2605.22097v1 Announce Type: cross Abstract: Photonic quantum computing is a promising platform for scalable quantum machine learning, but designing effective hybrid architectures remains challen

Quantitative coronary calcification analysis for prediction of myocardial ischemia using non-contrast CT calcium scoring

ResearchDGX agent

arXiv:2605.21745v1 Announce Type: new Abstract: Non-contrast computed tomography calcium scoring (CTCS) is widely recognized as an effective tool for cardiovascular risk stratification. This study aim

RADAR: Defending RAG Dynamically against Retrieval Corruption

ResearchDGX agent

arXiv:2605.22041v1 Announce Type: cross Abstract: While RAG systems are increasingly deployed in dynamic web search, temporal volatility amplifies their vulnerability to adversarial attacks. Existing

[Re] FairDICE: A Fair Tradeoff in Multi-objective Offline RL

SafetyDGX agent

arXiv:2603.03454v2 Announce Type: replace Abstract: Offline Reinforcement Learning (RL) is an emerging field of RL in which policies are learned solely from demonstrations. Within offline RL, some env

Reading Task Failure Off the Activations: A Sparse-Feature Audit of GPT-2 Small on Indirect Object Identification

HardwareDGX agent

arXiv:2605.22719v1 Announce Type: new Abstract: We report a small, reproducible audit of which sparse-autoencoder (SAE) features of GPT-2 small fire differently on failed versus successful trials of t

Reasoning through Verifiable Forecast Actions: Consistency-Grounded RL for Financial LLMs

Model ReleasesDGX agent

arXiv:2605.21975v1 Announce Type: new Abstract: Financial markets are characterized by extreme non-stationarity, low signal-to-noise ratios, and strong dependence on external information such as news,

Regret-Based (epsilon,elta)-optimal Stopping Criteria for Bayesian Optimization

ResearchDGX agent

arXiv:2605.22561v1 Announce Type: new Abstract: Bayesian optimization (BO) is a widely used iterative black-box optimization method that utilizes Gaussian process (GP) surrogate models. In practice, B

Reinforced Graph of Thoughts: RL-Driven Adaptive Prompting for LLMs

ResearchDGX agent

arXiv:2605.22195v1 Announce Type: new Abstract: Graph of Thoughts (GoT), a generalized form of recent prompting paradigms for large language models (LLMs), has been shown to be useful for elaborate pr

Reinforcement learning for ion shuttling on trapped-ion quantum computers

ResearchDGX agent

arXiv:2605.22463v1 Announce Type: cross Abstract: Scalable trapped-ion quantum computing is commonly realized with modular chips that feature distinct zones with specific functionalities, such as stor

Relational Linear Properties in Language Models: An Empirical Investigation

ResearchDGX agent

arXiv:2605.22532v1 Announce Type: new Abstract: Linear properties are ubiquitous in the representations of language models; however, testing them experimentally remains a challenging task. This work f

Reliable Wireless Indoor Localization via Cross-Validated Prediction-Powered Calibration

Local AiDGX agent

arXiv:2507.20268v3 Announce Type: replace Abstract: Wireless indoor localization using predictive models with received signal strength information (RSSI) requires proper calibration for reliable posit

Remember to be Curious: Episodic Context and Persistent Worlds for 3D Exploration

Local AiDGX agent

arXiv:2605.22814v1 Announce Type: new Abstract: Exploration is a prerequisite for learning useful behaviors in sparse-reward, long-horizon tasks, particularly within 3D environments. Curiosity-driven

Representation Gap: Explaining the Unreasonable Effectiveness of Neural Networks from a Geometric Perspective

Model ReleasesDGX agent

arXiv:2605.21692v1 Announce Type: new Abstract: Characterizing precisely the asymptotic generalization error of neural networks using parameters that can be estimated efficiently is a crucial problem

Represented Is Not Computed: A Causal Test of Candidate Algorithmic Intermediates in a Transformer

Local AiDGX agent

arXiv:2605.22488v1 Announce Type: new Abstract: Structured prompts require integrating components according to task-relevant relations. How a network implements this integration is often hard to judge

Rethinking Forward Processes for Score-Based Nonlinear Data Assimilation in High Dimensions

Model ReleasesDGX agent

arXiv:2604.02889v2 Announce Type: replace-cross Abstract: Data assimilation is the process of estimating the state of a dynamical system over time by combining model predictions with measurements. Thi

Revisiting Regularized Policy Optimization for Stable and Efficient Reinforcement Learning in Two-Player Games

SafetyDGX agent

arXiv:2602.10894v2 Announce Type: replace Abstract: Two-player games such as board games have long been used as traditional benchmarks for reinforcement learning. This work revisits a policy optimizat

Revisiting Robustness for LLM Safety Alignment via Selective Geometry Control

Model ReleasesDGX agent

arXiv:2602.07340v2 Announce Type: replace Abstract: Safety alignment of large language models remains brittle under domain shift and noisy preference supervision. Most existing robust alignment method

Richer Bayesian Last Layers with Subsampled NTK Features

ResearchDGX agent

arXiv:2602.01279v2 Announce Type: replace Abstract: Bayesian Last Layers (BLLs) provide a convenient and computationally efficient way to estimate uncertainty in neural networks. However, they underes

Riemannian geometry meets fMRI: the advantages of modeling correlation manifolds and eigenvector subspaces

ApplicationsDGX agent

arXiv:2605.22334v1 Announce Type: new Abstract: Correlation matrices are fundamental summaries of functional brain networks, yet standard analyses often treat entries independently, ignoring the curve

RobustSpeechFlow: Learning Robust Text-to-Speech Trajectories via Augmentation-based Contrastive Flow Matching

Model ReleasesDGX agent

arXiv:2605.22083v1 Announce Type: cross Abstract: While flow-matching text-to-speech (TTS) achieves strong zero-shot speaker similarity and naturalness, it remains susceptible to content fidelity issu

Rule-State Inference (RSI): A Bayesian Framework for Compliance Monitoring in Rule-Governed Domains

Model ReleasesDGX agent

arXiv:2603.21610v2 Announce Type: replace Abstract: Compliance monitoring in rule-governed domains (tax administration, clinical protocol adherence, environmental regulation) faces three structural ob

Same Architecture, Different Capacity: Optimizer-Induced Spectral Scaling Laws

ResearchDGX agent

arXiv:2605.21803v1 Announce Type: new Abstract: Scaling laws have made language-model performance predictable from model size, data, and compute, but they typically treat the optimizer as a fixed trai

Scalable On-Policy Reinforcement Learning via Adaptive Batch Scaling

Model ReleasesDGX agent

arXiv:2605.21557v1 Announce Type: cross Abstract: Conventional wisdom holds that large-batch training is fundamentally incompatible with Reinforcement Learning (RL) - beyond a modest threshold, increa

SCI-Defense: Defending Manipulation Attacks from Generative Engine Optimization

ResearchDGX agent

arXiv:2605.21948v1 Announce Type: new Abstract: LLM-based ranking systems are vulnerable to Generative Engine Optimization (GEO) attacks, where adversaries inject semantic signals into product descrip

SDPM: Survival Diffusion Probabilistic Model for Continuous-Time Survival Analysis

ResearchDGX agent

arXiv:2605.22776v1 Announce Type: new Abstract: Survival analysis aims to estimate a time-to-event distribution from data with censored observations. Many existing methods either impose structural ass

Self-orthogonalizing attractor neural networks emerging from the free energy principle

ResearchDGX agent

arXiv:2505.22749v2 Announce Type: replace-cross Abstract: Attractor dynamics are a hallmark of many complex systems, including the brain. Understanding how such self-organizing dynamics emerge from fi

Self-Supervised ConvLSTM for Fermi Large Area Telescope Transient Detection

Model ReleasesDGX agent

arXiv:2605.22112v1 Announce Type: cross Abstract: We present a framework for detecting transient gamma-ray phenomena in a controlled environment by combining end-to-end simulations of the Fermi-LAT sk

SepsisAI Orchestrator: A Containerized and Scalable Platform for Deploying AI Models and Real-Time Monitoring in Early Sepsis Detection

ResearchDGX agent

arXiv:2605.22331v1 Announce Type: new Abstract: Despite strong predictive results in the clinical machine learning literature, the translation of these models into bedside use remains limited by syste

SeqLoRA: Bilevel Orthogonal Adaptation for Continual Multi-Concept Generation

Model ReleasesDGX agent

arXiv:2605.22743v1 Announce Type: new Abstract: Parameter-efficient fine-tuning enables fast personalization of text-to-image diffusion models, but composing multiple custom concepts remains challengi

Short-Term-to-Long-Term Memory Transfer for Knowledge Graphs under Partial Observability

Model ReleasesDGX agent

arXiv:2605.22142v1 Announce Type: new Abstract: Reinforcement learning under partial observability requires deciding what information to retain, yet most memory-based approaches do not explicitly mode

Skill Weaving: Efficient LLM Improvement via Modular Skillpacks

AgentsDGX agent

arXiv:2605.22205v1 Announce Type: cross Abstract: Large language models increasingly require specialization across diverse domains, yet existing approaches struggle to balance multi-domain capacities

Soft Bayesian Context Tree Models for Real-Valued Time Series

ResearchDGX agent

arXiv:2601.11079v2 Announce Type: replace Abstract: This paper proposes the soft Bayesian context tree model (Soft-BCT), which is a novel BCT model for real-valued time series. The Soft-BCT considers

Sparse Orthogonal Parameters Tuning for Continual Learning

ResearchDGX agent

arXiv:2411.02813v3 Announce Type: replace Abstract: Continual learning methods based on pre-trained models (PTM) have recently gained attention which adapt to successive downstream tasks without catas

Spectra as Language: Large Language Models for Scalable Stellar Parameter and Abundance Inference

Model ReleasesDGX agent

arXiv:2605.22162v1 Announce Type: cross Abstract: Stellar spectra encode key information on the physical properties and chemical compositions of stars. Accurate stellar parameter determination is esse

← Previous
1…135136137138139…243
Next →