AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,557 results
Model Releases

Quantum reservoir computing in Jaynes-Cummings models: Nonlinear memory and time-series prediction

DGX agent

arXiv:2510.00171v2 Announce Type: replace-cross Abstract: We investigate quantum reservoir computing (QRC) using a hybrid qubit-boson system described by the Jaynes-Cummings (JC) Hamiltonian and its d

model-releasesarxiv-cs-lg
21 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

Reasoning-Trace Collapse: Evaluating the Loss of Explicit Reasoning During Fine-Tuning

DGX agent

arXiv:2605.21127v1 Announce Type: new Abstract: Explicit reasoning models are trained to produce intermediate reasoning traces before final answers, but downstream fine-tuning is often performed on or

researcharxiv-cs-lg
21 May 2026
Safety

REFLECTOR: Internalizing Step-wise Reflection against Indirect Jailbreak

DGX agent

arXiv:2605.20654v1 Announce Type: new Abstract: While Large Language Models (LLMs) demonstrate remarkable capabilities, they remain susceptible to sophisticated, multi-step jailbreak attacks that circ

safetyarxiv-cs-lg
21 May 2026
Research

Reinforcement Learning-based Control via Y-wise Affine Neural Networks: Comparative Case Studies for Chemical Processes

DGX agent

arXiv:2605.21211v1 Announce Type: cross Abstract: In this work we present an efficient and practically implementable approach for the application of reinforcement learning (RL)-based control in chemic

researcharxiv-cs-lg
21 May 2026
Safety

Reinforcement Learning with Discrete Diffusion Policies for Combinatorial Action Spaces

DGX agent

arXiv:2509.22963v3 Announce Type: replace Abstract: Reinforcement learning (RL) struggles to scale to large, combinatorial action spaces common in many real-world problems. This paper introduces a nov

safetyarxiv-cs-lg
21 May 2026
Safety

rePIRL: Learn PRM with Inverse RL for LLM Reasoning

DGX agent

arXiv:2602.07832v2 Announce Type: replace Abstract: Process rewards have been widely used in deep reinforcement learning to improve training efficiency, reduce variance, and prevent reward hacking. In

safetyarxiv-cs-lg
21 May 2026
Model Releases

Residual Paving: Diagnosing the Routing Bottleneck in Selective Refusal Editing

DGX agent

arXiv:2605.20262v1 Announce Type: new Abstract: We study selective refusal editing as a three-way control problem: induce non-refusal on designated edit prompts while preserving benign behavior and ha

model-releasesarxiv-cs-lg
21 May 2026
Research

ReversedQ: Opportunities for Faster Q-Learning in Episodic Online Reinforcement Learning

DGX agent

arXiv:2605.20592v1 Announce Type: new Abstract: We study model-free Q-learning in finite-horizon episodic Markov Decision Processes (MDPs) with stationary dynamics across episodes. We identify a centr

researcharxiv-cs-lg
21 May 2026
Research

Reviving Error Correction in Modern Deep Time-Series Forecasting

DGX agent

arXiv:2605.21088v1 Announce Type: new Abstract: Modern deep-learning models have achieved remarkable success in time-series forecasting. Yet, their performance degrades in long-term prediction due to

researcharxiv-cs-lg
21 May 2026
Research

Riemannian MeanFlow for One-Step Generation on Manifolds

DGX agent

arXiv:2603.10718v2 Announce Type: replace Abstract: Flow Matching enables simulation-free training of generative models on Riemannian manifolds, yet sampling typically still relies on numerically inte

researcharxiv-cs-lg
21 May 2026
Model Releases

Robust Personalized Recommendation under Hidden Confounding in MNAR

DGX agent

arXiv:2605.21066v1 Announce Type: new Abstract: Recommender systems often rely on observational user--item interaction data, which is prone to selection bias due to users' selective interactions with

model-releasesarxiv-cs-lg
21 May 2026
Safety

Robust Recommendation from Noisy Implicit Feedback: A GMM-Weighted Bayes-label Transition Matrix Framework

DGX agent

arXiv:2605.20721v1 Announce Type: new Abstract: Learning from implicit feedback in recommender systems is fundamentally challenged by pervasive label noise. While conventional denoising approaches oft

safetyarxiv-cs-lg
21 May 2026
Research

Robust Subspace-Constrained Quadratic Models for Low-Dimensional Structure Learning

DGX agent

arXiv:2605.20300v1 Announce Type: new Abstract: In this paper, we propose a robust subspace-constrained quadratic model (SCQM) for learning low-dimensional structure from high-dimensional data. Buildi

researcharxiv-cs-lg
21 May 2026
Model Releases

roto 2.0: The Robot Tactile Olympiad

DGX agent

arXiv:2605.21429v1 Announce Type: cross Abstract: Tactile-based reinforcement learning (RL) is currently hindered by fragmented research and a focus on over-saturated orientation tasks. We introduce v

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

Runtime-Certified Bounded-Error Quantized Attention

DGX agent

arXiv:2605.20868v1 Announce Type: new Abstract: KV cache quantization reduces the memory cost of long-context LLM inference, but introduces approximation error that is typically validated only empiric

model-releasesarxiv-cs-lg
21 May 2026
Research

Same Target, Different Basins: Hard vs. Soft Labels for Annotator Distributions

DGX agent

arXiv:2605.20642v1 Announce Type: new Abstract: When annotators disagree, that disagreement can reflect epistemic uncertainty rather than simple label noise. We study hard-label delivery as an alterna

researcharxiv-cs-lg
21 May 2026
Research

Sample Complexity of Transfer Learning: An Optimal Transport Approach

DGX agent

arXiv:2605.20545v1 Announce Type: cross Abstract: Transfer learning is an essential technique for many machine learning/AI models of complex structures such as large language models and generative AI.

researcharxiv-cs-lg
21 May 2026
Local Ai

Scale-Calibrated Median-of-Means for Robust Distributed Principal Component Analysis

DGX agent

arXiv:2605.20681v1 Announce Type: cross Abstract: Distributed principal component analysis (PCA) produces node-level estimates of both a mean vector and a principal subspace. Robustly aggregating thes

local-aiarxiv-cs-lg
21 May 2026
Research

Score-Based Causal Discovery of Latent Variable Causal Models

DGX agent

arXiv:2605.20396v1 Announce Type: new Abstract: Identifying latent variables and the causal structure involving them is essential across various scientific fields. While many existing works fall under

researcharxiv-cs-lg
21 May 2026
Safety

Secure, Verifiable, and Scalable Multi-Client Data Sharing via Consensus-Based Privacy-Preserving Data Distribution

DGX agent

arXiv:2601.00418v2 Announce Type: replace-cross Abstract: We propose the Consensus-Based Privacy-Preserving Data Distribution (CPPDD) framework, a lightweight and post-setup autonomous protocol for se

safetyarxiv-cs-lg
21 May 2026
Research

Self-Improving Skill Learning for Robust Skill-based Meta-Reinforcement Learning

DGX agent

arXiv:2502.03752v5 Announce Type: replace Abstract: Meta-reinforcement learning (Meta-RL) facilitates rapid adaptation to unseen tasks but faces challenges in long-horizon environments. Skill-based ap

researcharxiv-cs-lg
21 May 2026
Model Releases

Semiparametric Efficient Bilevel Gradient Estimation

DGX agent

arXiv:2605.21341v1 Announce Type: cross Abstract: Functional bilevel methods estimate a lower-level function and plug it into a hypergradient, but this plug-in gradient can retain first-order bias whe

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

Sequential Data Augmentation for Generative Recommendation

DGX agent

arXiv:2509.13648v3 Announce Type: replace Abstract: Generative recommendation plays a crucial role in personalized systems, predicting users' future interactions from their historical behavior sequenc

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

ShapeBench: A Scalable Benchmark and Diagnostic Suite for Standardized Evaluation in Aerodynamic Shape Optimization

DGX agent

arXiv:2605.20763v1 Announce Type: new Abstract: Rapid progress in aerodynamic shape optimization (ASO) has outpaced currently-available standardized evaluation frameworks. Fair comparison requires a u

model-releasesarxiv-cs-lg
21 May 2026
Safety

SMA-DP: Spectral Memory-Aware Differential Privacy for Deep Learning

DGX agent

arXiv:2605.20450v1 Announce Type: new Abstract: Differentially private stochastic gradient descent (DP-SGD) enables private deep learning through per-example clipping and calibrated Gaussian noise, bu

safetyarxiv-cs-lg
21 May 2026
Agents

Smaller Abstract State Spaces Enable Cross-Scale Generalization in Reinforcement Learning

DGX agent

arXiv:2605.20272v1 Announce Type: new Abstract: While humans readily generalize abstract concepts to more complex or larger tasks, building Reinforcement Learning (RL) systems with this ability remain

agentsarxiv-cs-lg
21 May 2026
Model Releases

SOLAR: A Self-Optimizing Open-Ended Autonomous Agent for Lifelong Learning and Continual Adaptation

DGX agent

arXiv:2605.20189v1 Announce Type: cross Abstract: Despite the remarkable success of large language models (LLMs), they still face bottlenecks while deploying in dynamic, real-world settings with prima

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

SOPE: Stabilizing Off-Policy Evaluation for Online RL with Prior Data

DGX agent

arXiv:2605.05863v2 Announce Type: replace Abstract: Incorporating prior data into online reinforcement learning accelerates training but typically forces a difficult trade-off between high computation

model-releasesarxiv-cs-lg
21 May 2026
Applications

Spectral bandits for smooth graph functions with applications in recommender systems

DGX agent

arXiv:2605.20552v1 Announce Type: cross Abstract: Smooth functions on graphs have wide applications in manifold and semi-supervised learning. In this paper, we study a bandit problem where the payoffs

applicationsarxiv-cs-lg
21 May 2026
Safety

Spectral Souping: A Unified Framework for Online Preference Alignment

DGX agent

arXiv:2605.20408v1 Announce Type: new Abstract: Reinforcement Learning from Human Feedback (RLHF) effectively aligns Large Language Models (LLMs) with aggregate human preferences but often fails to ad

safetyarxiv-cs-lg
21 May 2026
Model Releases

Spectral Unforgetting: Post-Hoc Recovery of Damaged Capabilities Without Retraining

DGX agent

arXiv:2605.20296v1 Announce Type: new Abstract: Fine-tuning a language model for a target task routinely degrades capabilities the training data never explicitly threatened. We study this phenomenon,

model-releasesarxiv-cs-lg
21 May 2026
Safety

Statistical Guarantees in the Search for Less Discriminatory Algorithms

DGX agent

arXiv:2512.23943v2 Announce Type: replace-cross Abstract: U.S. discrimination law can impose liability on firms that fail to adopt a less discriminatory alternative (LDA): a decision policy that achie

safetyarxiv-cs-lg
21 May 2026
Research

Stimulus symmetries can confound representational similarity analyses

DGX agent

arXiv:2605.21324v1 Announce Type: cross Abstract: What can representational similarity matrices (RSMs) tell us about a neural code? As the popularity of these summary statistics grows, so too does the

researcharxiv-cs-lg
21 May 2026
Applications

STM3: Mixture of Multiscale Mamba for Long-Term Spatio-Temporal Time-Series Prediction

DGX agent

arXiv:2508.12247v2 Announce Type: replace Abstract: Recently, spatio-temporal time-series prediction has developed rapidly, yet existing deep learning methods struggle with learning complex long-term

applicationsarxiv-cs-lg
21 May 2026
Safety

Strict Subgoal Execution: Reliable Long-Horizon Planning in Hierarchical Reinforcement Learning

DGX agent

arXiv:2506.21039v3 Announce Type: replace Abstract: Long-horizon goal-conditioned tasks pose fundamental challenges for reinforcement learning (RL), particularly when goals are distant and rewards are

safetyarxiv-cs-lg
21 May 2026
Safety

Supervised Latent Restructuring for Small-Data Quantum Learning in Plant Phenomics

DGX agent

arXiv:2605.20413v1 Announce Type: new Abstract: High-dimensional biological data often exhibit a severe mismatch between feature dimensionality and sample size, making reliable classification difficul

safetyarxiv-cs-lg
21 May 2026
Safety

SURF: Steering the Scalarization Weight to Uniformly Traverse the Pareto Front

DGX agent

arXiv:2605.20619v1 Announce Type: new Abstract: Scalarization is widely used in multi-objective optimization owing to its simplicity and scalability. In many applications, the goal is to generate solu

safetyarxiv-cs-lg
21 May 2026
Local Ai

Sustainability Is Not Linear: Quantifying Performance, Energy, and Privacy Trade-offs in On-Device Intelligence

DGX agent

arXiv:2603.26603v2 Announce Type: replace-cross Abstract: The migration of Large Language Models (LLMs) from cloud clusters to edge devices promises enhanced privacy and offline accessibility, but thi

local-aiarxiv-cs-lg
21 May 2026
Research

Sutra: Tensor-Op RNNs as a Compilation Target for Vector Symbolic Architectures

DGX agent

arXiv:2605.20919v1 Announce Type: new Abstract: Sutra is a typed, purely functional programming language whose compiled forward pass is a PyTorch neural network. The compiler beta-reduces the whole pr

researcharxiv-cs-lg
21 May 2026
Research

Symmetrization of Loss Functions for Robust Training of Neural Networks in the Presence of Noisy Labels

DGX agent

arXiv:2605.20347v1 Announce Type: new Abstract: Labeling a training set is often expensive and susceptible to errors, making the design of robust loss functions for label noise an important problem. T

researcharxiv-cs-lg
21 May 2026
Model Releases

TabPFN Extensions for Interpretable Geotechnical Modelling

DGX agent

arXiv:2603.21033v2 Announce Type: replace-cross Abstract: Geotechnical site characterisation relies on sparse, heterogeneous borehole data, where uncertainty quantification and interpretability matter

model-releasesarxiv-cs-lg
21 May 2026
Research

TabPFN-MT: A Natively Multitask In-Context Learner for Tabular Data

DGX agent

arXiv:2605.20234v1 Announce Type: new Abstract: Prior-Data Fitted networks (PFNs) have been very successful in tabular contexts, handling prediction tasks in context. However, they are designed for si

researcharxiv-cs-lg
21 May 2026
Applications

TelecomTS: A Multi-Modal Observability Dataset for Time Series and Language Analysis

DGX agent

arXiv:2510.06063v2 Announce Type: replace-cross Abstract: Modern enterprises generate vast streams of time series metrics when monitoring complex systems, known as observability data. Unlike conventio

applicationsarxiv-cs-lg
21 May 2026
Research

Testing Support Size More Efficiently Than Learning Histograms

DGX agent

arXiv:2410.18915v4 Announce Type: replace-cross Abstract: Consider two problems about an unknown probability distribution p: 1. How many samples from p are required to test if p is supported on n elem

researcharxiv-cs-lg
21 May 2026
Research

The Devil is in the Condition Numbers: Why is GLU Better than non-GLU Structure?

DGX agent

arXiv:2605.20749v1 Announce Type: new Abstract: Gated Linear Units (GLU) and their variants are widely adopted in modern open-source large language model architectures and consistently outperform thei

researcharxiv-cs-lg
21 May 2026
Safety

The Economics of AI Inference: Inflation Dynamics, Welfare Costs, and Optimal Monetary Policy under the Inference-Cost Phillips Curve

DGX agent

arXiv:2605.20281v1 Announce Type: cross Abstract: We develop a unified microeconomic and monetary theory of artificial intelligence inference costs and their pass-through to inflation, welfare, and op

safetyarxiv-cs-lg
21 May 2026
Model Releases

The Economics of Model Collapse: Equilibrium, Welfare, and Optimal Provenance Subsidies in Synthetic Data Markets

DGX agent

arXiv:2605.20279v1 Announce Type: cross Abstract: Generative artificial intelligence is rapidly transforming the supply side of training data: an increasing share of new tokens, images, and structured

model-releasesarxiv-cs-lg
21 May 2026
Local Ai

The General Theory of Localization Methods

DGX agent

arXiv:2605.20635v1 Announce Type: new Abstract: This paper proposes a general machine learning framework called the localization method, which is fundamentally built on two core concepts: localization

local-aiarxiv-cs-lg
21 May 2026
← Previous
1…176177178179180…304
Next →