AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlog
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,329 results
Agents

Regret, equilibrium, and learning in games: A guided tour

DGX agent

arXiv:2608.09389v1 Announce Type: cross Abstract: This note aims to serve as an entry point to the literature on learning in games, a topic with significant theoretical appeal and a wide range of appl

agentsarxiv-cs-lg
11 Aug 2026
Safety

Regret of exploratory policy improvement and q-learning

X Post
Paper
YouTube
Reddit
GitHub
Clear filters
DGX agent

arXiv:2411.01302v2 Announce Type: replace Abstract: We study the convergence of q-learning and related algorithms introduced by Jia and Zhou (J. Mach. Learn. Res., 24 (2023), 161) for controlled diffu

safetyarxiv-cs-lg
11 Aug 2026
Model Releases

ReliableNet: A Chance-Constrained Approach to Trustworthy Classification in Deep Learning

DGX agent

arXiv:2608.09768v1 Announce Type: new Abstract: A prediction that is both confident and wrong is a critical reliability failure because it can bypass abstention and human review precisely when the mod

model-releasesarxiv-cs-lg
11 Aug 2026
Research

Rethinking Learning-Based Influence Maximization: Simple Neural Surrogates and Native Discrete Search

DGX agent

arXiv:2608.08406v1 Announce Type: new Abstract: Existing learning-based influence maximization frameworks rely heavily on complex neural architectures and continuous optimization over seed representat

researcharxiv-cs-lg
11 Aug 2026
Research

Rethinking Reasoning with MDLMs: Early Exits, Post-hoc Reasoning, and Beyond

DGX agent

arXiv:2510.19990v2 Announce Type: replace Abstract: The reasoning paradigm, where language models reason before answering, has enabled breakthroughs on tasks such as mathematical problem-solving. Whil

researcharxiv-cs-lg
11 Aug 2026
Tutorials

Robust Reputation-Driven Crowdsourced Federated Learning

DGX agent

arXiv:2608.08574v1 Announce Type: new Abstract: Crowdsourced Federated Learning (CrowdFL) extends traditional federated learning by enabling open and heterogeneous participation through a crowdsourcin

tutorialsarxiv-cs-lg
11 Aug 2026
Model Releases

RotaryQuant: Fitting 120B MoE Models on Consumer Hardware via Fused Compressed-Space Attention

DGX agent

arXiv:2608.08081v1 Announce Type: cross Abstract: Large mixture-of-experts (MoE) language models with 26--120 billion parameters exceed the memory capacity of consumer devices through three simultaneo

model-releasesarxiv-cs-lg
11 Aug 2026
Model Releases

RouteGuard: Certifying Routing Gain in LLM Multi-Agent Systems When Complementarity Is Not Enough

DGX agent

arXiv:2608.07583v1 Announce Type: cross Abstract: Multi-agent LLM systems route among model-backed advisors, yet a deployer rarely knows before shipping whether routing will help at all. Prevailing ro

model-releasesarxiv-cs-lg
11 Aug 2026
Safety

SAFE-CHEM: Uncertainty-Aware Policy Switching for Robust Robotic Chemistry

DGX agent

arXiv:2608.09303v1 Announce Type: cross Abstract: The deployment of autonomous robotic systems in chemistry laboratories is accelerating experimental workflows and providing the foundational data for

safetyarxiv-cs-lg
11 Aug 2026
Safety

Satellite Trajectory Optimization via Proximal Policy Optimization for Space Debris Avoidance

DGX agent

arXiv:2608.09628v1 Announce Type: new Abstract: Collision avoidance systems are commonly used to avoid fragmentation events occurring in Low-Earth Orbit (LEO) and Geosynchronous Equatorial Orbit (GEO)

safetyarxiv-cs-lg
11 Aug 2026
Safety

Scalable extensions to given-data Sobol' index estimators

DGX agent

arXiv:2509.09078v3 Announce Type: replace-cross Abstract: Given-data methods for variance-based sensitivity analysis have significantly advanced the feasibility of Sobol' index computation for computa

safetyarxiv-cs-lg
11 Aug 2026
Research

SoftMCC: An MCC-Brier Calibration Bridge for Threshold-Free Model Selection under Class Imbalance

DGX agent

arXiv:2608.08984v1 Announce Type: new Abstract: Model selection for imbalanced binary classification often uses the Matthews correlation coefficient (MCC), but thresholding makes validation rankings t

researcharxiv-cs-lg
11 Aug 2026
Model Releases

Sparse corruption in low-rank matrix inference: the PCA benchmark

DGX agent

arXiv:2511.11927v2 Announce Type: replace-cross Abstract: Principal Component Analysis (PCA) is a standard tool for extracting a low-rank signal from noisy observations. It is known that applying PCA

model-releasesarxiv-cs-lg
11 Aug 2026
Research

Spatial Heterogeneity-Aware Multi-Hazard Susceptibility and Risk Mapping at Regional Scale

DGX agent

arXiv:2608.08321v1 Announce Type: new Abstract: Floods and landslides often co-occur, but their relationships with environmental controls vary spatially. This study develops a spatial heterogeneity-aw

researcharxiv-cs-lg
11 Aug 2026
Tutorials

SPD Learn: A Geometric Deep Learning Python Library for Neural Decoding Through Trivialization

DGX agent

arXiv:2602.22895v2 Announce Type: replace-cross Abstract: Implementations of symmetric positive definite (SPD) matrix-based neural networks for neural decoding remain fragmented across research codeba

tutorialsarxiv-cs-lg
11 Aug 2026
Model Releases

SPECTRA: Pushing the KV Cache Beyond the 2-Bit Cliff via Spectral Transform Coding

DGX agent

arXiv:2608.07915v1 Announce Type: new Abstract: Large language models (LLMs) increasingly read long inputs in the agentic era, from whole documents and codebases to conversations across many turns. Th

model-releasesarxiv-cs-lg
11 Aug 2026
Model Releases

Stateful CARS: Exact Cross-History Reuse for Policy-Constrained LLM Agents

DGX agent

arXiv:2608.08282v1 Announce Type: new Abstract: Tool-using language-model agents face constraints whose meaning changes with observations and prior actions. We study exact sampling from the model dist

model-releasesarxiv-cs-lg
11 Aug 2026
Research

Stochastic gradient descent with discontinuity across a manifold

DGX agent

arXiv:2608.07618v1 Announce Type: cross Abstract: Stochastic gradient descent for a loss function discontinuous across lower dimensional manifolds is analyzed by studying its differential equation lim

researcharxiv-cs-lg
11 Aug 2026
Research

Support Selection Beyond Smooth DAG Exactness: Completion Geometry,Score Margins, and Selective Certificates

DGX agent

arXiv:2608.08103v1 Announce Type: new Abstract: Smooth acyclicity constraints answer whether a weighted support is a DAG, whereas structure learning asks which support change should be made. Existing

researcharxiv-cs-lg
11 Aug 2026
Hardware

SwiftQK: Fast and Communication-Efficient Tensor Parallelism for Query-Key Normalization

DGX agent

arXiv:2608.09160v1 Announce Type: new Abstract: Query-Key Normalization (QK-Norm) improves the training stability and quality of modern Large Language Models (LLMs). However, under Tensor Parallelism

hardwarearxiv-cs-lg
11 Aug 2026
Local Ai

Targeted Label-Flipping and Oversampling Attacks on Federated Conditional GANs

DGX agent

arXiv:2608.09314v1 Announce Type: new Abstract: In a federated learning setup for GANs, several adversarial attacks are possible. One such attack is label flipping, in which malicious clients delibera

local-aiarxiv-cs-lg
11 Aug 2026
Model Releases

Task-to-Model Optimization for Enterprise LLM Coding Assistants: A Data-Driven Framework for Cost-Optimal Routing

DGX agent

arXiv:2608.08528v1 Announce Type: new Abstract: Enterprise AI coding assistants incur substantial inference spend, and naive token-cost minimization often fails to reduce end-to-end cost once retries,

model-releasesarxiv-cs-lg
11 Aug 2026
Model Releases

Test-time Generalization for Physics through Neural Operator Splitting

DGX agent

arXiv:2602.00884v2 Announce Type: replace Abstract: Neural operators have shown promise in learning solution maps of partial differential equations (PDEs), but they often struggle to generalize when t

model-releasesarxiv-cs-lg
11 Aug 2026
Research

Test-Time Scaling for CAD Generation via Verifier-Free Consensus Selection

DGX agent

arXiv:2608.09706v1 Announce Type: cross Abstract: Large language models can write parametric CAD programs from a natural-language description (text-to-CAD generation), but a single sample is often wro

researcharxiv-cs-lg
11 Aug 2026
Model Releases

The Cost of Adaptivity: Matching Lower Bounds Across Learning Problems

DGX agent

arXiv:2608.08826v1 Announce Type: new Abstract: Adaptive procedures must work without nuisance information an oracle may use, such as a gradient scale or smoothness index, and robust procedures may ha

model-releasesarxiv-cs-lg
11 Aug 2026
Research

The Neural Division of Labor: Biologically-Inspired Modular Architectures for Robust Neuromorphic Computing

DGX agent

arXiv:2608.08317v1 Announce Type: new Abstract: Biological neural systems achieve high efficiency and robustness through compartmentalized architectures. In contrast, modern artificial neural networks

researcharxiv-cs-lg
11 Aug 2026
Safety

The Sample Complexity of Policy Learning with Mu-Resets

DGX agent

arXiv:2608.07772v1 Announce Type: new Abstract: We study policy-based reinforcement learning under the mu-resets interaction protocol of Kakade and Langford [KL02]. This interaction protocol enables t

safetyarxiv-cs-lg
11 Aug 2026
Local Ai

The Spectral Neuron

DGX agent

arXiv:2608.08003v1 Announce Type: cross Abstract: As machine learned models increase in complexity and expressive power, features of simpler models, such as interpretability and control over the shape

local-aiarxiv-cs-lg
11 Aug 2026
Applications

Tracing sources of epistemic uncertainty in deep learning predictions: homo- and hetero-scedastic linearized estimators

DGX agent

arXiv:2608.07630v1 Announce Type: new Abstract: We adapt two classical statistical estimators for quantifying uncertainty to modern deep learning, in order to provide clearer insights into uncertainty

applicationsarxiv-cs-lg
11 Aug 2026
Model Releases

Tracking the Best Strategy in an Extensive-Form Game

DGX agent

arXiv:2608.09501v1 Announce Type: new Abstract: We consider the extensive-form bandit problem where on each trial the learner plays an extensive-form game against an oblivious adversary. We focus on t

model-releasesarxiv-cs-lg
11 Aug 2026
Research

Training-Free Universal Approximation by Prompting Random Transformers

DGX agent

arXiv:2608.09558v1 Announce Type: new Abstract: How expressive is prompting a transformer? Answering this question is important for separating the roles of prompting, architecture, and pretraining in

researcharxiv-cs-lg
11 Aug 2026
Model Releases

Trajectory Design and Budgeted Querying for Digital Twin Calibration

DGX agent

arXiv:2608.08631v1 Announce Type: new Abstract: Digital-twin calibration requires interaction data that is expensive to collect. We study two acquisition decisions: which trajectories to generate, and

model-releasesarxiv-cs-lg
11 Aug 2026
Research

Transfer Learning-Enabled Distortion Compensation for Amplitude-Phase-Time Block Modulation-Based Nonlinear Single-Carrier Wireless Communications

DGX agent

arXiv:2608.08554v1 Announce Type: cross Abstract: Power amplifier (PA) nonlinearity and memory effects significantly limit the spectral compliance, reliability, and energy efficiency of communication

researcharxiv-cs-lg
11 Aug 2026
Research

Transformers for Multimodal Brain State Decoding: Integrating Functional Magnetic Resonance Imaging Data and Medical Metadata

DGX agent

arXiv:2512.08462v2 Announce Type: replace Abstract: Decoding brain states from functional magnetic resonance imaging (fMRI) data is vital for advancing neuroscience and clinical applications. While tr

researcharxiv-cs-lg
11 Aug 2026
Tutorials

TS-Mob: Social and Geographical-Aware Time Series Foundation-Model Framework for Human Mobility Prediction

DGX agent

arXiv:2507.00945v2 Announce Type: replace Abstract: Short-term forecasting of aggregated human mobility flows supports urban planning, intelligent transportation systems, and emergency response, yet e

tutorialsarxiv-cs-lg
11 Aug 2026
Research

TSDS-Toolbox: A Toolbox for Measuring Time-Series Dataset Similarity

DGX agent

arXiv:2608.08119v1 Announce Type: new Abstract: The rapid advancement of artificial intelligence (AI) has significantly accelerated research in time-series analysis, particularly in forecasting, class

researcharxiv-cs-lg
11 Aug 2026
Research

Twin Rollouts: Noise-Coupled Counterfactual Branching in Interactive Video World Models

DGX agent

arXiv:2608.08982v1 Announce Type: new Abstract: Interactive video world models generate rollouts autoregressively under an action stream, yet they are trained and evaluated almost exclusively on factu

researcharxiv-cs-lg
11 Aug 2026
Research

Understanding Alternating Minimization for Matrix Completion

DGX agent

arXiv:1312.0925v4 Announce Type: replace Abstract: Alternating Minimization is a widely used and empirically successful heuristic for matrix completion and related low-rank optimization problems. Theo

researcharxiv-cs-lg
11 Aug 2026
Safety

Unimodality-Promoting Regularized Learning for Ordinal Regression

DGX agent

arXiv:2608.08359v1 Announce Type: new Abstract: Ordinal regression, also called ordinal classification, is classification of ordinal data, in which the underlying target variable is categorical and co

safetyarxiv-cs-lg
11 Aug 2026
Applications

V-Simba: Unleashing the Architectural Potential of RL in Visual Continuous Control

DGX agent

arXiv:2608.07870v1 Announce Type: new Abstract: Improving sample efficiency remains a core challenge in reinforcement learning (RL), especially in real-world settings like robotics, where data collect

applicationsarxiv-cs-lg
11 Aug 2026
Research

Variance reduction in lattice QCD observables via normalizing flows

DGX agent

arXiv:2603.02984v2 Announce Type: replace-cross Abstract: Normalizing flows can be used to construct unbiased, reduced-variance estimators for lattice field theory observables that are defined by a de

researcharxiv-cs-lg
11 Aug 2026
Research

Walk-on-Spheres Monte Carlo and deep neural network approximations of elliptic PDEs with drift and killing

DGX agent

arXiv:2608.09494v1 Announce Type: cross Abstract: In this paper we provide Monte Carlo and deep neural network approximations for stochastic representations of solutions to linear elliptic partial dif

researcharxiv-cs-lg
11 Aug 2026
Model Releases

Weak Correlations as the Underlying Principle for Linearization of Gradient-Based Learning Systems

DGX agent

arXiv:2401.04013v2 Announce Type: replace Abstract: Deep learning models, such as wide neural networks, can be conceptualized as nonlinear dynamical physical systems characterized by a multitude of in

model-releasesarxiv-cs-lg
11 Aug 2026
Model Releases

What Would Fix This RAG Failure? Auditing Counterfactual Response with Paired Evidence Interventions

DGX agent

arXiv:2608.08944v1 Announce Type: cross Abstract: A failed retrieval-augmented generation (RAG) answer can be consistent with several unseen responses to evidence repair. We introduce Pair-ID, an offl

model-releasesarxiv-cs-lg
11 Aug 2026
Research

When Can Fraud Operations Authorize Automation? A Decision-Support Framework for Fresh Audit Evidence and Review Workload

DGX agent

arXiv:2608.08577v1 Announce Type: new Abstract: Fraud operations must allocate events among automatic approval, analyst review, and automatic blocking even though the labels needed to evaluate these a

researcharxiv-cs-lg
11 Aug 2026
Model Releases

When Counterbalancing Hides the Bias: Access-Conditioned Position Lock in Forced-Choice LLM Evaluation

DGX agent

arXiv:2607.10202v2 Announce Type: replace Abstract: Forced-choice probes with counterbalanced orientations are a standard tool for measuring language-model 'value dispositions,' and a concentration/ex

model-releasesarxiv-cs-lg
11 Aug 2026
Model Releases

When Do Task Vectors Interfere? Mapping the Validity Boundaries of Weight-Space Composition

DGX agent

arXiv:2608.09490v1 Announce Type: new Abstract: Task arithmetic treats fine-tuning displacements as composable directions in weight space, yet it remains unclear when parameter addition reflects predi

model-releasesarxiv-cs-lg
11 Aug 2026
Safety

When Does Trace-Driven Evaluation Mislead MoE Expert Caching? Replay Semantics, Workload Contamination, and Operating Regimes

DGX agent

arXiv:2608.07911v1 Announce Type: new Abstract: Mixture-of-Experts (MoE) models have outgrown accelerator memory, and offloading expert weights to host memory is now standard. This makes expert cache

safetyarxiv-cs-lg
11 Aug 2026
← Previous
1…56789…299
Next →