AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,557 results
21 May 2026

Adaptive Signal Resuscitation: Channel-wise Post-Pruning Repair for Sparse Vision Networks

ResearchDGX agent

arXiv:2605.21426v1 Announce Type: new Abstract: One-shot magnitude pruning can cause severe accuracy collapse in the high-sparsity regime, even when the pruning mask preserves the largest weights. We

Advanced Scientific Methodology Plays Rossini

ResearchDGX agent

arXiv:2605.20220v1 Announce Type: cross Abstract: A musical score provides the essential instructions for its performance while containing indications - at times implicit - regarding the composer's in

Advantage Collapse in Group Relative Policy Optimization: Diagnosis and Mitigation

SafetyDGX agent

arXiv:2605.21125v1 Announce Type: new Abstract: Group Relative Policy Optimization (GRPO), a prominent algorithm within the Reinforcement Learning from Verifiable Rewards (RLVR) framework, has achieve


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Adversarial Robustness in One-Stage Learning-to-Defer

Model ReleasesDGX agent

arXiv:2510.10988v2 Announce Type: replace-cross Abstract: Learning-to-Defer (L2D) enables hybrid decision-making by routing inputs either to a predictor or to external experts. While promising, L2D is

Agent JIT Compilation for Latency-Optimizing Web Agent Planning and Scheduling

AgentsDGX agent

arXiv:2605.21470v1 Announce Type: new Abstract: Computer-use agents (CUA) automate tasks specified with natural language such as 'order the cheapest item from Taco Bell' by generating sequences of cal

Agentic Physical AI toward a Domain-Specific Foundation Model for Nuclear Reactor Control

Model ReleasesDGX agent

arXiv:2512.23292v3 Announce Type: replace-cross Abstract: The prevailing paradigm in AI for physical systems (scaling general-purpose foundation models toward universal multimodal reasoning) confronts

AGPO: Adaptive Group Policy Optimization with Dual Statistical Feedback

Model ReleasesDGX agent

arXiv:2605.20722v1 Announce Type: new Abstract: Reinforcement learning improves LLM reasoning, but PPO/GRPO typically use fixed clipping and decoding temperature, which makes training brittle and tuni

AI-based Prediction of Independent Construction Safety Outcomes from Universal Attributes

SafetyDGX agent

arXiv:1908.05972v3 Announce Type: replace Abstract: This paper significantly improves on, and finishes to validate, an approach proposed in previous research in which safety outcomes were predicted fr

AIMBio-Mat: An AI-Native FAIR Platform for Closed-Loop Materials Discovery and Biomedical Translation

SafetyDGX agent

arXiv:2605.21083v1 Announce Type: cross Abstract: Materials discovery and biomedical translation increasingly require models that can reason across composition, processing, structure, biological respo

AirfoilGen: A valid-by-construction and performance-aware latent diffusion model for airfoil generation

ResearchDGX agent

arXiv:2605.20303v1 Announce Type: new Abstract: Airfoil shape design is a fundamental task in aerospace engineering, with a direct impact on flight stability and fuel consumption. Deep learning has re

AMAR: Lightweight Attention-Based Multi-User Activity Recognition from Wi-Fi CSI

Model ReleasesDGX agent

arXiv:2605.20649v1 Announce Type: cross Abstract: Wi-Fi-based human activity recognition (HAR) has emerged as a promising approach for contactless sensing, leveraging channel state information (CSI) c

An exponential mechanism based on quadratic approximations for fine-tuning machine learning models with privacy guarantees

Model ReleasesDGX agent

arXiv:2605.20521v1 Announce Type: new Abstract: Fine-tuning adapts a pretrained machine learning model to a small, sensitive dataset, but this process risks memorizing individual new data points, maki

APEX: Autonomous Policy Exploration for Self-Evolving LLM Agents

Model ReleasesDGX agent

arXiv:2605.21240v1 Announce Type: new Abstract: LLM agents have shown strong performance across a wide range of complex tasks, including interactive environments that require long-horizon decision mak

Approximation Theory for Neural Networks: Old and New

Model ReleasesDGX agent

arXiv:2605.21451v1 Announce Type: new Abstract: Universal approximation theorems provide a mathematical explanation for the expressive power of neural networks. They assert that, under mild conditions

Augmented Analytics and Decision Quality: The Role of Trust among Non-Technical BI Users

ResearchDGX agent

arXiv:2605.20198v1 Announce Type: cross Abstract: Augmented analytics has transformed how business intelligence (BI) systems support managerial decision-making. This is especially true for users witho

Automated Byzantine-Resilient Clustered Decentralized Federated Learning for Battery Intelligence in Connected EVs

SafetyDGX agent

arXiv:2605.21115v1 Announce Type: cross Abstract: Federated learning (FL) has emerged as a promising paradigm for managing electric vehicle (EV) battery data in intelligent transportation systems (ITS

Automated Kernel Discovery Towards Understanding High-dimensional Bayesian Optimization

ResearchDGX agent

arXiv:2605.20249v1 Announce Type: new Abstract: Gaussian Process (GP) kernels are central to Bayesian optimization (BO), yet designing effective kernels for high-dimensional problems still relies on e

Axiomatizing Neural Networks via Pursuit of Subspaces

ResearchDGX agent

arXiv:2605.20534v1 Announce Type: new Abstract: While deep neural networks have achieved remarkable success across a wide range of domains, their underlying mechanisms remain poorly understood, and th

BALLAST: Bayesian Active Learning with Look-ahead Amendment for Sea-drifter Trajectories under Spatio-Temporal Vector Fields

ResearchDGX agent

arXiv:2509.26005v3 Announce Type: replace-cross Abstract: We introduce a formal active learning methodology for guiding the placement of Lagrangian observers to infer time-dependent vector fields -- a

Batched Single-Index Global Multi-Armed Bandits with Covariates

Model ReleasesDGX agent

arXiv:2503.00565v3 Announce Type: replace-cross Abstract: The multi-armed bandits (MAB) framework is a widely used approach for sequential decision-making, where a decision-maker selects an arm in eac

Bayesian Optimization by Kernel Regression and Density-based Exploration

TutorialsDGX agent

arXiv:2502.06178v5 Announce Type: replace-cross Abstract: Bayesian optimization is highly effective for optimizing expensive-to-evaluate black-box functions, but it faces significant computational cha

Behavior-Consistent Deep Reinforcement Learning

SafetyDGX agent

arXiv:2605.21214v1 Announce Type: new Abstract: Reinforcement learning (RL) often exhibits high variance across training runs, leading to unreliable performance and posing a major challenge to deploym

Beyond Numerical Features: CNN-Driven Algorithm Selection via Contour Plots for Continuous Black-Box Optimization

ResearchDGX agent

arXiv:2605.20797v1 Announce Type: new Abstract: The present paper introduces a new representation-driven approach to per-instance algorithm selection, applied to black-box optimization, for automatica

Beyond the Bellman Recursion: A Pontryagin-Guided Framework for Non-Exponential Discounting

SafetyDGX agent

arXiv:2605.20996v1 Announce Type: new Abstract: Most value-based and actor--critic reinforcement learning methods rely on Bellman-style recursions, yet these recursions collapse under non-exponential

C^2FG: Control Classifier-Free Guidance via Score Discrepancy Analysis

ResearchDGX agent

arXiv:2603.08155v3 Announce Type: replace Abstract: Classifier-Free Guidance (CFG) is a cornerstone of modern conditional diffusion models, yet its reliance on the fixed or heuristic dynamic guidance

CAdam: Context-Adaptive Moment Estimation for 3D Gaussian Densification in Generative Distillation

ResearchDGX agent

arXiv:2605.20872v1 Announce Type: new Abstract: Adaptive densification is the engine of 3D Gaussian Splatting (3DGS). However, when transposed to the optimization-based Generative Distillation paradig

Can Conversational XAI Improve User Performance? An Experimental Study

ResearchDGX agent

arXiv:2605.20439v1 Announce Type: new Abstract: Explainable AI (XAI) techniques aim to provide insights into predictive models and enhance user performance, yet they often fall short of these expectat

Can Microcanonical Langevin Dynamics Leverage Mini-Batch Gradient Noise?

SafetyDGX agent

arXiv:2602.06500v2 Announce Type: replace Abstract: Scaling inference methods such as Markov chain Monte Carlo to high-dimensional models remains a central challenge in Bayesian deep learning. A promi

CASCADE Conformal Prediction: Uncertainty-Adaptive Prediction Intervals for Two-Stage Clinical Decision Support

ResearchDGX agent

arXiv:2605.20468v1 Announce Type: new Abstract: Effective medication management in Parkinson's Disease (PD) is challenging due to heterogeneous disease progression, variable patient response, and medi

Catching a Moving Subspace: Low-Rank Bandits Beyond Stationarity

ApplicationsDGX agent

arXiv:2605.20269v1 Announce Type: new Abstract: Many bandit deployments (recommendation, clinical dosing, ad targeting) share two facts prior work handles only in isolation: rewards live on a low-dime

Causal Discovery from Heteroscedastic Stochastic Dynamical Systems under Imperfect Physical Models

ApplicationsDGX agent

arXiv:2602.04907v2 Announce Type: replace Abstract: Causal discovery is a data-driven paradigm for analyzing complex systems, while physics-based models, such as ordinary differential equations (ODEs)

Causal Machine Learning Is Not a Panacea: A Roadmap for Observational Causal Inference in Health

ResearchDGX agent

arXiv:2605.20782v1 Announce Type: new Abstract: Objective: The growing availability of large-scale observational clinical datasets and challenges in conducting randomized controlled trials have spurre

Causal Unlearning in Collaborative Optimization: Exact and Approximate Influence Reversal under Adversarial Contributions

Model ReleasesDGX agent

arXiv:2605.20341v1 Announce Type: new Abstract: Federated learning systems must support data deletion requests to comply with privacy regulations, yet retraining from scratch after each deletion is co

Choose Wisely and Privately: Proactive Client Selection for Fair and Efficient Federated Learning

SafetyDGX agent

arXiv:2605.20975v1 Announce Type: new Abstract: Federated Learning enables collaborative model training across decentralized data sources without data transfer. Averaging-based FL is limited by the pr

CIG: Exploration via Conditional Information Gain

ResearchDGX agent

arXiv:2605.20878v1 Announce Type: new Abstract: Intrinsic rewards for exploration in reinforcement learning condition on different contexts: lifelong rewards score each transition against accumulated

Classification of Single and Mixed Partial Discharges under Switching Voltage Using an AWA-CNN Framework

ResearchDGX agent

arXiv:2605.21352v1 Announce Type: new Abstract: The growing use of fast-switching power electronics has made partial discharge (PD) analysis under switching-voltage excitation increasingly important,

Closed-form predictive coding via hierarchical Gaussian filters

Local AiDGX agent

arXiv:2605.20293v1 Announce Type: new Abstract: Predictive coding (PC) offers a local and biologically grounded alternative to backpropagation in the training of artificial neural networks, yet to dat

Cluster-Based Generalized Additive Models Informed by Random Fourier Features

Model ReleasesDGX agent

arXiv:2512.19373v3 Announce Type: replace-cross Abstract: In developing data-driven modeling methodologies, there is an ongoing need to reconcile the strong predictive performance of opaque black-box

CoarseSoundNet: Building a reliable model for ecological soundscape analysis

ApplicationsDGX agent

arXiv:2605.21143v1 Announce Type: cross Abstract: A soundscape is composed of three types of sound: biophony (sounds made by animals), geophony (natural abiotic sounds) and anthropophony (sounds made

Code Generation by Differential Test Time Scaling

AgentsDGX agent

arXiv:2605.20473v1 Announce Type: cross Abstract: Test-time scaling has emerged as a promising approach for improving code generation by exploring large solution spaces at inference time. However, exi

Comparing Explanations is Not Enough, Explain the Change: New Standards are Needed to Explain Behavioral Shifts in Large Language Models

SafetyDGX agent

arXiv:2602.02304v2 Announce Type: replace-cross Abstract: Large-scale foundation models exhibit behavioral shifts when subjected to interventions such as scaling, fine-tuning, reinforcement learning w

CompilerKV: Risk-Adaptive KV Compression via Offline Experience Compilation

SafetyDGX agent

arXiv:2602.08686v2 Announce Type: replace Abstract: Prefill-only KV compression freezes a token subset at the end of prefill and decodes from it without further eviction. The retention decision is the

Complementing reinforcement learning with SFT through logit averaging in the post training of LLMs

SafetyDGX agent

arXiv:2605.20555v1 Announce Type: new Abstract: We introduce a novel method that averages the logits of a frozen reference policy (e.g., SFT) and a trainable policy, and incorporate the method into Gr

Compositional Transduction with Latent Analogies for Offline Goal-Conditioned Reinforcement Learning

AgentsDGX agent

arXiv:2605.20609v1 Announce Type: new Abstract: Compositional generalization is essential for reaching unseen goals under novel contextual variations in offline goal-conditioned reinforcement learning

Computational-Statistical Trade-off in Kernel Two-Sample Testing with Random Fourier Features

ResearchDGX agent

arXiv:2407.08976v2 Announce Type: replace-cross Abstract: Recent years have seen a surge in methods for two-sample testing, among which the Maximum Mean Discrepancy (MMD) test has emerged as an effect

Compute Only Once: UG-Separation for Efficient Large Recommendation Models

ResearchDGX agent

arXiv:2602.10455v2 Announce Type: replace-cross Abstract: Driven by scaling laws, recommender systems increasingly rely on larger-scale models to capture complex feature interactions and user behavior

Concentration of General Stochastic Approximation Under Heavy-Tailed Markovian Noise

ResearchDGX agent

arXiv:2605.20999v1 Announce Type: cross Abstract: We establish maximal concentration bounds for the iterates generated by stochastic approximation algorithms with general step sizes, where the noise h

Conditional Equivalence of DPO and RLHF: Implicit Assumption, Failure Modes, and Provable Alignment

SafetyDGX agent

arXiv:2605.20834v1 Announce Type: cross Abstract: Direct Preference Optimization (DPO) has emerged as a popular alternative to Reinforcement Learning from Human Feedback (RLHF), offering theoretical e

Conditioning Gaussian Processes on Almost Anything

ApplicationsDGX agent

arXiv:2605.21041v1 Announce Type: cross Abstract: Gaussian processes (GPs) offer a principled probabilistic model over functions, but exact inference is restricted to the linear-Gaussian regime. We es

Conformal Selective Acting: Anytime-Valid Risk Control for RLVR-Trained LLMs

Local AiDGX agent

arXiv:2605.20270v1 Announce Type: new Abstract: A local specialist LLM, fine-tuned with reinforcement learning from verifiable rewards (RLVR) on operator-local data, is installed in a regulated organi

Consistent Geometric Deep Learning via Hilbert Bundles and Cellular Sheaves

ApplicationsDGX agent

arXiv:2605.06395v2 Announce Type: replace Abstract: Modern deep learning architectures increasingly contend with sophisticated signals that are natively infinite-dimensional, such as time series, prob

Consistently Informative Soft-Label Temperature for Knowledge Distillation

SafetyDGX agent

arXiv:2605.20357v1 Announce Type: new Abstract: Knowledge distillation (KD) transfers knowledge from a high-capacity teacher to a compact student by matching their predictive distributions, with tempe

Contradiction Graphs Determine VC Dimension

ResearchDGX agent

arXiv:2605.20434v1 Announce Type: cross Abstract: We study the contradiction graphs associated with binary concept classes. For a class H subseteq {0,1}^X, the order-m contradiction graph G_m(H) has a

Control and optimization for Neural Partial Differential Equations in Supervised Learning

ResearchDGX agent

arXiv:2506.20764v2 Announce Type: replace-cross Abstract: Although there is a substantial body of literature on control and optimization problems for parabolic and hyperbolic systems, the specific pro

Control, Optimal Transport and Neural Differential Equations in Supervised Learning

ResearchDGX agent

arXiv:2503.15105v4 Announce Type: replace-cross Abstract: We study the fundamental computational problem of approximating optimal transport (OT) equations using neural differential equations (Neural O

Corrected Integrated Laplace Approximation for Bayesian Inference in Latent Gaussian Models

ResearchDGX agent

arXiv:2605.20345v1 Announce Type: cross Abstract: Latent Gaussian models (LGMs) are a popular class of Bayesian hierarchical models that include Gaussian processes, as well as certain spatial models a

Correcting Stochastic Update Bias in Preconditioned Language Model Optimizers

SafetyDGX agent

arXiv:2605.20756v1 Announce Type: new Abstract: Preconditioned optimizers are central to language model training, but their stochastic update rules are usually treated as direct approximations to popu

CRAFT: Conflict-Resolved Aggregation for Federated Training

SafetyDGX agent

arXiv:2605.21317v1 Announce Type: new Abstract: The aggregation of conflicting client updates remains a fundamental bottleneck in federated learning (FL) under heterogeneous data distributions. Naive

CRANE: Correcting Errors in Raw Nanopore Signals Using Hidden Markov Models

ResearchDGX agent

arXiv:2603.20420v2 Announce Type: replace-cross Abstract: Nanopore sequencing can read substantially longer sequences of nucleic acid molecules, called reads, than other sequencing methods, which has

CT-OT Flow: Estimating Continuous-Time Dynamics from Discrete Temporal Snapshots

ApplicationsDGX agent

arXiv:2505.17354v3 Announce Type: replace Abstract: In many real-world settings--e.g., single-cell RNA sequencing, mobility sensing, and environmental monitoring--data are observed only as temporally

← Previous
1…137138139140141…243
Next →