AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,557 results
3 Jun 2026

Critical evaluation of PINN for FWD inverse analysis and differentiable FEM as an alternative

Model ReleasesDGX agent

arXiv:2606.03210v1 Announce Type: cross Abstract: Automatic-differentiation-based inverse analysis methods, including physics-informed neural networks (PINNs) and differentiable programming, have rece

Data- and Variance-dependent Regret Bounds for Online Tabular MDPs

SafetyDGX agent

arXiv:2602.01903v2 Announce Type: replace Abstract: This work studies online episodic tabular Markov decision processes (MDPs) with known transitions and develops best-of-both-worlds algorithms that a

Data-Driven Forecasting of three-Component Seismograms Using Transformer Architectures

TutorialsDGX agent

arXiv:2606.02912v1 Announce Type: cross Abstract: Forecasting seismic waveforms beyond observed data remains challenging due to the nonlinear, dispersive, and multi-scale nature of seismic wave propag


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

DECA: Decentralizing Block-Wise Adam for Efficient LLM Full-Parameter Fine-Tuning on Non-IID Data

Model ReleasesDGX agent

arXiv:2606.03209v1 Announce Type: new Abstract: Fine-tuning large language models (LLMs) in privacy-sensitive and resource-constrained environments remains challenging. Since training data are often d

Decentralized Stochastic Nonconvex Optimization under the (L_0,L_1)-Smoothness

AgentsDGX agent

arXiv:2509.08726v3 Announce Type: replace-cross Abstract: This paper focuses on the decentralized stochastic optimization problem f(mathbf{x})=frac{1}{m}sum_{i=1}^m f_i(mathbf{x}) over a connected net

Demystifying Pipeline Parallelism: First Theory for PipeDream

Local AiDGX agent

arXiv:2606.03498v1 Announce Type: new Abstract: Training modern machine learning models increasingly requires computation to be distributed across many accelerators. Data parallelism remains the defau

Denoise First, Orthogonalize Later: Understanding Momentum in Muon via Spectral Filtering

SafetyDGX agent

arXiv:2606.03899v1 Announce Type: new Abstract: Muon has recently demonstrated strong empirical performance in large language model training, but the theoretical role of momentum in Muon remains uncle

DiffUNet^2: Bidirectional Prediction, Probabilistic Generation and Collaborative Visual Discovery for Scientific Data

ResearchDGX agent

arXiv:2606.03926v1 Announce Type: cross Abstract: Modeling temporal evolution is important to analyzing and reasoning about scientific phenomena, yet most machine learning methods provide deterministi

Discovering autonomous quantum error correction via deep reinforcement learning

SafetyDGX agent

arXiv:2511.12482v2 Announce Type: replace-cross Abstract: Quantum error correction is essential for fault-tolerant quantum computing. However, standard methods relying on active measurements may intro

DRAN: A Distribution and Relation Adaptive Network for Spatio-temporal Forecasting

ResearchDGX agent

arXiv:2504.01531v4 Announce Type: replace Abstract: Accurate predictions of spatio-temporal systems are crucial for tasks such as system management, control, and crisis prevention. However, the inhere

DriftSched: Adaptive QoS-Aware Scheduling under Runtime Token Drift for Multi-Tenant GPU Inference

SafetyDGX agent

arXiv:2606.02982v1 Announce Type: cross Abstract: The rapid growth of large language model (LLM) inference services has increased the demand for efficient multi-tenant GPU scheduling. While modern inf

Easy-to-Use Shielding for Reinforcement Learning

SafetyDGX agent

arXiv:2606.03804v1 Announce Type: new Abstract: Safe exploration is a key challenge in Reinforcement Learning (RL) that aims to prevent agents from making harmful decisions while exploring their envir

Enhanced Renewable Energy Forecasting using Context-Aware Conformal Prediction

Local AiDGX agent

arXiv:2510.15780v2 Announce Type: replace-cross Abstract: Artificial intelligence (AI) is increasingly used to support renewable energy forecasting and grid operations. As renewable penetration grows,

ERP-XTTN: Interpretable Prototype-Guided Cross-Attention for Cross-Subject ERP Classification

Model ReleasesDGX agent

arXiv:2606.02939v1 Announce Type: new Abstract: Interpretable brain-computer interface classifiers that generalize across subjects without calibration remain an open challenge. We test whether prototy

Estimating Bidirectional Causal Effects with Large Scale Online Kernel Learning

SafetyDGX agent

arXiv:2511.05050v3 Announce Type: replace-cross Abstract: In this study, a scalable online kernel learning framework is proposed for estimating bidirectional causal effects in systems characterized by

Explainable Forecasting of Scientific Breakthroughs from Concept Network Dynamics

SafetyDGX agent

arXiv:2606.03864v1 Announce Type: cross Abstract: We introduce an explainable machine-learning approach that forecasts the structural precursors of scientific breakthroughs -- the emergence and intens

Fairness Definitions and Metrics in Deep Reinforcement Learning for Drug Discovery in Healthcare: A Rapid Evidence Review

SafetyDGX agent

arXiv:2606.02902v1 Announce Type: cross Abstract: Deep reinforcement learning (DRL) is increasingly applied to de novo molecular design, but choices in data, rewards, and evaluation can yield uneven p

Fast and Expressive Multi-Byte Prediction with Probabilistic Circuits

ResearchDGX agent

arXiv:2511.11346v2 Announce Type: replace Abstract: Multi-token prediction (MTP) is a prominent strategy to significantly speed up generation in large language models (LLMs), especially in byte-level

Fast Organic Crystal Structure Prediction with Unit Cell Flow Matching

SafetyDGX agent

arXiv:2606.03199v1 Announce Type: new Abstract: Organic crystal structure prediction (CSP) is a requirement for computational modelling of organic solids, but traditionally costs several CPU-years per

Fast Unlearning at Scale via Margin Self-Correction

ResearchDGX agent

arXiv:2606.02920v1 Announce Type: new Abstract: Language-model unlearning updates a trained model to behave as if it had not seen selected training examples, while preserving utility and avoiding cost

Few-Shot Prediction for Pulsar Noise with Long Short-Term Memory Network

AgentsDGX agent

arXiv:2606.03574v1 Announce Type: cross Abstract: This work proposes a novel solution to predict pulsar timing residuals with limited data, addressing the critical challenge of data scarcity across sp

FGRPO: Federated GRPO with Adaptive Aggregation on Non-IID Data

SafetyDGX agent

arXiv:2606.03094v1 Announce Type: new Abstract: Recent advances in language models have established reinforcement learning as the primary paradigm for eliciting self-correction and long-chain reasonin

Finding Needles in the Haystack: Transductive Active Labeling in Ecology

ResearchDGX agent

arXiv:2606.03821v1 Announce Type: new Abstract: Active learning is now standard practice in labeling ecological data, enabling ecologists to quickly process large volumes of field data to understand a

FinStressTS: A Parametric Synthetic Benchmark for Time-Series Forecasting in Finance

Model ReleasesDGX agent

arXiv:2606.03184v1 Announce Type: cross Abstract: Financial forecasting is difficult due to low signal-to-noise ratios, latent factors, heavy tails, regime shifts, and jumps. Real-world benchmarks off

Flicker-DDPM: Accelerating Denoising Diffusion via 1/f Colored Noise Injection

ResearchDGX agent

arXiv:2606.03393v1 Announce Type: new Abstract: We propose a novel diffusion model, Flicker-DDPM, which incorporates flicker (1/f) noise inspired by self-organized criticality (SOC), a widely observed

Flow Learners for PDEs: Toward a Physics-to-Physics Paradigm for Scientific Computing

SafetyDGX agent

arXiv:2604.07366v2 Announce Type: replace Abstract: Partial differential equations (PDEs) govern nearly every physical process in science and engineering, but solving them at scale remains prohibitive

Forecasting Conceptual Diffusion in Science: The Case of Quantum Computing

Model ReleasesDGX agent

arXiv:2606.03919v1 Announce Type: cross Abstract: Understanding and anticipating scientific change requires models that distinguish between endogenous consolidation and exogenous diffusion of scientif

From Non-Convex to Strongly Convex: Curvature-Adaptive FTPL for Online Optimization

ResearchDGX agent

arXiv:2606.02948v1 Announce Type: new Abstract: Curvature adaptivity is a classical theme in online optimization: for convex Lipschitz losses, adaptive methods interpolate between the optimal O(sqrt{T

Gate AI: LLM Security Benchmark Evaluation Methodology and Results

Model ReleasesDGX agent

arXiv:2606.02959v1 Announce Type: new Abstract: Published evaluations of prompt-injection and jailbreak detectors for Large Language Models often suffer from two systematic weaknesses: per-dataset thr

Generating Rectifiable Measures through Neural Networks

Model ReleasesDGX agent

arXiv:2412.05109v2 Announce Type: replace Abstract: We derive universal approximation results for the class of (countably) m-rectifiable measures. Specifically, we prove that m-rectifiable measures ca

Grounding Functional Similarity by Invariance-Aware Model Stitching

TutorialsDGX agent

arXiv:2505.20142v2 Announce Type: replace Abstract: In deep learning, functional similarity evaluation quantifies the extent to which independently trained models learn similar input--output relations

HARVE: Hacking-Aware Reward-Head Vector Editing for Robust Reward Models

SafetyDGX agent

arXiv:2606.03131v1 Announce Type: new Abstract: Reward models are central to large language model (LLM) alignment, but they remain vulnerable to reward hacking. To evaluate reward-model robustness, we

Hierarchical RBF-KAN and RBF-SKAN Architectures for Multidimensional Function Approximation and Random Field Learning

TutorialsDGX agent

arXiv:2606.02936v1 Announce Type: new Abstract: In this manuscript, we propose and analyze hierarchical Kolmogorov--Arnold neural network architectures employing radial basis functions as activation f

Hierarchies of Calibration: Classification meets Regression

ResearchDGX agent

arXiv:2606.03245v1 Announce Type: cross Abstract: Concepts of calibration formalize the compatibility between probabilistic predictions and the respective outcomes. In a nutshell, the outcomes ought t

High-Dimensional Latents Should Be Diagnosed Through Phase Structure

ApplicationsDGX agent

arXiv:2606.02600v1 Announce Type: cross Abstract: We study autoencoder and variational-autoencoder latent spaces through the lens of spin-glass theory. The paper has two components. First, we formaliz

HiSE: A Lightweight Hierarchical Semantic Explainer for Heterogeneous Graph Neural Networks

TutorialsDGX agent

arXiv:2606.03495v1 Announce Type: new Abstract: Heterogeneous graph neural networks (HGNNs) have demonstrated remarkable performance in modeling complex relational data, however their interpretability

Honesty in Causal Forests: When It Helps and When It Hurts

Model ReleasesDGX agent

arXiv:2506.13107v4 Announce Type: replace Abstract: Causal forests estimate how treatment effects vary across individuals, guiding personalized interventions in areas like marketing, operations, and p

How Many Trees in a Random Forest? A Revisited Approach with Plateau Search and Optuna Integration

Model ReleasesDGX agent

arXiv:2606.03549v1 Announce Type: new Abstract: Hyperparameter optimization (HPO) for Random Forest faces a specific difficulty in tuning the number of trees: the predictive score typically improves m

How Visible Are Silent Manipulation Failures? An Observability Study of False-Success Detection in Simulated Robot Episodes

ResearchDGX agent

arXiv:2606.03134v1 Announce Type: cross Abstract: Imitation-learning policies for robot manipulation inherit the quality of the success labels attached to their training episodes, and those labels are

Human-in-the-Loop Contextual Bandits for Short-Term Rental Dynamic Pricing: Structural Equivalence of Historical Warm-Up and Approval-Gated Live Learning

SafetyDGX agent

arXiv:2606.02595v1 Announce Type: new Abstract: Dynamic pricing in short-term rental (STR) markets presents a distinctive challenge for online learning algorithms: pricing decisions carry significant

Hybrid Adaptive Kalman Filtering for Data-Efficient Joint Tracking and Classification

ApplicationsDGX agent

arXiv:2606.02767v1 Announce Type: cross Abstract: Kalman filtering performance is highly sensitive to model mismatch and noise covariance tuning. Learning-based approaches address these limitations bu

Impact of Graph Structure on Membership-Inference Risk for Graph Neural Networks

SafetyDGX agent

arXiv:2601.17130v2 Announce Type: replace Abstract: Graph neural networks (GNNs) are widely used for tasks such as node classification and link prediction, but their use in sensitive settings raises c

Jailbreak Attack Initializations as Extractors of Compliance Directions

SafetyDGX agent

arXiv:2502.09755v4 Announce Type: replace-cross Abstract: Safety-aligned LLMs respond to prompts with either compliance or refusal, each corresponding to distinct directions in the model's activation

KForge: LLM-Driven Cross-Platform Kernel Generation for AI Accelerators

Model ReleasesDGX agent

arXiv:2606.02963v1 Announce Type: new Abstract: Production inference increasingly targets a heterogeneous mix of accelerators. Agentic pipelines interleave reasoning, tool calls, and multi-agent coord

KVarN: Variance-Normalized KV-Cache Quantization Mitigates Error Accumulation in Reasoning Tasks

ResearchDGX agent

arXiv:2606.03458v1 Announce Type: new Abstract: Test-time scaling is a powerful approach to obtain better reasoning in large language models, but it becomes memory-bottlenecked during long-horizon dec

Laplacian Representations for Decision-Time Planning

Model ReleasesDGX agent

arXiv:2602.05031v2 Announce Type: replace Abstract: Planning with a learned model remains a key challenge in model-based reinforcement learning (RL). In decision-time planning, state representations a

LC-SAC: Lyapunov-Constrained Soft Actor-Critic via Koopman Operator Theory for Trajectory Tracking and Stabilization

SafetyDGX agent

arXiv:2602.04132v4 Announce Type: replace-cross Abstract: Reinforcement Learning (RL) has achieved remarkable success in solving complex sequential decision-making problems. However, its application t

Learning Coherent Representations: A Topological Approach to Interpretability

TutorialsDGX agent

arXiv:2606.02841v1 Announce Type: new Abstract: Deep neural networks learn representations where individual features often lack interpretable meaning; a single neuron may activate for scattered, unrel

Learning DNF through Generalized Fourier Representations

ResearchDGX agent

arXiv:2506.01075v2 Announce Type: replace-cross Abstract: The Boolean Fourier representation has been widely used in learning theory, particularly for learning Disjunctive Normal Form (DNF) under unif

Learning Power Flow with Confidence: A Probabilistic Guarantee Framework for Voltage Risk

SafetyDGX agent

arXiv:2308.07867v4 Announce Type: replace-cross Abstract: The absence of formal performance guarantees in machine learning (ML) has limited its adoption for safety-critical power system applications,

Learning Temporal Causal Structure via Smooth Differentiable Optimization

Model ReleasesDGX agent

arXiv:2606.03227v1 Announce Type: new Abstract: Causal discovery with instantaneous effects in multivariate time series is challenging, as the instantaneous structure must be acyclic. Prior methods en

Learning to Bet for Horizon-Aware Anytime-Valid Testing

SafetyDGX agent

arXiv:2603.19551v2 Announce Type: replace-cross Abstract: We develop horizon-aware anytime-valid tests and confidence sequences for bounded means under a strict deadline N. Using the betting/e-process

Learning to Solve, Forgetting to Retain: Correct-Set Turnover in RLVR

ResearchDGX agent

arXiv:2606.03087v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) improves the ability of large language model, yet headline accuracy gains often conceal a hidden c

Learning Unmasking Policies for Diffusion Language Models

SafetyDGX agent

arXiv:2512.09106v4 Announce Type: replace Abstract: Diffusion (Large) Language Models (dLLMs) now match the downstream performance of their autoregressive counterparts on many tasks, while holding the

Let There Be Light: Reflection, Refraction and Scattering for Neural Operators

SafetyDGX agent

arXiv:2606.03262v1 Announce Type: new Abstract: Neural operators learn mappings between infinite-dimensional function spaces and provide a data-driven surrogate modeling paradigm for parametric partia

Lethe: Adapter-Augmented Dual-Stream Update for Persistent Knowledge Erasure in Federated Unlearning

TutorialsDGX agent

arXiv:2601.22601v2 Announce Type: replace Abstract: Federated unlearning (FU) aims to erase designated client-level, class-level, or sample-level knowledge from a global model. Existing studies common

Limit Analysis of Graph Neural Networks with Wireless Conflict Graphs

ResearchDGX agent

arXiv:2606.03794v1 Announce Type: new Abstract: Graph Neural Networks (GNNs) have emerged as a powerful tool for wireless resource allocation that leverages the underlying graph structure of communica

Link Prediction or Perdition: the Seeds of Instability in Knowledge Graph Embeddings

ResearchDGX agent

arXiv:2606.03365v1 Announce Type: new Abstract: Embedding models (KGEMs) constitute the main link prediction approach to complete knowledge graphs. Standard evaluation protocols emphasize rank-based m

Locality Does Not Imply Reachability: Boundary Repair in Block-Sparse Causal Attention

Model ReleasesDGX agent

arXiv:2606.02680v1 Announce Type: new Abstract: Sparse causal attention is usually described by sequence locality: nearby tokens should remain easy to access, while distant tokens may be dropped to re

Localized, High-resolution Geographic Representations with Slepian Functions

SafetyDGX agent

arXiv:2602.00392v2 Announce Type: replace Abstract: Geographic data is fundamentally local. Disease outbreaks cluster in population centers, ecological patterns emerge along coastlines, and economic a

← Previous
1…101102103104105…243
Next →