AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,432 results
26 Jun 2026

Decentralized Best-Response-Based Learning in Two-Player Zero-Sum Stochastic Games: A Finite-Sample Analysis

SafetyDGX agent

arXiv:2409.01447v3 Announce Type: replace Abstract: We present a finite-sample analysis of decentralized learning in two-player zero-sum matrix games and stochastic games, with a focus on best-respons

Designing Reward Signals for Portable Query Generation: A Case Study in Industrial Semantic Job Search

SafetyDGX agent

arXiv:2606.27291v1 Announce Type: new Abstract: Job-search platforms rely on low-bandwidth query interfaces that often fail to capture the high-dimensional complexity of candidate profiles. We present

DMuon: Efficient Distributed Muon Training with Near-Adam Overhead

ResearchDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2606.27153v1 Announce Type: cross Abstract: Matrix-orthogonalization-based optimizers, exemplified by Muon, have demonstrated strong convergence behavior across a wide range of modern deep learn

Does Aurora Encode Atmospheric Structure? Latent Regime Analysis and Attribution

ResearchDGX agent

arXiv:2606.26361v1 Announce Type: new Abstract: ML foundation models are able to emulate atmospheric dynamics accurately and efficiently but operate as opaque ``black boxes''. We investigate the inter

DroidBreaker: Practical and Functional Problem-Space Attacks on Machine-Learning Android Malware Detectors

ResearchDGX agent

arXiv:2606.26707v1 Announce Type: cross Abstract: Adversarial APKs are Android applications modified in the problem space to evade machine-learning malware detectors. In this work, we first show that,

Effective Covariance Dynamics in Solvable High-Dimensional GANs

SafetyDGX agent

arXiv:2606.27246v1 Announce Type: new Abstract: We study a solvable high-dimensional model of generative adversarial network (GAN) training in which a linear generator learns a low-dimensional subspac

Efficient learning of bosonic Gaussian unitaries

ResearchDGX agent

arXiv:2510.05531v2 Announce Type: replace-cross Abstract: Bosonic Gaussian unitaries are fundamental building blocks of central continuous-variable quantum technologies such as quantum-optic interfero

EMA-FS: Accelerating GBDT Training via Gain-Informed Feature Screening

Model ReleasesDGX agent

arXiv:2606.26337v1 Announce Type: new Abstract: Gradient Boosted Decision Trees (GBDT), exemplified by LightGBM, spend a dominant fraction of training time -- typically 65-70% -- constructing per-feat

Embedding Foundation Model Predictions in Discrete-Choice Models with Structural Guarantees

ResearchDGX agent

arXiv:2606.26432v1 Announce Type: new Abstract: Tabular foundation models achieve strong accuracy on choice prediction tasks, but their predictions often violate the economic logic those tasks require

Empirical Software Engineering TerraProbe: A Layered-Oracle Framework for Detecting Deceptive Fixes in LLM-Assisted Terraform

Model ReleasesDGX agent

arXiv:2606.26590v1 Announce Type: new Abstract: Security misconfigurations in Terraform Infrastructure-as-Code are a growing risk in cloud deployments, and large language models are increasingly used

Enabling self-supervised learned primal dual with Noise2Inverse

ResearchDGX agent

arXiv:2606.26991v1 Announce Type: cross Abstract: X-ray computed tomography reconstruction is an ill-posed inverse problem, particularly in low-dose and sparse-angle settings where measurements are no

Equivariance and Augmentation for Bayesian Neural Networks

TutorialsDGX agent

arXiv:2606.26273v1 Announce Type: new Abstract: Symmetries are important for many deep learning tasks, ranging from applications in the sciences to medical imaging. However, there is an ongoing debate

Escaping Iterative Parameter-Space Noise: Differentially Private Learning with a Hypernetwork

Model ReleasesDGX agent

arXiv:2606.26772v1 Announce Type: new Abstract: Differentially private (DP) training of neural networks is often hindered by the large amount of noise required by gradient-based methods such as DP-SGD

Estimating Orbital Parameters of Direct Imaging Exoplanet Using Neural Network

Model ReleasesDGX agent

arXiv:2510.17459v3 Announce Type: replace-cross Abstract: In this work, we propose a flow-matching Markov chain Monte Carlo (FM-MCMC) algorithm for estimating the orbital parameters of exoplanetary sy

Explaining Temporal Graph Neural Networks via Feature-induced Information Flow

ApplicationsDGX agent

arXiv:2606.27201v1 Announce Type: new Abstract: Event-based Temporal Graph Neural Networks (ETGNNs) have demonstrated strong performance across a wide range of applications, including social network a

Fast algorithms for learning a Gaussian under halfspace truncation with optimal sample complexity

Model ReleasesDGX agent

arXiv:2606.27298v1 Announce Type: cross Abstract: We study the fundamental problem of learning a high-dimensional Gaussian truncated to an unknown halfspace. Lee, Mehrotra and Zampetakis (FOCS'24) rec

Federated Hash Projected Latent Factor Learning

ApplicationsDGX agent

arXiv:2606.26192v1 Announce Type: new Abstract: Hash Learning (HL) is an efficient representation learning approach that maps real-valued data into compact binary representations. Traditional HL metho

Finding Stationary Points by Comparisons

ResearchDGX agent

arXiv:2606.27082v1 Announce Type: new Abstract: We study the problem of finding stationary points of non-convex functions when access to the objective is provided only through a comparison oracle that

Finding the Time to Think: Learning Planning Budgets in Real-Time RL

SafetyDGX agent

arXiv:2606.26463v1 Announce Type: new Abstract: Deliberating takes time. In real-time settings, that time is not free. Standard reinforcement learning (RL) sidesteps this as the environment waits inde

fTNN: a tensor neural network for fractional PDEs

ResearchDGX agent

arXiv:2606.27140v1 Announce Type: new Abstract: We develop the fTNN, a deterministic tensor neural network subspace method for problems involving the fractional Laplacian on bounded domains, taking th

Generating Special Triangulations with Transformers

ResearchDGX agent

arXiv:2606.26660v1 Announce Type: cross Abstract: Triangulations, i.e., well-structured decompositions of geometric objects into triangle-like pieces, are central objects in many domains of mathematic

Generative Models on Analog Hardware with Dynamics

ResearchDGX agent

arXiv:2606.27294v1 Announce Type: cross Abstract: Analog hardware platforms such as coupled oscillators and Analog Ising Machines naturally solve differential equations at a fraction of the energy cos

Gradient Testing and Estimation by Comparisons

ResearchDGX agent

arXiv:2405.11454v3 Announce Type: replace Abstract: We study gradient testing and gradient estimation of smooth functions using only a comparison oracle that, given two points, indicates which one has

Graph Neural Networks Applications Across Domains: All Insights You Need

SafetyDGX agent

arXiv:2606.27202v1 Announce Type: new Abstract: Graph neural networks have moved from a niche representation-learning technique to the default model class wherever data carry relational structure. The

Graph Reinforcement Learning for Calibration-Aware Quantum Circuit Routing

SafetyDGX agent

arXiv:2606.12816v3 Announce Type: replace-cross Abstract: Quantum circuit routing is a key step in compiling programs for noisy intermediate-scale quantum processors. Routes that appear efficient by s

Hierarchical Muon: Tiled Newton-Schulz Updates for Efficient Muon Optimization

HardwareDGX agent

arXiv:2606.27216v1 Announce Type: cross Abstract: Muon-type optimizers construct update directions for dense neural-network weights by applying a finite Newton-Schulz map to momentum-gradient matrices

High-Probability PL-SGD with Markovian Noise: Optimal Mixing and Tail Dependence

SafetyDGX agent

arXiv:2606.26316v1 Announce Type: new Abstract: We study first-order methods for smooth objectives satisfying the Polyak-L{}ojasiewicz (PL) condition when gradient samples are generated by an exogenou

hisao{}: A GPU-Native Parallel Optimizer for Multimodal Black-Box Functions via Convergence-Anticonvergence Oscillation

Model ReleasesDGX agent

arXiv:2606.26164v1 Announce Type: new Abstract: Finding all modes of a multimodal black-box function is a fundamental challenge in optimization, Bayesian inference, and scientific computing. Existing

HOB: A Holistically Optimized Bidding Strategy under Heterogeneous Bidding Environments

Model ReleasesDGX agent

arXiv:2510.15238v2 Announce Type: replace-cross Abstract: Optimizing a single advertising campaign across heterogeneous channels is a central challenge in industrial autobidding. Auction mechanisms va

How Good Can Linear Models Be for Time-Series Forecasting?

ResearchDGX agent

arXiv:2606.27282v1 Announce Type: new Abstract: Time-series forecasting research has been moving steadily toward larger architectures, from specialized transformers to general-purpose foundation model

Huracan: A skillful end-to-end data-driven system for ensemble data assimilation and weather prediction

ResearchDGX agent

arXiv:2508.18486v2 Announce Type: replace-cross Abstract: Over the past few years, machine learning-based data-driven weather prediction has been transforming operational weather forecasting by provid

Implementation of reinforcement learning in chemical reaction networks: application to phototaxis as curiosity-driven exploration

Model ReleasesDGX agent

arXiv:2606.26168v1 Announce Type: new Abstract: Living systems navigate environments using noisy and incomplete sensory signals. In unicellular algae, phototaxis is often modeled as a mechanistic run-

Interpreting 'Interpretability' and Explaining 'Explainability' in Machine Learning in Physics

ResearchDGX agent

arXiv:2606.26228v1 Announce Type: cross Abstract: We review the concepts of interpretability and explainability as they apply to machine learning in physics. We define interpretability as concerning t

Kolmogorov Arnold networks (KAN) for aerodynamic prediction: a comparison with MLPs and GNNs

ResearchDGX agent

arXiv:2606.27126v1 Announce Type: new Abstract: Kolmogorov Arnold networks (KAN) have recently been introduced as a (deep) neural network architecture whose trainable parameters adapt the activation f

Latent Diffusion Posterior Sampling with Surrogate Likelihood Guidance for PDE Inverse Problems

Model ReleasesDGX agent

arXiv:2606.26592v1 Announce Type: cross Abstract: We propose latent-space diffusion posterior sampling (L-DPS), an approximate Bayesian framework for high-dimensional inverse problems governed by part

Learning from a Biased Sample

SafetyDGX agent

arXiv:2209.01754v5 Announce Type: replace-cross Abstract: The empirical risk minimization approach to data-driven decision making requires access to training data drawn under the same conditions as th

Learning from Equivalence Queries, Revisited

ResearchDGX agent

arXiv:2604.04535v2 Announce Type: replace Abstract: Modern machine learning systems, such as generative models and recommendation systems, often evolve through a cycle of deployment, user interaction,

Learning Long-Range Dependencies with Temporal Predictive Coding

Model ReleasesDGX agent

arXiv:2602.18131v2 Announce Type: replace Abstract: Temporal Predictive Coding provides a layer-local, parallelisable mechanism for learning in recurrent systems, making it an attractive candidate for

Learning Probabilistic Filters with Strictly Proper Scoring Rules

SafetyDGX agent

arXiv:2606.26497v1 Announce Type: new Abstract: Bayesian filtering of partially and noisily observed dynamical systems seeks to infer the evolving conditional distribution of the state of a dynamical

Learning Robust Penetration Testing Policies under Partial Observability: A systematic evaluation

SafetyDGX agent

arXiv:2509.20008v2 Announce Type: replace Abstract: Penetration testing, the simulation of cyberattacks to identify security vulnerabilities, presents a sequential decision-making problem well-suited

Learning to Explain Air Traffic Situation

AgentsDGX agent

arXiv:2502.10764v4 Announce Type: replace Abstract: Understanding how air traffic controllers construct a mental 'picture' of complex air traffic situations is crucial but remains a challenge due to t

Listening Like a Judge: A Music-Aware Framework for Automatic Singing Performance Evaluation

SafetyDGX agent

arXiv:2606.26451v1 Announce Type: cross Abstract: Automatic singing quality assessment (SQA) requires evaluating lyrical correctness and musical fidelity while handling expressive variations. However,

Mean-Field PhiBE: Continuous-Time Mean-Field Reinforcement Learning from Discrete-Time Data

Model ReleasesDGX agent

arXiv:2606.26498v1 Announce Type: cross Abstract: This paper addresses model-free continuous-time mean-field control in a setting where the population dynamics evolve continuously according to an unkn

Mesh-RL: Coupled subgrid reinforcement learning

Local AiDGX agent

arXiv:2606.26333v1 Announce Type: new Abstract: Reinforcement learning in large or sparse-reward environments suffers from slow temporal-difference reward propagation, as value information spreads onl

MetaboNet-Bench: A Multi-modal Benchmark for Glucose Forecasting in Type 1 Diabetes

Model ReleasesDGX agent

arXiv:2606.18640v2 Announce Type: replace Abstract: Glucose forecasting algorithms are an important aspect of glycemic control management in type 1 diabetes. So far, the research community has develop

NASimJax: A GPU-Accelerated Policy Learning Framework for Penetration Testing

SafetyDGX agent

arXiv:2603.19864v2 Announce Type: replace Abstract: Penetration testing, the practice of simulating cyberattacks to identify vulnerabilities, is a complex sequential decision-making task that is inher

Necessary but Not Sufficient: Temperature Control and Reproducibility in LLM-as-Judge Safety Evaluations

Model ReleasesDGX agent

arXiv:2606.26185v1 Announce Type: new Abstract: LLM-as-judge ('grader') components are now standard in evaluation harnesses, including safety evaluations where a pass/fail verdict may gate downstream

NervePool: A Simplicial Pooling Layer

ResearchDGX agent

arXiv:2305.06315v3 Announce Type: replace-cross Abstract: For deep learning problems on graph-structured data, pooling layers are important for down sampling, reducing computational cost, and to minim

No Free Lunch: Non-Asymptotic Analysis of Prediction-Powered Inference

ApplicationsDGX agent

arXiv:2505.20178v2 Announce Type: replace-cross Abstract: Prediction-Powered Inference (PPI) is a popular strategy for combining gold-standard and possibly noisy pseudo-labels to perform statistical e

Normalizing Flows are Capable Models for Continuous Control

SafetyDGX agent

arXiv:2505.23527v4 Announce Type: replace Abstract: Modern reinforcement learning (RL) algorithms have found success by using powerful probabilistic models, such as transformers, energy-based models,

Optimizing CUDA like a Human: Micro-Profiling Tools as Expert Surrogates for LLM-Based GPU Kernel Optimization

HardwareDGX agent

arXiv:2606.26453v1 Announce Type: new Abstract: We present KernelPro, a closed-loop multi-agent system that automatically generates, profiles, and iteratively optimizes GPU kernel code by integrating

Otter Weather: Skillful and Computationally Efficient Medium-Range Weather Forecasting

HardwareDGX agent

arXiv:2606.26421v1 Announce Type: new Abstract: State-of-the-art medium-range AI weather models can outperform traditional Numerical Weather Prediction (NWP) but require massive training budgets. This

Over-parameterization and Adversarial Robustness in Neural Networks: An Overview and Empirical Analysis

Model ReleasesDGX agent

arXiv:2406.10090v3 Announce Type: replace Abstract: Thanks to their extensive capacity, over-parameterized neural networks exhibit superior predictive capabilities and generalization. However, having

PersistentKV: Page-Aware Decode Scheduling for Long-Context LLM Serving on Commodity GPUs

Model ReleasesDGX agent

arXiv:2606.26666v1 Announce Type: new Abstract: Autoregressive large language model (LLM) serving is increasingly limited by key-value (KV) cache movement rather than dense matrix multiplication. Mode

Physics-guided Convolutional Neural Network for Domain Growth Prediction in Systems with Conserved Kinetics

TutorialsDGX agent

arXiv:2606.26128v1 Announce Type: new Abstract: The spatiotemporal evolution of many physical, chemical, and biological systems is described by nonlinear partial differential equations (PDEs). Recentl

Quantization in Federated Learning: Methods, Challenges and Future Directions

Local AiDGX agent

arXiv:2606.26822v1 Announce Type: new Abstract: Federated Learning (FL) has become a foundational paradigm for privacy-preserving distributed intelligence, yet its scalability remains fundamentally co

Reasoning Quality Emerges Early: Data Curation for Reasoning Models

ResearchDGX agent

arXiv:2606.26797v1 Announce Type: new Abstract: Supervised fine-tuning (SFT) on a small, high-quality set of long reasoning traces is an effective approach for eliciting strong reasoning capabilities

RecallRisk-BERT: A Multi-Task Framework for Post-Report Medical Device Recall Triage

SafetyDGX agent

arXiv:2606.27174v1 Announce Type: new Abstract: Medical device recalls are a critical regulatory mechanism for protecting patient safety. The growing volume of FDA recall records presents challenges i

Recovering Governing Equations from Solution Data: Identifiability Bounds for Linear and Nonlinear ODEs

ResearchDGX agent

arXiv:2606.27285v1 Announce Type: new Abstract: Learning governing equations from observed solution data is a fundamental challenge in scientific machine learning ite{bruntonDiscoveringGoverningEquati

Reinforcement Fine-Tuning of Flow-Matching Policies for Vision-Language-Action Models

Model ReleasesDGX agent

arXiv:2510.09976v2 Announce Type: replace Abstract: Vision-Language-Action (VLA) models such as OpenVLA, Octo, and pi_0 have shown strong generalization by leveraging large-scale demonstrations, yet t

← Previous
1…6566676869…241
Next →