AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,432 results
31 Jul 2026

Foundation-Model Earth Representations Enable Regional-Scale Forest Aboveground Biomass Monitoring Across the Northeastern United States

SafetyDGX agent

arXiv:2607.27217v1 Announce Type: cross Abstract: Forest aboveground biomass (AGB) is a critical indicator of ecosystem productivity and terrestrial carbon storage, yet regional carbon monitoring rema

From Expert Reduction to Behavioral Divergence: Tracing Numerical State through Sparse MoE Inference

Model ReleasesDGX agent

arXiv:2607.28097v1 Announce Type: new Abstract: Mathematically equivalent expert-reduction orders can produce observably different sparse-MoE executions. We isolate this effect in native DeepSeek-V4-F

Fully Inductive Cardinality Estimation

ApplicationsDGX agent

arXiv:2607.28311v1 Announce Type: cross Abstract: Query optimization of Basic Graph Patterns (BGP) SPARQL queries over Knowledge Graphs (KG) requires accurate cardinality estimation. Recently publishe


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

FunL2O: LLM-Guided Feature Function Design for Learning to Optimize

ResearchDGX agent

arXiv:2607.27389v1 Announce Type: new Abstract: Learning-to-optimize (L2O) methods accelerate repeated optimization by training models to predict solutions, warm starts, branching decisions, or other

Generalization and Trade-off in Adversarial Training: An RKHS Perspective via Kernel Integral Operators

Model ReleasesDGX agent

arXiv:2607.27995v1 Announce Type: cross Abstract: Adversarial training has emerged as a powerful approach for protecting models against adversarial attacks in a broad range of real-world applications.

Generalization Bounds on Optimal Control for Transformer Training and Wasserstein Distributional Robustness

ResearchDGX agent

arXiv:2607.27975v1 Announce Type: new Abstract: We derive finite-sample generalization bounds for Transformers trained with dynamic programming recursions. Building on the doubly lifted, measure-value

Good Rankers, Bad Objectives: Bilinear Contrastive Critics under Expressive Policy Search

Model ReleasesDGX agent

arXiv:2607.27422v1 Announce Type: new Abstract: Good action rankings do not make a contrastive critic safe to maximize. These critics increasingly act as value-like objectives for best-of-K selection,

Graph Neural Multilevel Preconditioners for Iterative Solvers

Model ReleasesDGX agent

arXiv:2607.28456v1 Announce Type: cross Abstract: Solving large, sparse linear systems is a core task in scientific computing, and efficient iterative solvers rely critically on effective and robust p

Graph Neural Network Force Fields for Spin Dynamics in Metallic Magnets

Model ReleasesDGX agent

arXiv:2607.28537v1 Announce Type: cross Abstract: Metallic magnets exhibit complex spin dynamics governed by electronically generated interactions. Predictive simulations of such dynamics typically re

Group-Reflective Self-Distillation for Agentic Reinforcement Learning

SafetyDGX agent

arXiv:2607.28076v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) is effective for training large language model agents. However, terminal rewards provide only co

GyRot: Leveraging Hidden Synergy between Rotation and Fine-grained Group Quantization for Low-bit LLM Inference

Model ReleasesDGX agent

arXiv:2607.27694v1 Announce Type: cross Abstract: Low-bit quantization is essential for efficient LLM inference, and both rotation and fine-grained group quantization have shown individual promise. Ho

HARGO: Heterogeneity-Aware Reward-Guided Optimization for RL Post-Training of LLMs on HPC Tasks

Model ReleasesDGX agent

arXiv:2607.28301v1 Announce Type: new Abstract: Supervised fine-tuning (SFT) can equip large language models (LLMs) with domain knowledge for high-performance computing (HPC) tasks such as data race d

Harnessing the Potential of Optimizing Data Mixtures via Bayesian Domain Reweighting

SafetyDGX agent

arXiv:2607.27928v1 Announce Type: new Abstract: The performance of Large Language Models (LLMs) is fundamentally influenced by the distributional composition of multi-domain pre-training data. While m

HealthCAT: An Interpretable Encoder-only Transformer Framework for Health Indicator Prediction and Temporal Interpretation of Wearable Sensor Data

ApplicationsDGX agent

arXiv:2607.27635v1 Announce Type: cross Abstract: Wearable sensors continuously capture fine-grained multivariate time-series data, providing opportunities to model behavioural patterns associated wit

Heterogeneous Ranking in Industrial-Scale Recommender Systems: A Case Study

SafetyDGX agent

arXiv:2607.27577v1 Announce Type: cross Abstract: Heterogeneous recommendation feeds present complex challenges that extend beyond those found in highly homogeneous environments (e.g., music-only or v

Hierarchical Multilevel Monte Carlo for Order-Optimal Neural Actor-Critic in Average-Reward CMDPs

SafetyDGX agent

arXiv:2607.28390v1 Announce Type: new Abstract: Constrained Markov Decision Processes (CMDPs) provide a natural framework for reinforcement learning in safety-critical applications, where agents maxim

HOMER: Huber-of-Means for Efficient and Robust Estimation in Hilbert Spaces

ResearchDGX agent

arXiv:2607.27532v1 Announce Type: cross Abstract: Heavy tails weaken high-confidence control for the empirical mean. Geometric median-of-means (MOM) also lacks a threshold that moves toward mean effic

Improving the Robustness/Accuracy Tradeoff Against Adversarial Attacks Using Information Bottleneck Distillation Through Dual Teachers

ResearchDGX agent

arXiv:2607.27737v1 Announce Type: new Abstract: Deep neural networks (DNNs) have achieved remarkable success in classical machine learning problems. However, they are known to be vulnerable to adversa

Information Bottleneck Learning for Faithful Time Series Forecasting Explanations

ApplicationsDGX agent

arXiv:2607.28124v1 Announce Type: new Abstract: As forecasts increasingly drive decisions in fields such as energy, transportation, and healthcare, understanding the historical data behind these predi

Integrating Contextual Embeddings into Evaluation of Expressive MIDI Piano Performances

SafetyDGX agent

arXiv:2607.27909v1 Announce Type: cross Abstract: Objective evaluation of expressive MIDI piano performances typically relies on attribute statistics such as timing, velocity, and duration of individu

Interpreting learning dynamics of autoencoders: Transient scaling and emerging concepts of the Ising model

TutorialsDGX agent

arXiv:2607.10285v2 Announce Type: replace Abstract: We study how unsupervised autoencoders trained on microscopic spin configurations from the Ising model learn macroscopic, theory-relevant variables

It's All Just Vectorization: einx, a Universal Notation for Tensor Operations

ResearchDGX agent

arXiv:2607.27987v1 Announce Type: new Abstract: Tensor operations represent a cornerstone of modern scientific computing. However, the Numpy-like notation adopted by predominant tensor frameworks is o

KAISEN: Reproducible Subgroup Fairness Auditing for Clinical Risk Models

Model ReleasesDGX agent

arXiv:2607.28608v1 Announce Type: new Abstract: Clinical risk models routinely achieve strong aggregate performance while producing materially different error rates across patient subgroups. Audit pip

Kalman Meets Curriculum: Efficient Dynamic Prompt Selection for Adaptive RL Finetuning

Model ReleasesDGX agent

arXiv:2607.27610v1 Announce Type: new Abstract: Reinforcement learning (RL) finetuning significantly enhances the reasoning capabilities of large language models (LLMs), yet its effectiveness critical

KernelGenBench: A Multi-Source and Multi-Chip Benchmark for LLM-based Kernel Generation

Model ReleasesDGX agent

arXiv:2607.27231v1 Announce Type: cross Abstract: Large language models (LLMs) have significantly increased the demand for efficient accelerator kernels, but kernel development remains a highly specia

Latent-Kernel Discrete Flow Maps for Few-Step Generation

ResearchDGX agent

arXiv:2607.27529v1 Announce Type: new Abstract: Discrete diffusion and flow-matching models denoise a sequence over many steps, but to keep each step cheap, they factorize the transition across positi

Latent Matters: Learning Deep State-Space Models

TutorialsDGX agent

arXiv:2602.23050v2 Announce Type: replace Abstract: Deep state-space models (DSSMs) enable temporal predictions by learning the underlying dynamics of observed sequence data. They are often trained by

Learning-Augmented and Randomized Algorithms for Line Aggregation with Delays

Model ReleasesDGX agent

arXiv:2607.27807v1 Announce Type: new Abstract: This paper studies learning-augmented and randomized online aggregation with delays on a line metric. We consider advice given as online suggested servi

Learning features from Newton's algorithm: a way to accelerate nonlinear parametrized PDE solvers

Model ReleasesDGX agent

arXiv:2607.28036v1 Announce Type: new Abstract: It is well known that Newton's method converges faster when the initial guess is closer to a root of a system of nonlinear equations. In this paper, a t

Learning to Detect Cyber Attacks: Neural Anomaly Detection for Cybersecurity with Theoretical Insights

TutorialsDGX agent

arXiv:2409.08521v2 Announce Type: replace-cross Abstract: In cybersecurity practice, new forms of cyberattacks continuously emerge, deliberately designed to evade defense systems that rely on previous

Learning to Trace Seiberg Dualities

Model ReleasesDGX agent

arXiv:2607.28628v1 Announce Type: cross Abstract: Dualities play an important role in establishing both microscopic and emergent phenomena in a wide range of physical systems. In practice, though, it

LEDGERMIND: Provenance-Constrained Multimodal Agentic Reasoning with a Structured Evidence Ledger

AgentsDGX agent

arXiv:2607.28374v1 Announce Type: new Abstract: Multimodal agents for visual question answering increasingly operate as multi-step trajectories that interleave perception, retrieval, and reasoning, ye

LightRot: A Light-Weighted Rotation Scheme and Architecture for Accurate Low-Bit Large Language Model Inference

Local AiDGX agent

arXiv:2607.27704v1 Announce Type: cross Abstract: As large language models (LLMs) continue to demonstrate exceptional capabilities across various domains, the challenge of achieving energy-efficient a

LLM-Guided Initialization for Accelerated Hybrid Quantum-Classical Medical Image Classification

Model ReleasesDGX agent

arXiv:2607.27262v1 Announce Type: cross Abstract: Variational quantum algorithms often encounter barren plateaus, where cost gradients decay rapidly with increasing circuit depth, undermining the trai

LM-GRASP: Instance-Specific Language Models for Combinatorial Construction via Online Imitation Learning

Model ReleasesDGX agent

arXiv:2607.28135v1 Announce Type: new Abstract: Machine learning for combinatorial optimization typically relies on neural constructors trained via reinforcement learning on large offline datasets for

LoRA Scaffolded Policy Optimization (LSPO): A Sampling-Time Low-Rank Scaffold for Recovering Reinforcement-Learning Gradient on Zero-Reward Cliff Prompts

Model ReleasesDGX agent

arXiv:2607.27787v1 Announce Type: new Abstract: Reinforcement learning from verifiable rewards (RLVR) for mathematical reasoning suffers from a structural blind spot: on 'cliff' prompts-those on which

MatCreatioNN: Machine learning-guided computational discovery of photocatalysts for environmental applications

Model ReleasesDGX agent

arXiv:2607.27295v1 Announce Type: cross Abstract: The rational design of photocatalysts for environmental remediation and CO2 conversion remains limited by the high computational cost and sparse exper

Measuring Distortion in the Empty Regions of Dimensionality Reduction Scatterplots with the Gap Index

ResearchDGX agent

arXiv:2607.28324v1 Announce Type: new Abstract: Quality metrics play a crucial role in the proper use of dimensionality reduction projections for visual analysis of high-dimensional data. They quantif

Memory Efficient Tabular Foundation Models

ResearchDGX agent

arXiv:2607.27546v1 Announce Type: new Abstract: Tabular Foundation Models, such as TabPFN, have received a large amount of recent attention due to their performance on in-context tabular machine learn

Meteosat Third Generation imagery improves CNN-based SSI retrieval

ResearchDGX agent

arXiv:2607.28093v1 Announce Type: cross Abstract: Accurate Surface Solar Irradiance (SSI) estimation is increasingly important for photovoltaic energy monitoring and forecasting. The recently introduc

Modeling Decisions in Blockchain Analytics: A Leakage-Aware Evaluation of Tree-Based vs. Sequential Models

ResearchDGX agent

arXiv:2607.27350v1 Announce Type: new Abstract: Sybil bots are Ethereum actors that imitate legitimate users to extract airdrop rewards or influence governance. Recent Sybil detection methods increasi

More Data, Worse Decisions? Preference Reversals in Neural Networks under Gram Incompatibility

ResearchDGX agent

arXiv:2607.27255v1 Announce Type: cross Abstract: Neural networks increasingly combine data across populations, time periods, and operating conditions to improve generalization. This raises a reliabil

MSGNN: A Spectral Graph Neural Network Based on a Novel Magnetic Signed Laplacian

ApplicationsDGX agent

arXiv:2209.00546v5 Announce Type: replace-cross Abstract: Signed and directed networks are ubiquitous in real-world applications. However, there has been relatively little work proposing spectral grap

MUGEN: A Unified Framework for Efficient Motion Understanding and Generation

SafetyDGX agent

arXiv:2607.27581v1 Announce Type: new Abstract: Grounding human motion in language, and language in motion, is a central step toward physical AI systems that can understand, generate, and communicate

Multi-channel Uplift Policy Learning

SafetyDGX agent

arXiv:2607.28182v1 Announce Type: new Abstract: E-commerce platforms must allocate fixed marketing budgets across multiple channels to maximize business utility. However, standard predict-then-optimiz

Nanoparticle Networks for Neuromorphic Computing

ResearchDGX agent

arXiv:2607.27844v1 Announce Type: cross Abstract: Physical computing leverages complex dynamical systems for energy-efficient data processing. In this work, we present a neuromorphic architecture base

Neural Network Approximation of Solutions to Fractional Parabolic Partial Differential Equations

ResearchDGX agent

arXiv:2607.27781v1 Announce Type: cross Abstract: We establish a dimension-efficient neural network approximation theory for solutions to fractional parabolic equations with lower-order drift and pote

Neural Network-Assisted CLEAN for Channel Modeling in Low-SNR Regimes

Model ReleasesDGX agent

arXiv:2607.27450v1 Announce Type: new Abstract: Accurate multipath parameter estimation is critical for modern wireless communication systems, particularly in challenging low-SNR environments. Traditi

NMINE: Normalized Mutual Information Neural Estimation

ResearchDGX agent

arXiv:2607.27710v1 Announce Type: new Abstract: Mutual information is a general measure of statistical dependence that captures both linear and nonlinear relationships between random variables. For co

Noisy Data is Destructive to Reinforcement Learning with Verifiable Rewards

TutorialsDGX agent

arXiv:2603.16140v2 Announce Type: replace Abstract: Reinforcement learning with verifiable rewards (RLVR) has driven recent capability advances of large language models across various domains. Recent

On a joint simultaneous learning of relevant feature subsets and subspaces in regression-like problems

Model ReleasesDGX agent

arXiv:2607.28080v1 Announce Type: cross Abstract: We extend a recently introduced Entropy-Optimal Manifold Clustering (EOMC) to allow for a joint simultaneous identification of subsets and subspaces o

On-Policy and Off-Policy Learning for Large Action Spaces

SafetyDGX agent

arXiv:2607.28408v1 Announce Type: new Abstract: This thesis studies policy learning in interactive systems where an agent observes a context, selects an action from a very large set, and receives part

On the Rate of Convergence of Kolmogorov-Arnold Network Regression Estimators

ResearchDGX agent

arXiv:2509.19830v3 Announce Type: replace Abstract: Kolmogorov-Arnold Networks (KANs) approximate multivariate functions by composing univariate transformations through additive or multiplicative aggr

OneShot: Index-in-Ranking with Neural Scoring for Large-Scale Retrieval

SafetyDGX agent

arXiv:2607.27475v1 Announce Type: cross Abstract: In modern recommendation systems, retrieval serves as a primary stage responsible for filtering billions of candidate items down to thousands prior to

Optimizing Regret

SafetyDGX agent

arXiv:2607.18866v2 Announce Type: replace-cross Abstract: Building on the identity that expected regret equals the covariance between costs and decisions, this paper develops a derivative theory of th

Oracle-Budgeted Molecular Optimization with Short-Term Graph Memory

Model ReleasesDGX agent

arXiv:2607.28437v1 Announce Type: new Abstract: Molecular optimization is commonly performed under a limited oracle budget, which makes deciding what to evaluate as important as deciding what to gener

Persistent Gaussian Perturbations Prevent Oversmoothing in Recurrent Graph Neural Networks

ResearchDGX agent

arXiv:2607.28185v1 Announce Type: new Abstract: Oversmoothing is a fundamental limitation of deep graph neural networks (GNNs), where repeated message passing causes node representations to become inc

PlantBGC: Transformer for Plant BGC Discovery via Label-Free Domain Adaptation and Weak Supervision

ResearchDGX agent

arXiv:2607.27258v1 Announce Type: cross Abstract: Plant biosynthetic gene clusters (BGCs) encode specialized-metabolite pathways, yet curated plant BGC labels remain scarce, hindering supervised disco

PlatformBid: An Auto-Bidding Benchmark from a Unified Advertising Platform's Perspective

Model ReleasesDGX agent

arXiv:2607.27265v1 Announce Type: new Abstract: Real-time bidding is central to computational advertising, comprising three elements: Supply Side Platform (SSP) selling ad impressions, Demand Side Pla

Policy Gradient Steering: Interventions from Behavioral Objectives

SafetyDGX agent

arXiv:2607.27574v1 Announce Type: new Abstract: Activation steering has emerged in large language models as a lightweight alternative for dynamically changing a model's behavior at inference time. How

← Previous
1…2223242526…241
Next →