AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,557 results
21 May 2026

Group-Algebraic Tensors: Provably-optimal Equivariant Learning and Physical Symmetry Discovery

Model ReleasesDGX agent

arXiv:2605.20440v1 Announce Type: new Abstract: We introduce the star_G tensor algebra, in which any finite group G defines the multiplication rule, making equivariance an intrinsic algebraic property

Group-Aware Matrix Estimation and Latent Subspace Recovery

ResearchDGX agent

arXiv:2605.20559v1 Announce Type: cross Abstract: Modern matrix completion problems often involve heterogeneous data whose rows simultaneously belong to many meta-categories, such as demographic and a

GROW: Aligning GRPO with State-Action Modeling for Open-World VLM Agents

SafetyDGX agent

arXiv:2605.20246v1 Announce Type: new Abstract: Recently, vision-language model (VLM) agents have shown promising progress in open-world tasks, where successful task completion often requires multiple


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Hack-Verifiable Environments: Towards Evaluating Reward Hacking at Scale

Model ReleasesDGX agent

arXiv:2605.20744v1 Announce Type: new Abstract: Aligning autonomous agents with human intent remains a central challenge in modern AI. A key manifestation of this challenge is reward hacking, whereby

HiRes: Inspectable Precedent Memory for Reaction Condition Recommendation

ResearchDGX agent

arXiv:2605.21420v1 Announce Type: new Abstract: Reaction condition recommendation sits immediately after retrosynthetic disconnection selection, and in practice, chemists require both accurate predict

HORST: Composing Optimizer Geometries for Sparse Transformer Training

SafetyDGX agent

arXiv:2605.21104v1 Announce Type: new Abstract: Sparsifying transformers remains a fundamental challenge, as standard optimizers fail to simultaneously encourage sparsity and maintain training stabili

How Many Human Survey Respondents is a Large Language Model Worth? An Uncertainty Quantification Perspective

ResearchDGX agent

arXiv:2502.17773v5 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly used to simulate survey responses, but synthetic data can be misaligned with the human populatio

How Much Online RL is Enough? Informative Rollouts for Offline Preference Optimization in RLVR

Model ReleasesDGX agent

arXiv:2605.21266v1 Announce Type: new Abstract: Reinforcement Learning from Verifiable Rewards (RLVR) has emerged as a powerful paradigm for reasoning in language models, with GRPO as its primary exam

Improved convergence rate of kNN graph Laplacians: differentiable self-tuned affinity

SafetyDGX agent

arXiv:2410.23212v2 Announce Type: replace-cross Abstract: In graph-based data analysis, k-nearest neighbor (kNN) graphs are widely used due to their adaptivity to local data densities. Allowing weight

Improved Guarantees for Constrained Online Convex Optimization via Self-Contraction

ResearchDGX agent

arXiv:2605.21107v1 Announce Type: new Abstract: We consider Constrained Online Convex Optimization (COCO) with adversarially chosen constraints. At each round, the learner chooses an action before obs

Inference Time Policy Optimization for Offline RL with Differentiable World Models

SafetyDGX agent

arXiv:2603.22430v2 Announce Type: replace Abstract: Offline Reinforcement Learning (RL) learns optimal policies from fixed datasets, training a policy once and deploying it at inference time without f

Informationally Compressive Anonymization: Non-Degrading Sensitive Input Protection for Privacy-Preserving Supervised Machine Learning

ApplicationsDGX agent

arXiv:2603.15842v2 Announce Type: replace Abstract: Modern machine learning systems increasingly rely on sensitive data, creating significant privacy, security, and regulatory risks that existing priv

Insights Generator: Systematic Corpus-Level Trace Diagnostics for LLM Agents

AgentsDGX agent

arXiv:2605.21347v1 Announce Type: cross Abstract: Diagnosing failures in LLM agents remains largely manual. Practitioners inspect a small subset of execution traces, form ad-hoc hypotheses, and iterat

Instance Discrimination for Link Prediction

ResearchDGX agent

arXiv:2605.20257v1 Announce Type: new Abstract: Recently, instance discrimination models have emerged as a major solution for self-supervised learning. Having already demonstrated its effectiveness in

Instant GPU Efficiency Visibility at Fleet Scale

HardwareDGX agent

arXiv:2605.20799v1 Announce Type: cross Abstract: We present Overall FLOP Utilization (OFU), a hardware-level, precision-agnostic GPU efficiency metric for AI workloads on HPC systems, derived from tw

Interaction Locality in Hierarchical Recursive Reasoning

Local AiDGX agent

arXiv:2605.20784v1 Announce Type: cross Abstract: Spatial reasoning requires both location-bound computation and location-invariant structure: agents must make local moves while preserving route, obje

Introspective X Training: Feedback Conditioning Improves Scaling Across all LLM Training Stages

TutorialsDGX agent

arXiv:2605.20285v1 Announce Type: new Abstract: We tackle the question of how to scale more efficiently across the many, ever-growing stages of current LLM training pipelines. Our guiding intuition st

Is Fixing Schema Graphs Necessary? Full-Resolution Graph Structure Learning for Relational Deep Learning

ApplicationsDGX agent

arXiv:2605.21475v1 Announce Type: new Abstract: Relational prediction tasks are fundamental in many real-world applications, where data are naturally stored in relational databases (RDBs). Relational

It Takes Two: Complementary Self-Distillation for Contextual Integrity in LLMs

SafetyDGX agent

arXiv:2605.20258v1 Announce Type: new Abstract: Contextual Integrity (CI) defines privacy not merely as keeping information hidden, but as governing information flows according to the norms of a given

Large-Step Training Dynamics of a Two-Factor Linear Transformer Model

Model ReleasesDGX agent

arXiv:2605.21292v1 Announce Type: cross Abstract: Gradient-flow analyses show that simplified linear transformers can learn the in-context linear-regression algorithm, but they do not explain the fini

Latent Geometry as a Structural Monitor: Eigenspace Alignment for Anomaly Detection in Anonymity Networks

SafetyDGX agent

arXiv:2605.20391v1 Announce Type: cross Abstract: Traditional anomaly detection marks events when measured signals cross predefined thresholds. This captures the moment of transition but not the struc

Latent Process Generator Matching

TutorialsDGX agent

arXiv:2605.20547v1 Announce Type: new Abstract: Many recent flow-matching and diffusion-style generative models rely on auxiliary stochastic dynamics during training: a richer process is simulated to

LEAP: A closed-loop framework for perovskite precursor additive discovery

Model ReleasesDGX agent

arXiv:2605.20242v1 Announce Type: new Abstract: Efficient discovery of precursor additives is essential for improving the performance of perovskite solar cells, yet the large chemical space makes conv

Learning Dynamics from Infrequent Output Measurements for Uncertainty-Aware Optimal Control

ApplicationsDGX agent

arXiv:2512.08013v2 Announce Type: replace-cross Abstract: Reliable optimal control is challenging when the dynamics of a nonlinear system are unknown and only infrequent, noisy output measurements are

Learning First Integrals via Backward-Generated Data and Guided Reinforcement Learning

ResearchDGX agent

arXiv:2605.21160v1 Announce Type: new Abstract: The discovery of first integrals is of fundamental scientific importance for understanding conservation laws in dynamical systems. However, existing sym

Learning fMRI activations dictionaries across individual geometries via optimal transport

Model ReleasesDGX agent

arXiv:2605.20883v1 Announce Type: new Abstract: Dictionary learning is a powerful tool for creating interpretable representations. When applied to functional magnetic resonance imaging (fMRI) data, th

Learning Incentive Structures for Cooperative Resilience in Multi-Agent Systems under Social Dilemmas

AgentsDGX agent

arXiv:2601.22292v2 Announce Type: replace-cross Abstract: Multi-agent social dilemmas, such as the tragedy of the commons, capture settings where individual incentives conflict with collective well-be

Learning to Defer in Non-Stationary Time Series via Switching State-Space Models

TutorialsDGX agent

arXiv:2601.22538v2 Announce Type: replace Abstract: Learning-to-defer (L2D) routes each decision to a system's own predictor or to an external expert. Streaming time-series settings break the offline-

Learning-to-Defer with Expert-Conditional Advice

Model ReleasesDGX agent

arXiv:2603.14324v3 Announce Type: replace-cross Abstract: Learning-to-Defer routes each input to the expert that minimizes expected cost, but it assumes that the information available to every expert

Less Data, Faster Training: repeating smaller datasets speeds up learning via sampling biases

ResearchDGX agent

arXiv:2605.20314v1 Announce Type: new Abstract: This work investigates the ``small-vs-large gap'', where repeating on fewer samples can lead to compute saving during training compared to using a large

Listwise Policy Optimization: Group-based RLVR as Target-Projection on the LLM Response Simplex

SafetyDGX agent

arXiv:2605.06139v2 Announce Type: replace Abstract: Reinforcement learning with verifiable rewards (RLVR) has become a standard approach for large language models (LLMs) post-training to incentivize r

Llamas on the Web: Memory-Efficient, Performance-Portable, and Multi-Precision LLM Inference with WebGPU

Model ReleasesDGX agent

arXiv:2605.20706v1 Announce Type: cross Abstract: Running language models in the browser presents a unique opportunity to build efficient, private, and portable AI applications, but requires contendin

LLM Pretraining Shapes a Generalizable Manifold: Insights into Cross-Modal Transfer to Time Series

SafetyDGX agent

arXiv:2605.20449v1 Announce Type: new Abstract: Can language-pretrained transformers become effective time-series forecasters, and why? In this paper, we show that cross-modal transfer arises because

LOSCAR-SGD: Local SGD with Communication-Computation Overlap and Delay-Corrected Sparse Model Averaging

Local AiDGX agent

arXiv:2605.20866v1 Announce Type: new Abstract: Communication is a major bottleneck in distributed learning, especially in large-scale settings and in federated learning environments with slow links.

LT2: Linear-Time Looped Transformers

TutorialsDGX agent

arXiv:2605.20670v1 Announce Type: new Abstract: Looped Transformers (LT) have emerged as a powerful architecture by iterating their layers multiple times before decoding the final token. However, pair

Machine-Learned Force Fields for Lattice Dynamics at Coupled-Cluster Level Accuracy

ResearchDGX agent

arXiv:2507.06929v2 Announce Type: replace-cross Abstract: We investigate Machine-Learned Force Fields (MLFFs) trained on approximate Density Functional Theory (DFT) and Coupled Cluster (CC) level pote

Machine-Learning-Enhanced Non-Invasive Testing for MASLD Fibrosis: Shallow-Deep Neural Networks Versus FIB-4, Tabular Foundation Models, and Large Language Models

ResearchDGX agent

arXiv:2605.20523v1 Announce Type: new Abstract: Advanced fibrosis is a major determinant of liver-related morbidity in metabolic dysfunction-associated steatotic liver disease (MASLD). FIB-4 is widely

MagBridge-Battery: A Synthetic Bridge Dataset for Li-ion Magnetometry and State-of-Health Diagnostics

Model ReleasesDGX agent

arXiv:2605.20240v1 Announce Type: new Abstract: Battery health diagnostics today rely overwhelmingly on electrochemical signals measured at the cell terminals. A parallel literature has shown that mag

Mahjax: A GPU-Accelerated Mahjong Simulator for Reinforcement Learning in JAX

SafetyDGX agent

arXiv:2605.20577v1 Announce Type: cross Abstract: Riichi Mahjong is a multi-player, imperfect-information game characterized by stochasticity and high-dimensional state spaces. These attributes presen

Markovian Circuit Tracing for Transformer State Dynamic

Model ReleasesDGX agent

arXiv:2605.20824v1 Announce Type: new Abstract: Many sequence computations are easier to study as movement through internal states than as isolated local circuits. We introduce Markovian Circuit Traci

Matryoshka Concept Bottleneck Models

ResearchDGX agent

arXiv:2605.20612v1 Announce Type: new Abstract: Concept Bottleneck Models (CBMs) have emerged as a prominent paradigm for interpretable deep learning, learning by grounding predictions in human-unders

Maxitive Donsker-Varadhan Formulation for Possibilistic Variational Inference

ResearchDGX agent

arXiv:2511.21223v2 Announce Type: replace-cross Abstract: Variational inference (VI) is a cornerstone of modern Bayesian learning, enabling approximate inference in complex models. However, its formul

Mechanisms of Misgeneralization in Physical Sequence Modeling

Local AiDGX agent

arXiv:2605.20299v1 Announce Type: new Abstract: Generative sequence models are often trained to plan motion in physical domains, from robotics to mechanical simulations. When constructing a dataset to

Memorisation, convergence and generalisation in generative models

TutorialsDGX agent

arXiv:2605.21402v1 Announce Type: cross Abstract: Generative neural networks learn how to produce highly realistic images from a large, but finite number of examples - or do they simply memorise their

Memory-Efficient Partitioned DNN Inference on Resource-Constrained Android Crowds

Model ReleasesDGX agent

arXiv:2605.20723v1 Announce Type: new Abstract: Deploying large deep neural networks on memory-constrained mobile devices is a central challenge in edge ML. While compression, pruning, and quantizatio

Mercer Large-Scale Kernel Machines from Ridge Function Perspective

ResearchDGX agent

arXiv:2307.11925v3 Announce Type: replace Abstract: To present Mercer large-scale kernel machines from a ridge function perspective, we recall the results by Lin and Pinkus from {it Fundamentality of

Miller-Index-Based Latent Crystallographic Fracture Plane Reasoning with Vision-Language Models

ApplicationsDGX agent

arXiv:2605.20416v1 Announce Type: new Abstract: We study whether multimodal large language models (MLLMs) can leverage crystallographic plane indices (Miller indices) as a structured latent representa

Mind the Sim-to-Real Gap & Think Like a Scientist

SafetyDGX agent

arXiv:2605.21458v1 Announce Type: cross Abstract: Suppose a planner has a pre-trained simulator of a sequential decision problem and the option to run real experiments in the field. The simulator is c

Mitigating Label Bias with Interpretable Rubric Embeddings

SafetyDGX agent

arXiv:2605.21455v1 Announce Type: new Abstract: Statistical decision algorithms are increasingly deployed in domains where ground-truth labels are hard to obtain, such as hiring, university admissions

Modality-Decoupled Online Recursive Editing

Local AiDGX agent

arXiv:2605.20273v1 Announce Type: new Abstract: Online model editing for multimodal large language models (MLLMs) requires assimilating a stream of corrections under tight compute and memory budgets.

Modeling Temporal scRNA-seq Data with Latent Gaussian Process and Optimal Transport

ResearchDGX agent

arXiv:2605.20989v1 Announce Type: new Abstract: Single-cell RNA sequencing provides insights into gene expression at single-cell resolution, yet inferring temporal processes from these static snapshot

Modular Multimodal Classification Without Fine-Tuning: A Simple Compositional Approach

ResearchDGX agent

arXiv:2605.20674v1 Announce Type: new Abstract: We introduce CoMET, extit{extbf{C}omposing extbf{M}odality extbf{E}ncoders with extbf{T}abular foundation models}, a simple yet highly competitive metho

Motion-Robust Deep Reconstruction for Free-Breathing Cardiac Cine MRI

ResearchDGX agent

arXiv:2605.20687v1 Announce Type: cross Abstract: Conventional cardiac cine MRI relies on breath-hold Cartesian acquisitions, which are vulnerable to motion artifacts and can be uncomfortable or infea

Multi-Agent Reinforcement Learning for Safe Autonomous Driving Under Pedestrian Behavioral Uncertainty

SafetyDGX agent

arXiv:2605.20255v1 Announce Type: new Abstract: Simulation-based testing of self-driving cars (SDCs) typically relies on scripted or simplified pedestrian models that do not capture the heterogeneity

Multi-Channel Replay Speech Detection using Acoustic Maps

ResearchDGX agent

arXiv:2602.16399v2 Announce Type: replace-cross Abstract: Replay attacks remain a critical vulnerability for automatic speaker verification systems, particularly in real-time voice assistant applicati

Multi-Head Attention as Ensemble Nadaraya-Watson Estimation: Variance Reduction, Decorrelation, and Optimal Head Diversity

SafetyDGX agent

arXiv:2605.20271v1 Announce Type: cross Abstract: We develop a rigorous statistical theory of multi-head attention (MHA) as an ensemble of Nadaraya-Watson (NW) kernel regression estimators. Building o

Multi-Step Likelihood-Ratio Correction for Reinforcement Learning with Verifiable Rewards

SafetyDGX agent

arXiv:2605.20865v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) plays a pivotal role in improving the reasoning ability of large language models. However, widely

Musical Attention Transformer: Music Generation Using a Music-Specific Attention Model

ResearchDGX agent

arXiv:2605.21081v1 Announce Type: cross Abstract: This study aims to enhance the quality of music generation using Transformers by incorporating meta-information. While Transformer-based approaches ar

NaP-Control: Navigating Diffusion Prior for Versatile and Fast Character Control

SafetyDGX agent

arXiv:2605.20209v1 Announce Type: cross Abstract: Achieving precise, versatile whole-body character control in physics-based animation remains challenging. Recent diffusion-based policies generate ric

NeighborDiv: Training-free Zero-shot Generalist Graph Anomaly Detection via Neighbor Diversity

ResearchDGX agent

arXiv:2605.20879v1 Announce Type: new Abstract: Graph Anomaly Detection (GAD) is increasingly shifting to Generalist GAD (GGAD) for cross-domain 'one-for-all' detection, but existing GGAD methods pred

← Previous
1…139140141142143…243
Next →