AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlog
84,570Total entries
1Added by human
84,569Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,557 results
Model Releases

Group-Algebraic Tensors: Provably-optimal Equivariant Learning and Physical Symmetry Discovery

DGX agent

arXiv:2605.20440v1 Announce Type: new Abstract: We introduce the star_G tensor algebra, in which any finite group G defines the multiplication rule, making equivariance an intrinsic algebraic property

model-releasesarxiv-cs-lg
21 May 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Research

Group-Aware Matrix Estimation and Latent Subspace Recovery

DGX agent

arXiv:2605.20559v1 Announce Type: cross Abstract: Modern matrix completion problems often involve heterogeneous data whose rows simultaneously belong to many meta-categories, such as demographic and a

researcharxiv-cs-lg
21 May 2026
Safety

GROW: Aligning GRPO with State-Action Modeling for Open-World VLM Agents

DGX agent

arXiv:2605.20246v1 Announce Type: new Abstract: Recently, vision-language model (VLM) agents have shown promising progress in open-world tasks, where successful task completion often requires multiple

safetyarxiv-cs-lg
21 May 2026
Model Releases

Hack-Verifiable Environments: Towards Evaluating Reward Hacking at Scale

DGX agent

arXiv:2605.20744v1 Announce Type: new Abstract: Aligning autonomous agents with human intent remains a central challenge in modern AI. A key manifestation of this challenge is reward hacking, whereby

model-releasesarxiv-cs-lg
21 May 2026
Research

HiRes: Inspectable Precedent Memory for Reaction Condition Recommendation

DGX agent

arXiv:2605.21420v1 Announce Type: new Abstract: Reaction condition recommendation sits immediately after retrosynthetic disconnection selection, and in practice, chemists require both accurate predict

researcharxiv-cs-lg
21 May 2026
Safety

HORST: Composing Optimizer Geometries for Sparse Transformer Training

DGX agent

arXiv:2605.21104v1 Announce Type: new Abstract: Sparsifying transformers remains a fundamental challenge, as standard optimizers fail to simultaneously encourage sparsity and maintain training stabili

safetyarxiv-cs-lg
21 May 2026
Research

How Many Human Survey Respondents is a Large Language Model Worth? An Uncertainty Quantification Perspective

DGX agent

arXiv:2502.17773v5 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly used to simulate survey responses, but synthetic data can be misaligned with the human populatio

researcharxiv-cs-lg
21 May 2026
Model Releases

How Much Online RL is Enough? Informative Rollouts for Offline Preference Optimization in RLVR

DGX agent

arXiv:2605.21266v1 Announce Type: new Abstract: Reinforcement Learning from Verifiable Rewards (RLVR) has emerged as a powerful paradigm for reasoning in language models, with GRPO as its primary exam

model-releasesarxiv-cs-lg
21 May 2026
Safety

Improved convergence rate of kNN graph Laplacians: differentiable self-tuned affinity

DGX agent

arXiv:2410.23212v2 Announce Type: replace-cross Abstract: In graph-based data analysis, k-nearest neighbor (kNN) graphs are widely used due to their adaptivity to local data densities. Allowing weight

safetyarxiv-cs-lg
21 May 2026
Research

Improved Guarantees for Constrained Online Convex Optimization via Self-Contraction

DGX agent

arXiv:2605.21107v1 Announce Type: new Abstract: We consider Constrained Online Convex Optimization (COCO) with adversarially chosen constraints. At each round, the learner chooses an action before obs

researcharxiv-cs-lg
21 May 2026
Safety

Inference Time Policy Optimization for Offline RL with Differentiable World Models

DGX agent

arXiv:2603.22430v2 Announce Type: replace Abstract: Offline Reinforcement Learning (RL) learns optimal policies from fixed datasets, training a policy once and deploying it at inference time without f

safetyarxiv-cs-lg
21 May 2026
Applications

Informationally Compressive Anonymization: Non-Degrading Sensitive Input Protection for Privacy-Preserving Supervised Machine Learning

DGX agent

arXiv:2603.15842v2 Announce Type: replace Abstract: Modern machine learning systems increasingly rely on sensitive data, creating significant privacy, security, and regulatory risks that existing priv

applicationsarxiv-cs-lg
21 May 2026
Agents

Insights Generator: Systematic Corpus-Level Trace Diagnostics for LLM Agents

DGX agent

arXiv:2605.21347v1 Announce Type: cross Abstract: Diagnosing failures in LLM agents remains largely manual. Practitioners inspect a small subset of execution traces, form ad-hoc hypotheses, and iterat

agentsarxiv-cs-lg
21 May 2026
Research

Instance Discrimination for Link Prediction

DGX agent

arXiv:2605.20257v1 Announce Type: new Abstract: Recently, instance discrimination models have emerged as a major solution for self-supervised learning. Having already demonstrated its effectiveness in

researcharxiv-cs-lg
21 May 2026
Hardware

Instant GPU Efficiency Visibility at Fleet Scale

DGX agent

arXiv:2605.20799v1 Announce Type: cross Abstract: We present Overall FLOP Utilization (OFU), a hardware-level, precision-agnostic GPU efficiency metric for AI workloads on HPC systems, derived from tw

hardwarearxiv-cs-lg
21 May 2026
Local Ai

Interaction Locality in Hierarchical Recursive Reasoning

DGX agent

arXiv:2605.20784v1 Announce Type: cross Abstract: Spatial reasoning requires both location-bound computation and location-invariant structure: agents must make local moves while preserving route, obje

local-aiarxiv-cs-lg
21 May 2026
Tutorials

Introspective X Training: Feedback Conditioning Improves Scaling Across all LLM Training Stages

DGX agent

arXiv:2605.20285v1 Announce Type: new Abstract: We tackle the question of how to scale more efficiently across the many, ever-growing stages of current LLM training pipelines. Our guiding intuition st

tutorialsarxiv-cs-lg
21 May 2026
Applications

Is Fixing Schema Graphs Necessary? Full-Resolution Graph Structure Learning for Relational Deep Learning

DGX agent

arXiv:2605.21475v1 Announce Type: new Abstract: Relational prediction tasks are fundamental in many real-world applications, where data are naturally stored in relational databases (RDBs). Relational

applicationsarxiv-cs-lg
21 May 2026
Safety

It Takes Two: Complementary Self-Distillation for Contextual Integrity in LLMs

DGX agent

arXiv:2605.20258v1 Announce Type: new Abstract: Contextual Integrity (CI) defines privacy not merely as keeping information hidden, but as governing information flows according to the norms of a given

safetyarxiv-cs-lg
21 May 2026
Model Releases

Large-Step Training Dynamics of a Two-Factor Linear Transformer Model

DGX agent

arXiv:2605.21292v1 Announce Type: cross Abstract: Gradient-flow analyses show that simplified linear transformers can learn the in-context linear-regression algorithm, but they do not explain the fini

model-releasesarxiv-cs-lg
21 May 2026
Safety

Latent Geometry as a Structural Monitor: Eigenspace Alignment for Anomaly Detection in Anonymity Networks

DGX agent

arXiv:2605.20391v1 Announce Type: cross Abstract: Traditional anomaly detection marks events when measured signals cross predefined thresholds. This captures the moment of transition but not the struc

safetyarxiv-cs-lg
21 May 2026
Tutorials

Latent Process Generator Matching

DGX agent

arXiv:2605.20547v1 Announce Type: new Abstract: Many recent flow-matching and diffusion-style generative models rely on auxiliary stochastic dynamics during training: a richer process is simulated to

tutorialsarxiv-cs-lg
21 May 2026
Model Releases

LEAP: A closed-loop framework for perovskite precursor additive discovery

DGX agent

arXiv:2605.20242v1 Announce Type: new Abstract: Efficient discovery of precursor additives is essential for improving the performance of perovskite solar cells, yet the large chemical space makes conv

model-releasesarxiv-cs-lg
21 May 2026
Applications

Learning Dynamics from Infrequent Output Measurements for Uncertainty-Aware Optimal Control

DGX agent

arXiv:2512.08013v2 Announce Type: replace-cross Abstract: Reliable optimal control is challenging when the dynamics of a nonlinear system are unknown and only infrequent, noisy output measurements are

applicationsarxiv-cs-lg
21 May 2026
Research

Learning First Integrals via Backward-Generated Data and Guided Reinforcement Learning

DGX agent

arXiv:2605.21160v1 Announce Type: new Abstract: The discovery of first integrals is of fundamental scientific importance for understanding conservation laws in dynamical systems. However, existing sym

researcharxiv-cs-lg
21 May 2026
Model Releases

Learning fMRI activations dictionaries across individual geometries via optimal transport

DGX agent

arXiv:2605.20883v1 Announce Type: new Abstract: Dictionary learning is a powerful tool for creating interpretable representations. When applied to functional magnetic resonance imaging (fMRI) data, th

model-releasesarxiv-cs-lg
21 May 2026
Agents

Learning Incentive Structures for Cooperative Resilience in Multi-Agent Systems under Social Dilemmas

DGX agent

arXiv:2601.22292v2 Announce Type: replace-cross Abstract: Multi-agent social dilemmas, such as the tragedy of the commons, capture settings where individual incentives conflict with collective well-be

agentsarxiv-cs-lg
21 May 2026
Tutorials

Learning to Defer in Non-Stationary Time Series via Switching State-Space Models

DGX agent

arXiv:2601.22538v2 Announce Type: replace Abstract: Learning-to-defer (L2D) routes each decision to a system's own predictor or to an external expert. Streaming time-series settings break the offline-

tutorialsarxiv-cs-lg
21 May 2026
Model Releases

Learning-to-Defer with Expert-Conditional Advice

DGX agent

arXiv:2603.14324v3 Announce Type: replace-cross Abstract: Learning-to-Defer routes each input to the expert that minimizes expected cost, but it assumes that the information available to every expert

model-releasesarxiv-cs-lg
21 May 2026
Research

Less Data, Faster Training: repeating smaller datasets speeds up learning via sampling biases

DGX agent

arXiv:2605.20314v1 Announce Type: new Abstract: This work investigates the ``small-vs-large gap'', where repeating on fewer samples can lead to compute saving during training compared to using a large

researcharxiv-cs-lg
21 May 2026
Safety

Listwise Policy Optimization: Group-based RLVR as Target-Projection on the LLM Response Simplex

DGX agent

arXiv:2605.06139v2 Announce Type: replace Abstract: Reinforcement learning with verifiable rewards (RLVR) has become a standard approach for large language models (LLMs) post-training to incentivize r

safetyarxiv-cs-lg
21 May 2026
Model Releases

Llamas on the Web: Memory-Efficient, Performance-Portable, and Multi-Precision LLM Inference with WebGPU

DGX agent

arXiv:2605.20706v1 Announce Type: cross Abstract: Running language models in the browser presents a unique opportunity to build efficient, private, and portable AI applications, but requires contendin

model-releasesarxiv-cs-lg
21 May 2026
Safety

LLM Pretraining Shapes a Generalizable Manifold: Insights into Cross-Modal Transfer to Time Series

DGX agent

arXiv:2605.20449v1 Announce Type: new Abstract: Can language-pretrained transformers become effective time-series forecasters, and why? In this paper, we show that cross-modal transfer arises because

safetyarxiv-cs-lg
21 May 2026
Local Ai

LOSCAR-SGD: Local SGD with Communication-Computation Overlap and Delay-Corrected Sparse Model Averaging

DGX agent

arXiv:2605.20866v1 Announce Type: new Abstract: Communication is a major bottleneck in distributed learning, especially in large-scale settings and in federated learning environments with slow links.

local-aiarxiv-cs-lg
21 May 2026
Tutorials

LT2: Linear-Time Looped Transformers

DGX agent

arXiv:2605.20670v1 Announce Type: new Abstract: Looped Transformers (LT) have emerged as a powerful architecture by iterating their layers multiple times before decoding the final token. However, pair

tutorialsarxiv-cs-lg
21 May 2026
Research

Machine-Learned Force Fields for Lattice Dynamics at Coupled-Cluster Level Accuracy

DGX agent

arXiv:2507.06929v2 Announce Type: replace-cross Abstract: We investigate Machine-Learned Force Fields (MLFFs) trained on approximate Density Functional Theory (DFT) and Coupled Cluster (CC) level pote

researcharxiv-cs-lg
21 May 2026
Research

Machine-Learning-Enhanced Non-Invasive Testing for MASLD Fibrosis: Shallow-Deep Neural Networks Versus FIB-4, Tabular Foundation Models, and Large Language Models

DGX agent

arXiv:2605.20523v1 Announce Type: new Abstract: Advanced fibrosis is a major determinant of liver-related morbidity in metabolic dysfunction-associated steatotic liver disease (MASLD). FIB-4 is widely

researcharxiv-cs-lg
21 May 2026
Model Releases

MagBridge-Battery: A Synthetic Bridge Dataset for Li-ion Magnetometry and State-of-Health Diagnostics

DGX agent

arXiv:2605.20240v1 Announce Type: new Abstract: Battery health diagnostics today rely overwhelmingly on electrochemical signals measured at the cell terminals. A parallel literature has shown that mag

model-releasesarxiv-cs-lg
21 May 2026
Safety

Mahjax: A GPU-Accelerated Mahjong Simulator for Reinforcement Learning in JAX

DGX agent

arXiv:2605.20577v1 Announce Type: cross Abstract: Riichi Mahjong is a multi-player, imperfect-information game characterized by stochasticity and high-dimensional state spaces. These attributes presen

safetyarxiv-cs-lg
21 May 2026
Model Releases

Markovian Circuit Tracing for Transformer State Dynamic

DGX agent

arXiv:2605.20824v1 Announce Type: new Abstract: Many sequence computations are easier to study as movement through internal states than as isolated local circuits. We introduce Markovian Circuit Traci

model-releasesarxiv-cs-lg
21 May 2026
Research

Matryoshka Concept Bottleneck Models

DGX agent

arXiv:2605.20612v1 Announce Type: new Abstract: Concept Bottleneck Models (CBMs) have emerged as a prominent paradigm for interpretable deep learning, learning by grounding predictions in human-unders

researcharxiv-cs-lg
21 May 2026
Research

Maxitive Donsker-Varadhan Formulation for Possibilistic Variational Inference

DGX agent

arXiv:2511.21223v2 Announce Type: replace-cross Abstract: Variational inference (VI) is a cornerstone of modern Bayesian learning, enabling approximate inference in complex models. However, its formul

researcharxiv-cs-lg
21 May 2026
Local Ai

Mechanisms of Misgeneralization in Physical Sequence Modeling

DGX agent

arXiv:2605.20299v1 Announce Type: new Abstract: Generative sequence models are often trained to plan motion in physical domains, from robotics to mechanical simulations. When constructing a dataset to

local-aiarxiv-cs-lg
21 May 2026
Tutorials

Memorisation, convergence and generalisation in generative models

DGX agent

arXiv:2605.21402v1 Announce Type: cross Abstract: Generative neural networks learn how to produce highly realistic images from a large, but finite number of examples - or do they simply memorise their

tutorialsarxiv-cs-lg
21 May 2026
Model Releases

Memory-Efficient Partitioned DNN Inference on Resource-Constrained Android Crowds

DGX agent

arXiv:2605.20723v1 Announce Type: new Abstract: Deploying large deep neural networks on memory-constrained mobile devices is a central challenge in edge ML. While compression, pruning, and quantizatio

model-releasesarxiv-cs-lg
21 May 2026
Research

Mercer Large-Scale Kernel Machines from Ridge Function Perspective

DGX agent

arXiv:2307.11925v3 Announce Type: replace Abstract: To present Mercer large-scale kernel machines from a ridge function perspective, we recall the results by Lin and Pinkus from {it Fundamentality of

researcharxiv-cs-lg
21 May 2026
Applications

Miller-Index-Based Latent Crystallographic Fracture Plane Reasoning with Vision-Language Models

DGX agent

arXiv:2605.20416v1 Announce Type: new Abstract: We study whether multimodal large language models (MLLMs) can leverage crystallographic plane indices (Miller indices) as a structured latent representa

applicationsarxiv-cs-lg
21 May 2026
Safety

Mind the Sim-to-Real Gap & Think Like a Scientist

DGX agent

arXiv:2605.21458v1 Announce Type: cross Abstract: Suppose a planner has a pre-trained simulator of a sequential decision problem and the option to run real experiments in the field. The simulator is c

safetyarxiv-cs-lg
21 May 2026
← Previous
1…174175176177178…304
Next →