AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,557 results
Research

Sobolev Regularized MMD Gradient Flow

DGX agent

arXiv:2605.11884v1 Announce Type: new Abstract: We propose Sobolev-regularized Maximum Mean Discrepancy (SrMMD) gradient flow, a regularized variant of maximum mean discrepancy (MMD) gradient flow bas

researcharxiv-cs-lg
13 May 2026
Research

SoK: Unlearnability and Unlearning for Model Dememorization

DGX agent
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2605.11592v1 Announce Type: new Abstract: Advanced model dememorization methods, including availability poisoning (unlearnability) and machine unlearning, are emerging as key safeguards against

researcharxiv-cs-lg
13 May 2026
Safety

Sparse Offline Reinforcement Learning with Corruption Robustness

DGX agent

arXiv:2512.24768v3 Announce Type: replace-cross Abstract: We investigate robustness to strong data corruption in offline sparse reinforcement learning (RL). In our setting, an adversary may arbitraril

safetyarxiv-cs-lg
13 May 2026
Safety

Sparsity and Out-of-Distribution Generalization

DGX agent

arXiv:2603.07388v2 Announce Type: replace Abstract: Explaining out-of-distribution generalization has been a central problem in epistemology since Goodman's 'grue' puzzle in 1946. Today it's a central

safetyarxiv-cs-lg
13 May 2026
Model Releases

Sparsity-Constraint Optimization via Splicing Iteration

DGX agent

arXiv:2406.12017v2 Announce Type: replace-cross Abstract: Sparsity-constrained optimization underlies many problems in signal processing, statistics, and machine learning. State-of-the-art hard-thresh

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

Spatial Adapter: Structured Spatial Decomposition and Closed-Form Covariance for Frozen Predictors

DGX agent

arXiv:2605.11394v1 Announce Type: cross Abstract: We present the Spatial Adapter, a parameter-efficient post-hoc layer that equips any frozen first-stage predictor with a structured spatial representa

model-releasesarxiv-cs-lg
13 May 2026
Research

Split the Differences, Pool the Rest: Provably Efficient Multi-Objective Imitation

DGX agent

arXiv:2605.12000v1 Announce Type: new Abstract: This work investigates multi-objective imitation learning: the problem of recovering policies that lie on the Pareto front given demonstrations from mul

researcharxiv-cs-lg
13 May 2026
Safety

Spurious Correlation Learning in Preference Optimization: Mechanisms, Consequences, and Mitigation via Tie Training

DGX agent

arXiv:2605.11134v1 Announce Type: new Abstract: Preference learning methods such as Direct Preference Optimization (DPO) are known to induce reliance on spurious correlations, leading to sycophancy an

safetyarxiv-cs-lg
13 May 2026
Research

SRG: Score-based Relaxation-guided Generation for Mixed Integer Linear Programming

DGX agent

arXiv:2603.24033v2 Announce Type: replace Abstract: We propose Score-based Relaxation-guided Generation (SRG), a generative framework based on an approximate formulation of relaxation-guided stochasti

researcharxiv-cs-lg
13 May 2026
Model Releases

STAGE: Tackling Semantic Drift in Multimodal Federated Graph Learning

DGX agent

arXiv:2605.11919v1 Announce Type: new Abstract: Federated graph learning (FGL) enables collaborative training on graph data across multiple clients. As graph data increasingly contain multimodal node

model-releasesarxiv-cs-lg
13 May 2026
Research

Stationary MMD Points

DGX agent

arXiv:2505.20754v3 Announce Type: replace-cross Abstract: Approximation of a target probability distribution using a finite set of points is a problem of fundamental importance in numerical integratio

researcharxiv-cs-lg
13 May 2026
Research

Steerable Neural ODEs on Homogeneous Spaces

DGX agent

arXiv:2605.11133v1 Announce Type: new Abstract: We introduce steerable neural ordinary differential equations on homogeneous spaces M=G/H. These models constitute a novel geometric extension of manifo

researcharxiv-cs-lg
13 May 2026
Agents

Stochastic Minimum-Cost Reach-Avoid Reinforcement Learning

DGX agent

arXiv:2605.11975v1 Announce Type: new Abstract: We study stochastic minimum-cost reach-avoid reinforcement learning, where an agent must satisfy a reach-avoid specification with probability at least p

agentsarxiv-cs-lg
13 May 2026
Applications

STRABLE: Benchmarking Tabular Machine Learning with Strings

DGX agent

arXiv:2605.12292v1 Announce Type: new Abstract: Benchmarking tabular learning has revealed the benefit of dedicated architectures, pushing the state of the art. But real-world tables often contain str

applicationsarxiv-cs-lg
13 May 2026
Tutorials

Strategically Deceptive Model Deployment in Performative Prediction

DGX agent

arXiv:2506.09044v2 Announce Type: replace Abstract: Machine Learning systems are increasingly deployed in decision-making settings that shape user behavior and, in turn, the data on which future decis

tutorialsarxiv-cs-lg
13 May 2026
Local Ai

Structural Interpretations of Protein Language Model Representations via Differentiable Graph Partitioning

DGX agent

arXiv:2605.10985v1 Announce Type: new Abstract: Protein language models such as ESM-2 learn rich residue representations that achieve strong performance on protein function prediction, but their featu

local-aiarxiv-cs-lg
13 May 2026
Model Releases

STRUM: A Spectral Transcription and Rhythm Understanding Model for End-to-End Generation of Playable Rhythm-Game Charts

DGX agent

arXiv:2605.12135v1 Announce Type: cross Abstract: We present STRUM (Spectral Transcription and Rhythm Understanding Model), an audio-to-chart pipeline that converts raw recordings into playable Clone

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

Support-Proximity Augmented Diffusion Estimation for Offline Black-Box Optimization

DGX agent

arXiv:2605.11246v1 Announce Type: new Abstract: Offline black-box optimization aims to discover novel designs with high property scores using only a static dataset, a task fundamentally challenged by

model-releasesarxiv-cs-lg
13 May 2026
Safety

SURGE: Surrogate Gradient Adaptation in Binary Neural Networks

DGX agent

arXiv:2605.10989v1 Announce Type: new Abstract: The training of Binary Neural Networks (BNNs) is fundamentally based on gradient approximation for non-differentiable binarization operations (e.g., sig

safetyarxiv-cs-lg
13 May 2026
Research

SurvBench: A Standardised Preprocessing Pipeline for Multi-Modal Electronic Health Record Survival Analysis

DGX agent

arXiv:2511.11935v2 Announce Type: replace Abstract: Deep-learning survival models for electronic health record (EHR) data are hard to compare across papers because the upstream preprocessing step, whi

researcharxiv-cs-lg
13 May 2026
Research

Tackling Fake Forgetting through Uncertainty Quantification

DGX agent

arXiv:2501.19403v3 Announce Type: replace Abstract: Machine unlearning seeks to remove the influence of specified data from a trained model. While the unlearning accuracy provides a widely used metric

researcharxiv-cs-lg
13 May 2026
Research

Taking the Road Less Scheduled with Adaptive Polyak Steps

DGX agent

arXiv:2511.07767v2 Announce Type: replace Abstract: Schedule-Free SGD, proposed in [Defazio et al., 2024], achieves optimal convergence rates without requiring the training horizon in advance, by repl

researcharxiv-cs-lg
13 May 2026
Model Releases

Targeted Neuron Modulation via Contrastive Pair Search

DGX agent

arXiv:2605.12290v1 Announce Type: new Abstract: Language models are instruction-tuned to refuse harmful requests, but the mechanisms underlying this behavior remain poorly understood. Popular steering

model-releasesarxiv-cs-lg
13 May 2026
Research

Targeted Tests for LLM Reasoning: An Audit-Constrained Protocol

DGX agent

arXiv:2605.11599v1 Announce Type: new Abstract: Fixed reasoning benchmarks evaluate canonical prompts, but semantically valid changes in presentation can still change model behavior. Studies of prompt

researcharxiv-cs-lg
13 May 2026
Safety

Test-Time Personalization: A Diagnostic Framework and Probabilistic Fix for Scaling Failures

DGX agent

arXiv:2605.10991v1 Announce Type: new Abstract: Existing approaches to LLM personalization focus on constructing better personalized models or inputs, while treating inference as a single-shot process

safetyarxiv-cs-lg
13 May 2026
Research

Testing General Relativity Through Gravitational Wave Classification: A Convolutional Neural Network Framework

DGX agent

arXiv:2605.02453v1 Announce Type: cross Abstract: We present a machine learning framework for testing general relativity (GR) with gravitational wave signals from binary black hole mergers. Using the

researcharxiv-cs-lg
13 May 2026
Tutorials

The Confusion is Real: GRAPHIC -- A Network Science Approach to Confusion Matrices in Deep Learning

DGX agent

arXiv:2602.19770v2 Announce Type: replace Abstract: Explainable artificial intelligence has emerged as a promising field of research to address reliability concerns in artificial intelligence. Despite

tutorialsarxiv-cs-lg
13 May 2026
Safety

The Evaluation Differential: When Frontier AI Models Recognise They Are Being Tested

DGX agent

arXiv:2605.11496v1 Announce Type: cross Abstract: Recent published evidence from frontier laboratories shows that contemporary AI models can recognise evaluation contexts, latently represent them, and

safetyarxiv-cs-lg
13 May 2026
Hardware

The Illusion of Power Capping in LLM Decode: A Phase-Aware Energy Characterisation Across Attention Architectures

DGX agent

arXiv:2605.11999v1 Announce Type: cross Abstract: Power capping is the standard GPU energy lever in LLM serving, and it appears to work: throughput drops, power readings fall, and energy budgets are m

hardwarearxiv-cs-lg
13 May 2026
Applications

The Luna Bound Propagator for Formal Analysis of Neural Networks

DGX agent

arXiv:2603.23878v2 Announce Type: replace Abstract: The parameterized CROWN analysis, a.k.a., alpha-CROWN has emerged as a practically successful abstract interpretation method for neural network veri

applicationsarxiv-cs-lg
13 May 2026
Research

The Offline-Frontier Shift: Diagnosing Distributional Limits in Generative Multi-Objective Optimization

DGX agent

arXiv:2602.11126v2 Announce Type: replace Abstract: Offline multi-objective optimization (MOO) aims to recover Pareto-optimal designs given a finite, static dataset. Recent generative approaches, incl

researcharxiv-cs-lg
13 May 2026
Model Releases

The Price of Proportional Representation in Temporal Voting

DGX agent

arXiv:2605.11157v1 Announce Type: cross Abstract: We study proportional representation in the temporal voting model, where collective decisions are made repeatedly over time over a fixed horizon. Prio

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

The Scaling Law of Evaluation Failure: Why Simple Averaging Collapses Under Data Sparsity and Item Difficulty Gaps, and How Item Response Theory Recovers Ground Truth Across Domains

DGX agent

arXiv:2605.11205v1 Announce Type: new Abstract: Benchmark evaluation across AI and safety-critical domains overwhelmingly relies on simple averaging. We demonstrate that this practice produces substan

model-releasesarxiv-cs-lg
13 May 2026
Safety

The tractability landscape of diffusion alignment: regularization, rewards, and computational primitives

DGX agent

arXiv:2605.11361v1 Announce Type: new Abstract: Inference-time reward alignment asks how to turn a pre-trained diffusion model with base law p into a sampler that favors a reward r while remaining clo

safetyarxiv-cs-lg
13 May 2026
Safety

TMRL: Diffusion Timestep-Modulated Pretraining Enables Exploration for Efficient Policy Finetuning

DGX agent

arXiv:2605.12236v1 Announce Type: cross Abstract: Fine-tuning pre-trained robot policies with reinforcement learning (RL) often inherits the bottlenecks introduced by pre-training with behavioral clon

safetyarxiv-cs-lg
13 May 2026
Model Releases

TOPPO: Rethinking PPO for Multi-Task Reinforcement Learning with Critic Balancing

DGX agent

arXiv:2605.11473v1 Announce Type: cross Abstract: Soft Actor-Critic (SAC) and its variants dominate Multi-Task Reinforcement Learning (MTRL) due to their off-policy sample efficiency, while on-policy

model-releasesarxiv-cs-lg
13 May 2026
Applications

Towards Affordable Energy: A Gymnasium Environment for Electric Utility Demand-Response Programs

DGX agent

arXiv:2605.12462v1 Announce Type: cross Abstract: Extreme weather and volatile wholesale electricity markets expose residential consumers to catastrophic financial risks, yet demand response at the di

applicationsarxiv-cs-lg
13 May 2026
Safety

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization

DGX agent

arXiv:2605.11974v1 Announce Type: new Abstract: Large Language Models (LLMs) suffer from order bias, where their performance is affected by the arrangement order of input elements. This unfairness lim

safetyarxiv-cs-lg
13 May 2026
Applications

Towards Uncertainty-Aware Federated Granger Causal Learning

DGX agent

arXiv:2602.13004v2 Announce Type: replace Abstract: Granger causality recovers directed interactions from time-series data, but in many distributed systems, the data are vertically partitioned across

applicationsarxiv-cs-lg
13 May 2026
Research

TRACE: Temporal Routing with Autoregressive Cross-channel Experts for EEG Representation Learning

DGX agent

arXiv:2605.11380v1 Announce Type: new Abstract: Learning transferable representations for electroencephalography (EEG) remains challenging because EEG signals are inherently multi-channel and non-stat

researcharxiv-cs-lg
13 May 2026
Safety

Training Transformers for KV Cache Compressibility

DGX agent

arXiv:2605.05971v2 Announce Type: replace Abstract: Long-context language modeling is increasingly constrained by the Key-Value (KV) cache, whose memory and decode-time access costs scale linearly wit

safetyarxiv-cs-lg
13 May 2026
Model Releases

Trajectory-Agnostic Asteroid Detection in TESS with Deep Learning

DGX agent

arXiv:2605.12391v1 Announce Type: cross Abstract: We present a novel method for extracting moving objects from TESS data using machine learning. Our approach uses two stacked 3D U-Nets with skip conne

model-releasesarxiv-cs-lg
13 May 2026
Safety

Trajectory First: A Curriculum for Discovering Diverse Policies

DGX agent

arXiv:2506.01568v3 Announce Type: replace Abstract: Being able to solve a task in diverse ways makes agents more robust to task variations and less prone to local optima. In this context, constrained

safetyarxiv-cs-lg
13 May 2026
Safety

Transferable Delay-Aware Reinforcement Learning via Implicit Causal Graph Modeling

DGX agent

arXiv:2605.12312v1 Announce Type: new Abstract: Random delays weaken the temporal correspondence between actions and subsequent state feedback, making it difficult for agents to identify the true prop

safetyarxiv-cs-lg
13 May 2026
Safety

Trust Region Inverse Reinforcement Learning: Explicit Dual Ascent using Local Policy Updates

DGX agent

arXiv:2605.11020v1 Announce Type: new Abstract: Inverse reinforcement learning (IRL) is typically formulated as maximizing entropy subject to matching the distribution of expert trajectories. Classica

safetyarxiv-cs-lg
13 May 2026
Safety

Trust the Batch, On- or Off-Policy: Adaptive Policy Optimization for RL Post-Training

DGX agent

arXiv:2605.12380v1 Announce Type: new Abstract: Reinforcement learning is structurally harder than supervised learning because the policy changes the data distribution it learns from. The resulting fr

safetyarxiv-cs-lg
13 May 2026
Model Releases

U-STS-LLM A Unified Spatio-Temporal Steered Large Language Model for Traffic Prediction and Imputation

DGX agent

arXiv:2605.11735v1 Announce Type: new Abstract: The efficient operation of modern cellular networks hinges on the accurate analysis of spatio-temporal traffic data. Mastering these patterns is essenti

model-releasesarxiv-cs-lg
13 May 2026
Safety

Understanding and Preventing Entropy Collapse in RLVR with On-Policy Entropy Flow Optimization

DGX agent

arXiv:2605.11491v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) has become an effective paradigm for improving the reasoning ability of large language models. How

safetyarxiv-cs-lg
13 May 2026
← Previous
1…208209210211212…304
Next →