AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,557 results
13 May 2026

Semi-Supervised Bayesian GANs with Log-Signatures for Uncertainty-Aware Credit Card Fraud Detection

ResearchDGX agent

arXiv:2509.00931v3 Announce Type: replace-cross Abstract: We present a novel deep generative semi-supervised framework for credit card fraud detection, formulated as time series classification task. A

Sequential Off-Policy Learning with Logarithmic Smoothing

SafetyDGX agent

arXiv:2506.10664v2 Announce Type: replace-cross Abstract: Off-policy learning enables training policies from logged interaction data. Most prior work considers the batch setting, where a policy is lea

Shaping Zero-Shot Coordination via State Blocking

AgentsDGX agent

arXiv:2605.11688v1 Announce Type: new Abstract: Zero-shot coordination (ZSC) aims to enable agents to cooperate with independently trained partners without prior interaction, a key requirement for rea


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Shapley Value Approximation Based on k-Additive Games

ResearchDGX agent

arXiv:2502.04763v2 Announce Type: replace-cross Abstract: The Shapley value is the prevalent solution for fair division problems in which a payout is to be divided among multiple agents. By adopting a

ShardTensor: Domain Parallelism for Scientific Machine Learning

ResearchDGX agent

arXiv:2605.11111v1 Announce Type: cross Abstract: Scientific Machine Learning (SciML) faces unique challenges for extreme-resolution data, with mitigations that often fail to scale or degrade the accu

Sharpen Your Flow: Sharpness-Aware Sampling for Flow Matching

TutorialsDGX agent

arXiv:2605.11547v1 Announce Type: new Abstract: Flow matching models generate samples by numerically integrating a learned velocity field, with each integration step requiring a neural network evaluat

Simpson's Paradox in Behavioral Curves: How Aggregation Distorts Parametric Models of User Dynamics

SafetyDGX agent

arXiv:2605.11017v1 Announce Type: new Abstract: Behavioral curve modeling -- fitting parametric functions to engagement-versus-exposure data -- is standard practice in recommendation, advertising, and

Simulation Distillation: Pretraining World Models in Simulation for Rapid Real-World Adaptation

SafetyDGX agent

arXiv:2603.15759v2 Announce Type: replace-cross Abstract: Robot learning requires adaptation methods that improve reliably from limited, mixed-quality interaction data. This is especially challenging

SkillGen: Verified Inference-Time Agent Skill Synthesis

AgentsDGX agent

arXiv:2605.10999v1 Announce Type: new Abstract: Skills are a promising way to improve LLM agent capabilities without retraining, while keeping the added procedure reusable and controllable. However, h

Smoothed Analysis of Learning from Positive Samples

TutorialsDGX agent

arXiv:2504.10428v2 Announce Type: replace-cross Abstract: Binary classification from positive-only samples is a variant of PAC learning where the learner receives i.i.d. positive samples and aims to l

Smoothness Errors in Dynamics Models and How to Avoid Them

TutorialsDGX agent

arXiv:2602.05352v2 Announce Type: replace Abstract: Modern neural networks have shown promise for solving partial differential equations over surfaces, often by discretizing the surface as a mesh and

SOAR: Scale Optimization for Accurate Reconstruction in NVFP4 Quantization

ResearchDGX agent

arXiv:2605.12245v1 Announce Type: new Abstract: NVFP4 has recently emerged as an efficient 4-bit microscaling format for large language models (LLMs), offering superior numerical fidelity with native

Sobolev Regularized MMD Gradient Flow

ResearchDGX agent

arXiv:2605.11884v1 Announce Type: new Abstract: We propose Sobolev-regularized Maximum Mean Discrepancy (SrMMD) gradient flow, a regularized variant of maximum mean discrepancy (MMD) gradient flow bas

SoK: Unlearnability and Unlearning for Model Dememorization

ResearchDGX agent

arXiv:2605.11592v1 Announce Type: new Abstract: Advanced model dememorization methods, including availability poisoning (unlearnability) and machine unlearning, are emerging as key safeguards against

Sparse Offline Reinforcement Learning with Corruption Robustness

SafetyDGX agent

arXiv:2512.24768v3 Announce Type: replace-cross Abstract: We investigate robustness to strong data corruption in offline sparse reinforcement learning (RL). In our setting, an adversary may arbitraril

Sparsity and Out-of-Distribution Generalization

SafetyDGX agent

arXiv:2603.07388v2 Announce Type: replace Abstract: Explaining out-of-distribution generalization has been a central problem in epistemology since Goodman's 'grue' puzzle in 1946. Today it's a central

Sparsity-Constraint Optimization via Splicing Iteration

Model ReleasesDGX agent

arXiv:2406.12017v2 Announce Type: replace-cross Abstract: Sparsity-constrained optimization underlies many problems in signal processing, statistics, and machine learning. State-of-the-art hard-thresh

Spatial Adapter: Structured Spatial Decomposition and Closed-Form Covariance for Frozen Predictors

Model ReleasesDGX agent

arXiv:2605.11394v1 Announce Type: cross Abstract: We present the Spatial Adapter, a parameter-efficient post-hoc layer that equips any frozen first-stage predictor with a structured spatial representa

Split the Differences, Pool the Rest: Provably Efficient Multi-Objective Imitation

ResearchDGX agent

arXiv:2605.12000v1 Announce Type: new Abstract: This work investigates multi-objective imitation learning: the problem of recovering policies that lie on the Pareto front given demonstrations from mul

Spurious Correlation Learning in Preference Optimization: Mechanisms, Consequences, and Mitigation via Tie Training

SafetyDGX agent

arXiv:2605.11134v1 Announce Type: new Abstract: Preference learning methods such as Direct Preference Optimization (DPO) are known to induce reliance on spurious correlations, leading to sycophancy an

SRG: Score-based Relaxation-guided Generation for Mixed Integer Linear Programming

ResearchDGX agent

arXiv:2603.24033v2 Announce Type: replace Abstract: We propose Score-based Relaxation-guided Generation (SRG), a generative framework based on an approximate formulation of relaxation-guided stochasti

STAGE: Tackling Semantic Drift in Multimodal Federated Graph Learning

Model ReleasesDGX agent

arXiv:2605.11919v1 Announce Type: new Abstract: Federated graph learning (FGL) enables collaborative training on graph data across multiple clients. As graph data increasingly contain multimodal node

Stationary MMD Points

ResearchDGX agent

arXiv:2505.20754v3 Announce Type: replace-cross Abstract: Approximation of a target probability distribution using a finite set of points is a problem of fundamental importance in numerical integratio

Steerable Neural ODEs on Homogeneous Spaces

ResearchDGX agent

arXiv:2605.11133v1 Announce Type: new Abstract: We introduce steerable neural ordinary differential equations on homogeneous spaces M=G/H. These models constitute a novel geometric extension of manifo

Stochastic Minimum-Cost Reach-Avoid Reinforcement Learning

AgentsDGX agent

arXiv:2605.11975v1 Announce Type: new Abstract: We study stochastic minimum-cost reach-avoid reinforcement learning, where an agent must satisfy a reach-avoid specification with probability at least p

STRABLE: Benchmarking Tabular Machine Learning with Strings

ApplicationsDGX agent

arXiv:2605.12292v1 Announce Type: new Abstract: Benchmarking tabular learning has revealed the benefit of dedicated architectures, pushing the state of the art. But real-world tables often contain str

Strategically Deceptive Model Deployment in Performative Prediction

TutorialsDGX agent

arXiv:2506.09044v2 Announce Type: replace Abstract: Machine Learning systems are increasingly deployed in decision-making settings that shape user behavior and, in turn, the data on which future decis

Structural Interpretations of Protein Language Model Representations via Differentiable Graph Partitioning

Local AiDGX agent

arXiv:2605.10985v1 Announce Type: new Abstract: Protein language models such as ESM-2 learn rich residue representations that achieve strong performance on protein function prediction, but their featu

STRUM: A Spectral Transcription and Rhythm Understanding Model for End-to-End Generation of Playable Rhythm-Game Charts

Model ReleasesDGX agent

arXiv:2605.12135v1 Announce Type: cross Abstract: We present STRUM (Spectral Transcription and Rhythm Understanding Model), an audio-to-chart pipeline that converts raw recordings into playable Clone

Support-Proximity Augmented Diffusion Estimation for Offline Black-Box Optimization

Model ReleasesDGX agent

arXiv:2605.11246v1 Announce Type: new Abstract: Offline black-box optimization aims to discover novel designs with high property scores using only a static dataset, a task fundamentally challenged by

SURGE: Surrogate Gradient Adaptation in Binary Neural Networks

SafetyDGX agent

arXiv:2605.10989v1 Announce Type: new Abstract: The training of Binary Neural Networks (BNNs) is fundamentally based on gradient approximation for non-differentiable binarization operations (e.g., sig

SurvBench: A Standardised Preprocessing Pipeline for Multi-Modal Electronic Health Record Survival Analysis

ResearchDGX agent

arXiv:2511.11935v2 Announce Type: replace Abstract: Deep-learning survival models for electronic health record (EHR) data are hard to compare across papers because the upstream preprocessing step, whi

Tackling Fake Forgetting through Uncertainty Quantification

ResearchDGX agent

arXiv:2501.19403v3 Announce Type: replace Abstract: Machine unlearning seeks to remove the influence of specified data from a trained model. While the unlearning accuracy provides a widely used metric

Taking the Road Less Scheduled with Adaptive Polyak Steps

ResearchDGX agent

arXiv:2511.07767v2 Announce Type: replace Abstract: Schedule-Free SGD, proposed in [Defazio et al., 2024], achieves optimal convergence rates without requiring the training horizon in advance, by repl

Targeted Neuron Modulation via Contrastive Pair Search

Model ReleasesDGX agent

arXiv:2605.12290v1 Announce Type: new Abstract: Language models are instruction-tuned to refuse harmful requests, but the mechanisms underlying this behavior remain poorly understood. Popular steering

Targeted Tests for LLM Reasoning: An Audit-Constrained Protocol

ResearchDGX agent

arXiv:2605.11599v1 Announce Type: new Abstract: Fixed reasoning benchmarks evaluate canonical prompts, but semantically valid changes in presentation can still change model behavior. Studies of prompt

Test-Time Personalization: A Diagnostic Framework and Probabilistic Fix for Scaling Failures

SafetyDGX agent

arXiv:2605.10991v1 Announce Type: new Abstract: Existing approaches to LLM personalization focus on constructing better personalized models or inputs, while treating inference as a single-shot process

Testing General Relativity Through Gravitational Wave Classification: A Convolutional Neural Network Framework

ResearchDGX agent

arXiv:2605.02453v1 Announce Type: cross Abstract: We present a machine learning framework for testing general relativity (GR) with gravitational wave signals from binary black hole mergers. Using the

The Confusion is Real: GRAPHIC -- A Network Science Approach to Confusion Matrices in Deep Learning

TutorialsDGX agent

arXiv:2602.19770v2 Announce Type: replace Abstract: Explainable artificial intelligence has emerged as a promising field of research to address reliability concerns in artificial intelligence. Despite

The Evaluation Differential: When Frontier AI Models Recognise They Are Being Tested

SafetyDGX agent

arXiv:2605.11496v1 Announce Type: cross Abstract: Recent published evidence from frontier laboratories shows that contemporary AI models can recognise evaluation contexts, latently represent them, and

The Illusion of Power Capping in LLM Decode: A Phase-Aware Energy Characterisation Across Attention Architectures

HardwareDGX agent

arXiv:2605.11999v1 Announce Type: cross Abstract: Power capping is the standard GPU energy lever in LLM serving, and it appears to work: throughput drops, power readings fall, and energy budgets are m

The Luna Bound Propagator for Formal Analysis of Neural Networks

ApplicationsDGX agent

arXiv:2603.23878v2 Announce Type: replace Abstract: The parameterized CROWN analysis, a.k.a., alpha-CROWN has emerged as a practically successful abstract interpretation method for neural network veri

The Offline-Frontier Shift: Diagnosing Distributional Limits in Generative Multi-Objective Optimization

ResearchDGX agent

arXiv:2602.11126v2 Announce Type: replace Abstract: Offline multi-objective optimization (MOO) aims to recover Pareto-optimal designs given a finite, static dataset. Recent generative approaches, incl

The Price of Proportional Representation in Temporal Voting

Model ReleasesDGX agent

arXiv:2605.11157v1 Announce Type: cross Abstract: We study proportional representation in the temporal voting model, where collective decisions are made repeatedly over time over a fixed horizon. Prio

The Scaling Law of Evaluation Failure: Why Simple Averaging Collapses Under Data Sparsity and Item Difficulty Gaps, and How Item Response Theory Recovers Ground Truth Across Domains

Model ReleasesDGX agent

arXiv:2605.11205v1 Announce Type: new Abstract: Benchmark evaluation across AI and safety-critical domains overwhelmingly relies on simple averaging. We demonstrate that this practice produces substan

The tractability landscape of diffusion alignment: regularization, rewards, and computational primitives

SafetyDGX agent

arXiv:2605.11361v1 Announce Type: new Abstract: Inference-time reward alignment asks how to turn a pre-trained diffusion model with base law p into a sampler that favors a reward r while remaining clo

TMRL: Diffusion Timestep-Modulated Pretraining Enables Exploration for Efficient Policy Finetuning

SafetyDGX agent

arXiv:2605.12236v1 Announce Type: cross Abstract: Fine-tuning pre-trained robot policies with reinforcement learning (RL) often inherits the bottlenecks introduced by pre-training with behavioral clon

TOPPO: Rethinking PPO for Multi-Task Reinforcement Learning with Critic Balancing

Model ReleasesDGX agent

arXiv:2605.11473v1 Announce Type: cross Abstract: Soft Actor-Critic (SAC) and its variants dominate Multi-Task Reinforcement Learning (MTRL) due to their off-policy sample efficiency, while on-policy

Towards Affordable Energy: A Gymnasium Environment for Electric Utility Demand-Response Programs

ApplicationsDGX agent

arXiv:2605.12462v1 Announce Type: cross Abstract: Extreme weather and volatile wholesale electricity markets expose residential consumers to catastrophic financial risks, yet demand response at the di

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization

SafetyDGX agent

arXiv:2605.11974v1 Announce Type: new Abstract: Large Language Models (LLMs) suffer from order bias, where their performance is affected by the arrangement order of input elements. This unfairness lim

Towards Uncertainty-Aware Federated Granger Causal Learning

ApplicationsDGX agent

arXiv:2602.13004v2 Announce Type: replace Abstract: Granger causality recovers directed interactions from time-series data, but in many distributed systems, the data are vertically partitioned across

TRACE: Temporal Routing with Autoregressive Cross-channel Experts for EEG Representation Learning

ResearchDGX agent

arXiv:2605.11380v1 Announce Type: new Abstract: Learning transferable representations for electroencephalography (EEG) remains challenging because EEG signals are inherently multi-channel and non-stat

Training Transformers for KV Cache Compressibility

SafetyDGX agent

arXiv:2605.05971v2 Announce Type: replace Abstract: Long-context language modeling is increasingly constrained by the Key-Value (KV) cache, whose memory and decode-time access costs scale linearly wit

Trajectory-Agnostic Asteroid Detection in TESS with Deep Learning

Model ReleasesDGX agent

arXiv:2605.12391v1 Announce Type: cross Abstract: We present a novel method for extracting moving objects from TESS data using machine learning. Our approach uses two stacked 3D U-Nets with skip conne

Trajectory First: A Curriculum for Discovering Diverse Policies

SafetyDGX agent

arXiv:2506.01568v3 Announce Type: replace Abstract: Being able to solve a task in diverse ways makes agents more robust to task variations and less prone to local optima. In this context, constrained

Transferable Delay-Aware Reinforcement Learning via Implicit Causal Graph Modeling

SafetyDGX agent

arXiv:2605.12312v1 Announce Type: new Abstract: Random delays weaken the temporal correspondence between actions and subsequent state feedback, making it difficult for agents to identify the true prop

Trust Region Inverse Reinforcement Learning: Explicit Dual Ascent using Local Policy Updates

SafetyDGX agent

arXiv:2605.11020v1 Announce Type: new Abstract: Inverse reinforcement learning (IRL) is typically formulated as maximizing entropy subject to matching the distribution of expert trajectories. Classica

Trust the Batch, On- or Off-Policy: Adaptive Policy Optimization for RL Post-Training

SafetyDGX agent

arXiv:2605.12380v1 Announce Type: new Abstract: Reinforcement learning is structurally harder than supervised learning because the policy changes the data distribution it learns from. The resulting fr

U-STS-LLM A Unified Spatio-Temporal Steered Large Language Model for Traffic Prediction and Imputation

Model ReleasesDGX agent

arXiv:2605.11735v1 Announce Type: new Abstract: The efficient operation of modern cellular networks hinges on the accurate analysis of spatio-temporal traffic data. Mastering these patterns is essenti

Understanding and Preventing Entropy Collapse in RLVR with On-Policy Entropy Flow Optimization

SafetyDGX agent

arXiv:2605.11491v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) has become an effective paradigm for improving the reasoning ability of large language models. How

← Previous
1…166167168169170…243
Next →