AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,557 results
21 May 2026

SMA-DP: Spectral Memory-Aware Differential Privacy for Deep Learning

SafetyDGX agent

arXiv:2605.20450v1 Announce Type: new Abstract: Differentially private stochastic gradient descent (DP-SGD) enables private deep learning through per-example clipping and calibrated Gaussian noise, bu

Smaller Abstract State Spaces Enable Cross-Scale Generalization in Reinforcement Learning

AgentsDGX agent

arXiv:2605.20272v1 Announce Type: new Abstract: While humans readily generalize abstract concepts to more complex or larger tasks, building Reinforcement Learning (RL) systems with this ability remain

SOLAR: A Self-Optimizing Open-Ended Autonomous Agent for Lifelong Learning and Continual Adaptation

Model ReleasesDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2605.20189v1 Announce Type: cross Abstract: Despite the remarkable success of large language models (LLMs), they still face bottlenecks while deploying in dynamic, real-world settings with prima

SOPE: Stabilizing Off-Policy Evaluation for Online RL with Prior Data

Model ReleasesDGX agent

arXiv:2605.05863v2 Announce Type: replace Abstract: Incorporating prior data into online reinforcement learning accelerates training but typically forces a difficult trade-off between high computation

Spectral bandits for smooth graph functions with applications in recommender systems

ApplicationsDGX agent

arXiv:2605.20552v1 Announce Type: cross Abstract: Smooth functions on graphs have wide applications in manifold and semi-supervised learning. In this paper, we study a bandit problem where the payoffs

Spectral Souping: A Unified Framework for Online Preference Alignment

SafetyDGX agent

arXiv:2605.20408v1 Announce Type: new Abstract: Reinforcement Learning from Human Feedback (RLHF) effectively aligns Large Language Models (LLMs) with aggregate human preferences but often fails to ad

Spectral Unforgetting: Post-Hoc Recovery of Damaged Capabilities Without Retraining

Model ReleasesDGX agent

arXiv:2605.20296v1 Announce Type: new Abstract: Fine-tuning a language model for a target task routinely degrades capabilities the training data never explicitly threatened. We study this phenomenon,

Statistical Guarantees in the Search for Less Discriminatory Algorithms

SafetyDGX agent

arXiv:2512.23943v2 Announce Type: replace-cross Abstract: U.S. discrimination law can impose liability on firms that fail to adopt a less discriminatory alternative (LDA): a decision policy that achie

Stimulus symmetries can confound representational similarity analyses

ResearchDGX agent

arXiv:2605.21324v1 Announce Type: cross Abstract: What can representational similarity matrices (RSMs) tell us about a neural code? As the popularity of these summary statistics grows, so too does the

STM3: Mixture of Multiscale Mamba for Long-Term Spatio-Temporal Time-Series Prediction

ApplicationsDGX agent

arXiv:2508.12247v2 Announce Type: replace Abstract: Recently, spatio-temporal time-series prediction has developed rapidly, yet existing deep learning methods struggle with learning complex long-term

Strict Subgoal Execution: Reliable Long-Horizon Planning in Hierarchical Reinforcement Learning

SafetyDGX agent

arXiv:2506.21039v3 Announce Type: replace Abstract: Long-horizon goal-conditioned tasks pose fundamental challenges for reinforcement learning (RL), particularly when goals are distant and rewards are

Supervised Latent Restructuring for Small-Data Quantum Learning in Plant Phenomics

SafetyDGX agent

arXiv:2605.20413v1 Announce Type: new Abstract: High-dimensional biological data often exhibit a severe mismatch between feature dimensionality and sample size, making reliable classification difficul

SURF: Steering the Scalarization Weight to Uniformly Traverse the Pareto Front

SafetyDGX agent

arXiv:2605.20619v1 Announce Type: new Abstract: Scalarization is widely used in multi-objective optimization owing to its simplicity and scalability. In many applications, the goal is to generate solu

Sustainability Is Not Linear: Quantifying Performance, Energy, and Privacy Trade-offs in On-Device Intelligence

Local AiDGX agent

arXiv:2603.26603v2 Announce Type: replace-cross Abstract: The migration of Large Language Models (LLMs) from cloud clusters to edge devices promises enhanced privacy and offline accessibility, but thi

Sutra: Tensor-Op RNNs as a Compilation Target for Vector Symbolic Architectures

ResearchDGX agent

arXiv:2605.20919v1 Announce Type: new Abstract: Sutra is a typed, purely functional programming language whose compiled forward pass is a PyTorch neural network. The compiler beta-reduces the whole pr

Symmetrization of Loss Functions for Robust Training of Neural Networks in the Presence of Noisy Labels

ResearchDGX agent

arXiv:2605.20347v1 Announce Type: new Abstract: Labeling a training set is often expensive and susceptible to errors, making the design of robust loss functions for label noise an important problem. T

TabPFN Extensions for Interpretable Geotechnical Modelling

Model ReleasesDGX agent

arXiv:2603.21033v2 Announce Type: replace-cross Abstract: Geotechnical site characterisation relies on sparse, heterogeneous borehole data, where uncertainty quantification and interpretability matter

TabPFN-MT: A Natively Multitask In-Context Learner for Tabular Data

ResearchDGX agent

arXiv:2605.20234v1 Announce Type: new Abstract: Prior-Data Fitted networks (PFNs) have been very successful in tabular contexts, handling prediction tasks in context. However, they are designed for si

TelecomTS: A Multi-Modal Observability Dataset for Time Series and Language Analysis

ApplicationsDGX agent

arXiv:2510.06063v2 Announce Type: replace-cross Abstract: Modern enterprises generate vast streams of time series metrics when monitoring complex systems, known as observability data. Unlike conventio

Testing Support Size More Efficiently Than Learning Histograms

ResearchDGX agent

arXiv:2410.18915v4 Announce Type: replace-cross Abstract: Consider two problems about an unknown probability distribution p: 1. How many samples from p are required to test if p is supported on n elem

The Devil is in the Condition Numbers: Why is GLU Better than non-GLU Structure?

ResearchDGX agent

arXiv:2605.20749v1 Announce Type: new Abstract: Gated Linear Units (GLU) and their variants are widely adopted in modern open-source large language model architectures and consistently outperform thei

The Economics of AI Inference: Inflation Dynamics, Welfare Costs, and Optimal Monetary Policy under the Inference-Cost Phillips Curve

SafetyDGX agent

arXiv:2605.20281v1 Announce Type: cross Abstract: We develop a unified microeconomic and monetary theory of artificial intelligence inference costs and their pass-through to inflation, welfare, and op

The Economics of Model Collapse: Equilibrium, Welfare, and Optimal Provenance Subsidies in Synthetic Data Markets

Model ReleasesDGX agent

arXiv:2605.20279v1 Announce Type: cross Abstract: Generative artificial intelligence is rapidly transforming the supply side of training data: an increasing share of new tokens, images, and structured

The General Theory of Localization Methods

Local AiDGX agent

arXiv:2605.20635v1 Announce Type: new Abstract: This paper proposes a general machine learning framework called the localization method, which is fundamentally built on two core concepts: localization

Theoretical guidelines for annealed Langevin dynamics in compositional simulation-based inference

Model ReleasesDGX agent

arXiv:2605.21253v1 Announce Type: cross Abstract: Compositional score-based approaches to simulation-based inference (SBI) approximate the posterior over a shared parameter given n independent observa

Time-Dependent PDE-Constrained Optimization via Weak-Form Latent Dynamics

Model ReleasesDGX agent

arXiv:2605.20639v1 Announce Type: cross Abstract: Optimization problems constrained by high-dimensional, time-dependent partial differential equations require repeated forward and sensitivity solves,

Time-Prompt: Integrated Heterogeneous Prompts for Unlocking LLMs in Time Series Forecasting

SafetyDGX agent

arXiv:2506.17631v4 Announce Type: replace Abstract: Time series forecasting aims to model temporal dependencies among variables for future state inference, holding significant importance and widesprea

TimeSRL: Generalizable Time-Series Behavioral Modeling via Semantic RL-Tuned LLMs -- A Case Study in Mental Health

Model ReleasesDGX agent

arXiv:2605.21295v1 Announce Type: new Abstract: Longitudinal passive sensing enables continuous health prediction, yet models often fail under cross-dataset distribution shifts. Traditional ML overfit

torchtune: PyTorch native post-training library

ResearchDGX agent

arXiv:2605.21442v1 Announce Type: new Abstract: Modern LLMs typically require multistage training pipelines to achieve strong downstream performance, with post-training serving as the main interface f

Towards Resilient and Autonomous Networks: A BlueSky Vision on AI-Native 6G

AgentsDGX agent

arXiv:2605.21395v1 Announce Type: cross Abstract: The proliferation of emerging applications, such as autonomous driving and immersive experiences, demands cellular networks that are not only faster,

Towards Understanding Self-Pretraining for Sequence Classification

TutorialsDGX agent

arXiv:2605.21070v1 Announce Type: new Abstract: Amos et al. (2024) showed that the accuracy of Transformer models in sequence classification can be significantly improved by first pretraining with a m

Training distribution determines the ceiling of drug-blind cancer sensitivity prediction

Model ReleasesDGX agent

arXiv:2605.20885v1 Announce Type: new Abstract: Precision oncology requires predicting which drugs will suppress a specific tumor from its molecular profile, but drug-blind sensitivity prediction has

TRAM: Test-Time Risk Adaptation with Mixture of Agents

Model ReleasesDGX agent

arXiv:2408.08812v2 Announce Type: replace Abstract: Deployed reinforcement learning agents often face safety requirements that are specified only after training, such as new hazard maps, revised risk

TreeText-CTS: Compact, Source-Traceable Tree-Path Evidence for Irregular Clinical Time-Series Prediction

ResearchDGX agent

arXiv:2605.20292v1 Announce Type: new Abstract: Numerical time-series models can effectively process irregular electronic health record (EHR) trajectories, but they do not naturally expose the measure

TriForces: Augmenting Atomistic GNNs for Transferable Representations

ResearchDGX agent

arXiv:2605.20581v1 Announce Type: new Abstract: Machine learning interatomic potentials (MLIPs) achieve excellent accuracy when trained on large Density Functional Theory (DFT) data. To be useful in p

Trusted Weights, Treacherous Optimizations? Optimization-Triggered Backdoor Attacks on LLMs

SafetyDGX agent

arXiv:2605.20641v1 Announce Type: cross Abstract: Inference optimization is a vital technique for deploying LLMs at scale. Compilation is the most widely adopted optimization technique for LLMs. While

Tunable MAGMAX: Preference-Aware Model Merging for Continual Learning

Model ReleasesDGX agent

arXiv:2605.20803v1 Announce Type: new Abstract: Continual learning (CL) aims to train models sequentially on multiple tasks while mitigating catastrophic forgetting of previously learned knowledge. Re

Understanding and Improving Communication Performance in Multi-node LLM Inference

Model ReleasesDGX agent

arXiv:2511.09557v4 Announce Type: replace-cross Abstract: As large language models (LLMs) continue to grow in size, distributed inference has become increasingly important. Model-parallel strategies m

Understanding Deterioration Random Effects for Causal Discovery in Infrastructure Management

HardwareDGX agent

arXiv:2605.20400v1 Announce Type: cross Abstract: Infrastructure deterioration poses significant challenges for asset management, yet existing approaches rely on population-averaged models that overlo

Unsupervised clustering and classification of upper limb EMG signals during functional movements: a data-driven

ResearchDGX agent

arXiv:2605.20599v1 Announce Type: new Abstract: This study presents a comprehensive approach for the clustering and classification of upper-limb surface electromyography (sEMG) signals during function

UOTIP: Unbalanced Optimal Transport Map for Unpaired Inverse Problems

ResearchDGX agent

arXiv:2605.21094v1 Announce Type: new Abstract: We investigate unpaired image inverse problems, a challenging setting where only independent, non-paired sets of noisy measurements and clean target sig

Velocityformer: Broken-Symmetry-Matched Equivariant Graph Transformers for Cosmological Velocity Reconstruction

SafetyDGX agent

arXiv:2605.21483v1 Announce Type: cross Abstract: Precise measurement of the kinematic Sunyaev-Zel'dovich (kSZ) effect - a probe of the large-scale distribution of baryonic matter, a key observable fo

Verifiable Error Bounds for Physics-Informed Neural Network Solutions of Lyapunov and Hamilton-Jacobi-Bellman Equations

SafetyDGX agent

arXiv:2603.19545v2 Announce Type: replace-cross Abstract: Many core problems in nonlinear systems analysis and control can be recast as solving partial differential equations (PDEs) such as Lyapunov a

Verification of Unknown Dynamical Systems via Autoencoder Latent Space

TutorialsDGX agent

arXiv:2512.13593v4 Announce Type: replace Abstract: Formal verification provides a powerful framework for proving that dynamical systems satisfy their specifications. However, these techniques face sc

WaveGraphNet: Physics-Consistent Guided-Wave Damage Localization through Coupled Inverse-Forward Graph Learning

Model ReleasesDGX agent

arXiv:2605.20311v1 Announce Type: new Abstract: Guided-wave structural health monitoring enables damage localization in composite plates using sparse networks of bonded piezoelectric transducers. Howe

Weasel: Out-of-Domain Generalization for Web Agents via Importance-Diversity Data Selection

ResearchDGX agent

arXiv:2605.20291v1 Announce Type: new Abstract: Large language models (LLMs) have enabled web agents that follow natural language goals through multi-step browser interactions. However, agents fine-tu

Weight Decay Regimes in Grokking Transformers: Cheap Online Diagnostics

Model ReleasesDGX agent

arXiv:2605.20441v1 Announce Type: new Abstract: Transformers trained on modular arithmetic exhibit sharp transitions between memorization, generalization, and collapse. We show that weight decay acts

WestWorld: A Knowledge-Encoded Scalable Trajectory World Model for Diverse Robotic Systems

ApplicationsDGX agent

arXiv:2603.14392v2 Announce Type: replace Abstract: Trajectory world models play a crucial role in robotic dynamics learning, planning, and control. While recent works have explored trajectory world m

What Twelve LLM Agent Benchmark Papers Disclose About Themselves: A Pilot Audit and an Open Scoring Schema

Model ReleasesDGX agent

arXiv:2605.21404v1 Announce Type: new Abstract: We read twelve well-known LLM agent benchmark papers and recorded, dimension by dimension, what each paper actually says about how its evaluation was ru

When AI Gets it Wrong: Reliability and Risk in AI-Assisted Medication Decision Systems

SafetyDGX agent

arXiv:2604.01449v3 Announce Type: replace-cross Abstract: Artificial intelligence (AI) systems are increasingly integrated into healthcare and pharmacy workflows, supporting tasks such as medication r

When Does Adaptation Win? Scaling Laws for Meta-Learning in Quantum Control

ResearchDGX agent

arXiv:2601.18973v4 Announce Type: replace Abstract: Quantum hardware suffers from intrinsic device heterogeneity and environmental drift, forcing practitioners to choose between suboptimal non-adaptiv

When to Retrain after Drift: A Data-Only Test of Post-Drift Data Size Sufficiency

Model ReleasesDGX agent

arXiv:2603.09024v2 Announce Type: replace Abstract: Sudden concept drift makes previously trained predictors unreliable, yet deciding when to retrain and what post-drift data size is sufficient is rar

Why Ask One When You Can Ask k? Learning-to-Defer to the Top-k Experts

SafetyDGX agent

arXiv:2504.12988v5 Announce Type: replace Abstract: Existing Learning-to-Defer (L2D) frameworks are limited to single-expert deferral, forcing each query to rely on only one expert and preventing the

ZEBRA: Zero-shot Budgeted Resource Allocation for LLM Orchestration

Model ReleasesDGX agent

arXiv:2605.20485v1 Announce Type: new Abstract: As autonomous agents increasingly execute end-to-end tasks under fixed monetary budgets, the pressing open question shifts from whether the budget is re

20 May 2026

A Cloud-Based Tool for Meteorite Recovery Using Drones and Machine Learning

ResearchDGX agent

arXiv:2605.19179v1 Announce Type: cross Abstract: We present a cloud-based tool that uses drones and machine learning to help recover instrumentally observed meteorite falls. We showcase a collection

A Derandomization Framework for Structure Discovery: Applications in Neural Networks and Beyond

ResearchDGX agent

arXiv:2510.19382v2 Announce Type: replace-cross Abstract: Understanding the dynamics of feature learning in neural networks (NNs) remains a significant challenge. The work of (Mousavi-Hosseini et al.,

A Family of Divergence Measures for Evaluating the Reconstruction Quality of Explainable Ensemble Trees

Model ReleasesDGX agent

arXiv:2605.19618v1 Announce Type: new Abstract: Validating interpretable surrogate models for ensemble learners requires measuring agreement between the ensemble's internal representation and its surr

A first-order method for nonconvex-nonconcave minimax problems under a local Kurdyka-Lojasiewicz condition

ResearchDGX agent

arXiv:2507.01932v2 Announce Type: replace-cross Abstract: We study a class of nonconvex-nonconcave minimax problems in which the inner maximization problem satisfies a local Kurdyka-Lojasiewicz (KL) c

A Geometric Analysis of Sign-Magnitude Asymmetry in a ReLU + RMSNorm Block under Ternary Quantization

SafetyDGX agent

arXiv:2605.18933v1 Announce Type: new Abstract: Pre-norm Transformers with RMSNorm tolerate ternary {-1,0,+1} weight quantization with surprisingly small loss (Ma et al., 2024). We give a geometric ex

A Heuristic Approach for Performance Tuning in RL-based Quadrotor Control via Reward Design and Termination Conditions

SafetyDGX agent

arXiv:2605.19166v1 Announce Type: cross Abstract: Reinforcement learning (RL)-based quadrotor control policies have achieved impressive performance in tasks such as fast navigation in cluttered enviro

← Previous
1…141142143144145…243
Next →