AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,557 results
14 May 2026

Runtime Monitoring of Perception-Based Autonomous Systems via Embedding Temporal Logic

SafetyDGX agent

arXiv:2605.12651v1 Announce Type: new Abstract: Runtime monitoring of autonomous systems traditionally relies on mapping continuous sensor observations to discrete logical propositions defined over lo

Safe Bayesian Optimization for Uncertain Correlations Matrices in Linear Models of Co-Regionalization

Model ReleasesDGX agent

arXiv:2605.13302v1 Announce Type: new Abstract: This paper extends safety guarantees for multi-task Bayesian optimization with uncertain correlation matrices from intrinsic co-reginalization models to

Sample-Efficient Optimisation over the Outputs of Generative Models

ResearchDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2509.23800v3 Announce Type: replace-cross Abstract: Modern generative AI models, such as diffusion and flow matching models, can sample from rich data distributions. However, many applications,

Sampling from Flow Language Models via Marginal-Conditioned Bridges

ResearchDGX agent

arXiv:2605.13681v1 Announce Type: new Abstract: Flow Language Models (FLMs) are a recently introduced class of language models which adapt continuous flow matching for one-hot encoded token sequences.

Scale-Sensitive Shattering: Learnability and Evaluability at Optimal Scale

ResearchDGX agent

arXiv:2605.13684v1 Announce Type: new Abstract: We study the optimal scale at which real-valued function classes exhibit uniform convergence and learnability. Our main result establishes a scale-sensi

Scaling Laws for Mixture Pretraining Under Data Constraints

ResearchDGX agent

arXiv:2605.12715v1 Announce Type: new Abstract: As language models scale, the amount of data they require grows -- yet many target data sources, such as low-resource languages or specialized domains,

scShapeBench: Discovering geometry from high dimensional scRNAseq data

Model ReleasesDGX agent

arXiv:2605.12662v1 Announce Type: new Abstract: High-dimensional point cloud data arise across many scientific domains, especially single-cell biology. The shapes or topologies of these datasets deter

Separating Shortcut Transition from Cross-Family OOD Failure in a Minimal Model

ResearchDGX agent

arXiv:2605.12945v1 Announce Type: new Abstract: Shortcut features are often invoked to explain out-of-distribution (OOD) failure, but training correlation, learned shortcut use, and test-time failure

Sharpness-Guided Group Relative Policy Optimization via Probability Shaping

SafetyDGX agent

arXiv:2511.00066v4 Announce Type: replace Abstract: Reinforcement learning with verifiable rewards (RLVR) has become a practical route to improve large language model reasoning, and Group Relative Pol

SHM-Agents: A Generalist-Specialist Integrated Agent System for Structural Health Monitoring

AgentsDGX agent

arXiv:2605.12916v1 Announce Type: cross Abstract: Artificial intelligence is increasingly used to simplify complex tasks. In engineering applications of structural health monitoring (SHM), existing sp

Shortcut Mitigation via Spurious-Positive Samples

TutorialsDGX agent

arXiv:2605.13340v1 Announce Type: new Abstract: Shortcut mitigation strategies commonly rely on training data annotations, group-balanced held-out data or the presence of all groups, i.e., all combina

SMA: Submodular Modality Aligner For Data Efficient Multimodal Learning

Model ReleasesDGX agent

arXiv:2605.12872v1 Announce Type: new Abstract: Despite the recent success of Multimodal Foundation Models (FMs), their reliance on massive paired datasets limits their applicability in low-data and r

Small Area Estimation of Case Growths for Timely COVID-19 Outbreak Detection

Model ReleasesDGX agent

arXiv:2312.04110v2 Announce Type: replace-cross Abstract: The COVID-19 pandemic has exerted a profound impact on the global economy and continues to exact a significant toll on human lives. The COVID-

Sockpuppetting: Jailbreaking LLMs by Combining Prefilling with Optimization

Model ReleasesDGX agent

arXiv:2601.13359v2 Announce Type: replace-cross Abstract: Prefill attacks are an effective and low-cost jailbreaking method, as they directly insert an acceptance sequence (e.g., 'Sure, here is how to

SoK: A Comprehensive Analysis of the Current Status of Neural Tangent Generalization Attacks with Research Directions

ResearchDGX agent

arXiv:2605.12792v1 Announce Type: new Abstract: There is recently a serious issue that Deep Neural Networks (DNNs) training uses more and more unauthorized data. A clean-label generalization attack, o

Spatiotemporal downscaling and nowcasting of urban land surface temperatures with deep neural networks

SafetyDGX agent

arXiv:2605.13566v1 Announce Type: new Abstract: Land Surface Temperature (LST) is a key variable for various applications, such as urban climate and ecology studies. Yet, existing satellite-derived LS

Spectral Energy Centroid: a Metric for Improving Performance and Analyzing Spectral Bias in Implicit Neural Representations

SafetyDGX agent

arXiv:2605.12709v1 Announce Type: new Abstract: Implicit Neural Representations (INRs) model continuous signals using multilayer perceptrons (MLPs), enabling compact, differentiable, and high-fidelity

SpikeProphecy: A Large-Scale Benchmark for Autoregressive Neural Population Forecasting

Model ReleasesDGX agent

arXiv:2605.12992v1 Announce Type: cross Abstract: Neural population models, which predict the joint firing of many simultaneously recorded neurons forward in time, are typically evaluated by a single

State-of-art minibatches via novel DPP kernels: discretization, wavelets, and rough objectives

ResearchDGX agent

arXiv:2605.13127v1 Announce Type: cross Abstract: Determinantal point processes (DPPs) have emerged as a kernelized alternative to vanilla independent sampling for generating efficient minibatches, co

State-Space NTK Collapse Near Bifurcations

Model ReleasesDGX agent

arXiv:2605.12763v1 Announce Type: new Abstract: Rich feature learning in tasks that unfold over time often requires the model to pass through bifurcations, constituting qualitative changes in the unde

Steer-to-Detect: Probing Hidden Representations for Detection of LLM-Generated Texts

ResearchDGX agent

arXiv:2605.12890v1 Announce Type: cross Abstract: The rapid advancement of large language models (LLMs) has made machine-generated text increasingly difficult to distinguish from human-written text. W

Stochastic Dimension-Free Zeroth-Order Estimator for High-Dimensional and High-Order PINNs

Model ReleasesDGX agent

arXiv:2603.24002v2 Announce Type: replace Abstract: Physics-Informed Neural Networks (PINNs) for high-dimensional and high-order partial differential equations (PDEs) are primarily constrained by the

Strategic PAC Learnability via Geometric Definability

ResearchDGX agent

arXiv:2605.13426v1 Announce Type: new Abstract: Strategic classification studies learning settings in which individuals can modify their features, at a cost, in order to influence the classifier's dec

Supervised Deep Multimodal Matrix Factorization for Interpretable Brain Network Analysis

ResearchDGX agent

arXiv:2605.13312v1 Announce Type: new Abstract: We present Supervised Deep Multimodal Matrix Factorization (SD3MF), an interpretable framework for integrative brain network analysis that generalizes S

Support-Conditioned Flow Matching Is Kernel Smoothing

ResearchDGX agent

arXiv:2605.13386v1 Announce Type: new Abstract: Generative models are often conditioned on a small set of examples via cross-attention. Under the Gaussian optimal-transport path, we show that the exac

Survival In-Context: Amortized Bayesian Survival Analysis via Prior-Fitted Networks

ApplicationsDGX agent

arXiv:2603.29475v2 Announce Type: replace Abstract: Survival analysis is crucial for many medical applications, but remains challenging for modern machine learning due to limited data, censoring, and

Switching Successor Measures for Hierarchical Zero-shot Reinforcement Learning

SafetyDGX agent

arXiv:2605.13207v1 Announce Type: new Abstract: Hierarchical reinforcement learning can improve generalization by decomposing long-horizon decision-making into simpler subproblems. However, existing a

Teaching and Learning under Deductive Errors

ResearchDGX agent

arXiv:2605.13384v1 Announce Type: new Abstract: Most models of machine teaching and learning assume the learner makes no errors in its internal deductive inference. However, humans and large language

Test-time Offline Reinforcement Learning on Goal-related Experience

SafetyDGX agent

arXiv:2507.18809v2 Announce Type: replace Abstract: Foundation models compress a large amount of information in a single, large neural network, which can then be queried for individual tasks. There ar

The Diffusion Encoder

ResearchDGX agent

arXiv:2605.13399v1 Announce Type: new Abstract: We construct a new kind of encoder, leveraging the expressive power of diffusion models. In a traditional variational autoencoder, the encoder and decod

The Efficiency Gap in Byte Modeling

ResearchDGX agent

arXiv:2605.12928v1 Announce Type: new Abstract: Modern language models have historically relied on two dominant design choices: subword tokenization and autoregressive (AR) ordering. These design deci

The Geometry of LLM Quantization: GPTQ as Babai's Nearest Plane Algorithm

Model ReleasesDGX agent

arXiv:2507.18553v4 Announce Type: replace Abstract: Quantizing the weights of large language models (LLMs) from 16-bit to lower bitwidth is the de facto approach to deploy massive transformers onto mo

The Horizon Threshold in Cooperative Multi-Agent Reward-Free Exploration

SafetyDGX agent

arXiv:2602.01453v3 Announce Type: replace Abstract: We study cooperative multi-agent reinforcement learning in the setting of reward-free exploration, where multiple agents jointly explore an unknown

The Mechanism of Weak-to-Strong Generalization: Feature Elicitation from Latent Knowledge

ResearchDGX agent

arXiv:2605.12908v1 Announce Type: cross Abstract: Weak-to-strong (W2S) generalization, in which a strong model is fine-tuned on outputs of a weaker, task-specialized model, has been proposed as an app

The Payment Heterogeneity Index: An Integrated Unsupervised Framework for High-Volume Procurement Oversight and Decision Support

ResearchDGX agent

arXiv:2605.12547v1 Announce Type: cross Abstract: Public procurement is vulnerable to error, fraud and corruption, yet high transaction volumes overwhelm oversight. While research often focuses on ten

The Sample Complexity of Multiple Change Point Identification under Bandit Feedback

ResearchDGX agent

arXiv:2605.13252v1 Announce Type: cross Abstract: We study multiple change point localization under bandit feedback. An unknown piecewise-constant function on a compact interval can be queried sequent

The Score-Difference Flow for Implicit Generative Modeling

ResearchDGX agent

arXiv:2304.12906v4 Announce Type: replace Abstract: Implicit generative modeling (IGM) aims to produce samples of synthetic data matching the characteristics of a target data distribution. Recent work

Three-Stage Learning Unlocks Strong Performance in Simple Models for Long-Term Time Series Forecasting

ResearchDGX agent

arXiv:2605.13678v1 Announce Type: new Abstract: Recent studies on long-term time series forecasting have shown that simple linear models and MLP-based predictors can achieve strong performance without

Tight Sample Complexity Bounds for Entropic Best Policy Identification

SafetyDGX agent

arXiv:2605.13717v1 Announce Type: new Abstract: We study best-policy identification for finite-horizon risk-sensitive reinforcement learning under the entropic risk measure. Recent work established a

Tighter Learning Guarantees on Digital Computers via Concentration of Measure on Finite Spaces

Model ReleasesDGX agent

arXiv:2402.05576v4 Announce Type: replace Abstract: Machine learning models with inputs in a Euclidean space R^d, when implemented on digital computers, generalize, and their generalization gap conver

ToolMol: Evolutionary Agentic Framework for Multi-objective Drug Discovery

AgentsDGX agent

arXiv:2605.12784v1 Announce Type: new Abstract: Advances in large language models (LLMs) have recently opened new and promising avenues for small-molecule drug discovery. Yet existing LLM-based approa

Toward AI-Driven Digital Twins for Metropolitan Floods: A Conditional Latent Dynamics Network Surrogate of the Shallow Water Equations

Model ReleasesDGX agent

arXiv:2605.13761v1 Announce Type: new Abstract: AI-driven flood digital twins demand fast hydrodynamic surrogates for ensemble forecasting and observation assimilation. Yet even GPU-accelerated two-di

Towards Generalizable Reasoning: Group Causal Counterfactual Policy Optimization for LLM Reasoning

SafetyDGX agent

arXiv:2602.06475v2 Announce Type: replace Abstract: Large language models (LLMs) excel at complex tasks with advances in reasoning capabilities. However, existing reward mechanisms remain tightly coup

Trajectory-Level Data Augmentation for Offline Reinforcement Learning

SafetyDGX agent

arXiv:2605.13401v1 Announce Type: new Abstract: We propose a data augmentation method for offline reinforcement learning, motivated by active positioning problems. Particularly, our approach enables t

TurboGR: An Accelerated Training System for Large-Scale Generative Recommendation

ResearchDGX agent

arXiv:2605.13433v1 Announce Type: cross Abstract: Generative recommendation (GR) has emerged as a promising paradigm that replaces fragmented, scenario-specific architectures with unified Transformer-

Twincher: Bijective Representation Learning for Robust Inversion of Continuous Systems

TutorialsDGX agent

arXiv:2605.13470v1 Announce Type: new Abstract: Recent advances in AI have been primarily driven by large-scale neural architectures that excel at function approximation, rather than by tailored induc

U-HNO: A U-shaped Hybrid Neural Operator with Sparse-Point Adaptive Routing for Non-stationary PDE Dynamics

ResearchDGX agent

arXiv:2605.12965v1 Announce Type: new Abstract: Solutions to many partial differential equations (PDEs) display coexisting smooth global transport and localized sharp features within a single trajecto

UFO: A Domain-Unification-Free Operator Framework for Generalized Operator Learning

ResearchDGX agent

arXiv:2605.12700v1 Announce Type: new Abstract: Neural operators have become an effective framework for learning mappings between function spaces, yet most existing architectures realize operators wit

Uncertainty-Aware Prediction of Lung Tumor Growth from Sparse Longitudinal CT Data via Bayesian Physics-Informed Neural Networks

Model ReleasesDGX agent

arXiv:2605.13560v1 Announce Type: new Abstract: This work studies lung tumor growth prediction from sparse and irregular longitudinal computed tomography (CT) observations with measurement variability

Uncertainty-Driven Anomaly Detection for Psychotic Relapse Using Smartwatches: Forecasting and Multi-Task Learning Fusion

Model ReleasesDGX agent

arXiv:2605.13816v1 Announce Type: new Abstract: Digital phenotyping enables continuous passive monitoring of behavior and physiology, offering a promising paradigm for early detection of psychotic rel

Understanding Catastrophic Forgetting In LoRA via Mean-Field Attention Dynamics

Model ReleasesDGX agent

arXiv:2402.15415v2 Announce Type: replace Abstract: Low-Rank Adaptation (LoRA) is the dominant parameter-efficient fine-tuning method due to its favorable compute-performance trade-off, yet it suffers

Unified generalization analysis for physics informed neural networks

ResearchDGX agent

arXiv:2605.13260v1 Announce Type: new Abstract: Physics-Informed Neural Networks (PINNs) and their variational counterparts (VPINNs) are neural networks that incorporate physical laws, making them use

Unifying Entropy Regularization in Optimal Control: From and Back to Classical Objectives via Iterated Soft Policies and Path Integral Solutions

SafetyDGX agent

arXiv:2512.06109v3 Announce Type: replace-cross Abstract: This paper develops a unified perspective on several optimal control formulations through the lens of Kullback-Leibler (KL) regularization. We

Universal Representation of Generalized Convex Functions and their Gradients

Model ReleasesDGX agent

arXiv:2509.04477v3 Announce Type: replace-cross Abstract: A wide range of optimization problems can often be written in terms of generalized convex functions (GCFs). When this structure is present, it

Vector-Quantized Discrete Latent Factors Meet Financial Priors: Dynamic Cross-Sectional Stock Ranking Prediction for Portfolio Construction

ResearchDGX agent

arXiv:2605.13407v1 Announce Type: new Abstract: Predicting cross-sectional stock returns is challenging due to low signal-to-noise ratios and evolving market regimes. Classical factor models offer int

VectorSmuggle: Steganographic Exfiltration in Embedding Stores and a Cryptographic Provenance Defense

Model ReleasesDGX agent

arXiv:2605.13764v1 Announce Type: cross Abstract: Modern retrieval-augmented generation (RAG) systems convert sensitive content into high-dimensional embeddings and store them in vector databases that

VIP-COP: Context Optimization for Tabular Foundation Models

ResearchDGX agent

arXiv:2605.12904v1 Announce Type: new Abstract: Tabular foundation models (TFMs) have emerged as a powerful paradigm for in-context learning on structured data, enabling direct prediction on new tabul

What Information Matters? Graph Out-of-Distribution Detection via Tri-Component Information Decomposition

ResearchDGX agent

arXiv:2605.13032v1 Announce Type: new Abstract: Graph neural networks are widely used for node classification, but they remain vulnerable to out-of-distribution (OOD) shifts in node features and graph

What is Learnable in Valiant's Theory of the Learnable?

ResearchDGX agent

arXiv:2605.13840v1 Announce Type: cross Abstract: Valiant's 1984 paper is widely credited with introducing the PAC learning model, but it, in fact, introduced a different model: unlike PAC learning, t

When and Why is Optimistic Multiplicative Weights Slow? The Geometry of Energy Dissipation

ResearchDGX agent

arXiv:2605.13242v1 Announce Type: cross Abstract: This paper studies the convergence of the Optimistic Multiplicative Weights Update algorithm (OMWU) in two player zero-sum games. Recent works have id

← Previous
1…160161162163164…243
Next →