AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,557 results
26 May 2026

Localization then Neutralization: Gradient-guided Token Suppression against Visual Prompt Injection Attack

Local AiDGX agent

arXiv:2605.25194v1 Announce Type: new Abstract: Adversarial images pose a severe security threat to multimodal large language models through prompt injection. Existing defenses largely lack a principl

Logic-Guided Vector Fields for Constrained Generative Modeling

ResearchDGX agent

arXiv:2602.02009v2 Announce Type: replace Abstract: Neuro-symbolic systems aim to combine the expressive structure of symbolic logic with the flexibility of neural learning; yet, generative models typ

Looped Diffusion Language Models

Model ReleasesDGX agent

arXiv:2605.26106v1 Announce Type: new Abstract: Masked diffusion models (MDMs) have emerged as a promising alternative to autoregressive models for language modeling, yet the effective design of trans


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

LWM-CDE: A Representation Space for Wireless Data Reasoning and Transferability

ApplicationsDGX agent

arXiv:2605.24077v1 Announce Type: cross Abstract: Machine learning deployments in real-world wireless communication tasks face significant generalization challenges due to location and environment-spe

Machine Learning Multiscale Interactions

ResearchDGX agent

arXiv:2605.25710v1 Announce Type: cross Abstract: Realistic physical systems are characterised by emergent interactions across multiple length and time scales, posing a significant challenge for predi

MathOptAI.jl: Embed trained machine learning predictors into JuMP models

HardwareDGX agent

arXiv:2507.03159v2 Announce Type: replace Abstract: We present exttt{MathOptAI.jl}, an open-source Julia library for embedding trained machine learning predictors into a JuMP model. exttt{MathOptAI.jl

Mean-Shift PCA by Knockoff Mean

TutorialsDGX agent

arXiv:2605.25460v1 Announce Type: cross Abstract: Removing noise is difficult, but adding noise is easy. In this work, we show how to eliminate mean-shift noisy components from PCA by deliberately int

MEDAL: Manifold Embedding Distillation via Autoencoder Learning

Model ReleasesDGX agent

arXiv:2605.24244v1 Announce Type: cross Abstract: Low-dimensional embeddings are widely used as visual summaries of high-dimensional data and to enable downstream scientific discoveries. Yet, popular

MedMamba: Multi-View State Space Models with Adaptive Graph Learning for Medical Time Series Classification

ApplicationsDGX agent

arXiv:2605.24961v1 Announce Type: new Abstract: Medical time series are central to healthcare, enabling continuous monitoring and supporting timely clinical decisions. Despite recent progress, existin

Memory-Induced Tool-Drift in LLM Agents

Model ReleasesDGX agent

arXiv:2605.24941v1 Announce Type: cross Abstract: Modern LLM agents combine long-term memory for personalization with tool-calling interfaces for taking actions in the world -- a combination underpinn

Merge-Bench: Resolve Merge Conflicts with Large Language Models

Model ReleasesDGX agent

arXiv:2605.25890v1 Announce Type: new Abstract: This paper applies machine learning to the difficult and important task of version control merging. (1) We constructed a dataset, Merge-Bench, of 7938 r

MGVQ: Synergizing Multi-dimensional Sensitivity-Aware and Gradient-Hessian Fusion for Vector Quantization

ResearchDGX agent

arXiv:2605.24019v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) achieve outstanding performance, yet their huge model size severely hinders deployment on edge devices with limited reso

MimirRAG: A Multi-Agent RAG Framework for Financial Data Retrieval with Metadata Integration

Model ReleasesDGX agent

arXiv:2605.25030v1 Announce Type: new Abstract: Retrieval-augmented generation (RAG) systems offer a promising approach to reduce hallucinations and improve answer accuracy in large language models (L

Minimax Limits of k-Fold Cross-Validation via Majority

Model ReleasesDGX agent

arXiv:2605.25859v1 Announce Type: cross Abstract: We study the mean-squared error of k-fold cross-validation as a risk estimator, with particular emphasis on how its accuracy depends on the number of

Missing Pattern Recognized Diffusion Imputation Model for Missing Not At Random

ApplicationsDGX agent

arXiv:2605.25439v1 Announce Type: new Abstract: Missing data frequently arises across diverse domains, including time-series and image domains. In the real world, missing occurrences often depend on t

Mitigating Gradient Pathology in PINNs through Aligned Constraint

ResearchDGX agent

arXiv:2605.25001v1 Announce Type: new Abstract: While Physics-Informed Neural Networks (PINNs) are powerful for solving Partial Differential Equations (PDEs), their training is often paralyzed by grad

Multi-Agent Systems are Mixtures of Experts: Who Becomes an Influencer?

SafetyDGX agent

arXiv:2605.25929v1 Announce Type: cross Abstract: The effectiveness of multi-agent LLM deliberation depends not only on the agents' individual predictions, but also on how they communicate and collabo

Multi-Alignment Contrastive Learning for Enzyme--Reaction Retrieval

SafetyDGX agent

arXiv:2512.08508v2 Announce Type: replace-cross Abstract: Identifying enzymes that catalyze target biochemical reactions is a key step in computational enzyme discovery and biocatalyst design. Recent

Multi-Level Strategic Classification: Incentivizing Improvement through Promotion and Relegation Dynamics

AgentsDGX agent

arXiv:2602.11439v2 Announce Type: replace Abstract: Strategic classification studies the problem where self-interested individuals or agents manipulate their response to obtain favorable decision outc

Multicalibration Boosting: Theory, Convergence, and Transferability

SafetyDGX agent

arXiv:2605.24364v1 Announce Type: cross Abstract: Multicalibration extends classical calibration by requiring predictions to be unbiased over a rich collection of functions, encompassing both predicti

Multimodality Stacking with Blockwise missing values and application to the PIONeeR biomarkers study for prediction of resistance to immunotherapy

ResearchDGX agent

arXiv:2605.25050v1 Announce Type: cross Abstract: Integrating multimodal datasets in clinical oncology is frequently hindered by high dimensionality and blockwise missingness, where entire data source

Multitask learning with semiempirical orbital charges enables sample-efficient MLIPs

TutorialsDGX agent

arXiv:2605.24073v1 Announce Type: cross Abstract: Machine learning interatomic potentials (MLIPs) require generating computationally expensive, large-scale training datasets to accurately simulate mat

Muon in Associative Memory Learning: Training Dynamics and Scaling Laws

Model ReleasesDGX agent

arXiv:2602.05725v2 Announce Type: replace Abstract: Muon updates matrix parameters via the matrix sign of the gradient and has shown strong empirical gains, yet its dynamics and scaling behavior remai

Muon in Vision Transformers: Optimizer-Recipe Interactions and Gradient Spectra

ResearchDGX agent

arXiv:2605.24770v1 Announce Type: new Abstract: Muon is a recently developed matrix-aware optimizer that has shown strong results in transformer training, but its behavior in vision transformers (ViTs

Music Transcription with (Almost) No Supervision

SafetyDGX agent

arXiv:2605.24193v1 Announce Type: cross Abstract: Competitive music transcription models require large amounts of paired audio-score data, which is scarce due to collection costs, alignment difficulty

MVR-cache: Optimizing Semantic Caching via Multi-Vector Retrieval and Learned Prompt Segmentation

ResearchDGX agent

arXiv:2605.24914v1 Announce Type: cross Abstract: To reduce LLM costs and latency, semantic caching systems must accurately identify when a new prompt matches a cached one. Current methods often rely

Near-Optimal Nonconvex-Strongly-Convex Bilevel Optimization with Fully First-Order Oracles

ResearchDGX agent

arXiv:2306.14853v5 Announce Type: replace-cross Abstract: In this work, we consider bilevel optimization when the lower-level problem is strongly convex. Recent works show that with a Hessian-vector p

NEST: Network- and Memory-Aware Device Placement For Distributed Deep Learning

ResearchDGX agent

arXiv:2603.06798v2 Announce Type: replace Abstract: The growing scale of deep learning demands distributed training frameworks that jointly reason about parallelism, memory, and network topology. Prio

Neural Integral Operators for Inverse Problems: An Operator-Learning Framework for Small-Sample Spectroscopic Classification

Model ReleasesDGX agent

arXiv:2505.03677v3 Announce Type: replace Abstract: Learning maps between function spaces with a strong inductive bias is a central challenge in soft computing, especially when training data are scarc

Neural Stochastic Differential Equations on Compact State Spaces: Theory, Methods, and Application to Suicide Risk Modeling

ResearchDGX agent

arXiv:2508.17090v4 Announce Type: replace-cross Abstract: Ecological Momentary Assessment (EMA) studies enable the collection of high-frequency self-reports of suicidal thoughts and behaviors (STBs) v

Nonconvex Decentralized Stochastic Bilevel Optimization under Heavy-Tailed Noise

ApplicationsDGX agent

arXiv:2509.15543v2 Announce Type: replace Abstract: Existing decentralized stochastic optimization methods assume the lower-level loss function is strongly convex and the stochastic gradient noise has

Nonstationary Generalized Linear Bandits with Discounted Online Mirror Descent

Model ReleasesDGX agent

arXiv:2605.25590v1 Announce Type: cross Abstract: We study nonstationary generalized linear bandits (GLBs), where the expected reward is modeled through a nonlinear link function with an unknown time-

NormimesDirection: Restoring the Missing Query Norm in Vision Linear Attention

Model ReleasesDGX agent

arXiv:2506.21137v3 Announce Type: replace Abstract: Linear attention mitigates the quadratic complexity of softmax attention but suffers from a critical loss of expressiveness. We identify two primary

Not only where, But when: Temporal Scheduling for RLVR

SafetyDGX agent

arXiv:2605.25381v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) has become a core technique for post-training of Large Language Models (LLMs). While policy optimi

Nystrom Kernel Stein Discrepancy Tests

ResearchDGX agent

arXiv:2605.25173v1 Announce Type: cross Abstract: Kernel Stein discrepancy (KSD) is among the most popular goodness-of-fit (GoF) measures on general domains with a large number of successful deploymen

Omissive Bias in Religious Representation: Benchmarking LLM Answers to Everyday Ethical Decision-making

Model ReleasesDGX agent

arXiv:2605.24319v1 Announce Type: new Abstract: As large language models become a default source of guidance on personal, moral, and existential questions, it matters whether they draw on the religiou

On Reliability of Efficient Membership Inference Vulnerability Evaluation

SafetyDGX agent

arXiv:2605.25819v1 Announce Type: new Abstract: Membership inference attacks (MIAs) are popular methods for empirically assessing the leakage of sensitive information in the training data through mode

On the Communication Complexity of Decentralized Stochastic Bilevel Optimization

TutorialsDGX agent

arXiv:2311.11342v5 Announce Type: replace Abstract: Stochastic bilevel optimization finds widespread applications in machine learning, including meta-learning, hyperparameter optimization, and neural

On the Interaction of Batch Noise, Adaptivity, and Compression, under (L_0,L_1)-Smoothness: An SDE Approach

ResearchDGX agent

arXiv:2506.00181v2 Announce Type: replace Abstract: Distributed stochastic optimization intertwines (i) stochastic gradient noise, (ii) communication compression, and (iii) adaptive/normalized updates

On the Sample Complexity of Robust Binary Hypothesis Testing

Model ReleasesDGX agent

arXiv:2605.24741v1 Announce Type: cross Abstract: We study the sample complexity of robust binary hypothesis testing under three standard contamination models: arepsilon-additive (Huber), arepsilon-su

One-for-All Model Initialization with Frequency-Domain Knowledge

Model ReleasesDGX agent

arXiv:2603.07523v2 Announce Type: replace Abstract: Transferring knowledge by fine-tuning large-scale pre-trained networks has become a standard paradigm for downstream tasks, yet the knowledge of a p

One-shot Conditional Sampling: MMD meets Nearest Neighbors

ResearchDGX agent

arXiv:2509.25507v2 Announce Type: replace-cross Abstract: How can we generate samples from a conditional distribution that we never fully observe? This question arises across a broad range of applicat

One-Step Bellman Alignment Enables Provably Efficient Transfer in Online RL

SafetyDGX agent

arXiv:2601.21924v2 Announce Type: replace Abstract: We study online transfer reinforcement learning (RL) in episodic Markov decision processes, where experience from related source tasks is available

Opportunistic Target Selection: Early Directional Commitment for Query-Efficient Black-Box Adversarial Attacks

ResearchDGX agent

arXiv:2605.25663v1 Announce Type: new Abstract: Black-box adversarial attacks that minimize only the ground-truth confidence suffer from class drift: perturbations wander through the feature space wit

Optimal and Order-optimal Gated Priority-based Greedy Policies for Two-layer Multi-item Order Fulfillment

Local AiDGX agent

arXiv:2605.25888v1 Announce Type: new Abstract: We study how an e-commerce firm should make real-time fulfillment decisions in a two-layer distribution network when multi-item customer orders arrive s

Optimal Design for Multinomial Logit Model with Applications to Best Assortment Identification

AgentsDGX agent

arXiv:2605.25592v1 Announce Type: cross Abstract: We study optimal experimental design for multinomial logit (MNL) bandits, where an agent repeatedly selects a subset of K items from a ground set of s

Optimal Non-Asymptotic Edgeworth Expansions for Multivariate Neural Network Outputs

ResearchDGX agent

arXiv:2605.24072v1 Announce Type: cross Abstract: Finite-width fully connected neural networks with Gaussian-initialized weights deviate from their infinite-width Gaussian limit, exhibiting non-vanish

Optimal uncertainty bounds for multivariate kernel regression under bounded noise: A Gaussian process-based dual function

ResearchDGX agent

arXiv:2603.16481v2 Announce Type: replace Abstract: Non-conservative uncertainty bounds are essential for making reliable predictions about latent functions from noisy data, and thus, a key enabler fo

Optimizing Digital Therapeutic Interventions: Online Learning under Endogenous Adherence

Model ReleasesDGX agent

arXiv:2605.24261v1 Announce Type: new Abstract: A critical challenge facing clinicians managing chronic disease interventions is sustaining long-run patient health given limited information and resour

Optimizing Multidimensional Scaling in Gini Metric Spaces

HardwareDGX agent

arXiv:2605.25124v1 Announce Type: new Abstract: The Gini Multidimensional Scaling (Gini MDS) framework extends the Euclidean multidimensional scaling. We introduce a Gini pseudo-distance based on valu

ORACAL: A Robust and Explainable Multimodal Framework for Smart Contract Vulnerability Detection with Causal Graph Enrichment

Model ReleasesDGX agent

arXiv:2603.28128v2 Announce Type: replace Abstract: Although Graph Neural Networks (GNNs) have shown promise for smart contract vulnerability detection, they still face significant limitations. Homoge

PAC Learning with Bandit Feedback: Sharp Sample Complexity in the Realizable Setting

ResearchDGX agent

arXiv:2605.25678v1 Announce Type: cross Abstract: We study the problem of multiclass PAC learning with bandit feedback in the realizable setting. In this framework, there is an unknown data distributi

PairFlow: Closed-Form Source-Target Coupling for Few-Step Generation in Discrete Flow Models

ResearchDGX agent

arXiv:2512.20063v3 Announce Type: replace Abstract: We introduce exttt{PairFlow}, a lightweight preprocessing step for training Discrete Flow Models (DFMs) to achieve few-step sampling without requiri

Paris 2.0: A Decentralized Diffusion Model for Video Generation

HardwareDGX agent

arXiv:2605.26064v1 Announce Type: cross Abstract: We present Paris 2.0, the first video generation model pre-trained through decentralized computation. Its training recipe builds upon Paris 1.0 (arXiv

Partition of Unity Neural Networks for Interpretable Classification with Explicit Class Regions

Model ReleasesDGX agent

arXiv:2602.00511v2 Announce Type: replace Abstract: Despite their empirical success, neural network classifiers remain difficult to interpret. In softmax-based models, class regions are defined implic

PDEInvBench: A Comprehensive Dataset and Design Space Exploration of Neural Networks for PDE Inverse Problems

Model ReleasesDGX agent

arXiv:2605.25353v1 Announce Type: new Abstract: Inverse problems in partial differential equations (PDEs) involve estimating the physical parameters of a system from observed spatiotemporal solution f

Personalized Federated Learning by Energy-Efficient UAV Communications

Model ReleasesDGX agent

arXiv:2605.25212v1 Announce Type: new Abstract: Federated learning (FL) is an effective paradigm for enhancing the learning capability of edge devices while preserving data privacy. In geographically

Physen-Noise2Noise: Physics-Guided Self-Supervised Defocus Deblurring with Bias Correction under Low-Light Conditions

Model ReleasesDGX agent

arXiv:2605.24590v1 Announce Type: cross Abstract: Low-light, long-exposure defocus deblurring remains a challenging problem due to the simultaneous presence of severe blur and complex biased noise. Ex

Physical Analogue Kolmogorov-Arnold Networks based on Reconfigurable Nonlinear-Processing Units

HardwareDGX agent

arXiv:2602.07518v2 Announce Type: replace-cross Abstract: Kolmogorov-Arnold Networks (KANs) shift neural computation from linear layers to learnable nonlinear edge functions, but implementing these no

Physics-Guided Concentration Inference from Resistance Transients in a Mixed-Phase SnO-SnO_2 Carbon Monoxide Sensor with p-n Switching

ResearchDGX agent

arXiv:2605.23971v1 Announce Type: cross Abstract: This work presents a physics-guided machine-learning framework for carbon monoxide concentration inference from experimentally measured resistance tra

← Previous
1…127128129130131…243
Next →