AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,432 results
7 May 2026

Malliavin Calculus for Counterfactual Gradient Estimation in Adaptive Inverse Reinforcement Learning

ResearchDGX agent

arXiv:2604.01345v2 Announce Type: replace Abstract: Inverse reinforcement learning (IRL) recovers the loss function of a forward learner from its observed responses. Adaptive IRL aims to reconstruct t

MalPurifier: Enhancing Android Malware Detection with Adversarial Purification against Evasion Attacks

ResearchDGX agent

arXiv:2312.06423v3 Announce Type: replace-cross Abstract: Machine learning (ML) has gained significant adoption in Android malware detection to address the escalating threats posed by the rapid prolif

Manifold of Failure: Behavioral Attraction Basins in Language Models

Model ReleasesDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2602.22291v3 Announce Type: replace Abstract: While prior work has focused on projecting adversarial examples back onto the manifold of natural data to restore safety, we argue that a comprehens

Manifold Steering Reveals the Shared Geometry of Neural Network Representation and Behavior

ResearchDGX agent

arXiv:2605.05115v1 Announce Type: new Abstract: Neural representations carry rich geometric structure; but does that structure causally shape behavior? To address this question, we intervene along pat

Massively Parallel Exact Inference for Hawkes Processes

HardwareDGX agent

arXiv:2604.01342v2 Announce Type: replace Abstract: Multivariate Hawkes processes are a widely used class of self-exciting point processes, but maximum likelihood estimation naively scales as O(N^2) i

Membership Inference Attacks for Retrieval Based In-Context Learning for Document Question Answering

ResearchDGX agent

arXiv:2605.04116v1 Announce Type: cross Abstract: We show that remotely hosted applications employing in-context learning when augmented with a retrieval function to select in-context examples can be

Memory as a Markov Matrix: Sample Efficient Knowledge Expansion via Token-to-Dictionary Mapping

Model ReleasesDGX agent

arXiv:2605.04308v1 Announce Type: new Abstract: Continual incorporation of new knowledge is essential for the long-term evolution of large language models (LLMs). Existing approaches typically rely on

Meta-Learning and Meta-Reinforcement Learning -- Tracing the Path towards DeepMind's Adaptive Agent

AgentsDGX agent

arXiv:2602.19837v3 Announce Type: replace-cross Abstract: Humans are highly effective at utilizing prior knowledge to adapt to novel tasks, a capability that standard machine learning models struggle

Meta-LegNet: A Transferable and Interpretable Framework for Surface Adsorption Prediction via Self-Defined Adsorption-Environment Learning

TutorialsDGX agent

arXiv:2605.04102v1 Announce Type: cross Abstract: A central challenge in computational catalysis is the identification of low-energy and chemically plausible adsorption configurations, as these direct

Mitigating Label Shift in Tabular In-Context Learning via Test-Time Posterior Adjustment

ResearchDGX agent

arXiv:2605.04363v1 Announce Type: new Abstract: TabPFN has recently gained attention as a foundation model for tabular datasets, achieving strong performance by leveraging in-context learning on synth

MixINN: Accelerating Plant Breeding by Combining Mixed Models and Deep Learning for Interaction Prediction

Local AiDGX agent

arXiv:2605.04744v1 Announce Type: new Abstract: Plant breeding underpins global food security through incremental, accumulating improvements in crop yield, quality and sustainability, achieved via rep

Model synthesis and identifiability analysis of stiff chemical reaction systems with inVAErt networks

ResearchDGX agent

arXiv:2605.04134v1 Announce Type: new Abstract: We consider the problem of learning data-driven replicas for stiff systems of ordinary differential equations arising in chemical kinetics that can be e

MoLF: Mixture-of-Latent-Flow for Pan-Cancer Spatial Gene Expression Prediction from Histology

ResearchDGX agent

arXiv:2602.02282v2 Announce Type: replace Abstract: Inferring spatial transcriptomics (ST) from histology enables scalable histogenomic profiling, yet current methods are largely restricted to single-

MP-ISMoE: Mixed-Precision Interactive Side Mixture-of-Experts for Efficient Transfer Learning

Model ReleasesDGX agent

arXiv:2605.04058v1 Announce Type: new Abstract: Parameter-efficient transfer learning (PETL) has emerged as a pivotal paradigm for adapting pre-trained foundation models to downstream tasks, significa

Multi-Scale Wavelet Transformers for Operator Learning of Dynamical Systems

SafetyDGX agent

arXiv:2602.01486v2 Announce Type: replace Abstract: Recent years have seen a surge in data-driven surrogates for dynamical systems that can be orders of magnitude faster than numerical solvers. Howeve

Multi-site modelling and reconstruction of past extreme skew surges along the French Atlantic coast

ResearchDGX agent

arXiv:2505.00835v2 Announce Type: replace-cross Abstract: Appropriate modelling of extreme skew surges is crucial, particularly for coastal risk management. Our study focuses on modelling extreme skew

MULTIBENCH++: A Unified and Comprehensive Multimodal Fusion Benchmarking Across Specialized Domains

Model ReleasesDGX agent

arXiv:2511.06452v3 Announce Type: replace Abstract: Although multimodal fusion has made significant progress, its advancement is severely hindered by the lack of adequate evaluation benchmarks. Curren

Multiscale Euclidean Network Trajectories: Second-Moment Geometry, Attribution, and Change Points

ResearchDGX agent

arXiv:2605.04589v1 Announce Type: cross Abstract: A central challenge in dynamic network analysis is to represent temporal evolution in a way that is both geometrically meaningful and statistically id

Multivariate Time Series Data Imputation via Distributionally Robust Regularization

SafetyDGX agent

arXiv:2602.00844v2 Announce Type: replace-cross Abstract: Multivariate time series imputation is often compromised by mismatch between the observed and true data distributions, a bias induced by the c

NEAT: Neighborhood-Guided, Efficient, Autoregressive Set Transformer for 3D Molecular Generation

SafetyDGX agent

arXiv:2512.05844v3 Announce Type: replace Abstract: Transformer-based autoregressive models offer an efficient alternative to diffusion- and flow-matching-based approaches for generating 3D molecules.

Neural Discovery of Strichartz Extremizers

ResearchDGX agent

arXiv:2605.04918v1 Announce Type: cross Abstract: Strichartz inequalities are a cornerstone of the modern theory of dispersive PDEs, but their extremizers are known explicitly only in a handful of sha

Neural-Guided Domain Restriction to Accelerate Pseudospectra Computation for Structured Non-normal Banded Matrices

ResearchDGX agent

arXiv:2605.04550v1 Announce Type: cross Abstract: Computing pseudospectra of non-normal matrices is essential for understanding the stability and transient behavior of dynamical systems. Such analysis

Norm Anchors Make Model Edits Last

ResearchDGX agent

arXiv:2602.02543v3 Announce Type: replace Abstract: Sequential Locate-and-Edit (L&E) model editing can fail abruptly after many edits. We identify and formalize this failure as a positive norm-feedbac

NSL-MT: Linguistically Informed Negative Samples for Efficient Machine Translation in Low-Resource Languages

ResearchDGX agent

arXiv:2511.09537v2 Announce Type: replace Abstract: We introduce negative space learning machine translation (NSL-MT), a training method for underresourced languages, that augments limited parallel da

On-line Learning in Tree MDPs by Treating Policies as Bandit Arms

SafetyDGX agent

arXiv:2605.04979v1 Announce Type: cross Abstract: A Tree Markov Decision Problem (T-MDP) is a finite-horizon MDP with a starting state s_{1}, in which every state is reachable from s_{1} through exact

On the Architectural Complexity of Neural Networks

ResearchDGX agent

arXiv:2605.04325v1 Announce Type: new Abstract: We introduce a unified theoretical framework for the rigorous analysis and systematic construction of deep neural networks (DNNs). This framework addres

On the Hardness of Junking LLMs

SafetyDGX agent

arXiv:2605.05116v1 Announce Type: new Abstract: Large language models (LLMs) are known to be vulnerable to jailbreak attacks, which typically rely on carefully designed prompts containing explicit sem

On the Influence of the Feature Computation Budget on Per-Instance Algorithm Selection for Black-Box Optimization

ResearchDGX agent

arXiv:2605.04954v1 Announce Type: cross Abstract: Per-instance algorithm selection (PIAS) takes advantage of complementarity between a set of algorithms by deciding which algorithm to run on a given i

On the Non-decoupling of Supervised Fine-tuning and Reinforcement Learning in Post-training

ResearchDGX agent

arXiv:2601.07389v2 Announce Type: replace Abstract: Post-training of large language models routinely interleaves supervised fine-tuning (SFT) with reinforcement learning (RL). These two methods have d

On the Wasserstein Gradient Flow Interpretation of Drifting Models

ResearchDGX agent

arXiv:2605.05118v1 Announce Type: new Abstract: Recently, Deng et al. (2026) proposed Generative Modeling via Drifting (GMD), a novel framework for generative tasks. This note presents an analysis of

One Pool, Two Caches: Adaptive HBM Partitioning for Accelerating Generative Recommender Serving

SafetyDGX agent

arXiv:2605.04450v1 Announce Type: cross Abstract: Generative Recommender (GR) inference places embedding hot caches (EMB) and KV caches in direct competition for limited GPU HBM: allocating more memor

Online Continual Learning on Intel Loihi 2 via a Co-designed Spiking Neural Network

Local AiDGX agent

arXiv:2511.01553v2 Announce Type: replace Abstract: AI systems on edge devices require online continual learning -- adapting to non-stationary streams and unfamiliar classes without catastrophic forge

Online Nonstochastic Prediction: Logarithmic Regret via Predictive Online Least Squares

ResearchDGX agent

arXiv:2605.04364v1 Announce Type: new Abstract: We study online prediction for marginally stable, partially observed linear dynamical systems under nonstochastic disturbances. Our objective is to mini

Optimal Control with Natural Images: Efficient Reinforcement Learning using Overcomplete Sparse Codes

Model ReleasesDGX agent

arXiv:2412.08893v3 Announce Type: replace Abstract: Optimal control and sequential decision making are widely used in many complex tasks. Optimal control over a sequence of natural images is a first s

Order-based Rehearsal Learning

TutorialsDGX agent

arXiv:2605.04955v1 Announce Type: new Abstract: When a machine learning (ML) model forecasts an undesired event, one often seeks a decision to avoid it, known as the avoiding undesired future (AUF) pr

Order Matters: Improving Domain Adaptation by Reordering Data

SafetyDGX agent

arXiv:2605.05084v1 Announce Type: new Abstract: Domain shift remains a key challenge in deploying machine learning models to the real world. Unsupervised domain adaptation (UDA) aims to address this b

OSAQ: Outlier Self-Absorption for Accurate Low-bit LLM Quantization

Model ReleasesDGX agent

arXiv:2605.04738v1 Announce Type: new Abstract: Large Language Models (LLMs) have demonstrated remarkable capabilities. However, their massive parameter scale leads to significant resource consumption

PAIR-CI: Calibrated Conditional Independence Testing for Causal Discovery with Incomplete Data

ResearchDGX agent

arXiv:2605.04838v1 Announce Type: cross Abstract: The standard constraint-based paradigm for causal discovery with incomplete data -- impute first, test second -- is frequently miscalibrated: any cons

Personalized Spiking Neural Networks with Ferroelectric Synapses for EEG Signal Processing

Local AiDGX agent

arXiv:2601.00020v3 Announce Type: replace-cross Abstract: Electroencephalography (EEG)-based brain-computer interfaces (BCIs) are strongly affected by non-stationary neural signals that vary across se

Perturbation is All You Need for Extrapolating Language Models

ApplicationsDGX agent

arXiv:2605.04344v1 Announce Type: cross Abstract: We introduce a simple yet powerful framework for training large language models. In contrast to the standard autoregressive next-token prediction base

Physiologically Grounded Driver Behavior Classification: SHAP-Driven Elite Feature Selection and Hybrid Gradient Boosting for Multimodal Physiological Signals

ResearchDGX agent

arXiv:2605.05120v1 Announce Type: new Abstract: An interpretable and scalable framework for decoding driving behaviors from multimodal physiological signals is proposed in this study. We utilize multi

Piper: Efficient Large-Scale MoE Training via Resource Modeling and Pipelined Hybrid Parallelism

HardwareDGX agent

arXiv:2605.05049v1 Announce Type: cross Abstract: Frontier models increasingly adopt Mixture-of-Experts (MoE) architectures to achieve large-model performance at reduced cost. However, training MoE mo

Positional Encoding in Transformer-Based Time Series Models: A Survey

ResearchDGX agent

arXiv:2502.12370v3 Announce Type: replace Abstract: Recent advancements in transformer-based models have greatly improved time series analysis, providing robust solutions for tasks such as forecasting

Power Distribution Bridges Sampling, Self-Reward RL, and Self-Distillation

Local AiDGX agent

arXiv:2605.04542v1 Announce Type: new Abstract: Recent analyses question whether reinforcement learning (RL) is responsible for strong reasoning in large language models (LLMs). At the same time, dist

Predict-then-Diffuse: Adaptive Response Length for Compute-Budgeted Inference in Diffusion LLMs

SafetyDGX agent

arXiv:2605.04215v1 Announce Type: new Abstract: Diffusion-based Large Language Models (D-LLMs) represent a promising frontier in generative AI, offering fully parallel token generation that can lead t

Predictive and Prescriptive AI toward Optimizing Wildfire Suppression

ApplicationsDGX agent

arXiv:2605.04510v1 Announce Type: cross Abstract: Intense wildfire seasons require critical prioritization decisions to allocate scarce suppression resources over a dispersed geographical area. This p

Preference-Based Self-Distillation: Beyond KL Matching via Reward Regularization

SafetyDGX agent

arXiv:2605.05040v1 Announce Type: new Abstract: On-policy distillation is an efficient alternative to reinforcement learning, offering dense token-level training signals. However, its reliance on a st

Probabilistic Classification and Uncertainty Quantification of Sahara Desert Climate Using Feedforward Neural Networks

ResearchDGX agent

arXiv:2605.04286v1 Announce Type: new Abstract: Climate classification plays a vital role in agricultural planning, hydrological studies, and climate science. One of the most widely used systems for c

Probing Structural Mathematical Reasoning in Language Models with Algebraic Trapdoors

Model ReleasesDGX agent

arXiv:2605.04352v1 Announce Type: new Abstract: We introduce a benchmark suite for evaluating structural mathematical reasoning in language models, built on subgroup-construction problems in SL(3, Z)

Provable imitation learning for control of instability in partially-observed Vlasov--Poisson equations

SafetyDGX agent

arXiv:2605.05081v1 Announce Type: new Abstract: We consider the stabilization of Vlasov--Poisson plasma dynamics, a central control problem in nuclear fusion. Our focus is the gap between what an idea

Provable Non-Convex Euclidean Distance Matrix Completion: Geometry, Reconstruction, and Robustness

Model ReleasesDGX agent

arXiv:2508.00091v3 Announce Type: replace-cross Abstract: The problem of recovering the configuration of points from their partial pairwise distances, referred to as the Euclidean Distance Matrix Comp

Proximal Projection for Doubly Sparse Regularized Models

ApplicationsDGX agent

arXiv:2605.05093v1 Announce Type: cross Abstract: Regularization is often used in high-dimensional regression settings to generate a sparse model, which can save tremendous computing resources and ide

Quadrature-TreeSHAP: Depth-Independent TreeSHAP and Shapley Interactions

HardwareDGX agent

arXiv:2605.04497v1 Announce Type: new Abstract: Shapley values are a standard tool for explaining predictions of tree ensembles, with Path-Dependent SHAP being the most widely used variant. Despite su

Quantile-Free Uncertainty Quantification in Graph Neural Networks

ApplicationsDGX agent

arXiv:2605.04847v1 Announce Type: new Abstract: Uncertainty quantification (UQ) in graph neural networks (GNNs) is crucial in high-stakes domains but remains a significant challenge. In graph settings

Quantum-inspired Reinforcement Learning for Synthesizable Drug Design

Model ReleasesDGX agent

arXiv:2409.09183v2 Announce Type: replace Abstract: Synthesizable molecular design (also known as synthesizable molecular optimization) is a fundamental problem in drug discovery, and involves designi

QUIVER: Cost-Aware Adaptive Preference Querying in Surrogate-Assisted Evolutionary Multi-Objective Optimization

ResearchDGX agent

arXiv:2605.04267v1 Announce Type: new Abstract: Interactive multi-objective optimization systems face a budget allocation dilemma: one can spend resources on expensive objective evaluations or on elic

Regime-Conditioned Evaluation in Multi-Context Bayesian Optimization

Model ReleasesDGX agent

arXiv:2605.04895v1 Announce Type: new Abstract: Published transfer-BO comparisons often estimate an average treatment effect of acquisition choice over hidden regime variables, while practitioners nee

Reliable Modeling of Distribution Shifts via Displacement-Reshaped Optimal Transport

ApplicationsDGX agent

arXiv:2605.04965v1 Announce Type: new Abstract: Optimal transport (OT) is a central framework for modeling distribution shifts. Because OT compares distributions directly in input space, a well-design

Replay-Based Continual Learning for Physics-Informed Neural Operators

ResearchDGX agent

arXiv:2605.04832v1 Announce Type: new Abstract: Neural operators generally demonstrate strong predictive performance on in-distribution (ID) problems. However, a critical limitation of existing method

Rethinking Convolutional Networks for Attribute-Aware Sequential Recommendation

ApplicationsDGX agent

arXiv:2605.04723v1 Announce Type: cross Abstract: Attribute-aware sequential recommendation entails predicting the next item a user will interact with based on a chronologically ordered history of pas

← Previous
1…181182183184185…241
Next →