Language-Induced Priors for Domain Adaptation
arXiv:2605.14301v1 Announce Type: new Abstract: Domain adaptation faces a fundamental paradox in the cold-start regime. When target data is scarce, statistical methods fail to distinguish relevant sou
Knowledge catalogue
arXiv:2605.14301v1 Announce Type: new Abstract: Domain adaptation faces a fundamental paradox in the cold-start regime. When target data is scarce, statistical methods fail to distinguish relevant sou
arXiv:2605.14524v1 Announce Type: cross Abstract: Recent studies have reported extit{saturation effects} and extit{multiple descent behavior} in large dimensional kernel ridge regression (KRR). Howeve
arXiv:2605.14241v1 Announce Type: new Abstract: Tool-augmented LLM agents increasingly access the same tool type through multiple functionally equivalent providers, such as web-search APIs, retrievers
arXiv:2510.00757v3 Announce Type: replace Abstract: Graph neural networks (GNNs) largely rely on the message-passing paradigm, where nodes iteratively aggregate information from their neighbors. Yet,
arXiv:2605.15113v1 Announce Type: new Abstract: Reinforcement learning from verifiable rewards (RLVR) suffers from sparse outcome signals, creating severe exploration bottlenecks on complex reasoning
arXiv:2601.21151v2 Announce Type: replace Abstract: Recent machine-learning approaches to weather forecasting often employ a monolithic architecture in which distinct physical mechanisms-advection (lo
arXiv:2605.14927v1 Announce Type: new Abstract: The success of deep learning in high-dimensional settings is often attributed to the presence of low-dimensional structure in real-world data. While sta
arXiv:2605.14571v1 Announce Type: cross Abstract: Observing touch on another's body can elicit corresponding tactile sensations in the observer, a phenomenon termed mirror touch that supports empathy
arXiv:2605.14186v1 Announce Type: new Abstract: Large language models (LLMs) often expose useful signals of self-monitoring: before solving a problem, they can estimate whether they are likely to succ
arXiv:2601.21929v2 Announce Type: replace Abstract: Training data attribution (TDA) identifies which training examples most influenced a model's prediction. Influence function methods are a theoretica
arXiv:2510.15141v4 Announce Type: replace-cross Abstract: Most existing manifold dimension estimators rely on the assumption that the underlying manifold is locally flat within the neighborhoods under
arXiv:2601.20173v2 Announce Type: replace Abstract: We present a new nonlinear dimensionality reduction method, MAPLE, that enhances UMAP by improving manifold modeling. MAPLE employs a self-supervise
arXiv:2509.01416v2 Announce Type: replace Abstract: The computational overhead of traditional numerical solvers for partial differential equations (PDEs) remains a critical bottleneck for large-scale
arXiv:2605.13930v1 Announce Type: new Abstract: EEG foundation models achieve state-of-the-art clinical performance, yet the internal computations driving their predictions remain opaque: a barrier to
arXiv:2605.14364v1 Announce Type: new Abstract: Continual learning requires models to adapt to new data while preserving previously acquired knowledge. At its core, this challenge can be viewed as pri
arXiv:2605.15032v1 Announce Type: cross Abstract: Intelligent Reflecting Surfaces (IRSs) are a promising technology for enhancing the spectral and energy efficiency of millimeter-wave (mmWave) multipl
arXiv:2605.14550v1 Announce Type: new Abstract: Artificial intelligence in high-stakes tabular domains cannot be evaluated by predictive performance alone, yet current practice still assesses explaina
arXiv:2602.21545v3 Announce Type: replace Abstract: Muon has recently emerged as a strong optimizer for large language model pre-training, orthogonalizing the momentum matrix via Newton--Schulz polar
arXiv:2605.14941v1 Announce Type: cross Abstract: Electroencephalogram (EEG) signals are highly susceptible to artifacts, resulting in a low signal-to-noise ratio which makes extraction of meaningful
arXiv:2605.15131v1 Announce Type: new Abstract: Reactive synthesis, the problem of automatically constructing a hardware circuit from a logical specification, is a long-standing challenge in formal ve
arXiv:2605.14343v1 Announce Type: new Abstract: Nearest-neighbor methods are fundamental to classical and modern machine learning, yet their geometric properties are typically analyzed under independe
arXiv:2602.00520v3 Announce Type: replace Abstract: Event stream data often exhibit hierarchical structure in which multiple events co-occur, resulting in a sequence of multisets (i.e., bags of events
arXiv:2605.13988v1 Announce Type: new Abstract: Inverse problems in scientific sensing are often solved with either hand-designed regularizers or supervised networks trained on simulated labels, yet b
arXiv:2602.13770v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have demonstrated strong semantic reasoning across multimodal domains. However, their integration with graph-base
arXiv:2605.13863v1 Announce Type: cross Abstract: Anomaly detection in dynamic networks is critical for applications from cybersecurity to industrial monitoring, yet existing methods face challenges i
arXiv:2510.07086v2 Announce Type: replace Abstract: Online structured prediction, including online classification as a special case, is the task of sequentially predicting labels from input features.
arXiv:2605.14260v1 Announce Type: cross Abstract: Conformal prediction is often calibrated with a single pooled threshold, but this can hide cross-group heterogeneity in score distributions and distor
arXiv:2512.16768v3 Announce Type: replace-cross Abstract: Flow matching (FM) constructs continuous-time ODE samplers by prescribing probability paths between a base distribution and a target distribut
arXiv:2510.13583v4 Announce Type: replace-cross Abstract: Causal discovery from i.i.d. observational data is known to be generally ill-posed. We demonstrate that if we have access to the distribution
arXiv:2512.01766v2 Announce Type: replace Abstract: Last-layer retraining (LLR) methods -- wherein the last layer of a neural network is reinitialized and retrained on a held-out set following ERM tra
arXiv:2507.21023v2 Announce Type: replace Abstract: Recent publications have suggested using the Shapley value for anomaly localization for sensor data systems. Using a reasonable mathematical anomaly
arXiv:2602.07519v3 Announce Type: replace Abstract: In contrast to static formalisms, computational definitions describe the operational mechanisms of a model. Simulations are an essential part of the
arXiv:2605.14240v1 Announce Type: new Abstract: The recent large-scale emergence of LLMs has left an open space for dealing with their consequences, such as plagiarism or the spread of false informati
arXiv:2605.14779v1 Announce Type: new Abstract: We propose a model-free offline multi-step reinforcement learning (RL) algorithm, Conservative Peng's Q(lambda) (CPQL). Our algorithm adapts the Peng's
arXiv:2511.19289v2 Announce Type: replace-cross Abstract: Estimating quantum entropies and divergences is an important problem in quantum physics, information theory, and machine learning. Quantum neu
arXiv:2605.13894v1 Announce Type: cross Abstract: In this work, we introduce a Tropical Axial Attention neural reasoning architecture that replaces vanilla softmax dot-product attention with max-plus
arXiv:2603.23129v2 Announce Type: replace Abstract: Godel agent realize recursive self-improvement: an agent inspects its own policy and traces and then modifies that policy in a tested loop. We intro
arXiv:2605.14888v1 Announce Type: cross Abstract: Speech-based analysis offers a scalable and non-invasive approach for detecting cognitive decline, yet progress has been constrained by the limited av
arXiv:2412.14291v2 Announce Type: replace-cross Abstract: We present a novel class of projected gradient (PG) methods for minimizing a smooth but not necessarily convex function over a convex compact
arXiv:2605.14235v1 Announce Type: new Abstract: We present an empirical evaluation of quantum entanglement in agent coordination within quantum multi agent reinforcement learning (QMARL). While QMARL
arXiv:2511.17367v2 Announce Type: replace Abstract: Computing worst-case robust strategies in pursuit-evasion games (PEGs) is time-consuming, especially when real-world factors like partial observabil
arXiv:2605.14351v1 Announce Type: cross Abstract: We present a physics-informed framework for system identification based on randomized stable atomic features. Impulse responses are represented as ran
arXiv:2605.13900v1 Announce Type: cross Abstract: In large-scale multi-agent systems with shared resource constraints, an upstream planner must iteratively evaluate candidate resource plans -- assessi
arXiv:2605.14939v1 Announce Type: cross Abstract: Reliable position and shape control in tokamak plasmas requires accurate real-time regulation of several strongly coupled shape parameters. The contro
arXiv:2605.14019v1 Announce Type: cross Abstract: Regret is the cost of uncertainty in algorithmic decision-making. Quantifying regret typically requires computationally expensive simulation via Sampl
arXiv:2605.14063v1 Announce Type: new Abstract: Continual test-time adaptation (CTTA) updates a pretrained model online on an unlabeled, non-stationary stream while anchoring it to a frozen source che
arXiv:2605.14686v1 Announce Type: new Abstract: Tabular data sharing under privacy constraints is increasingly important for research and collaboration. Synthetic data generators (SDGs) are a promisin
arXiv:2605.13932v1 Announce Type: new Abstract: Robust prediction of molecular properties under extreme out-of-distribution (OOD) scenarios is a pivotal bottleneck in AI-driven drug discovery. Current
arXiv:2512.21651v2 Announce Type: replace Abstract: Large Language Models (LLMs) deliver strong performance across a wide range of NLP tasks, but their massive sizes hinder deployment on resource-cons
arXiv:2602.05285v2 Announce Type: replace Abstract: A core challenge in structural biophysics is generating biomolecular conformations that are both physically plausible and consistent with experiment
arXiv:2605.15154v1 Announce Type: cross Abstract: Feature attribution analysis is critical for interpreting machine learning models and supporting reliable data-driven decisions. However, feature attr
arXiv:2505.09552v3 Announce Type: replace-cross Abstract: Mixed-effects models are widely used to model data with hierarchical grouping structures and high-cardinality categorical predictor variables.
arXiv:2605.14662v1 Announce Type: cross Abstract: The multi-path Traveling Salesman Problem with stochastic travel costs arises in hybrid vehicle routing applications designed for Smart City and City
arXiv:2506.20425v3 Announce Type: replace-cross Abstract: Linear mixed models (LMMs), which incorporate fixed and random effects, are key tools for analyzing heterogeneous data, such as in personalize
arXiv:2605.14567v1 Announce Type: cross Abstract: We propose a simple mechanism by which scaling laws emerge from feature learning in multi-layer networks. We study a high-dimensional hierarchical tar
arXiv:2510.23818v2 Announce Type: replace Abstract: As large language models (LLMs) continue to scale in size, the computational overhead has become a major bottleneck for task-specific fine-tuning. W
arXiv:2604.02482v2 Announce Type: replace Abstract: This paper aims to address the challenge of data generation beyond the training data and proposes a framework for Structural Extrapolated Data GEner
arXiv:2605.14551v1 Announce Type: new Abstract: Instance normalization (IN) is widely used in non-stationary multivariate time series forecasting to reduce distribution shifts and highlight common pat
arXiv:2605.14746v1 Announce Type: new Abstract: While large language models (LLMs) are trained to align with human values, their generations may still violate safety constraints. A growing line of wor
arXiv:2605.14228v1 Announce Type: cross Abstract: Background: Abilities for effective self-regulated learning (SRL) are critical for lifelong learning, particularly during adolescence when these skill