AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,557 results
10 Jun 2026

From Confident Closing to Silent Failure: Characterizing False Success in LLM Agents

AgentsDGX agent

arXiv:2606.09863v1 Announce Type: new Abstract: LLM agents can fail silently by asserting task completion when the environment state shows otherwise. We study this failure mode, false success, across

Generalized Conformal Predictive Systems Under Distributional Shifts

ResearchDGX agent

arXiv:2606.11044v1 Announce Type: cross Abstract: Conformal predictive systems (CPS) output calibrated bands of CDFs under exchangeability. We extend generalized CPS to non-exchangeable settings by en

Gradient-Guided Furthest Point Sampling for Robust Training Set Selection

TutorialsDGX agent

arXiv:2510.08906v2 Announce Type: replace-cross Abstract: Training set sampling methods are used to improve model performance and lower data costs in machine learning problems relevant to chemistry. W


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Gradient-Guided Reward Optimization for Inference-time Alignment

SafetyDGX agent

arXiv:2606.09635v1 Announce Type: cross Abstract: Ensuring the reliability of Large Language Models (LLMs) under distribution drift requires inference-time adaptation. While inference-time alignment m

GRAFT: Gain-Recalibrated Adapters for Transformer-Based Neural Population Activity Modeling

ResearchDGX agent

arXiv:2606.11066v1 Announce Type: new Abstract: Neural population activity models can recover rich temporal structure from binned spikes, but their read-in and readout layers often remain tied to a fi

Hasse Diagrams for Attention: A Partial Order Framework for Designing Transformer Masks

ResearchDGX agent

arXiv:2606.09951v1 Announce Type: new Abstract: During the training of large Transformer models, attention masks regulate the scope and direction of information flow across a sequence. Numerous mask v

Hyperparameter Learning for Latent Factorization of Tensors for Representation Learning to Large-scale Dynamic Weighted Directed Network

Model ReleasesDGX agent

arXiv:2606.09880v1 Announce Type: new Abstract: Large-scale dynamic weighted directed networks (DWDNs) are widely used to model time-varying interactions among nodes. Latent factorization of tensors (

In Defense of Cosine Similarity: Normalization Eliminates the Gauge Freedom

ResearchDGX agent

arXiv:2602.19393v2 Announce Type: replace Abstract: Steck, Ekanadham, and Kallus [arXiv:2403.05440] demonstrate that cosine similarity of learned embeddings from matrix factorization models can be ren

Influence Dynamics and Stagewise Data Attribution

TutorialsDGX agent

arXiv:2510.12071v2 Announce Type: replace Abstract: Current training data attribution (TDA) methods treat the influence one sample has on another as static, but neural networks learn in distinct stage

Informed Asymmetric Actor-Critic: Leveraging Privileged Signals Beyond Full-State Access

SafetyDGX agent

arXiv:2509.26000v3 Announce Type: replace Abstract: Asymmetric reinforcement learning leverages privileged information available during training to improve learning under partial observability. Existi

Integrating Biological-Informed Recurrent Neural Networks for Glucose-Insulin Dynamics Modeling

ResearchDGX agent

arXiv:2503.19158v3 Announce Type: replace Abstract: Type 1 Diabetes (T1D) management is a complex task due to many variability factors. Artificial Pancreas (AP) systems have alleviated patient burden

Integrating Out, Twice:The Open-System Case That Neural-Network Ensemble Theory Is Missing

ResearchDGX agent

arXiv:2606.09950v1 Announce Type: new Abstract: Averaging a neural network over its random parameters and marginalizing a Gaussian sector are the same operation, the Schur complement of the eliminated

Interpretable deep convolutional model for nonlinear multivariate time series in complex systems

Model ReleasesDGX agent

arXiv:2501.04339v2 Announce Type: replace-cross Abstract: We introduce the Deep Convolutional Interpreter for Time Series (DCIts), a deep-learning architecture for nonlinear multivariate time series t

Inverse Probability Weighting and Age-of-Information Aggregation for Decentralized Federated Learning under Partial Reception

SafetyDGX agent

arXiv:2606.10774v1 Announce Type: new Abstract: Decentralized Federated Learning (DFL) over lossy wireless networks faces two key challenges: selection bias, where updates from poor-quality links are

Ito maps for any-step SDEs

TutorialsDGX agent

arXiv:2606.11156v1 Announce Type: cross Abstract: Recent one-step generative models accelerate sampling by learning deterministic flow maps of the underlying dynamics. These methods rely on learning f

JGRA: Jacobian Geometry Robustness Assessment in NISQ Noise-Aware Quantum Neural Networks

Model ReleasesDGX agent

arXiv:2606.09964v1 Announce Type: cross Abstract: The NISQ era places stringent constraints on quantum computation, where noise and decoherence fundamentally limit performance. In classical deep learn

k-Nearest Neighbors in Gromov--Wasserstein Space

ResearchDGX agent

arXiv:2606.10295v1 Announce Type: cross Abstract: The Gromov--Wasserstein (GW) distance provides a framework for comparing metric measure spaces, regardless of their underlying structure or geometry.

Latent Guided Sampling for Combinatorial Optimization

Model ReleasesDGX agent

arXiv:2506.03672v2 Announce Type: replace-cross Abstract: Combinatorial Optimization problems are widespread in domains such as logistics, manufacturing, and drug discovery, yet their NP-hard nature m

Learning Doubly Sparse Explicitly Conditioned Transforms

ResearchDGX agent

arXiv:2606.10975v1 Announce Type: new Abstract: Finding convenient spaces in which certain hypotheses regarding an assumed sparse structure of natural signals hold true has become a desirable result i

Learning Entropy and Spatial Adaptation Dynamics of Multilayer Perceptrons for Structural Point Extraction

ApplicationsDGX agent

arXiv:2606.10170v1 Announce Type: new Abstract: This paper extends the concept of Learning Entropy (LE) from temporal adaptive systems to spatial learning in multilayer perceptron networks (MLPs) appl

Learning the Universe: Posterior Reliability of Neural Generative Models in High-Dimensional Field-Level Inference of Cosmic Initial Conditions

ApplicationsDGX agent

arXiv:2606.10023v1 Announce Type: cross Abstract: Accurate posterior estimation is central to scientific inference, as uncertainties determine what can be reliably learned from observational data. Whi

Limitations of Learning Tanh Neural Networks with Finite Precision

ResearchDGX agent

arXiv:2606.11104v1 Announce Type: new Abstract: We investigate limitations of learning anh neural networks from point evaluations under finite-precision computations and L^p accuracy guarantees, build

LLM-as-a-Discriminator: When Synthetic Tables Still Look Real

Model ReleasesDGX agent

arXiv:2606.09865v1 Announce Type: new Abstract: Privacy and data sharing are often in tension. Many organizations use synthetic data to reduce privacy risk and still share useful data. For tabular dat

LMT: A Bayesian Framework for Causal Discovery from Textual Alarm Records in Manufacturing Systems

ApplicationsDGX agent

arXiv:2606.09892v1 Announce Type: new Abstract: Textual event records, such as alarm logs, have become an increasingly common data source in engineering and manufacturing systems. Beyond identifying c

MAD: Manifold Attracted Diffusion

ResearchDGX agent

arXiv:2509.24710v2 Announce Type: replace-cross Abstract: Score-based diffusion models are a highly effective method for generating samples from a distribution of images. We consider scenarios where t

Magnetic HIP-NN for spin dynamics in disordered itinerant magnets

Model ReleasesDGX agent

arXiv:2606.10349v1 Announce Type: cross Abstract: We present a magnetic extension of the Hierarchically Interacting Particle Neural Network (HIP-NN) that enables large-scale simulations of electron-me

MemVenom: Triggered Poisoning of Multimodal Memories in Web Agents

Model ReleasesDGX agent

arXiv:2606.10742v1 Announce Type: cross Abstract: External memory has become a core component of modern web agents, enabling long-horizon reasoning through the retrieval of past experiences. However,

mlr3mbo: Bayesian Optimization in R

Model ReleasesDGX agent

arXiv:2603.29730v2 Announce Type: replace-cross Abstract: We present mlr3mbo, a modular toolbox for Bayesian optimization in R. mlr3mbo supports single- and multi-objective optimization, multi-point p

MODIP: Efficient Model-Based Optimization for Diffusion Policies

SafetyDGX agent

arXiv:2606.10825v1 Announce Type: new Abstract: Diffusion policies (DPs) have emerged as expressive policy representations for robot learning, often used with imitation learning methods such as behavi

Multi-task LLMs for Bug Classification: Efficient Inference with Auxiliary Decoding Heads

Local AiDGX agent

arXiv:2606.09956v1 Announce Type: cross Abstract: The rapid adoption of LLM-powered code generation has dramatically accelerated software development, yet effective verification methods remain severel

nCMD: Benign-Anchored Feature Selection for Imbalanced Network Intrusion Detection

Model ReleasesDGX agent

arXiv:2606.09934v1 Announce Type: new Abstract: Feature selection is critical for network intrusion detection systems (NIDS) operating under high-dimensional, highly imbalanced traffic, as found in op

Near-Exponential Convergence Rates for kNN Classification based on Boltzmann Margin

ResearchDGX agent

arXiv:2606.10361v1 Announce Type: cross Abstract: Convergence-rate analysis for classifiers is often conducted under either Tsybakov margin or Massart margin. The former is a relatively weak condition

Non-linear mechanical field reconstruction coupling recurrent neural networks with physics-informed graph neural networks

Local AiDGX agent

arXiv:2606.10909v1 Announce Type: cross Abstract: Reconstructing local stress fields in heterogeneous microstructures under non-linear, history-dependent loading remains a major computational bottlene

Nonlinear Estimator: Dual Bayesian Affine Estimators for Parameter Learning

Model ReleasesDGX agent

arXiv:2606.10111v1 Announce Type: new Abstract: This paper presents a nonlinear parameter estimator for Wiener-type state-space models obtained as a fixed-point architecture that couples two affine mi

Offline Reinforcement Learning for Rotation Profile Control in Tokamaks

SafetyDGX agent

arXiv:2605.05857v2 Announce Type: replace Abstract: Tokamaks remain leading candidates for achieving practical fusion energy, yet many important control problems inside these devices are still difficu

On-sky demonstration of reinforcement learning for adaptive optics control

SafetyDGX agent

arXiv:2606.10771v1 Announce Type: cross Abstract: Reinforcement learning (RL)-based algorithms have recently emerged as a promising approach for adaptive optics (AO) control. In simulations and labora

OncoTraj: a public benchmark for longitudinal resistance prediction in EGFR-mutant non-small-cell lung cancer on osimertinib

Model ReleasesDGX agent

arXiv:2606.11144v1 Announce Type: new Abstract: Resistance to first-line osimertinib in EGFR-mutant non-small-cell lung cancer (NSCLC) is the canonical example of predictable clonal evolution under th

One Step Closer to Ground Truth: A Multi-Scale Residual-Aware Representation Learning Pipeline for Predicting Time Series Data

Model ReleasesDGX agent

arXiv:2606.10678v1 Announce Type: new Abstract: Transformer-based models have emerged as leading paradigms in time-series forecasting in recent years, employing self-attention mechanisms to capture lo

Operator Fusion for LLM Inference on the Tensix Architecture

Local AiDGX agent

arXiv:2606.09879v1 Announce Type: new Abstract: This study addresses on-device inference bottlenecks of Transformer models on Tenstorrent's Tensix architecture and proposes an operator fusion strategy

Optimization-based Online Conformal Prediction for Multi-step Forecasting

Model ReleasesDGX agent

arXiv:2508.13362v2 Announce Type: replace Abstract: Conformal prediction (CP) is well-suited for uncertainty quantification in time series forecasting due to its distribution-free coverage guarantees.

Optuna Constrained Tree-Structured Parzen Estimator Is a Joint Density Generalization of c-TPE

ResearchDGX agent

arXiv:2606.09889v1 Announce Type: new Abstract: Constrained hyperparameter optimization (HPO) is common in practice, yet Optuna's widely used constrained TPE lacks algorithmic analysis. While c-TPE pr

Overcoming Rank Collapse in Feedback Alignment

Model ReleasesDGX agent

arXiv:2606.11123v1 Announce Type: new Abstract: Backpropagation (BP) is widely viewed as biologically implausible, in part because it requires feedback weights to be the transpose of forward weights f

Parity Cross-Resonance: A Multiqubit Gate

ResearchDGX agent

arXiv:2508.10807v2 Announce Type: replace-cross Abstract: We present a native three-qubit entangling gate that exploits engineered interactions to realize control-control-target and control-target-tar

PhysMetrics.Weather: An Evaluation Framework for Physical Consistency in ML Weather Models

ResearchDGX agent

arXiv:2606.10642v1 Announce Type: new Abstract: Machine learning weather prediction (MLWP) models have achieved impressive forecasting performance at a small fraction of the computational costs requir

PL-KKT-hPINN: Enforcing Nonlinear Equality Constraints on Neural Networks via Piecewise-Linear Projection

ApplicationsDGX agent

arXiv:2606.10682v1 Announce Type: new Abstract: While physics-informed neural networks (PINNs) have shown strong potential for process modeling, physical equations are only enforced as soft constraint

Population-Aware Physics-Informed Neural Particle Flow for Bayesian Update

ResearchDGX agent

arXiv:2606.10959v1 Announce Type: new Abstract: Physics-informed neural particle flow (PINPF) learns a deterministic transport field that moves particles from a prior distribution toward a Bayesian po

Predicting Future Behaviors in Reasoning Models Enables Better Steering

ResearchDGX agent

arXiv:2606.11172v1 Announce Type: new Abstract: Deployed large reasoning models (LRMs) often behave unexpectedly. Test-time steering controls LRM outputs by intervening on their hidden representations

PRISM: Parallel Residual Iterative Sequence Model

ResearchDGX agent

arXiv:2602.10796v3 Announce Type: replace Abstract: Generative sequence modeling faces a fundamental tension between the expressivity of Transformers and the efficiency of linear sequence models. Exis

Privacy-Preserving Credit Risk Prediction with Alternative Data

ApplicationsDGX agent

arXiv:2606.10333v1 Announce Type: new Abstract: Credit risk prediction is a critical problem in the consumer credit industry. Traditionally, financial institutions construct credit risk prediction mod

Profy: Interpretable Visualization of Expertise-Dependent Motor Skills Toward Supporting Piano Practice

ResearchDGX agent

arXiv:2606.10627v1 Announce Type: cross Abstract: The quality of piano performance depends on nuanced timing, articulation, and dynamic control, but practice feedback is often summary-based and hard t

Pushing the limits of one-dimensional NMR spectroscopy for automated structure elucidation using artificial intelligence

ResearchDGX agent

arXiv:2512.18531v2 Announce Type: replace-cross Abstract: One-dimensional NMR spectroscopy is one of the most widely used techniques for the characterization of organic compounds and natural products.

Quality Is Not a Safety Proxy Under Quantization

Model ReleasesDGX agent

arXiv:2606.10154v1 Announce Type: new Abstract: Quantized checkpoints are often screened first with quality metrics and only later, if at all, with direct safety tests. This paper audits that shortcut

Range Penalization: Theoretical Insights with Applications in Federated Learning

ResearchDGX agent

arXiv:2606.10916v1 Announce Type: cross Abstract: This paper introduces range regularization for federated learning with linear systematic components to enhance statistical accuracy and induce cross-c

Rank Collapse, Fixed Points, and the Renormalization Group Structure of MLP Residual Networks

Model ReleasesDGX agent

arXiv:2606.10324v1 Announce Type: new Abstract: The analogy between deep neural network forward passes and renormalization group (RG) flows has been repeatedly noted in the literature, but existing tr

Rethinking the Flow-Based Gradual Domain Adaptation: A Semi-Dual Optimal Transport Perspective

ResearchDGX agent

arXiv:2602.01179v2 Announce Type: replace Abstract: Gradual domain adaptation (GDA) aims to mitigate domain shift by progressively adapting models from the source domain to the target domain via inter

Revisiting Positive Samples in Graph Contrastive Learning: From the Perspective of Message Passing

ResearchDGX agent

arXiv:2606.10284v1 Announce Type: new Abstract: Graph Contrastive Learning (GCL), which trains graph encoders by maximizing similarity between positive samples and minimizing it between negative ones,

Risk Comparisons in Linear Regression: Implicit Regularization Dominates Explicit Regularization

ResearchDGX agent

arXiv:2509.17251v2 Announce Type: replace-cross Abstract: Existing theory suggests that for linear regression problems categorized by capacity and source conditions, gradient descent (GD) is always mi

Robust Active Learning for Few-Shot Example Selection in Text-to-SQL

ResearchDGX agent

arXiv:2606.10125v1 Announce Type: cross Abstract: Few-shot example retrieval is the dominant paradigm for grounding large language models (LLMs) in domain-specific text-to-SQL systems. However, the qu

Robust Regression of General ReLUs with Queries

SafetyDGX agent

arXiv:2606.11130v1 Announce Type: new Abstract: We study the task of agnostically learning general (as opposed to homogeneous) ReLUs under the Gaussian distribution with respect to the squared loss. I

Sampling Triangulations and Calabi-Yau Threefolds with Autoregressive GNNs

HardwareDGX agent

arXiv:2605.27770v2 Announce Type: cross Abstract: We introduce `dualGNN', an autoregressive message-passing GNN for sampling fine, regular triangulations (FRTs) of convex polytopes. dualGNN operates o

← Previous
1…8889909192…243
Next →