AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,432 results
7 May 2026

Breaking the Quality-Privacy Tradeoff in Tabular Data Generation via In-Context Learning

ApplicationsDGX agent

arXiv:2605.04911v1 Announce Type: new Abstract: Tabular data synthesis aims to generate high-quality data while preserving privacy. However, we find that existing tabular generative models exhibit a c

Bridging Input Feature Spaces Towards Graph Foundation Models

ResearchDGX agent

arXiv:2605.04834v1 Announce Type: new Abstract: Unlike vision and language domains, graph learning lacks a shared input space, as input features differ across graph datasets not only in semantics, but

Budget-aware Auto Optimizer Configurator

HardwareDGX agent

arXiv:2605.04711v1 Announce Type: cross Abstract: Optimizer states occupy massive GPU memory in large-scale model training. However, gradients in different network blocks exhibit distinct behaviors, s


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Building informative materials datasets beyond targeted objectives

ResearchDGX agent

arXiv:2605.05104v1 Announce Type: cross Abstract: Materials science data collection can be expensive, making the reuse and long-term utility of datasets critical important for future discovery campaig

Calibrating Tabular Anomaly Detection via Optimal Transport

ResearchDGX agent

arXiv:2602.06810v2 Announce Type: replace Abstract: Tabular anomaly detection (TAD) remains challenging due to the heterogeneity of tabular data: features lack natural relationships, vary widely in di

Capabilities of Auto-encoders and Principal Component Analysis of the Reduction of Microstructural Images; Application on the Acceleration of Phase-Field Simulations

ResearchDGX agent

arXiv:2605.04229v1 Announce Type: new Abstract: In this work, a data-driven framework based on Phase-Field simulations data is proposed to highlight the capabilities of neural networks to ensure accur

Capacity-Aware Mixture Law Enables Efficient LLM Data Optimization

Model ReleasesDGX agent

arXiv:2603.08022v2 Announce Type: replace Abstract: A data mixture refers to how different data sources are combined to train large language models, and selecting an effective mixture is crucial for o

Causal discovery under mean independence and linearity

SafetyDGX agent

arXiv:2605.04381v1 Announce Type: cross Abstract: Causal discovery methods such as LiNGAM identify causal structure from observational data by assuming mutually independent disturbances. This assumpti

Centrality-Based Pruning for Efficient Echo State Networks

ResearchDGX agent

arXiv:2603.20684v2 Announce Type: replace Abstract: Echo State Networks (ESNs) are a reservoir computing framework widely used for nonlinear time-series prediction. However, despite their effectivenes

Climate-based Pre-screening of Self-sustaining Regreening Opportunities in Drylands: A Case Study for Saudi Arabia

ApplicationsDGX agent

arXiv:2605.04206v1 Announce Type: new Abstract: Large-scale restoration in drylands is widely promoted to address land degradation and biodiversity loss, yet many efforts rely on long-term irrigation,

Cognitive Twins: Investigating Personalized Thinking Model Building and Its Performance Enhancement with Human-in-the-Loop

Model ReleasesDGX agent

arXiv:2605.04761v1 Announce Type: new Abstract: This paper presents the Personalized Thinking Model (PTM), a hierarchical and interpretable learner representation designed for AI supported education.

Combining Abstract Argumentation and Machine Learning for Efficiently Analyzing Low-Level Process Event Streams

ResearchDGX agent

arXiv:2505.05880v2 Announce Type: replace-cross Abstract: Monitoring and analyzing process traces is a critical task for modern companies and organizations. In scenarios where there is a gap between t

Concurrence of Symmetry Breaking and Nonlocality Phase Transitions in Diffusion Models

Local AiDGX agent

arXiv:2605.04830v1 Announce Type: new Abstract: Diffusion models undergo a phase transition in a critical time window during generation dynamics, with two complementary diagnoses of criticality. The s

Conditional Flow-VAE for Safety-Critical Traffic Scenario Generation

SafetyDGX agent

arXiv:2605.04366v1 Announce Type: cross Abstract: Safety-critical scenarios are essential for the development of autonomous vehicles (AVs) but are rare in real-world driving data. While simulation off

Conditional outlier detection for clinical alerting

ResearchDGX agent

arXiv:2605.05124v1 Announce Type: new Abstract: We develop and evaluate a data-driven approach for detecting unusual (anomalous) patient-management actions using past patient cases stored in an electr

Confronting Label Indeterminacy in Automated Bail Decisions

SafetyDGX agent

arXiv:2605.04073v1 Announce Type: new Abstract: Bail decisions present a fundamental challenge for data-driven decision support systems. When bail is denied, the counterfactual outcome of whether the

Constrained Extreme Gradient Boosting for Adapting Reduced-Order Models

Model ReleasesDGX agent

arXiv:2605.04130v1 Announce Type: new Abstract: High-fidelity simulations, such as computational fluid dynamics and finite element analysis, are essential for modeling complex engineering systems but

Constraint-Enhanced Reinforcement Learning Based on Dynamic Decoupled Spherical Radial Squashing

Model ReleasesDGX agent

arXiv:2605.04185v1 Announce Type: new Abstract: When deploying reinforcement learning policies to physical robots, actuator rate constraints -- hard limits on how fast each joint can move per control

ContextPilot: Fast Long-Context Inference via Context Reuse

AgentsDGX agent

arXiv:2511.03475v4 Announce Type: replace Abstract: AI applications increasingly depend on long-context inference, where LLMs consume substantial context to support stronger reasoning. Common examples

Contextual Memory-Enhanced Source Coding for Low-SNR Communications

ResearchDGX agent

arXiv:2605.04400v1 Announce Type: cross Abstract: While Separate Source-Channel Coding (SSCC) retains the practical benefits of modular system design, its effectiveness in noisy text transmission is f

Counter-Dyna: Data-Efficient RL-Based HVAC Control using Counterfactual Building Models

SafetyDGX agent

arXiv:2605.04555v1 Announce Type: new Abstract: Model-based reinforcement learning (MBRL) offers a promising approach for data-efficient energy management in buildings, combining the strengths of pred

Counterfactual identifiability beyond global monotonicity: non-monotone triangular structural causal models

Local AiDGX agent

arXiv:2605.04413v1 Announce Type: new Abstract: Structural causal models provide a unified semantics for interventions and counterfactuals, but most identifiability results rely on restrictive assumpt

CRAFT: Counterfactual-to-Interactive Reinforcement Fine-Tuning for Driving Policies

SafetyDGX agent

arXiv:2605.04470v1 Announce Type: new Abstract: Open-loop imitation learning has advanced modern autonomous driving policy architectures, but closed-loop deployment remains vulnerable to policy-induce

Critical Windows of Complexity Control: When Transformers Decide to Reason or Memorize

ResearchDGX agent

arXiv:2605.04396v1 Announce Type: new Abstract: Recent work has shown that Transformers' compositional generalization is governed by complexity control, initialization scale and weight decay, which st

Cross-Model Consistency of Feature Importance in Electrospinning: Separating Robust from Model-Dependent Features

Model ReleasesDGX agent

arXiv:2605.04905v1 Announce Type: new Abstract: Electrospinning is a highly sensitive fabrication process in which small variations in operating parameters can significantly influence fiber morphology

CuBridge: An LLM-Based Framework for Understanding and Reconstructing High-Performance Attention Kernels

HardwareDGX agent

arXiv:2605.05023v1 Announce Type: new Abstract: Efficient CUDA implementations of attention mechanisms are critical to modern deep learning systems, yet supporting diverse and evolving attention varia

Data-dependent Exploration for Online Reinforcement Learning from Human Feedback

SafetyDGX agent

arXiv:2605.04477v1 Announce Type: new Abstract: Online reinforcement learning from human feedback (RLHF) has emerged as a promising paradigm for aligning large language models (LLMs) by continuously c

Dataset-Driven Channel Masks in Transformers for Multivariate Time Series

ResearchDGX agent

arXiv:2410.23222v3 Announce Type: replace Abstract: Recent advancements in foundation models have been successfully extended to the time series (TS) domain, facilitated by the emergence of large-scale

Deep Wave Network for Modeling Multi-Scale Physical Dynamics

HardwareDGX agent

arXiv:2605.04198v1 Announce Type: new Abstract: Performance of deep learning models is strongly governed by architectural capacity, with width and depth as primary controls. However, in physical-scien

DeFed-GMM-DaDiL: A Decentralized Federated Framework for Domain Adaptation

ResearchDGX agent

arXiv:2605.04324v1 Announce Type: new Abstract: Decentralized multi-source domain adaptation seeks to transfer knowledge from multiple heterogeneous and related source domains to an unlabeled target d

Delving into Non-Exchangeability for Conformal Prediction in Graph-Structured Multivariate Time Series

ApplicationsDGX agent

arXiv:2605.04957v1 Announce Type: new Abstract: Point forecasting for graph-structured multivariate time series is a fundamental problem, but rigorous uncertainty quantification for such predictions i

Demystifying Manifold Constraints in LLM Pre-training

ResearchDGX agent

arXiv:2605.04418v1 Announce Type: new Abstract: The empirical success of large language model (LLM) pre-training relies heavily on heuristic stabilization techniques, such as explicit normalization la

Denoising Particle Filters: Learning State Estimation with Single-Step Objectives

ResearchDGX agent

arXiv:2602.19651v2 Announce Type: replace-cross Abstract: Learning-based methods commonly treat state estimation in robotics as a sequence modeling problem. While this paradigm can be effective at max

Deployment-Relevant Alignment Cannot Be Inferred from Model-Level Evaluation Alone

Model ReleasesDGX agent

arXiv:2605.04454v1 Announce Type: cross Abstract: Alignment evaluation in machine learning has largely become evaluation of models. Influential benchmarks score model outputs under fixed inputs, such

Designing a double deep reinforcement learning selection tool for resilient demand prediction

AgentsDGX agent

arXiv:2605.04068v1 Announce Type: new Abstract: The use of artificial intelligence in supply chain forecasting has attracted many scientific studies for several decades. However, the process of select

Differentiable Chemistry in PINNs for Solving Parameterized and Stiff Reaction Systems

Model ReleasesDGX agent

arXiv:2605.04708v1 Announce Type: new Abstract: From neural ODEs to continuous-time machine learning, differentiable solvers allow physics, optimization, and simulation to become trainable components

Discovering New Theorems via LLMs with In-Context Proof Learning in Lean

Model ReleasesDGX agent

arXiv:2509.14274v2 Announce Type: replace Abstract: Large Language Models (LLMs) have demonstrated significant promise in formal theorem proving. In this study, we investigate the ability of LLMs to d

Discovering Sparse Counterfactual Factors via Latent Adjustment for Survey-based Community Intervention

SafetyDGX agent

arXiv:2605.04460v1 Announce Type: new Abstract: Transportation surveys are widely used to understand travel preferences and adoption barriers, yet most survey-based analyses remain descriptive or pred

DistributedEstimator: Distributed Training of Quantum Neural Networks via Circuit Cutting

ResearchDGX agent

arXiv:2602.16233v2 Announce Type: replace-cross Abstract: Circuit cutting decomposes a large quantum circuit into smaller subcircuits whose outputs are classically reconstructed to recover original ex

Distributional Principal Autoencoders

ResearchDGX agent

arXiv:2404.13649v2 Announce Type: replace-cross Abstract: Dimension reduction techniques usually lose information in the sense that reconstructed data are not identical to the original data. However,

DPD-Cancer: Explainable Graph-Based Deep Learning for Small Molecule Anti-Cancer Activity Prediction

ResearchDGX agent

arXiv:2603.26114v2 Announce Type: replace Abstract: DPD-Cancer is a graph-attention deep learning framework for predicting small-molecule DPD-Cancer is a graph-attention deep learning framework for pr

Dream-MPC: Gradient-Based Model Predictive Control with Latent Imagination

SafetyDGX agent

arXiv:2605.04568v1 Announce Type: new Abstract: State-of-the-art model-based Reinforcement Learning (RL) approaches either use gradient-free, population-based methods for planning, learned policy netw

DualTCN: A Physics-Constrained Temporal Convolutional Network for 2 Time-Domain Marine CSEM Inversion

HardwareDGX agent

arXiv:2605.04997v1 Announce Type: new Abstract: DualTCN is the first deep-learning framework for inverting time-domain marine controlled-source electromagnetic (MCSEM) transient data. Moving away from

Dynamic Hyperparameter Importance for Efficient Multi-Objective Optimization

SafetyDGX agent

arXiv:2601.03166v2 Announce Type: replace Abstract: Choosing a suitable ML model is a complex task that can depend on several objectives, e.g., accuracy, fairness, or energy consumption. In practice,

DyWPE: Signal-Aware Dynamic Wavelet Positional Encoding for Time Series Transformers

ResearchDGX agent

arXiv:2509.14640v2 Announce Type: replace Abstract: Existing positional encoding methods in transformers are fundamentally signal-agnostic, deriving positional information solely from sequence indices

Echoes of the Past: A Unified Perspective on Fading memory and Echo States

ResearchDGX agent

arXiv:2508.19145v3 Announce Type: replace-cross Abstract: Recurrent neural networks (RNNs) have become increasingly popular in information processing tasks involving time series and temporal data. A f

EdgeRazor: A Lightweight Framework for Large Language Models via Mixed-Precision Quantization-Aware Distillation

ResearchDGX agent

arXiv:2605.04062v1 Announce Type: new Abstract: Recent years have witnessed an increasing interest in deploying LLMs on resource-constrained devices, among which quantization has emerged as a promisin

Efficiency of Parallel and Restart Exploration Strategies in Model Free Stochastic Simulations

SafetyDGX agent

arXiv:2503.03565v3 Announce Type: replace-cross Abstract: We analyze the efficiency of parallelization and restart mechanisms for stochastic simulations in model-free settings, where the underlying sy

Efficient Handwriting-Based Alzheimer,s Disease Diagnosis Using a Low-Rank Mixture of Experts Deep Learning Framework

ResearchDGX agent

arXiv:2605.04079v1 Announce Type: new Abstract: Early and reliable detection of Alzheimer's disease (AD) is crucial for timely clinical intervention and improved patient management. It also supports t

Efficiently Aligning Language Models with Online Natural Language Feedback

SafetyDGX agent

arXiv:2605.04356v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards has been used to elicit impressive performance from language models in many domains. But, broadly benefic

ELVIS: Ensemble-Calibrated Latent Imagination for Long-Horizon Visual MPC

ApplicationsDGX agent

arXiv:2605.04709v1 Announce Type: new Abstract: A central challenge of visual control with model-based reinforcement learning (RL) is reliable long-horizon planning: long rollouts with learned latent

Empirical Study of Pop and Jazz Mix Ratios for Genre-Adaptive Chord Generation

Model ReleasesDGX agent

arXiv:2605.04998v1 Announce Type: cross Abstract: Chord progression generation is practically important but understudied. Most large-scale symbolic music systems target melody, multi-track arrangement

Enabling Real-Time Training of a Wildfire-to-Smoke Map with Multilinear Operators

ResearchDGX agent

arXiv:2605.04164v1 Announce Type: new Abstract: Wildfires are a major producer of fine particulate matter, impacting human health and the electrical grid. Accurately forecasting smoke impacts over lon

Encoding Predictability and Legibility for Style-Conditioned Diffusion Policy

SafetyDGX agent

arXiv:2603.16368v2 Announce Type: replace-cross Abstract: Striking a balance between efficiency and transparent motion is a core challenge in human-robot collaboration, as highly expressive movements

Endogenous Regime Switching Driven by Scalar-Irreducible Learning Dynamics

AgentsDGX agent

arXiv:2605.04054v1 Announce Type: new Abstract: Achieving endogenous regime switching is crucial for the emergence of autonomous intelligence, yet remains a central challenge for existing machine lear

Enhancing the interpretability of spatially variable N2O model predictions with soft sensors during wastewater treatment

SafetyDGX agent

arXiv:2605.04082v1 Announce Type: new Abstract: Model-based solutions for nitrous oxide (N2O) emissions from wastewater treatment plants (WWTP) are informed by operational datasets designed to control

Ensuring Reliability in Programming Knowledge Tracing: A Re-evaluation of Attention-augmented Models and Experimental Protocols

ResearchDGX agent

arXiv:2605.04727v1 Announce Type: new Abstract: Programming Knowledge Tracing (PKT) has recently advanced through hybrid approaches that integrate attention-based feature modeling for code representat

Entropic Riemannian Neural Optimal Transport

ResearchDGX agent

arXiv:2605.04255v1 Announce Type: cross Abstract: Many machine learning problems involve data supported on curved spaces such as spheres, rotation groups, hyperbolic spaces, and general Riemannian man

EP-GRPO: Entropy-Progress Aligned Group Relative Policy Optimization with Implicit Process Guidance

SafetyDGX agent

arXiv:2605.04960v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR), particularly Group Relative Policy Optimization (GRPO), has advanced LLM reasoning. However, GRPO

Estimating the expected output of wide random MLPs more efficiently than sampling

TutorialsDGX agent

arXiv:2605.05179v1 Announce Type: new Abstract: By far the most common way to estimate an expected loss in machine learning is to draw samples, compute the loss on each one, and take the empirical ave

← Previous
1…179180181182183…241
Next →