AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,557 results
21 May 2026

Cumulative Meta-Learning from Active Learning Queries for Robustness to Spurious Correlations

SafetyDGX agent

arXiv:2605.20771v1 Announce Type: new Abstract: Spurious correlations in real-world datasets cause machine learning models to rely on irrelevant patterns, undermining reliability, generalization, and

Data-Efficient Neural Operator Training via Physics-Based Active Learning

SafetyDGX agent

arXiv:2605.21348v1 Announce Type: new Abstract: Solving partial differential equations with neural operators significantly reduces computational costs but remains bottlenecked by high training data re

DC-LA: Difference-of-Convex Langevin Algorithm

ApplicationsDGX agent

arXiv:2601.22932v2 Announce Type: replace Abstract: We study a sampling problem whose target distribution is pi propto exp(-f-r) where the data fidelity term f is Lipschitz smooth while the regularize


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Decision-Path Patterns as Tree Reliability Signals: Path-based Adaptive Weighting for Random Forest Classification

SafetyDGX agent

arXiv:2605.20716v1 Announce Type: new Abstract: Random forests aggregate tree votes by simple majority, treating all trees as equally informative. We observe that the topological pattern along each tr

Decomposing MXFP4 quantization error for LLM reinforcement learning: reducible bias, recoverable deadzone, and an irreducible floor

SafetyDGX agent

arXiv:2605.20402v1 Announce Type: new Abstract: MXFP4 arithmetic can dramatically accelerate reinforcement learning (RL) post-training of large language models (LLMs), yet the quantization error intro

DeCoR: Design and Control Co-Optimization for Urban Streets Using Reinforcement Learning

SafetyDGX agent

arXiv:2605.21311v1 Announce Type: new Abstract: Modern vision systems can detect, track, and forecast urban actors at scale, yet translating perception outputs to urban design remains limited. We intr

Decoupling Communication from Policy: Robust MARL under Bandwidth Constraints

SafetyDGX agent

arXiv:2605.21085v1 Announce Type: cross Abstract: Communication enables coordination in multi-agent reinforcement learning (MARL), but many real-world applications, e.g., search-and-rescue with drone

Deep Learning Surrogates for Emulating Stochastic Climate Tipping Dynamics

ResearchDGX agent

arXiv:2605.20580v1 Announce Type: new Abstract: This work explores a dynamics-informed Temporal Fusion Transformer (TFT) as a data-driven surrogate for computationally intensive Earth system simulatio

Deep Neural Networks as Discrete Dynamical Systems: Implications for Physics-Informed Learning

Model ReleasesDGX agent

arXiv:2601.00473v3 Announce Type: replace Abstract: We revisit the analogy between feed-forward deep neural networks (DNNs) and discrete dynamical systems derived from neural integral equations and th

Design for Manufacturing: A Manufacturability Knowledge-Integrated Reinforcement Learning Framework for Free-Form Pipe Routing in Aeroengines

SafetyDGX agent

arXiv:2605.20644v1 Announce Type: new Abstract: Design for manufacturing plays a critical role in advanced aeroengine development, where complex components necessitate careful consideration of manufac

Diagnosing Overhead in Dispatch Operations: Cross-architecture Observatory

Model ReleasesDGX agent

arXiv:2605.20982v1 Announce Type: cross Abstract: AlltoAll dispatch is the dominant bottleneck of MoE expert parallelism, and the interconnect community has responded with four families of mitigations

Diffusion Models Memorize in Training -- and Generalize in Inference

Local AiDGX agent

arXiv:2603.13419v2 Announce Type: replace Abstract: Diffusion models generalize well in practice. However, an optimal diffusion model fully memorizes the training data and therefore fails to generaliz

DISC: Decoupling Instruction from State-Conditioned Control via Policy Generation

Model ReleasesDGX agent

arXiv:2605.20856v1 Announce Type: cross Abstract: Language-conditioned manipulation policies typically process instructions and observations through shared network parameters. This task-state entangle

Disentangling Bias by Modeling Intra- and Inter-modal Causal Attention for Multimodal Sentiment Analysis

SafetyDGX agent

arXiv:2508.04999v2 Announce Type: replace Abstract: Multimodal sentiment analysis (MSA) aims to understand human emotions by integrating information from multiple modalities, such as text, audio, and

Distributed Direct Preference Optimization

SafetyDGX agent

arXiv:2605.20696v1 Announce Type: new Abstract: Preference-based reinforcement learning (RL) is a key paradigm for aligning policies with human judgments, yet its theoretical behavior in distributed s

Divide and Contrast: Learning Robust Temporal Features without Augmentation

ApplicationsDGX agent

arXiv:2605.21241v1 Announce Type: new Abstract: Self-supervised learning for time-series representation aims to reduce reliance on labeled data while maintaining strong downstream performance, yet man

Divide et Calibra: Multiclass Local Calibration via Vector Quantization

Model ReleasesDGX agent

arXiv:2605.21060v1 Announce Type: new Abstract: Accurate and well-calibrated Machine Learning (ML) models are mandatory in high-stakes settings, yet effective multiclass calibration remains challengin

Domain-Adaptable Reinforcement Learning for Code Generation with Dense Rewards

SafetyDGX agent

arXiv:2605.21180v1 Announce Type: new Abstract: Large language models show strong potential for automated code generation, but lack guarantees for correctness, quality, safety, and domain-specific con

Dynamic Shapley Computation

ResearchDGX agent

arXiv:2605.20620v1 Announce Type: new Abstract: Shapley-based data valuation provides a principled way to quantify the contribution of training data, but its high computational cost makes it impractic

Dynamic TMoE: A Drift-Aware Dynamic Mixture of Experts Framework for Non-Stationary Time Series Forecasting

ResearchDGX agent

arXiv:2605.20678v1 Announce Type: new Abstract: Non-stationary time series forecasting is challenged by evolving distribution shifts that static models struggle to capture. While Mixture-of-Experts (M

E-PCN: Jet Tagging with Explainable Particle Chebyshev Networks Using Kinematic Features

ResearchDGX agent

arXiv:2512.07420v2 Announce Type: cross Abstract: The identification and classification of collimated particle sprays, or jets, are essential for interpreting data from high-energy collider experiment

ECUAS{n}: A family of metrics for principled evaluation of uncertainty-augmented systems

Model ReleasesDGX agent

arXiv:2605.20490v1 Announce Type: cross Abstract: In high-stakes automated decision-making, access to predictive uncertainty is essential for enabling users -- human or downstream systems -- to accept

Effective Model Pruning: Measure The Redundancy of Model Components

ResearchDGX agent

arXiv:2509.25606v3 Announce Type: replace Abstract: This article initiates the study of a basic question about model pruning. Given a vector s of importance scores assigned to model components, how ma

Efficient Banzhaf-Based Data Valuation for k-Nearest Neighbors Classification

ApplicationsDGX agent

arXiv:2605.21033v1 Announce Type: new Abstract: Data valuation, the task of quantifying the contribution of individual data points to model performance, has emerged as a fundamental challenge in machi

Efficient Learning of Deep State Space Models via Importance Smoothing

ResearchDGX agent

arXiv:2605.21108v1 Announce Type: new Abstract: Latent state space systems are ubiquitous in statistical modelling, arising naturally when a time series is observed through a noisy measurement functio

Efficient numeracy in language models through single-token number embeddings

TutorialsDGX agent

arXiv:2510.06824v2 Announce Type: replace Abstract: To drive progress in science and engineering, large language models (LLMs) must be able to process large amounts of numerical data and solve long ca

Enhanced Reinforcement Learning-based Process Synthesis via Quantum Computing

Model ReleasesDGX agent

arXiv:2605.21213v1 Announce Type: cross Abstract: In this work, we present quantum reinforcement learning (RL) as a solution strategy for process synthesis problems. Building on our prior work, we dev

Ensemble RL through Classifier Models: Enhancing Risk-Return Trade-offs in Trading Strategies

ResearchDGX agent

arXiv:2502.17518v2 Announce Type: replace Abstract: This paper presents a comprehensive study on the use of ensemble Reinforcement Learning (RL) models in financial trading strategies, leveraging clas

Epistemic Uncertainty Quantification for Pre-trained VLMs via Riemannian Flow Matching

ResearchDGX agent

arXiv:2601.21662v2 Announce Type: replace Abstract: Vision-Language Models (VLMs) are typically deterministic in nature and lack intrinsic mechanisms to quantify epistemic uncertainty, which reflects

Equilibrium Reasoners: Learning Attractors Enables Scalable Reasoning

ResearchDGX agent

arXiv:2605.21488v1 Announce Type: new Abstract: Scaling test-time compute by iteratively updating a latent state has emerged as a powerful paradigm for reasoning. Yet the internal mechanisms that enab

Everywhere Valid Bounds on False Discovery Proportions in Conformal Inference

ResearchDGX agent

arXiv:2605.20726v1 Announce Type: cross Abstract: Modern applications of conformal inference to multiple testing problems, such as outlier detection and candidate selection, often involve selecting te

Evolutionary Generation of Multi-Agent Systems

Model ReleasesDGX agent

arXiv:2602.06511v3 Announce Type: replace Abstract: Large language model (LLM)-based multi-agent systems (MAS) show strong promise for complex reasoning, planning, and tool-augmented tasks, but design

EvoStruct: Bridging Evolutionary and Structural Priors for Antibody CDR Design via Protein Language Model Adaptation

ResearchDGX agent

arXiv:2605.21485v1 Announce Type: new Abstract: Equivariant graph neural network (GNN) methods for antibody complementarity-determining region (CDR) design achieve the highest sequence recovery but su

Explainability Methods for Hardware Trojan Detection: A Systematic Comparison

Model ReleasesDGX agent

arXiv:2601.18696v4 Announce Type: replace Abstract: Hardware trojans are malicious circuits which compromise the functionality and security of an integrated circuit (IC). These circuits are manufactur

extit{Stochastic} MeanFlow Policies: One-Step Generative Control with Entropic Mirror Descent

SafetyDGX agent

arXiv:2605.21282v1 Announce Type: new Abstract: Online off-policy reinforcement learning (RL) is shaped by two coupled choices: the policy class and the update rule. Gaussian policies are fast and hav

FAIR-Pruner: A Flexible Framework for Automatic Layer-Wise Pruning via Tolerance of Difference

Model ReleasesDGX agent

arXiv:2508.02291v3 Announce Type: replace Abstract: Structured pruning is a standard tool for compressing deep neural networks, but its practical performance depends on how sparsity is allocated acros

Fast and Stable Triangular Inversion for Delta-Rule Linear Transformers

ResearchDGX agent

arXiv:2605.21325v1 Announce Type: new Abstract: Linear attention has emerged as a cornerstone for efficient long-context architectures, as evidenced by its integration into state-of-the-art open-sourc

Fast Reconstruction of Exact Maxwell Dynamics from Sparse Data

ResearchDGX agent

arXiv:2605.20514v1 Announce Type: new Abstract: We introduce FLASH-MAX, a shallow, exact-by-construction neural network architecture for predicting homogeneous electromagnetic fields from sparse point

FBOS-RL: Feedback-Driven Bi-Objective Synergistic Reinforcement Learning

Model ReleasesDGX agent

arXiv:2605.20256v1 Announce Type: new Abstract: Reinforcement learning has become a cornerstone for aligning and unlocking the reasoning capabilities of large-scale models. At its core, the training l

FEAT: A Linear-Complexity Foundation Model for Extremely Large Structured Data

SafetyDGX agent

arXiv:2603.16513v3 Announce Type: replace Abstract: Structured data is widely used in domains such as healthcare, finance, and scientific data management. Recent studies on structured data foundation

FedCoE: Bridging Generalization and Personalization via Federated Coordinated Dual-level MoEs

Model ReleasesDGX agent

arXiv:2605.21264v1 Announce Type: new Abstract: Federated Learning (FL) has emerged as a promising paradigm for privacy-preserving distributed learning. However, existing FL methods face a fundamental

Federated Learning of Nonlinear Temporal Dynamics with Graph Attention-based Cross-Client Interpretability

Local AiDGX agent

arXiv:2602.13485v2 Announce Type: replace Abstract: Networks of modern industrial systems are increasingly monitored by distributed sensors, where each system comprises multiple subsystems generating

Federated LoRA Fine-Tuning for LLMs via Collaborative Alignment

Model ReleasesDGX agent

arXiv:2605.21217v1 Announce Type: cross Abstract: Low-rank adaptation (LoRA) has emerged as a powerful tool for parameter-efficient fine-tuning of large language models (LLMs). This paper studies LoRA

Finite-Time Regret Analysis of Retry-Aware Bandits

ResearchDGX agent

arXiv:2605.20854v1 Announce Type: new Abstract: We study a stochastic bandit algorithm motivated by retry-aware objectives that value the best outcome among multiple attempts, such as pass@k and max@k

FLUME-FNO: data-efficient and scalable prediction of 3D wind and temperature fields in unseen urban morphologies

Local AiDGX agent

arXiv:2503.19708v2 Announce Type: replace-cross Abstract: Urban microclimate, encompassing wind and temperature fields shaped by building geometry, significantly impacts energy consumption, pedestrian

For How Long Should We Be Punching? Learning Action Duration in Fighting Games

AgentsDGX agent

arXiv:2605.20911v1 Announce Type: cross Abstract: Fighting games such as Street Fighter II present unique challenges to reinforcement learning (RL) agents due to their fast-paced, real-time nature. In

From Circuit Evidence to Mechanistic Theory: An Inductive Logic Approach

ResearchDGX agent

arXiv:2605.21303v1 Announce Type: new Abstract: Mechanistic interpretability produces circuit-level causal analyses of neural network behaviour, but discovered circuits often remain isolated experimen

Frontier: Towards Comprehensive and Accurate LLM Inference Simulation

HardwareDGX agent

arXiv:2605.21312v1 Announce Type: cross Abstract: Modern LLM serving is no longer homogeneous or monolithic. Production systems now combine disaggregated execution, complex parallelism, runtime optimi

FT-Dojo: Towards Autonomous LLM Fine-Tuning with Language Agents

Model ReleasesDGX agent

arXiv:2603.01712v2 Announce Type: replace-cross Abstract: Fine-tuning large language models for vertical domains remains labor-intensive, requiring practitioners to curate data, configure training, an

Gaussian Sheaf Neural Networks

ApplicationsDGX agent

arXiv:2605.21435v1 Announce Type: new Abstract: Graph Neural Networks (GNNs) have become the de facto standard for learning on relational data. While traditional GNNs' message passing is well suited f

GenAI-Driven Threat Detection with Microsoft Security Copilot

Model ReleasesDGX agent

arXiv:2605.20896v1 Announce Type: cross Abstract: Defending against today's increasingly sophisticated cyberattacks requires security analysts to continuously translate evolving attacker tradecraft in

Genetic Programming with Transformer-Based Mutation for Approximate Circuit Design

ResearchDGX agent

arXiv:2605.21055v1 Announce Type: cross Abstract: A recent trend is to leverage machine learning models to improve the evolutionary design and optimization process. We propose a novel transformer-base

GeoPT: Scaling Physics Simulation via Lifted Geometric Pre-Training

ResearchDGX agent

arXiv:2602.20399v2 Announce Type: replace Abstract: Neural simulators promise efficient surrogates for physics simulation, but scaling them is bottlenecked by the prohibitive cost of generating high-f

Gradient Scalability and Taylor Surrogation of Quantum Cost Landscapes

ResearchDGX agent

arXiv:2507.06344v3 Announce Type: replace-cross Abstract: Variational Quantum Algorithms are promising candidates for near-term quantum computing, yet they face scalability challenges due to barren pl

GradPower: Powering Gradients for Faster Language Model Pre-Training

Model ReleasesDGX agent

arXiv:2505.24275v3 Announce Type: replace Abstract: We propose GradPower, a lightweight gradient-transformation technique for accelerating language model pre-training. Given a gradient vector g=(g_i)_

Graph Autoencoder for Process Monitoring

ApplicationsDGX agent

arXiv:2602.03004v2 Announce Type: replace Abstract: To improve the reliability and interpretability of industrial process monitoring, this article proposes a Causal Graph Spatial-Temporal Autoencoder

Graph Navier Stokes Networks

ApplicationsDGX agent

arXiv:2605.21247v1 Announce Type: new Abstract: Graph Neural Networks (GNNs) have emerged as a cornerstone of deep learning, with most existing methods rooted in graph signal processing and diffusion

Graph Transductive Sharpening: Leveraging Unlabeled Predictions in Node Classification

SafetyDGX agent

arXiv:2605.20248v1 Announce Type: new Abstract: In the transductive setting, where the full graph is observed but node labels are only partially available, progress in semi-supervised node classificat

GraphCSVAE: Graph Categorical Structured Variational Autoencoder for Spatiotemporal Auditing of Physical Vulnerability Towards Sustainable Post-Disaster Risk Reduction

ResearchDGX agent

arXiv:2509.10308v2 Announce Type: replace Abstract: In the aftermath of disasters, many institutions worldwide face challenges in monitoring changes in disaster risk, limiting assessment of progress t

GraphDiffMed: Knowledge-Constrained Differential Attention with Pharmacological Graph Priors for Medication Recommendation

SafetyDGX agent

arXiv:2605.20188v1 Announce Type: new Abstract: Recommending safe and effective medication combinations from electronic health records (EHRs) is a core clinical AI problem, yet it remains difficult be

← Previous
1…138139140141142…243
Next →