AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,432 results
21 Apr 2026

Bi-LoRA: Efficient Sharpness-Aware Minimization for Fine-Tuning Large-Scale Models

Model ReleasesDGX agent

arXiv:2508.19564v2 Announce Type: replace Abstract: Fine-tuning large-scale pre-trained models with limited data presents significant challenges for generalization. While Sharpness-Aware Minimization

Bilinear Input Modulation for Mamba: Koopman Bilinear Forms for Memory Retention and Multiplicative Computation

ResearchDGX agent

arXiv:2604.17221v1 Announce Type: cross Abstract: Selective State Space Models (SSMs), notably Mamba, employ diagonal state transitions that limit both memory retention and bilinear computational capa

Bit-Flip Vulnerability of Shared KV-Cache Blocks in LLM Serving Systems

HardwareDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2604.17249v1 Announce Type: cross Abstract: Rowhammer on GPU DRAM has enabled adversarial bit flips in model weights; shared KV-cache blocks in LLM serving systems present an analogous but previ

Block-encodings as programming abstractions: The Eclipse Qrisp BlockEncoding Interface

TutorialsDGX agent

arXiv:2604.18276v1 Announce Type: cross Abstract: Block-encoding is a foundational technique in modern quantum algorithms, enabling the implementation of non-unitary operations by embedding them into

BOIL: Learning Environment Personalized Information

AgentsDGX agent

arXiv:2604.17137v1 Announce Type: new Abstract: Navigating complex environments poses challenges for multi-agent systems, requiring efficient extraction of insights from limited information. In this p

Boltzmann Machine Learning with a Parallel, Persistent Markov chain Monte Carlo method for Estimating Evolutionary Fields and Couplings from a Protein Multiple Sequence Alignment

SafetyDGX agent

arXiv:2604.18022v1 Announce Type: cross Abstract: The inverse Potts problem for estimating evolutionary single-site fields and pairwise couplings in homologous protein sequences from their single-site

Bounded Graph Clustering with Graph Neural Networks

ResearchDGX agent

arXiv:2512.05623v2 Announce Type: replace Abstract: In community detection, many methods require the user to specify the number of clusters in advance since an exhaustive search over all possible valu

Bounded Ratio Reinforcement Learning

SafetyDGX agent

arXiv:2604.18578v1 Announce Type: new Abstract: Proximal Policy Optimization (PPO) has become the predominant algorithm for on-policy reinforcement learning due to its scalability and empirical robust

Bridge-Centered Metapath Classification Using R-GCN-VGAE for Disaster-Resilient Maintenance Decisions

ApplicationsDGX agent

arXiv:2604.18399v1 Announce Type: new Abstract: Daily infrastructure management in preparation for disasters is critical for urban resilience. When bridges remain resilient against disaster-induced ex

CAARL: In-Context Learning for Interpretable Co-Evolving Time Series Forecasting

TutorialsDGX agent

arXiv:2604.18305v1 Announce Type: new Abstract: In this paper we investigate forecasting coevolving time series that feature intricate dependencies and nonstationary dynamics by using an LLM Large Lan

Can Explicit Physical Feasibility Benefit VLA Learning? An Empirical Study

SafetyDGX agent

arXiv:2604.17896v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models map multimodal inputs directly to robot actions and are typically trained through large-scale imitation learning. Wh

Can we generate portable representations for clinical time series data using LLMs?

ApplicationsDGX agent

arXiv:2603.23987v2 Announce Type: replace Abstract: Deploying clinical ML is slow and brittle: models that work at one hospital often degrade under distribution shifts at the next. In this work, we st

CAPO: Counterfactual Credit Assignment in Sequential Cooperative Teams

SafetyDGX agent

arXiv:2604.17693v1 Announce Type: new Abstract: In cooperative teams where agents act in a fixed order and share a single team reward, it is hard to know how much each agent contributed, and harder st

Capture Timing-Attention of Events in Clinical Time Series

Model ReleasesDGX agent

arXiv:2602.10385v2 Announce Type: replace Abstract: Automatically discovering personalized sequential events from large-scale time-series data is crucial for enabling precision medicine in clinical re

Causally-Constrained Probabilistic Forecasting for Time-Series Anomaly Detection

Local AiDGX agent

arXiv:2604.17998v1 Announce Type: new Abstract: Anomaly detection in multivariate time series is a central challenge in industrial monitoring, as failures frequently arise from complex temporal dynami

Central Limit Theorems for Asynchronous Averaged Q-Learning

ResearchDGX agent

arXiv:2509.18964v3 Announce Type: replace Abstract: This paper establishes central limit theorems for Polyak-Ruppert averaged Q-learning under asynchronous updates. We prove a non-asymptotic central l

Centre manifold theorem for maps along manifolds of fixed points

ResearchDGX agent

arXiv:2604.18202v1 Announce Type: cross Abstract: We prove a centre manifold theorem for a map along a manifold-with-boundary of fixed points, and provide an application to the study of gradient desce

CGCMA: Conditionally-Gated Cross-Modal Attention for Event-Conditioned Asynchronous Fusion

SafetyDGX agent

arXiv:2604.16411v1 Announce Type: new Abstract: We study asynchronous alignment, a first-class multimodal learning setting in which a dense primary stream must be fused with sporadic external context

Chronax: A Jax Library for Univariate Statistical Forecasting and Conformal Inference

ApplicationsDGX agent

arXiv:2604.16719v1 Announce Type: new Abstract: Time-series forecasting is central to many scientific and industrial domains, such as energy systems, climate modeling, finance, and retail. While forec

CLASP: Training-Free LLM-Assisted Source Code Watermarking via Semantic-Preserving Transformations

Local AiDGX agent

arXiv:2510.11251v2 Announce Type: replace-cross Abstract: The proliferation of open-source code and large language models (LLMs) for code generation has amplified the risks of unauthorized reuse and i

Clusterability-Based Assessment of Potentially Noisy Views for Multi-View Clustering

ApplicationsDGX agent

arXiv:2604.18024v1 Announce Type: new Abstract: In multi-view clustering, the quality of different views may vary substantially, and low-quality or degraded views can impair overall clustering perform

CoLLM: A Unified Framework for Co-execution of LLMs Federated Fine-tuning and Inference

Model ReleasesDGX agent

arXiv:2604.16400v1 Announce Type: cross Abstract: As Large Language Models (LLMs) are increasingly adopted in edge intelligence to power domain-specific applications and personalized services, the qua

Complex normalizing flows can be information Kahler-Ricci flows

Model ReleasesDGX agent

arXiv:2604.17954v1 Announce Type: cross Abstract: We develop interconnections between the complex normalizing flow for data drawn from Borel probability measures on the twofold realification of the co

Conditional Attribution for Root Cause Analysis in Time-Series Anomaly Detection

Model ReleasesDGX agent

arXiv:2604.17616v1 Announce Type: new Abstract: Root cause analysis (RCA) for time-series anomaly detection is critical for the reliable operation of complex real-world systems. Existing explanation m

Conformal Risk Control under Non-Monotone Losses: Theory and Finite-Sample Guarantees

Model ReleasesDGX agent

arXiv:2604.01502v2 Announce Type: replace-cross Abstract: Conformal risk control (CRC) provides distribution-free guarantees for controlling the expected loss at a user-specified level. Existing theor

ConforNets: Latents-Based Conformational Control in OpenFold3

ResearchDGX agent

arXiv:2604.18559v1 Announce Type: cross Abstract: Models from the AlphaFold (AF) family reliably predict one dominant conformation for most well-ordered proteins but struggle to capture biologically r

ConMeZO: Adaptive Descent-Direction Sampling for Gradient-Free Finetuning of Large Language Models

Model ReleasesDGX agent

arXiv:2511.02757v2 Announce Type: replace Abstract: Zeroth-order or derivative-free optimization (MeZO) is an attractive strategy for finetuning large language models (LLMs) because it eliminates the

Continual Safety Alignment via Gradient-Based Sample Selection

SafetyDGX agent

arXiv:2604.17215v1 Announce Type: new Abstract: Large language models require continuous adaptation to new tasks while preserving safety alignment. However, fine-tuning on even benign data often compr

Continuous ageing trajectory representations for knee-aware lifetime prediction of lithium-ion batteries across heterogeneous dataset

ResearchDGX agent

arXiv:2604.16580v1 Announce Type: new Abstract: Accurate assessment of lithium-ion battery ageing is challenged by cell-to-cell variability, heterogeneous cycling protocols, and limited transferabilit

Continuous Limits of Coupled Flows in Representation Learning

Model ReleasesDGX agent

arXiv:2604.16801v1 Announce Type: new Abstract: While modern representation learning relies heavily on global error signals, decentralized algorithms driven by local interactions offer a fundamental d

Contraction and Hourglass Persistence for Learning on Graphs, Simplices, and Cells

ApplicationsDGX agent

arXiv:2604.17548v1 Announce Type: new Abstract: Persistent homology (PH) encodes global information, such as cycles, and is thus increasingly integrated into graph neural networks (GNNs). PH methods i

Convergence theory for Hermite approximations under adaptive coordinate transformations

ResearchDGX agent

arXiv:2604.16975v1 Announce Type: cross Abstract: Recent work has shown that parameterizing and optimizing coordinate transformations using normalizing flows, i.e., invertible neural networks, can sig

Cooperative Coevolution versus Monolithic Evolutionary Search for Semi-Supervised Tabular Classification

SafetyDGX agent

arXiv:2604.16412v1 Announce Type: cross Abstract: This paper studies semi-supervised tabular classification in the extreme low-label regime using lightweight base learners. The paper proposes a cooper

Correction and Corruption: A Two-Rate View of Error Flow in LLM Protocols

SafetyDGX agent

arXiv:2604.18245v1 Announce Type: new Abstract: Large language models are increasingly deployed as protocols: structured multi-call procedures that spend additional computation to transform a baseline

Counterfactual Modeling with Fine-Tuned LLMs for Health Intervention Design and Sensor Data Augmentation

Model ReleasesDGX agent

arXiv:2601.14590v2 Announce Type: replace Abstract: Counterfactual explanations (CFEs) provide human-centric interpretability by identifying the minimal, actionable changes required to alter a machine

Covariance-Based Structural Equation Modeling in Small-Sample Settings with p>n

ApplicationsDGX agent

arXiv:2604.16894v1 Announce Type: new Abstract: Factor-based Structural Equation Modeling (SEM) relies on likelihood-based estimation assuming a nonsingular sample covariance matrix, which breaks down

Cross-Modal Bayesian Low-Rank Adaptation for Uncertainty-Aware Multimodal Learning

Model ReleasesDGX agent

arXiv:2604.16657v1 Announce Type: new Abstract: Large pre-trained language models are increasingly adapted to downstream tasks using parameter-efficient fine-tuning (PEFT), but existing PEFT methods a

Cross-Modal Generation: From Commodity WiFi to High-Fidelity mmWave and RFID Sensing

TutorialsDGX agent

arXiv:2604.16558v1 Announce Type: new Abstract: AIGC has shown remarkable success in CV and NLP, and has recently demonstrated promising potential in the wireless domain. However, significant data imb

D-QRELO: Training- and Data-Free Delta Compression for Large Language Models via Quantization and Residual Low-Rank Approximation

Model ReleasesDGX agent

arXiv:2604.16940v1 Announce Type: new Abstract: Supervised Fine-Tuning (SFT) accelerates taskspecific large language models (LLMs) development, but the resulting proliferation of finetuned models incu

DARLING: Detection Augmented Reinforcement Learning with Non-Stationary Guarantees

ResearchDGX agent

arXiv:2604.16684v1 Announce Type: new Abstract: We study model-free reinforcement learning (RL) in non-stationary finite-horizon episodic Markov decision processes (MDPs) without prior knowledge of th

Debate as Reward: A Multi-Agent Reward System for Scientific Ideation via RL Post-Training

SafetyDGX agent

arXiv:2604.16723v1 Announce Type: cross Abstract: Large Language Models (LLMs) have demonstrated potential in automating scientific ideation, yet current approaches relying on iterative prompting or c

Decoding AI Tutor Effects for Educational Measurement: Temporal, Multi-Outcome, and Behavior-Cognitive Analysis

SafetyDGX agent

arXiv:2604.16366v1 Announce Type: cross Abstract: Artificial intelligence (AI) tutors have become increasingly popular in learning environments. In this study, we propose an AI agent prototype framewo

Decoding RWA Tokenized U.S. Treasuries: Functional Dissection and Address Role Inference

ApplicationsDGX agent

arXiv:2507.14808v3 Announce Type: replace-cross Abstract: Tokenized U.S. Treasuries have emerged as a prominent subclass of real-world assets (RWAs), offering cryptographically secured, yield-bearing

Decomposing the Depth Profile of Fine-Tuning

ResearchDGX agent

arXiv:2604.17177v1 Announce Type: new Abstract: Fine-tuning adapts pretrained networks to new objectives. Whether the resulting depth profile of representational change reflects an intrinsic property

Deep Learning-Enhanced Calibration of the Heston Model: A Unified Framework

Model ReleasesDGX agent

arXiv:2510.24074v2 Announce Type: replace-cross Abstract: The Heston stochastic volatility model is a widely used tool in financial mathematics for pricing European options. However, its calibration r

DeepRitzSplit Neural Operator for Phase-Field Models via Energy Splitting

ResearchDGX agent

arXiv:2604.18261v1 Announce Type: cross Abstract: The multi-scale and non-linear nature of phase-field models of solidification requires fine spatial and temporal discretization, leading to long compu

DeepThinkVLA: Enhancing Reasoning Capability of Vision-Language-Action Models

SafetyDGX agent

arXiv:2511.15669v2 Announce Type: replace Abstract: Does Chain-of-Thought (CoT) reasoning genuinely improve Vision-Language-Action (VLA) models, or does it merely add overhead? Existing CoT-VLA system

Demonstrating Real Advantage of Machine-Learning-Enhanced Monte Carlo for Combinatorial Optimization

ResearchDGX agent

arXiv:2510.19544v2 Announce Type: replace-cross Abstract: Combinatorial optimization problems are central to both practical applications and the development of optimization methods. While classical an

DFedReweighting: A Unified Framework for Objective-Oriented Reweighting in Decentralized Federated Learning

SafetyDGX agent

arXiv:2512.12022v2 Announce Type: replace Abstract: Decentralized federated learning (DFL) has emerged as a promising paradigm that enables multiple clients to collaboratively train machine learning m

Differential Privacy in Two-Layer Networks: How DP-SGD Harms Fairness and Robustness

SafetyDGX agent

arXiv:2603.04881v2 Announce Type: replace Abstract: Differentially private learning is essential for training models on sensitive data, but empirical studies consistently show that it can degrade perf

Dimensional Criticality at Grokking Across MLPs and Transformers

ResearchDGX agent

arXiv:2604.16431v1 Announce Type: new Abstract: Abrupt transitions between distinct dynamical regimes are a hallmark of complex systems. Grokking in deep neural networks provides a striking example --

Dissipative Latent Residual Physics-Informed Neural Networks for Modeling and Identification of Electromechanical Systems

TutorialsDGX agent

arXiv:2604.18277v1 Announce Type: new Abstract: Accurate dynamical modeling is essential for simulation and control of embodied systems, yet first-principles models of electromechanical systems often

Distributional Off-Policy Evaluation with Deep Quantile Process Regression

SafetyDGX agent

arXiv:2604.18143v1 Announce Type: cross Abstract: This paper investigates the off-policy evaluation (OPE) problem from a distributional perspective. Rather than focusing solely on the expectation of t

Distributionally Robust Regret Optimal Control Under Moment-Based Ambiguity Sets

ResearchDGX agent

arXiv:2512.10906v2 Announce Type: replace-cross Abstract: We consider a class of finite-horizon, linear-quadratic stochastic control problems, where the probability distribution governing the noise pr

Diverse Dictionary Learning

SafetyDGX agent

arXiv:2604.17568v1 Announce Type: new Abstract: Given only observational data X = g(Z), where both the latent variables Z and the generating process g are unknown, recovering Z is ill-posed without ad

DMax: Aggressive Parallel Decoding for dLLMs

SafetyDGX agent

arXiv:2604.08302v2 Announce Type: replace Abstract: We present DMax, a new paradigm for efficient diffusion language models (dLLMs). It mitigates error accumulation in parallel decoding, enabling aggr

Do LLM-derived graph priors improve multi-agent coordination?

Model ReleasesDGX agent

arXiv:2604.17191v1 Announce Type: new Abstract: Multi-agent reinforcement learning (MARL) is crucial for AI systems that operate collaboratively in distributed and adversarial settings, particularly i

Does 'Do Differentiable Simulators Give Better Policy Gradients?'' Give Better Policy Gradients?

SafetyDGX agent

arXiv:2604.18161v1 Announce Type: new Abstract: In policy gradient reinforcement learning, access to a differentiable model enables 1st-order gradient estimation that accelerates learning compared to

Downgrade to Upgrade: Optimizer Simplification Enhances Robustness in LLM Unlearning

SafetyDGX agent

arXiv:2510.00761v5 Announce Type: replace Abstract: Large language model (LLM) unlearning aims to surgically remove the influence of undesired data or knowledge from an existing model while preserving

DR-SAC: Distributionally Robust Soft Actor-Critic for Reinforcement Learning under Uncertainty

SafetyDGX agent

arXiv:2506.12622v2 Announce Type: replace Abstract: Deep reinforcement learning (RL) has achieved remarkable success, yet its deployment in real-world scenarios is often limited by vulnerability to en

← Previous
1…214215216217218…241
Next →