AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlog
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,329 results
Model Releases

Entropy After </Think> for reasoning model early exiting

DGX agent

arXiv:2509.26522v3 Announce Type: replace Abstract: Reasoning LLMs show improved performance with longer chains of thought. However, recent work has highlighted their tendency to overthink, continuing

model-releasesarxiv-cs-lg
10 Apr 2026
Model Releases

Epistemic Robust Offline Reinforcement Learning

X Post
Paper
YouTube
Reddit
GitHub
Clear filters
DGX agent

arXiv:2604.07072v1 Announce Type: new Abstract: Offline reinforcement learning learns policies from fixed datasets without further environment interaction. A key challenge in this setting is epistemic

model-releasesarxiv-cs-lg
10 Apr 2026
Safety

Equivariant Multi-agent Reinforcement Learning for Multimodal Vehicle-to-Infrastructure Systems

DGX agent

arXiv:2604.06914v1 Announce Type: new Abstract: In this paper, we study a vehicle-to-infrastructure (V2I) system where distributed base stations (BSs) acting as road-side units (RSUs) collect multimod

safetyarxiv-cs-lg
10 Apr 2026
Research

Evaluating PQC KEMs, Combiners, and Cascade Encryption via Adaptive IND-CPA Testing Using Deep Learning

DGX agent

arXiv:2604.06942v1 Announce Type: cross Abstract: Ensuring ciphertext indistinguishability is fundamental to cryptographic security, but empirically validating this property in real implementations an

researcharxiv-cs-lg
10 Apr 2026
Research

EvoFlows: Evolutionary Edit-Based Flow-Matching for Protein Engineering

DGX agent

arXiv:2603.11703v2 Announce Type: replace Abstract: We introduce EvoFlows, a variable-length protein sequence-to-sequence modeling approach designed for protein engineering. Existing protein language

researcharxiv-cs-lg
10 Apr 2026
Safety

Explainable AI to Improve Machine Learning Reliability for Industrial Cyber-Physical Systems

DGX agent

arXiv:2601.16074v2 Announce Type: replace Abstract: Industrial Cyber-Physical Systems (CPS) are sensitive infrastructure from both safety and economics perspectives, making their reliability criticall

safetyarxiv-cs-lg
10 Apr 2026
Tutorials

ExplainFuzz: Explainable and Constraint-Conditioned Test Generation with Probabilistic Circuits

DGX agent

arXiv:2604.06559v1 Announce Type: cross Abstract: Understanding and explaining the structure of generated test inputs is essential for effective software testing and debugging. Existing approaches--in

tutorialsarxiv-cs-lg
10 Apr 2026
Research

Extraction of linearized models from pre-trained networks via knowledge distillation

DGX agent

arXiv:2604.06732v1 Announce Type: new Abstract: Recent developments in hardware, such as photonic integrated circuits and optical devices, are driving demand for research on constructing machine learn

researcharxiv-cs-lg
10 Apr 2026
Local Ai

Fast reconstruction-based ROI triggering via anomaly detection in the CYGNO optical TPC

DGX agent

arXiv:2512.24290v2 Announce Type: replace-cross Abstract: Optical-readout Time Projection Chambers (TPCs) produce megapixel-scale images whose fine-grained topological information is essential for rar

local-aiarxiv-cs-lg
10 Apr 2026
Research

Fast Spatial Memory with Elastic Test-Time Training

DGX agent

arXiv:2604.07350v1 Announce Type: cross Abstract: Large Chunk Test-Time Training (LaCT) has shown strong performance on long-context 3D reconstruction, but its fully plastic inference-time updates rem

researcharxiv-cs-lg
10 Apr 2026
Local Ai

FedDetox: Robust Federated SLM Alignment via On-Device Data Sanitization

DGX agent

arXiv:2604.06833v1 Announce Type: cross Abstract: As high quality public data becomes scarce, Federated Learning (FL) provides a vital pathway to leverage valuable private user data while preserving p

local-aiarxiv-cs-lg
10 Apr 2026
Model Releases

FedSpy-LLM: Towards Scalable and Generalizable Data Reconstruction Attacks from Gradients on LLMs

DGX agent

arXiv:2604.06297v1 Announce Type: cross Abstract: Given the growing reliance on private data in training Large Language Models (LLMs), Federated Learning (FL) combined with Parameter-Efficient Fine-Tu

model-releasesarxiv-cs-lg
10 Apr 2026
Model Releases

FlowAdam: Implicit Regularization via Geometry-Aware Soft Momentum Injection

DGX agent

arXiv:2604.06652v1 Announce Type: new Abstract: Adaptive moment methods such as Adam use a diagonal, coordinate-wise preconditioner based on exponential moving averages of squared gradients. This diag

model-releasesarxiv-cs-lg
10 Apr 2026
Hardware

ForkKV: Scaling Multi-LoRA Agent Serving via Copy-on-Write Disaggregated KV Cache

DGX agent

arXiv:2604.06370v1 Announce Type: cross Abstract: The serving paradigm of large language models (LLMs) is rapidly shifting towards complex multi-agent workflows where specialized agents collaborate ov

hardwarearxiv-cs-lg
10 Apr 2026
Hardware

Foundry: Template-Based CUDA Graph Context Materialization for Fast LLM Serving Cold Start

DGX agent

arXiv:2604.06664v1 Announce Type: cross Abstract: Modern LLM service providers increasingly rely on autoscaling and parallelism reconfiguration to respond to rapidly changing workloads, but cold-start

hardwarearxiv-cs-lg
10 Apr 2026
Local Ai

From Synthetic Data to Real Restorations: Diffusion Model for Patient-specific Dental Crown Completion

DGX agent

arXiv:2603.26588v2 Announce Type: replace-cross Abstract: We present ToothCraft, a diffusion-based model for the contextual generation of tooth crowns, trained on artificially created incomplete teeth

local-aiarxiv-cs-lg
10 Apr 2026
Research

Gaussian Approximation for Asynchronous Q-learning

DGX agent

arXiv:2604.07323v1 Announce Type: cross Abstract: In this paper, we derive rates of convergence in the high-dimensional central limit theorem for Polyak-Ruppert averaged iterates generated by the asyn

researcharxiv-cs-lg
10 Apr 2026
Safety

Geometric Properties of the Voronoi Tessellation in Latent Semantic Manifolds of Large Language Models

DGX agent

arXiv:2604.06767v1 Announce Type: new Abstract: Language models operate on discrete tokens but compute in continuous vector spaces, inducing a Voronoi tessellation over the representation manifold. We

safetyarxiv-cs-lg
10 Apr 2026
Safety

GIFT: Group-Relative Implicit Fine-Tuning Integrates GRPO with DPO and UNA

DGX agent

arXiv:2510.23868v4 Announce Type: replace Abstract: This paper proposes extit{Group-relative Implicit Fine-Tuning (GIFT)}, a reinforcement learning framework for aligning large language models (LLMs

safetyarxiv-cs-lg
10 Apr 2026
Hardware

Graph Neural ODE Digital Twins for Control-Oriented Reactor Thermal-Hydraulic Forecasting Under Partial Observability

DGX agent

arXiv:2604.07292v1 Announce Type: new Abstract: Real-time supervisory control of advanced reactors requires accurate forecasting of plant-wide thermal-hydraulic states, including locations where physi

hardwarearxiv-cs-lg
10 Apr 2026
Applications

GraphWalker: Graph-Guided In-Context Learning for Clinical Reasoning on Electronic Health Records

DGX agent

arXiv:2604.06684v1 Announce Type: new Abstract: Clinical Reasoning on Electronic Health Records (EHRs) is a fundamental yet challenging task in modern healthcare. While in-context learning (ICL) offer

applicationsarxiv-cs-lg
10 Apr 2026
Model Releases

Holistic Optimal Label Selection for Robust Prompt Learning under Partial Labels

DGX agent

arXiv:2604.06614v1 Announce Type: cross Abstract: Prompt learning has gained significant attention as a parameter-efficient approach for adapting large pre-trained vision-language models to downstream

model-releasesarxiv-cs-lg
10 Apr 2026
Research

How Does Machine Learning Manage Complexity?

DGX agent

arXiv:2604.07233v1 Announce Type: new Abstract: We provide a computational complexity lens to understand the power of machine learning models, particularly their ability to model complex systems. Mach

researcharxiv-cs-lg
10 Apr 2026
Tutorials

How to sketch a learning algorithm

DGX agent

arXiv:2604.07328v1 Announce Type: new Abstract: How does the choice of training data influence an AI model? This question is of central importance to interpretability, privacy, and basic science. At i

tutorialsarxiv-cs-lg
10 Apr 2026
Safety

Improving Semantic Uncertainty Quantification in Language Model Question-Answering via Token-Level Temperature Scaling

DGX agent

arXiv:2604.07172v1 Announce Type: new Abstract: Calibration is central to reliable semantic uncertainty quantification, yet prior work has largely focused on discrimination, neglecting calibration. As

safetyarxiv-cs-lg
10 Apr 2026
Tutorials

Inference-Time Scaling of Diffusion Language Models via Trajectory Refinement

DGX agent

arXiv:2507.08390v4 Announce Type: replace Abstract: Discrete diffusion models have recently emerged as strong alternatives to autoregressive language models, matching their performance through large-s

tutorialsarxiv-cs-lg
10 Apr 2026
Research

Interventional Time Series Priors for Causal Foundation Models

DGX agent

arXiv:2603.11090v2 Announce Type: replace Abstract: Prior-data fitted networks (PFNs) have emerged as powerful foundation models for tabular causal inference, yet their extension to time series remain

researcharxiv-cs-lg
10 Apr 2026
Model Releases

Learning Debt and Cost-Sensitive Bayesian Retraining: A Forecasting Operations Framework

DGX agent

arXiv:2604.06438v1 Announce Type: cross Abstract: Forecasters often choose retraining schedules by convention rather than by an explicit decision rule. This paper gives that decision a posterior-space

model-releasesarxiv-cs-lg
10 Apr 2026
Local Ai

Learning to Query History: Nonstationary Classification via Learned Retrieval

DGX agent

arXiv:2604.07027v1 Announce Type: new Abstract: Nonstationarity is ubiquitous in practical classification settings, leading deployed models to perform poorly even when they generalize well to holdout

local-aiarxiv-cs-lg
10 Apr 2026
Safety

Limits of Difficulty Scaling: Hard Samples Yield Diminishing Returns in GRPO-Tuned SLMs

DGX agent

arXiv:2604.06298v1 Announce Type: new Abstract: Recent alignment work on Large Language Models (LLMs) suggests preference optimization can improve reasoning by shifting probability mass toward better

safetyarxiv-cs-lg
10 Apr 2026
Model Releases

LNN-PINN: A Unified Physics-Only Training Framework with Liquid Residual Blocks

DGX agent

arXiv:2508.08935v4 Announce Type: replace Abstract: Physics-informed neural networks (PINNs) have attracted considerable attention for their ability to integrate partial differential equation priors i

model-releasesarxiv-cs-lg
10 Apr 2026
Model Releases

LoFT: Parameter-Efficient Fine-Tuning for Long-tailed Semi-Supervised Learning in Open-World Scenarios

DGX agent

arXiv:2509.09926v5 Announce Type: replace Abstract: Long-tailed semi-supervised learning (LTSSL) presents a formidable challenge where models must overcome the scarcity of tail samples while mitigatin

model-releasesarxiv-cs-lg
10 Apr 2026
Research

Low-Rank Key Value Attention

DGX agent

arXiv:2601.11471v3 Announce Type: replace Abstract: The key-value (KV) cache is a primary memory bottleneck in Transformers. We propose Low-Rank Key-Value (LRKV) attention, which reduces KV cache memo

researcharxiv-cs-lg
10 Apr 2026
Model Releases

Lumbermark: Resistant Clustering by Chopping Up Mutual Reachability Minimum Spanning Trees

DGX agent

arXiv:2604.07143v1 Announce Type: new Abstract: We introduce Lumbermark, a robust divisive clustering algorithm capable of detecting clusters of varying sizes, densities, and shapes. Lumbermark iterat

model-releasesarxiv-cs-lg
10 Apr 2026
Safety

LUMINA: Foundation Models for Topology Transferable ACOPF

DGX agent

arXiv:2603.04300v2 Announce Type: replace Abstract: Foundation models in general promise to accelerate scientific computation by learning reusable representations across problem instances, yet constra

safetyarxiv-cs-lg
10 Apr 2026
Model Releases

Matrix Profile for Time-Series Anomaly Detection: A Reproducible Open-Source Benchmark on TSB-AD

DGX agent

arXiv:2604.02445v2 Announce Type: replace Abstract: Matrix Profile (MP) methods are an interpretable and scalable family of distance-based methods for time-series anomaly detection, but strong benchma

model-releasesarxiv-cs-lg
10 Apr 2026
Safety

MDP modeling for multi-stage stochastic programs

DGX agent

arXiv:2509.22981v2 Announce Type: replace Abstract: We study a class of multi-stage stochastic programs, which incorporate modeling features from Markov decision processes (MDPs). This class includes

safetyarxiv-cs-lg
10 Apr 2026
Hardware

Measurement of Generative AI Workload Power Profiles for Whole-Facility Data Center Infrastructure Planning

DGX agent

arXiv:2604.07345v1 Announce Type: cross Abstract: The rapid growth of generative artificial intelligence (AI) has introduced unprecedented computational demands, driving significant increases in the e

hardwarearxiv-cs-lg
10 Apr 2026
Agents

MedRoute: RL-Based Dynamic Specialist Routing in Multi-Agent Medical Diagnosis

DGX agent

arXiv:2604.06180v1 Announce Type: cross Abstract: Medical diagnosis using Large Multimodal Models (LMMs) has gained increasing attention due to capability of these models in providing precise diagnose

agentsarxiv-cs-lg
10 Apr 2026
Research

MENO: MeanFlow-Enhanced Neural Operators for Dynamical Systems

DGX agent

arXiv:2604.06881v1 Announce Type: new Abstract: Neural operators have emerged as powerful surrogates for dynamical systems due to their grid-invariant properties and computational efficiency. However,

researcharxiv-cs-lg
10 Apr 2026
Model Releases

MF-GLaM: A multifidelity stochastic emulator using generalized lambda models

DGX agent

arXiv:2507.10303v2 Announce Type: replace-cross Abstract: Stochastic simulators exhibit intrinsic stochasticity due to unobservable, uncontrollable, or unmodeled input variables, resulting in random o

model-releasesarxiv-cs-lg
10 Apr 2026
Research

MICA: Multivariate Infini Compressive Attention for Time Series Forecasting

DGX agent

arXiv:2604.06473v1 Announce Type: new Abstract: Multivariate forecasting with Transformers faces a core scalability challenge: modeling cross-channel dependencies via attention compounds attention's q

researcharxiv-cs-lg
10 Apr 2026
Applications

Mining Electronic Health Records to Investigate Effectiveness of Ensemble Deep Clustering

DGX agent

arXiv:2604.07085v1 Announce Type: new Abstract: In electronic health records (EHRs), clustering patients and distinguishing disease subtypes are key tasks to elucidate pathophysiology and aid clinical

applicationsarxiv-cs-lg
10 Apr 2026
Research

MoE Routing Testbed: Studying Expert Specialization and Routing Behavior at Small Scale

DGX agent

arXiv:2604.07030v1 Announce Type: new Abstract: Sparse Mixture-of-Experts (MoE) architectures are increasingly popular for frontier large language models (LLM) but they introduce training challenges d

researcharxiv-cs-lg
10 Apr 2026
Safety

Multi-Turn Reasoning LLMs for Task Offloading in Mobile Edge Computing

DGX agent

arXiv:2604.07148v1 Announce Type: new Abstract: Emerging computation-intensive applications impose stringent latency requirements on resource-constrained mobile devices. Mobile Edge Computing (MEC) ad

safetyarxiv-cs-lg
10 Apr 2026
Local Ai

NativeTernary: A Self-Delimiting Binary Encoding with Unary Run-Length Hierarchy Markers for Ternary Neural Network Weights, Structured Data, and General Computing Infrastructure

DGX agent

arXiv:2604.03336v2 Announce Type: replace Abstract: BitNet b1.58 (Ma et al., 2024) demonstrates that large language models can operate entirely on ternary weights {-1, 0, +1}, yet no native binary wir

local-aiarxiv-cs-lg
10 Apr 2026
Model Releases

Negative Binomial Variational Autoencoders for Overdispersed Latent Modeling

DGX agent

arXiv:2508.05423v2 Announce Type: replace Abstract: Although artificial neural networks are often described as brain-inspired, their representations typically rely on continuous activations, such as t

model-releasesarxiv-cs-lg
10 Apr 2026
Hardware

NestPipe: Large-Scale Recommendation Training on 1,500+ Accelerators via Nested Pipelining

DGX agent

arXiv:2604.06956v1 Announce Type: cross Abstract: Modern recommendation models have increased to trillions of parameters. As cluster scales expand to O(1k), distributed training bottlenecks shift from

hardwarearxiv-cs-lg
10 Apr 2026
← Previous
1…295296297298299
Next →