AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,557 results
2 Jun 2026

Benchmarking Waitlist Mortality Prediction in Heart Transplantation Through Time-to-Event Modeling using New Longitudinal UNOS Dataset

Model ReleasesDGX agent

arXiv:2507.07339v2 Announce Type: replace-cross Abstract: Decisions about managing patients on the heart transplant waitlist are currently made by committees of doctors who consider multiple factors,

BERT4beam: Large AI Model Enabled Generalized Beamforming Optimization

ResearchDGX agent

arXiv:2509.11056v2 Announce Type: replace-cross Abstract: Artificial intelligence (AI) is anticipated to emerge as a pivotal enabler for the forthcoming sixth-generation (6G) wireless communication sy

Beyond Discreteness: Sample Complexity Analysis of Straight-Through Estimator for 1-bit Quantization

ResearchDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2505.18113v2 Announce Type: replace Abstract: Training quantized neural networks requires addressing the non-differentiable and discrete nature of the underlying optimization problem. To tackle

Beyond ell_2-norm and ell_infty-norm: A Curvature-Inspired ell_p-Norm Scheme for Deep Neural Networks

Model ReleasesDGX agent

arXiv:2606.02078v1 Announce Type: new Abstract: The existing optimizers for deep neural networks (DNNs) typically rely on either the ell_2 norm or the ell_infty norm, resulting in optimizers that do n

Beyond Procedure: Substantive Fairness in Conformal Prediction

SafetyDGX agent

arXiv:2602.16794v2 Announce Type: replace-cross Abstract: Conformal prediction (CP) offers distribution-free uncertainty quantification for machine learning models, yet its interplay with fairness in

Bit-Exact AI Inference Verification Without Performance Tradeoffs

HardwareDGX agent

arXiv:2606.00279v1 Announce Type: cross Abstract: Verifying claims about AI workloads is a pre- requisite for credible AI governance of covert adversaries (who comply with monitoring only when detecti

BLISS: A Lightweight Bilevel Influence Scoring Method for Data Selection in Language Model Pretraining

Model ReleasesDGX agent

arXiv:2510.06048v4 Announce Type: replace Abstract: Effective data selection is essential for pretraining large language models (LLMs), enhancing efficiency and improving generalization to downstream

BlockGen: Flexible Blockwise Sequence Modeling with Hybrid Samplers

ResearchDGX agent

arXiv:2606.02241v1 Announce Type: new Abstract: Is the uniform-state diffusion framework a more powerful paradigm for discrete diffusion? Recent studies indicate that this may be the case. In combinat

Both Topology and Text Matter: Revisiting LLM-guided Out-of-Distribution Detection on Text-attributed Graphs

SafetyDGX agent

arXiv:2602.11641v2 Announce Type: replace Abstract: Text-attributed graphs (TAGs) associate nodes with textual attributes and graph structure, enabling GNNs to jointly model semantic and structural in

Byte Pair Encoding for Efficient Time Series Forecasting

ResearchDGX agent

arXiv:2505.14411v4 Announce Type: replace Abstract: Existing time series tokenization methods predominantly encode a constant number of samples into individual tokens. This inflexible approach can gen

Can Vision Language Models Learn Intuitive Physics from Interaction?

TutorialsDGX agent

arXiv:2602.06033v2 Announce Type: replace Abstract: Pre-trained vision language models do not have good intuitions about the physical world. Recent work has shown that supervised fine-tuning can impro

CANARY: Zero-Label Detection of Fine-Tuning Contamination in Language Models

ResearchDGX agent

arXiv:2606.01695v1 Announce Type: new Abstract: Adversaries can implant latent harmful behavior by poisoning as few as 1% of fine-tuning examples. The contamination is invisible to every output-level

Canonicalized Stable-List Replay for Private Federated Continual Learning over Language-Model Embeddings

ResearchDGX agent

arXiv:2606.00426v1 Announce Type: new Abstract: Federated continual learning (FCL) lets distributed clients adapt language-model heads to evolving NLP tasks without sharing raw text. Under user-level

Cellular Sheaf Neural Operators for Structure-Preserving Surrogate Modeling of Constrained PDEs

SafetyDGX agent

arXiv:2606.00937v1 Announce Type: new Abstract: Neural operators provide fast surrogate models for PDE simulations, but standard architectures often treat geometry and discretization as secondary to f

Cellwise and Casewise Robust Covariance in High Dimensions

ApplicationsDGX agent

arXiv:2505.19925v2 Announce Type: replace-cross Abstract: The sample covariance matrix is a cornerstone of multivariate statistics, but it is highly sensitive to outliers. These can be casewise outlie

Chaining 2-FWL GNNs for Combinatorial Graph Alignment

SafetyDGX agent

arXiv:2510.03086v2 Announce Type: replace Abstract: For the combinatorial graph alignment problem (GAP) -- finding the node correspondence that maximizes the number of common edges (nce) between two u

Challenges in the calibration of tree-based models for imbalanced classification

SafetyDGX agent

arXiv:2412.16209v5 Announce Type: replace Abstract: When using machine learning for imbalanced binary classification problems, it is common to subsample the majority class to create a (more) balanced

CHAM-net: A Contrastive Hierarchical Adaptive Meta-network for Robust Global Methane Flux Prediction

ResearchDGX agent

arXiv:2606.00338v1 Announce Type: new Abstract: Methane is a potent greenhouse gas that significantly contributes to global warming. However, accurately estimating global methane emissions and consump

Cluster Analysis with Resampling for Validation and Exploration (CARVE)

TutorialsDGX agent

arXiv:2606.00327v1 Announce Type: cross Abstract: Clustering is widely used across the sciences as the foundation for downstream data-driven scientific discoveries. However, clustering results are hig

Coarse-to-Fine Compositional Diffusion for Long-Horizon Planning

Local AiDGX agent

arXiv:2606.00837v1 Announce Type: cross Abstract: Diffusion models provide strong priors for generating structured data, but many tasks require outputs beyond the scale on which these models are typic

Coherent Off-Policy Improvement of Large Behavior Models with Learned Rewards

SafetyDGX agent

arXiv:2606.02194v1 Announce Type: new Abstract: Distilling expert demonstration data into large generative models using behavioral cloning is a scalable approach to learning capable policies for robot

COLLIE: Guiding Skill Discovery in Semantically Coherent Latent Space

TutorialsDGX agent

arXiv:2606.00950v1 Announce Type: new Abstract: Unsupervised skill discovery (USD) aims to learn diverse behaviors without reward functions, but often results in task-irrelevant or hazardous behaviors

Conditioned free-energy density of proteins using unbalanced solutions to constraint satisfaction problems

ResearchDGX agent

arXiv:2606.01329v1 Announce Type: new Abstract: We show that computing the log-partition function (free-energy) of conditioned inhomogeneous Curie--Weiss spin Hamiltonians reduces to an unbalanced 2 o

Context-aware child-directed speech detection from long-form recordings

ResearchDGX agent

arXiv:2606.01134v1 Announce Type: cross Abstract: Automatically distinguishing child-directed speech from adult-directed speech in long-form recordings is key to scalable analyses of children's langua

Continual Learning as a Multiphase Moving-Boundary Problem

ResearchDGX agent

arXiv:2606.01863v1 Announce Type: new Abstract: Continual learning struggles to balance retaining past knowledge with absorbing new tasks. Stefan-CL elegantly resolves this stability-plasticity dilemm

Controllable Value Alignment in Large Language Models through Neuron-Level Editing

Model ReleasesDGX agent

arXiv:2602.07356v2 Announce Type: replace Abstract: Aligning large language models (LLMs) with human values has become increasingly important as their influence on human behavior and decision-making e

Convex Distance Operator Transport: A Convex and Geometry-Preserving Formulation

ResearchDGX agent

arXiv:2606.02047v1 Announce Type: cross Abstract: We introduce Convex Distance Operator Transport (CDOT), the first convex optimal transport framework that aligns distributions across heterogeneous do

Cortex and subcortex play distinct roles over learning when cortical memory is limited

TutorialsDGX agent

arXiv:2606.00667v1 Announce Type: cross Abstract: It has been proposed that the brain integrates flexible, computationally expensive cortical processing with simpler, lower-cost subcortical mechanisms

CRAFT: Fine-Grained Cost-Aware Expert Replication For Efficient Mixture-of-Experts Serving

HardwareDGX agent

arXiv:2603.28768v2 Announce Type: replace-cross Abstract: Mixture-of-Experts (MoE) has recently emerged as the mainstream architecture for efficiently scaling large language models while maintaining n

CRePE: Convolution-aware Relative Importance in Post-training Pruning with Efficient Search

Local AiDGX agent

arXiv:2606.01544v1 Announce Type: new Abstract: Deploying Large Language Models (LLMs) in practice incurs substantial memory and computational costs. Post-training pruning (PTP) is an effective approa

CRMA: A Spectrally-Bounded Backbone for Modular Continual Fine-Tuning of LLMs

Model ReleasesDGX agent

arXiv:2606.00382v1 Announce Type: new Abstract: Sequential fine-tuning of large language models forces a choice: let the shared substrate keep learning and accept catastrophic forgetting, or freeze it

CryoProt: A Protein Pretraining Framework with Cross-Box Interactions on Cryo-EM Density Maps

Local AiDGX agent

arXiv:2606.00955v1 Announce Type: new Abstract: Despite the growing availability of cryo-electron microscopy (cryo-EM) density maps, effectively leveraging them for protein representation remains chal

CUPID in the Model Zoo: Online Matchmaking for Selecting Your Dream LLM

SafetyDGX agent

arXiv:2606.00846v1 Announce Type: new Abstract: Users increasingly face the challenge of selecting an appropriate LLM for a given task from a rapidly growing pool of LLMs, each with distinct but often

d2: Improving Reasoning in Diffusion Language Models via Trajectory Likelihood Estimation

SafetyDGX agent

arXiv:2509.21474v4 Announce Type: replace Abstract: While diffusion language models (DLMs) have achieved competitive performance in text generation, improving their reasoning ability with reinforcemen

DAGGER: Gradient-Free Construction of Transiently Amplifying Networks under Hard Connectivity Constraints

ResearchDGX agent

arXiv:2606.01227v1 Announce Type: new Abstract: Many networks not only support but also rely on transient non-normal amplification, an orders-of-magnitude increase in the activity of an otherwise stab

DAPD: Dependency-Aware Parallel Decoding via Attention for Diffusion LLMs

ResearchDGX agent

arXiv:2603.12996v2 Announce Type: replace Abstract: Parallel decoding for Diffusion LLMs (dLLMs) is difficult because each denoising step provides only token-wise marginal distributions, while unmaski

Data-Driven Spectral Prediction for Accelerating Large-Scale Electronic Structure Calculations

ResearchDGX agent

arXiv:2606.00401v1 Announce Type: cross Abstract: Simulating large molecular systems comprising thousands of atoms requires highly scalable methodologies. While modern Density Functional Theory (DFT)

Data Enrichment for Symbolic Regression Using Diffusion Models

ResearchDGX agent

arXiv:2606.00988v1 Announce Type: new Abstract: Symbolic regression (SR) offers a route to scientific discovery by converting observations into interpretable governing equations. However, despite its

Decentralized Instruction Tuning: Conflict-Aware Splitting and Weight Merging

Model ReleasesDGX agent

arXiv:2606.01717v1 Announce Type: new Abstract: Instruction tuning aligns large language models, including multimodal ones, with diverse user intents, but scaling to heterogeneous mixtures is hindered

Decision-Focused On-Policy Learning for Contextual Linear Optimization with Partial Feedback

Model ReleasesDGX agent

arXiv:2606.01081v1 Announce Type: new Abstract: Decision-focused learning (DFL) trains predictive models by optimizing downstream decision quality rather than standalone prediction accuracy. For conte

Design-MLLM: A Reinforcement Alignment Framework for Verifiable and Aesthetic Interior Design

Model ReleasesDGX agent

arXiv:2603.13312v2 Announce Type: replace-cross Abstract: Interior design is a requirements-to-visual-plan generation process that must simultaneously satisfy verifiable spatial feasibility and compar

Design Space Exploration of DMA based Finer-Grain Compute Communication Overlap

HardwareDGX agent

arXiv:2512.10236v2 Announce Type: replace-cross Abstract: Modern ML workloads demand distributing training and inference across multiple GPUs. However, these parallelization techniques often suffer fr

Dialectics of Alignment: Harnessing Unsafe Knowledge for Dynamic Safety Routing

SafetyDGX agent

arXiv:2606.00686v1 Announce Type: new Abstract: The prevailing paradigm in large language model (LLM) alignment operates via erasure, filtering unsafe data or training models to strictly refuse harmfu

Differentially Private Datastore Generation for Retrieval-Augmented Inference

Model ReleasesDGX agent

arXiv:2606.01413v1 Announce Type: cross Abstract: It is crucial for modern on-device AI systems that rely on retrieval-augmented inference to release and share datastores without compromising individu

Dimension Reduction via Sum-of-Squares and Improved Clustering Algorithms for Non-Spherical Mixtures

ResearchDGX agent

arXiv:2411.12438v2 Announce Type: replace-cross Abstract: We develop a new approach for clustering non-spherical (i.e., arbitrary component covariances) Gaussian mixture models via a subroutine, based

Discovering Nonlinear Static Relationships in Unlabeled Dataset using Autoencoder with Ordered Variance

ApplicationsDGX agent

arXiv:2402.14031v2 Announce Type: replace-cross Abstract: This paper presents an autoencoder with ordered variance (AEO), in which the conventional reconstruction loss is augmented by a variance-based

DistMatch: Adaptive Binning via Distribution Matching for Robust Sequential Conformal Prediction

Local AiDGX agent

arXiv:2606.00690v1 Announce Type: new Abstract: Sequential conformal prediction (CP) provides valid uncertainty quantification under the assumption of residual exchangeability. However, this assumptio

Distributed GNEP Algorithms without Multiplier Sharing and Applications to Multi-Robot Coordination and Contextual Bandit-Based Active Learning

ApplicationsDGX agent

arXiv:2606.00759v1 Announce Type: new Abstract: Recent advances in artificial intelligence have expanded the focus from classical optimization to include equilibrium analysis in noncooperative games.

Distribution-free changepoint localization after sequential change detection

ResearchDGX agent

arXiv:2606.01256v1 Announce Type: cross Abstract: This paper introduces a distribution-free framework for constructing post-detection confidence sets for changepoints after stopping a sequential chang

Doing well with less! On Sampling Techniques for Empirical Pairwise Loss Estimation/Minimization

ResearchDGX agent

arXiv:2606.02345v1 Announce Type: cross Abstract: Many machine learning problems, including similarity learning, ranking, and clustering, rely on empirical pairwise loss functions whose quadratic comp

Don't Let a Few Network Failures Slow the Entire AllReduce

HardwareDGX agent

arXiv:2606.01680v1 Announce Type: cross Abstract: Network failures are among the most frequent hardware faults in large-scale GPU clusters and a leading cause of training-job interruptions. Modern col

DREAM-S: Speculative Decoding with Searchable Drafting and Target-Aware Refinement for Multimodal Generation

ResearchDGX agent

arXiv:2606.00535v1 Announce Type: new Abstract: Speculative decoding (SD) has proven to be an effective technique for accelerating autoregressive generation in large language models (LLMs) however, it

DuetServe: Harmonizing Prefill and Decode for LLM Serving via Adaptive GPU Multiplexing

HardwareDGX agent

arXiv:2511.04791v2 Announce Type: replace Abstract: Modern LLM serving systems must sustain high throughput while meeting strict latency SLOs across two distinct inference phases: compute-intensive pr

Dynamic Proxy-Mixing: Transferring Replay Controllers from Small to Large Models for Continual Instruction Tuning

Model ReleasesDGX agent

arXiv:2606.00400v1 Announce Type: new Abstract: Continual instruction tuning updates a language model through a sequence of new domains, yet each update can progressively erode previously learned capa

Early Prediction of Liver Cirrhosis Up to Two Years in Advance: A Machine Learning Study Benchmarking Against the FIB-4 and APRI Scores

Model ReleasesDGX agent

arXiv:2601.00175v2 Announce Type: replace Abstract: Objective: Develop and evaluate machine learning (ML) models for predicting incident liver cirrhosis (LC) one and two years prior to diagnosis using

Easy, robust approximate message passing for planted spike models

ResearchDGX agent

arXiv:2606.00500v1 Announce Type: cross Abstract: We present a simple and efficient algorithm for robust approximate message passing (AMP) in the spiked matrix setting. In particular, let arepsilon be

Echo State Networks for Time Series Forecasting: Hyperparameter Sweep and Benchmarking

Model ReleasesDGX agent

arXiv:2602.03912v4 Announce Type: replace Abstract: This paper investigates the performance of Echo State Networks (ESNs) for univariate forecasting of monthly and quarterly time series from the M4 Fo

Edge-aware Decoding for Neural Asymmetric Routing

ResearchDGX agent

arXiv:2606.02136v1 Announce Type: new Abstract: Neural asymmetric routing models increasingly encode directionality through matrix representations and asymmetry-aware attention. The final routing acti

EEG-FuseFormer: A Transformer-Driven Feature Fusion Framework for Seizure Onset Prediction

ResearchDGX agent

arXiv:2606.02166v1 Announce Type: new Abstract: Epilepsy is one of the most common neurological disorders globally, characterized by recurring seizures and significantly impacting the quality of life.

Efficient Approximation for Encoder--Decoder Neural Operators via Variation Spaces

ResearchDGX agent

arXiv:2606.01244v1 Announce Type: cross Abstract: We study operator learning using encoder--decoder neural networks. Inspired by the function-space theory of neural networks, we introduce a variation

← Previous
1…104105106107108…243
Next →