AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,432 results
21 Apr 2026

Q-SINDy: Quantum-Kernel Sparse Identification of Nonlinear Dynamics with Provable Coefficient Debiasing

SafetyDGX agent

arXiv:2604.16779v1 Announce Type: cross Abstract: Quantum feature maps offer expressive embeddings for classical learning tasks, and augmenting sparse identification of nonlinear dynamics (SINDy) with

Quantifying how AI Panels improve precision

ResearchDGX agent

arXiv:2604.16432v1 Announce Type: cross Abstract: AI in applications like screening job applicants had become widespread, and may contribute to unemployment especially among the young. Biases in the A

RACE Attention: A Strictly Linear-Time Attention Layer for Training on Outrageously Large Contexts

HardwareDGX agent

arXiv:2510.04008v5 Announce Type: replace Abstract: Softmax Attention has a quadratic time complexity in sequence length, which becomes prohibitive to run at long contexts, even with highly optimized


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Random Matrix Theory of Early-Stopped Gradient Flow: A Transient BBP Scenario

ResearchDGX agent

arXiv:2604.18450v1 Announce Type: cross Abstract: Empirical studies of trained models often report a transient regime in which signal is detectable in a finite gradient descent time window before over

Randomized Antipodal Search Done Right for Data Pareto Improvement of LLM Unlearning

ResearchDGX agent

arXiv:2604.16591v1 Announce Type: new Abstract: Large language models (LLMs) sometimes memorize undesirable knowledge, which must be removed after deployment. Prior work on machine unlearning has focu

Randomly Initialized Networks Can Learn from Peer-to-Peer Consensus

TutorialsDGX agent

arXiv:2604.18390v1 Announce Type: new Abstract: In self-supervised learning, self-distilled methods have shown impressive performance, learning representations useful for downstream tasks and even dis

Ranking Abuse via Strategic Pairwise Data Perturbations

ApplicationsDGX agent

arXiv:2604.17805v1 Announce Type: new Abstract: Pairwise ranking systems based on Maximum Likelihood Estimation (MLE), such as the Bradley-Terry model, are widely used to aggregate preferences from pa

RASP-Tuner: Retrieval-Augmented Soft Prompts for Context-Aware Black-Box Optimization in Non-Stationary Environments

ApplicationsDGX agent

arXiv:2604.18026v1 Announce Type: new Abstract: Many deployed systems expose black-box objectives whose minimizing configuration shifts with an externally observed context. When contexts revisit a sma

Rate-Distortion Optimization for Transformer Inference

ResearchDGX agent

arXiv:2601.22002v3 Announce Type: replace Abstract: Transformers achieve superior performance on many tasks, but impose heavy compute and memory requirements during inference. This inference can be ma

RAYEN: Imposition of Hard Convex Constraints on Neural Networks

SafetyDGX agent

arXiv:2307.08336v2 Announce Type: replace Abstract: Despite the numerous applications of convex constraints in Robotics, enforcing them within learning-based frameworks remains an open challenge. Exis

REALM: Reliable Expertise-Aware Language Model Fine-Tuning from Noisy Annotations

ResearchDGX agent

arXiv:2604.17289v1 Announce Type: new Abstract: Supervised fine-tuning of large language models relies on human-annotated data, yet annotation pipelines routinely involve multiple crowdworkers of hete

Reasoning on the Manifold: Bidirectional Consistency for Self-Verification in Diffusion Language Models

SafetyDGX agent

arXiv:2604.16565v1 Announce Type: new Abstract: While Diffusion Large Language Models (dLLMs) offer structural advantages for global planning, efficiently verifying that they arrive at correct answers

Recovery Guarantees for Continual Learning of Dependent Tasks: Memory, Data-Dependent Regularization, and Data-Dependent Weights

ResearchDGX agent

arXiv:2604.17578v1 Announce Type: new Abstract: Continual learning (CL) is concerned with learning multiple tasks sequentially without forgetting previously learned tasks. Despite substantial empirica

Reference-state System Reliability method for scalable uncertainty quantification of coherent systems

ResearchDGX agent

arXiv:2604.17066v1 Announce Type: new Abstract: Coherent systems are representative of many practical applications, ranging from infrastructure networks to supply chains. Probabilistic evaluation of s

RefineStat: Efficient Exploration for Probabilistic Program Synthesis

ResearchDGX agent

arXiv:2509.01082v3 Announce Type: replace Abstract: Probabilistic programming offers a powerful framework for modeling uncertainty, yet statistical model discovery in this domain entails navigating an

ReGA: Model-Based Safeguard for LLMs via Representation-Guided Abstraction

SafetyDGX agent

arXiv:2506.01770v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have achieved tremendous success in various tasks, yet concerns about their safety and security have emerged. In

Representation Before Training: A Fixed-Budget Benchmark for Generative Medical Event Models

Model ReleasesDGX agent

arXiv:2604.16775v1 Announce Type: new Abstract: Every prediction from a generative medical event model is bounded by how clinical events are tokenized, yet input representation is rarely isolated from

Rethinking Cross-Modal Fine-Tuning: Optimizing the Interaction Between Feature Alignment and Target Fitting

Model ReleasesDGX agent

arXiv:2601.18231v4 Announce Type: replace Abstract: Adapting pre-trained models to unseen feature modalities has become increasingly important due to the growing need for cross-disciplinary knowledge

Rethinking the Comparison Unit in Sequence-Level Reinforcement Learning: An Equal-Length Paired Training Framework from Loss Correction to Sample Construction

SafetyDGX agent

arXiv:2604.17328v1 Announce Type: new Abstract: This paper investigates the length problem in sequence-level relative reinforcement learning. We observe that, although existing methods partially allev

Rethinking Uncertainty Estimation in LLMs: A Principled Single-Sequence Measure

ApplicationsDGX agent

arXiv:2412.15176v3 Announce Type: replace Abstract: Large Language Models (LLMs) are increasingly employed in real-world applications, driving the need to evaluate the trustworthiness of their generat

Revisiting Active Sequential Prediction-Powered Mean Estimation

Model ReleasesDGX agent

arXiv:2604.18569v1 Announce Type: cross Abstract: In this work, we revisit the problem of active sequential prediction-powered mean estimation, where at each round one must decide the query probabilit

Revisiting Auxiliary Losses for Conditional Depth Routing: An Empirical Study

Model ReleasesDGX agent

arXiv:2604.17228v1 Announce Type: new Abstract: Conditional depth execution routes a subset of tokens through a lightweight cheap FFN while the remainder execute the standard full FFN at each controll

Revisiting Forest Proximities via Sparse Leaf-Incidence Kernels

ResearchDGX agent

arXiv:2601.02735v2 Announce Type: replace Abstract: Decision forests induce supervised similarities through the partition structure of their trees. Yet forest proximity computation is still often trea

R&F-Inventory: A Large-Scale Dataset for Monotonic Inventory Estimation in Reach and Frequency Advertising

Model ReleasesDGX agent

arXiv:2604.16821v1 Announce Type: new Abstract: Reach and Frequency (R&F) contract advertising is an important form of widely used brand advertising. Unlike performance advertising, R&F contracts emph

RISC-V Functional Safety for Autonomous Automotive Systems: An Analytical Framework and Research Roadmap for ML-Assisted Certification

SafetyDGX agent

arXiv:2604.17391v1 Announce Type: cross Abstract: RISC-V is emerging as a viable platform for automotive-grade embedded computing, with recent ISO 26262 ASIL-D certifications demonstrating readiness f

Robust Tool Use via Fission-GRPO: Learning to Recover from Execution Errors

SafetyDGX agent

arXiv:2601.15625v2 Announce Type: replace Abstract: Large language models (LLMs) can call tools effectively, yet they remain brittle in multi-turn execution: after a tool-call error, smaller models of

RosettaSearch: Multi-Objective Inference-Time Search for Protein Sequence Design

Model ReleasesDGX agent

arXiv:2604.17175v1 Announce Type: new Abstract: We introduce RosettaSearch, an inference-time multi-objective optimization approach for protein sequence optimization. We use large language models (LLM

Saddle-To-Saddle Dynamics in Deep ReLU Networks: Low-Rank Bias in the First Saddle Escape

Model ReleasesDGX agent

arXiv:2505.21722v2 Announce Type: replace Abstract: When a deep ReLU network is initialized with small weights, gradient descent (GD) is at first dominated by the saddle at the origin in parameter spa

Safe Control using Learned Safety Filters and Adaptive Conformal Inference

Model ReleasesDGX agent

arXiv:2604.18482v1 Announce Type: cross Abstract: Safety filters have been shown to be effective tools to ensure the safety of control systems with unsafe nominal policies. To address scalability chal

SafeAnchor: Preventing Cumulative Safety Erosion in Continual Domain Adaptation of Large Language Models

Model ReleasesDGX agent

arXiv:2604.17691v1 Announce Type: new Abstract: Safety alignment in large language models is remarkably shallow: it is concentrated in the first few output tokens and reversible by fine-tuning on as f

SafeLM: Unified Privacy-Aware Optimization for Trustworthy Federated Large Language Models

SafetyDGX agent

arXiv:2604.16606v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly deployed in high-stakes domains, yet a unified treatment of their overlapping safety challenges remains

Sampling for Quality: Training-Free Reward-Guided LLM Decoding via Sequential Monte Carlo

ResearchDGX agent

arXiv:2604.16453v1 Announce Type: new Abstract: We introduce a principled probabilistic framework for reward-guided decoding in large language models, addressing the limitations of standard decoding m

Sampling Matters: The Effect of ECG Frequency on Deep Learning-Based Atrial Fibrillation Detection

Model ReleasesDGX agent

arXiv:2604.16437v1 Announce Type: cross Abstract: Deep learning models for atrial fibrillation (AF) detection are increasingly trained on heterogeneous electrocardiogram (ECG) datasets with varying sa

Scalable and Adaptive Parallel Training of Graph Transformer on Large Graphs

HardwareDGX agent

arXiv:2604.16715v1 Announce Type: cross Abstract: Graph foundation models have demonstrated remarkable adaptability across diverse downstream tasks through large-scale pretraining on graphs. However,

Scalable Neighborhood-Based Multi-Agent Actor-Critic

SafetyDGX agent

arXiv:2604.18190v1 Announce Type: new Abstract: We propose MADDPG-K, a scalable extension to Multi-Agent Deep Deterministic Policy Gradient (MADDPG) that addresses the computational limitations of cen

Scalable Physics-Informed Neural Differential Equations and Data-Driven Algorithms for HVAC Systems

SafetyDGX agent

arXiv:2604.18438v1 Announce Type: new Abstract: We present a scalable, data-driven simulation framework for large-scale heating, ventilation, and air conditioning (HVAC) systems that couples physics-i

Scalable Quantum Error Mitigation with Physically Informed Graph Neural Networks

Local AiDGX agent

arXiv:2604.16815v1 Announce Type: cross Abstract: Quantum error mitigation (QEM) provides a practical route for estimating reliable observables on noisy intermediate-scale quantum (NISQ) devices. Trad

Scale-free adaptive planning for deterministic dynamics & discounted rewards

ResearchDGX agent

arXiv:2604.18312v1 Announce Type: new Abstract: We address the problem of planning in an environment with deterministic dynamics and stochastic rewards with discounted returns. The optimal value funct

Scaling Human-AI Coding Collaboration Requires a Governable Consensus Layer

Model ReleasesDGX agent

arXiv:2604.17883v1 Announce Type: cross Abstract: Vibe coding produces correct, executable code at speed, but leaves no record of the structural commitments, dependencies, or evidence behind it. Revie

Scaling Recurrence-aware Foundation Models for Clinical Records via Next-Visit Prediction

Model ReleasesDGX agent

arXiv:2603.24562v2 Announce Type: replace Abstract: While large-scale pretraining has revolutionized language modeling, its potential remains underexplored in healthcare with structured electronic hea

SCATR: Simple Calibrated Test-Time Ranking

ResearchDGX agent

arXiv:2604.16535v1 Announce Type: new Abstract: Test-time scaling (TTS) improves large language models (LLMs) by allocating additional compute at inference time. In practice, TTS is often achieved thr

SeekerGym: A Benchmark for Reliable Information Seeking

Model ReleasesDGX agent

arXiv:2604.17143v1 Announce Type: new Abstract: Despite their substantial successes, AI agents continue to face fundamental challenges in terms of trustworthiness. Consider deep research agents, taske

Self-Predictive Representations for Combinatorial Generalization in Behavioral Cloning

ResearchDGX agent

arXiv:2506.10137v3 Announce Type: replace Abstract: While goal-conditioned behavior cloning (GCBC) methods can perform well on in-distribution training tasks, they do not necessarily generalize zero-s

Self-Reinforcing Controllable Synthesis of Rare Relational Data via Bayesian Calibration

TutorialsDGX agent

arXiv:2604.16817v1 Announce Type: new Abstract: Imbalanced data is commonly present in real-world applications. While data synthesis can effectively mitigate the data scarcity problem of rare-classes,

Semantic-based Distributed Learning for Diverse and Discriminative Representations

ResearchDGX agent

arXiv:2604.18237v1 Announce Type: new Abstract: In large-scale distributed scenarios, increasingly complex tasks demand more intelligent collaboration across networks, requiring the joint extraction o

Semantic Step Prediction: Multi-Step Latent Forecasting in LLM Reasoning Trajectories via Step Sampling

Local AiDGX agent

arXiv:2604.18464v1 Announce Type: new Abstract: Semantic Tube Prediction (STP) leverages representation geometric to regularize LLM hidden-state trajectories toward locally linear geodesics during fin

Shifting the Gradient: Understanding How Defensive Training Methods Protect Language Model Integrity

ApplicationsDGX agent

arXiv:2604.16423v1 Announce Type: new Abstract: Defensive training methods such as positive preventative steering (PPS) and inoculation prompting (IP) offer surprising results through seemingly simila

SigGate-GT: Taming Over-Smoothing in Graph Transformers via Sigmoid-Gated Attention

Model ReleasesDGX agent

arXiv:2604.17324v1 Announce Type: new Abstract: Graph transformers achieve strong results on molecular and long-range reasoning tasks, yet remain hampered by over-smoothing (the progressive collapse o

SIGMA: A Semantic-Grounded Instruction-Driven Generative Multi-Task Recommender at AliExpress

ApplicationsDGX agent

arXiv:2602.22913v2 Announce Type: replace-cross Abstract: With the rapid evolution of Large Language Models (LLMs), generative recommendation is gradually reshaping the paradigm of recommender systems

Singularity Formation: Synergy in Theoretical, Numerical and Machine Learning Approaches

ResearchDGX agent

arXiv:2604.16842v1 Announce Type: cross Abstract: This thesis develops numerical and theoretical approaches for understanding and analyzing singularity formation in Partial Differential Equations (PDE

SinkRouter: Sink-Aware Routing for Efficient Long-Context Decoding in Large Language and Multimodal Models

Model ReleasesDGX agent

arXiv:2604.16883v1 Announce Type: new Abstract: In long-context decoding for LLMs and LMMs, attention becomes increasingly memory-bound because each decoding step must load a large amount of KV-cache

SLO-Guard: Crash-Aware, Budget-Consistent Autotuning for SLO-Constrained LLM Serving

HardwareDGX agent

arXiv:2604.17627v1 Announce Type: new Abstract: Serving large language models under latency service-level objectives (SLOs) is a configuration-heavy systems problem with an unusually failure-prone sea

Sobolev Gradient Ascent for Optimal Transport: Barycenter Optimization and Convergence Analysis

ResearchDGX agent

arXiv:2505.13660v2 Announce Type: replace-cross Abstract: This paper introduces a new constraint-free concave dual formulation for the Wasserstein barycenter. Tailoring the vanilla dual gradient ascen

Sonata: A Hybrid World Model for Inertial Kinematics under Clinical Data Scarcity

Model ReleasesDGX agent

arXiv:2604.18058v1 Announce Type: new Abstract: We introduce Sonata, a compact latent world model for six-axis trunk IMU representation learning under clinical data scarcity. Clinical cohorts typicall

SparrowSNN: A Hardware/software Co-design for Energy Efficient ECG Classification

ResearchDGX agent

arXiv:2406.06543v2 Announce Type: replace-cross Abstract: Deep learning has driven significant technological advancements, but its high energy consumption limits its use on battery-operated edge devic

SPaRSe-TIME: Saliency-Projected Low-Rank Temporal Modeling for Efficient and Interpretable Time Series Prediction

ApplicationsDGX agent

arXiv:2604.17350v1 Announce Type: cross Abstract: Time series forecasting is traditionally dominated by sequence-based architectures such as recurrent neural networks and attention mechanisms, which p

Spectral bandits for smooth graph functions

SafetyDGX agent

arXiv:2604.18420v1 Announce Type: cross Abstract: Smooth functions on graphs have wide applications in manifold and semi-supervised learning. In this paper, we study a bandit problem where the payoffs

SpiralFormer: Looped Transformers Can Learn Hierarchical Dependencies via Multi-Resolution Recursion

Model ReleasesDGX agent

arXiv:2602.11698v2 Announce Type: replace Abstract: Recursive (looped) Transformers decouple computational depth from parameter depth by repeatedly applying shared layers, providing an explicit archit

SPOT: Single-Shot Positioning via Trainable Near-Field Rainbow Beamforming

ResearchDGX agent

arXiv:2511.11391v3 Announce Type: replace Abstract: Phase-time arrays, which integrate phase shifters (PSs) and true-time delays (TTDs), have emerged as a cost-effective architecture for generating fr

Stable On-Policy Distillation through Adaptive Target Reformulation

Model ReleasesDGX agent

arXiv:2601.07155v2 Announce Type: replace Abstract: Knowledge distillation (KD) is a widely adopted technique for transferring knowledge from large language models to smaller student models; however,

← Previous
1…218219220221222…241
Next →