AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,329 results
20 Apr 2026

Teaching Language Models Mechanistic Explainability Through MechSMILES

ResearchDGX agent

arXiv:2512.05722v2 Announce Type: replace Abstract: Chemical reaction mechanisms are the foundation of how chemists evaluate reactivity and feasibility, yet current Computer-Assisted Synthesis Plannin

The Harder Path: Last Iterate Convergence for Uncoupled Learning in Zero-Sum Games with Bandit Feedback

SafetyDGX agent

arXiv:2604.16087v1 Announce Type: new Abstract: We study the problem of learning in zero-sum matrix games with repeated play and bandit feedback. Specifically, we focus on developing uncoupled algorit

The Machine Learning Approach to Moment Closure Relations for Plasma: A Review

ResearchDGX agent

arXiv:2511.22486v2 Announce Type: replace-cross Abstract: The requirement for large-scale global simulations of plasma is an ongoing challenge in both space and laboratory plasma physics. Any simulati


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

The Spectral Geometry of Thought: Phase Transitions, Instruction Reversal, Token-Level Dynamics, and Perfect Correctness Prediction in How Transformers Reason

Model ReleasesDGX agent

arXiv:2604.15350v1 Announce Type: new Abstract: We discover that large language models exhibit spectral phase transitions in their hidden activation spaces when engaging in reasoning versus factual re

TopFeaRe: Locating Critical State of Adversarial Resilience for Graphs Regarding Topology-Feature Entanglement

TutorialsDGX agent

arXiv:2604.15370v1 Announce Type: cross Abstract: Graph adversarial attacks are usually produced from the two perspectives of topology/structure and node feature, both of them represent the paramount

Towards Robust Endogenous Reasoning: Unifying Drift Adaptation in Non-Stationary Tuning

SafetyDGX agent

arXiv:2604.15705v1 Announce Type: new Abstract: Reinforcement Fine-Tuning (RFT) has established itself as a critical paradigm for the alignment of Multi-modal Large Language Models (MLLMs) with comple

Towards Universal Convergence of Backward Error in Linear System Solvers

Model ReleasesDGX agent

arXiv:2604.16075v1 Announce Type: cross Abstract: The quest for an algorithm that solves an nimes n linear system in O(n^2) time complexity, or O(n^2 ext{poly}(1/epsilon)) when solving up to epsilon r

Truncated Kernel Stochastic Gradient Descent with General Losses and Spherical Radial Basis Functions

SafetyDGX agent

arXiv:2510.04237v5 Announce Type: replace Abstract: In this paper, we propose a novel kernel stochastic gradient descent (SGD) algorithm for large-scale supervised learning with general losses. Compar

TwinTrack: Post-hoc Multi-Rater Calibration for Medical Image Segmentation

Model ReleasesDGX agent

arXiv:2604.15950v1 Announce Type: new Abstract: Pancreatic ductal adenocarcinoma (PDAC) segmentation on contrast-enhanced CT is inherently ambiguous: inter-rater disagreement among experts reflects ge

Two-Dimensional Deep ReLU CNN Approximation for Korobov Functions: A Constructive Approach

Model ReleasesDGX agent

arXiv:2503.07976v2 Announce Type: replace-cross Abstract: This paper investigates approximation capabilities of two-dimensional (2D) deep convolutional neural networks (CNNs), with Korobov functions s

Univariate Channel Fusion for Multivariate Time Series Classification

ResearchDGX agent

arXiv:2604.16119v1 Announce Type: new Abstract: Multivariate time series classification (MTSC) plays a crucial role in various domains, including biomedical signal analysis and motion monitoring. Howe

Unsupervised domain adaptation for radioisotope identification in gamma spectroscopy

SafetyDGX agent

arXiv:2603.05719v2 Announce Type: replace Abstract: Training machine learning models for radioisotope identification using gamma spectroscopy remains an elusive challenge for many practical applicatio

Verification Modulo Tested Library Contracts

ResearchDGX agent

arXiv:2604.15533v1 Announce Type: cross Abstract: We consider the problem of verification modulo tested library contracts as a step towards automating the verification of client programs that use comp

(Weighted) Adaptive Radius Near Neighbor Search: Evaluation for WiFi Fingerprint-based Positioning

ResearchDGX agent

arXiv:2604.15940v1 Announce Type: new Abstract: Fixed Radius Near Neighbor (FRNN) search is an alternative to the widely used k Nearest Neighbors (kNN) search. Unlike kNN, FRNN determines a label or a

What Makes LLMs Effective Sequential Recommenders? A Study on Preference Intensity and Temporal Context

SafetyDGX agent

arXiv:2506.02261v3 Announce Type: replace-cross Abstract: What enables large language models (LLMs) to effectively model user preferences in sequential recommendation? Our investigation reveals that e

Why Colors Make Clustering Harder:Global Integrality Gaps, the Price of Fairness, and Color-Coupled Algorithms in Chromatic Correlation Clustering

SafetyDGX agent

arXiv:2604.15738v1 Announce Type: new Abstract: Chromatic Correlation Clustering (CCC) extends Correlation Clustering by assigning semantic colors to edges and requiring each cluster to receive a sing

Zero-Shot Scalable Resilience in UAV Swarms: A Decentralized Imitation Learning Framework with Physics-Informed Graph Interactions

SafetyDGX agent

arXiv:2604.15762v1 Announce Type: new Abstract: Large-scale Unmanned Aerial Vehicle (UAV) failures can split an unmanned aerial vehicle swarm network into disconnected sub-networks, making decentraliz

17 Apr 2026

A Mechanistic Account of Attention Sinks in GPT-2: One Circuit, Broader Implications for Mitigation

SafetyDGX agent

arXiv:2604.14722v1 Announce Type: new Abstract: Transformers commonly exhibit an attention sink: disproportionately high attention to the first position. We study this behavior in GPT-2-style models w

A Nonlinear Separation Principle: Applications to Neural Networks, Control and Learning

Model ReleasesDGX agent

arXiv:2604.15238v1 Announce Type: cross Abstract: This paper investigates continuous-time and discrete-time firing-rate and Hopfield recurrent neural networks (RNNs), with applications in nonlinear co

A Synonymous Variational Perspective on the Rate-Distortion-Perception Tradeoff

ResearchDGX agent

arXiv:2604.14603v1 Announce Type: cross Abstract: The fundamental limit of natural signal compression has traditionally been characterized by classical rate-distortion (RD) theory through the tradeoff

Active Learning with Selective Time-Step Acquisition for PDEs

Model ReleasesDGX agent

arXiv:2511.18107v2 Announce Type: replace Abstract: Accurately solving partial differential equations (PDEs) is critical to understanding complex scientific and engineering phenomena, yet traditional

Adaptive Canonicalization with Application to Invariant Anisotropic Geometric Networks

ResearchDGX agent

arXiv:2509.24886v3 Announce Type: replace Abstract: Canonicalization is a widely used strategy in equivariant machine learning, enforcing symmetry in neural networks by mapping each input to a standar

Adaptive Test-Time Compute Allocation for Reasoning LLMs via Constrained Policy Optimization

Model ReleasesDGX agent

arXiv:2604.14853v1 Announce Type: new Abstract: Test-time compute scaling, the practice of spending extra computation during inference via repeated sampling, search, or extended reasoning, has become

AgentGA: Evolving Code Solutions in Agent-Seed Space

Model ReleasesDGX agent

arXiv:2604.14655v1 Announce Type: cross Abstract: We present AgentGA, a framework that evolves autonomous code-generation runs by optimizing the agent seed: the task prompt plus optional parent archiv

AIPC: Agent-Based Automation for AI Model Deployment with Qualcomm AI Runtime

Local AiDGX agent

arXiv:2604.14661v1 Announce Type: cross Abstract: Edge AI model deployment is a multi-stage engineering process involving model conversion, operator compatibility handling, quantization calibration, r

Amortized Optimal Transport from Sliced Potentials

ResearchDGX agent

arXiv:2604.15114v1 Announce Type: cross Abstract: We propose a novel amortized optimization method for predicting optimal transport (OT) plans across multiple pairs of measures by leveraging Kantorovi

An Intelligent Robotic and Bio-Digestor Framework for Smart Waste Management

AgentsDGX agent

arXiv:2604.14882v1 Announce Type: cross Abstract: Rapid urbanization and continuous population growth have made municipal solid waste management increasingly challenging. These challenges highlight th

An unsupervised decision-support framework for multivariate biomarker analysis in athlete monitoring

SafetyDGX agent

arXiv:2604.14534v1 Announce Type: new Abstract: Purpose. Athlete monitoring is constrained by small cohorts, heterogeneous biomarker scales, limited feasibility of repeated sampling, and the lack of r

Anomaly Detection in IEC-61850 GOOSE Networks: Evaluating Unsupervised and Temporal Learning for Real-Time Intrusion Detection

ResearchDGX agent

arXiv:2604.14233v1 Announce Type: cross Abstract: The IEC-61850 GOOSE protocol underpins time-critical communication in modern digital substations but lacks native security mechanisms, leaving it vuln

Assessing the Performance-Efficiency Trade-off of Foundation Models in Probabilistic Electricity Price Forecasting

ResearchDGX agent

arXiv:2604.14739v1 Announce Type: new Abstract: Large-scale renewable energy deployment introduces pronounced volatility into the electricity system, turning grid operation into a complex stochastic o

Assessing the Potential of Masked Autoencoder Foundation Models in Predicting Downhole Metrics from Surface Drilling Data

ResearchDGX agent

arXiv:2604.15169v1 Announce Type: new Abstract: Oil and gas drilling operations generate extensive time-series data from surface sensors, yet accurate real-time prediction of critical downhole metrics

Asynchronous Probability Ensembling for Federated Disaster Detection

ResearchDGX agent

arXiv:2604.14450v1 Announce Type: new Abstract: Quick and accurate emergency handling in Disaster Decision Support Systems (DDSS) is often hampered by network latency and suboptimal application accura

Atropos: Improving Cost-Benefit Trade-off of LLM-based Agents under Self-Consistency with Early Termination and Model Hotswap

Local AiDGX agent

arXiv:2604.15075v1 Announce Type: cross Abstract: Open-weight Small Language Models(SLMs) can provide faster local inference at lower financial cost, but may not achieve the same performance level as

AutoRAN: Automated Hijacking of Safety Reasoning in Large Reasoning Models

Model ReleasesDGX agent

arXiv:2505.10846v3 Announce Type: replace Abstract: This paper presents AutoRAN, the first framework to automate the hijacking of internal safety reasoning in large reasoning models (LRMs). At its cor

Auxiliary Finite-Difference Residual-Gradient Regularization for PINNs

Model ReleasesDGX agent

arXiv:2604.14472v1 Announce Type: new Abstract: Physics-informed neural networks (PINNs) are often selected by a single scalar loss even when the quantity of interest is more specific. We study a hybr

Awakening Dormant Experts:Counterfactual Routing to Mitigate MoE Hallucinations

ResearchDGX agent

arXiv:2604.14246v1 Announce Type: new Abstract: Sparse Mixture-of-Experts (MoE) models have achieved remarkable scalability, yet they remain vulnerable to hallucinations, particularly when processing

Bayesian Optimization with Gaussian Processes to Accelerate Stationary Point Searches

ApplicationsDGX agent

arXiv:2603.10992v3 Announce Type: replace-cross Abstract: Building local surrogates to accelerate stationary point searches on potential energy surfaces spans decades of effort. Done correctly, surrog

Benchmarking Optimizers for MLPs in Tabular Deep Learning

Model ReleasesDGX agent

arXiv:2604.15297v1 Announce Type: new Abstract: MLP is a heavily used backbone in modern deep learning (DL) architectures for supervised learning on tabular data, and AdamW is the go-to optimizer used

Best of both worlds: Stochastic & adversarial best-arm identification

Model ReleasesDGX agent

arXiv:2604.14860v1 Announce Type: cross Abstract: We study bandit best-arm identification with arbitrary and potentially adversarial rewards. A simple random uniform learner obtains the optimal rate o

Beyond Importance Sampling: Rejection-Gated Policy Optimization

SafetyDGX agent

arXiv:2604.14895v1 Announce Type: new Abstract: We propose a new perspective on policy optimization: rather than reweighting all samples by their importance ratios, an optimizer should select which sa

Beyond the Laplacian: Doubly Stochastic Matrices for Graph Neural Networks

ResearchDGX agent

arXiv:2604.15069v1 Announce Type: new Abstract: Graph Neural Networks (GNNs) conventionally rely on standard Laplacian or adjacency matrices for structural message passing. In this work, we substitute

Bias in Surface Electromyography Features across a Demographically Diverse Cohort

SafetyDGX agent

arXiv:2604.14460v1 Announce Type: cross Abstract: Neuromotor decoding from upper-limb electromyography (sEMG) can enhance human-machine interfaces and offer a more natural means of controlling prosthe

Bit-Accurate Modeling of GPU Matrix Multiply-Accumulate Units: Demystifying Numerical Discrepancy and Accuracy

HardwareDGX agent

arXiv:2511.10909v2 Announce Type: replace-cross Abstract: Modern AI accelerators rely on matrix multiply-accumulate units (MMAUs), such as NVIDIA Tensor Cores and AMD Matrix Cores, to accelerate deep

Blazing the trails before beating the path: Sample-efficient Monte-Carlo planning

ResearchDGX agent

arXiv:2604.14974v1 Announce Type: new Abstract: You are a robot and you live in a Markov decision process (MDP) with a finite or an infinite number of transitions from state-action to next states. You

Bridging the Gap between Learning and Inference for Diffusion-Based Molecule Generation

Model ReleasesDGX agent

arXiv:2411.05472v2 Announce Type: replace Abstract: The paradigm shift toward structure-driven molecule generation has been propelled by advances in deep generative models, such as variational auto-en

Calibrate-Then-Delegate: Safety Monitoring with Risk and Budget Guarantees via Model Cascades

SafetyDGX agent

arXiv:2604.14251v1 Announce Type: new Abstract: Monitoring LLM safety at scale requires balancing cost and accuracy: a cheap latent-space probe can screen every input, but hard cases should be escalat

Calibration-Gated LLM Pseudo-Observations for Online Contextual Bandits

SafetyDGX agent

arXiv:2604.14961v1 Announce Type: new Abstract: Contextual bandit algorithms suffer from high regret during cold-start, when the learner has insufficient data to distinguish good arms from bad. We pro

Can LLMs Score Medical Diagnoses and Clinical Reasoning as well as Expert Panels?

SafetyDGX agent

arXiv:2604.14892v1 Announce Type: new Abstract: Evaluating medical AI systems using expert clinician panels is costly and slow, motivating the use of large language models (LLMs) as alternative adjudi

Catching Every Ripple: Enhanced Anomaly Awareness via Dynamic Concept Adaptation

Model ReleasesDGX agent

arXiv:2604.14726v1 Announce Type: new Abstract: Online anomaly detection (OAD) plays a pivotal role in real-time analytics and decision-making for evolving data streams. However, existing methods ofte

Certified and accurate computation of function space norms of deep neural networks

ResearchDGX agent

arXiv:2603.06431v2 Announce Type: replace-cross Abstract: Neural network methods for PDEs require reliable error control in function space norms. However, trained neural networks can typically only be

CLion: Efficient Cautious Lion Optimizer with Enhanced Generalization

Model ReleasesDGX agent

arXiv:2604.14587v1 Announce Type: new Abstract: Lion optimizer is a popular learning-based optimization algorithm in machine learning, which shows impressive performance in training many deep learning

Cloning is as Hard as Learning for Stabilizer States

TutorialsDGX agent

arXiv:2604.15269v1 Announce Type: cross Abstract: The impossibility of simultaneously cloning non-orthogonal states lies at the foundations of quantum theory. Even when allowing for approximation erro

Combining Bayesian and Frequentist Inference for Laboratory-Specific Performance Guarantees in Copy Number Variation Detection

ApplicationsDGX agent

arXiv:2604.14305v1 Announce Type: cross Abstract: Targeted amplicon panels are widely used in oncology diagnostics, but providing per-gene performance guarantees for copy number variant (CNV) detectio

Conformal Policy Control

SafetyDGX agent

arXiv:2603.02196v2 Announce Type: replace-cross Abstract: An agent must try new behaviors to explore and improve. In high-stakes environments, an agent that violates safety constraints may cause harm

Constrained Decoding for Safe Robot Navigation Foundation Models

Model ReleasesDGX agent

arXiv:2509.01728v4 Announce Type: replace-cross Abstract: Recent advances in the development of robotic foundation models have led to promising end-to-end and general-purpose capabilities in robotic s

Constraint-based Pre-training: From Structured Constraints to Scalable Model Initialization

ResearchDGX agent

arXiv:2604.14769v1 Announce Type: new Abstract: The pre-training and fine-tuning paradigm has become the dominant approach for model adaptation. However, conventional pre-training typically yields mod

Continual Learning for fMRI-Based Brain Disorder Diagnosis via Functional Connectivity Matrices Generative Replay

ApplicationsDGX agent

arXiv:2604.14259v1 Announce Type: cross Abstract: Functional magnetic resonance imaging (fMRI) is widely used for studying and diagnosing brain disorders, with functional connectivity (FC) matrices pr

Continuous-time reinforcement learning: ellipticity enables model-free value function approximation

SafetyDGX agent

arXiv:2602.06930v2 Announce Type: replace Abstract: We study off-policy reinforcement learning for controlling continuous-time Markov diffusion processes with discrete-time observations and actions. W

Cornfigurator: Automated Planning for Any-to-Any Multimodal Model Serving

ResearchDGX agent

arXiv:2512.14098v3 Announce Type: replace Abstract: Any-to-Any models are an emerging class of multimodal models that accept combinations of text and multimodal data as input and generate them as outp

CSRA: Controlled Spectral Residual Augmentation for Robust Sepsis Prediction

ResearchDGX agent

arXiv:2604.14532v1 Announce Type: new Abstract: Accurate prediction of future risk and disease progression in sepsis is clinically important for early warning and timely intervention in intensive care

← Previous
1…220221222223224…239
Next →