AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,329 results
16 Apr 2026

Robust Low-Rank Tensor Completion based on M-product with Weighted Correlated Total Variation and Sparse Regularization

Model ReleasesDGX agent

arXiv:2604.13525v1 Announce Type: cross Abstract: The robust low-rank tensor completion problem addresses the challenge of recovering corrupted high-dimensional tensor data with missing entries, outli

Robust Ultra Low-Bit Post-Training Quantization via Stable Diagonal Curvature Estimate

ResearchDGX agent

arXiv:2604.13806v1 Announce Type: new Abstract: Large Language Models (LLMs) are widely used across many domains, but their scale makes deployment challenging. Post-Training Quantization (PTQ) reduces

Robust Verification of Controllers under State Uncertainty via Hamilton-Jacobi Reachability Analysis

SafetyDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2511.14755v2 Announce Type: replace-cross Abstract: As perception-based controllers for autonomous systems become increasingly popular in the real world, it is important that we can formally ver

RPS: Information Elicitation with Reinforcement Prompt Selection

Model ReleasesDGX agent

arXiv:2604.13817v1 Announce Type: new Abstract: Large language models (LLMs) have shown remarkable capabilities in dialogue generation and reasoning, yet their effectiveness in eliciting user-known bu

Sandpile Economics: Theory, Identification, and Evidence

AgentsDGX agent

arXiv:2604.13890v1 Announce Type: cross Abstract: Why do capitalist economies recurrently generate crises whose severity is disproportionate to the size of the triggering shock? This paper proposes a

Scalable Spatiotemporal Inference with Biased Scan Attention Transformer Neural Processes

HardwareDGX agent

arXiv:2506.09163v2 Announce Type: replace Abstract: Neural Processes (NPs) are a rapidly evolving class of models designed to directly model the posterior predictive distribution of stochastic process

Scalable unsupervised feature selection via weight stability

ResearchDGX agent

arXiv:2506.06114v4 Announce Type: replace Abstract: Unsupervised feature selection is critical for improving clustering performance in high-dimensional data, where irrelevant features can obscure mean

Selecting Feature Interactions for Generalized Additive Models by Distilling Foundation Models

TutorialsDGX agent

arXiv:2604.13332v1 Announce Type: new Abstract: Identifying meaningful feature interactions is a central challenge in building accurate and interpretable models for tabular data. Generalized additive

Self-Organizing Maps with Optimized Latent Positions

Model ReleasesDGX agent

arXiv:2604.13622v1 Announce Type: new Abstract: Self-Organizing Maps (SOM) are a classical method for unsupervised learning, vector quantization, and topographic mapping of high-dimensional data. Howe

SFT-GRPO Data Overlap as a Post-Training Hyperparameter for Autoformalization

SafetyDGX agent

arXiv:2604.13515v1 Announce Type: new Abstract: Supervised fine-tuning (SFT) followed by Group Relative Policy Optimization (GRPO) is a common post-training recipe. We conduct a controlled ablation ov

SHARe-KAN: Post-Training Vector Quantization for Cache-Resident KAN Inference

Model ReleasesDGX agent

arXiv:2512.15742v2 Announce Type: replace Abstract: Pre-trained Vision Kolmogorov-Arnold Networks (KANs) store a dense B-spline grid on every edge, inflating prediction-head parameter counts by more t

Simulation-Based Optimisation of Batting Order and Bowling Plans in T20 Cricket

ResearchDGX agent

arXiv:2604.13861v1 Announce Type: new Abstract: This paper develops a unified Markov Decision Process (MDP) framework for optimising two recurring in-match decisions in T20 cricket namely batting orde

Soft Q(lambda): A multi-step off-policy method for entropy regularised reinforcement learning using eligibility traces

SafetyDGX agent

arXiv:2604.13780v1 Announce Type: new Abstract: Soft Q-learning has emerged as a versatile model-free method for entropy-regularised reinforcement learning, optimising for returns augmented with a pen

Some Theoretical Limitations of t-SNE

ResearchDGX agent

arXiv:2604.13295v1 Announce Type: new Abstract: t-SNE has gained popularity as a dimension reduction technique, especially for visualizing data. It is well-known that all dimension reduction technique

Sparse Goodness: How Selective Measurement Transforms Forward-Forward Learning

TutorialsDGX agent

arXiv:2604.13081v1 Announce Type: new Abstract: The Forward-Forward (FF) algorithm is a biologically plausible alternative to backpropagation that trains neural networks layer by layer using a local g

SparseBalance: Load-Balanced Long Context Training with Dynamic Sparse Attention

Model ReleasesDGX agent

arXiv:2604.13847v1 Announce Type: new Abstract: While sparse attention mitigates the computational bottleneck of long-context LLM training, its distributed training process exhibits extreme heterogene

Spectral Entropy Collapse as an Empirical Signature of Delayed Generalisation in Grokking

Model ReleasesDGX agent

arXiv:2604.13123v1 Announce Type: new Abstract: Grokking -- delayed generalisation long after memorisation -- lacks a predictive mechanistic explanation. We identify the normalised spectral entropy il

Spectral methods: crucial for machine learning, natural for quantum computers?

SafetyDGX agent

arXiv:2603.24654v2 Announce Type: replace-cross Abstract: This article presents an argument for why quantum computers could unlock new methods for machine learning. We argue that spectral methods, in

Spectral Thompson sampling

ApplicationsDGX agent

arXiv:2604.13739v1 Announce Type: new Abstract: Thompson Sampling (TS) has attracted a lot of interest due to its good empirical performance, in particular in the computational advertising. Though suc

Stochastic Trust-Region Methods for Over-parameterized Models

Model ReleasesDGX agent

arXiv:2604.14017v1 Announce Type: cross Abstract: Under interpolation-type assumptions such as the strong growth condition, stochastic optimization methods can attain convergence rates comparable to f

Structure- and Stability-Preserving Learning of Port-Hamiltonian Systems

ResearchDGX agent

arXiv:2604.13297v1 Announce Type: cross Abstract: This paper investigates the problem of data-driven modeling of port-Hamiltonian systems while preserving their intrinsic Hamiltonian structure and sta

Swap Regret Minimization Through Response-Based Approachability

ResearchDGX agent

arXiv:2602.06264v2 Announce Type: replace Abstract: We consider the problem of minimizing different notions of swap regret in online optimization. These forms of regret are tightly connected to correl

Synthetic Tabular Generators Fail to Preserve Behavioral Fraud Patterns: A Benchmark on Temporal, Velocity, and Multi-Account Signals

Model ReleasesDGX agent

arXiv:2604.13125v1 Announce Type: new Abstract: We introduce behavioral fidelity -- a third evaluation dimension for synthetic tabular data that measures whether generated data preserves the temporal,

Text-Attributed Knowledge Graph Enrichment with Large Language Models for Medical Concept Representation

Model ReleasesDGX agent

arXiv:2604.13331v1 Announce Type: new Abstract: In electronic health record (EHR) mining, learning high-quality representations of medical concepts (e.g., standardized diagnosis, medication, and proce

The cognitive companion: a lightweight parallel monitoring architecture for detecting and recovering from reasoning degradation in LLM agents

Model ReleasesDGX agent

arXiv:2604.13759v1 Announce Type: cross Abstract: Large language model (LLM) agents on multi-step tasks suffer reasoning degradation, looping, drift, stuck states, at rates up to 30% on hard tasks. Cu

The Long Delay to Arithmetic Generalization: When Learned Representations Outrun Behavior

SafetyDGX agent

arXiv:2604.13082v1 Announce Type: new Abstract: Grokking in transformers trained on algorithmic tasks is characterized by a long delay between training-set fit and abrupt generalization, but the sourc

The Signal is in the Steps: Local Scoring for Reasoning Data Selection

ResearchDGX agent

arXiv:2510.03988v2 Announce Type: replace Abstract: Distilling long-form reasoning from teacher models into smaller students requires selecting which candidate solutions to train on. Recent work argue

Think Outside the Policy: In-Context Steered Policy Optimization

SafetyDGX agent

arXiv:2510.26519v3 Announce Type: replace Abstract: Existing Reinforcement Learning from Verifiable Rewards (RLVR) methods, such as Group Relative Policy Optimization (GRPO), have achieved remarkable

TIP: Token Importance in On-Policy Distillation

Model ReleasesDGX agent

arXiv:2604.14084v1 Announce Type: new Abstract: On-policy knowledge distillation (OPD) trains a student on its own rollouts under token-level supervision from a teacher. Not all token positions matter

Transcriptomic Models for Immunotherapy Response Prediction Show Limited Cross-cohort Generalisability

Model ReleasesDGX agent

arXiv:2604.05478v2 Announce Type: replace-cross Abstract: Immune checkpoint inhibitors (ICIs) have transformed cancer therapy; yet substantial proportion of patients exhibit intrinsic or acquired resi

UI-Copilot: Advancing Long-Horizon GUI Automation via Tool-Integrated Policy Optimization

Model ReleasesDGX agent

arXiv:2604.13822v1 Announce Type: new Abstract: MLLM-based GUI agents have demonstrated strong capabilities in complex user interface interaction tasks. However, long-horizon scenarios remain challeng

Universality of Gaussian-Mixture Reverse Kernels in Conditional Diffusion

ResearchDGX agent

arXiv:2604.13470v1 Announce Type: new Abstract: We prove that conditional diffusion models whose reverse kernels are finite Gaussian mixtures with ReLU-network logits can approximate suitably regular

Unsupervised Anomaly Detection in Process-Complex Industrial Time Series: A Real-World Case Study

Model ReleasesDGX agent

arXiv:2604.13928v1 Announce Type: new Abstract: Industrial time-series data from real production environments exhibits substantially higher complexity than commonly used benchmark datasets, primarily

Unsupervised domain transfer: Overcoming signal degradation in sleep monitoring by increasing scoring realism

TutorialsDGX agent

arXiv:2604.13988v1 Announce Type: new Abstract: Objective: Investigate whether hypnogram 'realism' can be used to guide an unsupervised method for handling arbitrary types of signal degradation in mob

VIGILant: an automatic classification pipeline for glitches in the Virgo detector

ResearchDGX agent

arXiv:2604.13687v1 Announce Type: cross Abstract: Glitches frequently contaminate data in gravitational-wave detectors, complicating the observation and analysis of astrophysical signals. This work in

When Less Latent Leads to Better Relay: Information-Preserving Compression for Latent Multi-Agent LLM Collaboration

AgentsDGX agent

arXiv:2604.13349v1 Announce Type: new Abstract: Communication in Large Language Model (LLM)-based multi-agent systems is moving beyond discrete tokens to preserve richer context. Recent work such as L

WIN-U: Woodbury-Informed Newton-Unlearning as a retain-free Machine Unlearning Framework

ResearchDGX agent

arXiv:2604.13438v1 Announce Type: new Abstract: Privacy concerns in LLMs have led to the rapidly growing need to enforce a data's 'right to be forgotten'. Machine unlearning addresses precisely this t

Zero-Shot Function Encoder-Based Differentiable Predictive Control

ResearchDGX agent

arXiv:2511.05757v3 Announce Type: replace-cross Abstract: We introduce a differentiable framework for zero-shot adaptive control over parametric families of nonlinear dynamical systems. Our approach i

ZK-APEX: Zero-Knowledge Approximate Personalized Unlearning with Executable Proofs

SafetyDGX agent

arXiv:2512.09953v2 Announce Type: replace-cross Abstract: Machine unlearning aims to remove the influence of specific data points from a trained model to satisfy privacy, copyright, and safety require

15 Apr 2026

A Bayesian Perspective on the Role of Epistemic Uncertainty for Delayed Generalization in In-Context Learning

ResearchDGX agent

arXiv:2604.12434v1 Announce Type: cross Abstract: In-context learning enables transformers to adapt to new tasks from a few examples at inference time, while grokking highlights that this generalizati

A Bipartite Graph Approach to U.S.-China Cross-Market Return Forecasting

ResearchDGX agent

arXiv:2603.10559v2 Announce Type: replace Abstract: This paper studies cross-market return predictability through a machine learning framework that preserves economic structure. Exploiting the non-ove

A DeepONet for inverting the Neumann-to-Dirichlet Operator in Electrical Impedance Tomography: An approximation theoretic perspective and numerical results

ResearchDGX agent

arXiv:2407.17182v4 Announce Type: replace Abstract: In this work, we consider the non-invasive medical imaging modality of Electrical Impedance Tomography (EIT), where the goal is to recover the condu

A Geometric Algebra-informed NeRF Framework for Generalizable Wireless Channel Prediction

ApplicationsDGX agent

arXiv:2604.11983v1 Announce Type: cross Abstract: In this paper, we propose the geometric algebra-informed neural radiance fields (GAI-NeRF), a novel framework for wireless channel prediction that lev

A Large-Scale Comparative Analysis of Imputation Methods for Single-Cell RNA Sequencing Data

Model ReleasesDGX agent

arXiv:2603.24626v2 Announce Type: replace-cross Abstract: Background: Single-cell RNA sequencing (scRNA-seq) enables gene expression profiling at cellular resolution but is inherently affected by spar

A Nonparametric Adaptive EWMA Control Chart for Binary Monitoring of Multiple Stream Processes

ApplicationsDGX agent

arXiv:2604.12095v1 Announce Type: cross Abstract: Monitoring binomial proportions across multiple independent streams is a critical challenge in Statistical Process Control (SPC), with applications fr

A Residual-Shell-Based Lower Bound for Ollivier-Ricci Curvature

ResearchDGX agent

arXiv:2604.12211v1 Announce Type: new Abstract: Ollivier-Ricci curvature (ORC), defined via the Wasserstein distance that captures rich geometric information, has received growing attention in both th

A Theoretical Comparison of No-U-Turn Sampler Variants: Necessary and Sufficient Convergence Conditions and Mixing Time Analysis under Gaussian Targets

ResearchDGX agent

arXiv:2603.18640v3 Announce Type: replace-cross Abstract: The No-U-Turn Sampler (NUTS) is the computational workhorse of modern Bayesian software libraries, yet its qualitative and quantitative conver

A Two-Timescale Primal-Dual Framework for Reinforcement Learning via Online Dual Variable Guidance

SafetyDGX agent

arXiv:2505.04494v3 Announce Type: replace-cross Abstract: We study reinforcement learning by combining recent advances in regularized linear programming formulations with the classical theory of stoch

A unified data format for managing diabetes time-series data: DIAbetes eXchange (DIAX)

ResearchDGX agent

arXiv:2604.11944v1 Announce Type: new Abstract: Diabetes devices, including Continuous Glucose Monitoring (CGM), Smart Insulin Pens, and Automated Insulin Delivery systems, generate rich time-series d

Active Imitation Learning for Thermal- and Kernel-Aware LFM Inference on 3D S-NUCA Many-Cores

SafetyDGX agent

arXiv:2604.11948v1 Announce Type: new Abstract: Large Foundation Model (LFM) inference is both memory- and compute-intensive, traditionally relying on GPUs. However, the limited availability and high

Adaptive Budget Allocation in LLM-Augmented Surveys

ResearchDGX agent

arXiv:2604.12497v1 Announce Type: new Abstract: Large language models (LLMs) can generate survey responses at low cost, but their reliability varies substantially across questions and is unknown befor

Agentic Control in Variational Language Models

Local AiDGX agent

arXiv:2604.12513v1 Announce Type: new Abstract: We study whether a variational language model can support a minimal and measurable form of agentic control grounded in its own internal evidence. Our mo

Agentic LLM Reasoning in a Self-Driving Laboratory for Air-Sensitive Lithium Halide Spinel Conductors

AgentsDGX agent

arXiv:2604.11957v1 Announce Type: cross Abstract: Self-driving laboratories promise to accelerate materials discovery. Yet current automated solid-state synthesis platforms are limited to ambient cond

An Optimal Sauer Lemma Over k-ary Alphabets

ResearchDGX agent

arXiv:2604.12952v1 Announce Type: new Abstract: The Sauer-Shelah-Perles Lemma is a cornerstone of combinatorics and learning theory, bounding the size of a binary hypothesis class in terms of its Vapn

Analyzing the Effect of Noise in LLM Fine-tuning

Model ReleasesDGX agent

arXiv:2604.12469v1 Announce Type: new Abstract: Fine-tuning is the dominant paradigm for adapting pretrained large language models (LLMs) to downstream NLP tasks. In practice, fine-tuning datasets may

Beyond Weather Correlation: A Comparative Study of Static and Temporal Neural Architectures for Fine-Grained Residential Energy Consumption Forecasting in Melbourne, Australia

ApplicationsDGX agent

arXiv:2604.12304v1 Announce Type: new Abstract: Accurate short-term residential energy consumption forecasting at sub-hourly resolution is critical for smart grid management, demand response programme

BLOSSOM: Block-wise Federated Learning Over Shared and Sparse Observed Modalities

AgentsDGX agent

arXiv:2603.27552v2 Announce Type: replace Abstract: Multimodal federated learning (FL) is essential for real-world applications such as autonomous systems and healthcare, where data is distributed acr

Can LLMs Beat Classical Hyperparameter Optimization Algorithms? A Study on autoresearch

Model ReleasesDGX agent

arXiv:2603.24647v4 Announce Type: replace Abstract: The autoresearch repository enables an LLM agent to optimize hyperparameters by editing training code directly. We use it as a testbed to compare cl

Causal Diffusion Models for Counterfactual Outcome Distributions in Longitudinal Data

SafetyDGX agent

arXiv:2604.12992v1 Announce Type: cross Abstract: Predicting counterfactual outcomes in longitudinal data, where sequential treatment decisions heavily depend on evolving patient states, is critical y

CLAD: Efficient Log Anomaly Detection Directly on Compressed Representations

ResearchDGX agent

arXiv:2604.13024v1 Announce Type: new Abstract: The explosive growth of system logs makes streaming compression essential, yet existing log anomaly detection (LAD) methods incur severe pre-processing

← Previous
1…226227228229230…239
Next →