AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,557 results
18 May 2026

Rule2DRC: Benchmarking LLM Agents for DRC Script Synthesis with Execution-Guided Test Generation

Model ReleasesDGX agent

arXiv:2605.15669v1 Announce Type: new Abstract: Manufacturable chip layouts must satisfy thousands of geometry-based design rules, and design rule checking (DRC) enforces them by running executable DR

Runtime-Orchestrated Second-Order Optimization for Scalable LLM Training

Model ReleasesDGX agent

arXiv:2605.16184v1 Announce Type: cross Abstract: Second-order methods offer an attractive path toward more sample-efficient LLM training, but their practical use is often blocked by the systems cost

SAFE Quantum Machine Learning with Variational Quantum Classifiers

SafetyDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2605.16067v1 Announce Type: new Abstract: We propose a variational quantum classifier operating on high dimensional deep representations via amplitude encoding, stabilized by a learnable classic

Sampling-Free Privacy Accounting for Matrix Mechanisms under Random Allocation

ResearchDGX agent

arXiv:2601.21636v3 Announce Type: replace Abstract: We study privacy amplification for differentially private model training with matrix factorization under random allocation (also known as the balls-

Scalable neuromorphic computing from autonomous spiking dynamics in a clockless reconfigurable chip

AgentsDGX agent

arXiv:2605.16114v1 Announce Type: cross Abstract: We propose a scalable neuromorphic architecture based on spiking dynamics emerging from the autonomous time-continuous evolution of clockless (asynchr

Searching on a Budget: HW-NAS with 10 Latency Probes

Model ReleasesDGX agent

arXiv:2504.00663v2 Announce Type: replace Abstract: Existing hardware-aware NAS (HW-NAS) methods typically assume access to precise information circa the target device, either via analytical approxima

SEED: Targeted Data Selection by Weighted Independent Set

Local AiDGX agent

arXiv:2605.15691v1 Announce Type: new Abstract: Data selection seeks to identify a compact yet informative subset from large-scale training corpora, balancing sample quality against collection diversi

SGNO: Spectral Generator Neural Operators for Stable Long Horizon PDE Rollouts

ResearchDGX agent

arXiv:2602.18801v2 Announce Type: replace Abstract: Autoregressive neural PDE surrogates predict future states by repeatedly applying a learned one-step operator. This is a simple and widely used meth

Skew-adaptive conformal prediction

TutorialsDGX agent

arXiv:2605.16145v1 Announce Type: cross Abstract: We develop a skew-adaptive extension of split conformal prediction for regression. The method starts from an asymmetric interval family centered at a

SKILL0: In-Context Agentic Reinforcement Learning for Skill Internalization

SafetyDGX agent

arXiv:2604.02268v2 Announce Type: replace Abstract: Agent skills, structured packages of procedural knowledge and executable resources that agents dynamically load at inference time, have become a rel

Spectral Priors vs. Attention: Investigating the Utility of Attention Mechanisms in EEG-Based Diagnosis

ResearchDGX agent

arXiv:2605.15433v1 Announce Type: new Abstract: Electroencephalograph (EEG) timeseries signals are characterized by significant noise and coarse spatial resolution, which complicates the classificatio

Stochastic Compositional Optimization via Hybrid Momentum Frank--Wolfe

ResearchDGX agent

arXiv:2605.15350v1 Announce Type: cross Abstract: Stochastic compositional optimization minimizes objectives of the form min_{m{x} in X} F(m{f}(m{x}), m{x}), where m{f} is accessible only through nois

Stochastic Non-Smooth Convex Optimization with Unbounded Gradients

ResearchDGX agent

arXiv:2605.15522v1 Announce Type: cross Abstract: Much of the existing theory on first-order non-smooth optimization is built on a restrictive assumption that the gradients of the objective function a

SurvivalPFN: Amortizing Survival Prediction via In-Context Bayesian Inference

Model ReleasesDGX agent

arXiv:2605.15488v1 Announce Type: new Abstract: Survival analysis provides a powerful statistical framework for modeling time-to-event outcomes in the presence of censoring. However, selecting an appr

SwAIther-Precip: Lead-Time-Aware Bias Correction Enables Kilometer-Scale Downscaling of Global AI Precipitation Forecasts over Switzerland

Local AiDGX agent

arXiv:2605.16163v1 Announce Type: cross Abstract: Skillful medium-range precipitation forecasting at kilometer scale remains challenging over complex terrain because precipitation arises from multisca

T2T-LA: A Topology-to-Topology LLM Agent for Graph Learning with Neither Feature Access nor Task Knowledge

Model ReleasesDGX agent

arXiv:2512.08964v4 Announce Type: replace Abstract: Graph learning aims to convert data into graph representations, which are fundamental to many problems in machine learning for CAD, where circuits,

Tadpole: Autoencoders as Foundation Models for 3D PDEs with Online Learning

Model ReleasesDGX agent

arXiv:2605.15284v1 Announce Type: new Abstract: We introduce Tadpole, a novel foundation model for three-dimensional partial differential equations (PDEs) that addresses key challenges in transferabil

Talking Trees: Reasoning-Assisted Induction of Decision Trees for Tabular Data

AgentsDGX agent

arXiv:2509.21465v3 Announce Type: replace Abstract: Tabular foundation models are becoming increasingly popular for low-resource tabular problems. These models make up for small training datasets by p

TeamTR: Trust-Region Fine-Tuning for Multi-Agent LLM Coordination

AgentsDGX agent

arXiv:2605.15207v1 Announce Type: new Abstract: Multi-agent LLM systems have shown promise for complex reasoning, yet recent evaluations reveal they often underperform single-model baselines. We ident

Testing properties of trees in graphical models with covariance queries

ResearchDGX agent

arXiv:2605.15996v1 Announce Type: cross Abstract: We consider the problem of testing properties of graphs underlying high-dimensional graphical models. We adopt the model of covariance queries introdu

The Privacy Price of Tail-Risk Learning: Effective Tail Sample Size in Differentially Private CVaR Optimization

ResearchDGX agent

arXiv:2605.16219v1 Announce Type: new Abstract: Differential privacy changes the effective sample size governing CVaR learning. For tail mass au, the privacy-relevant sample size is not n, but nau; eq

Ti-iLSTM: A TinyDL Approach for Logic-Level Anomaly Detection in Industrial Water Treatment Systems

Local AiDGX agent

arXiv:2605.15874v1 Announce Type: new Abstract: Industrial Water Treatment Systems (IWTS) are safety critical cyber-physical infrastructures and due to increased connectivity, these systems are expose

Tighter Regret Bounds for Contextual Action-Set Reinforcement Learning

ResearchDGX agent

arXiv:2605.15692v1 Announce Type: new Abstract: We study episodic reinforcement learning with fixed reward and transition functions, but with episode-dependent admissible action sets that are observed

Time-Varying Deep State Space Models for Sequences with Switching Dynamics

ResearchDGX agent

arXiv:2605.15311v1 Announce Type: new Abstract: The identification and modeling of time-varying systems is a fundamental challenge in signal processing and system identification. To address this chall

Toward World Modeling of Physiological Signals with Chaos-Theoretic Balancing and Latent Dynamics

ApplicationsDGX agent

arXiv:2605.15465v1 Announce Type: new Abstract: Physiological time series signals reflect complex, multi-scale dynamical processes of the human body. Existing modeling studies focus on static tasks su

Towards a more realistic evaluation of machine learning models for bearing fault diagnosis

SafetyDGX agent

arXiv:2509.22267v4 Announce Type: replace Abstract: Reliable detection of bearing faults is essential for maintaining the safety and operational efficiency of rotating machinery. While recent advances

Towards Code-Oriented LM Embeddings for Surrogate-Assisted Neural Architecture Search

SafetyDGX agent

arXiv:2605.15649v1 Announce Type: new Abstract: Developing effective surrogates (performance predictors) for Neural Architecture Search (NAS) typically requires expensive fine-tuning or the engineerin

Towards Efficient Large Language Reasoning Models via Extreme-Ratio Chain-of-Thought Compression

Model ReleasesDGX agent

arXiv:2602.08324v3 Announce Type: replace Abstract: Chain-of-Thought (CoT) reasoning successfully enhances the reasoning capabilities of Large Language Models (LLMs), yet it incurs substantial computa

Training on Documents About Monitoring Leads to CoT Obfuscation

AgentsDGX agent

arXiv:2605.15257v1 Announce Type: new Abstract: Chain-of-thought (CoT) monitoring is one of the most promising tools we have for detecting model misbehavior, but its effectiveness depends on models fa

Transformer-like Inference from Optimal Control

ResearchDGX agent

arXiv:2605.15608v1 Announce Type: new Abstract: Decoder-only transformers compute the conditional probability of the next token from a sequence of past observations. This paper derives, from first pri

Transformer Scalability Crisis: The First Comprehensive Empirical Analysis of Performance Walls in Modern Language Models

Model ReleasesDGX agent

arXiv:2605.15413v1 Announce Type: new Abstract: Despite the remarkable success of transformer architectures in natural language processing, their scalability limitations remain poorly understood throu

Transformers are Inherently Succinct

ResearchDGX agent

arXiv:2510.19315v3 Announce Type: replace-cross Abstract: We study succinctness as a measure of the expressive power of transformers. Succinctness -- how compactly a formalism can describe a language

Unified High-Probability Analysis of Stochastic Variance-Reduced Estimation

SafetyDGX agent

arXiv:2605.15388v1 Announce Type: new Abstract: Stochastic estimators are fundamental to large-scale optimization, where population quantities must be inferred from noisy oracle observations. Although

Unified Simulation of Lagrangian Particle Dynamics via Transformer

ApplicationsDGX agent

arXiv:2605.15305v1 Announce Type: cross Abstract: A unified simulator that can model diverse physical phenomena without solver-specific redesign is a long-standing goal across simulation science. We p

Universal Magnetic Structure Prediction from Atomic Coordinates with Near-Experimental Accuracy

ResearchDGX agent

arXiv:2605.16230v1 Announce Type: cross Abstract: Magnetic order is a fundamental property of materials, governing collective behavior and enabling a broad range of functionalities. Yet magnetic struc

Unsupervised Domain Shift Detection with Interpretable Subspace Attribution

ResearchDGX agent

arXiv:2605.15920v1 Announce Type: cross Abstract: We developed a tool for detecting domain shifts, namely subtle differences in the probability distributions of datasets. We identify these shifts usin

Unveiling the Black Box: A Multi-Layer Framework for Explaining Reinforcement Learning-Based Cyber Agents

SafetyDGX agent

arXiv:2505.11708v3 Announce Type: replace-cross Abstract: Reinforcement Learning (RL) agents are increasingly used to simulate sophisticated cyberattacks, but their decision-making processes remain op

Variational Autoregressive Networks with probability priors

ResearchDGX agent

arXiv:2605.16020v1 Announce Type: new Abstract: Monte Carlo methods are essential across diverse scientific fields, yet their efficiency is frequently hampered by critical slowing down-a sharp increas

Vector-valued self-normalized concentration inequalities beyond sub-Gaussianity

ResearchDGX agent

arXiv:2511.03606v2 Announce Type: replace-cross Abstract: The study of self-normalized processes plays a crucial role in a wide range of applications, from sequential decision-making to econometrics.

Viability of perturbative expansion for quantum field theories on neurons

ResearchDGX agent

arXiv:2508.03810v4 Announce Type: replace-cross Abstract: Neural Network (NN) architectures that break statistical independence of parameters have been proposed as a new approach for simulating local

What Is Preference Optimization Doing, and Why?

SafetyDGX agent

arXiv:2512.00778v2 Announce Type: replace Abstract: Preference optimization (PO) is indispensable for large language models (LLMs), with methods such as direct preference optimization (DPO) and proxim

15 May 2026

A Block Coordinate Descent Method for Nonsmooth Composite Optimization under Orthogonality Constraints

ResearchDGX agent

arXiv:2304.03641v4 Announce Type: replace-cross Abstract: Nonsmooth composite optimization with orthogonality constraints has a wide range of applications in statistical learning and data science. How

A Hardware-Aware, Per-Layer Methodology for Post-Training Quantization of Large Language Models

ResearchDGX agent

arXiv:2605.14929v1 Announce Type: new Abstract: Scaled Outer Product (SOP) is a post-training quantization methodology for large language model weights, designed to deliver near-lossless fidelity at 4

A Mutual Information Lower Bound for Multimodal Regression Active Learning

ResearchDGX agent

arXiv:2605.14917v1 Announce Type: new Abstract: Active learning for continuous regression has lacked an acquisition function that targets epistemic uncertainty when the predictive distribution is mult

A New Framework for Convex Clustering in Kernel Spaces: Finite Sample Bounds, Consistency and Performance Insights

ApplicationsDGX agent

arXiv:2511.05159v2 Announce Type: replace-cross Abstract: Convex clustering is a well-regarded clustering method, resembling the similar centroid-based approach of Lloyd's k-means, without requiring a

A Non-Monotone Preconditioned Trust-Region Method for Neural Network Training

ResearchDGX agent

arXiv:2605.14860v1 Announce Type: cross Abstract: Training deep neural networks at scale can benefit from domain decomposition, where the network is split into subdomains trained in parallel and coupl

A Novel Schur-Decomposition-Based Weight Projection Method for Stable State-Space Neural-Network Architectures

ApplicationsDGX agent

arXiv:2605.14489v1 Announce Type: new Abstract: Building black-box models for dynamical systems from data is a challenging problem in machine learning, especially when asymptotic stability guarantees

A Survey on Data-Dependent Worst-Case Generalization Bounds

Model ReleasesDGX agent

arXiv:2605.13913v1 Announce Type: cross Abstract: Deep neural networks generalize well despite being heavily overparameterized, in apparent contradiction with classical learning theory based on unifor

A Systematic Evaluation of Imbalance Handling Methods in Biomedical Binary Classification

ResearchDGX agent

arXiv:2605.14147v1 Announce Type: new Abstract: Objective: The primary goal of this study was to systematically examine the impact of commonly used imbalance handling methods (IHMs) on predictive perf

A Unified Geometric Framework for Weighted Contrastive Learning

ResearchDGX agent

arXiv:2605.13943v1 Announce Type: new Abstract: Contrastive learning (CL) aims to preserve relational structure between samples by learning representations that reflect a similarity graph. Yet, the ge

AaSP: Aliasing-aware Self-Supervised Pre-Training for Audio Spectrogram Transformers

TutorialsDGX agent

arXiv:2512.03637v2 Announce Type: replace-cross Abstract: Transformer-based audio self-supervised learning (SSL) models commonly use spectrograms, vision-style Transformers, and masked modeling object

Accelerated Sequential Flow Matching: A Bayesian Filtering Perspective

ResearchDGX agent

arXiv:2602.05319v3 Announce Type: replace Abstract: Sequential probabilistic inference from streaming observations requires modeling distributions over future trajectories as new observations arrive.

ActivePusher: Active Learning and Planning with Residual Physics for Nonprehensile Manipulation

SafetyDGX agent

arXiv:2506.04646v4 Announce Type: replace-cross Abstract: Planning with learned dynamics models offers a promising approach toward versatile real-world manipulation, particularly in nonprehensile sett

AIMing for Standardised Explainability Evaluation in GNNs: A Framework and Case Study on Graph Kernel Networks

ApplicationsDGX agent

arXiv:2605.14884v1 Announce Type: new Abstract: Graph Neural Networks (GNNs) have advanced significantly in handling graph-structured data, but a comprehensive framework for evaluating explainability

All-atomistic Transferable Neural Potentials for Protein Solvation

ResearchDGX agent

arXiv:2605.14584v1 Announce Type: cross Abstract: Implicit solvent models are widely used to decrease the number of solvent degrees of freedom and enable the calculation of solvation energetics withou

An Interpretable Latency Model for Speculative Decoding in LLM Serving

ApplicationsDGX agent

arXiv:2605.15051v1 Announce Type: new Abstract: Speculative decoding (SD) accelerates large language model (LLM) inference by using a smaller draft model to propose multiple tokens that are verified b

AQKA: Active Quantum Kernel Acquisition Under a Shot Budget

SafetyDGX agent

arXiv:2605.14672v1 Announce Type: new Abstract: Estimating an N imes N quantum kernel from circuit fidelities requires Theta(N^2 S) measurement shots, the dominant bottleneck for deployment on near-te

Attention-Based Multimodal Survival Prediction with Cross-Modal Bilinear Fusion

Model ReleasesDGX agent

arXiv:2605.13897v1 Announce Type: cross Abstract: We propose a novel multimodal deep learning framework for patient-level survival prediction, which integrates whole-slide histology features, RNA-seq

Average Gradient Outer Product in kernel regression provably recovers the central subspace for multi-index models

ResearchDGX agent

arXiv:2605.15082v1 Announce Type: cross Abstract: We study a prototypical situation when a learned predictor can discover useful low-dimensional structure in data, while using fewer samples than are n

BCI-Based Assessment of Ocular Response Time Using Dynamic Time Warping Leveraging an RDWT-Driven Deep Neural Framework

ResearchDGX agent

arXiv:2605.14883v1 Announce Type: cross Abstract: Mild traumatic brain injury (mTBI) is a prevalent condition that remains difficult to diagnose in its early stages. Oculomotor dysfunction is a well-e

← Previous
1…153154155156157…243
Next →