AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,432 results
Tutorials

Stable GFlowNets with Probabilistic Guarantees

DGX agent

arXiv:2605.01729v1 Announce Type: new Abstract: Generative Flow Networks (GFlowNets) learn to sample states proportional to an unnormalized reward. Despite their theoretical promise, practical trainin

tutorialsarxiv-cs-lg
5 May 2026
Research

Stable Localized Conformal Prediction via Transduction

AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2605.01452v1 Announce Type: cross Abstract: Existing evaluations of conformal prediction, such as prediction efficiency and test-conditional coverage, are defined in expectation over the calibra

researcharxiv-cs-lg
5 May 2026
Safety

STABLEVAL: Disagreement-Aware and Stable Evaluation of AI Systems

DGX agent

arXiv:2605.02122v1 Announce Type: new Abstract: Human evaluation remains the primary standard for assessing modern AI systems, yet annotator disagreement, bias, and variability make system rankings fr

safetyarxiv-cs-lg
5 May 2026
Model Releases

Standing on the Shoulders of Giants: Stabilized Knowledge Distillation for Cross--Language Code Clone Detection

DGX agent

arXiv:2605.02860v1 Announce Type: cross Abstract: Cross-language code clone detection (X-CCD) is challenging because semantically equivalent programs written in different languages often share little

model-releasesarxiv-cs-lg
5 May 2026
Research

STAR: Decode-Phase Rescheduling for LLM Inference

DGX agent

arXiv:2510.13668v2 Announce Type: replace-cross Abstract: Large Language Model (LLM) inference has emerged as a fundamental paradigm, however, variations in output length cause severe workload imbalan

researcharxiv-cs-lg
5 May 2026
Research

Statistical Consistency and Generalization of Contrastive Representation Learning

DGX agent

arXiv:2605.02116v1 Announce Type: new Abstract: Contrastive representation learning (CRL) underpins many modern foundation models. Despite recent theoretical progress, existing analyses suffer from se

researcharxiv-cs-lg
5 May 2026
Model Releases

Statistically-Lossless Quantization of Large Language Models

DGX agent

arXiv:2605.02404v1 Announce Type: new Abstract: Model quantization has become essential for efficient large language model deployment, yet existing approaches involve clear trade-offs: methods such as

model-releasesarxiv-cs-lg
5 May 2026
Applications

Stochastic Modeling of Human-Machine Authentication Channels under Partial Information Leakage

DGX agent

arXiv:2605.02102v1 Announce Type: cross Abstract: Reliable and secure human-machine communication is fundamental to IoT and cyber-physical ecosystems, where smartphones and wearables commonly serve as

applicationsarxiv-cs-lg
5 May 2026
Hardware

Stochastic Sparse Attention for Memory-Bound Inference

DGX agent

arXiv:2605.01910v1 Announce Type: new Abstract: Autoregressive decoding becomes bandwidth-limited at long contexts, as generating each token requires reading all n_k key and value vectors from KV cach

hardwarearxiv-cs-lg
5 May 2026
Model Releases

StreamIndex: Memory-Bounded Compressed Sparse Attention via Streaming Top-k

DGX agent

arXiv:2605.02568v1 Announce Type: new Abstract: DeepSeek-V3.2 and V4 introduce Compressed Sparse Attention (CSA): a lightning indexer (a learned scoring projection over compressed keys) scores them, t

model-releasesarxiv-cs-lg
5 May 2026
Research

Structured Analytic Coherent Point Drift for Non-Rigid Point Set Registration

DGX agent

arXiv:2605.00934v1 Announce Type: new Abstract: We introduce Analytic-CPD, a structured analytic variant of coherent point drift for non-rigid point set registration. The method retains the CPD poster

researcharxiv-cs-lg
5 May 2026
Model Releases

StyleShield: Exposing the Fragility of AIGC Detectors through Continuous Controllable Style Transfer

DGX agent

arXiv:2605.00924v1 Announce Type: new Abstract: AI-generated content (AIGC) detectors are increasingly deployed in high-stakes settings such as academic integrity screening, yet their reliability rest

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

Submodular Benchmark Selection

DGX agent

arXiv:2605.02209v1 Announce Type: cross Abstract: Evaluating large language models across many benchmarks is expensive, yet many benchmarks are highly correlated. We formalize the selection of a small

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

SURGE: SuperBatch Unified Resource-efficient GPU Encoding for Heterogeneous Partitioned Data

DGX agent

arXiv:2605.01060v1 Announce Type: cross Abstract: We present SURGE, a streaming GPU encoding system deployed in production to generate embeddings for over 800 million texts across 40,000 logical parti

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

SwiftChannel: Algorithm-Hardware Co-Design for Deep Learning-Based 5G Channel Estimation

DGX agent

arXiv:2605.01931v1 Announce Type: cross Abstract: Channel estimation is crucial in 5G communication networks for optimizing transmission parameters and ensuring reliable, high-speed communication. How

model-releasesarxiv-cs-lg
5 May 2026
Research

TetraJet-v2: Accurate NVFP4 Training for Large Language Models with Oscillation Suppression and Outlier Control

DGX agent

arXiv:2510.27527v2 Announce Type: replace Abstract: Large Language Models (LLMs) training is prohibitively expensive, driving interest in low-precision fully-quantized training (FQT). While novel 4-bi

researcharxiv-cs-lg
5 May 2026
Research

The Banach-Butterfly Invariant: Influence-Adaptive Walsh Geometry for Ternary Polynomial Threshold Functions

DGX agent

arXiv:2605.01637v1 Announce Type: new Abstract: We introduce the Banach-Butterfly Invariant (BBT), an influence-adaptive Banach geometry on the Walsh-Hadamard butterfly factorization. For a Boolean fu

researcharxiv-cs-lg
5 May 2026
Safety

The Case for ESM3 as a General-Purpose AI Model with Systemic Risk Under the EU AI Act

DGX agent

arXiv:2605.01611v1 Announce Type: cross Abstract: Due to ambiguity in the wording of the EU AI Act, we examine the question of to what extent frontier biological foundation models such as ESM3 are sub

safetyarxiv-cs-lg
5 May 2026
Research

The Causal Description Gap: Information-Theoretic Separations Across Pearl's Hierarchy

DGX agent

arXiv:2605.02177v1 Announce Type: cross Abstract: Pearl's causal hierarchy shows that observational, interventional, and counterfactual queries are qualitatively distinct. We ask a quantitative versio

researcharxiv-cs-lg
5 May 2026
Research

The elbow statistic: Multiscale clustering statistical significance

DGX agent

arXiv:2603.03235v2 Announce Type: replace-cross Abstract: Selecting the number of clusters remains a fundamental challenge in unsupervised learning. Existing approaches typically focus on identifying

researcharxiv-cs-lg
5 May 2026
Safety

The Geometric Inductive Bias of Grokking: Bypassing Phase Transitions via Architectural Topology

DGX agent

arXiv:2603.05228v3 Announce Type: replace Abstract: Mechanistic interpretability typically relies on post-hoc analysis of trained networks. We instead adopt an interventional approach: testing hypothe

safetyarxiv-cs-lg
5 May 2026
Safety

The Geometric Mechanics of Contrastive Representation Learning: Alignment Potentials, Entropic Dispersion, and Cross-modal Divergence

DGX agent

arXiv:2601.19597v3 Announce Type: replace Abstract: While InfoNCE underlies modern contrastive learning, its geometric mechanisms remain under-characterized beyond the canonical alignment--uniformity

safetyarxiv-cs-lg
5 May 2026
Model Releases

The Good, the Bad, and the Sampled: a No-Regret Approach to Safe Online Classification

DGX agent

arXiv:2510.01020v2 Announce Type: replace Abstract: We study sequential testing for a binary disease outcome when risk follows an unknown logistic model. At each round, the decision maker may either p

model-releasesarxiv-cs-lg
5 May 2026
Research

The (Marginal) Value of a Search Ad: An Online Causal Framework for Repeated Second-price Auctions

DGX agent

arXiv:2605.01756v1 Announce Type: cross Abstract: Existing auto-bidding algorithms in digital advertising often treat the value of an ad opportunity as the revenue obtained when an ad is shown and/or

researcharxiv-cs-lg
5 May 2026
Research

The Measure of Deception: An Analysis of Data Forging in Machine Unlearning

DGX agent

arXiv:2509.05865v2 Announce Type: replace Abstract: Motivated by privacy regulations and the need to mitigate the effects of harmful data, machine unlearning seeks to modify trained models so that the

researcharxiv-cs-lg
5 May 2026
Safety

The Model Knows, the Decoder Finds: Future Value Guided Particle Power Sampling

DGX agent

arXiv:2605.02427v1 Announce Type: cross Abstract: A recurring pattern in 'reasoning without training' is that base LLMs already assign non-trivial probability mass to correct multi-step solutions; the

safetyarxiv-cs-lg
5 May 2026
Research

The Norm-Separation Delay Law of Grokking: A First-Principles Theory of Delayed Generalization

DGX agent

arXiv:2603.13331v2 Announce Type: replace-cross Abstract: Grokking -- the sudden generalisation that appears long after a model has perfectly memorised its training data -- has been widely observed bu

researcharxiv-cs-lg
5 May 2026
Safety

The Partial Testimony of Logs: Evaluation of Language Model Generation under Confounded Model Choice

DGX agent

arXiv:2605.01311v1 Announce Type: new Abstract: Offline evaluation of language models from usage logs is biased when model choice is confounded: the same user-side factors that influence which model i

safetyarxiv-cs-lg
5 May 2026
Safety

The Pragmatic Frames of Spurious Correlations in Machine Learning: Interpreting How and Why They Matter

DGX agent

arXiv:2411.04696v5 Announce Type: replace Abstract: Learning correlations from data forms the foundation of today's machine learning (ML) and artificial intelligence research. While contemporary metho

safetyarxiv-cs-lg
5 May 2026
Model Releases

Think2SQL: Reinforce LLM Reasoning Capabilities for Text2SQL

DGX agent

arXiv:2504.15077v5 Announce Type: replace Abstract: Large Language Models (LLMs) can translate natural language into SQL, but small models struggle with multi-table and complex queries in Zero-Shot Le

model-releasesarxiv-cs-lg
5 May 2026
Applications

TIJERE: A Novel Threat Intelligence Joint Extraction Model Based on Analyst Expert Knowledge

DGX agent

arXiv:2605.02041v1 Announce Type: new Abstract: The extraction of entities and relationships from threat intelligence reports into structured formats, such as cybersecurity knowledge graphs, is essent

applicationsarxiv-cs-lg
5 May 2026
Tutorials

Time-series forecasting through the lens of dynamics

DGX agent

arXiv:2507.15774v2 Announce Type: replace Abstract: While deep learning is facing an homogenization across modalities led by Transformers, they are still challenged by shallow linear models in the tim

tutorialsarxiv-cs-lg
5 May 2026
Research

Token-Efficient Change Detection in LLM APIs

DGX agent

arXiv:2602.11083v2 Announce Type: replace Abstract: Remote change detection in LLMs is a difficult problem. Existing methods are either too expensive for deployment at scale, or require initial white-

researcharxiv-cs-lg
5 May 2026
Safety

Topological Neural Tangent Kernel

DGX agent

arXiv:2605.01110v1 Announce Type: new Abstract: Graph neural tangent kernels give a principled infinite-width theory for graph neural networks, but inherit a basic limitation of graph models: they see

safetyarxiv-cs-lg
5 May 2026
Research

Toward a foundational thermal model for residential buildings

DGX agent

arXiv:2605.01364v1 Announce Type: new Abstract: The building energy community lacks a foundational thermal model, i.e., a single pretrained model capable of generalizing across diverse buildings, clim

researcharxiv-cs-lg
5 May 2026
Research

Toward Resilient 5G Networks: Comparative Analysis of Federated and Centralized Learning for RF Jamming Detection

DGX agent

arXiv:2605.01705v1 Announce Type: cross Abstract: Jamming attacks are proliferating and pose a significant threat to the security of 5G and beyond networks. These attacks target 5G radio frequency (RF

researcharxiv-cs-lg
5 May 2026
Safety

Towards Efficient and Expressive Offline RL via Flow-Anchored Noise-conditioned Q-Learning

DGX agent

arXiv:2605.01663v1 Announce Type: new Abstract: We propose Flow-Anchored Noise-conditioned Q-Learning (FAN), a highly efficient and high-performing offline reinforcement learning (RL) algorithm. Recen

safetyarxiv-cs-lg
5 May 2026
Research

Towards Systematic Generalization for Power Grid Optimization Problems

DGX agent

arXiv:2605.02026v1 Announce Type: new Abstract: AC Optimal Power Flow (ACOPF) and Security-Constrained Unit Commitment (SCUC) are fundamental optimization problems in power system operations. ACOPF se

researcharxiv-cs-lg
5 May 2026
Model Releases

TRACED: In vivo imaging of extracellular intrinsic diffusivity, tortuosity, cell size distribution and cell density in human glioma patients

DGX agent

arXiv:2605.02615v1 Announce Type: cross Abstract: The lack of analytical models describing diffusion time dependence at intermediate time scales in complex tissue microstructure limits the accurate qu

model-releasesarxiv-cs-lg
5 May 2026
Safety

Training Non-Differentiable Networks via Optimal Transport

DGX agent

arXiv:2605.01928v1 Announce Type: new Abstract: Neural networks increasingly embed non-differentiable components (spiking neurons, quantized layers, discrete routing, blackbox simulators, etc.) where

safetyarxiv-cs-lg
5 May 2026
Research

Transfer Learning for Tonal Noise Prediction in VRF Units Using Thermodynamic and Vibration Signals

DGX agent

arXiv:2605.00895v1 Announce Type: cross Abstract: The second-order harmonic (2f) component generated by twin-rotary compressor is a dominant low-frequency noise source of variable refrigerant flow (VR

researcharxiv-cs-lg
5 May 2026
Safety

TRAP: Tail-aware Ranking Attack for World-Model Planning

DGX agent

arXiv:2605.01950v1 Announce Type: new Abstract: World models enable long-horizon planning by internally generating and evaluating imagined trajectories, making them a promising foundation for generali

safetyarxiv-cs-lg
5 May 2026
Research

Trees and Graphs with Non Log-concave Dominating Set Sequence via AI Tools

DGX agent

arXiv:2605.02193v1 Announce Type: cross Abstract: We give new examples of graphs and trees with dominating set sequences that are not log-concave. These examples were generated by PatternBoost, a tran

researcharxiv-cs-lg
5 May 2026
Local Ai

Trust, but Verify: Peeling Low-Bit Transformer Networks for Training Monitoring

DGX agent

arXiv:2605.02853v1 Announce Type: new Abstract: Understanding whether deep neural networks are effectively optimized remains challenging, as training occurs in highly nonconvex landscapes and standard

local-aiarxiv-cs-lg
5 May 2026
Tutorials

U-Define: Designing User Workflows for Hard and Soft Constraints in LLM-Based Planning

DGX agent

arXiv:2605.02765v1 Announce Type: cross Abstract: LLMs are increasingly used for end-user task planning, yet their black-box nature limits users' ability to ensure reliability and control. While recen

tutorialsarxiv-cs-lg
5 May 2026
Local Ai

Ultrafast On-chip Online Learning via Spline Locality in Kolmogorov-Arnold Networks

DGX agent

arXiv:2602.02056v2 Announce Type: replace-cross Abstract: Ultrafast online learning is essential for high-frequency systems, such as controls for quantum computing and nuclear fusion, where adaptation

local-aiarxiv-cs-lg
5 May 2026
Safety

Understanding Adversarial Imitation Learning in Small Sample Regime: A Stage-coupled Analysis

DGX agent

arXiv:2208.01899v2 Announce Type: replace Abstract: Imitation learning learns a policy from expert trajectories. While the expert data is believed to be crucial for imitation quality, it was found tha

safetyarxiv-cs-lg
5 May 2026
Model Releases

Understanding Emergent Misalignment via Feature Superposition Geometry

DGX agent

arXiv:2605.00842v1 Announce Type: cross Abstract: Emergent misalignment, where fine-tuning on narrow, non-harmful tasks induces harmful behaviors, poses a key challenge for AI safety in LLMs. Despite

model-releasesarxiv-cs-lg
5 May 2026
← Previous
1…240241242243244…301
Next →