AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,432 results
5 May 2026

Submodular Benchmark Selection

Model ReleasesDGX agent

arXiv:2605.02209v1 Announce Type: cross Abstract: Evaluating large language models across many benchmarks is expensive, yet many benchmarks are highly correlated. We formalize the selection of a small

SURGE: SuperBatch Unified Resource-efficient GPU Encoding for Heterogeneous Partitioned Data

Model ReleasesDGX agent

arXiv:2605.01060v1 Announce Type: cross Abstract: We present SURGE, a streaming GPU encoding system deployed in production to generate embeddings for over 800 million texts across 40,000 logical parti

SwiftChannel: Algorithm-Hardware Co-Design for Deep Learning-Based 5G Channel Estimation

Model ReleasesDGX agent

arXiv:2605.01931v1 Announce Type: cross Abstract: Channel estimation is crucial in 5G communication networks for optimizing transmission parameters and ensuring reliable, high-speed communication. How


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

TetraJet-v2: Accurate NVFP4 Training for Large Language Models with Oscillation Suppression and Outlier Control

ResearchDGX agent

arXiv:2510.27527v2 Announce Type: replace Abstract: Large Language Models (LLMs) training is prohibitively expensive, driving interest in low-precision fully-quantized training (FQT). While novel 4-bi

The Banach-Butterfly Invariant: Influence-Adaptive Walsh Geometry for Ternary Polynomial Threshold Functions

ResearchDGX agent

arXiv:2605.01637v1 Announce Type: new Abstract: We introduce the Banach-Butterfly Invariant (BBT), an influence-adaptive Banach geometry on the Walsh-Hadamard butterfly factorization. For a Boolean fu

The Case for ESM3 as a General-Purpose AI Model with Systemic Risk Under the EU AI Act

SafetyDGX agent

arXiv:2605.01611v1 Announce Type: cross Abstract: Due to ambiguity in the wording of the EU AI Act, we examine the question of to what extent frontier biological foundation models such as ESM3 are sub

The Causal Description Gap: Information-Theoretic Separations Across Pearl's Hierarchy

ResearchDGX agent

arXiv:2605.02177v1 Announce Type: cross Abstract: Pearl's causal hierarchy shows that observational, interventional, and counterfactual queries are qualitatively distinct. We ask a quantitative versio

The elbow statistic: Multiscale clustering statistical significance

ResearchDGX agent

arXiv:2603.03235v2 Announce Type: replace-cross Abstract: Selecting the number of clusters remains a fundamental challenge in unsupervised learning. Existing approaches typically focus on identifying

The Geometric Inductive Bias of Grokking: Bypassing Phase Transitions via Architectural Topology

SafetyDGX agent

arXiv:2603.05228v3 Announce Type: replace Abstract: Mechanistic interpretability typically relies on post-hoc analysis of trained networks. We instead adopt an interventional approach: testing hypothe

The Geometric Mechanics of Contrastive Representation Learning: Alignment Potentials, Entropic Dispersion, and Cross-modal Divergence

SafetyDGX agent

arXiv:2601.19597v3 Announce Type: replace Abstract: While InfoNCE underlies modern contrastive learning, its geometric mechanisms remain under-characterized beyond the canonical alignment--uniformity

The Good, the Bad, and the Sampled: a No-Regret Approach to Safe Online Classification

Model ReleasesDGX agent

arXiv:2510.01020v2 Announce Type: replace Abstract: We study sequential testing for a binary disease outcome when risk follows an unknown logistic model. At each round, the decision maker may either p

The (Marginal) Value of a Search Ad: An Online Causal Framework for Repeated Second-price Auctions

ResearchDGX agent

arXiv:2605.01756v1 Announce Type: cross Abstract: Existing auto-bidding algorithms in digital advertising often treat the value of an ad opportunity as the revenue obtained when an ad is shown and/or

The Measure of Deception: An Analysis of Data Forging in Machine Unlearning

ResearchDGX agent

arXiv:2509.05865v2 Announce Type: replace Abstract: Motivated by privacy regulations and the need to mitigate the effects of harmful data, machine unlearning seeks to modify trained models so that the

The Model Knows, the Decoder Finds: Future Value Guided Particle Power Sampling

SafetyDGX agent

arXiv:2605.02427v1 Announce Type: cross Abstract: A recurring pattern in 'reasoning without training' is that base LLMs already assign non-trivial probability mass to correct multi-step solutions; the

The Norm-Separation Delay Law of Grokking: A First-Principles Theory of Delayed Generalization

ResearchDGX agent

arXiv:2603.13331v2 Announce Type: replace-cross Abstract: Grokking -- the sudden generalisation that appears long after a model has perfectly memorised its training data -- has been widely observed bu

The Partial Testimony of Logs: Evaluation of Language Model Generation under Confounded Model Choice

SafetyDGX agent

arXiv:2605.01311v1 Announce Type: new Abstract: Offline evaluation of language models from usage logs is biased when model choice is confounded: the same user-side factors that influence which model i

The Pragmatic Frames of Spurious Correlations in Machine Learning: Interpreting How and Why They Matter

SafetyDGX agent

arXiv:2411.04696v5 Announce Type: replace Abstract: Learning correlations from data forms the foundation of today's machine learning (ML) and artificial intelligence research. While contemporary metho

Think2SQL: Reinforce LLM Reasoning Capabilities for Text2SQL

Model ReleasesDGX agent

arXiv:2504.15077v5 Announce Type: replace Abstract: Large Language Models (LLMs) can translate natural language into SQL, but small models struggle with multi-table and complex queries in Zero-Shot Le

TIJERE: A Novel Threat Intelligence Joint Extraction Model Based on Analyst Expert Knowledge

ApplicationsDGX agent

arXiv:2605.02041v1 Announce Type: new Abstract: The extraction of entities and relationships from threat intelligence reports into structured formats, such as cybersecurity knowledge graphs, is essent

Time-series forecasting through the lens of dynamics

TutorialsDGX agent

arXiv:2507.15774v2 Announce Type: replace Abstract: While deep learning is facing an homogenization across modalities led by Transformers, they are still challenged by shallow linear models in the tim

Token-Efficient Change Detection in LLM APIs

ResearchDGX agent

arXiv:2602.11083v2 Announce Type: replace Abstract: Remote change detection in LLMs is a difficult problem. Existing methods are either too expensive for deployment at scale, or require initial white-

Topological Neural Tangent Kernel

SafetyDGX agent

arXiv:2605.01110v1 Announce Type: new Abstract: Graph neural tangent kernels give a principled infinite-width theory for graph neural networks, but inherit a basic limitation of graph models: they see

Toward a foundational thermal model for residential buildings

ResearchDGX agent

arXiv:2605.01364v1 Announce Type: new Abstract: The building energy community lacks a foundational thermal model, i.e., a single pretrained model capable of generalizing across diverse buildings, clim

Toward Resilient 5G Networks: Comparative Analysis of Federated and Centralized Learning for RF Jamming Detection

ResearchDGX agent

arXiv:2605.01705v1 Announce Type: cross Abstract: Jamming attacks are proliferating and pose a significant threat to the security of 5G and beyond networks. These attacks target 5G radio frequency (RF

Towards Efficient and Expressive Offline RL via Flow-Anchored Noise-conditioned Q-Learning

SafetyDGX agent

arXiv:2605.01663v1 Announce Type: new Abstract: We propose Flow-Anchored Noise-conditioned Q-Learning (FAN), a highly efficient and high-performing offline reinforcement learning (RL) algorithm. Recen

Towards Systematic Generalization for Power Grid Optimization Problems

ResearchDGX agent

arXiv:2605.02026v1 Announce Type: new Abstract: AC Optimal Power Flow (ACOPF) and Security-Constrained Unit Commitment (SCUC) are fundamental optimization problems in power system operations. ACOPF se

TRACED: In vivo imaging of extracellular intrinsic diffusivity, tortuosity, cell size distribution and cell density in human glioma patients

Model ReleasesDGX agent

arXiv:2605.02615v1 Announce Type: cross Abstract: The lack of analytical models describing diffusion time dependence at intermediate time scales in complex tissue microstructure limits the accurate qu

Training Non-Differentiable Networks via Optimal Transport

SafetyDGX agent

arXiv:2605.01928v1 Announce Type: new Abstract: Neural networks increasingly embed non-differentiable components (spiking neurons, quantized layers, discrete routing, blackbox simulators, etc.) where

Transfer Learning for Tonal Noise Prediction in VRF Units Using Thermodynamic and Vibration Signals

ResearchDGX agent

arXiv:2605.00895v1 Announce Type: cross Abstract: The second-order harmonic (2f) component generated by twin-rotary compressor is a dominant low-frequency noise source of variable refrigerant flow (VR

TRAP: Tail-aware Ranking Attack for World-Model Planning

SafetyDGX agent

arXiv:2605.01950v1 Announce Type: new Abstract: World models enable long-horizon planning by internally generating and evaluating imagined trajectories, making them a promising foundation for generali

Trees and Graphs with Non Log-concave Dominating Set Sequence via AI Tools

ResearchDGX agent

arXiv:2605.02193v1 Announce Type: cross Abstract: We give new examples of graphs and trees with dominating set sequences that are not log-concave. These examples were generated by PatternBoost, a tran

Trust, but Verify: Peeling Low-Bit Transformer Networks for Training Monitoring

Local AiDGX agent

arXiv:2605.02853v1 Announce Type: new Abstract: Understanding whether deep neural networks are effectively optimized remains challenging, as training occurs in highly nonconvex landscapes and standard

U-Define: Designing User Workflows for Hard and Soft Constraints in LLM-Based Planning

TutorialsDGX agent

arXiv:2605.02765v1 Announce Type: cross Abstract: LLMs are increasingly used for end-user task planning, yet their black-box nature limits users' ability to ensure reliability and control. While recen

Ultrafast On-chip Online Learning via Spline Locality in Kolmogorov-Arnold Networks

Local AiDGX agent

arXiv:2602.02056v2 Announce Type: replace-cross Abstract: Ultrafast online learning is essential for high-frequency systems, such as controls for quantum computing and nuclear fusion, where adaptation

Understanding Adversarial Imitation Learning in Small Sample Regime: A Stage-coupled Analysis

SafetyDGX agent

arXiv:2208.01899v2 Announce Type: replace Abstract: Imitation learning learns a policy from expert trajectories. While the expert data is believed to be crucial for imitation quality, it was found tha

Understanding Emergent Misalignment via Feature Superposition Geometry

Model ReleasesDGX agent

arXiv:2605.00842v1 Announce Type: cross Abstract: Emergent misalignment, where fine-tuning on narrow, non-harmful tasks induces harmful behaviors, poses a key challenge for AI safety in LLMs. Despite

Universality in Deep Neural Networks: An approach via the Lindeberg exchange principle

ResearchDGX agent

arXiv:2605.02771v1 Announce Type: cross Abstract: We consider the infinite-width limit of a fully connected deep neural network with general weights, and we prove quantitative general bounds on the 2-

Unsupervised full-field Bayesian inference of orthotropic hyperelasticity from a single biaxial test: a myocardial case study

Model ReleasesDGX agent

arXiv:2510.09498v3 Announce Type: replace-cross Abstract: Cardiac muscle tissue exhibits highly non-linear hyperelastic and orthotropic material behavior during passive deformation. Traditional consti

Unsupervised Machine Learning for Detecting Structural Anomalies in European Regional Statistics

SafetyDGX agent

arXiv:2605.02884v1 Announce Type: new Abstract: Ensuring the coherence of regional socio-economic statistics is a central task for national statistical institutes. Traditional validation tools, such a

Value Functions for Temporal Logic: Optimal Policies and Safety Filters

SafetyDGX agent

arXiv:2605.01051v1 Announce Type: cross Abstract: While Bellman equations for basic reach, avoid, and reach-avoid problems are well studied, the relationship between value optimality and policy optima

Variational Matrix-Learning Fourier Networks for Parametric Multiphysics Surrogates

Model ReleasesDGX agent

arXiv:2605.02280v1 Announce Type: new Abstract: Multiphysics simulation is critical for system-technology co-optimization (STCO) in chiplet-based design, but repeated finite-element solutions of PDE-g

Visual Latents Know More Than They Say: Unsilencing Latent Reasoning in MLLMs

Model ReleasesDGX agent

arXiv:2605.02735v1 Announce Type: new Abstract: Continuous latent-space reasoning offers a compact alternative to textual chain-of-thought for multimodal models, enabling high-dimensional visual evide

Visualizing Critic Match Loss Landscapes for Interpretation of Online Reinforcement Learning Control Algorithms

Model ReleasesDGX agent

arXiv:2603.14535v2 Announce Type: replace Abstract: Reinforcement learning has proven its power on various occasions. However, its performance is not always guaranteed when system dynamics change. Ins

Weight Clipping for Robust Conformal Inference under Unbounded Covariate Shifts

ApplicationsDGX agent

arXiv:2605.02072v1 Announce Type: new Abstract: Conformal prediction (CP) provides powerful, distribution-free prediction sets, but its guarantees rely on the exchangeability of training and test data

What price to pay? Auto-tuning a building MPC controller for optimal economic cost

ApplicationsDGX agent

arXiv:2501.10859v2 Announce Type: replace-cross Abstract: Demand-side management (DSM) programs introduce complex pricing, requiring advanced control for cost minimization. Model Predictive Control (M

When Attention Collapses: Residual Evidence Modeling for Compositional Inference

SafetyDGX agent

arXiv:2605.02323v1 Announce Type: new Abstract: Compositional inference - the decomposition of observations into an unknown number of latent components - is central to perception and scientific data a

When Embedding-Based Defenses Fail: Rethinking Safety in LLM-Based Multi-Agent Systems

SafetyDGX agent

arXiv:2605.01133v1 Announce Type: cross Abstract: Large language model (LLM)-powered multi-agent systems (MAS) enable agents to communicate and share information, achieving strong performance on compl

When RL Meets Adaptive Speculative Training: A Unified Training-Serving System

Model ReleasesDGX agent

arXiv:2602.06932v3 Announce Type: replace Abstract: Speculative decoding can significantly accelerate LLM serving, yet most deployments today disentangle speculator training from serving, treating spe

Zero-Shot Adaptation of Behavioral Foundation Models to Unseen Dynamics

SafetyDGX agent

arXiv:2505.13150v2 Announce Type: replace Abstract: Behavioral Foundation Models (BFMs) proved successful in producing policies for arbitrary tasks in a zero-shot manner, requiring no test-time traini

Zero-Shot, Safe and Time-Efficient UAV Navigation via Potential-Based Reward Shaping, Control Lyapunov and Barrier Functions

SafetyDGX agent

arXiv:2605.01787v1 Announce Type: cross Abstract: Autonomous navigation and obstacle avoidance remain a core challenge of modern Unmanned Aerial Vehicles (UAVs). While traditional control methods stru

ZNO: Stable Rational Neural Operators in the Z-Domain for Discrete-Time Dynamic

Model ReleasesDGX agent

arXiv:2605.02356v1 Announce Type: new Abstract: We introduce the Z-Domain Neural Operator (ZNO), a causal neural operator whose layers are stable low-rank multiple-input multiple-output (MIMO) rationa

4 May 2026

A Comparative Analysis of Machine Learning Models for Intrusion Detection in Intelligent Transport Systems

Local AiDGX agent

arXiv:2605.00279v1 Announce Type: cross Abstract: AI-powered edge computing security is moving Intelligent Transportation Systems (ITS) from passive, rule-based protections to proactive, smart, zero-t

A Comparative Study of QSPR Methods on a Unique Multitask PAMPA dataset

ResearchDGX agent

arXiv:2605.00508v1 Announce Type: new Abstract: We present a unique, multitask dataset comprising 143 drug and drug candidate molecules, each evaluated on in vitro, parallel artificial-membrane permea

A Comparative Study of UMAP and Other Dimensionality Reduction Methods

ResearchDGX agent

arXiv:2603.02275v2 Announce Type: replace Abstract: Uniform Manifold Approximation and Projection (UMAP) is a widely used manifold learning technique for dimensionality reduction. This paper studies U

A Dirac-Frenkel-Onsager principle: Instantaneous residual minimization with gauge momentum for nonlinear parametrizations of PDE solutions

Model ReleasesDGX agent

arXiv:2605.00284v1 Announce Type: new Abstract: Dirac-Frenkel instantaneous residual minimization evolves nonlinear parametrizations of PDE solutions in time, but ill-conditioning can render the param

A Policy-Driven DRL Framework for System-Level Tradeoff Control in NR-U/Wi-Fi Coexistence

SafetyDGX agent

arXiv:2605.00457v1 Announce Type: cross Abstract: The coexistence of NR-U and Wi-Fi in unlicensed spectrum introduces a system-level resource coordination problem, where heterogeneous channel access m

A unified perspective on fine-tuning and sampling with diffusion and flow models

SafetyDGX agent

arXiv:2605.00229v1 Announce Type: cross Abstract: We study the problem of training diffusion and flow generative models to sample from target distributions defined by an exponential tilting of a base

AdaMeZO: Adam-style Zeroth-Order Optimizer for LLM Fine-tuning Without Maintaining the Moments

HardwareDGX agent

arXiv:2605.00650v1 Announce Type: new Abstract: Fine-tuning LLMs is necessary for various dedicated downstream tasks, but classic backpropagation-based fine-tuning methods require substantial GPU memo

Adaptive Node Feature Selection For Graph Neural Networks

ResearchDGX agent

arXiv:2510.03096v2 Announce Type: replace Abstract: We propose an adaptive node feature selection approach for graph neural networks (GNNs) that identifies and removes unnecessary features during trai

Adaptive Norm-Based Regularization for Neural Networks

ResearchDGX agent

arXiv:2605.00171v1 Announce Type: cross Abstract: In this paper, we study norm-based regularization methods for neural networks. We compare existing penalization approaches and introduce two regulariz

← Previous
1…192193194195196…241
Next →