AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,570Total entries
1Added by human
84,569Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,557 results
28 May 2026

Ariel-ML: Computing Parallelization with Embedded Rust for Neural Networks on Heterogeneous Multi-core Microcontrollers

Model ReleasesDGX agent

arXiv:2512.09800v2 Announce Type: replace Abstract: Low-power microcontroller (MCU) hardware is currently evolving from single-core architectures to predominantly multi-core architectures. In parallel

AtomComposer: Discovering Chemical Space from First Principles with Reinforcement Learning

AgentsDGX agent

arXiv:2605.28287v1 Announce Type: new Abstract: Discovering novel stable molecules without training data remains a grand scientific challenge. Current molecular generative models are trained on large,

Augmenting Attention with Exponentially Decaying Memory Improves Query-Aware KV Sparsity

Model ReleasesDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2605.28640v1 Announce Type: new Abstract: Efficient inference is critical for long-context language models, where attention computation and KV-cache access dominate the cost. Recent work RAT+, i

Automating Formal Verification with Agent-Guided Tree Search

Model ReleasesDGX agent

arXiv:2605.27485v1 Announce Type: cross Abstract: Formal verification offers a path to provably correct software, but writing verified code remains expensive enough that the technique is rarely used i

Bayesian Deployment Approval for Learned Landing Controllers under Finite Rollout Validation

SafetyDGX agent

arXiv:2605.27720v1 Announce Type: new Abstract: Reinforcement learning and data-driven autonomous controllers are commonly evaluated using cumulative reward and empirical success frequency under finit

Benchmarking Inductive Biases for Multivariate Time-Series Anomaly Detection with a Robust Multi-View Channel-Graph Detector

Model ReleasesDGX agent

arXiv:2605.28103v1 Announce Type: new Abstract: We present a unified experiment, analysis, and benchmark study of multivariate time-series (MTS) anomaly detection. Ten family-representative detectors

Beyond Lipschitz: Data-Driven Robustness via Discrete Modulus of Continuity

Local AiDGX agent

arXiv:2605.28729v1 Announce Type: cross Abstract: Robustness of neural networks is commonly quantified via local or global Lipschitz constants. However, Lipschitz continuity can be overly coarse or ov

Bilinear Coordinate Alignment for Training-Free Task-Vector Transfer

Model ReleasesDGX agent

arXiv:2605.28444v1 Announce Type: new Abstract: Fine-tuning large-scale pre-trained models is a recent prevalent paradigm for adapting general representations to specialized tasks. However, when a new

Bio-Inspired Self-Supervised Learning for Wrist-worn Accelerometer Data

ResearchDGX agent

arXiv:2603.10961v2 Announce Type: replace Abstract: Wearable accelerometers enable large-scale health monitoring, yet learning robust human-activity representations has been constrained by scarce labe

BPPO: Binary Prefix Policy Optimization for Efficient GRPO-Style Reasoning RL with Concise Responses

SafetyDGX agent

arXiv:2605.28028v1 Announce Type: new Abstract: Group Relative Policy Optimization (GRPO) is widely used for training reasoning models, but updating all sampled completions in each group incurs substa

Bridging Maximum Likelihood and Optimal Transport for Efficient Inference and Model Selection in Stochastic Block Models

ResearchDGX agent

arXiv:2605.28488v1 Announce Type: cross Abstract: We study inference in stochastic block models (SBMs) through the lens of optimal transport (OT). We first establish that maximum likelihood variationa

Bullet Trains: Parallelizing Training of Temporally Precise Spiking Neural Networks

ResearchDGX agent

arXiv:2603.13283v2 Announce Type: replace-cross Abstract: Continuous-time, event-native spiking neural networks (SNNs) operate strictly on spike events, treating spike timing and ordering as the repre

Calibrated Inference for the Conditional Average Treatment Effect in the Few-Placebo Regime via Gaussian Processes

SafetyDGX agent

arXiv:2605.27473v1 Announce Type: cross Abstract: Estimating how much an intervention helps a given individual the conditional average treatment effect (CATE) is increasingly central to decision-makin

Can Decision Trees Teach Large Language Models? Distilling Verbalized Knowledge for Molecular Property Prediction

Model ReleasesDGX agent

arXiv:2603.12344v2 Announce Type: replace Abstract: Molecular Property Prediction (MPP) is a fundamental problem in drug discovery that has recently attracted growing attention. Large Language Models

Can Entry-Wise Clipping Give Spectral Control of Stochastic Gradients?

ResearchDGX agent

arXiv:2605.27733v1 Announce Type: new Abstract: Training instabilities such as loss spikes are frequently the result of stochastic gradient noise. Because of rare expressions in language training data

CANDOR: Counterfactual ANnotated DOubly Robust Off-Policy Evaluation

SafetyDGX agent

arXiv:2412.08052v2 Announce Type: replace Abstract: Off-policy evaluation (OPE) is critical for applying contextual bandit algorithms to high-stakes decision-making settings such as healthcare, where

Causal Machine Learning: A Survey and Open Problems

SafetyDGX agent

arXiv:2206.15475v3 Announce Type: replace Abstract: Causal Machine Learning (CausalML) is an umbrella term for machine learning methods that formalize the data-generation process as a structural causa

CFDTwin: An open-source GUI and Python toolkit for POD-NN surrogate modeling of ANSYS Fluent simulations

Model ReleasesDGX agent

arXiv:2605.27725v1 Announce Type: cross Abstract: High-fidelity computational fluid dynamics (CFD) is widely used for thermal-fluid design, but repeated CFD solves remain expensive for design optimiza

Chreode: A Cell World Model for One-Step Temporal Dynamics and Perturbation Prediction

ResearchDGX agent

arXiv:2605.28111v1 Announce Type: new Abstract: Predicting how a cell will change its transcriptional state under a developmental signal or a genetic perturbation is the computational core of in-silic

Commit to the Bit: Reactive Reinforcement Learning Done Right

SafetyDGX agent

arXiv:2605.28276v1 Announce Type: new Abstract: Reinforcement learning algorithms are commonly analyzed (and designed) under the Markov assumption. This is unrealistic, as most environments encountere

Compositional Generalization in Autoregressive Models via Logit Composition

ResearchDGX agent

arXiv:2605.28304v1 Announce Type: new Abstract: Composing autoregressive models remains a core challenge in understanding how large language models can combine behaviors or skills learned across tasks

Conditionally Site-Independent Neural Evolution of Antibody Sequences

ResearchDGX agent

arXiv:2602.18982v4 Announce Type: replace Abstract: Common deep learning approaches for antibody engineering focus on modeling the marginal distribution of sequences. By treating sequences as independ

Conformal Prediction for Hierarchical Data

ResearchDGX agent

arXiv:2411.13479v4 Announce Type: replace-cross Abstract: We consider conformal prediction for multivariate data and focus on hierarchical data, where some components are linear combinations of others

Conservative neural posterior estimation via distributionally robust training

Model ReleasesDGX agent

arXiv:2605.28516v1 Announce Type: cross Abstract: Simulation-based inference with neural posterior estimation (NPE) often yields overconfident and unreliable posteriors under limited simulation budget

Context Features Are Cheap: Rank-Aware Decomposition for Efficient Feature Interaction in Recommender Systems

ApplicationsDGX agent

arXiv:2605.27450v1 Announce Type: cross Abstract: Modern industrial recommender systems use a deep ranking model to score N candidates against the same user and context features. Standard implementati

Continual Learning in Modern Hopfield Networks with an Application to Diffusion Models

ResearchDGX agent

arXiv:2605.27975v1 Announce Type: new Abstract: Generative models, including diffusion models, are increasingly used as foundation models and adapted through sequential fine-tuning, making continual l

Continuous Diffusion Models Can Obey Formal Syntax

TutorialsDGX agent

arXiv:2602.12468v2 Announce Type: replace Abstract: Diffusion language models offer a promising alternative to autoregressive models due to their global, non-causal generation process, but their conti

Conveyance: A Versatile Framework for Learning in Structured Class Spaces

ApplicationsDGX agent

arXiv:2605.28420v1 Announce Type: new Abstract: While machine learning (ML) architectures have evolved rapidly to account for complex data, loss functions like cross-entropy remain mostly structure-ag

Cost-Sensitive Evaluation for Binary Classifiers

Model ReleasesDGX agent

arXiv:2510.22016v2 Announce Type: replace Abstract: Selecting an appropriate evaluation metric for classifiers is crucial for model comparison, parameter optimization, and deployment decisions, yet th

Counterfactually Fair Regression via Optimal Transport

SafetyDGX agent

arXiv:2605.28251v1 Announce Type: cross Abstract: We consider the problem of learning a counterfactually fair regressor. We adopt a causal uncertainty view in which counterfactual fairness is defined

Cyclical Entropy Eruption: Entropy Dynamics in Agent Reinforcement Learning

AgentsDGX agent

arXiv:2605.27954v1 Announce Type: new Abstract: Agentic large language models are increasingly used to solve real-world tasks by reasoning over goals, invoking tools, and interacting with external env

DAISI: Data Assimilation with Inverse Sampling using Stochastic Interpolants

ResearchDGX agent

arXiv:2512.00252v4 Announce Type: replace-cross Abstract: Data assimilation (DA) is a cornerstone of scientific and engineering applications, combining model forecasts with sparse and noisy observatio

Dark Quest II: A Wide-Coverage Neural Network Emulator of the Nonlinear Matter Power Spectrum Across Extended Cosmologies

Model ReleasesDGX agent

arXiv:2605.28596v1 Announce Type: cross Abstract: extsc{DarkEmulator2} is a neural network emulator of the nonlinear matter power spectrum in a nine-dimensional w_0 w_a nu o CDM parameter space, devel

Decentralized Parameter-Free Online Learning with Compressed Gossip

Model ReleasesDGX agent

arXiv:2605.27831v1 Announce Type: new Abstract: We study decentralized online convex optimization when agents communicate over a graph and messages may be compressed. Classical decentralized online me

Decision-focused learning for optimal PV-Battery scheduling

ResearchDGX agent

arXiv:2605.28340v1 Announce Type: cross Abstract: The use of residential photovoltaics has increased dramatically in recent years. With battery systems becoming more affordable, the optimal operation

Decoupling Variance and Scale-Invariant Updates in Adaptive Gradient Descent for Unified Vector and Matrix Optimization

ResearchDGX agent

arXiv:2602.06880v2 Announce Type: replace Abstract: Adaptive methods like Adam have become the extit{de facto} standard for large-scale vector and Euclidean optimization due to their coordinate-wise a

Deep Neural Network Training as Random Effects: An Optimization-Inference Duality

ResearchDGX agent

arXiv:2605.27991v1 Announce Type: cross Abstract: Deep neural networks (DNNs) have achieved remarkable empirical success, yet their training dynamics remain understood mainly from optimization rather

DeepC4: Deep Conditional Census-Constrained Clustering for Large-scale Multitask Spatial Disaggregation of Urban Morphology

Local AiDGX agent

arXiv:2507.22554v3 Announce Type: replace Abstract: To understand our global progress for sustainable development and disaster risk reduction in many developing economies, two recent major initiatives

Density-aware Sample-specific Attack

ResearchDGX agent

arXiv:2605.27809v1 Announce Type: new Abstract: Despite recent progress in backdoor attacks, existing methods remain susceptible to post-training defenses that erase the backdoor through fine-tuning o

Detecting Diffusion-Generated Time Series Under Generator Shift

ResearchDGX agent

arXiv:2605.28355v1 Announce Type: new Abstract: The boundary between real and diffusion-generated time series is becoming increasingly difficult to draw, yet detection in this domain remains underexpl

Dimensionality Reduction for Robust Federated Learning: A Theoretical Analysis and Convergence Guarantee

Model ReleasesDGX agent

arXiv:2605.28335v1 Announce Type: new Abstract: Federated Learning (FL) enables multiple clients to collaboratively train models without sharing raw data, but it is highly vulnerable to Byzantine atta

Do Audio LLMs Listen or Read? Analyzing and Mitigating Paralinguistic Failures with VoxParadox

Model ReleasesDGX agent

arXiv:2605.27772v1 Announce Type: cross Abstract: Audio large language models (Audio LLMs) demonstrate strong performance on speech understanding tasks, yet their ability to understand paralinguistic

Dynamic Topic Modeling with a Higher-Order Hypergraphical Representation

ResearchDGX agent

arXiv:2605.28269v1 Announce Type: new Abstract: Dynamic topic modeling is widely used to analyze evolving trends in scientific literature, medical records, and social media. Traditional topic models r

E^3-Agent: An Executable and Evolving Agent for Resource Management of Edge Generative Inference

AgentsDGX agent

arXiv:2605.27428v1 Announce Type: new Abstract: Edge deployments of generative inference increasingly face two practical realities: per-device per-model performance is often unknown at deployment time

Evaluating Local Explainability Metrics for Machine Learning Models on Tabular Data

Model ReleasesDGX agent

arXiv:2605.27618v1 Announce Type: new Abstract: Despite the wide use of explainability techniques to attempt to understand the behavior of Artificial Intelligence (AI), the generated explanations may

Evolving and Detecting Multi-Turn Deception using Geometric Signatures

SafetyDGX agent

arXiv:2605.27671v1 Announce Type: cross Abstract: Safety defenses for large language models (LLMs) are typically trained and evaluated on single-turn prompts, yet real attacks often unfold as indirect

Excited Pfaffians: Generalized Neural Wave Functions Across Structure and State

ResearchDGX agent

arXiv:2603.14515v2 Announce Type: replace Abstract: Neural-network wave functions in Variational Monte Carlo (VMC) have achieved great success in accurately representing both ground and excited states

Exploratory Experience Shapes the Geometry of Predictive Representations

Model ReleasesDGX agent

arXiv:2605.27929v1 Announce Type: cross Abstract: Active sensing links behavior and learning through an action-perception loop: actions determine the observations used to update internal predictive mo

Expressive Power of Floating-Point Neural Networks with Arbitrary Reduction Orders and Inexact Activation Implementations

ResearchDGX agent

arXiv:2605.28704v1 Announce Type: new Abstract: Most existing expressivity theories for neural networks assume exact real arithmetic, whereas practical neural networks are executed under finite-precis

Extensions of Robbins-Siegmund Theorem with Applications in Reinforcement Learning

ResearchDGX agent

arXiv:2509.26442v2 Announce Type: replace Abstract: The Robbins-Siegmund theorem establishes the convergence of stochastic processes that are almost supermartingales and is one of the most commonly us

Falsification-driven reinforcement learning for maritime motion planning

AgentsDGX agent

arXiv:2510.06970v2 Announce Type: replace-cross Abstract: Compliance with maritime traffic rules is essential for the safe operation of autonomous vessels, yet training reinforcement learning (RL) age

Fast KV Compaction via Attention Matching

ResearchDGX agent

arXiv:2602.16284v2 Announce Type: replace Abstract: Scaling language models to long contexts is often bottlenecked by the size of the key-value (KV) cache. In deployed settings, long contexts are typi

Faster Thermal Profiling of a Lunar Rover with Machine Learning Adapted Finite Difference Model

AgentsDGX agent

arXiv:2605.27651v1 Announce Type: new Abstract: Autonomous space systems operating in extreme thermal environments require accurate and efficient thermal modeling to support both pre-mission system de

FedEHR-Gen: Federated Synthetic Time-Series EHR Generation via Latent Space Alignment and Distribution-Aware Aggregation

SafetyDGX agent

arXiv:2605.27892v1 Announce Type: new Abstract: Synthetic Electronic Health Record (EHR) generation provides a promising avenue for data augmentation and cross-hospital modeling in privacy-constrained

Federated Learning for Multivariate Time Series Anomaly Detection in Industrial Automation

Model ReleasesDGX agent

arXiv:2605.27486v1 Announce Type: new Abstract: Federated learning (FL) has broadened the horizon for multivariate time series anomaly detection (MTSAD). However, benchmarking such anomaly detection m

Fine-Tuning Dynamics of In-Context Factual Recall in Transformers

ResearchDGX agent

arXiv:2605.27774v1 Announce Type: new Abstract: In-context learning -- performing tasks based on examples given in the prompt -- is an important capability that has emerged in large language models an

Fitting Unknown Number of Hyperplanes with Manifold Optimization

ResearchDGX agent

arXiv:2605.28501v1 Announce Type: new Abstract: Fitting an unknown number of hyperplanes to data is a fundamental yet challenging problem in machine learning, characterized by its non-convexity, non-d

Flatness-Aware Stochastic Gradient Langevin Dynamics

ResearchDGX agent

arXiv:2510.02174v3 Announce Type: replace Abstract: Flatness of the loss landscape has been widely studied as an important perspective for understanding the behavior and generalization of deep learnin

Frequency-Guided Action Diffusion via Sub-Frequency Manifold Traversal

ResearchDGX agent

arXiv:2605.27919v1 Announce Type: cross Abstract: Learning visuomotor policies via behavior cloning typically involves mimicking expert demonstrations collected by human operators. However, natural hu

Genetic algorithm vs. gradient descent for training a neural network architecture dedicated to low data regimes in small medical datasets

ResearchDGX agent

arXiv:2605.27411v1 Announce Type: cross Abstract: Aim/Introduction: Distance-encoding biomorphic-informational neural network (DEBI-NN) is a recently proposed architecture in which connection weights

← Previous
1…118119120121122…243
Next →