AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,557 results
14 May 2026

Beyond Oversquashing: Understanding Signal Propagation in GNNs Via Observables

ResearchDGX agent

arXiv:2605.13383v1 Announce Type: new Abstract: Graph Neural Networks (GNNs) perform computations on graphs by routing the signal between graph regions using a graph shift operator or a message passin

Beyond Softmax: A Natural Parameterization for Categorical Random Variables

ResearchDGX agent

arXiv:2509.24728v2 Announce Type: replace Abstract: Latent categorical variables are frequently found in deep learning architectures. They can model actions in discrete reinforcement-learning environm

BioSEN: A Bio-acoustic Signal Enhancement Network for Animal Vocalizations

ResearchDGX agent

arXiv:2605.12534v1 Announce Type: cross Abstract: Most work in audio enhancement targets human speech, while bioacoustics is less studied due to noisy recordings and the distinct traits of animal soun


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Building Interactive Real-Time Agents with Asynchronous I/O and Speculative Tool Calling

Model ReleasesDGX agent

arXiv:2605.13360v1 Announce Type: new Abstract: There is a growing demand for agentic AI technologies for a range of downstream applications like customer service and personal assistants. For applicat

Byzantine-Robust Distributed Sparse Learning Revisited

ResearchDGX agent

arXiv:2605.13283v1 Announce Type: new Abstract: We revisit Byzantine robust distributed estimation for high-dimensional sparse linear models. By combining local ell_1-regularized robust estimation wit

Causal Fine-Tuning under Latent Confounded Shift

SafetyDGX agent

arXiv:2410.14375v3 Announce Type: replace Abstract: Adapting to latent confounded shift remains a core challenge in modern AI. This setting is driven by hidden variables that induce spurious correlati

Causal Learning with the Invariance Principle

ResearchDGX agent

arXiv:2605.13589v1 Announce Type: cross Abstract: Causal discovery, the problem of inferring the direction of causality, is generally ill-posed. We use the language of structural causal models (SCM) t

CAWI: Copula-Aligned Weight Initialization for Randomized Neural Networks

ResearchDGX agent

arXiv:2605.12580v1 Announce Type: new Abstract: Randomized neural networks (RdNNs) enable efficient, backpropagation-free training by freezing randomly initialized input-to-hidden weights, which permi

Centralized Adaptive Sampling for Reliable Co-Training of Independent Multi-Agent Policies

SafetyDGX agent

arXiv:2508.01049v2 Announce Type: replace Abstract: Independent on-policy policy gradient algorithms are widely used for multi-agent reinforcement learning (MARL) in cooperative and no-conflict games,

Certified Robustness under Heterogeneous Perturbations via Hybrid Randomized Smoothing

SafetyDGX agent

arXiv:2605.12876v1 Announce Type: new Abstract: Randomized smoothing provides strong, model-agnostic robustness certificates, but existing guarantees are limited to single modalities, treating continu

Chem-GMNet: A Sphere-Native Geometric Transformer for Molecular Property Prediction

ResearchDGX agent

arXiv:2605.13262v1 Announce Type: new Abstract: Modern SMILES-based chemical language models obtain strong MoleculeNet performance by treating SMILES as generic text and compensating with multi-millio

Clustering in pure-attention hardmax transformers and its role in sentiment analysis

ResearchDGX agent

arXiv:2407.01602v2 Announce Type: replace-cross Abstract: Transformers are extremely successful machine learning models whose mathematical properties remain poorly understood. Here, we rigorously char

CO-MAP: A Reinforcement Learning Approach to the Qubit Allocation Problem

SafetyDGX agent

arXiv:2605.13638v1 Announce Type: cross Abstract: A quantum compiler is a critical piece in the quantum computing pipeline since it allows an abstract quantum circuit to be run on a physical quantum c

Code-Centric Detection of Vulnerability-Fixing Commits: A Unified Benchmark and Empirical Study

Model ReleasesDGX agent

arXiv:2605.13138v1 Announce Type: cross Abstract: Automated detection of vulnerability-fixing commits (VFCs) is critical for timely security patch deployment, as advisory databases lag patch releases

Collaborating in Multi-Armed Bandits with Strategic Agents

AgentsDGX agent

arXiv:2605.13145v1 Announce Type: new Abstract: We study collaborative learning in multi-agent Bayesian bandit problems, where strategic agents collectively solve the same bandit instance. While multi

Collaborative Parameter Learning: Mitigating Forgetting via Parameter-Level Gradient Analysis

Model ReleasesDGX agent

arXiv:2601.21577v2 Announce Type: replace Abstract: Catastrophic forgetting during knowledge injection impairs the ability of large language models to acquire new knowledge without overwriting previou

Conformal Anomaly Detection in Python: Moving Beyond Heuristic Thresholds with 'nonconform'

TutorialsDGX agent

arXiv:2605.13642v1 Announce Type: cross Abstract: Most anomaly detection systems output scores rather than calibrated decisions, leaving practitioners to choose thresholds heuristically and without cl

Connecting the Dots: A Machine Learning Ready Dataset for Ionospheric Forecasting Models

Model ReleasesDGX agent

arXiv:2511.15743v2 Announce Type: replace Abstract: Operational forecasting of the ionosphere remains a critical space weather challenge due to sparse observations, complex coupling across geospatial

ConRetroBert: EMA Stabilized Dual Encoders for Template-Based Single-Step Retrosynthesis

Model ReleasesDGX agent

arXiv:2605.12736v1 Announce Type: new Abstract: Template based single step retrosynthesis predicts reactants by selecting and applying an explicit reaction template, making each prediction traceable t

Constraint-Aware Flow Matching: Decision Aligned End-to-End Training for Constrained Sampling

ApplicationsDGX agent

arXiv:2605.12754v1 Announce Type: new Abstract: Deep generative models provide state-of-the-art performance across a wide array of applications, with recent studies showing increasing applicability fo

Constraints-of-Thought: A Framework for Constrained Reasoning in Language-Model-Guided Search

SafetyDGX agent

arXiv:2510.08992v3 Announce Type: replace Abstract: While researchers have made significant progress in enabling large language models (LLMs) to perform multi-step planning, LLMs struggle to ensure th

Context-Aware Web Attack Detection in Open-Source SIEM Systems via MITRE ATT&CK-Enriched Behavioral Profiling

ResearchDGX agent

arXiv:2605.13337v1 Announce Type: cross Abstract: Security Information and Event Management (SIEM) systems aggregate log data from heterogeneous sources to detect coordinated attacks. Traditional rule

Contextual Bandits for Resource-Constrained Devices using Probabilistic Learning

Local AiDGX agent

arXiv:2605.13346v1 Announce Type: new Abstract: Contextual bandits (CB) are online sequential decision-making problems under partial feedback that underpin many adaptive services. There is a growing d

Continual Fine-Tuning of Large Language Models via Program Memory

Model ReleasesDGX agent

arXiv:2605.13162v1 Announce Type: new Abstract: Parameter-Efficient Fine-Tuning (PEFT), particularly Low-Rank Adaptation (LoRA), has become a standard approach for adapting Large Language Models (LLMs

Convergent Differential Privacy Analysis for General Federated Learning

ResearchDGX agent

arXiv:2408.15621v3 Announce Type: replace Abstract: The powerful cooperation of federated learning (FL) and differential privacy~(DP) provides a promising paradigm for the large-scale private clients.

Coreset-Induced Conditional Velocity Flow Matching

ResearchDGX agent

arXiv:2605.12951v1 Announce Type: cross Abstract: We propose Coreset-Induced Conditional Velocity Flow Matching (CCVFM), a generative model that augments hierarchical rectified flow with a data-inform

Coupling-Informed Transport Maps for Bayesian Filtering in Nonlinear Dynamical Systems

ResearchDGX agent

arXiv:2605.13174v1 Announce Type: cross Abstract: A likelihood-free transport filtering method is proposed based on the couplings between state and observation variables. By exploiting a block-triangu

CR-Net: Scaling Parameter-Efficient Training with Cross-Layer Low-Rank Structure

Model ReleasesDGX agent

arXiv:2509.18993v3 Announce Type: replace Abstract: Low-rank architectures have become increasingly important for efficient large language model (LLM) pre-training, providing substantial reductions in

Data-Driven Integration Kernels for Interpretable Nonlocal Operator Learning

ResearchDGX agent

arXiv:2603.10305v3 Announce Type: replace Abstract: Machine learning models can represent climate processes that are nonlocal in horizontal space, height, and time, often by combining information acro

Decision Support for Marketplace Policies under Incomplete Evidence: From Replay to Launch Readiness

SafetyDGX agent

arXiv:2605.12840v1 Announce Type: cross Abstract: Marketplace platforms routinely evaluate pricing and allocation policies using logged observational data, yet strong offline performance does not impl

Decision Tree Learning on Product Spaces

Model ReleasesDGX agent

arXiv:2605.12983v1 Announce Type: new Abstract: Decision tree learning has long been a central topic in theoretical computer science, driven by its practical importance. A fundamental and widely used

Decoupling Exploration and Policy Optimization: Uncertainty Guided Tree Search for Hard Exploration

SafetyDGX agent

arXiv:2603.22273v4 Announce Type: replace Abstract: The process of discovery requires active exploration -- the act of collecting new and informative data. However, efficient autonomous exploration re

Deep Learning as Neural Low-Degree Filtering: A Spectral Theory of Hierarchical Feature Learning

TutorialsDGX agent

arXiv:2605.13612v1 Announce Type: new Abstract: Understanding how deep neural networks learn useful internal representations from data remains a central open problem in the theory of deep learning. We

Dense vs Sparse Pretraining at Tiny Scale: Active-Parameter vs Total-Parameter Matching

Model ReleasesDGX agent

arXiv:2605.13769v1 Announce Type: cross Abstract: We study dense and mixture-of-experts (MoE) transformers in a tiny-scale pretraining regime under a shared LLaMA-style decoder training recipe. The sp

Descriptive Collision in Sparse Autoencoder Auto-Interpretability: When One Explanation Describes Many Features

Model ReleasesDGX agent

arXiv:2605.12874v1 Announce Type: new Abstract: Sparse autoencoders (SAEs) are now standard tools for decomposing language model activations into interpretable features, and automated interpretability

Differentially Private Nonparametric Confidence Intervals Under Minimal Distributional Assumptions

ResearchDGX agent

arXiv:2511.01303v2 Announce Type: replace-cross Abstract: We consider the problem of constructing differentially private nonparametric confidence intervals (CIs) for an arbitrary quantity using resamp

Diffusion-Inspired Reconfiguration of Transformers for Uncertainty Calibration

ResearchDGX agent

arXiv:2602.08920v2 Announce Type: replace Abstract: Uncertainty calibration in pre-trained transformers is critical for their reliable deployment in risk-sensitive applications. Yet, most existing pre

Diffusion Model's Generalization Can Be Characterized by Inductive Biases toward a Data-Dependent Ridge Manifold

SafetyDGX agent

arXiv:2602.06021v2 Announce Type: replace-cross Abstract: We study a data-dependent notion of diffusion-model generalization: when a model does not memorize the training set, where do its generated sa

DiffusionHijack: Supply-Chain PRNG Backdoor Attack on Diffusion Models and Quantum Random Number Defense

SafetyDGX agent

arXiv:2605.13115v1 Announce Type: cross Abstract: Diffusion models depend on pseudo-random number generators (PRNGs) for latent noise sampling. We present DiffusionHijack, a supply-chain backdoor atta

Digital Twins as Synthetic Controls in Single-Arm Trials

SafetyDGX agent

arXiv:2605.12832v1 Announce Type: cross Abstract: Single-arm trials are an important study design for evaluating drug efficacy and safety without enrolling patients into a control arm. Although they d

DisAgg: Distributed Aggregators for Efficient Secure Aggregation in Federated Learning

ResearchDGX agent

arXiv:2605.13708v1 Announce Type: cross Abstract: Federated learning enables collaborative model training across distributed clients, yet vanilla FL exposes client updates to the central server. Secur

Discrete Stochastic Localization for Non-autoregressive Generation

ResearchDGX agent

arXiv:2605.12836v1 Announce Type: new Abstract: Continuous diffusion is a natural framework for non-autoregressive generation but has generally lagged behind masked discrete diffusion models (MDMs) on

Distinguishing performance gains from learning when using generative AI

ApplicationsDGX agent

arXiv:2605.13731v1 Announce Type: new Abstract: Generative artificial intelligence (AI) is increasingly being integrated into education, where it can boost learners' performance. However, these uses d

Distribution Shift in Missing Data Imputation: A Risk-Based Perspective and Importance-Weighted Correction under MAR

TutorialsDGX agent

arXiv:2602.06713v2 Announce Type: replace-cross Abstract: Missing data imputation, where a model is trained on observed data to estimate unobserved values, is a fundamental problem in machine learning

Do Activation Verbalization Methods Convey Privileged Information?

ResearchDGX agent

arXiv:2509.13316v4 Announce Type: replace-cross Abstract: Recent interpretability methods have proposed to translate LLM internal representations into natural language descriptions using a second verb

Do Heavy Tails Help Diffusion? On the Subtle Trade-off Between Initialization and Training

ApplicationsDGX agent

arXiv:2605.13175v1 Announce Type: new Abstract: Recent works have proposed incorporating heavy-tailed (HT) noise into diffusion- and flow-based generative models, with the goals of better recovering t

DP-KFC: Data-Free Preconditioning for Privacy-Preserving Deep Learning

ResearchDGX agent

arXiv:2605.13418v1 Announce Type: new Abstract: Differentially private optimization suffers from a fundamental geometric mismatch: deep networks have highly anisotropic loss landscapes, yet DP-SGD inj

DP-Muon: Differentially Private Optimization via Matrix-Orthogonalized Momentum

SafetyDGX agent

arXiv:2605.12994v1 Announce Type: new Abstract: We study differentially private (DP) training with Muon, a matrix-valued optimizer that updates hidden-layer weights using momentum followed by Newton--

DRIFT: A Benchmark for Task-Free Continual Graph Learning with Continuous Distribution Shifts

Model ReleasesDGX agent

arXiv:2605.12998v1 Announce Type: new Abstract: Continual graph learning (CGL) aims to learn from dynamically evolving graphs while mitigating catastrophic forgetting. Existing CGL approaches typicall

Early Data Exposure Improves Robustness to Subsequent Fine-Tuning

ResearchDGX agent

arXiv:2605.12705v1 Announce Type: new Abstract: How can we train models whose post-trained capabilities survive subsequent fine-tuning? Rather than focusing on downstream interventions to mitigate for

Earth Science Foundation Models: From Perception to Reasoning and Discovery

AgentsDGX agent

arXiv:2605.12542v1 Announce Type: cross Abstract: Large foundation models (FMs) are transforming Earth science by integrating heterogeneous multimodal data, such as multi-platform imagery, gridded rea

Effective Context in Transformers: An Analysis of Fragmentation and Tokenization

Model ReleasesDGX agent

arXiv:2605.13485v1 Announce Type: new Abstract: Transformers predict over a representation of a sequence. The same data can be written as bytes, characters, or subword tokens, and these representation

Efficient distributional regression trees learning algorithms for calibrated non-parametric probabilistic forecasts

TutorialsDGX agent

arXiv:2502.05157v3 Announce Type: replace Abstract: The perspective of developing trustworthy AI for critical applications in science and engineering requires machine learning techniques that are capa

Efficient Generative Prediction for EHR Foundation Models: The SCOPE and REACH Estimators

ApplicationsDGX agent

arXiv:2602.03730v2 Announce Type: replace-cross Abstract: Generative foundation models trained on tokenized electronic health record (EHR) timelines show promise for clinical outcome prediction via Mo

Efficient Sensor Fusion for Gesture Recognition on Resource-Constrained Devices

Local AiDGX agent

arXiv:2605.13462v1 Announce Type: new Abstract: Gesture recognition is a cornerstone of Human-Computer Interaction (HCI) for smart eyewear, enabling natural and device-free control in augmented realit

EGSS: Entropy-guided Stepwise Scaling for Reliable Software Engineering

AgentsDGX agent

arXiv:2602.05242v1 Announce Type: cross Abstract: Agentic Test-Time Scaling (TTS) has delivered state-of-the-art (SOTA) performance on complex software engineering tasks such as code generation and bu

Embodied Neurocomputation: A Framework for Interfacing Biological Neural Cultures with Scaled Task-Driven Validation

Model ReleasesDGX agent

arXiv:2605.13315v1 Announce Type: cross Abstract: Biological neural networks (BNNs) have been established as a powerful and adaptive substrate that offer the potential for incredibly energy and data e

EMO: Frustratingly Easy Progressive Training of Extendable MoE

HardwareDGX agent

arXiv:2605.13247v1 Announce Type: new Abstract: Sparse Mixture-of-Experts (MoE) models offer a powerful way to scale model size without increasing compute, as per-token FLOPs depend only on k active e

Ergodic Trajectory Design by Learned Pushforward Maps: Provable Coverage via Conditional Flow Matching

SafetyDGX agent

arXiv:2605.13063v1 Announce Type: new Abstract: Designing continuous trajectories whose time-averaged occupancy provably matches a prescribed spatial density (the ergodic coverage problem) is central

ERPPO: Entropy Regularization-based Proximal Policy Optimization

SafetyDGX agent

arXiv:2605.13131v1 Announce Type: new Abstract: Multi-Agent Proximal Policy Optimization (MAPPO) is a variant of the Proximal Policy Optimization (PPO) algorithm, specifically tailored for multi-agent

← Previous
1…157158159160161…243
Next →