AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,557 results
12 May 2026

Beyond Red-Teaming: Formal Guarantees of LLM Guardrail Classifiers

Model ReleasesDGX agent

arXiv:2605.10901v1 Announce Type: new Abstract: Guardrail Classifiers defend production language models against harmful behavior, but although results seem promising in testing, they provide no formal

Beyond Static Bias: Adaptive Multi-Fidelity Bandits with Improving Proxies

SafetyDGX agent

arXiv:2605.08558v1 Announce Type: new Abstract: As an extension of the classical multi-armed bandit problem, multi-fidelity multi-armed bandits (MF-MAB) enable individual arms to be evaluated using di

Beyond the All-in-One Agent: Benchmarking Role-Specialized Multi-Agent Collaboration in Enterprise Workflows

Model ReleasesDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2605.08761v1 Announce Type: cross Abstract: Large language model (LLM) agents are increasingly expected to operate in enterprise environments, where work is distributed across specialized roles,

Beyond the Black Box: An Interpretable Machine Learning Framework for Predicting Electronic Structure Microdescriptors and Structure-Performance Relationships in Fe-based Catalytic Systems

AgentsDGX agent

arXiv:2605.08994v1 Announce Type: cross Abstract: The current catalyst discovery and development pipeline for energy-intensive applications like methane conversion remains bottlenecked by expensive tr

Bi-CoG: Bi-Consistency-Guided Self-Training for Vision-Language Models

SafetyDGX agent

arXiv:2510.20477v2 Announce Type: replace Abstract: Exploiting unlabeled data through semi-supervised learning (SSL) or leveraging pre-trained models via fine-tuning are two prevailing paradigms for a

Bilinear autoencoders find interpretable manifolds

Model ReleasesDGX agent

arXiv:2605.08891v1 Announce Type: new Abstract: Sparse autoencoders have become a standard tool for uncovering interpretable latent representations in neural networks. Yet salient concepts often span

Black-Box Detection of LLM-Generated Text Using Generalized Jensen-Shannon Divergence

ResearchDGX agent

arXiv:2510.07500v2 Announce Type: replace Abstract: We study black-box detection of machine-generated text under practical constraints: the scoring model (proxy LM) may mismatch the unknown source mod

BoostLLM: Boosting-inspired LLM Fine-tuning for Few-shot Tabular Classification

Model ReleasesDGX agent

arXiv:2605.06117v2 Announce Type: replace Abstract: Large language models (LLMs) have recently been adapted to tabular prediction by serializing structured features into natural language, but their pe

Breaking the Reward Barrier: Accelerating Tree-of-Thought Reasoning via Speculative Exploration

ResearchDGX agent

arXiv:2605.10195v1 Announce Type: new Abstract: Tree-of-Thought (ToT) reasoning structures Large Language Model (LLM) inference as a tree-based search, demonstrating strong potential for solving compl

BRIDGE: Building Representations In Domain Guided Program Synthesis

SafetyDGX agent

arXiv:2511.21104v3 Announce Type: replace Abstract: Large language models can generate plausible code, but remain brittle for formal verification in proof assistants such as Lean. A central scalabilit

Bridging Spectral Operator Learning and U-Net Hierarchies: SpectraNet for Stable Autoregressive PDE Surrogates

Model ReleasesDGX agent

arXiv:2605.09096v1 Announce Type: new Abstract: Neural operators for time-dependent PDEs face a structural tension: spectral architectures (FNO and descendants) inherit exponential rollout-error growt

BROS: Bias-Corrected Randomized Subspaces for Memory-Efficient Single-Loop Bilevel Optimization

SafetyDGX agent

arXiv:2605.10288v1 Announce Type: new Abstract: Stochastic bilevel optimization (SBO) has become a standard framework for hyperparameter learning, data reweighting, representation learning, and data-m

CALYREX: Cross-Attention LaYeR EXtended Transformers for System Prompt Anchoring

Model ReleasesDGX agent

arXiv:2605.09737v1 Announce Type: new Abstract: Modern large language models (LLMs) rely on system prompts to establish behavioral constraints and safety rules. Standard causal self-attention treats p

Can Muon Fine-tune Adam-Pretrained Models?

ResearchDGX agent

arXiv:2605.10468v1 Announce Type: new Abstract: Muon has emerged as an efficient alternative to Adam for pretraining, yet remains underused for fine-tuning. A key obstacle is that most open models are

Can Revealed Preferences Clarify LLM Alignment and Steering?

SafetyDGX agent

arXiv:2605.08556v1 Announce Type: new Abstract: LLMs are increasingly used to make or support high-stakes decisions under uncertainty, where alignment depends not only on factual accuracy but on how m

Causal Discovery Should Embrace the Wisdom of the Crowd

ApplicationsDGX agent

arXiv:2603.02678v3 Announce Type: replace Abstract: This paper argues for recognizing an emerging paradigm of causal learning by wisdom of the crowd. Recent developments in government, industry, and r

Causal Explanations from the Geometric Properties of ReLU Neural Networks

SafetyDGX agent

arXiv:2605.10396v1 Announce Type: new Abstract: Neural networks have proved an effective means of learning control policies for autonomous systems, but these learned policies are difficult to understa

Central Limit Theorem for Two-Time-Scale Approximate Distributionally Robust RL

TutorialsDGX agent

arXiv:2605.08417v1 Announce Type: new Abstract: Designing model-free algorithms for distributionally robust reinforcement learning (DRRL) poses fundamental challenges. The robust Bellman operator is n

Characterizing the Generalization Error of Random Feature Regression with Arbitrary Data-Augmentation

ResearchDGX agent

arXiv:2605.10290v1 Announce Type: cross Abstract: This paper aims at analyzing the regularization effect that data augmentation induces on supervised regression methods in the proportional regime, whe

Cheap Thrills: Effective Amortized Optimization Using Inexpensive Labels

ResearchDGX agent

arXiv:2603.05495v2 Announce Type: replace Abstract: To scale optimization and simulation, prior work has explored training machine-learning surrogates that map problem parameters to solutions inexpens

Chebyshev Center-Based Direction Selection for Multi-Objective Optimization and Training PINNs

ResearchDGX agent

arXiv:2605.09975v1 Announce Type: new Abstract: Physics-informed neural networks (PINNs) are a promising approach for solving partial differential equations (PDEs). Their training, however, is often d

Classification-Head Bias in Class-Level Machine Unlearning: Diagnosis, Mitigation, and Evaluation

Model ReleasesDGX agent

arXiv:2605.08730v1 Announce Type: new Abstract: Class-level machine unlearning aims to remove the influence of specified classes while preserving model utility on retained classes. Existing methods ar

CoDistill-GRPO: A Co-Distillation Recipe for Efficient Group Relative Policy Optimization

Model ReleasesDGX agent

arXiv:2605.08873v1 Announce Type: new Abstract: Group Relative Policy Optimization (GRPO) has emerged as a powerful algorithm for improving the reasoning capabilities of language models, but often fai

Compact SO(3) Equivariant Atomistic Foundation Models via Structural Pruning

ResearchDGX agent

arXiv:2605.08885v1 Announce Type: new Abstract: SO(3) equivariant graph neural networks have become the dominant paradigm for atomistic foundation models, achieving high accuracy and data efficiency b

Compander-Aligned Query Geometry for Quantized Zeroth-Order Optimization

Model ReleasesDGX agent

arXiv:2605.10673v1 Announce Type: new Abstract: Low-bit forward evaluation is an attractive route to memory-efficient zeroth-order (ZO) adaptation: the optimizer needs only scalar losses, and the mode

Comparing Two Proxy Methods for Causal Identification

ResearchDGX agent

arXiv:2512.00175v3 Announce Type: replace-cross Abstract: Identifying causal effects in the presence of unmeasured variables is a fundamental challenge in causal inference, for which proxy variable me

Complex-Valued Phase-Coherent Transformer

Model ReleasesDGX agent

arXiv:2605.10123v1 Announce Type: new Abstract: Complex-valued Transformers have largely inherited softmax attention from real-valued architectures. However, row-normalised token competition is not ne

Composing diffusion priors with explicit physical context via generative Gibbs sampling

ResearchDGX agent

arXiv:2605.10642v1 Announce Type: new Abstract: Pretrained diffusion models provide powerful learned priors, but in scientific sampling the target distribution often depends on physical context that i

Compute as Teacher: Turning Inference Compute Into Reference-Free Supervision

SafetyDGX agent

arXiv:2509.14234v3 Announce Type: replace Abstract: Where do learning signals come from when there is no ground truth in post-training? We show that inference compute itself can serve as supervision.

Concordia: Self-Improving Synthetic Tables for Federated LLMs

Model ReleasesDGX agent

arXiv:2605.09855v1 Announce Type: new Abstract: Federated learning (FL) enables training large language models (LLMs) without sharing raw data, but adapting LLMs under strict data isolation and non-II

Conditional anomaly detection methods for patient-management alert systems

ApplicationsDGX agent

arXiv:2605.10847v1 Announce Type: new Abstract: Anomaly detection methods can be very useful in identifying unusual or interesting patterns in data. A recently proposed conditional anomaly detection f

ConfoundingSHAP: Quantifying confounding strength in causal inference

ResearchDGX agent

arXiv:2605.10533v1 Announce Type: new Abstract: In causal inference, confounders are variables that influence both treatment decisions and outcomes. However, unlike as in randomized clinical trials, t

ConQuR: Corner Aligned Activation Quantization via Optimized Rotations for LLMs

Model ReleasesDGX agent

arXiv:2605.10793v1 Announce Type: new Abstract: Large language models (LLMs) are costly to deploy due to their large memory footprint and high inference cost. Weight-activation quantization can reduce

Consistent Projection of Langevin Dynamics: Preserving Thermodynamics and Kinetics in Coarse-Grained Models

ResearchDGX agent

arXiv:2512.03706v2 Announce Type: replace-cross Abstract: Coarse graining (CG) is an important task for efficient modeling and simulation of complex multi-scale systems, such as the conformational dyn

Consolidation-Expansion Operator Mechanics:A Unified Framework for Adaptive Learning

ResearchDGX agent

arXiv:2605.09968v1 Announce Type: new Abstract: Every adaptive learning system must alternate between two operations: consolidating what it already knows and expanding into new evidence. We propose Co

Constraint-Aware Reinforcement Learning via Adaptive Action Scaling

SafetyDGX agent

arXiv:2510.11491v3 Announce Type: replace-cross Abstract: Safe reinforcement learning (RL) seeks to mitigate unsafe behaviors that arise from exploration during training by reducing constraint violati

Constructive conditional normalizing flows

ResearchDGX agent

arXiv:2602.08606v2 Announce Type: replace-cross Abstract: Motivated by applications in conditional sampling, given a probability measure mu and a diffeomorphism phi, we consider the problem of simulta

CONTRA: Conformal Prediction Region via Normalizing Flow Transformation

ResearchDGX agent

arXiv:2605.08561v1 Announce Type: cross Abstract: Density estimation and reliable prediction regions for outputs are crucial in supervised and unsupervised learning. While conformal prediction effecti

Controllability in preference-conditioned multi-objective reinforcement learning

AgentsDGX agent

arXiv:2605.10585v1 Announce Type: new Abstract: Multi-objective reinforcement learning (MORL) allows a user to express preference over outcomes in terms of the relative importance of the objectives, b

Controlling Transient Amplification Improves Long-horizon Rollouts

ResearchDGX agent

arXiv:2605.08856v1 Announce Type: new Abstract: Autoregressive neural simulators now match classical solvers on short-horizon prediction of physical systems, yet their accuracy degrades rapidly when r

Convergence Analysis of Newton's Method for Neural Networks in the Overparameterized Limit

Model ReleasesDGX agent

arXiv:2605.08352v1 Announce Type: new Abstract: A convergence analysis is developed for the regularized Newton method for training neural networks (NNs) in the overparameterized limit. As the number o

Cosine-Gated Adam-Decay: Drop-In Staleness-Aware Outer Optimization for Decoupled DiLoCo

Model ReleasesDGX agent

arXiv:2605.09126v1 Announce Type: new Abstract: Asynchronous DiLoCo systems may receive pseudo-gradients computed several outer rounds earlier, yet the standard Nesterov outer optimizer does not expli

Cross-Domain Lossy Compression via Constrained Minimum Entropy Coupling

ResearchDGX agent

arXiv:2605.09833v1 Announce Type: cross Abstract: This paper studies cross-domain lossy compression through the lens of minimum entropy coupling (MEC) with rate and classification constraints. In this

Crowding Out The Noise: Algorithmic Collective Action Under Differential Privacy

SafetyDGX agent

arXiv:2505.05707v2 Announce Type: replace Abstract: The integration of AI into daily life has generated considerable attention and excitement, while also raising concerns about automating algorithmic

CrystalREPA: Transferring Physical Priors from Universal MLIPs to Crystal Generative Models

Model ReleasesDGX agent

arXiv:2605.08960v1 Announce Type: cross Abstract: Crystal generative models mainly learn what stable crystals look like, with little explicit supervision for what makes them stable. We reveal a substa

CUDABeaver: Benchmarking LLM-Based Automated CUDA Debugging

Model ReleasesDGX agent

arXiv:2605.08455v1 Announce Type: new Abstract: Debugging CUDA programs has long been challenging because failures often arise from subtle interactions among hardware behavior, compiler decisions, mem

CUDAHercules: Benchmarking Hardware-Aware Expert-level CUDA Optimization for LLMs

Model ReleasesDGX agent

arXiv:2605.08467v1 Announce Type: new Abstract: Large language models show promise for automated CUDA programming, however even the strongest coding models (e.g., Claude-Opus-4.6) may still fall short

D2ACE: Multi-Label Batch Selection Guided by Dual Dynamics and Adaptive Correlation Enhancement

ResearchDGX agent

arXiv:2605.09400v1 Announce Type: new Abstract: Batch selection is crucial for improving both training efficiency and predictive performance in deep multi-label classification (MLC). Existing batch se

DANCE: Detect and Classify Events in EEG

ApplicationsDGX agent

arXiv:2605.10688v1 Announce Type: new Abstract: Event identification in continuous neural recordings is a critical task in neuroscience. Decoding in EEG is dominated by classifying windows aligned to

Data-driven transport modelling without overfit

SafetyDGX agent

arXiv:2605.08801v1 Announce Type: new Abstract: Macroscopic transport modelling aims to predict traffic flows after proposed public policy interventions, such as a new road or railway section or a tem

DataArc-SynData-Toolkit: A Unified Closed-Loop Framework for Multi-Path, Multimodal, and Multilingual Data Synthesis

ApplicationsDGX agent

arXiv:2605.08138v1 Announce Type: new Abstract: Synthetic data has emerged as a crucial solution to the data scarcity bottleneck in large language models (LLMs), particularly for specialized domains a

Debiased Front-Door Learners for Heterogeneous Effects

ApplicationsDGX agent

arXiv:2509.22531v2 Announce Type: replace-cross Abstract: In observational settings where treatment and outcome share unmeasured confounders but an observed mediator remains unconfounded, the front-do

Decentralized Conformal Novelty Detection via Quantized Model Exchange

ResearchDGX agent

arXiv:2605.08263v1 Announce Type: cross Abstract: This work studies decentralized novelty detection with global false discovery rate (FDR) control across heterogeneous composite null distributions, wi

Decoding Islamophobic Discourse: Using LLMs to Identify Tropes and Semi-Coded Hate Speech

ResearchDGX agent

arXiv:2503.18273v3 Announce Type: replace Abstract: In recent years, Islamophobia has gained significant traction across Western societies, fueled by the rise of digital communication networks. This p

Deep-ICE: the first globally optimal algorithm for minimizing 0-1 loss in two-layer ReLU and maxout networks

ResearchDGX agent

arXiv:2505.05740v3 Announce Type: replace Abstract: This paper introduces the first globally optimal algorithm for the empirical risk minimization problem of two-layer maxout and ReLU networks, i.e.,

Deep Learning under Fractional-Order Differential Privacy

Model ReleasesDGX agent

arXiv:2605.09890v1 Announce Type: cross Abstract: Differentially private stochastic gradient descent (DP-SGD) is a standard approach to privacy-preserving learning based on per-example clipping, subsa

Deep set based operator learning with uncertainty quantification

ResearchDGX agent

arXiv:2509.25646v2 Announce Type: replace Abstract: Learning operators from data is central to scientific machine learning. While DeepONets are widely used for their ability to handle complex domains,

DeepLevy: Learning Heavy-Tailed Uncertainty in Highly Volatile Time Series

ResearchDGX agent

arXiv:2605.10364v1 Announce Type: new Abstract: Modeling uncertainty in heavy-tailed time series remains a critical challenge for deep probabilistic forecasting models, which often struggle to capture

DeepLog: A Software Framework for Modular Neurosymbolic AI

ResearchDGX agent

arXiv:2605.10279v1 Announce Type: new Abstract: DeepLog is an operational neurosymbolic framework that unifies logic and deep learning within standard PyTorch workflows. While existing neurosymbolic s

Dendritic Neural Networks with Equilibrium Propagation

ResearchDGX agent

arXiv:2605.08135v1 Announce Type: new Abstract: Equilibrium propagation (EP) is a biologically plausible alternative to backpropagation (BP), but its effectiveness can degrade in deeper and more chall

← Previous
1…168169170171172…243
Next →