AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,557 results
14 May 2026

When is Warmstarting Effective for Scaling Language Models?

Model ReleasesDGX agent

arXiv:2605.13405v1 Announce Type: new Abstract: Model growth from a given checkpoint aims to accelerate training of a larger model, offering potential resource savings. Despite recent interest, warmst

When to Act, Ask, or Learn: Uncertainty-Aware Policy Steering

SafetyDGX agent

arXiv:2602.22474v2 Announce Type: replace-cross Abstract: Policy steering is an emerging way to adapt robot behaviors at deployment-time: a learned verifier analyzes low-level action samples proposed

When to Transfer: Adaptive Source Selection for Positive Transfer in Linear Models

TutorialsDGX agent

arXiv:2510.16986v2 Announce Type: replace-cross Abstract: In many business settings, task-specific labeled data are scarce or costly to obtain, limiting supervised learning on a target task. A classic


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

When to Trust Confidence Thresholding: Calibration Diagnostics for Pseudo-Labelled Regression

SafetyDGX agent

arXiv:2605.12780v1 Announce Type: cross Abstract: Calibrated probability outputs of trained classifiers are increasingly used as inputs to downstream regression estimands such as effects, prevalences,

Why is 'Chicago' Predictive of Deceptive Reviews? Using LLMs to Discover Language Phenomena from Lexical Cues

TutorialsDGX agent

arXiv:2511.13658v2 Announce Type: replace-cross Abstract: Deceptive reviews mislead consumers, harm businesses, and undermine trust in online marketplaces. Machine learning classifiers can learn from

Yield Curves Dynamics Using Variational Autoencoders Under No-arbitrage

ResearchDGX agent

arXiv:2605.12764v1 Announce Type: cross Abstract: This paper introduces a physics-informed generative framework that resolves the fundamental conflict between the statistical flexibility of deep learn

ZKBoost: Zero-Knowledge Verifiable Training for XGBoost

ApplicationsDGX agent

arXiv:2602.04113v3 Announce Type: replace-cross Abstract: Gradient boosted decision trees, particularly XGBoost, are among the most effective methods for tabular data. As deployment in sensitive setti

13 May 2026

20/20 Vision Language Models: A Prescription for Better VLMs through Data Curation Alone

ResearchDGX agent

arXiv:2605.11405v1 Announce Type: new Abstract: Data curation has shifted the quality-compute frontier for language-model and contrastive image-text pretraining, but its role for vision-language model

A Boundary-Aware Non-parametric Granular-Ball Classifier Based on Minimum Description Length

Model ReleasesDGX agent

arXiv:2605.11406v1 Announce Type: new Abstract: Existing granular-ball classification methods are often driven by handcrafted quality measures, neighborhood rules, or heuristic splitting and stopping

A Comparative Study of Federated Learning Aggregation Strategies under Homogeneous and Heterogeneous Data Distributions

Model ReleasesDGX agent

arXiv:2605.11010v1 Announce Type: new Abstract: Federated Learning has emerged as a transformative paradigm for collaborative machine learning across distributed environments. However, its performance

A Comparative Study of Model Selection Criteria for Symbolic Regression

ResearchDGX agent

arXiv:2605.11233v1 Announce Type: new Abstract: Effective model selection is critical in symbolic regression (SR) to identify mathematical expressions that balance accuracy and complexity, and have lo

A Composite Activation Function for Learning Stable Binary Representations

ResearchDGX agent

arXiv:2605.11558v1 Announce Type: new Abstract: Activation functions play a central role in neural networks by shaping internal representations. Recently, learning binary activation representations ha

A Controlled Counterexample to Strong Proxy-Based Explanations of OOD Performance: in a Fixed Pretraining-and-Probing Setup

ResearchDGX agent

arXiv:2605.11554v1 Announce Type: new Abstract: Task-agnostic structure proxies are often used to interpret why one pretraining corpus transfers better than another, but such explanations require the

A Fast and Energy-Efficient Latch-Based Memristive Analog Content-Addressable Memory

Local AiDGX agent

arXiv:2605.11847v1 Announce Type: cross Abstract: Analog content-addressable memories (aCAMs) based on memristors provide a promising pathway toward energy-efficient large-scale associative computing

A New Technique for AI Explainability using Feature Association Map

Model ReleasesDGX agent

arXiv:2605.12350v1 Announce Type: new Abstract: Lack of transparency in AI systems poses challenges in critical real-life applications. It is important to be able to explain the decisions of an AI sys

A nonlinear extension of parametric model embedding for dimensionality reduction in parametric shape design

Model ReleasesDGX agent

arXiv:2605.11759v1 Announce Type: cross Abstract: Dimensionality reduction is essential in simulation-based shape design, where high-dimensional parameterizations hinder optimization, surrogate modeli

A Proof-of-Concept Simulation-Driven Digital Twin Framework for Decision-Aware Diabetes Modeling

Model ReleasesDGX agent

arXiv:2605.11247v1 Announce Type: new Abstract: This paper presents a proof-of-concept digital twin framework for simulation-driven diabetes modeling using benchmark clinical data, synthetic temporal

A proximal gradient algorithm for composite log-concave sampling

ResearchDGX agent

arXiv:2605.12461v1 Announce Type: cross Abstract: We propose an algorithm to sample from composite log-concave distributions over R^d, i.e., densities of the form pipropto e^{-f-g}, assuming access to

A Semi-Supervised Framework for Speech Confidence Detection using Whisper

ResearchDGX agent

arXiv:2605.12387v1 Announce Type: cross Abstract: Automatic detection of speaker confidence is critical for adaptive computing but remains constrained by limited labelled data and the subjectivity of

A Switching System Theory of Q-Learning with Linear Function Approximation

Model ReleasesDGX agent

arXiv:2605.11021v1 Announce Type: new Abstract: This paper develops a switching-system interpretation of Q-learning with linear function approximation (LFA) based on the joint spectral radius (JSR). W

A Unified Graph Language Model for Multi-Domain Multi-Task Graph Alignment Instruction Tuning

SafetyDGX agent

arXiv:2605.12197v1 Announce Type: new Abstract: Leveraging Graph Neural Networks (GNNs) as graph encoders and aligning the resulting representations with Large Language Models (LLMs) through alignment

Acceleration of horizontal numerical advection for atmospheric modeling through surrogate modeling with temporal coarse-graining

ResearchDGX agent

arXiv:2605.10956v1 Announce Type: cross Abstract: Machine-learned surrogate modeling of advection may accelerate geoscientific models, but existing approaches have either achieved limited speedup or h

ACSAC: Adaptive Chunk Size Actor-Critic with Causal Transformer Q-Network

SafetyDGX agent

arXiv:2605.11009v1 Announce Type: new Abstract: Long-horizon, sparse-reward tasks pose a fundamental challenge for reinforcement learning, since single-step TD learning suffers from bootstrapping erro

Adaptive Calibration in Non-Stationary Environments

ResearchDGX agent

arXiv:2605.11490v1 Announce Type: new Abstract: Making calibrated online predictions is a central challenge in modern AI systems. Much of the existing literature focuses on fully adversarial environme

Adaptive Policy Learning Under Unknown Network Interference

SafetyDGX agent

arXiv:2605.11191v1 Announce Type: cross Abstract: Adaptive experimentation under unknown network interference requires solving two coupled problems: (i) learning the underlying dynamics of interferenc

Adaptive TD-Lambda for Cooperative Multi-agent Reinforcement Learning

SafetyDGX agent

arXiv:2605.11880v1 Announce Type: new Abstract: TD(lambda) in value-based MARL algorithms or the Temporal Difference critic learning in Actor-Critic-based (AC-based) algorithms synergistically integra

ADMM-Q: An Improved Hessian-based Weight Quantizer for Post-Training Quantization of Large Language Models

Local AiDGX agent

arXiv:2605.11222v1 Announce Type: new Abstract: Quantization is an effective strategy to reduce the storage and computation footprint of large language models (LLMs). Post-training quantization (PTQ)

Adversarial Causal Tuning for Realistic Time-series Generation

ResearchDGX agent

arXiv:2506.02084v2 Announce Type: replace Abstract: We address the problem of generating simulated, yet realistic, time-series data from a causal model with the same observational and interventional d

Adversarial Effects on Expressibility and Trainability in Distributed Variational Quantum Algorithms

ResearchDGX agent

arXiv:2605.03629v1 Announce Type: cross Abstract: Distributed quantum algorithms offer a promising pathway to scale variational quantum algorithms beyond the constraints of noisy intermediate-scale qu

AESOP: Adversarial Execution-path Selection to Overload Deep Learning Pipelines

ApplicationsDGX agent

arXiv:2605.10987v1 Announce Type: new Abstract: Modern machine learning deployments increasingly compose specialized models into dynamic inference pipelines, where upstream components produce intermed

Agent-Based Post-Hoc Correction of Agricultural Yield Forecasts

Model ReleasesDGX agent

arXiv:2605.12375v1 Announce Type: new Abstract: Accurate crop yield forecasting in commercial soft fruit production is constrained by the data available in typical commercial farm records, which lack

Aligning Flow Map Policies with Optimal Q-Guidance

SafetyDGX agent

arXiv:2605.12416v1 Announce Type: new Abstract: Generative policies based on expressive model classes, such as diffusion and flow matching, are well-suited to complex control problems with highly mult

Analytical Provisioning for Attention-FFN Disaggregated LLM Serving under Stochastic Workloads

ResearchDGX agent

arXiv:2601.21351v3 Announce Type: replace Abstract: Attentio-FFN disaggregation (AFD) is an emerging architecture for LLM decoding that separates state-heavy, KV-cache-dominated Attention computation

Approximating Simple ReLU Networks based on Spectral Decomposition of Fisher Information

ResearchDGX agent

arXiv:2505.17907v2 Announce Type: replace-cross Abstract: Properties of Fisher information matrices of 2-layer neural ReLU networks with random hidden weights are studied. For these networks, it is kn

Approximation of Maximally Monotone Operators : A Graph Convergence Perspective

ResearchDGX agent

arXiv:2605.12301v1 Announce Type: new Abstract: Operator learning has been highly successful for continuous mappings between infinite-dimensional spaces, such as PDE solution operators. However, many

Approximation Theory of Laplacian-Based Neural Operators for Reaction-Diffusion System

Model ReleasesDGX agent

arXiv:2605.12025v1 Announce Type: new Abstract: Neural operators provide a framework for learning solution operators of partial differential equations (PDEs), enabling efficient surrogate modeling for

Arbitrated Indirect Treatment Comparisons

ResearchDGX agent

arXiv:2510.18071v2 Announce Type: replace-cross Abstract: Matching-adjusted indirect comparison (MAIC) has been increasingly employed in health technology assessments (HTA). By reweighting subjects fr

arepsilon-Good Action Identification in Fixed-Budget Monte Carlo Tree Search

ResearchDGX agent

arXiv:2605.11324v1 Announce Type: new Abstract: We study the fixed-budget max-min action identification problem in depth-2 max-min trees, an important special case of Monte Carlo Tree Search. A learne

ASD-Bench: A Four-Axis Comprehensive Benchmark of AI Models for Autism Spectrum Disorder

Model ReleasesDGX agent

arXiv:2605.11091v1 Announce Type: new Abstract: Automated ASD screening tools remain limited by single-architecture evaluations, axis-restricted assessment, and near-exclusive focus on adult cohorts,

Assessment of cloud and associated radiation fields from a GAN stochastic cloud subcolumn generator

SafetyDGX agent

arXiv:2605.11968v1 Announce Type: cross Abstract: Modern Earth System Models (ESMs) operate on horizontal scales far larger than typical cloud features, requiring stochastic subcolumn generators to re

Attacks and Mitigations for Distributed Governance of Agentic AI under Byzantine Adversaries

AgentsDGX agent

arXiv:2605.12364v1 Announce Type: cross Abstract: Agentic AI governance is a critical component of agentic AI infrastructure ensuring that agents follow their owner's communication and interaction pol

Augmented Lagrangian Method for Last-Iterate Convergence for Constrained MDPs

SafetyDGX agent

arXiv:2605.11694v1 Announce Type: new Abstract: We study policy optimization for infinite-horizon, discounted constrained Markov decision processes (CMDPs). While existing theoretical guarantees typic

Autoregressive Learning in Joint KL: Sharp Oracle Bounds and Lower Bounds

SafetyDGX agent

arXiv:2605.12316v1 Announce Type: new Abstract: We study the fundamental and timely problem of learning long sequences in autoregressive modeling and next-token prediction under model misspecification

Backbone-Equated Diffusion OOD via Sparse Internal Snapshots

Model ReleasesDGX agent

arXiv:2605.11014v1 Announce Type: new Abstract: Fair comparison between diffusion-based OOD detectors is challenging, as conclusions can vary with backbone choice, corruption parameterization, and tes

Bayesian Surrogate Training on Multiple Data Sources: A Hybrid Modeling Strategy

ApplicationsDGX agent

arXiv:2412.11875v3 Announce Type: replace-cross Abstract: Surrogate models are often used as computationally efficient approximations to complex simulation models, enabling tasks such as solving inver

Behavioral Mode Discovery for Fine-tuning Multimodal Generative Policies

ResearchDGX agent

arXiv:2605.11387v1 Announce Type: new Abstract: We address the problem of fine-tuning pre-trained generative policies with reinforcement learning (RL) while preserving the multimodality of their actio

Beyond GRPO and On-Policy Distillation: An Empirical Sparse-to-Dense Reward Principle for Language-Model Post-Training

Model ReleasesDGX agent

arXiv:2605.12483v1 Announce Type: new Abstract: In settings where labeled verifiable training data is the binding constraint, each checked example should be allocated carefully. The standard practice

Beyond Manual Curation: Augmenting Targeted Protein Degradation Databases via Agentic Literature Extraction Workflows

AgentsDGX agent

arXiv:2605.11221v1 Announce Type: cross Abstract: Predictive models in biomedicine depend on structured assay data locked in the text, tables, and supplements of primary publications. This bottleneck

Beyond Parameter Aggregation: Semantic Consensus for Federated Fine-Tuning of LLMs

Model ReleasesDGX agent

arXiv:2605.11857v1 Announce Type: new Abstract: Federated fine-tuning of large language models is commonly formulated as a parameter aggregation problem. However, even parameter-efficient methods requ

Beyond Point Estimates: Distributional Uncertainty in Machine Learning Performance Evaluation

ResearchDGX agent

arXiv:2501.16931v2 Announce Type: replace Abstract: Machine learning models are often evaluated using point estimates of performance metrics such as accuracy, F1 score, or mean squared error. Such sum

Beyond Prediction: Interval Neural Networks for Uncertainty-Aware System Identification

Model ReleasesDGX agent

arXiv:2605.11460v1 Announce Type: new Abstract: System identification (SysID) is critical for modeling dynamical systems from experimental data, yet traditional approaches often fail to capture nonlin

Beyond Similarity: Temporal Operator Attention for Time Series Analysis

ResearchDGX agent

arXiv:2605.11287v1 Announce Type: new Abstract: A persistent paradox in time-series forecasting is that structurally simple MLP and linear models often outperform high-capacity Transformers. We argue

Bin Latent Transformer (BiLT): A shift-invariant autoencoder for calibration-free spectral unmixing of turbid media

Model ReleasesDGX agent

arXiv:2605.11829v1 Announce Type: cross Abstract: The accurate recovery of constituent-level optical properties from integrating sphere measurements is a central analytical challenge in pharmaceutical

BLOCK-EM: Preventing Emergent Misalignment via Latent Blocking

ResearchDGX agent

arXiv:2602.00767v2 Announce Type: replace Abstract: Emergent misalignment can arise when a language model is fine-tuned on a narrowly scoped supervised objective: the model learns the target behavior,

Block-R1: Rethinking the Role of Block Size in Multi-domain Reinforcement Learning for Diffusion Large Language Models

Model ReleasesDGX agent

arXiv:2605.11726v1 Announce Type: new Abstract: Recently, reinforcement learning (RL) has been widely applied during post-training for diffusion large language models (dLLMs) to enhance reasoning with

BOOST: BOttleneck-Optimized Scalable Training Framework for Low-Rank Large Language Models

HardwareDGX agent

arXiv:2512.12131v2 Announce Type: replace Abstract: The scale of transformer model pre-training is constrained by the increasing computation and communication cost. Low-rank bottleneck architectures o

Breaking extit{Winner-Takes-All}: Cooperative Policy Optimization Improves Diverse LLM Reasoning

Model ReleasesDGX agent

arXiv:2605.11461v1 Announce Type: cross Abstract: Reinforcement learning with verifiers (RLVR) has become a central paradigm for improving LLM reasoning, yet popular group-based optimization algorithm

BSO: Safety Alignment Is Density Ratio Matching

SafetyDGX agent

arXiv:2605.12339v1 Announce Type: new Abstract: Aligning language models for both helpfulness and safety typically requires complex pipelines-separate reward and cost models, online reinforcement lear

CATS: Cascaded Adaptive Tree Speculation for Memory-Limited LLM Inference Acceleration

Model ReleasesDGX agent

arXiv:2605.11186v1 Announce Type: new Abstract: Auto-regressive decoding in Large Language Models (LLMs) is inherently memory-bound: every generation step requires loading the model weights and interm

Causal Algorithmic Recourse: Foundations and Methods

ApplicationsDGX agent

arXiv:2605.11373v1 Announce Type: cross Abstract: The trustworthiness of AI decision-making systems is increasingly important. A key feature of such systems is the ability to provide recommendations f

← Previous
1…161162163164165…243
Next →