AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,329 results
16 Apr 2026

Jump-Start Reinforcement Learning with Vision-Language-Action Regularization

SafetyDGX agent

arXiv:2604.13733v1 Announce Type: new Abstract: Reinforcement learning (RL) enables high-frequency, closed-loop control for robotic manipulation, but scaling to long-horizon tasks with sparse or imper

KV Packet: Recomputation-Free Context-Independent KV Caching for LLMs

Model ReleasesDGX agent

arXiv:2604.13226v1 Announce Type: new Abstract: Large Language Models (LLMs) rely heavily on Key-Value (KV) caching to minimize inference latency. However, standard KV caches are context-dependent: re

Learning-Based Estimation of Spatially Resolved Scatter Radiation Fields in Interventional Radiology

ResearchDGX agent

arXiv:2512.17654v3 Announce Type: replace Abstract: We present three variants of a lightweight, fully connected artificial neural network, suited for interactive estimation of three-dimensional, spati


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Learning Dynamics from Input-Output Data with Hamiltonian Gaussian Processes

ApplicationsDGX agent

arXiv:2511.05330v2 Announce Type: replace Abstract: Embedding non-restrictive prior knowledge, such as energy conservation laws, into learning methods is a key motive to construct physically consisten

Learning from Change: Predictive Models for Incident Prevention in a Regulated IT Environment

ApplicationsDGX agent

arXiv:2604.13462v1 Announce Type: cross Abstract: Effective IT change management is important for businesses that depend on software and services, particularly in highly regulated sectors such as fina

Learning Inference Concurrency in DynamicGate MLP Structural and Mathematical Justification

ResearchDGX agent

arXiv:2604.13546v1 Announce Type: new Abstract: Conventional neural networks strictly separate learning and inference because if parameters are updated during inference, outputs become unstable and ev

Learning Probabilistic Responsibility Allocations for Multi-Agent Interactions

SafetyDGX agent

arXiv:2604.13128v1 Announce Type: cross Abstract: Human behavior in interactive settings is shaped not only by individual objectives but also by shared constraints with others, such as safety. Underst

LEGO-MOF: Equivariant Latent Manipulation for Editable, Generative, and Optimizable MOF Design

ResearchDGX agent

arXiv:2604.13520v1 Announce Type: new Abstract: Metal-organic frameworks (MOFs) are highly promising for carbon capture, yet navigating their vast design space remains challenging. Recent deep generat

Linear Probe Accuracy Scales with Model Size and Benefits from Multi-Layer Ensembling

ResearchDGX agent

arXiv:2604.13386v1 Announce Type: new Abstract: Linear probes can detect when language models produce outputs they 'know' are wrong, a capability relevant to both deception and reward hacking. However

LongCoT: Benchmarking Long-Horizon Chain-of-Thought Reasoning

Model ReleasesDGX agent

arXiv:2604.14140v1 Announce Type: new Abstract: As language models are increasingly deployed for complex autonomous tasks, their ability to reason accurately over longer horizons becomes critical. An

LoRA-MME: Multi-Model Ensemble of LoRA-Tuned Encoders for Code Comment Classification

Model ReleasesDGX agent

arXiv:2603.03959v4 Announce Type: replace-cross Abstract: Code comment classification is a critical task for automated software documentation and analysis. In the context of the NLBSE'26 Tool Competit

MAny: Merge Anything for Multimodal Continual Instruction Tuning

Model ReleasesDGX agent

arXiv:2604.14016v1 Announce Type: new Abstract: Multimodal Continual Instruction Tuning (MCIT) is essential for sequential task adaptation of Multimodal Large Language Models (MLLMs) but is severely r

MDPs with a State Sensing Cost

Model ReleasesDGX agent

arXiv:2505.03280v3 Announce Type: replace Abstract: In many practical sequential decision-making problems, tracking the state of the environment incurs a sensing/communication/computation cost. In the

Minimax Optimality and Spectral Routing for Majority-Vote Ensembles under Markov Dependence

ResearchDGX agent

arXiv:2604.13414v1 Announce Type: new Abstract: Majority-vote ensembles achieve variance reduction by averaging over diverse, approximately independent base learners. When training data exhibits Marko

Mitigating Barren Plateaus in Quantum Denoising Diffusion Probabilistic Model

ResearchDGX agent

arXiv:2512.06695v2 Announce Type: replace Abstract: Quantum generative models exploit quantum superposition and entanglement to enhance learning efficiency for both classical and quantum data. Recentl

mLaSDI: Multi-stage latent space dynamics identification

ResearchDGX agent

arXiv:2506.09207v4 Announce Type: replace Abstract: Accurately solving partial differential equations (PDEs) is essential across many scientific disciplines. However, high-fidelity solvers can be comp

Mobius transforms and Shapley values for vector-valued functions on weighted directed acyclic multigraphs

ResearchDGX agent

arXiv:2510.05786v3 Announce Type: replace-cross Abstract: Mobius inversion and Shapley values are two mathematical tools for characterizing and decomposing higher-order structure in complex systems. T

Modeling Student Learning with 3.8 Million Program Traces

TutorialsDGX agent

arXiv:2510.05056v2 Announce Type: replace Abstract: As programmers write code, they often edit and retry multiple times, creating rich 'interaction traces' that reveal how they approach coding tasks a

MolCryst-MLIPs: A Machine-Learned Interatomic Potentials Database for Molecular Crystals

Model ReleasesDGX agent

arXiv:2604.13897v1 Announce Type: new Abstract: We present an open Molecular Crystal (MC) database of Machine-Learned Interatomic Potentials (MLIP) called MolCryst-MLIPs. The first release comprises f

Momentum Further Constrains Sharpness at the Edge of Stochastic Stability

ResearchDGX agent

arXiv:2604.14108v1 Announce Type: new Abstract: Recent work suggests that (stochastic) gradient descent self-organizes near an instability boundary, shaping both optimization and the solutions found.

Monthly Diffusion v0.9: A Latent Diffusion Model for the First AI-MIP

ResearchDGX agent

arXiv:2604.13481v1 Announce Type: new Abstract: Here, we describe Monthly Diffusion at 1.5-degree grid spacing (MD-1.5 version 0.9), a climate emulator that leverages a spherical Fourier neural operat

MOONSHOT : A Framework for Multi-Objective Pruning of Vision and Large Language Models

Model ReleasesDGX agent

arXiv:2604.13287v1 Announce Type: new Abstract: Weight pruning is a common technique for compressing large neural networks. We focus on the challenging post-training one-shot setting, where a pre-trai

Multi-Task LLM with LoRA Fine-Tuning for Automated Cancer Staging and Biomarker Extraction

Model ReleasesDGX agent

arXiv:2604.13328v1 Announce Type: new Abstract: Pathology reports serve as the definitive record for breast cancer staging, yet their unstructured format impedes large-scale data curation. While Large

Multistage Conditional Compositional Optimization

ResearchDGX agent

arXiv:2604.14075v1 Announce Type: cross Abstract: We introduce Multistage Conditional Compositional Optimization (MCCO) as a new paradigm for decision-making under uncertainty that combines aspects of

Nested Fourier-enhanced neural operator for efficient modeling of radiation transfer in fires

Local AiDGX agent

arXiv:2604.13919v1 Announce Type: cross Abstract: Computational fluid dynamics (CFD) has become an essential tool for predicting fire behavior, yet maintaining both efficiency and accuracy remains cha

Neural architectures for resolving references in program code

ApplicationsDGX agent

arXiv:2604.14073v1 Announce Type: new Abstract: Resolving and rewriting references is fundamental in programming languages. Motivated by a real-world decompilation task, we abstract reference rewritin

Neural Mean-Field Games: Extending Mean-Field Game Theory with Neural Stochastic Differential Equations

SafetyDGX agent

arXiv:2504.13228v4 Announce Type: replace Abstract: Mean-field game theory relies on approximating games that are intractable to model due to a very large to infinite population of players. While thes

node2vec or triangle-biased random walks: stationarity, regularity & recurrence

ResearchDGX agent

arXiv:2604.13681v1 Announce Type: cross Abstract: The node2vec random walk is a non-Markovian random walk on the vertex set of a graph, widely used for network embedding and exploration. This random w

Nonparametric Sparse Online Learning of the Koopman Operator

TutorialsDGX agent

arXiv:2405.07432v4 Announce Type: replace-cross Abstract: The Koopman operator provides a powerful framework for representing the dynamics of general nonlinear dynamical systems. However, existing dat

Numerical Instability and Chaos: Quantifying the Unpredictability of Large Language Models

AgentsDGX agent

arXiv:2604.13206v1 Announce Type: cross Abstract: As Large Language Models (LLMs) are increasingly integrated into agentic workflows, their unpredictability stemming from numerical instability has eme

On an L^2 norm for stationary ARMA processes

ResearchDGX agent

arXiv:2408.10610v5 Announce Type: replace Abstract: We propose an L^2 norm for stationary Autoregressive Moving Average (ARMA) models. We look at ARMA models within the Hilbert space of the past with

On the Fundamental Limitations of Dual Static CVaR Decompositions in Markov Decision Processes

SafetyDGX agent

arXiv:2507.14005v2 Announce Type: replace Abstract: It was recently shown that dynamic programming (DP) methods for finding static CVaR-optimal policies in Markov Decision Processes (MDPs) can fail wh

Online learning with noisy side observations

Model ReleasesDGX agent

arXiv:2604.13740v1 Announce Type: new Abstract: We propose a new partial-observability model for online learning problems where the learner, besides its own loss, also observes some noisy feedback abo

Optimization with SpotOptim

Model ReleasesDGX agent

arXiv:2604.13672v1 Announce Type: new Abstract: The `spotoptim` package implements surrogate-model-based optimization of expensive black-box functions in Python. Building on two decades of Sequential

Optimizing Earth Observation Satellite Schedules under Unknown Operational Constraints: An Active Constraint Acquisition Approach

TutorialsDGX agent

arXiv:2604.13283v1 Announce Type: cross Abstract: Earth Observation (EO) satellite scheduling (deciding which imaging tasks to perform and when) is a well-studied combinatorial optimization problem. E

Ordinary Least Squares is a Special Case of Transformer

Model ReleasesDGX agent

arXiv:2604.13656v1 Announce Type: new Abstract: The statistical essence of the Transformer architecture has long remained elusive: Is it a universal approximator, or a neural network version of known

Out of Context: Reliability in Multimodal Anomaly Detection Requires Contextual Inference

Model ReleasesDGX agent

arXiv:2604.13252v1 Announce Type: new Abstract: Anomaly detection aims to identify observations that deviate from expected behavior. Because anomalous events are inherently sparse, most frameworks are

Outperforming Self-Attention Mechanisms in Solar Irradiance Forecasting via Physics-Guided Neural Networks

ResearchDGX agent

arXiv:2604.13455v1 Announce Type: new Abstract: Accurate Global Horizontal Irradiance (GHI) forecasting is critical for grid stability, particularly in arid regions characterized by rapid aerosol fluc

Parameter-efficient Quantum Multi-task Learning

Model ReleasesDGX agent

arXiv:2604.13560v1 Announce Type: new Abstract: Multi-task learning (MTL) improves generalization and data efficiency by jointly learning related tasks through shared representations. In the widely us

Parameter-Free Non-Ergodic Extragradient Algorithms for Solving Monotone Variational Inequalities

Model ReleasesDGX agent

arXiv:2604.07662v2 Announce Type: replace-cross Abstract: Monotone variational inequalities (VIs) provide a unifying framework for convex minimization, equilibrium computation, and convex-concave sadd

Pareto-Optimal Offline Reinforcement Learning via Smooth Tchebysheff Scalarization

SafetyDGX agent

arXiv:2604.13175v1 Announce Type: new Abstract: Large language models can be aligned with human preferences through offline reinforcement learning (RL) on small labeled datasets. While single-objectiv

Physics-Informed Neural Networks for Methane Sorption: Cross-Gas Transfer Learning, Ensemble Collapse Under Physics Constraints, and Monte Carlo Dropout Uncertainty Quantification

ResearchDGX agent

arXiv:2604.13992v1 Announce Type: new Abstract: Accurate methane sorption prediction across heterogeneous coal ranks requires models that combine thermodynamic consistency, efficient knowledge transfe

Physics-Informed Neural Networks for Solving Derivative-Constrained PDEs

ResearchDGX agent

arXiv:2604.13723v1 Announce Type: new Abstract: Physics-Informed Neural Networks (PINNs) recast PDE solving as an optimisation problem in function space by minimising a residual-based objective, yet m

Physics-informed reservoir characterization from bulk and extreme pressure events with a differentiable simulator

ResearchDGX agent

arXiv:2604.13291v1 Announce Type: new Abstract: Accurate characterization of subsurface heterogeneity is challenging but essential for applications such as reservoir pressure management, geothermal en

Power Transform Revisited: Numerically Stable, and Federated

ApplicationsDGX agent

arXiv:2510.04995v3 Announce Type: replace Abstract: Power transforms are popular parametric methods for making data more Gaussian-like, and are widely used as preprocessing steps in statistical analys

Predicting Time Pressure of Powered Two-Wheeler Riders for Proactive Safety Interventions

Model ReleasesDGX agent

arXiv:2601.03173v2 Announce Type: replace Abstract: Time pressure critically influences risky maneuvers and crash proneness among powered two-wheeler riders, yet its prediction remains underexplored i

PRiMeFlow: Capturing Complex Expression Heterogeneity in Perturbation Response Modelling

ResearchDGX agent

arXiv:2604.13986v1 Announce Type: new Abstract: Predicting the effects of perturbations in-silico on cell state can identify drivers of cell behavior at scale and accelerate drug discovery. However, m

Provably Efficient Offline-to-Online Value Adaptation with General Function Approximation

ResearchDGX agent

arXiv:2604.13966v1 Announce Type: new Abstract: We study value adaptation in offline-to-online reinforcement learning under general function approximation. Starting from an imperfect offline pretraine

Quantifying and Understanding Uncertainty in Large Reasoning Models

ResearchDGX agent

arXiv:2604.13395v1 Announce Type: cross Abstract: Large Reasoning Models (LRMs) have recently demonstrated significant improvements in complex reasoning. While quantifying generation uncertainty in LR

Quantum Machine Learning for Colorectal Cancer Data: Anastomotic Leak Classification and Risk Factors

ResearchDGX agent

arXiv:2604.13951v1 Announce Type: new Abstract: This study evaluates colorectal risk factors and compares classical models against Quantum Neural Networks (QNNs) for anastomotic leak prediction. Analy

Random Walk Learning and the Pac-Man Attack

ApplicationsDGX agent

arXiv:2508.05663v4 Announce Type: replace-cross Abstract: Random walk (RW)-based algorithms have long been popular in distributed systems due to low overheads and scalability, with recent growing appl

Randomized Neural Networks for Integro-Differential Equations with Application to Neutron Transport

ResearchDGX agent

arXiv:2604.13830v1 Announce Type: cross Abstract: Integro-differential equations arise in a wide range of applications, including transport, kinetic theory, radiative transfer, and multiphysics modeli

RANDPOL: Parameter-Efficient End-to-End Quadruped Locomotion via Randomized Policy Learning

Model ReleasesDGX agent

arXiv:2505.19054v2 Announce Type: replace Abstract: Modern learning-based locomotion controllers typically rely on fully trainable deep neural networks with a large number of parameters. This paper st

Rare Event Analysis via Stochastic Optimal Control

Model ReleasesDGX agent

arXiv:2604.13213v1 Announce Type: cross Abstract: Rare events such as conformational changes in biomolecules, phase transitions, and chemical reactions are central to the behavior of many physical sys

Reachability Constraints in Variational Quantum Circuits: Optimization within Polynomial Group Module

ResearchDGX agent

arXiv:2604.13735v1 Announce Type: cross Abstract: This work identifies a necessary condition for any variational quantum approach to reach the exact ground state. Briefly, the norms of the projections

Representation over Routing: Overcoming Surrogate Hacking in Multi-Timescale PPO

SafetyDGX agent

arXiv:2604.13517v1 Announce Type: new Abstract: Temporal credit assignment in reinforcement learning has long been a central challenge. Inspired by the multi-timescale encoding of the dopamine system

ReproMIA: A Comprehensive Analysis of Model Reprogramming for Proactive Membership Inference Attacks

ResearchDGX agent

arXiv:2603.28942v3 Announce Type: replace Abstract: The pervasive deployment of deep learning models across critical domains has concurrently intensified privacy concerns due to their inherent propens

Restless Bandits with Individual Penalty Constraints: A New Near-Optimal Index Policy and How to Learn It

SafetyDGX agent

arXiv:2604.04101v2 Announce Type: replace Abstract: This paper investigates the Restless Multi-Armed Bandit (RMAB) framework under individual penalty constraints to address resource allocation challen

Reward Hacking in the Era of Large Models: Mechanisms, Emergent Misalignment, Challenges

Model ReleasesDGX agent

arXiv:2604.13602v1 Announce Type: new Abstract: Reinforcement Learning from Human Feedback (RLHF) and related alignment paradigms have become central to steering large language models (LLMs) and multi

RiskWebWorld: A Realistic Interactive Benchmark for GUI Agents in E-commerce Risk Management

Model ReleasesDGX agent

arXiv:2604.13531v1 Announce Type: cross Abstract: Graphical User Interface (GUI) agents show strong capabilities for automating web tasks, but existing interactive benchmarks primarily target benign,

← Previous
1…225226227228229…239
Next →