AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,432 results
30 Jun 2026

Entropy-Regularized Reinforcement Learning for Linear-Quadratic Stackelberg Differential Games in Regime-Switching Diffusion Models

ResearchDGX agent

arXiv:2606.28671v1 Announce Type: new Abstract: Stackelberg differential games (SDGs) provide a powerful framework for hierarchical decision-making in stochastic and continuous-time environments, yet

Entropy Regularized Reinforcement Learning for Zero-Sum Stochastic Differential Games in a Regime-Switching Jump-Diffusion Process

Model ReleasesDGX agent

arXiv:2606.28669v1 Announce Type: new Abstract: To address parameter misspecification and sudden structural environmental changes in conventional stochastic differential game (SDG) frameworks, this pa

Experience Augmented Policy Optimization for LLM Reasoning

Model Releases

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2606.30420v1 Announce Type: new Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) is a powerful paradigm for improving the reasoning capabilities of large language models (LLMs). H

Expert-guided Clinical Text Augmentation via Query-Based Model Collaboration

SafetyDGX agent

arXiv:2509.21530v2 Announce Type: replace Abstract: Data augmentation is a widely used strategy to improve model robustness and generalization by enriching training datasets with synthetic examples. W

Exploring Differences Between Tabular Enterprise Data and Public Benchmarks

ApplicationsDGX agent

arXiv:2606.30452v1 Announce Type: new Abstract: Tabular data dominate the landscape of data science, increasingly attracting innovative machine learning models and tailored benchmarks. Yet, little is

Exploring the Cryptographic Limits of Transformer Networks

ResearchDGX agent

arXiv:2606.29389v1 Announce Type: cross Abstract: In recent work it has been shown that colluding AI agents can use steganographic methods to exchange malicious information. Whether a transformer can

Exploring the Effects of Entanglement on Quantum Machine Learning of Pathogen Epitope-Receptor Binding

Model ReleasesDGX agent

arXiv:2606.28655v1 Announce Type: cross Abstract: Parameterized quantum circuits (PQCs) provide a flexible substrate for hybrid quantum machine learning (QML), but their practical value on Noisy Inter

Extrapolating from Regularised Solutions for Solving Ill-Conditioned Linear Systems in Machine Learning

ResearchDGX agent

arXiv:2606.30328v1 Announce Type: cross Abstract: Rapid prototyping of algorithms is a critical step in modern machine learning. Most algorithms exploit linear algebra, creating a need for lightweight

Factorizable Normalizing Flows for parameter-dependent density morphing

Model ReleasesDGX agent

arXiv:2606.30489v1 Announce Type: cross Abstract: Normalizing Flows excel at modeling a single fixed density, yet many problems across the sciences, such as high energy physics, instead require modeli

Falsifying Discriminant Validity of Predictive Algorithms

ResearchDGX agent

arXiv:2601.17146v2 Announce Type: replace-cross Abstract: Empirical investigations into unintended model behavior often show that the algorithm is predicting another outcome than what was intended. Th

Favorability of Loss Landscape with Weight Decay Requires Both Large Overparametrization and Initialization

ResearchDGX agent

arXiv:2505.22578v2 Announce Type: replace Abstract: The optimization of neural networks under weight decay remains poorly understood from a theoretical standpoint. While weight decay is standard pract

Federated Graph Learning for EV Charging Demand Forecasting with Personalization Against Cyberattacks

ResearchDGX agent

arXiv:2405.00742v2 Announce Type: replace-cross Abstract: Mitigating cybersecurity risk in electric vehicle (EV) charging demand forecasting plays a crucial role in the safe operation of collective EV

fev-bench: A Realistic Benchmark for Time Series Forecasting

Model ReleasesDGX agent

arXiv:2509.26468v3 Announce Type: replace Abstract: Benchmark quality is critical for meaningful evaluation and sustained progress in time series forecasting, particularly with the rise of pretrained

Few-Step Boltzmann Generators via Scalable Likelihood Flow Maps

ResearchDGX agent

arXiv:2606.29110v1 Announce Type: new Abstract: Recent progress in flow-based generative modeling has led to models that output high-quality samples while using only a small number of function evaluat

FlexTab: A Flexible Encoder-Decoder Architecture for In-Context Learning Across Diverse Tabular Tasks

ApplicationsDGX agent

arXiv:2606.30336v1 Announce Type: new Abstract: We introduce FlexTab, a flexible encoder-decoder architecture for in-context learning on tabular data that pairs a single, task-agnostic encoder with a

FlipGuard: Defending Large Language Models Against Quantization-Conditioned Backdoor Attacks

Model ReleasesDGX agent

arXiv:2606.28962v1 Announce Type: cross Abstract: Model quantization is essential for the efficient deployment of Large Language Models (LLMs), but introduces a critical vulnerability: Quantization-Co

Forensic Trajectory Signatures for Agent Memory Poisoning Detection

Model ReleasesDGX agent

arXiv:2606.30566v1 Announce Type: cross Abstract: We discover a behavioral invariant in LLM agents under persistent memory poisoning: in architectures where routing information is retrieved through ob

Fourier Neural Operators with Least-Squares Readout Refit for Learning Random Obstacle-to-Solution Maps

TutorialsDGX agent

arXiv:2606.29436v1 Announce Type: cross Abstract: We study operator learning for random obstacle-to-solution maps arising from elliptic variational inequalities with finite-band self-affine random obs

Fractional Stochastic Neural Networks

ResearchDGX agent

arXiv:2606.29438v1 Announce Type: cross Abstract: In this paper, we develop a fractional stochastic neural network with residual dynamics driven by fractional Brownian motion. By introducing a discret

Freeze, Prompt, and Adapt: A Framework for Source-free Unsupervised GNN Prompting

Model ReleasesDGX agent

arXiv:2505.16903v2 Announce Type: replace Abstract: Prompt tuning has become a key mechanism for adapting pre-trained Graph Neural Networks (GNNs) to new downstream tasks. However, existing approaches

Friend or Foe

ResearchDGX agent

arXiv:2509.00123v2 Announce Type: replace-cross Abstract: A fundamental challenge in microbial ecology is determining whether bacteria compete or cooperate in different environmental conditions. With

From Failure Taxonomy to Intervention: A Diagnostic Methodology for Industry-Scale AVLM in Video and Live-Streaming Platform Moderation

Model ReleasesDGX agent

arXiv:2606.30059v1 Announce Type: new Abstract: Industry-scale video and live-streaming moderation imposes requirements that are difficult to satisfy with generic pretrained public models or external

Generalization Analysis of Transformers in Distribution Regression

Model ReleasesDGX agent

arXiv:2606.29256v1 Announce Type: cross Abstract: In recent years, models based on the Transformer architecture have seen widespread applications and have become one of the core tools in the field of

Generalization error of min-norm interpolators in transfer learning

SafetyDGX agent

arXiv:2406.13944v2 Announce Type: replace-cross Abstract: This paper establishes the generalization error of pooled min-ell_2-norm interpolation in transfer learning, where data from diverse distribut

Generation of Uncertainty-Aware High-Level Spatial Concepts in Factorized 3D Scene Graphs via Graph Neural Networks

ResearchDGX agent

arXiv:2409.11972v4 Announce Type: replace-cross Abstract: Enabling robots to autonomously discover high-level spatial concepts (e.g., rooms and walls) from primitive geometric observations (e.g., plan

Generative Learning as a Tool to Improve Perception of Emotional Body Motion Expressions

SafetyDGX agent

arXiv:2606.28769v1 Announce Type: new Abstract: Emotional body motion expressions are an essential element of non-verbal communication. Effectively conveying these expressions through technology is of

Geometric Algebra Meets Cartesian Tensors: Higher-Order Equivariance for Interatomic Potentials

ResearchDGX agent

arXiv:2606.29584v1 Announce Type: cross Abstract: Cl(3,0) interatomic potentials, despite their algebraic elegance, predict force magnitudes accurately but force directions poorly. Across ten rMD17 mo

Geometrically Principled Randomized Optimization for Efficient LLM Training

Model ReleasesDGX agent

arXiv:2510.01878v2 Announce Type: replace Abstract: Low-rank gradient optimization for large language models is currently divided into two categories: structured methods that rigorously identify subsp

GLACIER: Rethinking Mass Spectrum Prediction as an Object Detection Problem

ResearchDGX agent

arXiv:2606.29161v1 Announce Type: new Abstract: Predicting tandem mass spectra (MS/MS) from molecular structures represents a central task in analytical chemistry with direct relevance to clinical met

GLIP: Graph and LLM Joint Pretraining for Graph-Level Tasks

Local AiDGX agent

arXiv:2606.29773v1 Announce Type: new Abstract: Graphs are widely used to model relational systems, with applications in domains such as social networks, finance, and biomedicine. Graph neural network

Golden Hour Divide: Trauma Care Accessibility and Resource Vulnerability in Sri Lanka

SafetyDGX agent

arXiv:2606.29889v1 Announce Type: new Abstract: Timely intensive care dictates survival, yet emergency infrastructure remains unevenly distributed across Sri Lanka. While pre-hospital services have ex

GPU Parallelization Strategies for Forward and Backward Propagation in Shallow Neural Networks: A CUDA-Based Comparative Study

HardwareDGX agent

arXiv:2606.30497v1 Announce Type: cross Abstract: We present a comparative study of CUDA optimization strategies applied to forward and backward propagation in a shallow neural network. Three stacked

Gradient Boosted Mixed Models: Flexible Estimation of Mean and Variance Components for Clustered Data

ApplicationsDGX agent

arXiv:2511.00217v2 Announce Type: replace-cross Abstract: We introduce Gradient Boosted Mixed Models (GBMixed), a framework which extends boosting to clustered data by jointly modeling the mean and va

Gradient boosting with vector-valued leafs

ResearchDGX agent

arXiv:2606.29326v1 Announce Type: cross Abstract: Gradient boosting in the form of decision tree ensembles has successfully been applied to a variety of problems using simple objective functions based

Guided Unconditional and Conditional Generative Models for Super-Resolution and Inference of Quasi-Geostrophic Turbulence

ResearchDGX agent

arXiv:2507.00719v3 Announce Type: replace-cross Abstract: Typically, numerical simulations of Earth systems are coarse, and Earth observations are sparse and gappy. We apply four generative diffusion

Harvesting AI Computation at the Edge via Generic Approximation

ResearchDGX agent

arXiv:2606.29518v1 Announce Type: cross Abstract: With the widespread adoption of AI in various IoT scenarios such as smart sensing and processing, AI chips have become a common component at the edge.

Heads, Not Backbones: Output Heads Dominate Architectures on Fat-Tailed Returns

ResearchDGX agent

arXiv:2606.30037v1 Announce Type: new Abstract: In a deep forecasting pipeline for fat-tailed financial returns at short horizons, which matters more - the backbone architecture or the output head? We

HieraMix: A Hierarchical MLP-Mixer for Large-Scale Traffic Forecasting

ApplicationsDGX agent

arXiv:2512.07854v2 Announce Type: replace Abstract: Traffic forecasting task is significant to modern urban management. Recently, there is growing attention on large-scale forecasting, as it better re

Hierarchical Decision Making with Structured Policies: A Principled Design via Inverse Optimization

SafetyDGX agent

arXiv:2606.28764v1 Announce Type: new Abstract: Hierarchical decision-making frameworks are pivotal for addressing complex control tasks, enabling agents to decompose intricate problems into manageabl

High-Resolution Climate Projections Using Diffusion-Based Downscaling of a Lightweight Climate Emulator

ResearchDGX agent

arXiv:2602.13416v2 Announce Type: replace Abstract: The proliferation of data-driven models in weather and climate sciences has marked a significant paradigm shift, with advanced models demonstrating

Highly Data Parallelizable Estimation of the Sliced-Wasserstein Distance Using Cumulative Distribution Functions

ResearchDGX agent

arXiv:2606.30310v1 Announce Type: cross Abstract: The Sliced Wasserstein (SW) distance has emerged as a computationally attractive alternative to the Wasserstein distance by leveraging one-dimensional

How Far Can Sharpness and Complexity Jointly Explain Generalization?

Model ReleasesDGX agent

arXiv:2606.29043v1 Announce Type: new Abstract: Sharpness and complexity are two central factors in the generalization analysis of deep neural networks. Existing quantitative evaluations of generaliza

How Should World Models Be Evaluated for Embodied Decision-Making? A Decision-Making-Centric Position

SafetyDGX agent

arXiv:2606.15032v2 Announce Type: replace Abstract: World models have become a central abstraction in modern AI. The term now refers to several different objects: action-conditioned environment models

How Token Influence Decays with Distance: A Green-Function View of Trained Language Models

ResearchDGX agent

arXiv:2606.29139v1 Announce Type: new Abstract: We study how the next-token prediction of an autoregressive Transformer language model changes under small perturbations of earlier input token embeddin

HSAP: A Hierachical Sequence-aware Parallelism for Hybrid-Context Generative Models

ResearchDGX agent

arXiv:2606.30460v1 Announce Type: new Abstract: In this paper, we aim to combine the advantages of existing sequence parallelism paradigms and overcomes their drawbacks, the most serious of which is t

Hybrid Active-Online Learning Framework for Label-Efficient Concept Drift Adaptation in Optical Network Failure Detection

ResearchDGX agent

arXiv:2606.30322v1 Announce Type: new Abstract: We propose a hybrid active-online learning framework for label-efficient concept drift adaptation in optical network failure detection. Using margin-bas

I-BBS: Coordinate-Free Inference of Latent Sub-Manifolds Using Random Distance Matrix Theory

Model ReleasesDGX agent

arXiv:2606.29675v1 Announce Type: new Abstract: Bogomolny, Bohigas and Schmit (BBS) found that the spectrum of the pairwise distance matrix on N points sampled from a smooth d-dimensional manifold enc

IG-Lens: Exact Additive Probability Attribution Across Transformer Layers via Telescoping Integrated Gradients

ResearchDGX agent

arXiv:2606.29693v1 Announce Type: new Abstract: We ask a simple question about decoder-only transformers: between which two layers is the probability of a predicted token actually produced? Existing l

Implementation of Hyperelastic Physics-Augmented Neural Networks in the Explicit Finite Element Codes Simcenter Radioss and OpenRadioss with Applications to Impact Events

ResearchDGX agent

arXiv:2606.29874v1 Announce Type: cross Abstract: Data-driven material modeling techniques have gained significant attention due to their ability to capture complex constitutive behaviors beyond the l

Improved Multi-Dimensional Forecasting for Swap Regret

AgentsDGX agent

arXiv:2606.29533v1 Announce Type: cross Abstract: We study the problem of forecasting for an arbitrary number of downstream agents with unknown objectives, each of whom best responds to the forecaster

Improved Predictive Performance and Interpretability for Mesomorphic Neural Networks Using Local Fidelity Regularization

Model ReleasesDGX agent

arXiv:2606.29951v1 Announce Type: new Abstract: Interpretable Mesomorphic Neural Networks (IMNs) offer a promising framework that combines the predictive power of deep neural networks with the interpr

Improving Coherence in Hierarchical Time Series Forecasting using Structured Temporal Fusion

Model ReleasesDGX agent

arXiv:2606.28553v1 Announce Type: new Abstract: In many real-world applications, such as retail sales, energy usage, and supply chain planning, forecasting is performed across hierarchical structures.

Improving Patient Subtyping on Longitudinal Data using Representations from Mamba-based Architecture

ApplicationsDGX agent

arXiv:2606.28623v1 Announce Type: new Abstract: Effective sub-typing (also known as grouping or clustering) of patients using their electronic health record (EHR) data can greatly inform precision med

In-Vehicle Digital Twin-Based Collision Warning Framework with Sybil Attack Detection

SafetyDGX agent

arXiv:2606.28625v1 Announce Type: cross Abstract: Connected Vehicles (CVs) rely extensively on communication technologies to enable data-driven predictive analyses for enhancing performance and safety

Inexact calculus of variations on the hyperspherical tangent bundle with connections to the attention mechanism

ResearchDGX agent

arXiv:2507.15431v4 Announce Type: replace Abstract: We offer a theoretical mathematical background through Lagrangian optimization on the unit hyperspherical manifold and its tangential structure. Our

Inference-time optimization for experiment-grounded protein ensemble generation

SafetyDGX agent

arXiv:2602.24007v3 Announce Type: replace-cross Abstract: Protein function relies on dynamic conformational ensembles, yet current generative models like AlphaFold3 often fail to produce ensembles tha

Internal-State Probes Read the Situation, Not the Action: Three Negative Results for Pre-Action Misalignment Monitoring

Model ReleasesDGX agent

arXiv:2606.30449v1 Announce Type: new Abstract: Probes on model internals could help monitor agentic systems if they identify harmful text or tool actions before those actions are generated. We ask wh

Interventional Flow Matching: Prospective Dose-Response Forecasting with Velocity-Field Jacobian Regularization

SafetyDGX agent

arXiv:2606.29386v1 Announce Type: new Abstract: Predicting a patient's physiological trajectory under a planned treatment sequence is a prospective interventional problem, not standard time-series ext

Investigating ECG Diagnosis with Ambiguous Labels using Partial Label Learning

TutorialsDGX agent

arXiv:2512.11095v2 Announce Type: replace Abstract: Label ambiguity is an inherent and largely unaddressed challenge in real-world electrocardiogram (ECG) diagnosis, arising from overlapping condition

ITSPACE: Monotone Gaussian Optimal Transport Updates

SafetyDGX agent

arXiv:2606.30523v1 Announce Type: new Abstract: Covariance matrices serve as compact descriptors of feature distributions in many machine-learning pipelines, including domain adaptation and Gaussian e

← Previous
1…6061626364…241
Next →