AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,557 results
2 Jun 2026

Taming the Loss Landscape of PINNs with Noisy Feynman-Kac Supervision: Operator Preconditioning and Non-Asymptotic Error Bounds

ResearchDGX agent

arXiv:2606.00643v1 Announce Type: cross Abstract: Physics-Informed Neural Networks (PINNs) often train slowly or fail to converge on challenging partial differential equations (PDEs), a behavior recen

Target localization, identification and sensing using latent symmetries

Local AiDGX agent

arXiv:2606.01421v1 Announce Type: new Abstract: We show that an array of scatterers which has been designed to have latent ('hidden') symmetries can be used as a sensor. We use the capacitance matrix

Task-Induced Representational Invariances Depend on Learning Objective in Deep RL

SafetyDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2606.01868v1 Announce Type: new Abstract: Reinforcement Learning (RL) has long served as a model for goal-directed animal behavior in neuroscience. Modern deep RL has shown remarkable success ac

Temporal Motif Signatures for Temporal Graph Neural Networks

ResearchDGX agent

arXiv:2606.01176v1 Announce Type: new Abstract: Real temporal interaction streams carry predictive structure in short-horizon motif patterns -- repetition, reciprocity, star diversity, triadic flow --

The Assistant as a Privileged Persona: A canonical reference in cross-persona self-recognition

Model ReleasesDGX agent

arXiv:2606.00545v1 Announce Type: new Abstract: Post-trained language models can recognize their own outputs from a sentence or two out of context. In a companion paper itep{jack2026twomodes} we showe

The Entropic Signature of Class Speciation in Diffusion Models

ResearchDGX agent

arXiv:2602.09651v2 Announce Type: replace-cross Abstract: Diffusion models do not recover semantic structure uniformly over time. Instead, samples transition from semantic ambiguity to class commitmen

The Ghost Couple: Correlated LLM Name Priors and Their Haunting of the Web and Academic Publishing

Model ReleasesDGX agent

arXiv:2606.02184v1 Announce Type: cross Abstract: These names do not exist. Elena Vasquez and Marcus Chen have appeared as volcano experts, astronauts, thriller protagonists, podcast hosts, and academ

The Lie We Tell: Correcting the Euclidean Fallacy in Vision Language Action Policies via Score Matching on Tangent Space

ResearchDGX agent

arXiv:2606.01847v1 Announce Type: cross Abstract: Diffusion-based Vision-Language-Action policies achieve remarkable success in robotic manipulation, yet commit a fundamental geometric error we term t

The Representation-Rationalizability Tradeoff in Reward Learning

ResearchDGX agent

arXiv:2606.00291v1 Announce Type: cross Abstract: In RLHF, each training example contains a prompt x and two candidate responses y,y', and annotators provide pairwise preferences between these respons

The role of class encoding in neural collapse

SafetyDGX agent

arXiv:2606.00344v1 Announce Type: new Abstract: Neural collapse is a structural property of the last-hidden-layer activations in neural network classification models, when trained beyond a zero classi

Theoretical Analysis of Engression and Reverse Markov Engression

ResearchDGX agent

arXiv:2606.01002v1 Announce Type: cross Abstract: Engression is a recently proposed and effective framework for conditional distribution learning. Its multi-step Reverse Markov extension further impro

TimeBlocks: Foundational and Continual Time-Series Blockbase -- Extended Version

ResearchDGX agent

arXiv:2606.02142v1 Announce Type: new Abstract: The ongoing digitization has led to a proliferation of time-series data streams that monitor a variety of processes, from which valuable insights may be

Tiny Recursive Models for Solving the J2-Perturbed Lambert Problem

Model ReleasesDGX agent

arXiv:2606.00895v1 Announce Type: cross Abstract: This paper presents a fast, recursive neural solver for the J2-perturbed Lambert problem based on Tiny Recursive Models (TRM), termed the TRM-Perturbe

Topology-Aware State Abstraction with Tangle Cores for Markov Decision Processes

ResearchDGX agent

arXiv:2606.00427v1 Announce Type: new Abstract: State abstraction in reinforcement learning is usually formulated as a partition of states based on reward and transition similarity. This excludes a co

Torus Graphs for Large Scale Neural Phase Analysis

Local AiDGX agent

arXiv:2606.00496v1 Announce Type: new Abstract: Oscillatory neural signals such as electroencephalography (EEG) and local field potentials (LFPs) show phase relationships that coordinate communication

Towards Automated Discovery: A Review of Generative Models, Multimodal Learning and Closed-Loop Workflows in Inverse Materials Design

TutorialsDGX agent

arXiv:2606.02507v1 Announce Type: cross Abstract: Inverse materials design is shifting materials discovery from forward prediction to targeted proposal of candidates that satisfy objectives under phys

Towards Optimal Robustness in Learning-Augmented Paging

TutorialsDGX agent

arXiv:2606.01342v1 Announce Type: cross Abstract: Learning-augmented paging has been extensively studied in recent years. A key advantage over naive ML-based approaches is bounded robustness, which gu

Towards Simple and Provable Parameter-Free Adaptive Gradient Methods

Model ReleasesDGX agent

arXiv:2412.19444v2 Announce Type: replace Abstract: Optimization algorithms such as AdaGrad and Adam have significantly advanced the training of deep models by dynamically adjusting the learning rate

Towards Stable, Globally Expressive Graph Representations with Laplacian Eigenvectors

TutorialsDGX agent

arXiv:2410.09737v2 Announce Type: replace Abstract: A popular way to improve the expressive power of graph neural networks (GNNs) is to use Laplacian eigenvectors as additional node features, since th

Tractable Shapley Values and Interactions via Tensor Networks

TutorialsDGX agent

arXiv:2510.22138v3 Announce Type: replace Abstract: We show how to replace the O(2^n) coalition enumeration over n features behind Shapley values and Shapley-style interaction indices with a few-evalu

Training-Free Imitation Learning with Closed-Form Diffusion Policies

SafetyDGX agent

arXiv:2606.01238v1 Announce Type: cross Abstract: While diffusion-based policies have impressive performance and expressivity, their long offline training slows down the data collection and policy dep

Trajectory Data Suffices for Statistically Efficient Policy Evaluation in Fixed-Horizon Offline RL with Linear q^pi-Realizability and Concentrability

SafetyDGX agent

arXiv:2510.03494v2 Announce Type: replace Abstract: We study finite-horizon offline reinforcement learning (RL) with function approximation for both policy evaluation and policy optimization. Prior wo

Tree-Guided Identify-Then-Exploit: A Unified Framework of Best Arm Identification and Regret Minimization for Dueling Bandits

ResearchDGX agent

arXiv:2606.01799v1 Announce Type: new Abstract: We study N-armed stochastic dueling bandits under the Condorcet-winner assumption, where three widely adopted objectives are considered: best-arm identi

Turning Back Without Forgetting: Selective Backward Refinement for Parameter-Efficient Continual Learning

Model ReleasesDGX agent

arXiv:2606.01379v1 Announce Type: new Abstract: While prompt-based parameter-efficient continual learning mitigates catastrophic forgetting by isolating task-specific prompts, this isolation also limi

UME: A Unified Meta-Generalization Framework for Cross-Domain ETA

ApplicationsDGX agent

arXiv:2606.00979v1 Announce Type: new Abstract: Accurate Estimated Time of Arrival (ETA) prediction on checkout page is crucial in instant logistics for enhancing user satisfaction, optimizing dispatc

Uncertainty-Aware Graph Neural Reconstruction of Urban Temperature Fields from Sparse Sensors under Deployment Constraints

ResearchDGX agent

arXiv:2606.02038v1 Announce Type: cross Abstract: Reconstructing spatially continuous daily temperature fields from sparse observations is important for urban climate monitoring and heat-risk analysis

Uncertainty-Calibrated Diffusion for Reliable 3D Molecular Graph Generation

ResearchDGX agent

arXiv:2606.01595v1 Announce Type: new Abstract: Bayesian inference provides a principled framework for modeling epistemic uncertainty in neural networks by treating predictions as distributions rather

UniPinRec: Unifying Generative Retrieval and Ranking at Pinterest Scale

ApplicationsDGX agent

arXiv:2606.00422v1 Announce Type: cross Abstract: Modern recommendation systems predominantly train retrieval and ranking as separate models despite both increasingly relying on large transformers enc

Unlearning Isn't Invisible: Detecting Unlearning Traces in LLMs from Model Outputs

ResearchDGX agent

arXiv:2506.14003v5 Announce Type: replace Abstract: Machine unlearning (MU) for large language models (LLMs), commonly referred to as LLM unlearning, seeks to remove specific undesirable data or knowl

Variance-sensitive Thompson sampling for generalised linear bandits, revisited

ResearchDGX agent

arXiv:2606.00431v1 Announce Type: new Abstract: We prove a variance-sensitive regret bound for Thompson sampling in stochastic generalised linear bandits. The argument assumes a warm-up, after which t

Vegas: Self-Speculative Decoding with Verification-Guided Sparse Attention

ResearchDGX agent

arXiv:2602.07223v2 Announce Type: replace Abstract: Long-context large language model (LLM) inference has become the norm for today's AI applications. However, it is severely bottlenecked by the incre

ViBE: Co-Optimizing Workload Skew and Hardware Variability for MoE Serving

HardwareDGX agent

arXiv:2606.00735v1 Announce Type: cross Abstract: In distributed Mixture-of-Experts (MoE) inference, input-dependent token routing interacts with GPU performance variability to create persistent strag

VMDNet: Temporal Leakage-Free Variational Mode Decomposition for Electricity Demand Forecasting

TutorialsDGX agent

arXiv:2509.15394v3 Announce Type: replace Abstract: Accurate electricity demand forecasting is challenging due to the strong multi-periodicity of real-world demand series, which makes effective modeli

Well-Posed KL-Regularized Control via Wasserstein and Kalman-Wasserstein KL Divergences

ResearchDGX agent

arXiv:2602.02250v2 Announce Type: replace-cross Abstract: Kullback-Leibler (KL) divergence regularization is widely used in reinforcement learning, but it becomes infinite under support mismatch and c

What Cosine Similarity of Label Representations Can and Cannot Tell us

ResearchDGX agent

arXiv:2603.29488v2 Announce Type: replace Abstract: Cosine similarity is often used to measure the similarity of vector representations of neural network models. However, the cosine similarity of repr

When Hard Negatives Hurt: Bridging the Generative-Discriminative Gap in Hard Negative Synthesis for Retrieval

ResearchDGX agent

arXiv:2606.01304v1 Announce Type: new Abstract: Hard negative mining has become the dominant strategy for training retrievers, yet it faces intrinsic limitations: negatives are bounded by corpus avail

When Parallelism Pays Off: Cohesion-Aware Task Partitioning for Multi-Agent Coding

Model ReleasesDGX agent

arXiv:2606.00953v1 Announce Type: new Abstract: Multi-agent Large Language Model (LLM) systems offer a way to decompose complex tasks, such as coding, through parallelization and context isolation. Ho

When Tabular Foundation Models Transfer Across Modalities: A Systematic Evaluation Across 95 Datasets, 7 Modalities, and Two Regimes

TutorialsDGX agent

arXiv:2606.02106v1 Announce Type: new Abstract: We present a single classification pipeline that combines an Equiangular Tight Frame (ETF) preprocessing stage with a tabular foundation model for in-co

Which Leakage Types Matter? A Quantitative Landscape Across 2,047 Benchmark Datasets

Model ReleasesDGX agent

arXiv:2604.04199v2 Announce Type: replace Abstract: Twenty-eight within-subject counterfactual experiments across 2,047 iid tabular datasets, plus a boundary experiment on 129 temporal datasets, measu

Why Are DMD Students Lazy? Understanding the Copying Behavior in Few-Step Distillation

ResearchDGX agent

arXiv:2606.02237v1 Announce Type: new Abstract: Distribution Matching Distillation (DMD) compresses pretrained diffusion models into efficient few-step generators by aligning their noised distribution

WildCat: Near-Linear Attention in Theory and Practice

Model ReleasesDGX agent

arXiv:2602.10056v2 Announce Type: replace Abstract: We introduce WildCat, a high-accuracy, low-cost approach to compressing the attention mechanism in neural networks. While attention is a staple of m

World Models: A Comprehensive Survey of Architectures, Methodologies, Reasoning Paradigms, and Applications

SafetyDGX agent

arXiv:2606.00133v1 Announce Type: new Abstract: World models, internal simulators that learn the structure and dynamics of an environment, have emerged as a central paradigm in the pursuit of artifici

World-Task Factorization for Robot Learning

SafetyDGX agent

arXiv:2606.02027v1 Announce Type: cross Abstract: Robot learning must produce policies that generalize to new combinations of constraints, teammates, and environments. To achieve this, we must structu

WUSH: Near-Optimal Adaptive Transforms for LLM Quantization

Model ReleasesDGX agent

arXiv:2512.00956v3 Announce Type: replace Abstract: Quantizing LLM weights and activations is a standard approach for efficient deployment, but a few extreme outliers can stretch the dynamic range and

1 Jun 2026

A hitchhiker's guide to Poisson gradient estimation

SafetyDGX agent

arXiv:2602.03896v2 Announce Type: replace-cross Abstract: Poisson-distributed latent variable models are widely used in computational neuroscience, but differentiating through discrete stochastic samp

A holomorphic neural network framework for 3D boundary value problems governed by harmonic potentials

ResearchDGX agent

arXiv:2605.31231v1 Announce Type: cross Abstract: We present a neural-network-based framework for the solution of three-dimensional boundary value problems where the solution is expressible in terms o

A Lecture Note on Offline RL and IRL, Part II: Foundations of Inverse Reinforcement Learning and Dynamic Discrete Choice Models

SafetyDGX agent

arXiv:2605.30843v1 Announce Type: new Abstract: In the forward reinforcement-learning problem, the reward is fixed and known; the learner is asked to find a good policy or value function. Here we turn

A Novel Computer Vision Approach for Assessing Fish Responses to Intrusive Objects in Aquaculture

ApplicationsDGX agent

arXiv:2605.30399v1 Announce Type: cross Abstract: The aquaculture industry needs to address several challenges to secure sustainable seafood production that can serve an increasing global demand. One

A Novel Evaluation Metric for Unsupervised Learning in AIS-Based Maritime Anomaly Detection: MADQI

ResearchDGX agent

arXiv:2605.30388v1 Announce Type: new Abstract: This paper introduces a new systematic framework for detecting anomalies in maritime Automatic Identification System (AIS) datasets. These anomalies inc

A Perturbation Approach to Unconstrained Linear Bandits

ResearchDGX agent

arXiv:2603.28201v2 Announce Type: replace Abstract: We revisit the standard perturbation-based approach of Abernethy et al. (2008) in the context of unconstrained Bandit Linear Optimization (uBLO). We

A Tight Theory of Error Feedback Algorithms in Distributed Optimization

AgentsDGX agent

arXiv:2605.31594v1 Announce Type: new Abstract: Communication costs are a major bottleneck in distributed learning and first-order optimization. A common approach to alleviate this issue is to compres

A Unifying View of Anchoring via Operator-Side Tikhonov Regularization

ResearchDGX agent

arXiv:2605.30905v1 Announce Type: cross Abstract: Anchored fixed point and monotone equation methods, including Halpern iteration, extra anchored gradient, and their relatives, add a vanishing pull to

AbstainGNN: Teaching Graph Neural Networks to Abstain for Graph Classification

Model ReleasesDGX agent

arXiv:2605.30786v1 Announce Type: new Abstract: Graph classification is a core task in graph data mining with widespread real-world applications. Recent advances in graph neural networks (GNNs) have l

Accelerated Multiple Wasserstein Gradient Flows for Multi-objective Distributional Optimization

ResearchDGX agent

arXiv:2601.19220v2 Announce Type: replace Abstract: We study multi-objective optimization over probability distributions in Wasserstein space. Recently, Nguyen et al. (2025) introduced Multiple Wasser

Adaptive NAD: Online and Self-adaptive Unsupervised Network Anomaly Detector

Model ReleasesDGX agent

arXiv:2410.22967v5 Announce Type: replace Abstract: The widespread usage of the Internet of Things (IoT) has raised the risks of cyber threats; thus, developing Anomaly Detection Systems (ADSs) that c

Adaptive Physics Transformer with Fused Global-Local Attention for Subsurface Energy Systems

Local AiDGX agent

arXiv:2602.11208v2 Announce Type: replace Abstract: The Earth's subsurface is a cornerstone of modern society, providing essential energy resources like hydrocarbons, geothermal, and minerals while se

Advances and Challenges in Meta-Learning: A Technical Review

ApplicationsDGX agent

arXiv:2307.04722v2 Announce Type: replace Abstract: Meta-learning empowers learning systems with the ability to acquire knowledge from multiple tasks, enabling faster adaptation and generalization to

Aggregation Buffer: Revisiting DropEdge with a New Parameter Block

Model ReleasesDGX agent

arXiv:2505.20840v2 Announce Type: replace Abstract: We revisit DropEdge, a data augmentation technique for GNNs which randomly removes edges to expose diverse graph structures during training. While b

Algorithmic Recourse of In-Context Learning for Tabular Data

ApplicationsDGX agent

arXiv:2605.31272v1 Announce Type: new Abstract: As predictive models are increasingly deployed in high-stakes settings such as credit approval, there is a growing need for post-hoc methods that provid

An Efficient and Scalable Graph Condensation with Structure-Preserving

ApplicationsDGX agent

arXiv:2605.31016v1 Announce Type: new Abstract: Graph condensation (GC) is pivotal for enabling Graph Neural Networks (GNNs) deployment in resource-constrained scenarios by compressing large-scale gra

← Previous
1…109110111112113…243
Next →