AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,557 results
23 Jun 2026

Two Layers of Instability in Causal Estimation

ResearchDGX agent

arXiv:2606.21185v1 Announce Type: cross Abstract: There is a precise sense in which drawing causal inferences from observational data is hard, even when identifiability is assumed. In particular, Robi

UBP2: Uncertainty-Balanced Preference Planning for Efficient Preference-based Reinforcement Learning

Model ReleasesDGX agent

arXiv:2606.19328v2 Announce Type: replace Abstract: Preference-based RL provides an approach to learning reward models from pairwise comparisons of behaviors, bypassing the need for explicit reward de

Ultra-Peripheral Collisions as a Nuclear-Structure Interferometer with Interpretable Multitask Deep Learning

ResearchDGX agent

arXiv:2606.23353v1 Announce Type: cross Abstract: Precise knowledge of nuclear structure is essential across fundamental physics, yet probing these structures is notoriously difficult. To address this


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

UltraQuant: 4-bit KV Caching for Context-Heavy Agents

AgentsDGX agent

arXiv:2606.20474v2 Announce Type: replace Abstract: Context-heavy agents place unusual pressure on the key-value (KV) cache: long prefixes are reused across many short turns, while concurrency determi

Understanding Latent Flow Models for Tabular Data Synthesis: Targets, Paths, and Sampling

ResearchDGX agent

arXiv:2606.20878v1 Announce Type: new Abstract: Synthetic tabular data enables microdata sharing in regulated domains, yet deploying continuous-time generative models requires balancing analytical uti

Understanding Parallel Samplers in Masked Diffusion via Random Walks on Graphs

Model ReleasesDGX agent

arXiv:2606.22976v1 Announce Type: new Abstract: In this paper, we propose using random walks on graphs as a verifiable sandbox to study different parallel sampling strategies in masked diffusion model

UniFFBench: Evaluating Universal Machine Learning Force Fields Against Experimental Measurements

ApplicationsDGX agent

arXiv:2508.05762v3 Announce Type: replace-cross Abstract: Universal machine learning force fields (UMLFFs) promise to revolutionize materials science by enabling rapid atomistic simulations across the

UniRank: Unified Rank Allocation for Low-Rank LLM Compression

Model ReleasesDGX agent

arXiv:2606.21847v1 Announce Type: new Abstract: Low-rank decomposition serves as a promising compression paradigm for large language models, however, rank allocation remains challenging: manual rules

Universal Encoders for Modular Relational Deep Learning

ResearchDGX agent

arXiv:2606.21434v1 Announce Type: new Abstract: Relational Deep Learning (RDL) models multi-tabular databases as temporal heterogeneous graphs for end-to-end representation learning. While RDL is evol

Universal priors: solving empirical Bayes via Bayesian inference and pretraining

ResearchDGX agent

arXiv:2602.15136v2 Announce Type: replace-cross Abstract: We theoretically justify the recent empirical finding of [Teh et al., 2025] that a transformer pretrained on synthetically generated data achi

Unlocking In-Context Learning in Audio-Language Models from Decentralized Medical Audio

ResearchDGX agent

arXiv:2606.23243v1 Announce Type: new Abstract: Clinical audio diagnosis in low-resource settings requires models that identify conditions from minimal examples without large annotated corpora. We pro

Unsupervised Disentanglement Without Compromises : How Functional Orthogonality Enforces Identifiability

ResearchDGX agent

arXiv:2606.21385v1 Announce Type: new Abstract: This paper explores unsupervised disentangled representation learning from a functional perspective. We define latent concepts as factors that influence

Urban Power Grid Topology and Hierarchy Identification from Open Data

ResearchDGX agent

arXiv:2606.21352v1 Announce Type: new Abstract: Understanding the complex topology and hierarchy of urban power grid is crucial for energy prognosis, power flow management, and system resilience analy

Using predictive multiplicity to measure individual performance within the AI Act

SafetyDGX agent

arXiv:2602.11944v2 Announce Type: replace Abstract: When building AI systems for decision support, one often encounters the phenomenon of predictive multiplicity: a single best model does not exist; i

Variance-Tilted Diffusion Models for Diverse Sampling

ResearchDGX agent

arXiv:2606.22239v1 Announce Type: cross Abstract: Diffusion models are typically sampled independently, even when the downstream objective is to obtain a diverse set of candidates. We introduce a vari

VDW-GNNs: Vector diffusion wavelets for geometric graph neural networks

ApplicationsDGX agent

arXiv:2510.01022v3 Announce Type: replace Abstract: We introduce vector diffusion wavelets (VDWs), a novel family of wavelets inspired by the vector diffusion maps algorithm that was introduced to ana

VegSim: A Geospatial World Model for Scenario-Conditioned Vegetation Simulation

ApplicationsDGX agent

arXiv:2606.21961v1 Announce Type: new Abstract: Vegetation monitoring under climate stress requires answering not only how it will evolve given the expected weather, but how it would respond to altern

VeriBound: PAC-Bayesian Generalization Bounds for Process Reward Models Trained with Formal Verification Tools

ResearchDGX agent

arXiv:2606.20740v1 Announce Type: cross Abstract: Process Reward Models (PRMs) provide step-level verification for Large Language Model (LLM) reasoning, yet their training data acquisition remains a b

VRPO: Rethinking Value Modeling for Robust RL under Noisy Supervision in LLM Post-Training

SafetyDGX agent

arXiv:2508.03058v2 Announce Type: replace Abstract: Reinforcement Learning (RL) in real-world environments often suffers from ambiguous or incomplete reward supervision, which undermines policy stabil

WarPGNN: A Parametric Thermal Warpage Analysis Framework with Physics-aware Graph Neural Network

ResearchDGX agent

arXiv:2603.18581v2 Announce Type: replace-cross Abstract: With the advent of system-in-package (SiP) chiplet-based design and heterogeneous 2.5D/3D integration, thermal-induced warpage has become a cr

Weighted Score-Oriented Losses for Temporally Localized Event Prediction

Model ReleasesDGX agent

arXiv:2606.23145v1 Announce Type: new Abstract: Operational event-detection systems are rarely assessed by pointwise accuracy alone. In anomaly detection, changepoint detection, and warning systems, t

What Accuracy and Gradient Cosine Miss: Evaluating Feedback Alignment via Scale Stability, Reference Validity, and Depth Utility

SafetyDGX agent

arXiv:2606.21126v1 Announce Type: new Abstract: Despite the success of deep learning, training deep networks in biologically plausible and hardware-efficient ways remains an open challenge. Feedback a

What Do Lorentz-Equivariant Jet Taggers Learn?

TutorialsDGX agent

arXiv:2606.21790v1 Announce Type: new Abstract: We study what Lorentz-equivariant jet taggers learn internally, using equivariance tests, linear probes and grade ablations across five models including

What Do Neural Networks Learn for TDOA Estimation? A Cross-Architecture Probing Study

TutorialsDGX agent

arXiv:2606.22020v1 Announce Type: cross Abstract: Neural networks outperform classical GCC-PHAT for Time-Difference-of-Arrival (TDOA) estimation in noise and reverberation, yet their internal strategy

What Does a Chemical Language Model Know About Molecules?

TutorialsDGX agent

arXiv:2606.23443v1 Announce Type: new Abstract: Chemical language models (cLMs) are widely assumed to learn surface-level syntactic patterns rather than learning meaningful molecular semantics. Here,

What Shapes Emergent Misalignment? Insights from Training Dynamics, Model Priors, and Data

Local AiDGX agent

arXiv:2606.20814v1 Announce Type: cross Abstract: Emergent misalignment (EM) is a phenomenon in which models generalize with narrow fine-tuning, leading to broad (yet uneven) misalignment across evalu

When AUC 0.998 Is Not Enough: A Candidate Evaluation Protocol for Hidden-State Probes of Indirect Prompt Injection in Multimodal Computer-Use Agents

Model ReleasesDGX agent

arXiv:2606.22864v1 Announce Type: new Abstract: Hidden-state probing -- a linear classifier on a frozen vision-language model's internal activations -- has emerged as an attractive evaluation tool for

When Is an LLM Worth It for Hyperparameter Optimization? A Budget-Matched Study on Tabular Data Finds the Warm-Start Is a Default Configuration, Not the Model

ResearchDGX agent

arXiv:2606.21641v1 Announce Type: new Abstract: Large language models (LLMs) have been proposed as hyperparameter-optimization (HPO) advisors that 'warm-start' search from prior knowledge, proposing s

When Web Agents Finish but Still Fail: Reproducible Triggers and Trace Diagnostics for Parallel Web Exploration

Model ReleasesDGX agent

arXiv:2606.20724v1 Announce Type: cross Abstract: Long-horizon web agents often fail in ways hidden by final-answer evaluation: they may visit useful pages, produce a well-formed answer, and terminate

Where Does the Signal Live? A Web Data Recipe for Medical Encoder Pretraining

Model ReleasesDGX agent

arXiv:2606.22079v1 Announce Type: cross Abstract: Web data curation has been widely studied for decoder Large Language Model (LLM) pretraining. Encoders for dense-terminology domains such as medicine,

Who Owns the AI Recommendation? A Multi-Industry Empirical Map of Brand Category Ownership Across Large Language Models

Model ReleasesDGX agent

arXiv:2606.23057v1 Announce Type: cross Abstract: Large language models now mediate how buyers discover products and services, making the competitive structure of AI-generated recommendations a strate

WiSP: A Working-Set View of Mixture-of-Experts Serving on Extremely Low-Resource Hardware

Local AiDGX agent

arXiv:2606.21868v1 Announce Type: new Abstract: Modern Mixture-of-Experts (MoE) models place most of their parameters in expert layers, yet only a small fraction of those experts are used for any toke

Words as Difference Makers: How Large Language Models Determine Causal Structure in Text

TutorialsDGX agent

arXiv:2606.22430v1 Announce Type: cross Abstract: Because large language models (LLMs) are impressively successful in predicting text, it appears that they must have access to a 'world model' represen

XConv: Low-memory stochastic backpropagation for convolutional layers

Local AiDGX agent

arXiv:2106.06998v4 Announce Type: replace Abstract: Training convolutional neural networks at scale demands substantial memory, largely because intermediate activations must be stored for backpropagat

Zero-order Parameter-free Optimization for LMO-based Methods: Novel Approach for Efficient Fine-tuning

Model ReleasesDGX agent

arXiv:2606.14970v2 Announce Type: replace Abstract: Fine-tuning large language models (LLMs) has become a central application of modern optimization, enabling pretrained models to adapt to diverse dow

11 Jun 2026

A Data-Centric Framework for Detecting and Correcting Corrupted Labels

ApplicationsDGX agent

arXiv:2606.11699v1 Announce Type: new Abstract: The performance of machine learning and deep learning models largely depends on the quality of the training data. However, the quality of the real-world

A Judge-Aware Ranking Framework for Evaluating Large Language Models without Ground Truth

ResearchDGX agent

arXiv:2601.21817v3 Announce Type: replace-cross Abstract: Evaluating large language models (LLMs) on open-ended tasks without ground-truth labels is increasingly done via the LLM-as-a-judge paradigm.

A Multi-Modal Sensor Fusion Instrument for Measuring Regional Human Mobility: The Distributed Human Data Engine (DHDE)

Local AiDGX agent

arXiv:2603.21639v2 Announce Type: replace-cross Abstract: Accurately estimating human mobility in peripheral regional economies presents a fundamental measurement challenge: physical ground-truth sens

A prior-free blind detection of information leakage from model predictions

ResearchDGX agent

arXiv:2606.11267v1 Announce Type: new Abstract: Data leakage -- contamination of a model with information unavailable at baseline -- is the dominant reproducibility failure in machine-learning-based s

A Riemannian Approach to Low-Rank Optimal Transport

ResearchDGX agent

arXiv:2606.12120v1 Announce Type: new Abstract: Low-rank optimal transport (OT) mitigates the quadratic scaling of classical solvers, yet existing approaches rely heavily on first-order mirror-descent

A theory of learning data statistics in diffusion models, from easy to hard

SafetyDGX agent

arXiv:2603.12901v2 Announce Type: replace-cross Abstract: While diffusion models have emerged as a powerful class of generative models, their learning dynamics remain poorly understood. We address thi

Accurate and Resource-Efficient Federated Continual Learning

TutorialsDGX agent

arXiv:2606.11480v1 Announce Type: new Abstract: Federated continual learning (FCL) must learn from distributed task streams under limited resources, such as communication, computation, memory, and lab

Adjoint Method versus Physics-Informed Neural Networks in PDE-Constrained Inverse Problems

ResearchDGX agent

arXiv:2606.12337v1 Announce Type: cross Abstract: Inverse problems governed by partial differential equations (PDEs) are central to computational mechanics and are commonly solved by adjoint-based opt

Analytic Bijections for Smooth and Interpretable Normalizing Flows

ResearchDGX agent

arXiv:2601.10774v2 Announce Type: replace Abstract: A key challenge in normalizing flows is finding expressive invertible scalar bijections. Existing approaches face trade-offs: affine transformations

Anatomy of Post-Training: Using Interpretability to Characterize Data and Shape the Learning Signal

TutorialsDGX agent

arXiv:2606.12360v1 Announce Type: new Abstract: Language-model post-training is the main stage at which model behavior is shaped, yet it still largely involves optimization of scalar rewards that summ

Annealed Entropic Allocation for Ranking and Selection

Model ReleasesDGX agent

arXiv:2606.11347v1 Announce Type: cross Abstract: We propose Annealed Entropic Allocation, an annealed weighted soft-min framework for sequential budget allocation in ranking and selection. The centra

APEX: A Network-Native Time-Series Foundation Model for Forecasting and Anomaly Detection for Wireless Edge Operations

Model ReleasesDGX agent

arXiv:2606.11553v1 Announce Type: new Abstract: Generic time-series foundation models transfer poorly to wireless network telemetry whose signals are bursty, zero-inflated, and coupled across protocol

AsFT: Anchoring Safety During LLM Fine-Tuning Within Narrow Safety Basin

Model ReleasesDGX agent

arXiv:2506.08473v4 Announce Type: replace Abstract: Fine-tuning large language models (LLMs) improves performance but introduces critical safety vulnerabilities: even minimal harmful data can severely

Attention by Synchronization in Coupled Oscillator Networks

ResearchDGX agent

arXiv:2606.12059v1 Announce Type: new Abstract: We address transformer attention on energy-constrained physical substrates. Softmax attention requires exponentiation and global reduction, operations w

Bergson: An Open Source Library for Data Attribution

ResearchDGX agent

arXiv:2606.11660v1 Announce Type: new Abstract: Data attribution is a promising field in interpretability that aims to explain model behavior through the influence of its training data, with applicati

Bernstein-Schur Kernels: Random Features by Sketched Modulation and Radial Randomization

ResearchDGX agent

arXiv:2606.11255v1 Announce Type: new Abstract: Bernstein--Schur kernels are products of a finite-feature kernel (one with an explicit finite-dimensional feature map) and a completely monotone shift-i

Beyond the Golden Teacher: Enhancing Graph Learning through LLM-GNN Co-teaching

TutorialsDGX agent

arXiv:2606.11583v1 Announce Type: new Abstract: Text-attributed graphs (TAGs) underlie real-world applications such as citation networks, social media, and e-commerce. Few-shot graph learning on TAGs

Bootstrapped Monitoring: Leveraging Transparent Reasoning to Oversee Stronger AI Agents

AgentsDGX agent

arXiv:2606.11998v1 Announce Type: new Abstract: Trusted monitoring is a cornerstone of AI control. However, as frontier models grow more capable, the increasing capabilities gap between trusted and un

Calibrating Decision Robustness via Inverse Conformal Risk Control

ResearchDGX agent

arXiv:2510.07750v3 Announce Type: replace-cross Abstract: Robust optimization safeguards decisions against uncertainty by optimizing against worst-case scenarios, yet their effectiveness hinges on a p

Capacity-Constrained Online Convex Optimization with Delayed Feedback

ResearchDGX agent

arXiv:2606.11711v1 Announce Type: new Abstract: Online learning with delayed feedback typically assumes that the learner can track all pending rounds until their feedback arrives. In practice, trackin

CaReTS: A Multi-Task Framework Unifying Classification and Regression for Time Series Forecasting

ApplicationsDGX agent

arXiv:2511.09789v2 Announce Type: replace Abstract: Recent advances in deep forecasting models have achieved remarkable performance, yet most approaches still struggle to provide both accurate predict

Categorical Robustness Assessment for Machine Learning based Network Intrusion Detection Systems

ApplicationsDGX agent

arXiv:2606.12075v1 Announce Type: cross Abstract: Network Intrusion Detection Systems (NIDS) heavily utlize Machine Learning (ML) but ML models can be manipulated via adversarial attacks. These attack

Composing Linear Layers from Irreducibles

ResearchDGX agent

arXiv:2507.11688v4 Announce Type: replace Abstract: Contemporary large models often exhibit behaviors suggesting the presence of low-level primitives that compose into modules with richer functionalit

Conformal Bayes under Label Shift: Post-Hoc Calibration vs. In-Training Adaptation

Model ReleasesDGX agent

arXiv:2606.11865v1 Announce Type: cross Abstract: Conformal Bayes combines Bayesian posterior predictives with conformal calibration to produce prediction sets that are both statistically valid and ge

Counterexample Guided Learning in the Large using Reasoning Agents

AgentsDGX agent

arXiv:2606.11521v1 Announce Type: new Abstract: LLMs and LLM agents should improve when given feedback, but identifying when they are able to do so is difficult: feedback is heterogeneous, domain-spec

← Previous
1…8485868788…243
Next →