AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent
84,433Total entries
1Added by human
84,432Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,557 results
26 Jun 2026

Quantization in Federated Learning: Methods, Challenges and Future Directions

Local AiDGX agent

arXiv:2606.26822v1 Announce Type: new Abstract: Federated Learning (FL) has become a foundational paradigm for privacy-preserving distributed intelligence, yet its scalability remains fundamentally co

Reasoning Quality Emerges Early: Data Curation for Reasoning Models

ResearchDGX agent

arXiv:2606.26797v1 Announce Type: new Abstract: Supervised fine-tuning (SFT) on a small, high-quality set of long reasoning traces is an effective approach for eliciting strong reasoning capabilities

RecallRisk-BERT: A Multi-Task Framework for Post-Report Medical Device Recall Triage

SafetyDGX agent

arXiv:2606.27174v1 Announce Type: new Abstract: Medical device recalls are a critical regulatory mechanism for protecting patient safety. The growing volume of FDA recall records presents challenges i


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Recovering Governing Equations from Solution Data: Identifiability Bounds for Linear and Nonlinear ODEs

ResearchDGX agent

arXiv:2606.27285v1 Announce Type: new Abstract: Learning governing equations from observed solution data is a fundamental challenge in scientific machine learning ite{bruntonDiscoveringGoverningEquati

Reinforcement Fine-Tuning of Flow-Matching Policies for Vision-Language-Action Models

Model ReleasesDGX agent

arXiv:2510.09976v2 Announce Type: replace Abstract: Vision-Language-Action (VLA) models such as OpenVLA, Octo, and pi_0 have shown strong generalization by leveraging large-scale demonstrations, yet t

Reinforcement Learning Enables Autonomous Microrobot Navigation and Intervention in Simulated Blood Capillaries

AgentsDGX agent

arXiv:2606.26154v1 Announce Type: cross Abstract: Autonomous microrobots navigating biological vasculature could enable targeted drug delivery and thrombolysis, yet training control policies for reali

Reinforcement Learning without Ground-Truth Solutions can Improve LLMs

Model ReleasesDGX agent

arXiv:2606.27369v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) for training LLMs typically rely on ground-truth answers to assign rewards, limiting their applica

Representation Costs in Data Science: Foundations and the Quasi-Banach Spaces of Deep Neural Networks

Model ReleasesDGX agent

arXiv:2606.14954v3 Announce Type: replace-cross Abstract: We develop a general framework for analyzing representation costs of parametric data-fitting methods through their parameter-space regularizer

Revisiting Action Factorization for Complex Action Spaces

AgentsDGX agent

arXiv:2606.26574v1 Announce Type: new Abstract: Many real-world control problems involve hybrid discrete-continuous action spaces. For example, steering and signaling in autonomous driving, and aiming

Ribbon: Scalable Approximation and Robust Uncertainty Quantification

Model ReleasesDGX agent

arXiv:2606.27269v1 Announce Type: cross Abstract: Reliably quantifying predictive uncertainty is difficult for complex, high-dimensional, or misspecified models. Both fully Bayesian and bootstrap resa

RolloutPipe: Overlapping Pipelined Rollout and Training in Disaggregated On-Policy LLM Reinforcement Learning

SafetyDGX agent

arXiv:2606.26997v1 Announce Type: cross Abstract: Large language model (LLM) post-training for reasoning increasingly relies on reinforcement learning with verifiable rewards (RLVR), where models lear

RSPC: A Benchmark for Modeling Stress and Psychiatric Conditions in Digitally Mediated Relationships using Psychiatrist Annotations

Model ReleasesDGX agent

arXiv:2606.27247v1 Announce Type: new Abstract: In NLP, mental health conditions are often modeled as isolated phenomena, without interpersonal context. We use Reddit posts about long-distance relatio

Running the Gauntlet: Re-evaluating the Capabilities of Agents Beyond Familiar Environments

Model ReleasesDGX agent

arXiv:2606.14397v2 Announce Type: replace Abstract: As agentic systems continue to evolve and are widely deployed in real-world scenarios, there is a growing demand to faithfully evaluate their capabi

Sample-efficient Transfer Reinforcement Learning via Adaptive Reward Shaping and Policy-Ratio Reweighting Strategy

SafetyDGX agent

arXiv:2606.26527v1 Announce Type: new Abstract: Transfer learning improves policy learning efficiency by reusing knowledge from source tasks, providing a feasible paradigm for safe and efficient auton

Scalable Message-Passing Quantum Graph Neural Networks in the Weisfeiler-Leman Hierarchy

ResearchDGX agent

arXiv:2606.26873v1 Announce Type: cross Abstract: Graphs provide a natural language for relational data in chemistry, biology and optimisation. Graph neural networks (GNNs) have driven much of the rec

Scoring Is Not Enough: Addressing Gaps in Utility-fairness Trade-offs for Ranking

SafetyDGX agent

arXiv:2606.26369v1 Announce Type: cross Abstract: Scoring functions are used to represent the relevance of individual documents. In modern information retrieval or recommendation systems, they are oft

Signature filtering: a lightweight enhancement for statistical watermark detection in large language models

Model ReleasesDGX agent

arXiv:2606.18430v2 Announce Type: replace Abstract: Statistical watermarks help organizations attribute large language model (LLM) outputs, yet existing detectors often struggle when watermark signals

Sketched Linear Contrastive Learning: Approximation, Optimization, and Statistical Scaling

SafetyDGX agent

arXiv:2606.26617v1 Announce Type: new Abstract: Scaling laws describe how learning performance varies with model size, data size, and compute. While recent theoretical work has established scaling law

State-Specific Respiratory Signatures for Affective and Stress Recognition: Interpretable Respiratory Markers, Autocorrelation Lags, and Compact CNN Models

ResearchDGX agent

arXiv:2606.26723v1 Announce Type: cross Abstract: Respiratory activity is a direct and interpretable physiological channel for wearable stress and affective-state recognition, yet many studies emphasi

Stochastic Gradient Optimization with Model-Assisted Sampling

Model ReleasesDGX agent

arXiv:2606.27171v1 Announce Type: new Abstract: This work addresses the problem of variance in stochastic gradient estimation for machine learning optimization. Deep learning relies on mini-batch meth

Symplectic Neural Networks for learning Generalized Hamiltonians

ResearchDGX agent

arXiv:2606.27029v1 Announce Type: new Abstract: Hamiltonian Neural Networks (HNNs) integrate physical priors into neural models by learning a system's Hamiltonian, improving generalization and sample

Target-Aware Bandit Allocation for Scalable Surrogate Optimization in Chemical Space

ApplicationsDGX agent

arXiv:2606.26657v1 Announce Type: new Abstract: Identifying high-utility candidates from massive discrete spaces under expensive evaluations is a recurring challenge across the sciences, with structur

The Role of Input Dimensionality in the Emergence and Targeted Control of Adversarial Examples

ResearchDGX agent

arXiv:2606.26207v1 Announce Type: cross Abstract: Several theoretical works have tried to explain the adversarial vulnerability of deep neural networks through properties of high-dimensional geometry.

Theory of the Frequency Principle for General Deep Neural Networks

TutorialsDGX agent

arXiv:1906.09235v3 Announce Type: replace Abstract: Along with fruitful applications of Deep Neural Networks (DNNs) to realistic problems, recently, some empirical studies of DNNs reported a universal

Theory-Scale Auto-Formalization of Logics for Computer Science

Model ReleasesDGX agent

arXiv:2606.26525v1 Announce Type: new Abstract: Auto-formalization is critical for scalable formal verification, but existing progress largely focuses on isolated statements, while theory-scale auto-f

Topology-Informed Neural Networks for Flood Detection in Optical and Synthetic Aperture Radar Imagery

SafetyDGX agent

arXiv:2606.26204v1 Announce Type: new Abstract: Floods frequently impact regions around the world. Rapid and accurate flood detection is crucial for emergency response and timely mitigation of human a

Training-Free Generation of Protein Sequences from Small Family Alignments via Stochastic Attention

SafetyDGX agent

arXiv:2603.14717v2 Announce Type: replace Abstract: Generating novel protein sequences that respect a family's statistical constraints typically requires training deep generative models on thousands t

Transformer-Based Classification of Bacterial Raman Spectra with LOOCV

ResearchDGX agent

arXiv:2606.27096v1 Announce Type: new Abstract: Transformer-based models have recently attracted increasing attention for Raman spectral classification. In this study, a transformer-based approach was

Uncertainty quantification via conformal prediction in data assimilation

ResearchDGX agent

arXiv:2606.27001v1 Announce Type: new Abstract: Quantifying the evolution of uncertainty is critical to both probabilistic forecasting and data assimilation in numerical weather prediction. In this st

Use What You Know: Causal Foundation Models with Partial Graphs

ResearchDGX agent

arXiv:2602.14972v2 Announce Type: replace Abstract: Estimating causal quantities traditionally relies on bespoke estimators tailored to specific assumptions. Recently proposed Causal Foundation Models

What Survives When You Compress a Recursive Reasoner for the Edge?

ResearchDGX agent

arXiv:2606.26488v1 Announce Type: new Abstract: Recursive reasoning models can solve complex structured tasks with only a few million parameters by repeatedly updating a latent state. Deploying these

When are likely answers right? On Sequence Probability and Correctness in LLMs

ResearchDGX agent

arXiv:2606.27359v1 Announce Type: cross Abstract: Many decoding methods for large language models can be understood as shifting probability mass toward outputs that are more likely under the model, ei

When Does Quality-Aware Multimodal Fusion Matter? A Leakage-Safe Diagnostic for Decision-Level Dependence

ResearchDGX agent

arXiv:2606.26473v1 Announce Type: new Abstract: Many multimodal systems estimate the reliability of each modality and weight their contributions to the final prediction. However, it remains unclear wh

When to Write and When to Suppress: Route-Specialized Dual Adapters for Memory-Assisted Knowledge Editing

Model ReleasesDGX agent

arXiv:2606.14668v3 Announce Type: replace Abstract: Knowledge editing systems must update selected facts while preserving nearby but irrelevant behavior. This paper studies this problem in a memory-as

25 Jun 2026

A 3D-Printable Dataset for Fair Testing and Comparisons of Tactile Sensors

Model ReleasesDGX agent

arXiv:2606.25886v1 Announce Type: cross Abstract: Existing texture datasets for tactile sensing primarily consist of sensor readings from a specific sensor interacting with available surfaces/objects

A Bregman Perspective on Classification and Regression Trees

ResearchDGX agent

arXiv:2606.13984v2 Announce Type: replace-cross Abstract: Classification and Regression Trees (CART) constitute one of the most influential paradigms in statistical learning. Although a variety of imp

A Flow-rate-conserving CNN-based Domain Decomposition Method for Blood Flow Simulations

Model ReleasesDGX agent

arXiv:2509.15900v2 Announce Type: replace-cross Abstract: This work aims to predict blood flow with non-Newtonian viscosity in stenosed arteries using convolutional neural network (CNN) surrogate mode

A Framework for Directed Hypergraph Signal Processing via tensor t-SVD

ResearchDGX agent

arXiv:2606.25112v1 Announce Type: new Abstract: We introduce Directed Hypergraph Signal Processing (DHGSP), a unified framework that extends graph signal processing to accommodate both higher-order (p

A Geometry-Aware Efficient Algorithm for Compositional Entropic Risk Minimization

ResearchDGX agent

arXiv:2602.02877v2 Announce Type: replace Abstract: This paper studies optimization for a family of problems termed extbf{compositional entropic risk minimization}, in which each data's loss is formul

A Hybrid CNN-LSTM Intrusion Detection Framework for Cybersecurity in Smart Renewable Energy Grids

Model ReleasesDGX agent

arXiv:2606.25200v1 Announce Type: new Abstract: The accelerated digitalization of renewable energy smart grids through IoT sensors, AMI, and SCADA systems has significantly expanded the attack surface

A Probabilistic Framework for LLM-Based Model Discovery

AgentsDGX agent

arXiv:2602.18266v2 Announce Type: replace Abstract: Automated methods for discovering mechanistic simulator models from observational data offer a promising path toward accelerating scientific progres

A Single Stepsize Suffices for Unprojected Linear TD(0): Simultaneous Robust and Fast Rates via Polyak--Ruppert Averaging

Model ReleasesDGX agent

arXiv:2606.24981v1 Announce Type: new Abstract: We study linear TD(0) under Markovian sampling, where data are generated along a single trajectory. We provide high-probability guarantees for a plain u

A Spectral Phase Diagram for Binary Few-Shot Classification: Intrinsic Dimensionality, Geometric Saturation, and Representational Diagnosis

ResearchDGX agent

arXiv:2606.24903v1 Announce Type: new Abstract: Deciding when to stop collecting labeled examples is a fundamental but undertheorized problem in applied machine learning. The saturation index S(K) = o

A Zeroth-Order Deep Learning Method for Fully Nonlinear Parabolic Partial Differential Equations with Unknown Coefficients

SafetyDGX agent

arXiv:2606.24999v1 Announce Type: new Abstract: High-dimensional partial differential equations (PDEs) with unknown coefficients arise widely in scientific machine learning, including continuous-time

ACT-JEPA: Novel Joint-Embedding Predictive Architecture for Efficient Policy Representation Learning

SafetyDGX agent

arXiv:2501.14622v5 Announce Type: replace Abstract: Learning efficient representations for decision-making policies is a challenge in imitation learning (IL). Current IL methods require expert demonst

Adapt Only When It Pays: Budgeted Decision-Loss Priority for Delayed Online Time-Series Adaptation

TutorialsDGX agent

arXiv:2606.25068v1 Announce Type: new Abstract: Online time-series forecasters receive labels only after horizon-dependent delays, while every adaptation step spends limited compute. We study when an

Adaptive Cumulative Mass Calibration with Conformal Prediction

ResearchDGX agent

arXiv:2505.15437v3 Announce Type: replace-cross Abstract: Reliable probability estimates by classifiers are essential in high-risk applications. In practice, however, predicted probabilities are often

Adaptive Joint Compression and Synchronisation in Federated Split Learning for IoT Rainfall Prediction

SafetyDGX agent

arXiv:2606.25003v1 Announce Type: new Abstract: Federated split learning (FSL) enables collaborative training across bandwidth-constrained IoT devices, but repeated activation and gradient exchange cr

AeroCast: Probabilistic 3D Trajectory Prediction for Non-Cooperative Aerial Obstacles via Transformer-MDN Architecture

AgentsDGX agent

arXiv:2606.25122v1 Announce Type: cross Abstract: Autonomous aerial vehicles operating in shared airspace must predict the future positions of non-cooperative obstacles to plan evasive maneuvers befor

Agentic evolution of physically constrained foundation models

Model ReleasesDGX agent

arXiv:2606.25532v1 Announce Type: cross Abstract: Artificial intelligence increasingly drives automated scientific discovery, yet contemporary generalist agents lack physical grounding, frequently hal

An Analysis of Posterior Collapse, Parameterization and Initialization in Variational Deep Gaussian Processes

ResearchDGX agent

arXiv:2606.25882v1 Announce Type: new Abstract: DGPs are probabilistic models with remarkable prediction performance that concatenate GPs across several layers. Exact inference in DGPs is intractable,

Are Tabular Foundation Models Robust to Realistic Query Distribution Shifts in Microbiome Data?

Model ReleasesDGX agent

arXiv:2606.24995v1 Announce Type: new Abstract: Tabular foundation models (TFMs) achieve strong performance on microbiome abundance data, yet their robustness under realistic distribution shift remain

ATMA: Length-Invariant Language Modeling via Polar Attention and Gated-Delta Compression Memory

Local AiDGX agent

arXiv:2606.25156v1 Announce Type: new Abstract: Modern large language models based on softmax scaled-dot-product attention are constrained by their training sequence length: as the key-value sequence

Auto-Configured Explainable Graph Neural Networks for Multi-Site Pollution Prediction

ResearchDGX agent

arXiv:2606.24978v1 Announce Type: new Abstract: Accurate particulate matter (PM) prediction is crucial for mitigating air pollution. Graph Neural Networks (GNNs) effectively model spatiotemporal depen

Auto-exploration for online reinforcement learning

Model ReleasesDGX agent

arXiv:2512.06244v2 Announce Type: replace Abstract: The exploration-exploitation dilemma in reinforcement learning (RL) is a fundamental challenge to efficient RL algorithms. Existing algorithms for f

Beyond Defensive Reporting: Machine Learning for Active Anti-Money Laundering Control in Insurance

ApplicationsDGX agent

arXiv:2606.16663v2 Announce Type: replace Abstract: Money laundering through insurance claims poses a threat to insurers both through fraudulent payouts and reputational and regulatory risk. Despite t

Beyond One-Size-Fits-All: Diagnosis-Driven Online Reinforcement Learning with Offline Priors

Model ReleasesDGX agent

arXiv:2606.25527v1 Announce Type: new Abstract: Online reinforcement learning (RL) agents increasingly depend on knowledge acquired offline to achieve practical efficiency. Originally studied in offli

Bias-Controlled Primal-Dual Natural Actor-Critic: Optimal Rates for Constrained Multi-Objective Average-Reward RL

SafetyDGX agent

arXiv:2606.25012v1 Announce Type: new Abstract: Many reinforcement learning (RL) problems in the infinite-horizon average-reward setting require optimizing multiple conflicting objectives while satisf

Bias Fitting to Mitigate Length Bias of Reward Model in RLHF

SafetyDGX agent

arXiv:2505.12843v2 Announce Type: replace Abstract: Reinforcement Learning from Human Feedback (RLHF) relies on reward models to align large language models with human preferences. However, RLHF often

Black-Box Assisted Regression: Phase Transitions and Minimax Optimality

ResearchDGX agent

arXiv:2606.25743v1 Announce Type: new Abstract: Foundation models are often used as fixed black-box predictors for downstream tasks with limited labeled data, but their predictions may be biased and u

← Previous
1…6869707172…243
Next →