AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent
84,433Total entries
1Added by human
84,432Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,557 results
12 May 2026

PINS: Proximal Iterations with Sparse Newton and Sinkhorn for Optimal Transport

Model ReleasesDGX agent

arXiv:2502.03749v2 Announce Type: replace Abstract: Optimal transport (OT) is a widely used tool in machine learning, but computing high-accuracy solutions for large instances remains costly. Entropic

Plan2Cleanse: Test-Time Backdoor Defense via Monte-Carlo Planning in Deep Reinforcement Learning

SafetyDGX agent

arXiv:2605.09638v1 Announce Type: new Abstract: Ensuring the security of reinforcement learning (RL) models is critical, particularly when they are trained by third parties and deployed in real-world

PMCTS: Particle Monte Carlo Tree Search for Principled Parallelized Inference Time Scaling

SafetyDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2605.08982v1 Announce Type: new Abstract: Monte Carlo Tree Search (MCTS) is a widely used approach for policy improvement through search with increasing popularity for real world applications. D

PoHAR: Understanding Hyperlocal Human Activities with Pollution Sensor Networks

Local AiDGX agent

arXiv:2605.09434v1 Announce Type: cross Abstract: Low-cost air quality sensors are becoming ubiquitous in our daily lives as public awareness of air pollution continues to grow, and people take measur

Positional LSH: Binary Block Matrix Approximation for Attention with Linear Biases

SafetyDGX agent

arXiv:2605.09472v1 Announce Type: new Abstract: Positional encoding in transformers is commonly implemented through positional embeddings, attention masks, or bias terms, but formal connections betwee

Practical Scaling Laws: Converting Compute into Performance in a Data-Constrained World

ResearchDGX agent

arXiv:2605.09189v1 Announce Type: new Abstract: The scaling laws guiding modern model training were calibrated for a single regime: data-rich, single-epoch pretraining. The dominant such scaling law f

Predicting Large Model Test Losses with a Noisy Quadratic System

ResearchDGX agent

arXiv:2605.09154v1 Announce Type: new Abstract: We introduce a predictive model that estimates the pre-training loss of large models from model size (N), batch size (B) and number of weight updates (K

Predicting Plasticity in Deep Continual Learning: A Theoretical Perspective

ResearchDGX agent

arXiv:2605.09044v1 Announce Type: new Abstract: Deep continual learning requires models to adapt to new tasks without retraining from scratch. However, neural networks can lose their ability to adapt

Predictive Radiomics for Evaluation of Cancer Immune SignaturE in Glioblastoma: the PRECISE-GBM study

ResearchDGX agent

arXiv:2605.10278v1 Announce Type: new Abstract: Background: Radiogenomics allows identification of radiological biomarkers for genomic phenotypes. In glioblastoma, these biomarkers could potentially c

Preventing Prompt Injection with Type-Directed Privilege Separation

AgentsDGX agent

arXiv:2509.25926v2 Announce Type: replace-cross Abstract: Modern language models have enabled the development of agentic systems that achieve strong performance on reasoning-intensive tasks. Unfortuna

Price of Quality: Sufficient Conditions for Sparse Recovery using Mixed-Quality Data

ResearchDGX agent

arXiv:2605.10713v1 Announce Type: cross Abstract: We study sparse recovery when observations come from mixed-quality sources: a small collection of high-quality measurements with small noise variance

PRIM: Meta-Learned Bayesian Root Cause Analysis

Model ReleasesDGX agent

arXiv:2605.08786v1 Announce Type: new Abstract: Root cause analysis (RCA) in complex systems is challenging due to error propagation across multiple variables, the need for structural causal knowledge

Priority-Driven Control and Communication in Decentralized Multi-Agent Systems via Reinforcement Learning

Model ReleasesDGX agent

arXiv:2605.10482v1 Announce Type: cross Abstract: Event-triggered control provides a mechanism for avoiding excessive use of constrained communication bandwidth in networked multi-agent systems. Howev

PRISM: Fast Online LLM Serving via Scheduling-Memory Co-design

AgentsDGX agent

arXiv:2605.08581v1 Announce Type: new Abstract: Modern online large language model (LLM) services, such as Retrieval-Augmented Generation (RAG) and agent systems, increasingly expose two prominent cha

Privacy Auditing Synthetic Data Release through Local Likelihood Attacks

Model ReleasesDGX agent

arXiv:2508.21146v2 Announce Type: replace Abstract: Auditing the privacy leakage of synthetic data is an important but unresolved problem. Existing privacy auditing frameworks for synthetic data rely

Privacy-Preserving Distributed Learning in IoT Systems: A Unified Threat Model and Evaluation Framework

Local AiDGX agent

arXiv:2605.09232v1 Announce Type: cross Abstract: The increasing deployment of Internet-of-Things (IoT) devices has accelerated the use of distributed learning frameworks, where data remains local whi

Private Vertical Federated Inference for Time-Series

ResearchDGX agent

arXiv:2605.08343v1 Announce Type: new Abstract: Institutions may benefit from collaborative inference on time-series data. In settings where privacy is necessary, multi-party computation (MPC) is a st

ProcVLM: Learning Procedure-Grounded Progress Rewards for Robotic Manipulation

SafetyDGX agent

arXiv:2605.08774v1 Announce Type: cross Abstract: Long-horizon robotic manipulation requires dense feedback that reflects how a task advances through its procedural stages, not merely whether the fina

Projection-Free Functional Constrained Optimization for Risk Aversion and Sparsity Control

ResearchDGX agent

arXiv:2210.05108v2 Announce Type: replace-cross Abstract: We study projection-free methods for functional constrained optimization with convex or smooth nonconvex objectives. Such problems arise in ap

Prophecy: Inferring Formal Properties from Neuron Activations

ResearchDGX agent

arXiv:2509.21677v2 Announce Type: replace Abstract: We present Prophecy, a tool for automatically inferring formal properties of feed-forward neural networks. Prophecy is based on the observation that

QT-Net: Rethinking Evaluation of AI Models in Atomic Chemical Space

ResearchDGX agent

arXiv:2605.10458v1 Announce Type: new Abstract: Atomic properties such as partial charges or multipoles encode chemically meaningful information that can inform downstream molecular property predictio

Quantifying Concentration Phenomena of Mean-Field Transformers in the Low-Temperature Regime

Model ReleasesDGX agent

arXiv:2605.10931v1 Announce Type: cross Abstract: Transformers with self-attention modules as their core components have become an integral architecture in modern large language and foundation models.

Quantile-Coupled Flow Matching for Distributional Reinforcement Learning

SafetyDGX agent

arXiv:2605.08515v1 Announce Type: new Abstract: Unlike standard expected-return Reinforcement Learning (RL), Distributional RL (DRL) models the full return distribution, making it better-suited for un

Quantitative Clustering in Mean-Field Transformer Models

ResearchDGX agent

arXiv:2504.14697v3 Announce Type: replace Abstract: The evolution of tokens through deep transformer models can be modeled as an interacting particle system that has been shown to exhibit an asymptoti

Quantitative Error Feedback for Quantization Noise Reduction of Filtering over Graphs

ResearchDGX agent

arXiv:2506.01404v2 Announce Type: replace Abstract: This paper introduces an innovative error feedback framework designed to mitigate quantization noise in distributed graph filtering, where communica

Quantitative Local Convergence of Mean-Field Stein Variational Gradient Flow

ResearchDGX agent

arXiv:2605.09456v1 Announce Type: cross Abstract: Stein Variational Gradient Descent (SVGD) is a deterministic interacting-particle method for sampling from a target probability measure given access t

Quantitative Sobolev Approximation Bounds for Neural Operators with Empirical Validation on Burgers Equation

Model ReleasesDGX agent

arXiv:2605.08170v1 Announce Type: new Abstract: Neural operators have emerged as a powerful tool for learning mappings between infinite-dimensional function spaces. However, their approximation proper

Quantum Circuit Simulation of Compartmental Drug Dynamics: Leveraging Variational Algorithms for Nonlinear Mixed-Effects Population Pharmacokinetics

Model ReleasesDGX agent

arXiv:2605.09691v1 Announce Type: new Abstract: Population pharmacokinetic/pharmacodynamic (PK/PD) modeling traditionally relies on classical ordinary differential equations to simulate drug dynamics.

Quantum Transfer Learning Shows Improved Robustness in Low-Data Regimes

ResearchDGX agent

arXiv:2605.09118v1 Announce Type: cross Abstract: Transfer learning under limited data is a challenging setting, where models must adapt to new tasks with minimal supervision. Prior work has primarily

Quasi-Linear ICA for Motor Unit Decomposition during Dynamic Contractions

Model ReleasesDGX agent

arXiv:2406.19581v2 Announce Type: replace-cross Abstract: Decomposing surface electromyography (EMG) into the spike trains of individual motor neurons is a long-standing inverse problem and a key step

RareCP: Regime-Aware Retrieval for Efficient Conformal Prediction

Model ReleasesDGX agent

arXiv:2605.08857v1 Announce Type: new Abstract: Recent advances in uncertainty quantification for time series forecasting show that conformal prediction can provide reliable prediction intervals, yet

Reconfigurable Computing Challenge: Real-Time Graph Neural Networks for Online Event Selection in Big Science

ResearchDGX agent

arXiv:2605.10612v1 Announce Type: cross Abstract: Graph neural networks are increasingly adopted in trigger systems for collider experiments, where strict latency and throughput constraints render dep

Reflective Prompted Policy Optimization: Trajectory-Grounded Revision and Salience Bias

SafetyDGX agent

arXiv:2605.08315v1 Announce Type: new Abstract: Existing LLM-based policy optimizers see only scalar rewards: that a policy scored 0.45, but not whether the agent got stuck in a loop, fell into a hole

Region Seeding via Pre-Activation Regularization: A Geometric View of Piecewise Affine Neural Networks

ResearchDGX agent

arXiv:2605.06300v2 Announce Type: replace Abstract: Deep networks with continuous piecewise affine activations induce polyhedral partitions of the input space, making the number of realized affine reg

Regret Analysis of Guided Diffusion for Black-Box Optimization over Structured Inputs

ResearchDGX agent

arXiv:2605.10385v1 Announce Type: cross Abstract: Guided-diffusion black-box optimization (BO) has shown strong empirical performance on structured design problems such as molecules and crystals, but

Regret Minimization in Bilateral Trade With Perturbed Markets

ResearchDGX agent

arXiv:2605.10475v1 Announce Type: cross Abstract: We address the problem of maximizing Gain from Trade (GFT) in repeated buyer-seller exchanges subject to global budget balance constraints. While this

Reinforcement learning for inverse structural design and rapid laser cutting of kirigami prototypes

SafetyDGX agent

arXiv:2605.08098v1 Announce Type: new Abstract: Kirigami is an increasingly useful fabrication method to produce shape-programmable metamaterial structures. However, inverse design remains difficult b

Reinforcement Learning Measurement Model

Model ReleasesDGX agent

arXiv:2605.09305v1 Announce Type: cross Abstract: Interactive assessments generate sequential process data that are not well handled by conventional item response models. Existing MDP-based measuremen

Relational reasoning and inductive bias in transformers and large language models

SafetyDGX agent

arXiv:2506.04289v3 Announce Type: replace Abstract: Transformer-based models have demonstrated remarkable reasoning abilities, but the mechanisms underlying relational reasoning remain poorly understo

RelBench v2: A Large-Scale Benchmark and Repository for Relational Data

Model ReleasesDGX agent

arXiv:2602.12606v2 Announce Type: replace Abstract: Relational deep learning (RDL) has emerged as a powerful paradigm for learning directly on relational databases by modeling entities and their relat

RelFlexformer: Efficient Attention 3D-Transformers for Integrable Relative Positional Encodings

ResearchDGX agent

arXiv:2605.10706v1 Announce Type: new Abstract: We present a new class of efficient attention mechanisms applying universal 3D Relative Positional Encoding (RPE) methods given by arbitrary integrable

Reliable LLM-Based Edge-Cloud-Expert Cascades for Telecom Knowledge Systems

Model ReleasesDGX agent

arXiv:2512.20012v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are emerging as key enablers of automation in domains such as telecommunications, assisting with tasks including

ReLibra: Routing-Replay-Guided Load Balancing for MoE Training in Reinforcement Learning

ResearchDGX agent

arXiv:2605.08639v1 Announce Type: new Abstract: Load imbalance is a long-standing challenge in Mixture-of-Experts (MoE) training and is exacerbated in reinforcement learning (RL) for LLMs, where hot e

Remember to Forget: Gated Adaptive Positional Encoding

SafetyDGX agent

arXiv:2605.10414v1 Announce Type: new Abstract: Rotary Positional Encoding (RoPE) is widely used in modern large language models. However, when sequences are extended beyond the range seen during trai

Rennala MVR: Improved Time Complexity for Parallel Stochastic Optimization via Momentum-Based Variance Reduction

Model ReleasesDGX agent

arXiv:2605.08871v1 Announce Type: cross Abstract: Large-scale machine learning models are trained on clusters of machines that exhibit heterogeneous performance due to hardware variability, network de

Representative Action Selection for Large Action Space Bandit Families

ResearchDGX agent

arXiv:2505.18269v5 Announce Type: replace Abstract: We study the problem of selecting a subset from a large action space shared by a family of bandits. In many natural situations, while the nominal se

Rethinking Ratio-Based Trust Regions for Policy Optimization in Multi-Agent Reinforcement Learning

SafetyDGX agent

arXiv:2605.09212v1 Announce Type: new Abstract: Centralized training with decentralized execution (CTDE) is a standard framework for cooperative multi-agent policy-gradient reinforcement learning, all

Rethinking the Global Knowledge of CLIP in Training-Free Open-Vocabulary Semantic Segmentation

TutorialsDGX agent

arXiv:2502.06818v3 Announce Type: replace Abstract: Recent works modify CLIP to perform open-vocabulary semantic segmentation in a training-free manner (TF-OVSS). In vanilla CLIP, patch-wise image rep

Retrieval Mechanisms Surpass Long-Context Scaling in Time Series Forecasting

Model ReleasesDGX agent

arXiv:2605.08217v1 Announce Type: new Abstract: Time Series Foundation Models (TSFMs) have borrowed the long context paradigm from natural language processing under the premise that feeding more histo

Revisiting Policy Gradients for Restricted Policy Classes: Escaping Myopic Local Optima with k-step Policy Gradients

SafetyDGX agent

arXiv:2605.10909v1 Announce Type: new Abstract: This work revisits standard policy gradient methods used on restricted policy classes, which are known to get stuck in suboptimal critical points. We id

Reward-Conditioned Reinforcement Learning

SafetyDGX agent

arXiv:2603.05066v2 Announce Type: replace Abstract: Single-task RL agents are typically trained under a fixed reward function, which limits their robustness to reward misspecification and their abilit

RIR-Former: Coordinate-Guided Transformer for Continuous Reconstruction of Room Impulse Responses

ApplicationsDGX agent

arXiv:2602.01861v3 Announce Type: replace-cross Abstract: Room impulse responses (RIRs) are essential for many acoustic signal processing tasks, yet measuring them densely across space is often imprac

Robust Remote Reinforcement Learning over Unreliable Communication Channels using Homomorphic State Encoding

AgentsDGX agent

arXiv:2508.07722v2 Announce Type: replace Abstract: Traditional Reinforcement Learning (RL) frameworks generally assume that the agent perceives the state of the underlying Markov process instantaneou

Robust Server Defense Against Unreliable Clients in One-Shot Fair Collaborative Machine Learning

Model ReleasesDGX agent

arXiv:2605.08616v1 Announce Type: new Abstract: Collaborative machine learning (CML) enables multiple clients to train a global model jointly in a data-distributed setting. To address data privacy and

Robust Spectral Watermark for Synthetic Tabular Data

Model ReleasesDGX agent

arXiv:2511.21600v2 Announce Type: replace-cross Abstract: The rise of generative AI has enabled the production of high-fidelity synthetic tabular data across fields such as healthcare, finance, and pu

Root Cause Analysis of Measurement and Mechanistic Anomalies

ApplicationsDGX agent

arXiv:2601.23026v2 Announce Type: replace Abstract: Root cause analysis of anomalies aims to identify how and why a sample deviates from the normal process. Existing methods primarily focus on telling

RubiConv -- Efficient Boundary-Respecting Convolutions

ApplicationsDGX agent

arXiv:2605.08451v1 Announce Type: new Abstract: Convolutional architectures have emerged as powerful alternatives to Transformers for sequence modeling. The primary advantage is that they offer improv

RubricRefine: Improving Tool-Use Agent Reliability with Training-Free Pre-Execution Refinement

Model ReleasesDGX agent

arXiv:2605.09730v1 Announce Type: new Abstract: Iterative self-refinement is a popular inference-time reliability technique, but its effectiveness in code-mode tool use depends heavily on the structur

SACHI: Structured Agent Coordination via Holistic Information Integration in Multi-Agent Reinforcement Learning

Model ReleasesDGX agent

arXiv:2605.08391v1 Announce Type: new Abstract: Cooperative multi-agent reinforcement learning agents that act on partial local observations face a fundamental information bottleneck: the knowledge ne

SAFA-SNN: Sparsity-Aware On-Device Few-Shot Class-Incremental Learning with Fast-Adaptive Structure of Spiking Neural Network

Model ReleasesDGX agent

arXiv:2510.03648v2 Announce Type: replace Abstract: Continuous learning of novel classes is crucial for edge devices to preserve data privacy and maintain reliable performance in dynamic environments.

← Previous
1…173174175176177…243
Next →