AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,619Total entries
1Added by human
84,618Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,557 results
Safety

Approximate Equivariance via Projection-based Regularisation

DGX agent

arXiv:2601.05028v2 Announce Type: replace Abstract: Equivariance is a powerful inductive bias in neural networks, improving generalisation and physical consistency. Recently, however, non-equivariant

safetyarxiv-cs-lg
27 May 2026
Model Releases
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

ARBITER: Reasoning Trajectory Basins and Majority Vote Failures in Test-Time Sampling

DGX agent

arXiv:2605.26172v1 Announce Type: new Abstract: When language models use test-time sampling, they generate multiple reasoning trajectories and select an answer by majority vote. We show that these tra

model-releasesarxiv-cs-lg
27 May 2026
Agents

ATOM: Instantiating Budget-Controllable Multi-Agent Collaboration via Nucleus-Electron Hierarchy

DGX agent

arXiv:2605.26178v1 Announce Type: cross Abstract: Large Language Model (LLM)-based multi-agent systems rely on optimized collaboration topologies to balance performance and communication costs. Howeve

agentsarxiv-cs-lg
27 May 2026
Applications

Balancing Plasticity and Stability with Fast and Slow Successor Features

DGX agent

arXiv:2605.26357v1 Announce Type: new Abstract: A hallmark of intelligence is the ability to adapt in non-stationary environments, yet deep Reinforcement Learning (RL) agents often struggle in such se

applicationsarxiv-cs-lg
27 May 2026
Safety

BASIS: Batchwise Advantage Estimation from Single-Rollout Information Sharing for LLM Reasoning

DGX agent

arXiv:2605.27293v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards has become a standard recipe for improving the reasoning abilities of large language models. Existing alg

safetyarxiv-cs-lg
27 May 2026
Model Releases

Benchmark Leakage Trap: Can We Trust LLM-based Recommendation?

DGX agent

arXiv:2602.13626v3 Announce Type: replace Abstract: The expanding integration of Large Language Models (LLMs) into recommender systems poses critical challenges to evaluation reliability. This paper i

model-releasesarxiv-cs-lg
27 May 2026
Research

Beyond Differences: Doubly Robust Meta-Learners for Ratio-Based Treatment Effects

DGX agent

arXiv:2605.26288v1 Announce Type: cross Abstract: When treatment effects are naturally expressed as ratios -- as in medicine, pricing, and marketing -- the ratio-based CATE au(x) = E[Y|W=1,X=x] / E[Y|

researcharxiv-cs-lg
27 May 2026
Model Releases

Beyond Holistic Models: Systematic Component-level Benchmarking of Deep Multivariate Time-Series Forecasting

DGX agent

arXiv:2605.26562v1 Announce Type: new Abstract: While previous research in multivariate time series forecasting has focused on developing complex holistic models, this work advocates for a shift towar

model-releasesarxiv-cs-lg
27 May 2026
Research

BUILD with Precision: Bottom-Up Inference of Linear DAGs

DGX agent

arXiv:2512.16111v2 Announce Type: replace Abstract: Learning the structure of directed acyclic graphs (DAGs) from observational data is a central problem in causal discovery, statistical signal proces

researcharxiv-cs-lg
27 May 2026
Local Ai

CART Random Forests as Sequential Allocation over Random Opportunity Sets: A Stochastic-Control Theory of Ensemble Risk

DGX agent

arXiv:2605.26675v1 Announce Type: cross Abstract: CART random forests are among the most widely used modern predictive methods, with well-documented empirical success. Yet, at the mechanistic level, t

local-aiarxiv-cs-lg
27 May 2026
Model Releases

Causal Representation Learning for Generalisable Recommendation

DGX agent

arXiv:2605.27043v1 Announce Type: cross Abstract: Predictive models trained on observational data often fail to generalise to the distributions they encounter when deployed, especially when the traini

model-releasesarxiv-cs-lg
27 May 2026
Tutorials

Causal Risk Minimization for High-Dimensional Treatments

DGX agent

arXiv:2605.27281v1 Announce Type: new Abstract: Predicting the effect of interventions with many possible variations, e.g., therapeutic content that affects mental health outcomes or an earnings call

tutorialsarxiv-cs-lg
27 May 2026
Model Releases

CktGen: Automated Analog Circuit Design with Generative Artificial Intelligence

DGX agent

arXiv:2410.00995v3 Announce Type: replace Abstract: The automatic synthesis of analog circuits presents significant challenges. Most existing approaches formulate the problem as a single-objective opt

model-releasesarxiv-cs-lg
27 May 2026
Research

Classification and detection of multiple UAVs using rational Gaussian wavelet neural networks

DGX agent

arXiv:2605.26310v1 Announce Type: new Abstract: The detection of unmanned aerial vehicles (UAVs) is important for the protection of civilian and military infrastructure. In this paper we propose a cos

researcharxiv-cs-lg
27 May 2026
Model Releases

CleanSurvival: Automated data preprocessing for time-to-event models using reinforcement learning

DGX agent

arXiv:2502.03946v5 Announce Type: replace Abstract: Data preprocessing is often paid little attention in machine learning, despite its potentially significant impact on model performance. While automa

model-releasesarxiv-cs-lg
27 May 2026
Safety

CompassDPO: Dynamics-Controlled Direct Preference Optimization for Robust Safety Alignment

DGX agent

arXiv:2603.07211v2 Announce Type: replace Abstract: Direct Preference Optimization (DPO) has become a standard framework for safety alignment, but its reliance on pairwise preference updates makes tra

safetyarxiv-cs-lg
27 May 2026
Safety

Constrained Bayesian Experimental Design via Online Planning

DGX agent

arXiv:2605.26990v1 Announce Type: cross Abstract: Bayesian experimental design (BED) is a principled framework for data-efficient design of sequential experiments. However, existing BED methods are un

safetyarxiv-cs-lg
27 May 2026
Safety

Constrained Meta Reinforcement Learning with Provable Test-Time Safety

DGX agent

arXiv:2601.21845v2 Announce Type: replace Abstract: Meta reinforcement learning (RL) allows agents to leverage experience across a distribution of tasks on which the agent can train at will, enabling

safetyarxiv-cs-lg
27 May 2026
Research

Convergence of Spectral Descent for Non-smooth Optimization

DGX agent

arXiv:2605.26977v1 Announce Type: new Abstract: The Muon optimizer has recently demonstrated remarkable empirical success in training large language models. However, the theoretical understanding of i

researcharxiv-cs-lg
27 May 2026
Tutorials

Corrected Samplers for Discrete Flow Models

DGX agent

arXiv:2601.22519v2 Announce Type: replace-cross Abstract: Discrete flow models (DFMs) have been proposed to learn the data distribution on finite state space, offering a flexible framework as an alter

tutorialsarxiv-cs-lg
27 May 2026
Safety

Cost of Structural Learning Under Censored Feedback: A Threshold-Bandit Approach

DGX agent

arXiv:2605.27076v1 Announce Type: cross Abstract: In many multi-agent applications, tasks yield rewards only when executed by a coalition meeting an unknown size threshold; otherwise, feedback is full

safetyarxiv-cs-lg
27 May 2026
Safety

Cross-Receiver Generalization for RF Fingerprint Identification via Feature Disentanglement and Adversarial Training

DGX agent

arXiv:2510.09405v2 Announce Type: replace Abstract: Radio frequency fingerprint identification (RFFI) is a key technique for wireless network security, leveraging intrinsic hardware imperfections to e

safetyarxiv-cs-lg
27 May 2026
Research

Data-driven sparse identification of governing PDEs via knockoff filters and multi-criteria trade-offs

DGX agent

arXiv:2605.26631v1 Announce Type: cross Abstract: We propose KO-PDE-IDENT, a data-driven framework for identifying parsimonious partial differential equations (PDEs) with false discovery rate (FDR) co

researcharxiv-cs-lg
27 May 2026
Research

Deep Learning-based Algebraic Reynolds Stress Closures for RANS Simulations of Turbulent Flows

DGX agent

arXiv:2605.26358v1 Announce Type: cross Abstract: Turbulence is ubiquitous in engineering and science, yet direct simulation is prohibitively expensive. The Reynolds-averaged Navier-Stokes (RANS) equa

researcharxiv-cs-lg
27 May 2026
Research

Detectability in Diversity: Improved Canary Crafting for Privacy Auditing in One Run

DGX agent

arXiv:2605.27292v1 Announce Type: new Abstract: Privacy auditing aims to empirically assess privacy leakage in machine learning models using membership inference attacks (MIAs), and to derive lower bo

researcharxiv-cs-lg
27 May 2026
Model Releases

Device Context Protocol: A Compact, Safety-First Architecture for LLM-Driven Control of Constrained Devices

DGX agent

arXiv:2605.26159v1 Announce Type: cross Abstract: Large language models are increasingly used as orchestrators of external tools via the Model Context Protocol (MCP), but MCP is built for software ser

model-releasesarxiv-cs-lg
27 May 2026
Agents

Disentangled Representation Learning through Unsupervised Symmetry Group Discovery

DGX agent

arXiv:2603.11790v3 Announce Type: replace Abstract: Symmetry-based disentangled representation learning leverages the group structure of environment transformations to uncover the latent factors of va

agentsarxiv-cs-lg
27 May 2026
Local Ai

Distributed Control of Network Systems in the Space of Stabilizing Graph Neural Network Policies

DGX agent

arXiv:2512.18540v2 Announce Type: replace-cross Abstract: We study distributed control of networked systems through reinforcement learning, where neural policies must be simultaneously scalable, expre

local-aiarxiv-cs-lg
27 May 2026
Model Releases

Distribution-Aware Conformal Prediction: A Framework for generating efficient prediction intervals for time series

DGX agent

arXiv:2605.26569v1 Announce Type: new Abstract: We present Distribution-aware Conformal Prediction (DCP), a unified framework integrating probabilistic predictors like Monte Carlo dropout, deep ensemb

model-releasesarxiv-cs-lg
27 May 2026
Applications

Dynamic Link Prediction with Temporally Enhanced Signed Graph Neural Networks

DGX agent

arXiv:2605.26290v1 Announce Type: new Abstract: Temporal signed networks (TSNs) model the time evolution of cooperative and adversarial relationships that arise in applications such as social media an

applicationsarxiv-cs-lg
27 May 2026
Model Releases

ECHO-2: A Large-Scale Distributed Rollout Framework for Cost-Efficient Reinforcement Learning

DGX agent

arXiv:2602.02192v5 Announce Type: replace Abstract: Reinforcement learning (RL) is a critical stage in post-training large language models (LLMs), involving repeated interaction between rollout genera

model-releasesarxiv-cs-lg
27 May 2026
Research

Efficient Learning of Mesh-Based Physical Simulation with BSMS-GNN

DGX agent

arXiv:2210.02573v5 Announce Type: replace Abstract: Learning the physical simulation on large-scale meshes with flat Graph Neural Networks (GNNs) and stacking Message Passings (MPs) is challenging due

researcharxiv-cs-lg
27 May 2026
Model Releases

Efficient Prediction of SO(3)-Equivariant Hamiltonian Matrices via SO(2) Local Frames

DGX agent

arXiv:2506.09398v3 Announce Type: replace Abstract: We consider the task of predicting Hamiltonian matrices to accelerate electronic structure calculations, which plays an important role in physics, c

model-releasesarxiv-cs-lg
27 May 2026
Research

Error Analysis of Discrete Flow with Generator Matching

DGX agent

arXiv:2509.21906v3 Announce Type: replace-cross Abstract: Discrete flow models offer a powerful framework for learning distributions over discrete state spaces and have demonstrated superior performan

researcharxiv-cs-lg
27 May 2026
Research

Explainable Comparison of Feature-Based and Deep Learning Models for TROPOMI Methane Plume Screening

DGX agent

arXiv:2605.27236v1 Announce Type: new Abstract: Continuous and global detection of large methane emissions is a crucial step for global warming mitigation. Satellite observations, such as from S5P/TRO

researcharxiv-cs-lg
27 May 2026
Research

Exploring the robustness of TractOracle methods in RL-based tractography

DGX agent

arXiv:2507.11486v2 Announce Type: replace Abstract: Tractography algorithms leverage diffusion MRI to reconstruct the fibrous architecture of the brain's white matter. Among machine learning approache

researcharxiv-cs-lg
27 May 2026
Model Releases

Extra-Merge: Tracing the Rank-1 Subspace of Model Merging in Language Model Pre-Training

DGX agent

arXiv:2605.26484v1 Announce Type: new Abstract: Model merging has emerged as a lightweight paradigm for enhancing Large Language Models (LLMs), yet its underlying mechanisms remain poorly understood.

model-releasesarxiv-cs-lg
27 May 2026
Safety

Flow Matching Policy Optimization with Mirror Descent and Entropy Constraints

DGX agent

arXiv:2603.17685v3 Announce Type: replace Abstract: Balancing policy expressiveness with the exploration-exploitation trade-off is a core challenge in online Reinforcement Learning (RL). While Stochas

safetyarxiv-cs-lg
27 May 2026
Research

FluxNet: Learning Capacity-Constrained Local Transport Operators for Conservative and Bounded PDE Surrogates

DGX agent

arXiv:2602.01941v2 Announce Type: replace-cross Abstract: Autoregressive learning of time-stepping operators provides an effective approach to data-driven partial differential equation (PDE) simulatio

researcharxiv-cs-lg
27 May 2026
Safety

FM-fMRI: Event Conditioned Flow Matching for Rest-to-Task fMRI Time-Series Synthesis

DGX agent

arXiv:2605.26423v1 Announce Type: new Abstract: Task-based fMRI provides a direct readout of task-evoked neural dynamics, but it is expensive and difficult to acquire at scale, motivating rest-to-task

safetyarxiv-cs-lg
27 May 2026
Model Releases

Focal Reward: Balanced Reinforcement Learning under Rubric-Based Rewards

DGX agent

arXiv:2605.26579v1 Announce Type: new Abstract: The open-ended generation in LLMs usually requires multi-dimensional rubrics to adequately assess quality and guide the improvement of reinforcement lea

model-releasesarxiv-cs-lg
27 May 2026
Research

From Privacy to Generalization: Linear Max-Information Bounds for DP-SGD

DGX agent

arXiv:2605.26222v1 Announce Type: new Abstract: Understanding the relationship between generalization and privacy remains a central challenge in modern machine learning theory, particularly for deep n

researcharxiv-cs-lg
27 May 2026
Research

From Scores to Gibbs Correctors: Accelerating Uniform-Rate Discrete Diffusion Models

DGX agent

arXiv:2605.27352v1 Announce Type: new Abstract: Discrete diffusion models have achieved strong empirical performance in text and other symbolic domains, but, especially for uniform-rate models, they o

researcharxiv-cs-lg
27 May 2026
Applications

Function-Valued Causal Influence in Nonlinear Time Series

DGX agent

arXiv:2605.26408v1 Announce Type: new Abstract: Causal discovery in time series is increasingly performed using nonlinear machine-learning models, yet the resulting causal relationships are almost alw

applicationsarxiv-cs-lg
27 May 2026
Applications

Gaussian Process-based learning with new MCMC-based implementation of Wishart prior on correlation matrix

DGX agent

arXiv:2605.27093v1 Announce Type: cross Abstract: In probabilstic supervised learning of an input-output relationship - as a sample function of a Gaussian Process (GP) - priors are typically specified

applicationsarxiv-cs-lg
27 May 2026
Safety

Generalist Graph Anomaly Detection via Prototype-Based Distillation

DGX agent

arXiv:2605.26857v1 Announce Type: new Abstract: Driven by the pressing demand for graph anomaly detection (GAD) in high-stakes domains, the generalist GAD paradigm, which trains a single detector tran

safetyarxiv-cs-lg
27 May 2026
Research

Generating realistic global precipitation fields from modelled atmospheric circulation

DGX agent

arXiv:2504.00307v2 Announce Type: replace Abstract: Improving the representation of precipitation in Earth system models (ESMs) is critical for assessing the impacts of climate change and especially o

researcharxiv-cs-lg
27 May 2026
Research

Greening AI Inference with Accuracy and Latency-aware User Incentives

DGX agent

arXiv:2605.27309v1 Announce Type: new Abstract: The widespread use of AI services has raised concerns for its environmental sustainability, towards which recent studies have identified carbon emission

researcharxiv-cs-lg
27 May 2026
← Previous
1…152153154155156…304
Next →