AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,557 results
23 Jun 2026

Revealing the Pitfalls and Re-Evaluating the Advancement of Heterophilic Graph Learning

Model ReleasesDGX agent

arXiv:2409.05755v3 Announce Type: replace Abstract: Over the past decade, Graph Neural Networks (GNNs) have achieved great success on machine learning tasks with relational data. However, recent studi

Revisiting OmniAnomaly for Anomaly Detection: performance metrics and comparison with PCA-based models

ResearchDGX agent

arXiv:2603.18985v2 Announce Type: replace-cross Abstract: Deep learning models have become the dominant approach for multivariate time series anomaly detection (MTSAD), often reporting substantial per

Revisiting the Neural Tangent Kernel: the role of large width and depth

Model ReleasesDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2511.07272v2 Announce Type: replace Abstract: Overparameterized fully-connected neural networks have been shown to behave like kernel models when trained with gradient descent, assuming standard

Reward-free Pretraining for Reinforcement Learning via Occupancy Coverage Maximization

ResearchDGX agent

arXiv:2606.21271v1 Announce Type: new Abstract: Sparse rewards pose a central challenge in reinforcement learning, since agents receive no informative signal until they reach their goal. Intrinsic-rew

Right Knowledge, Wrong Answer: Test-Time Steering for Temporal Fact Conflicts in Open-Weight Language Models

Model ReleasesDGX agent

arXiv:2606.20959v1 Announce Type: new Abstract: Large language models can store both outdated facts and newer superseding facts in their parameters, but standard prompting may still elicit the outdate

RIZZ: Routing Interactions to Near Zero-Interference Zones for Continual Adaptation of Black-Box Agents

Local AiDGX agent

arXiv:2606.20638v1 Announce Type: cross Abstract: Large language models are increasingly deployed as long-lived agents that must adapt across users, tasks, domains, modalities, and feedback regimes wi

RLM-Cascade: Response-Level Speculative Decoding for Cost-Efficient LLM API Serving

Model ReleasesDGX agent

arXiv:2606.22840v1 Announce Type: new Abstract: We present RLM-Cascade, a proxy-layer system that applies speculative decoding at the response level to reduce LLM API costs without requiring model arc

Robust and Differentially Private Principal Component Analysis

ResearchDGX agent

arXiv:2507.15232v2 Announce Type: replace-cross Abstract: Recent advances have sparked significant interest in the development of privacy-preserving Principal Component Analysis (PCA). However, many e

Robust Auto-associative Memory via Convolutional Restricted Hopfield Networks

ResearchDGX agent

arXiv:2606.20666v1 Announce Type: cross Abstract: Associative memory models play a fundamental role in pattern retrieval, but their performance often degrades under adversarial perturbations and sever

Robust Diffusion Models via Divergence-Induced Weighted Denoising

ResearchDGX agent

arXiv:2606.22521v1 Announce Type: cross Abstract: We show that replacing the standard MSE denoising loss in diffusion models with a nonlinear transformation induced by an f-divergence yields a simple

Robustness Cannot be Reduced to Regularization: Studying Adversarial Training Beyond the Linear Case

ResearchDGX agent

arXiv:2606.21488v1 Announce Type: new Abstract: The vulnerability of ML models to adversarial examples has recently emerged as a major concern. While adversarial training is one of the most effective

RocketPFN: Accurate Time Series Classification via In-Context Learning

ResearchDGX agent

arXiv:2606.21786v1 Announce Type: new Abstract: We introduce RocketPFN, a training-free pipeline for time series classification that combines random convolutional feature extraction (Rocket) with in-c

RouteJudge: An Open Platform for Reproducible and Preference-Aware LLM Routing

Model ReleasesDGX agent

arXiv:2606.18774v2 Announce Type: replace Abstract: We present RouteJudge, an online pairwise preference evaluation framework for LLM routing systems, with a public platform available at https://route

Sakana Fugu Technical Report

AgentsDGX agent

arXiv:2606.21228v1 Announce Type: new Abstract: The capabilities of frontier Large Language Models (LLMs) continue to advance, with different providers increasingly specializing in distinct domains. T

SamatNext v0.2-B: An Exploratory Study of RMS-Normalized Hybrid Decoders for Curriculum Retention in Small Code Models

Model ReleasesDGX agent

arXiv:2606.22248v1 Announce Type: new Abstract: Standard autoregressive Transformer decoders can often exhibit substantial forgetting under sequential fine-tuning on shifting curriculum distributions.

Scalable Bayesian Additive Models for Stellar Flare Detection via Amortized Gaussian Process Inference and Hidden Markov Models

TutorialsDGX agent

arXiv:2606.22601v1 Announce Type: cross Abstract: Gaussian Processes (GPs) are a powerful tool for Bayesian time-series modeling, yet their cubic computational cost remains a severe barrier for applic

Scalable Maximum Entropy Reinforcement Learning for Diffusion Policies via Adjoint Matching

ResearchDGX agent

arXiv:2606.22630v1 Announce Type: new Abstract: Diffusion policies have recently emerged as a powerful paradigm for representing complex action distributions in reinforcement learning (RL). However, t

Scalable Physics-Inspired Transformers for Spin Glasses

HardwareDGX agent

arXiv:2606.22984v1 Announce Type: cross Abstract: Efficient sampling of the Boltzmann distribution in frustrated spin glasses is central to statistical mechanics and combinatorial optimization. Despit

Scaling Linear Mode Connectivity and Merging to Billion Parameter Pretrained Transformers

Model ReleasesDGX agent

arXiv:2606.23607v1 Announce Type: new Abstract: Linear mode connectivity (LMC) provides a promising foundation for understanding and merging independently trained neural networks, but existing methods

SCENIC: Semantic-Conditioned Edge-Aware Neural Framework for Structured IoT Command Generation

HardwareDGX agent

arXiv:2606.22296v1 Announce Type: new Abstract: Edge Internet of Things (IoT) agents are often constrained by memory capacity, privacy requirements, communication latency, and recurring inference cost

Scheduling Thoughts: Learning the Order of Thought in Diffusion Language Models

SafetyDGX agent

arXiv:2606.23567v1 Announce Type: new Abstract: Masked diffusion language models decode by iteratively unmasking tokens, where the unmasking order defines an 'order of thought' that strongly influence

scLLM-DSC: LLM-Knowledge Enhanced Cross-Modal Deep Structural Clustering for Single-Cell RNA Sequencing

SafetyDGX agent

arXiv:2606.13007v2 Announce Type: replace Abstract: Clustering is fundamental to scRNA-seq analysis, serving as a cornerstone for identifying cell populations and resolving tissue heterogeneity. Howev

Sea-Scan: High-Accuracy, ML-based Dark Vessel Detection and Localisation via Weakly Supervised DAS Monitoring

ResearchDGX agent

arXiv:2606.21326v1 Announce Type: cross Abstract: We present an ML-based vessel detection and localization system, trained with weak supervision from imperfect AIS labels, that achieves a 97.8% detect

Select-to-Act: Hierarchical Reinforcement Learning via Adaptive Language Guidance

Model ReleasesDGX agent

arXiv:2606.22350v1 Announce Type: new Abstract: Reinforcement Learning (RL) has been widely applied to sequential decision-making, yet it often suffers from poor sample efficiency due to costly intera

Selective Ensemble Based on Preference-Directed Multi-Objective Bandits

TutorialsDGX agent

arXiv:2606.21929v1 Announce Type: new Abstract: Selective ensemble for modern machine learning systems requires choosing promising model candidates under limited evaluation budgets, while downstream t

Selective Time Series Forecasting via Metalearning

ResearchDGX agent

arXiv:2606.23448v1 Announce Type: new Abstract: Deep learning methods have achieved state-of-the-art in time series forecasting, yet their accuracy varies considerably across samples, as some instance

Self-Evolution for Multi-Turn Tool-Calling Agents via Divergence-Point Preference Learning

Model ReleasesDGX agent

arXiv:2606.23112v1 Announce Type: new Abstract: Multi-turn tool-using agents must coordinate long-horizon tool sequences while tracking dialogue state and policy constraints. Existing approaches often

Self-Improvement Can Self-Regress: The Rise-and-Collapse Failure Mode of LLM Self-Training

Model ReleasesDGX agent

arXiv:2606.21090v1 Announce Type: cross Abstract: Self-improvement can self-regress. In REINFORCE post-training for code, a model can quickly improve on its optimized metric and then collapse within t

Sequential Minimal Optimization Algorithm for One-Class Support Vector Machines With Privileged Information

Model ReleasesDGX agent

arXiv:2606.22210v1 Announce Type: new Abstract: One of the powerful techniques in data modeling is accounting for features that are available at the training stage, but are not available when the trai

Set-based v.s. Distribution-based Representations of Epistemic Uncertainty: A Comparative Study

Model ReleasesDGX agent

arXiv:2602.22747v2 Announce Type: replace Abstract: Epistemic uncertainty in neural networks is commonly modeled using two second-order paradigms: distribution-based representations, which rely on pos

SFT Overtraining Predicts Rank Inversion via Entropy Collapse Under RLVR

Model ReleasesDGX agent

arXiv:2606.18487v2 Announce Type: replace Abstract: The standard heuristic of selecting the SFT checkpoint with the highest pass@1 for GRPO can fail when SFT compresses the rollout distribution. For b

Short-Term Electricity Demand Forecasting for New England Using a Hybrid Transformer-XGBoost Framework with Weather, Calendar, and COVID-19 Indicators

ResearchDGX agent

arXiv:2606.20918v1 Announce Type: new Abstract: Accurate short-term electricity demand forecasting is critical for reliable power system operation, energy market planning, and infrastructure optimizat

Signed Evidence Flow: Conflict-Aware and Stability-Calibrated Data Analysis

ApplicationsDGX agent

arXiv:2606.21875v1 Announce Type: cross Abstract: Modern data analysis usually gives a prediction without showing whether the evidence behind it is clear, conflicting, or stable. Two cases can have th

SignVLA: Real-Time Sign Language-Guided Robotic Manipulation via Attention LSTM and Vision-Language-Action Models

SafetyDGX agent

arXiv:2606.20857v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models enable robots to execute manipulation tasks from natural-language instructions grounded in visual observations. Ho

Sim2O: Efficient Offline-to-Online MARL via Joint Action Composition

AgentsDGX agent

arXiv:2606.21085v1 Announce Type: new Abstract: Offline-to-online adaptation serves as a pivotal paradigm for mitigating the prohibitive cost of online exploration by bootstrapping reinforcement learn

Simplex-Constrained Sparse Bagging: Transitioning from Uniform Priors to Sparse Posteriors in Ensemble Learning

Local AiDGX agent

arXiv:2606.13589v2 Announce Type: replace Abstract: We present Simplex-Constrained Sparse Bagging (SCSB), a mathematically rigorous framework for post-training compression and probability calibration

Simulation-Free Estimation of Traffic Flows from Sparse Count Data

ResearchDGX agent

arXiv:2606.23536v1 Announce Type: new Abstract: We propose a method for estimating time-varying traffic flow patterns from sparse aggregated vehicle counts. The method partitions the study area into s

Skill Coverage: A Test Adequacy Metric for Agent Skills

Model ReleasesDGX agent

arXiv:2606.20659v1 Announce Type: cross Abstract: Agent skills encode reusable procedural knowledge that guides large language model agents across tasks and execution contexts. Existing evaluations pr

SkillHarness: Harnessing Safe Skills for Computer-Use Agents

SafetyDGX agent

arXiv:2606.20636v1 Announce Type: cross Abstract: Computer-Use Agents (CUAs) are increasingly deployed in dynamic interactive environments, creating a growing need for continual skill learning during

SkyJEPA: Learning Long-Horizon World Models for Zero-Shot Sim-to-Real Control of Quadrotors

ApplicationsDGX agent

arXiv:2606.23444v1 Announce Type: cross Abstract: Accurate dynamics models are critical for informed decision-making in robotic systems, particularly for agile aerial vehicles operating under uncertai

SLeDGe: Semi-Supervised Learning on Data Streams with Graph Structure Learning

ResearchDGX agent

arXiv:2606.21096v1 Announce Type: new Abstract: Semi-supervised learning (SSL) on data streams is challenging due to the continuous evolution of high-volume data and the scarcity of labels. Existing m

Small LLMs: Pruning vs. Training from Scratch

Model ReleasesDGX agent

arXiv:2606.14150v2 Announce Type: replace Abstract: Pruning promises a shortcut to strong small language models. In this work, we examine this promise by pruning Llama-3.1-8B at pruning ratios of 0.5-

SOAP-Bubbles: Structured Weight Uncertainty for Neural Networks

ResearchDGX agent

arXiv:2606.23357v1 Announce Type: new Abstract: Structured weight-uncertainty can improve many aspects of deep learning, but it remains costly to estimate and difficult to implement. Here, we show tha

SOHET: Sequence Of Heterogeneous Events Transformer with Self-Supervised Pre-Training

Model ReleasesDGX agent

arXiv:2606.21356v1 Announce Type: new Abstract: Many machine learning applications rely on heterogeneous event streams to make predictions, either causally as events arrive or bidirectionally over com

SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity

AgentsDGX agent

arXiv:2602.08690v2 Announce Type: replace Abstract: Deep Reinforcement Learning (DRL) has achieved remarkable success in domains requiring sequential decision-making, motivating its application to cyb

Solve for the Hyperparameter, Skip the Search: Kolmogorov-Optimal Scaling Laws for Spline Regression

SafetyDGX agent

arXiv:2606.23575v1 Announce Type: new Abstract: Hyperparameter tuning almost always means search: fit the model at every value on a grid, score each by cross-validation, and keep the winner. For splin

Sovereign Execution Broker: Enforcing Certificate-Bound Authority in Agentic Control Planes

SafetyDGX agent

arXiv:2606.20520v2 Announce Type: replace-cross Abstract: Autonomous agents are increasingly connected to cloud, deployment, and data-control workflows, but production mutation authority should not re

Spark: Strategic Policy-Aware Exploration via Dynamic Branching for Long-Horizon Agentic Learning

SafetyDGX agent

arXiv:2601.20209v2 Announce Type: replace Abstract: Reinforcement learning has empowered large language models to act as intelligent agents, yet training them for long-horizon tasks remains challengin

Specialize Roles, Mix Deployments: Pushing the Cost-Accuracy Frontier of LLM Agent Teams

Model ReleasesDGX agent

arXiv:2606.20629v1 Announce Type: cross Abstract: LLM agents are increasingly deployed as multi-role teams, where tasks are divided across specialized roles such as planner, executor, and verifier. In

Spectrally Safe Neural Operator Warm-Starts for Large-Scale Newton Solvers

ResearchDGX agent

arXiv:2606.21828v1 Announce Type: cross Abstract: Neural operators are increasingly used to warm-start Newton solvers for nonlinear PDEs, on the premise that a low test error places the initial guess

SpotAttention: Plug-In Block-Sparse Routing for Pretrained Long-Context Transformers

ResearchDGX agent

arXiv:2606.22874v1 Announce Type: new Abstract: Long contexts have become standard in pretrained LLMs, yet they remain expensive to run: prefill compute grows quadratically with sequence length, and e

SPOTR: Spatio-temporal Pooling One-Token Reconstruction for Universal Physiological Signal Self-supervised Learning

HardwareDGX agent

arXiv:2606.21973v1 Announce Type: new Abstract: Physiological signals such as EEG, ECG, and PPG are widely used in clinical monitoring. Recent self-supervised learning (SSL) methods offer an attractiv

SQLConductor: Search-to-Policy Learning for Step-wise Text-to-SQL Orchestration

SafetyDGX agent

arXiv:2606.23537v1 Announce Type: cross Abstract: Text-to-SQL enables users to access relational databases via natural language, but real-world settings remain challenging due to coordinated reasoning

Stage-dependent integer-binary encoding in factorization-machine black-box optimization

ResearchDGX agent

arXiv:2606.23188v1 Announce Type: new Abstract: Black-box optimization (BBO) deals with problems where objective functions lack explicit analytical forms and are expensive to evaluate. Factorization m

Stationary Robust Mean-Field Games under Model Mismatches

SafetyDGX agent

arXiv:2606.22579v1 Announce Type: new Abstract: Deploying multi-agent reinforcement learning (MARL) in the real world is often limited by model mismatches between the training simulators and the true

Statistical Inference for Misspecified Contextual Bandits

SafetyDGX agent

arXiv:2606.22639v1 Announce Type: cross Abstract: Contextual bandit algorithms have transformed modern experimentation by enabling real-time adaptation for personalized treatment. Yet these advantages

Statistical Matching via Schrodinger Bridge beyond Conditional Independence

ApplicationsDGX agent

arXiv:2606.22770v1 Announce Type: new Abstract: Statistical matching combines partially overlapping datasets that share covariates X but observe the target Y and auxiliary variables Z separately. Clas

Stealthy World Model Manipulation via Data Poisoning

ResearchDGX agent

arXiv:2606.18697v2 Announce Type: replace Abstract: Model-based learning agents use learned world models to predict future states, plan actions, and adapt to new environments. However, the process of

Steer, Don't Solve: Training Small Critic Models for Large Code Agents

Model ReleasesDGX agent

arXiv:2606.21811v1 Announce Type: cross Abstract: End-to-end code agent training is resource-intensive and plateaus on the strategy-level reasoning needed to resolve code issues, since jointly optimiz

Strengthening LLMs for Tabular Prediction with Structural Priors

Model ReleasesDGX agent

arXiv:2510.17385v5 Announce Type: replace Abstract: Tabular prediction has long been dominated by gradient-boosted decision trees and specialized deep tabular models, while large language models (LLMs

← Previous
1…8283848586…243
Next →