AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,557 results
3 Jun 2026

Making Brain-Computer Interfaces More Secure

ResearchDGX agent

arXiv:2606.02597v1 Announce Type: new Abstract: The development of brain-computer interfaces (BCIs) based on electroencephalograms (EEGs) has advanced significantly mainly to machine learning. Althoug

Minimax Optimal Strategy for Delayed Observations in Online Reinforcement Learning

AgentsDGX agent

arXiv:2603.03480v2 Announce Type: replace Abstract: We study reinforcement learning with delayed state observation, where the agent observes the current state after some random number of time steps. W

Mitigating False Credit Propagation: Probabilistic Graphical Reward Aggregation for Rubric-Based Reinforcement Learning

SafetyDGX agent

arXiv:2606.03361v1 Announce Type: new Abstract: Rubric-based rewards are increasingly used for open-ended language model post-training, but criterion-level scores are often aggregated as independent u


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Mitigating Spurious Correlations with Memorization-Guided Dataset De-Biasing

ApplicationsDGX agent

arXiv:2606.02830v1 Announce Type: new Abstract: Real-world datasets often contain spurious correlations that are not causally related to the target label. When such correlations dominate the majority

MLSkip: Data Skipping for ML Filters via Lightweight Metadata

Model ReleasesDGX agent

arXiv:2606.03946v1 Announce Type: cross Abstract: Database vendors recently released AI functions that can be used in filter predicates. As such functions often rely on costly, black-box ML models, th

MOSAIC: Efficient Mixture-of-Agent Scheduling via Adaptive Aggregation and Inference Concurrency

HardwareDGX agent

arXiv:2606.03014v1 Announce Type: new Abstract: Mixture-of-Agents (MoA) systems improve reasoning accuracy by routing each query to multiple expert LLMs and aggregating their outputs. Efficiently exec

MuLoCo: Muon is a practical inner optimizer for DiLoCo

ResearchDGX agent

arXiv:2505.23725v3 Announce Type: replace Abstract: DiLoCo is a powerful framework for training large language models (LLMs), enabling larger optimal batch sizes and increased accelerator utilization

Multi-Modal Machine Learning for Breast Cancer Recurrence Prediction

Model ReleasesDGX agent

arXiv:2606.02892v1 Announce Type: new Abstract: Breast cancer recurrence, a leading cause of long-term mortality among survivors, requires timely and accurate risk assessment to guide follow-up care a

Multi^2: Hierarchical Multi-Agent Decision-Making with LLM-Based Agents in Interactive Environments

Model ReleasesDGX agent

arXiv:2606.03698v1 Announce Type: new Abstract: A central goal of large language model (LLM) research is to build agentic systems that can plan, act, and adapt through sustained interaction with dynam

Names Don't Matter: Symbol-Invariant Transformer for Open-Vocabulary Learning

ResearchDGX agent

arXiv:2601.23169v2 Announce Type: replace Abstract: Current neural architectures lack a principled way to handle interchangeable tokens, i.e., symbols that are semantically equivalent yet distinguisha

Neural Navigation Functions for Zero-Shot Generalizable Motion Planning

Model ReleasesDGX agent

arXiv:2606.03756v1 Announce Type: cross Abstract: We introduce Neural Navigation Functions (Neural-NF), a learned reactive navigation function capable of zero-shot transfer across unseen environment g

Neural Networks Provably Learn Spectral Representations for Group Composition

SafetyDGX agent

arXiv:2606.02993v1 Announce Type: new Abstract: Understanding how structured internal structure emerges during neural network training is central to the study of deep learning. We investigate this phe

Neutrino Fingerprints: Image-Based Encodings of IceCube Events for CNN Direction Reconstruction

Model ReleasesDGX agent

arXiv:2606.02788v1 Announce Type: cross Abstract: Reconstructing the direction of incoming neutrinos in the IceCube Neutrino Observatory is an important problem in astrophysics. The public IceCube--Ne

OASIS: Outlier-Aware LUT-Based GEMM with Dual-Side Quantization for LLM Inference Acceleration

ResearchDGX agent

arXiv:2507.23035v4 Announce Type: replace Abstract: Large language models (LLMs) have demonstrated impressive capabilities across a wide range of applications, but demand substantial memory and comput

One Transit Is All You Need: Detecting Exoplanets Through Learned Stellar Behaviour with EXOVEIL

ResearchDGX agent

arXiv:2606.02778v1 Announce Type: cross Abstract: I present EXOVEIL, a transit detection system that learns what a star's brightness should look like and flags when reality disagrees. Unlike existing

Online Learning with Gradient-Variation Interval Regret

ResearchDGX agent

arXiv:2606.03831v1 Announce Type: new Abstract: This paper investigates non-stationary online learning using the metric of interval regret, which requires an online algorithm to perform well over ever

Optimal Initialization in Depth: Lyapunov Initialization and Limit Theorems for Deep Leaky ReLU Networks

Model ReleasesDGX agent

arXiv:2602.10949v2 Announce Type: replace-cross Abstract: Effective initialization in deep networks requires an understanding of random neural networks. In this work, a rigorous probabilistic analysis

Outsmarting the Chameleon: Counterfactual Decoupling for Tactical OOD Shifts in Live Streaming Risk Assessment

Model ReleasesDGX agent

arXiv:2606.02946v1 Announce Type: new Abstract: Live streaming has emerged as a primary medium for social interaction and digital commerce, yet it is increasingly plagued by sophisticated risks. A fun

ParaBlock: Communication-Computation Parallel Block Coordinate Federated Learning for Large Language Models

Local AiDGX agent

arXiv:2511.19959v2 Announce Type: replace Abstract: Federated learning (FL) has been extensively studied as a privacy-preserving training paradigm. Recently, federated block coordinate descent scheme

PerchRL: Vision-Based Agile Perching on Inclined Platforms under Rapid and Irregular Motion

Model ReleasesDGX agent

arXiv:2606.03441v1 Announce Type: cross Abstract: Autonomous vision-based perching of quadrotors on moving inclined platforms is critical for air-ground collaboration but remains challenging due to th

Position: Adversarial ML for LLMs Is Not Making Any Progress

ResearchDGX agent

arXiv:2502.02260v2 Announce Type: replace Abstract: In the past decade, considerable research effort has been devoted to securing machine learning (ML) models that operate in adversarial settings. Yet

Privacy-Robust Incrementality Measurement for Advertising Systems under Signal Loss

ResearchDGX agent

arXiv:2606.03878v1 Announce Type: cross Abstract: Advertising platforms use randomized lift tests to measure incrementality, but privacy-preserving reporting systems degrade the observed signal throug

Pruning Deep Neural Networks via the Marchenko--Pastur Distribution

HardwareDGX agent

arXiv:2606.02608v1 Announce Type: new Abstract: We study a Marchenko--Pastur (MP) random-matrix approach to pruning deep neural networks with very small post-pruning fine-tuning budgets. The main prac

Psi-Bench: Evaluating Persona-Sensitive Influencing in Persuasive Dialogues

Model ReleasesDGX agent

arXiv:2606.02754v1 Announce Type: new Abstract: Personalization is a crucial capability of modern language agents. However, current research primarily positions personalized agents as passive responde

Qift: Shift-Friendly No-Zero W2 Post-Training Quantization for Rotated W2A4/KV4 LLM Inference

Model ReleasesDGX agent

arXiv:2606.02823v1 Announce Type: new Abstract: Two-bit weight quantization is attractive for memory-efficient LLM inference, but the standard W2 level set {-2,-1,0,+1} often collapses under aggressiv

Quadratic integrate-and-fire neurons exhibit less fragmented loss landscapes and outperform leaky integrate-and-fire neurons in spike-based gradient descent

Model ReleasesDGX agent

arXiv:2606.03935v1 Announce Type: cross Abstract: The ability to train spiking neural networks is essential for modeling biological neural networks as well as for neuromorphic computing. However, for

QUIVER: Quantum-Informed Views for Enhanced Representations in Large ML Models

Model ReleasesDGX agent

arXiv:2606.02785v1 Announce Type: new Abstract: Large machine learning models benefit substantially from multimodal inputs that provide a complementary view of the same example. We introduce QUIVER (Q

R2DN: Scalable Parameterization of Contracting and Lipschitz Recurrent Deep Networks

ResearchDGX agent

arXiv:2504.01250v2 Announce Type: replace Abstract: This paper presents the Robust Recurrent Deep Network (R2DN), a scalable parameterization of robust recurrent neural networks for machine learning a

ReciNet: Reciprocal Space-Aware Long-Range Modeling for Crystalline Property Prediction

Local AiDGX agent

arXiv:2502.02748v4 Announce Type: replace Abstract: Predicting properties of crystals from their structures is a fundamental yet challenging task in materials science. Unlike molecules, crystal struct

Regime-Arrival Uncertainty in Generalization Bounds under Distribution Shift

ResearchDGX agent

arXiv:2606.02657v1 Announce Type: new Abstract: The standard generalization bounds assume that the training and deployment distributions are the same, or are static, and don't consider regime switchin

RESCAST-100K: A Comprehensive Dataset for Cross-Domain Residential Load and Indoor Temperature Forecasting

Model ReleasesDGX agent

arXiv:2606.02852v1 Announce Type: new Abstract: Accurate short-term forecasting of residential energy load and indoor temperature is essential for home energy management systems, grid-level demand res

Resource-Constrained Adaptive Inference for Sequential Pricing

Local AiDGX agent

arXiv:2606.03736v1 Announce Type: cross Abstract: Resource-constrained pricing controllers can make fixed-price inference impossible: the controller's resource state may remove the target price neighb

Rethinking Neural Width for Alternating Current Optimal Power Flow Proxies

SafetyDGX agent

arXiv:2606.03125v1 Announce Type: new Abstract: Deep learning proxies for Alternating Current Optimal Power Flow (ACOPF) lack systematic methods for determining architectural size. This paper conducts

Right Makes Might: Aligning Verified Hidden States Empowers RL Reasoning

SafetyDGX agent

arXiv:2606.03234v1 Announce Type: new Abstract: Reinforcement Learning from Verifiable Rewards (RLVR) has become the dominant approach for improving mathematical reasoning in large language models, ye

RMPrior: Bridging Propagation Priors and Diffusion Refinement for Efficient Radio Map Construction

ResearchDGX agent

arXiv:2606.03074v1 Announce Type: new Abstract: Diffusion models achieve high-fidelity radio map construction through iterative denoising, yet their sampling cost limits practicality in dynamic wirele

RogueMerge: Robust and Unified Attacks against LLM Model Merging

Model ReleasesDGX agent

arXiv:2606.03344v1 Announce Type: cross Abstract: Model merging composes specialized capabilities into a single LLM by aggregating task vectors sourced from unverified public platforms, exposing a cri

RRISE: Robust Radius Inference via a Surrogate Estimator

ResearchDGX agent

arXiv:2606.02876v1 Announce Type: new Abstract: Randomized smoothing (RS) uses a smoothed classifier to provide architecture-agnostic certificates of ell_2 classification robustness, but its dependenc

SAIL: Sound Abstract Interpreters with LLMs

TutorialsDGX agent

arXiv:2511.13663v2 Announce Type: replace-cross Abstract: How to construct globally sound abstract interpreters to safely approximate program behaviors remains a bottleneck in abstract interpretation.

Scalable Derivative Gaussian Processes via Exact Gradient Reduction

Local AiDGX agent

arXiv:2606.02909v1 Announce Type: cross Abstract: Gradient observations can substantially improve Gaussian process (GP) surrogates, particularly in high-dimensional settings where function evaluations

Scalable Single-Cell Gene Expression Generation with Latent Diffusion Models

ResearchDGX agent

arXiv:2511.02986v2 Announce Type: replace-cross Abstract: Computational modeling of single-cell gene expression is crucial for understanding cellular processes, but generating realistic expression pro

scBatchProx: Federated-Inspired Refinement for Stable Cell-Type Discriminability under Heterogeneous Batch Compositions

ResearchDGX agent

arXiv:2602.00423v3 Announce Type: replace Abstract: Single-cell integration workflows often construct low-dimensional cell embeddings and then refine them with post-hoc methods to reduce batch effects

ScoreStop: Gradient-based early stopping using functional score tests

Model ReleasesDGX agent

arXiv:2606.02740v1 Announce Type: cross Abstract: Gradient boosted decision trees require a stopping rule to avoid overfitting. The standard rule monitors a validation loss and stops if the loss fails

SeeTraceAct: Visibility-Aware Latent Planning from Cross-Embodiment Demonstration Videos

Model ReleasesDGX agent

arXiv:2606.02745v1 Announce Type: cross Abstract: Vision-language-action models (VLAs) are promising general-purpose robot policies, but adapting them to new tasks typically requires costly task-speci

Self-Soupervision: Cooking Model Soups without Labels

ResearchDGX agent

arXiv:2602.02890v2 Announce Type: replace Abstract: Model soups are strange and strangely effective combinations of parameters. They take a model (the stock), fine-tune it into multiple models (the in

Set-Preserving Calibration from Conformal P-Values to E-Values

ResearchDGX agent

arXiv:2606.03600v1 Announce Type: cross Abstract: Standard conformal prediction (CP) procedures are typically formulated in terms of p-values, but reliance on p-values alone limits flexibility, for ex

SketchSong: Hierarchical Song Generation with Sketch Planning and Fine-Grained Multi-Track Modeling

ResearchDGX agent

arXiv:2606.03169v1 Announce Type: cross Abstract: Recent song generation systems can synthesize realistic audio, yet generating complete songs remains challenging for two reasons. First, explicit song

Spatial Transcriptomics-Guided Alignment Enhances Molecular Profiling in Pathology Foundation Model

Model ReleasesDGX agent

arXiv:2606.03644v1 Announce Type: new Abstract: Comprehensive molecular profiling is essential for modern precision oncology but remains hindered by prohibitive costs, specimen exhaustion, and protrac

Spectral Asymptotics of Neural Network Loss Landscapes: An Exact Decomposition of the Curvature Exponent

SafetyDGX agent

arXiv:2606.02596v1 Announce Type: new Abstract: The curvature exponent alpha in h_k propto sigma_k^alpha -- governing how Hessian eigenvalues scale with gradient singular values -- varies systematical

Spectral-Progressive Thought Flow for Lightweight Multimodal Reasoning

ResearchDGX agent

arXiv:2606.02842v1 Announce Type: new Abstract: Multimodal spatial reasoning often relies on long chains of intermediate textual and visual thoughts, where accumulating visual tokens and dense cross-m

Speedrunning Tabular Foundation Model Pretraining

HardwareDGX agent

arXiv:2606.03681v1 Announce Type: new Abstract: Pretraining cost is a major bottleneck for research on tabular foundation models, slowing the iteration cycle for new architectures, priors, and optimiz

State-Coupled Volatility in Latent Dynamical Systems: Recovery Under Partial Observation

Model ReleasesDGX agent

arXiv:2606.02664v1 Announce Type: cross Abstract: Latent state-space models are widely used to study partially observed dynamical systems, yet most formulations assume that process variability is inde

Suboptimality bounds for trace-bounded SDPs enable a faster and scalable low-rank SDP solver SDPLR+

ResearchDGX agent

arXiv:2406.10407v3 Announce Type: replace-cross Abstract: Semidefinite programs (SDPs) and their solvers are powerful tools with many applications in machine learning and data science. Designing scala

Synthetic Hallucinations, Real Gains: Hard Negatives from Frontier Models for FIM Hallucination Mitigation

Model ReleasesDGX agent

arXiv:2606.03130v1 Announce Type: new Abstract: Small open-source code models that power IDE autocomplete still emit hallucinated Fill-in-the-Middle (FIM) completions: syntactically natural calls to m

Tailoring Strictly Proper Scoring Rules for Downstream Tasks: An Application to Causal Inference

Local AiDGX agent

arXiv:2606.03332v1 Announce Type: new Abstract: Probabilistic models are typically trained using task-agnostic objectives like log-loss, which can lead to significant errors in downstream estimation.

Testing Most Influential Sets

ResearchDGX agent

arXiv:2510.20372v4 Announce Type: replace-cross Abstract: Small influential data subsets can dramatically impact model conclusions, with a few data points overturning key findings. While recent work i

Testing the Test: Score-Direction Instability in Class-Split Anomaly Detection

ResearchDGX agent

arXiv:2606.02601v1 Announce Type: new Abstract: Within-dataset class-split evaluation is widely used as a proxy for fully unconditional out-of-distribution anomaly detection. We show that this protoco

Text-attributed Graph Condensation via Text Selection and Attribute Matching

ResearchDGX agent

arXiv:2606.03839v1 Announce Type: new Abstract: Text-Attributed Graph (TAG) is an important type of graph structured data, where each node has a text description. TAG models usually train a Graph Neur

The Efficiency vs. Accuracy Trade-off: Optimizing RAG-Enhanced LLM Recommender Systems Using Multi-Head Early Exit

ResearchDGX agent

arXiv:2501.02173v2 Announce Type: replace-cross Abstract: The deployment of Large Language Models (LLMs) in recommender systems for predicting Click-Through Rates (CTR) necessitates a delicate balance

The Impact of Temporal Granularity on Socio-Demographic Inference from Household Load Profiles

ResearchDGX agent

arXiv:2606.03358v1 Announce Type: new Abstract: Smart meter data can reveal sensitive socio-demographic characteristics of households, raising privacy concerns. While this risk has been demonstrated a

Theoretical Aspects of Lie Groupoid and Lie Algebroid Equivariant Convolutional Neural Networks

ResearchDGX agent

arXiv:2606.02758v1 Announce Type: cross Abstract: We introduce Lie groupoid equivariant neural networks as a specialization of recently proposed topological category-equivariant neural networks to the

← Previous
1…102103104105106…243
Next →