AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlog
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,557 results
Local Ai

Mesh-RL: Coupled subgrid reinforcement learning

DGX agent

arXiv:2606.26333v1 Announce Type: new Abstract: Reinforcement learning in large or sparse-reward environments suffers from slow temporal-difference reward propagation, as value information spreads onl

local-aiarxiv-cs-lg
26 Jun 2026
Model Releases

MetaboNet-Bench: A Multi-modal Benchmark for Glucose Forecasting in Type 1 Diabetes

X Post
Paper
YouTube
Reddit
GitHub
Clear filters
DGX agent

arXiv:2606.18640v2 Announce Type: replace Abstract: Glucose forecasting algorithms are an important aspect of glycemic control management in type 1 diabetes. So far, the research community has develop

model-releasesarxiv-cs-lg
26 Jun 2026
Safety

NASimJax: A GPU-Accelerated Policy Learning Framework for Penetration Testing

DGX agent

arXiv:2603.19864v2 Announce Type: replace Abstract: Penetration testing, the practice of simulating cyberattacks to identify vulnerabilities, is a complex sequential decision-making task that is inher

safetyarxiv-cs-lg
26 Jun 2026
Model Releases

Necessary but Not Sufficient: Temperature Control and Reproducibility in LLM-as-Judge Safety Evaluations

DGX agent

arXiv:2606.26185v1 Announce Type: new Abstract: LLM-as-judge ('grader') components are now standard in evaluation harnesses, including safety evaluations where a pass/fail verdict may gate downstream

model-releasesarxiv-cs-lg
26 Jun 2026
Research

NervePool: A Simplicial Pooling Layer

DGX agent

arXiv:2305.06315v3 Announce Type: replace-cross Abstract: For deep learning problems on graph-structured data, pooling layers are important for down sampling, reducing computational cost, and to minim

researcharxiv-cs-lg
26 Jun 2026
Applications

No Free Lunch: Non-Asymptotic Analysis of Prediction-Powered Inference

DGX agent

arXiv:2505.20178v2 Announce Type: replace-cross Abstract: Prediction-Powered Inference (PPI) is a popular strategy for combining gold-standard and possibly noisy pseudo-labels to perform statistical e

applicationsarxiv-cs-lg
26 Jun 2026
Safety

Normalizing Flows are Capable Models for Continuous Control

DGX agent

arXiv:2505.23527v4 Announce Type: replace Abstract: Modern reinforcement learning (RL) algorithms have found success by using powerful probabilistic models, such as transformers, energy-based models,

safetyarxiv-cs-lg
26 Jun 2026
Hardware

Optimizing CUDA like a Human: Micro-Profiling Tools as Expert Surrogates for LLM-Based GPU Kernel Optimization

DGX agent

arXiv:2606.26453v1 Announce Type: new Abstract: We present KernelPro, a closed-loop multi-agent system that automatically generates, profiles, and iteratively optimizes GPU kernel code by integrating

hardwarearxiv-cs-lg
26 Jun 2026
Hardware

Otter Weather: Skillful and Computationally Efficient Medium-Range Weather Forecasting

DGX agent

arXiv:2606.26421v1 Announce Type: new Abstract: State-of-the-art medium-range AI weather models can outperform traditional Numerical Weather Prediction (NWP) but require massive training budgets. This

hardwarearxiv-cs-lg
26 Jun 2026
Model Releases

Over-parameterization and Adversarial Robustness in Neural Networks: An Overview and Empirical Analysis

DGX agent

arXiv:2406.10090v3 Announce Type: replace Abstract: Thanks to their extensive capacity, over-parameterized neural networks exhibit superior predictive capabilities and generalization. However, having

model-releasesarxiv-cs-lg
26 Jun 2026
Model Releases

PersistentKV: Page-Aware Decode Scheduling for Long-Context LLM Serving on Commodity GPUs

DGX agent

arXiv:2606.26666v1 Announce Type: new Abstract: Autoregressive large language model (LLM) serving is increasingly limited by key-value (KV) cache movement rather than dense matrix multiplication. Mode

model-releasesarxiv-cs-lg
26 Jun 2026
Tutorials

Physics-guided Convolutional Neural Network for Domain Growth Prediction in Systems with Conserved Kinetics

DGX agent

arXiv:2606.26128v1 Announce Type: new Abstract: The spatiotemporal evolution of many physical, chemical, and biological systems is described by nonlinear partial differential equations (PDEs). Recentl

tutorialsarxiv-cs-lg
26 Jun 2026
Local Ai

Quantization in Federated Learning: Methods, Challenges and Future Directions

DGX agent

arXiv:2606.26822v1 Announce Type: new Abstract: Federated Learning (FL) has become a foundational paradigm for privacy-preserving distributed intelligence, yet its scalability remains fundamentally co

local-aiarxiv-cs-lg
26 Jun 2026
Research

Reasoning Quality Emerges Early: Data Curation for Reasoning Models

DGX agent

arXiv:2606.26797v1 Announce Type: new Abstract: Supervised fine-tuning (SFT) on a small, high-quality set of long reasoning traces is an effective approach for eliciting strong reasoning capabilities

researcharxiv-cs-lg
26 Jun 2026
Safety

RecallRisk-BERT: A Multi-Task Framework for Post-Report Medical Device Recall Triage

DGX agent

arXiv:2606.27174v1 Announce Type: new Abstract: Medical device recalls are a critical regulatory mechanism for protecting patient safety. The growing volume of FDA recall records presents challenges i

safetyarxiv-cs-lg
26 Jun 2026
Research

Recovering Governing Equations from Solution Data: Identifiability Bounds for Linear and Nonlinear ODEs

DGX agent

arXiv:2606.27285v1 Announce Type: new Abstract: Learning governing equations from observed solution data is a fundamental challenge in scientific machine learning ite{bruntonDiscoveringGoverningEquati

researcharxiv-cs-lg
26 Jun 2026
Model Releases

Reinforcement Fine-Tuning of Flow-Matching Policies for Vision-Language-Action Models

DGX agent

arXiv:2510.09976v2 Announce Type: replace Abstract: Vision-Language-Action (VLA) models such as OpenVLA, Octo, and pi_0 have shown strong generalization by leveraging large-scale demonstrations, yet t

model-releasesarxiv-cs-lg
26 Jun 2026
Agents

Reinforcement Learning Enables Autonomous Microrobot Navigation and Intervention in Simulated Blood Capillaries

DGX agent

arXiv:2606.26154v1 Announce Type: cross Abstract: Autonomous microrobots navigating biological vasculature could enable targeted drug delivery and thrombolysis, yet training control policies for reali

agentsarxiv-cs-lg
26 Jun 2026
Model Releases

Reinforcement Learning without Ground-Truth Solutions can Improve LLMs

DGX agent

arXiv:2606.27369v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) for training LLMs typically rely on ground-truth answers to assign rewards, limiting their applica

model-releasesarxiv-cs-lg
26 Jun 2026
Model Releases

Representation Costs in Data Science: Foundations and the Quasi-Banach Spaces of Deep Neural Networks

DGX agent

arXiv:2606.14954v3 Announce Type: replace-cross Abstract: We develop a general framework for analyzing representation costs of parametric data-fitting methods through their parameter-space regularizer

model-releasesarxiv-cs-lg
26 Jun 2026
Agents

Revisiting Action Factorization for Complex Action Spaces

DGX agent

arXiv:2606.26574v1 Announce Type: new Abstract: Many real-world control problems involve hybrid discrete-continuous action spaces. For example, steering and signaling in autonomous driving, and aiming

agentsarxiv-cs-lg
26 Jun 2026
Model Releases

Ribbon: Scalable Approximation and Robust Uncertainty Quantification

DGX agent

arXiv:2606.27269v1 Announce Type: cross Abstract: Reliably quantifying predictive uncertainty is difficult for complex, high-dimensional, or misspecified models. Both fully Bayesian and bootstrap resa

model-releasesarxiv-cs-lg
26 Jun 2026
Safety

RolloutPipe: Overlapping Pipelined Rollout and Training in Disaggregated On-Policy LLM Reinforcement Learning

DGX agent

arXiv:2606.26997v1 Announce Type: cross Abstract: Large language model (LLM) post-training for reasoning increasingly relies on reinforcement learning with verifiable rewards (RLVR), where models lear

safetyarxiv-cs-lg
26 Jun 2026
Model Releases

RSPC: A Benchmark for Modeling Stress and Psychiatric Conditions in Digitally Mediated Relationships using Psychiatrist Annotations

DGX agent

arXiv:2606.27247v1 Announce Type: new Abstract: In NLP, mental health conditions are often modeled as isolated phenomena, without interpersonal context. We use Reddit posts about long-distance relatio

model-releasesarxiv-cs-lg
26 Jun 2026
Model Releases

Running the Gauntlet: Re-evaluating the Capabilities of Agents Beyond Familiar Environments

DGX agent

arXiv:2606.14397v2 Announce Type: replace Abstract: As agentic systems continue to evolve and are widely deployed in real-world scenarios, there is a growing demand to faithfully evaluate their capabi

model-releasesarxiv-cs-lg
26 Jun 2026
Safety

Sample-efficient Transfer Reinforcement Learning via Adaptive Reward Shaping and Policy-Ratio Reweighting Strategy

DGX agent

arXiv:2606.26527v1 Announce Type: new Abstract: Transfer learning improves policy learning efficiency by reusing knowledge from source tasks, providing a feasible paradigm for safe and efficient auton

safetyarxiv-cs-lg
26 Jun 2026
Research

Scalable Message-Passing Quantum Graph Neural Networks in the Weisfeiler-Leman Hierarchy

DGX agent

arXiv:2606.26873v1 Announce Type: cross Abstract: Graphs provide a natural language for relational data in chemistry, biology and optimisation. Graph neural networks (GNNs) have driven much of the rec

researcharxiv-cs-lg
26 Jun 2026
Safety

Scoring Is Not Enough: Addressing Gaps in Utility-fairness Trade-offs for Ranking

DGX agent

arXiv:2606.26369v1 Announce Type: cross Abstract: Scoring functions are used to represent the relevance of individual documents. In modern information retrieval or recommendation systems, they are oft

safetyarxiv-cs-lg
26 Jun 2026
Model Releases

Signature filtering: a lightweight enhancement for statistical watermark detection in large language models

DGX agent

arXiv:2606.18430v2 Announce Type: replace Abstract: Statistical watermarks help organizations attribute large language model (LLM) outputs, yet existing detectors often struggle when watermark signals

model-releasesarxiv-cs-lg
26 Jun 2026
Safety

Sketched Linear Contrastive Learning: Approximation, Optimization, and Statistical Scaling

DGX agent

arXiv:2606.26617v1 Announce Type: new Abstract: Scaling laws describe how learning performance varies with model size, data size, and compute. While recent theoretical work has established scaling law

safetyarxiv-cs-lg
26 Jun 2026
Research

State-Specific Respiratory Signatures for Affective and Stress Recognition: Interpretable Respiratory Markers, Autocorrelation Lags, and Compact CNN Models

DGX agent

arXiv:2606.26723v1 Announce Type: cross Abstract: Respiratory activity is a direct and interpretable physiological channel for wearable stress and affective-state recognition, yet many studies emphasi

researcharxiv-cs-lg
26 Jun 2026
Model Releases

Stochastic Gradient Optimization with Model-Assisted Sampling

DGX agent

arXiv:2606.27171v1 Announce Type: new Abstract: This work addresses the problem of variance in stochastic gradient estimation for machine learning optimization. Deep learning relies on mini-batch meth

model-releasesarxiv-cs-lg
26 Jun 2026
Research

Symplectic Neural Networks for learning Generalized Hamiltonians

DGX agent

arXiv:2606.27029v1 Announce Type: new Abstract: Hamiltonian Neural Networks (HNNs) integrate physical priors into neural models by learning a system's Hamiltonian, improving generalization and sample

researcharxiv-cs-lg
26 Jun 2026
Applications

Target-Aware Bandit Allocation for Scalable Surrogate Optimization in Chemical Space

DGX agent

arXiv:2606.26657v1 Announce Type: new Abstract: Identifying high-utility candidates from massive discrete spaces under expensive evaluations is a recurring challenge across the sciences, with structur

applicationsarxiv-cs-lg
26 Jun 2026
Research

The Role of Input Dimensionality in the Emergence and Targeted Control of Adversarial Examples

DGX agent

arXiv:2606.26207v1 Announce Type: cross Abstract: Several theoretical works have tried to explain the adversarial vulnerability of deep neural networks through properties of high-dimensional geometry.

researcharxiv-cs-lg
26 Jun 2026
Tutorials

Theory of the Frequency Principle for General Deep Neural Networks

DGX agent

arXiv:1906.09235v3 Announce Type: replace Abstract: Along with fruitful applications of Deep Neural Networks (DNNs) to realistic problems, recently, some empirical studies of DNNs reported a universal

tutorialsarxiv-cs-lg
26 Jun 2026
Model Releases

Theory-Scale Auto-Formalization of Logics for Computer Science

DGX agent

arXiv:2606.26525v1 Announce Type: new Abstract: Auto-formalization is critical for scalable formal verification, but existing progress largely focuses on isolated statements, while theory-scale auto-f

model-releasesarxiv-cs-lg
26 Jun 2026
Safety

Topology-Informed Neural Networks for Flood Detection in Optical and Synthetic Aperture Radar Imagery

DGX agent

arXiv:2606.26204v1 Announce Type: new Abstract: Floods frequently impact regions around the world. Rapid and accurate flood detection is crucial for emergency response and timely mitigation of human a

safetyarxiv-cs-lg
26 Jun 2026
Safety

Training-Free Generation of Protein Sequences from Small Family Alignments via Stochastic Attention

DGX agent

arXiv:2603.14717v2 Announce Type: replace Abstract: Generating novel protein sequences that respect a family's statistical constraints typically requires training deep generative models on thousands t

safetyarxiv-cs-lg
26 Jun 2026
Research

Transformer-Based Classification of Bacterial Raman Spectra with LOOCV

DGX agent

arXiv:2606.27096v1 Announce Type: new Abstract: Transformer-based models have recently attracted increasing attention for Raman spectral classification. In this study, a transformer-based approach was

researcharxiv-cs-lg
26 Jun 2026
Research

Uncertainty quantification via conformal prediction in data assimilation

DGX agent

arXiv:2606.27001v1 Announce Type: new Abstract: Quantifying the evolution of uncertainty is critical to both probabilistic forecasting and data assimilation in numerical weather prediction. In this st

researcharxiv-cs-lg
26 Jun 2026
Research

Use What You Know: Causal Foundation Models with Partial Graphs

DGX agent

arXiv:2602.14972v2 Announce Type: replace Abstract: Estimating causal quantities traditionally relies on bespoke estimators tailored to specific assumptions. Recently proposed Causal Foundation Models

researcharxiv-cs-lg
26 Jun 2026
Research

What Survives When You Compress a Recursive Reasoner for the Edge?

DGX agent

arXiv:2606.26488v1 Announce Type: new Abstract: Recursive reasoning models can solve complex structured tasks with only a few million parameters by repeatedly updating a latent state. Deploying these

researcharxiv-cs-lg
26 Jun 2026
Research

When are likely answers right? On Sequence Probability and Correctness in LLMs

DGX agent

arXiv:2606.27359v1 Announce Type: cross Abstract: Many decoding methods for large language models can be understood as shifting probability mass toward outputs that are more likely under the model, ei

researcharxiv-cs-lg
26 Jun 2026
Research

When Does Quality-Aware Multimodal Fusion Matter? A Leakage-Safe Diagnostic for Decision-Level Dependence

DGX agent

arXiv:2606.26473v1 Announce Type: new Abstract: Many multimodal systems estimate the reliability of each modality and weight their contributions to the final prediction. However, it remains unclear wh

researcharxiv-cs-lg
26 Jun 2026
Model Releases

When to Write and When to Suppress: Route-Specialized Dual Adapters for Memory-Assisted Knowledge Editing

DGX agent

arXiv:2606.14668v3 Announce Type: replace Abstract: Knowledge editing systems must update selected facts while preserving nearby but irrelevant behavior. This paper studies this problem in a memory-as

model-releasesarxiv-cs-lg
26 Jun 2026
Model Releases

A 3D-Printable Dataset for Fair Testing and Comparisons of Tactile Sensors

DGX agent

arXiv:2606.25886v1 Announce Type: cross Abstract: Existing texture datasets for tactile sensing primarily consist of sensor readings from a specific sensor interacting with available surfaces/objects

model-releasesarxiv-cs-lg
25 Jun 2026
Research

A Bregman Perspective on Classification and Regression Trees

DGX agent

arXiv:2606.13984v2 Announce Type: replace-cross Abstract: Classification and Regression Trees (CART) constitute one of the most influential paradigms in statistical learning. Although a variety of imp

researcharxiv-cs-lg
25 Jun 2026
← Previous
1…8586878889…304
Next →