AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlog
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,432 results
Safety

Start Classifying: Categorical Critics for LLM Reinforcement Learning

DGX agent

arXiv:2608.02181v1 Announce Type: new Abstract: Proximal Policy Optimization (PPO) for large language models typically trains its critic by mean-squared-error (MSE) regression on scalar value targets.

safetyarxiv-cs-lg
4 Aug 2026
Research
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

Statistical comparisons of time-series feature sets on classification tasks

DGX agent

arXiv:2608.01586v1 Announce Type: cross Abstract: In recent years, numerous open-source software libraries have been developed for computing sets of features from univariate time series. The type and

researcharxiv-cs-lg
4 Aug 2026
Research

Statistical Mechanics of Learning on Product Wasserstein Manifolds

DGX agent

arXiv:2608.01434v1 Announce Type: new Abstract: Normally the statistical mechanics of learning treats constraints on weight distributions as restrictions that shrink the space of possible solutions. T

researcharxiv-cs-lg
4 Aug 2026
Agents

Stop When Memory Suffices: Evidence-Conditioned Progressive Execution for LLM Agents

DGX agent

arXiv:2608.01285v1 Announce Type: new Abstract: The continued development of LLMs toward persistent and adaptive intelligence increasingly requires long-term memory mechanisms that preserve and reuse

agentsarxiv-cs-lg
4 Aug 2026
Model Releases

Structured Memory for Edge Language Models: Persistent Context and Corpus Retrieval via O(1) SSM State Injection

DGX agent

arXiv:2608.02560v1 Announce Type: new Abstract: Retrieval-augmented generation (RAG) imposes a prefill cost proportional to retrieved context length, and -- with Transformer backbones -- a KV-cache th

model-releasesarxiv-cs-lg
4 Aug 2026
Research

Subtype Robustness Is Not Just Accuracy: Calibration Under Unseen Subtype Shift

DGX agent

arXiv:2608.00928v1 Announce Type: new Abstract: Subtype robustness asks whether a model keeps the correct coarse prediction when test examples come from fine-grained subtypes absent from training but

researcharxiv-cs-lg
4 Aug 2026
Research

Surrogate Modeling for the Design of Optimal Lattice Structures using Tensor Completion

DGX agent

arXiv:2510.07474v2 Announce Type: replace Abstract: When designing new materials, it is often necessary to design a material with specific desired properties. Unfortunately, as new design variables ar

researcharxiv-cs-lg
4 Aug 2026
Research

T-TAMER: Provably Taming Trade-offs in ML Serving

DGX agent

arXiv:2509.22992v2 Announce Type: replace Abstract: As machine learning models continue to grow in size and complexity, efficient serving faces increasingly broad trade-offs spanning accuracy, latency

researcharxiv-cs-lg
4 Aug 2026
Model Releases

TabDPT-Turbo: Efficient In-Context Learning for Tabular Prediction

DGX agent

arXiv:2608.01400v1 Announce Type: new Abstract: Tabular foundation models, driven by in-context learning, have rapidly grown in quality and popularity. However, recent approaches with either cell-base

model-releasesarxiv-cs-lg
4 Aug 2026
Model Releases

Tevatron Meets Megatron: Expert-Parallel LLM Reranker Training on an Academic Budget

DGX agent

arXiv:2608.00916v1 Announce Type: cross Abstract: Modern reranking recipes---billion-scale cross-encoders, mixture-of-experts (MoE) backbones, and distillation against strong teachers---have outpaced

model-releasesarxiv-cs-lg
4 Aug 2026
Research

tFUSOperator: Operator Learning for Transcranial Focused Ultrasound Digital Twins

DGX agent

arXiv:2608.01839v1 Announce Type: new Abstract: Transcranial focused ultrasound (tFUS) requires accurate estimation of the intracranial acoustic field, which is distorted by skull-induced aberrations.

researcharxiv-cs-lg
4 Aug 2026
Research

The Bayesian Reflex: A Predictive Coding Engine for Artificial Intelligence

DGX agent

arXiv:2608.00492v1 Announce Type: cross Abstract: Predictive coding offers a powerful theory of cortical computation, but corresponding scalable algorithmic implementations for artificial intelligence

researcharxiv-cs-lg
4 Aug 2026
Model Releases

The Condition-Number Barrier in Sparse Least Squares

DGX agent

arXiv:2608.02588v1 Announce Type: cross Abstract: In [AS21], Axiotis and Sviridenko conjectured that the linear dependence on the restricted condition number in sparse convex optimization cannot be im

model-releasesarxiv-cs-lg
4 Aug 2026
Research

The Elements of Differentiable Programming

DGX agent

arXiv:2403.14606v4 Announce Type: replace Abstract: Artificial intelligence has recently experienced remarkable advances, fueled by large models, vast datasets, accelerated hardware, and, last but not

researcharxiv-cs-lg
4 Aug 2026
Research

The Fourth Quadrant: A Stylized View of Benign Misfitting

DGX agent

arXiv:2608.01032v1 Announce Type: new Abstract: Training error is what we can observe on a training set; test error is the quantity we actually care about. We study linear regression with squared-erro

researcharxiv-cs-lg
4 Aug 2026
Research

The Label Defines the Timescale: Trait-State Limits of Temporal-Aggregate Learning

DGX agent

arXiv:2608.01587v1 Announce Type: cross Abstract: Machine-learning benchmarks often pair a label that aggregates a long temporal horizon with input observed through one or a few short windows. Their a

researcharxiv-cs-lg
4 Aug 2026
Research

The No-Clash Teaching Dimension is Bounded by VC Dimension

DGX agent

arXiv:2603.23561v4 Announce Type: replace-cross Abstract: In the realm of machine learning theory, to prevent unnatural coding schemes between teacher and learner, No-Clash Teaching Dimension was intr

researcharxiv-cs-lg
4 Aug 2026
Research

Thermalizing Stochastic Programs

DGX agent

arXiv:2608.01615v1 Announce Type: cross Abstract: We present a set of tools for mapping general stochastic programs to thermodynamic hardware designed for energy-efficient stochastic sampling. Given a

researcharxiv-cs-lg
4 Aug 2026
Model Releases

Towards Anomaly Detection on Relational Data

DGX agent

arXiv:2606.18621v2 Announce Type: replace Abstract: Relational databases are widely used for managing structured data in real-world systems. Detecting anomalies from such relational data is crucial fo

model-releasesarxiv-cs-lg
4 Aug 2026
Research

Towards Effective Federated Multimodal Graph Learning via Navigating Multifaceted Heterogeneity

DGX agent

arXiv:2608.00623v1 Announce Type: new Abstract: Multimodal-attributed graphs (MAGs), where nodes carry heterogeneous semantic content across multiple modalities while edges encode relational dependenc

researcharxiv-cs-lg
4 Aug 2026
Safety

Towards General Language-Conditioned Latent Safety Filters

DGX agent

arXiv:2608.00315v1 Announce Type: cross Abstract: Robot policies are becoming increasingly general, with vision-language-action (VLA) models enabling a single policy to execute diverse tasks specified

safetyarxiv-cs-lg
4 Aug 2026
Model Releases

Training Deep Morphological Neural Networks as Universal Approximators

DGX agent

arXiv:2505.09710v4 Announce Type: replace Abstract: We investigate deep morphological neural networks (DMNNs), studying how changes in algebraic structure affect the expressivity and trainability of d

model-releasesarxiv-cs-lg
4 Aug 2026
Model Releases

Training nGPT

DGX agent

arXiv:2608.01284v1 Announce Type: new Abstract: The normalized Transformer (nGPT) realizes hyperspherical representation learning by constraining model parameter vectors and activation vectors to the

model-releasesarxiv-cs-lg
4 Aug 2026
Safety

Training Small LLMs as Spatial Multi-Agent Policies

DGX agent

arXiv:2608.01425v1 Announce Type: cross Abstract: Training LLM-based multi-agent systems with multi-agent reinforcement learning is rapidly gaining traction, and a parallel line of work argues that su

safetyarxiv-cs-lg
4 Aug 2026
Agents

Trajectories That Segment Themselves: Agent-Declared Boundaries as a Training Unit

DGX agent

arXiv:2608.02302v1 Announce Type: cross Abstract: Long-horizon coding-agent trajectories are poorly matched to the credit units available to train on: a single action has no stable value, an episode l

agentsarxiv-cs-lg
4 Aug 2026
Applications

Transfer Learning of CATE with Kernel Ridge Regression

DGX agent

arXiv:2502.11331v4 Announce Type: replace-cross Abstract: The proliferation of data has sparked significant interest in leveraging findings from one study to estimate treatment effects in a different

applicationsarxiv-cs-lg
4 Aug 2026
Safety

Trust or Check? Understanding the (Evolutionary) Dynamics of User Trust in AI Systems

DGX agent

arXiv:2603.24742v2 Announce Type: replace-cross Abstract: As the capabilities and adoption of Artificial Intelligence (AI) systems grow, trust in these AI systems is an increasingly urgent concern. Mu

safetyarxiv-cs-lg
4 Aug 2026
Safety

Trustworthy AI in Digital Health: A Comprehensive Review of Robustness and Explainability

DGX agent

arXiv:2608.02238v1 Announce Type: cross Abstract: Ensuring trust in AI systems is essential for the safe and ethical integration of machine learning systems into high-stakes domains such as digital he

safetyarxiv-cs-lg
4 Aug 2026
Model Releases

Tunneling the Loss Landscape: Bypassing Memorization with Monte Carlo Parameter Swapping

DGX agent

arXiv:2608.01833v1 Announce Type: cross Abstract: Grokking is a striking phenomenon in neural network training, where a model can undergo a prolonged period of pure memorization before abrupt generali

model-releasesarxiv-cs-lg
4 Aug 2026
Local Ai

Uncertainty-Aware Simulation-Based Inference for Operations Research with Large Language Models

DGX agent

arXiv:2608.00019v1 Announce Type: new Abstract: Deploying large language models (LLMs) for operations research (OR) tasks remains challenging because correctness depends on a coherent modeling process

local-aiarxiv-cs-lg
4 Aug 2026
Research

Uncertainty-guided active learning for surrogate prediction of stream-finishing wear fields

DGX agent

arXiv:2608.00593v1 Announce Type: cross Abstract: In stream finishing, the wear experienced by a workpiece depends strongly on its orientation within the rotating abrasive media. Determining suitable

researcharxiv-cs-lg
4 Aug 2026
Model Releases

Uncertainty Is Not Enough: Value-of-Information Routing for Mixtures of LoRA Experts

DGX agent

arXiv:2608.02528v1 Announce Type: new Abstract: Mixtures of low-rank adaptation experts increase parameter-efficient capacity by routing each input through a subset of adapters. Recent dynamic routers

model-releasesarxiv-cs-lg
4 Aug 2026
Safety

Understanding and Correcting Low-Frequency Bias in EEG Foundation Model

DGX agent

arXiv:2608.01898v1 Announce Type: new Abstract: Increasing EEG pretraining data scale or model capacity does not consistently improve downstream performance. We identify a persistent low-frequency bia

safetyarxiv-cs-lg
4 Aug 2026
Research

UOT-IR: Structured Routing of High-Polyphony Symbolic Music into Fixed-Budget Representations

DGX agent

arXiv:2608.00576v1 Announce Type: cross Abstract: High-polyphony symbolic music is increasingly used in generation, analysis, and arrangement, yet many downstream tasks require bounded representations

researcharxiv-cs-lg
4 Aug 2026
Model Releases

UpliftBench: Revealing Outcome-Regime and Objective Mismatch in Uplift Evaluation

DGX agent

arXiv:2608.00915v1 Announce Type: new Abstract: Uplift modeling (conditional-average-treatment-effect estimation) drives personalized targeting, yet published uplift benchmarks frequently disagree on

model-releasesarxiv-cs-lg
4 Aug 2026
Safety

Upper-Expectile Multi-Step Q-Learning for Off-Policy Reinforcement Learning

DGX agent

arXiv:2608.02034v1 Announce Type: new Abstract: Multi-step returns accelerate reward propagation in off-policy reinforcement learning, but couple the evaluation of each decision to the suboptimal logg

safetyarxiv-cs-lg
4 Aug 2026
Applications

Using Lower-Bound Representations for Trajectory Similarity Learning

DGX agent

arXiv:2608.01039v1 Announce Type: cross Abstract: Trajectory similarity learning is fundamental to efficient trajectory retrieval under complex distance measures. Existing learning-based methods typic

applicationsarxiv-cs-lg
4 Aug 2026
Research

Using Non-Lipschitz Signum-based Functions for Distributed Optimization and Machine Learning: Trade-off Between Con-vergence Rate and Optimality Gap

DGX agent

arXiv:2608.01220v1 Announce Type: cross Abstract: In recent years, the prevalence of large-scale data-sets and the demand for sophisti-cated learning models have necessitated the development of effici

researcharxiv-cs-lg
4 Aug 2026
Safety

Wasserstein mixing time of the unadjusted Langevin algorithm

DGX agent

arXiv:2608.02430v1 Announce Type: cross Abstract: We provide new estimates in Wasserstein distance for the asymptotic bias of the unadjusted Langevin algorithm, in the classical setting of log-smooth

safetyarxiv-cs-lg
4 Aug 2026
Applications

WHALE: A Scalable Unified Model for Recommendation with Wukong-HSTU Architecture

DGX agent

arXiv:2607.17017v2 Announce Type: replace-cross Abstract: As scalability becomes increasingly important in recommendation modeling, recent architectures have advanced the modeling of two broad sources

applicationsarxiv-cs-lg
4 Aug 2026
Agents

What Could the Agent See at 19:05? Generating Temporal Enterprise Scenarios from Real Research and Replaying Them to Evaluate Agents

DGX agent

arXiv:2608.01042v1 Announce Type: cross Abstract: Enterprise AI agents act across many apps whose data changes continuously, so an answer is correct only relative to what data existed and who could se

agentsarxiv-cs-lg
4 Aug 2026
Agents

When Collaboration Becomes a Trigger: Collective Evidence-Threshold Backdoors in Multi-Agent Systems

DGX agent

arXiv:2608.01085v1 Announce Type: cross Abstract: LLM-based multi-agent systems (MAS) extend LLM capabilities through iterative communication and shared contexts. However, this collaboration introduce

agentsarxiv-cs-lg
4 Aug 2026
Tutorials

When Differential Privacy Meets Wireless Federated Learning: An Improved Analysis for Privacy and Convergence

DGX agent

arXiv:2603.19040v2 Announce Type: replace Abstract: Differentially private wireless federated learning (DPWFL) is a promising framework for protecting sensitive user data. However, foundational questi

tutorialsarxiv-cs-lg
4 Aug 2026
Research

When Do Surrogate Updates Improve Decisions? A Local Theory of Trajectory-Wise Transfer

DGX agent

arXiv:2608.01130v1 Announce Type: new Abstract: A broad range of models face the mismatch where they are updated through trajectory losses but are evaluated by downstream task reward. Here, a trajecto

researcharxiv-cs-lg
4 Aug 2026
Safety

When May a Model Replace the Experiment? Audits, Licenses, and the Price of Trust in Surrogate-Driven Design

DGX agent

arXiv:2608.01378v1 Announce Type: new Abstract: Design campaigns in chemistry, materials science, and machine learning share a bottleneck: determining how good a candidate truly is requires an expensi

safetyarxiv-cs-lg
4 Aug 2026
Research

When Replanning Becomes the Bottleneck: Budgeted Replanning for Embodied Agents

DGX agent

arXiv:2608.01428v1 Announce Type: cross Abstract: Embodied agents replan frequently to recover from execution drift, partial observability, and coordination hazards, but each LLM-based replanning call

researcharxiv-cs-lg
4 Aug 2026
Model Releases

Who Belongs in the Eval Set? A Capability-Taxonomy-Driven Pipeline for Curating Regression Eval Sets in Agent-Extensibility Platforms

DGX agent

arXiv:2608.01004v1 Announce Type: new Abstract: Platform teams hosting agent-extensibility surfaces face a regression-economics paradox: every onboarding customer ships an evaluation set tuned to thei

model-releasesarxiv-cs-lg
4 Aug 2026
Model Releases

Why Formal Monitors Fail: Attack Distribution Entropy as a Coverage Bound for LTL-Based LLM Agent Safety

DGX agent

arXiv:2608.01388v1 Announce Type: cross Abstract: Runtime safety monitors based on Linear Temporal Logic (LTL) and finite automata (FSA) are increasingly deployed to intercept unsafe tool-call sequenc

model-releasesarxiv-cs-lg
4 Aug 2026
← Previous
1…2324252627…301
Next →