AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,557 results
Research

SWaRL: Safeguard Code Watermarking via Reinforcement Learning

DGX agent

arXiv:2601.02602v2 Announce Type: replace-cross Abstract: We present SWaRL, a robust and fidelity-preserving watermarking framework designed to protect the intellectual property of code LLMs by embedd

researcharxiv-cs-lg
11 May 2026
Research
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Synergistic Benefits of Joint Molecule Generation and Property Prediction

DGX agent

arXiv:2504.16559v3 Announce Type: replace Abstract: Modeling the joint distribution of data samples and their properties allows to construct a single model for both data generation and property predic

researcharxiv-cs-lg
11 May 2026
Model Releases

Target-Aware Data Augmentation for SAT Prediction

DGX agent

arXiv:2605.06931v1 Announce Type: new Abstract: Learning-based approaches to NP-hard problems have shown increasing promise, but their progress is fundamentally constrained by the high cost of generat

model-releasesarxiv-cs-lg
11 May 2026
Safety

Temporal Attention for Adaptive Control of Euler-Lagrange Systems with Unobservable Memory

DGX agent

arXiv:2605.06877v1 Announce Type: new Abstract: Adaptive control of Euler-Lagrange systems is challenging when friction is governed by a finite-horizon internal state that is not directly observable f

safetyarxiv-cs-lg
11 May 2026
Research

Tessellations of Semi-Discrete Flow Matching

DGX agent

arXiv:2605.07513v1 Announce Type: new Abstract: We study Flow Matching in a semi-discrete setting where a Gaussian source is transported toward a discrete target supported on finitely many points. Thi

researcharxiv-cs-lg
11 May 2026
Local Ai

Test-Time Compositional Generalization in Diffusion Models via Concept Discovery

DGX agent

arXiv:2605.07078v1 Announce Type: new Abstract: Compositional generalization requires models to produce novel configurations from familiar parts. In diffusion models, prior compositional generation me

local-aiarxiv-cs-lg
11 May 2026
Research

Testing Noise Assumptions of Learning Algorithms

DGX agent

arXiv:2501.09189v3 Announce Type: replace Abstract: We pose a fundamental question in computational learning theory: can we efficiently test whether a training set satisfies the assumptions of a given

researcharxiv-cs-lg
11 May 2026
Model Releases

The Convergence Gap: Instruction-Tuned Language Models Stabilize Later in the Forward Pass

DGX agent

arXiv:2605.07282v1 Announce Type: new Abstract: Final outputs hide when a checkpoint commits to its next-token prediction. We introduce the convergence gap, a model-diffing diagnostic that decodes eac

model-releasesarxiv-cs-lg
11 May 2026
Model Releases

The Coupling Tax: How Shared Token Budgets Undermine Visible Chain-of-Thought Under Fixed Output Limits

DGX agent

arXiv:2605.07686v1 Announce Type: new Abstract: Chain-of-thought reasoning is often treated as a monotone way to improve language-model accuracy by letting a model think longer. We identify a counterv

model-releasesarxiv-cs-lg
11 May 2026
Research

The Minimax Rate of Second-Order Calibration

DGX agent

arXiv:2605.07808v1 Announce Type: new Abstract: We characterize the minimax rate of estimating the second-order calibration error for binary classification, which quantifies whether a higher-order pre

researcharxiv-cs-lg
11 May 2026
Model Releases

Theory of Optimal Learning Rate Schedules and Scaling Laws for a Random Feature Model

DGX agent

arXiv:2602.04774v2 Announce Type: replace-cross Abstract: Setting the learning rate (LR) for a deep learning model is a critical part of successful training. Choosing LRs is often done empirically wit

model-releasesarxiv-cs-lg
11 May 2026
Model Releases

ThinKV: Thought-Adaptive KV Cache Compression for Efficient Reasoning Models

DGX agent

arXiv:2510.01290v2 Announce Type: replace Abstract: The long-output context generation of large reasoning models enables extended chain of thought (CoT) but also drives rapid growth of the key-value (

model-releasesarxiv-cs-lg
11 May 2026
Safety

Toward Better Geometric Representations for Molecule Generative Models

DGX agent

arXiv:2605.07693v1 Announce Type: new Abstract: Geometric representation-conditioned molecule generation provides an effective paradigm that decouples molecule representation modeling from structure g

safetyarxiv-cs-lg
11 May 2026
Safety

TRACE: Transport Alignment Conformal Prediction via Diffusion and Flow Matching Models

DGX agent

arXiv:2605.07100v1 Announce Type: cross Abstract: Constructing valid and informative conformal prediction regions for multi-dimensional outputs remains a fundamental challenge. While conformal predict

safetyarxiv-cs-lg
11 May 2026
Model Releases

Training-Induced Escape from Token Clustering in a Mean-Field Formulation of Transformers

DGX agent

arXiv:2605.07772v1 Announce Type: new Abstract: Transformers perform inference by iteratively transforming token representations across layers. This layerwise computation has been studied empirically,

model-releasesarxiv-cs-lg
11 May 2026
Tutorials

Transfer Learning Across Fast- and Full-Simulation Domains in High-Energy Physics

DGX agent

arXiv:2605.07471v1 Announce Type: new Abstract: Machine-learning models in high-energy physics are often trained on simulated data, where fully simulated samples are computationally expensive while fa

tutorialsarxiv-cs-lg
11 May 2026
Research

Transformer-Based Wildlife Species Classification from Daily Movement Trajectories

DGX agent

arXiv:2605.06726v1 Announce Type: new Abstract: Inferring the identity of wildlife species from daily movement data alone is a challenging task. We train sequence models on large-scale, 7-species GPS

researcharxiv-cs-lg
11 May 2026
Applications

TraXion: Rethinking Pre-training Frameworks for Mobility and Beyond

DGX agent

arXiv:2605.06906v1 Announce Type: new Abstract: Human mobility differs from text and from generic time series in three structural ways: visits are tuple-valued events whose meaning depends on the join

applicationsarxiv-cs-lg
11 May 2026
Tutorials

Tree SAE: Learning Hierarchical Feature Structures in Sparse Autoencoders

DGX agent

arXiv:2605.07922v1 Announce Type: new Abstract: Learning hierarchical features in Sparse Autoencoders (SAEs) is essential for capturing the structured nature of real-world data and mitigating issues l

tutorialsarxiv-cs-lg
11 May 2026
Research

TUANDROMD-X: Advanced Entropy and Visual Analytics Dataset for Enhanced Malware Detection and Classification

DGX agent

arXiv:2605.06718v1 Announce Type: cross Abstract: Malware and malware-based attacks are becoming more prevalent and complex. Attackers regularly come up with new techniques that have the ability to ev

researcharxiv-cs-lg
11 May 2026
Research

Tyche: One Step Flow for Efficient Probabilistic Weather Forecasting

DGX agent

arXiv:2605.06916v1 Announce Type: new Abstract: Probabilistic weather forecasting requires not only accurate trajectories, but calibrated distributions over plausible atmospheric futures. Recent data-

researcharxiv-cs-lg
11 May 2026
Tutorials

Uncovering Hidden Systematics in Neural Network Models for High Energy Physics

DGX agent

arXiv:2605.07470v1 Announce Type: new Abstract: Neural networks (NNs) are inherently multidimensional classifiers that learn complex, non-linear relationships among input observables. While their flex

tutorialsarxiv-cs-lg
11 May 2026
Model Releases

Understanding Robustness of Model Editing in Code LLMs

DGX agent

arXiv:2511.03182v2 Announce Type: replace-cross Abstract: Large language models (LLMs) for code are increasingly used in software development, but they remain static after pretraining while APIs and s

model-releasesarxiv-cs-lg
11 May 2026
Research

Upper Generalization Bounds for Neural Oscillators

DGX agent

arXiv:2603.09742v2 Announce Type: replace Abstract: Neural oscillators that originate from second-order ordinary differential equations (ODEs) have shown competitive performance in learning mappings b

researcharxiv-cs-lg
11 May 2026
Hardware

Versatile yet Efficient Network Traffic Analysis: Offloading Network Foundation Model to SmartNIC

DGX agent

arXiv:2508.02001v2 Announce Type: replace-cross Abstract: Pervasive encryption makes large-scale labeling infeasible for traffic analysis, while security operations demand edge analysis to avert servi

hardwarearxiv-cs-lg
11 May 2026
Research

VNN-LIB 2.0: Rigorous Foundations for Neural Network Verification

DGX agent

arXiv:2605.07451v1 Announce Type: new Abstract: Neural network verification is an active and rapidly maturing research area, with a growing ecosystem of solvers and tools. The VNN-LIB standard was int

researcharxiv-cs-lg
11 May 2026
Safety

When Descent Is Too Stable: Event-Triggered Hamiltonian Learning to Optimize

DGX agent

arXiv:2605.06868v1 Announce Type: new Abstract: Fixed-budget nonconvex optimization can fail not because local descent is unstable, but because it is too stable: after reaching a nearby stationary poi

safetyarxiv-cs-lg
11 May 2026
Research

When Diffusion Model Can Ignore Dimension: An Entropy-Based Theory

DGX agent

arXiv:2605.07969v1 Announce Type: new Abstract: Diffusion models perform remarkably well on high-dimensional data such as images, often using only a modest number of reverse-time steps. Despite this p

researcharxiv-cs-lg
11 May 2026
Research

When Does Embedding Magnitude Matter? A Cross-Task Functional-Symmetry Framework

DGX agent

arXiv:2602.09229v3 Announce Type: replace Abstract: Cosine similarity normalizes both sides; dot product normalizes neither. We propose a 2x2 framework that independently controls query-side and docum

researcharxiv-cs-lg
11 May 2026
Tutorials

When Symbol Names Should Not Matter: A Logistic Theory of Fresh-Symbol Classification

DGX agent

arXiv:2605.07120v1 Announce Type: new Abstract: Template tasks have emerged as a clean testbed for asking whether transformers reason with abstract symbols rather than concrete token names. We study t

tutorialsarxiv-cs-lg
11 May 2026
Model Releases

Where to Spend Rollouts: Hit-Utility Optimal Rollout Allocation for Group-Based RLVR

DGX agent

arXiv:2605.07114v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) has emerged as a central paradigm for improving the reasoning capabilities of large language model

model-releasesarxiv-cs-lg
11 May 2026
Safety

Why Does Agentic Safety Fail to Generalize Across Tasks?

DGX agent

arXiv:2605.06992v1 Announce Type: new Abstract: AI agents are increasingly deployed in multi-task settings, where the task to perform is specified at test time, and the agent must generalize to unseen

safetyarxiv-cs-lg
11 May 2026
Applications

XDecomposer: Learning Prior-Free Set Decomposition for Multiphase X-ray Diffraction

DGX agent

arXiv:2605.05866v1 Announce Type: cross Abstract: Multiphase powder X-ray diffraction (PXRD) analysis remains a fundamental bottleneck in structure identification, as real-world synthesis often produc

applicationsarxiv-cs-lg
11 May 2026
Research

You Only Stack Once (YOSO): A Motion-Filtered, Deep-Learning Framework for Detecting Faint Moving Sources

DGX agent

arXiv:2605.06913v1 Announce Type: cross Abstract: We present You Only Stack Once (YOSO), an automated pipeline designed to detect faint, slow-moving Solar System objects in wide-field astronomical sur

researcharxiv-cs-lg
11 May 2026
Safety

Zero-Shot Imagined Speech Decoding via Imagined-to-Listened MEG Mapping

DGX agent

arXiv:2605.08075v1 Announce Type: new Abstract: Decoding imagined speech from non-invasive brain recordings is challenging because imagined datasets are scarce and difficult to align temporally across

safetyarxiv-cs-lg
11 May 2026
Hardware

Zero-Shot Neural Network Evaluation with Sample-Wise Activation Patterns

DGX agent

arXiv:2605.07378v1 Announce Type: new Abstract: Zero-shot proxies, also known as training-free metrics, are widely adopted to reduce the computational overhead in neural network evaluation for scenari

hardwarearxiv-cs-lg
11 May 2026
Model Releases

A Biased Nonnegative Block Term Tensor Decomposition Model for Dynamic QoS Prediction

DGX agent

arXiv:2605.04813v1 Announce Type: new Abstract: With the rapid development of cloud computing and Web services, Quality of Service (QoS) has become a key criterion for service selection and recommenda

model-releasesarxiv-cs-lg
7 May 2026
Applications

A Consistency-Centric Approach to Set-Based Optimization with Multiple Models of Unranked Fidelity

DGX agent

arXiv:2605.04051v1 Announce Type: cross Abstract: In complex real-world settings, optimization is challenged by the presence of diverse models of differing fidelity. In many optimization problems, a s

applicationsarxiv-cs-lg
7 May 2026
Applications

A Foundation Model for Zero-Shot Logical Rule Induction

DGX agent

arXiv:2605.04916v1 Announce Type: cross Abstract: Inductive Logic Programming (ILP) learns interpretable logical rules from data. Existing methods are transductive: their learned parameters are bound

applicationsarxiv-cs-lg
7 May 2026
Research

A foundation model of vision, audition, and language for in-silico neuroscience

DGX agent

arXiv:2605.04326v1 Announce Type: cross Abstract: Cognitive neuroscience is fragmented into specialized models, each tailored to specific experimental paradigms, hence preventing a unified model of co

researcharxiv-cs-lg
7 May 2026
Research

A geometric relation of the error introduced by sampling a language model's output distribution to its internal state

DGX agent

arXiv:2605.04899v1 Announce Type: new Abstract: GPT-style language models are sensitive to single-token changes at generation points where the predicted probability distribution is spread across multi

researcharxiv-cs-lg
7 May 2026
Research

A Harmonic Mean Formulation of Average Reward Reinforcement Learning in SMDPs

DGX agent

arXiv:2605.04880v1 Announce Type: new Abstract: Recent research has revived and amplified interest in algorithms for undiscounted average reward reinforcement learning in infinite-horizon, non-episodi

researcharxiv-cs-lg
7 May 2026
Tutorials

A Hybrid Quantum-Classical Framework for Financial Volatility Forecasting Based on Quantum Circuit Born Machines

DGX agent

arXiv:2603.09789v2 Announce Type: replace Abstract: Accurate financial volatility forecasting is crucial but challenged by the non-linear, highly correlated nature of market data. Recently, quantum co

tutorialsarxiv-cs-lg
7 May 2026
Applications

A Mean Curvature Approach to Boundary Detection: Geometric Insights for Unsupervised Learning

DGX agent

arXiv:2605.04274v1 Announce Type: new Abstract: Accurate boundary detection in high-dimensional data remains a central challenge in unsupervised learning, particularly in the presence of non-linear st

applicationsarxiv-cs-lg
7 May 2026
Hardware

A Physics-Aware Framework for Short-Term GPU Power Forecasting of AI Data Centers

DGX agent

arXiv:2605.04074v1 Announce Type: new Abstract: AI data centers experience rapid fluctuations in power demand due to the heterogeneity of computational tasks that they have to support. For example, th

hardwarearxiv-cs-lg
7 May 2026
Research

A Provably Convergent and Practical Algorithm for Gromov--Wasserstein Optimal Transport

DGX agent

arXiv:2605.04175v1 Announce Type: new Abstract: Gromov--Wasserstein optimal transport (GWOT) aligns metric measure spaces by matching their within-domain relational structures, but large-scale GWOT re

researcharxiv-cs-lg
7 May 2026
Hardware

A Queueing-Theoretic Framework for Stability Analysis of LLM Inference with KV Cache Memory Constraints

DGX agent

arXiv:2605.04595v1 Announce Type: new Abstract: The rapid adoption of large language models (LLMs) has created significant challenges for efficient inference at scale. Unlike traditional workloads, LL

hardwarearxiv-cs-lg
7 May 2026
Model Releases

A Regulatory Governance Framework for AI-Driven Financial Fraud Detection in U.S. Banking: Integrating OCC, SR 11-7, CFPB, and FinCEN Compliance Requirements for Model Development, Validation, and Monitoring Lifecycles

DGX agent

arXiv:2605.04076v1 Announce Type: new Abstract: U.S. financial institutions deploying AI-based fraud detection face a fragmented compliance landscape spanning four regulatory frameworks -- OCC Bulleti

model-releasesarxiv-cs-lg
7 May 2026
← Previous
1…225226227228229…304
Next →