AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “hardware”

GridTimelineEvolution
2,627 results
Hardware

PermuQuant: Lowering Per-Group Quantization Error by Reordering Channels for Diffusion Models

DGX agent

arXiv:2605.09503v1 Announce Type: new Abstract: Large-scale visual generative models have achieved remarkable performance. However, their high computational and memory costs make deployment challengin

hardwarearxiv-cs-cv
12 May 2026
Hardware
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

SkillEvolver: Skill Learning as a Meta-Skill

DGX agent

arXiv:2605.10500v1 Announce Type: new Abstract: Agent skills today are static artifact: authored once -- by human curation or one-shot generation from parametric knowledge -- and then consumed unchang

hardwarearxiv-cs-ai
12 May 2026
Hardware

SwiftI2V: Efficient High-Resolution Image-to-Video Generation via Conditional Segment-wise Generation

DGX agent

arXiv:2605.06356v2 Announce Type: replace Abstract: High-resolution image-to-video (I2V) generation aims to synthesize realistic temporal dynamics while preserving fine-grained appearance details of t

hardwarearxiv-cs-cv
12 May 2026
Hardware

An Efficient Hybrid Sparse Attention with CPU-GPU Parallelism for Long-Context Inference

DGX agent

arXiv:2605.07719v1 Announce Type: cross Abstract: Long-context inference increasingly operates over CPU-resident KV caches, either because decoding-time KV states exceed GPU memory capacity or because

hardwarearxiv-cs-ai
11 May 2026
Hardware

Closed-Form Linear-Probe Dataset Distillation for Pre-trained Vision Models

DGX agent

arXiv:2605.07194v1 Announce Type: cross Abstract: Dataset distillation compresses a large training set into a small synthetic set that preserves downstream training utility. While most existing method

hardwarearxiv-cs-ai
11 May 2026
Research

Code Generation and Conic Constraints for Model-Predictive Control on Microcontrollers with Conic-TinyMPC

DGX agent

arXiv:2403.18149v3 Announce Type: replace Abstract: Model-predictive control (MPC) is a state-of-the-art control method for constrained robotic systems, yet deployment on resource-limited hardware rem

researcharxiv-cs-ro
11 May 2026
Hardware

Direction-Preserving Number Representations

DGX agent

arXiv:2605.07662v1 Announce Type: new Abstract: Low-precision number formats are widely used in modern machine learning systems due to their efficiency. Accurate direction representation is key to the

hardwarearxiv-cs-lg
11 May 2026
Hardware

Don't Learn the Shape: Forecasting Periodic Time Series by Rank-1 Decomposition

DGX agent

arXiv:2605.07222v1 Announce Type: new Abstract: How few parameters do we really need to forecast a periodic time series? An hourly electricity series, reshaped as a 24-row matrix with one column per d

hardwarearxiv-cs-lg
11 May 2026
Hardware

MemSearcher: Training LLMs to Reason, Search and Manage Memory via End-to-End Reinforcement Learning

DGX agent

arXiv:2511.02805v2 Announce Type: replace-cross Abstract: LLM-based search agents often concatenate the full interaction history into the context, producing long and noisy inputs, and increasing compu

hardwarearxiv-cs-ai
11 May 2026
Hardware

Physics-Based Flow Matching for Full-Field Prediction of Silicon Photonic Devices

DGX agent

arXiv:2605.06929v1 Announce Type: cross Abstract: Designing photonic integrated circuits requires accurate electromagnetic field simulations, which remain computationally expensive even for simple dev

hardwarearxiv-cs-lg
11 May 2026
Hardware

RateQuant: Optimal Mixed-Precision KV Cache Quantization via Rate-Distortion Theory

DGX agent

arXiv:2605.06675v1 Announce Type: cross Abstract: Large language models cache all previously computed key-value (KV) pairs during generation, and this KV cache grows linearly with sequence length, mak

hardwarearxiv-cs-cl
11 May 2026
Hardware

Render, Don't Decode: Weight-Space World Models with Latent Structural Disentanglement

DGX agent

arXiv:2605.06298v2 Announce Type: replace-cross Abstract: Training world models on vast quantities of unlabelled videos is a critical step toward fully autonomous intelligence. However, the prevailing

hardwarearxiv-cs-ai
11 May 2026
Hardware

Semantic-Aware Adaptive Visual Memory for Streaming Video Understanding

DGX agent

arXiv:2605.07897v1 Announce Type: cross Abstract: Online streaming video understanding requires models to process continuous visual inputs and respond to user queries in real time, where the unbounded

hardwarearxiv-cs-ai
11 May 2026
Hardware

SOCKET: SOft Collision Kernel EsTimator for Sparse Attention

DGX agent

arXiv:2602.06283v2 Announce Type: replace Abstract: Exploiting sparsity during long-context inference is key to scaling large language models, as attention dominates the cost of autoregressive decodin

hardwarearxiv-cs-lg
11 May 2026
Hardware

Sparser, Faster, Lighter Transformer Language Models

DGX agent

arXiv:2603.23198v2 Announce Type: replace-cross Abstract: Scaling autoregressive large language models (LLMs) has driven unprecedented progress but comes with vast computational costs. In this work, w

hardwarearxiv-cs-cl
11 May 2026
Hardware

Towards Billion-scale Multi-modal Biometric Search

DGX agent

arXiv:2605.07655v1 Announce Type: cross Abstract: Searching a multi-biometric database of a billion records for a country-level identity system requires pushing the limits of all aspects of a biometri

hardwarearxiv-cs-ai
11 May 2026
Hardware

Zero-Shot Neural Network Evaluation with Sample-Wise Activation Patterns

DGX agent

arXiv:2605.07378v1 Announce Type: new Abstract: Zero-shot proxies, also known as training-free metrics, are widely adopted to reduce the computational overhead in neural network evaluation for scenari

hardwarearxiv-cs-lg
11 May 2026
Hardware

A Physics-Aware Framework for Short-Term GPU Power Forecasting of AI Data Centers

DGX agent

arXiv:2605.04074v1 Announce Type: new Abstract: AI data centers experience rapid fluctuations in power demand due to the heterogeneity of computational tasks that they have to support. For example, th

hardwarearxiv-cs-lg
7 May 2026
Hardware

A Queueing-Theoretic Framework for Stability Analysis of LLM Inference with KV Cache Memory Constraints

DGX agent

arXiv:2605.04595v1 Announce Type: new Abstract: The rapid adoption of large language models (LLMs) has created significant challenges for efficient inference at scale. Unlike traditional workloads, LL

hardwarearxiv-cs-lg
7 May 2026
Hardware

Budget-aware Auto Optimizer Configurator

DGX agent

arXiv:2605.04711v1 Announce Type: cross Abstract: Optimizer states occupy massive GPU memory in large-scale model training. However, gradients in different network blocks exhibit distinct behaviors, s

hardwarearxiv-cs-lg
7 May 2026
Hardware

CuBridge: An LLM-Based Framework for Understanding and Reconstructing High-Performance Attention Kernels

DGX agent

arXiv:2605.05023v1 Announce Type: new Abstract: Efficient CUDA implementations of attention mechanisms are critical to modern deep learning systems, yet supporting diverse and evolving attention varia

hardwarearxiv-cs-lg
7 May 2026
Hardware

Deep Wave Network for Modeling Multi-Scale Physical Dynamics

DGX agent

arXiv:2605.04198v1 Announce Type: new Abstract: Performance of deep learning models is strongly governed by architectural capacity, with width and depth as primary controls. However, in physical-scien

hardwarearxiv-cs-lg
7 May 2026
Hardware

DualTCN: A Physics-Constrained Temporal Convolutional Network for 2 Time-Domain Marine CSEM Inversion

DGX agent

arXiv:2605.04997v1 Announce Type: new Abstract: DualTCN is the first deep-learning framework for inverting time-domain marine controlled-source electromagnetic (MCSEM) transient data. Moving away from

hardwarearxiv-cs-lg
7 May 2026
Hardware

Massively Parallel Exact Inference for Hawkes Processes

DGX agent

arXiv:2604.01342v2 Announce Type: replace Abstract: Multivariate Hawkes processes are a widely used class of self-exciting point processes, but maximum likelihood estimation naively scales as O(N^2) i

hardwarearxiv-cs-lg
7 May 2026
Hardware

OptiLookUp: An Optical ROM-Based Loop up Table Engine for Photonic Accelerators

DGX agent

arXiv:2605.03241v1 Announce Type: cross Abstract: Read-only memory (ROM) provides deterministic access to predefined data mappings. Extending ROM concepts to the optical domain enables high-bandwidth,

hardwarearxiv-cs-ai
7 May 2026
Hardware

Quadrature-TreeSHAP: Depth-Independent TreeSHAP and Shapley Interactions

DGX agent

arXiv:2605.04497v1 Announce Type: new Abstract: Shapley values are a standard tool for explaining predictions of tree ensembles, with Path-Dependent SHAP being the most widely used variant. Despite su

hardwarearxiv-cs-lg
7 May 2026
Hardware

SemiConLens: Visual Analytics for 2D Semiconductor Discovery

DGX agent

arXiv:2605.04067v1 Announce Type: cross Abstract: The past few years have witnessed vibrant efforts in discovering new two-dimensional (2D) semiconductor materials from both academia and the industry,

hardwarearxiv-cs-lg
7 May 2026
Research

DARTH-PUM: A Hybrid Processing-Using-Memory Architecture

DGX agent

arXiv:2602.16075v2 Announce Type: replace-cross Abstract: Analog processing-using-memory (PUM; a.k.a. in-memory computing) makes use of electrical interactions inside memory arrays to perform bulk mat

researcharxiv-cs-lg
6 May 2026
Hardware

Learning Zero-Shot Subject-Driven Video Generation Using 1% Compute

DGX agent

arXiv:2504.17816v3 Announce Type: replace Abstract: Subject-driven video generation (SDV-Gen) aims to produce videos of a specific subject by adapting a pretrained video model, enabling personalized a

hardwarearxiv-cs-cv
6 May 2026
Hardware

MOOSE-Star: Unlocking Tractable Training for Scientific Discovery by Breaking the Complexity Barrier

DGX agent

arXiv:2603.03756v3 Announce Type: replace-cross Abstract: While large language models (LLMs) show promise in scientific discovery, existing research focuses on inference or feedback-driven training, l

hardwarearxiv-cs-cl
6 May 2026
Hardware

One Sequence to Segment Them All: Efficient Data Augmentation for CT and MRI Cross-Domain 3D Spine Segmentation

DGX agent

arXiv:2605.03098v1 Announce Type: new Abstract: Deep learning-based medical image segmentation is increasingly used to support clinical diagnosis and develop new treatment strategies. However, model p

hardwarearxiv-cs-cv
6 May 2026
Hardware

Robust Visual SLAM for UAV Navigation in GPS-Denied and Degraded Environments: A Multi-Paradigm Evaluation and Deployment Study

DGX agent

arXiv:2605.03678v1 Announce Type: new Abstract: Reliable localization in GPS-denied, visually degraded environments is critical for autonomous UAV opera- tions. This paper presents a systematic compar

hardwarearxiv-cs-ro
6 May 2026
Hardware

Test-Time Training with KV Binding Is Secretly Linear Attention

DGX agent

arXiv:2602.21204v3 Announce Type: replace-cross Abstract: Test-time training (TTT) with KV binding as sequence modeling layer is commonly interpreted as a form of online meta-learning that memorizes a

hardwarearxiv-cs-cv
6 May 2026
Hardware

The Polar Express: Optimal Matrix Sign Methods and Their Application to the Muon Algorithm

DGX agent

arXiv:2505.16932v5 Announce Type: replace-cross Abstract: Computing the polar decomposition and the related matrix sign function has been a well-studied problem in numerical analysis for decades. Rece

hardwarearxiv-cs-cl
6 May 2026
Hardware

ZeRO-Prefill: Zero Redundancy Overheads in MoE Prefill Serving

DGX agent

arXiv:2605.02960v1 Announce Type: new Abstract: Production LLM workloads increasingly serve discriminative tasks, such as classification, recommendation, and verification, whose answers are read from

hardwarearxiv-cs-lg
6 May 2026
Hardware

AgriKD: Cross-Architecture Knowledge Distillation for Efficient Leaf Disease Classification

DGX agent

arXiv:2605.01355v1 Announce Type: new Abstract: Automated leaf disease classification is critical for early disease detection in resource-constrained field environments. Vision Transformers (ViTs) pro

hardwarearxiv-cs-cv
5 May 2026
Hardware

Can AI Debias the News? LLM Interventions Improve Cross-Partisan Receptivity but LLMs Overestimate Their Own Effectiveness

DGX agent

arXiv:2605.01006v1 Announce Type: new Abstract: Partisan news media erode cross-partisan trust, but large language models (LLMs) offer a potential means of debiasing such content at scale. Across two

hardwarearxiv-cs-cl
5 May 2026
Hardware

DELTA: Dynamic Layer-Aware Token Attention for Efficient Long-Context Reasoning

DGX agent

arXiv:2510.09883v2 Announce Type: replace Abstract: Large reasoning models (LRMs) achieve state-of-the-art performance on challenging benchmarks by generating long chains of intermediate steps, but th

hardwarearxiv-cs-cl
5 May 2026
Hardware

Effective Capacitance Modeling Using Graph Neural Networks

DGX agent

arXiv:2507.03787v2 Announce Type: replace Abstract: Static timing analysis is a crucial stage in the VLSI design flow that verifies the timing correctness of circuits. Timing analysis depends on the p

hardwarearxiv-cs-lg
5 May 2026
Hardware

Fast Log-Domain Sinkhorn Optimal Transport with Warp-Level GPU Reductions

DGX agent

arXiv:2605.00837v1 Announce Type: new Abstract: Entropic regularized optimal transport (OT) via the Sinkhorn algorithm has become a fundamental tool in machine learning, yet existing implementations e

hardwarearxiv-cs-lg
5 May 2026
Hardware

InfiniteDiffusion: Bridging Learned Fidelity and Procedural Utility for Open-World Terrain Generation

DGX agent

arXiv:2512.08309v4 Announce Type: replace Abstract: For decades, procedural worlds have been built on procedural noise functions such as Perlin noise, which are fast and infinite, yet fundamentally li

hardwarearxiv-cs-cv
5 May 2026
Hardware

LEAP: Layer-wise Exit-Aware Pretraining for Efficient Transformer Inference

DGX agent

arXiv:2605.01058v1 Announce Type: cross Abstract: Layer-aligned distillation and convergence-based early exit represent two predominant computational efficiency paradigms for transformer inference; ye

hardwarearxiv-cs-cl
5 May 2026
Hardware

LLM Output Detectability and Task Performance Can be Jointly Optimized

DGX agent

arXiv:2605.01350v1 Announce Type: new Abstract: Detecting machine-generated text is essential for transparency and accountability when deploying large language models (LLMs). Among detection approache

hardwarearxiv-cs-cl
5 May 2026
Hardware

Maistros: A Greek Large Language Model Adapted Through Knowledge Distillation From Large Reasoning Models

DGX agent

arXiv:2605.01870v1 Announce Type: new Abstract: Large Language Models (LLMs) have substantially advanced the field of Natural Language Processing (NLP), achieving state-of-the-art performance across a

hardwarearxiv-cs-cl
5 May 2026
Hardware

MusicInfuser: Making Video Diffusion Listen and Dance

DGX agent

arXiv:2503.14505v3 Announce Type: replace Abstract: We introduce MusicInfuser, an approach that aligns pre-trained text-to-video diffusion models to generate high-quality dance videos synchronized wit

hardwarearxiv-cs-cv
5 May 2026
Hardware

Need for Speed: Zero-Shot Depth Completion with Single-Step Diffusion

DGX agent

arXiv:2603.10584v2 Announce Type: replace Abstract: We introduce Marigold-SSD, a single-step, late-fusion depth completion framework that leverages strong diffusion priors while eliminating the costly

hardwarearxiv-cs-cv
5 May 2026
Hardware

Rethinking Network Topologies for Cost-Effective Mixture-of-Experts LLM Serving

DGX agent

arXiv:2605.00254v1 Announce Type: cross Abstract: Mixture-of-experts (MoE) architectures have turned LLM serving into a cluster-scale workload in which communication consumes a considerable portion of

hardwarearxiv-cs-ai
5 May 2026
Hardware

SplitZip: Ultra Fast Lossless KV Compression for Disaggregated LLM Serving

DGX agent

arXiv:2605.01708v1 Announce Type: cross Abstract: Contemporary systems serving large language models (LLMs) have adopted prefill-decode disaggregation to better load-balance between the compute-bound

hardwarearxiv-cs-lg
5 May 2026
← Previous
1…1718192021…55
Next →