AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlog
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “hardware”

GridTimelineEvolution
4,450 results
Hardware

Bake It Till You Make It: Ultrafast Spatial Texture-Atlas Splatting

DGX agent

arXiv:2607.13808v1 Announce Type: new Abstract: Recent extensions of 3D Gaussian Splatting (3DGS) capture fine color details using hash-grid-based appearance parameterization but incur high computatio

hardwarearxiv-cs-cv
16 Jul 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Hardware

BREAKING: Dylan Patel (@dylan522p) of @SemiAnalysis_ says 'chips in Europe have less seasoning than chips in America & Mexico.' Plus: › Data…

DGX agent

BREAKING: Dylan Patel (@dylan522p) of @SemiAnalysis_ says 'chips in Europe have less seasoning than chips in America & Mexico.' Plus: › Data centers & France’s nuclear power › AI infrastructure over-o

hardwaredylan-patel--x
9 Jul 2026
Hardware

InferNet: Exploiting Aggregate GPU Profiles as Side-Channel for DNN Architecture Inference

DGX agent

arXiv:2304.03388v2 Announce Type: replace Abstract: Deep Neural Networks (DNNs) have become ubiquitous for their ability to solve problems across various domains, including computer vision, natural la

hardwarearxiv-cs-lg
9 Jul 2026
Hardware

Latency-Constrained DNN Architecture Learning for Edge Systems using Zerorized Batch Normalization

DGX agent

arXiv:2607.06922v1 Announce Type: cross Abstract: Deep learning applications have been widely adopted on edge devices, to mitigate the privacy and latency issues of accessing cloud servers. Deciding t

hardwarearxiv-cs-cv
9 Jul 2026
Hardware

C4N, now GA: Delivering cloud’s highest per vCPU network and block storage I/O for x86 workloads

DGX agent

As organizations scale modern workloads — from high-throughput databases and network/security appliances to real-time analytics and AI/ML inference — network and block storage performance can quickly

hardwaregoogle-cloud-ai
8 Jul 2026
Hardware

A Multi-Task Deep Learning Framework for Real-Time Intelligent Video Surveillance with Temporal Event Validation

DGX agent

arXiv:2607.03131v1 Announce Type: cross Abstract: Modern video surveillance systems generate far more video streams than human operators can effectively monitor, making automated analysis essential fo

hardwarearxiv-cs-ai
7 Jul 2026
Hardware

Generative wave propagator

DGX agent

arXiv:2607.04440v1 Announce Type: cross Abstract: Seismic wavefield simulation is fundamental to seismology, but conventional finite-difference (FD) methods remain limited by numerical dispersion and

hardwarearxiv-cs-ai
7 Jul 2026
Hardware

HyperVAttention: Efficient Sparse Attention with Spatio-Temporal Clustering for Video Diffusion

DGX agent

arXiv:2607.03012v1 Announce Type: cross Abstract: Video Diffusion Transformers (VDiTs) have demonstrated significant capabilities in high-fidelity video generation. However, their ability to produce l

hardwarearxiv-cs-ai
7 Jul 2026
Hardware

Scaling Weisfeiler-Leman Expressiveness Analysis to Massive Graphs with GPUs

DGX agent

arXiv:2607.02603v1 Announce Type: cross Abstract: The stable coloring of the Weisfeiler-Leman (1-WL) test is a cornerstone of Graph Neural Networks because it provides an upper bound to the expressive

hardwarearxiv-cs-lg
7 Jul 2026
Hardware

The S-ICDF Dataset: Sionna-Simulated Dynamic Interference Characterization and Direction Finding

DGX agent

arXiv:2607.03411v1 Announce Type: cross Abstract: Jamming and spoofing threaten wireless and satellite navigation by disrupting or manipulating radio frequency (RF) signals, undermining availability,

hardwarearxiv-cs-ai
7 Jul 2026
Hardware

Wan-Streamer v0.2: Higher Resolution, Same Latency

DGX agent

arXiv:2607.04443v1 Announce Type: cross Abstract: We present Wan-Streamer v0.2, a latency-preserving upgrade of the native-streaming, end-to-end audio-visual interaction model. v0.2 keeps the v0.1 mod

hardwarearxiv-cs-ai
7 Jul 2026
Hardware

183/365 of GPU Programming This 4.5 hour lesson on CUDA + ThunderKittens by @bfspector (TK co-author, Stanford PhD student) is one of the be…

DGX agent

183/365 of GPU Programming This 4.5 hour lesson on CUDA + ThunderKittens by @bfspector (TK co-author, Stanford PhD student) is one of the best educational videos on kernels out there (think @karpathy

hardwareswyx--x
5 Jul 2026
Hardware

DeadPool: Resilient LLM Training with Hot-Swapping via Zero-Overhead Checkpoint

DGX agent

arXiv:2607.01646v1 Announce Type: new Abstract: State-of-the-art large language model (LLM) training takes tens of thousands of graphics processing units (GPUs) for months and encounters failures acro

hardwarearxiv-cs-lg
3 Jul 2026
Hardware

GPUAlert: A Zero-Instrumentation Process-Boundary Monitor for Diagnosing GPU Training-Job Failures

DGX agent

arXiv:2607.01409v1 Announce Type: cross Abstract: GPU training jobs fail often, roughly two in five on large production clusters, yet the operator typically learns of a failure only by reconnecting ho

hardwarearxiv-cs-ai
3 Jul 2026
Hardware

“Within the next 18 months, you will be able to host GLM 5.2 equivalent intelligence on an RTX 5090 GPU.” -Ahmad Osman, AI World’s Fair

DGX agent

Ahmad Osman stated at AI World's Fair that within 18 months, GLM 5.2-equivalent AI intelligence will be deployable locally on a single RTX 5090 GPU, indicating rapid progress toward running advanced l

hardwareswyx--x
2 Jul 2026
Hardware

AgRefactor: Self-Evolving Agentic Workflow for HLS Compatibility and Performance

DGX agent

arXiv:2606.30949v1 Announce Type: new Abstract: High-Level Synthesis (HLS) provides a fast path from concepts to silicon, but converting real-world software into synthesizable HLS code remains challen

hardwarearxiv-cs-ai
1 Jul 2026
Hardware

Hierarchical Global Attention (HGA)

DGX agent

arXiv:2606.30709v1 Announce Type: cross Abstract: Hierarchical Global Attention (HGA) is a drop-in replacement for dense causal attention in pretrained long-context transformers. HGA preserves the ori

hardwarearxiv-cs-ai
1 Jul 2026
Hardware

America is about to lose the AI race, and it will not happen at the frontier. It will happen at the floor. Everyone is watching who ships th…

DGX agent

America is about to lose the AI race, and it will not happen at the frontier. It will happen at the floor. Everyone is watching who ships the smartest model. The actual war is over the 80% of tokens n

hardwareclem-delangue--x
27 Jun 2026
Hardware

EGG: An Expert-Guided Agent Framework for Kernel Generation

DGX agent

arXiv:2606.26758v1 Announce Type: new Abstract: High-performance GPU kernels are critical for reducing the exponentially growing computational costs of large language models (LLMs), but their developm

hardwarearxiv-cs-ai
26 Jun 2026
Hardware

https://huggingface.co/nvidia/GLM-5.2-NVFP4

DGX agent

NVIDIA's GLM-5.2-NVFP4 is a quantized version of a large language model optimized for inference efficiency using NVIDIA's proprietary quantization format. The model is hosted on Hugging Face and repre

hardwareclem-delangue--x
26 Jun 2026
Hardware

“If there is a deflation of the AI bubble, the optimists say that the new infrastructure will remain even if the companies do not — just as …

DGX agent

“If there is a deflation of the AI bubble, the optimists say that the new infrastructure will remain even if the companies do not — just as railways survived the 19th-century railway bust. However, th

hardwaregary-marcus--x
26 Jun 2026
Research

SOLAR: AI-Powered Speed-of-Light Performance Analysis

DGX agent

arXiv:2606.26383v1 Announce Type: cross Abstract: How fast could a deep-learning model run on target hardware, and how far is today's implementation from that limit? These questions are central to sof

researcharxiv-cs-ai
26 Jun 2026
Hardware

Accelerating Disaggregated RL for Visual Generative LLMs with Diffusion-Based Parallelism and Trainer-Assisted Generation

DGX agent

arXiv:2606.24369v1 Announce Type: new Abstract: Reinforcement learning (RL) has become a dominant post-training paradigm, driving the emergence of high-performance RL systems such as veRL for autoregr

hardwarearxiv-cs-ai
24 Jun 2026
Hardware

End-to-End Radar and Communication Modulation Recognition with Neuromorphic Computing

DGX agent

arXiv:2606.24075v1 Announce Type: cross Abstract: Although deep learning-based methods can achieve high accuracy in automatic modulation recognition (AMR) tasks, their high computational cost makes it

hardwarearxiv-cs-ai
24 Jun 2026
Hardware

OpenAI and Broadcom unveil LLM-optimized inference chip

DGX agent

OpenAI and Broadcom announced a collaboration on a specialized inference chip designed to optimize the performance and efficiency of large language model deployments. The chip, referred to as 'Jalapen

hardwareopenai
24 Jun 2026
Hardware

TurboMPC: Fast, Scalable, and Differentiable Model Predictive Control on the GPU

DGX agent

arXiv:2606.24039v1 Announce Type: new Abstract: Robotics increasingly relies on GPUs for parallel simulation, large-scale learning, and neural-network inference. For model predictive control (MPC) to

hardwarearxiv-cs-ro
24 Jun 2026
Hardware

CITADEL: CSI-Based Jamming Detection and Open-Set Classification for IIoT Networks

DGX agent

arXiv:2606.22939v1 Announce Type: cross Abstract: Radio frequency jamming poses a critical threat to the availability of wireless Industrial Internet of Things (IIoT) networks. Existing detection and

hardwarearxiv-cs-lg
23 Jun 2026
Hardware

Fast-TurboQuant: A Multiplier-Free Online Vector Quantization Approach

DGX agent

arXiv:2606.21448v1 Announce Type: new Abstract: As large language models scale, memory bandwidth for key-value caches and retrieval-augmented generation systems becomes a critical bottleneck. While 1-

hardwarearxiv-cs-lg
23 Jun 2026
Hardware

Real-World Deployment of Massively Parallel Sampling-Based MPC for Contact-Rich Manipulation

DGX agent

arXiv:2606.20712v1 Announce Type: new Abstract: Sampling-based Model Predictive Control (SMPC) is a promising strategy for contact-rich robotic manipulation, combining gradient-free optimization with

hardwarearxiv-cs-ro
23 Jun 2026
Hardware

SCENIC: Semantic-Conditioned Edge-Aware Neural Framework for Structured IoT Command Generation

DGX agent

arXiv:2606.22296v1 Announce Type: new Abstract: Edge Internet of Things (IoT) agents are often constrained by memory capacity, privacy requirements, communication latency, and recurring inference cost

hardwarearxiv-cs-lg
23 Jun 2026
Hardware

The Secret Life of Computers

DGX agent

This article likely explores the hidden operational processes and internal workings of computers that users typically don't observe, such as microprocessor functions, system-level operations, or the b

hardwarethe-chip-letter
22 Jun 2026
Hardware

MPK: A Compiler and Runtime for Mega-Kernelizing Tensor Programs

DGX agent

arXiv:2512.22219v2 Announce Type: replace-cross Abstract: We introduce Mirage Persistent Kernel (MPK), the first compiler and runtime system that automatically transforms multi-GPU model inference int

hardwarearxiv-cs-lg
11 Jun 2026
Hardware

SwiftVR: Real-Time One-Step Generative Video Restoration

DGX agent

arXiv:2606.09516v1 Announce Type: new Abstract: Real-time video restoration (VR) for live streams requires high-resolution outputs under strict per-frame latency constraints. Existing one-step diffusi

hardwarearxiv-cs-cv
9 Jun 2026
Hardware

Toward Compiler World Models: Learning Latent Dynamics for Efficient Tensor Program Search

DGX agent

arXiv:2606.09312v1 Announce Type: new Abstract: Tensor program optimization is essential for modern machine learning systems, but its search space is enormous. Existing auto-schedulers reduce measurem

hardwarearxiv-cs-lg
9 Jun 2026
Hardware

E2Former-V2: On-the-Fly Equivariant Attention with Linear Activation Memory

DGX agent

arXiv:2601.16622v2 Announce Type: replace-cross Abstract: Equivariant Graph Neural Networks (EGNNs) have become a widely used approach for modeling 3D atomistic systems. However, mainstream architectu

hardwarearxiv-cs-ai
8 Jun 2026
Research

ITP-STDP: An Intrinsic-Timing Power-of-Two Learning Engine for On-Chip SNN Training

DGX agent

arXiv:2606.06159v1 Announce Type: cross Abstract: Spiking neural networks (SNNs) have the potential to emerge as the third generation of neural networks and have attracted increasing attention across

researcharxiv-cs-ai
6 Jun 2026
Hardware

Accelerating and Scaling MPC-Guided Reinforcement Learning for Humanoid Locomotion and Manipulation

DGX agent

arXiv:2606.05687v1 Announce Type: new Abstract: In humanoid motion control, model predictive control (MPC) offers physically grounded prediction and constraint handling, while reinforcement learning (

hardwarearxiv-cs-ro
5 Jun 2026
Hardware

Staged Factorial Screening for Budget-Constrained Micro-Pretraining

DGX agent

arXiv:2606.05186v1 Announce Type: cross Abstract: Budget-constrained micro-pretraining often requires triaging many candidate recipes on a shared accelerator before larger search budgets are spent. We

hardwarearxiv-cs-cl
5 Jun 2026
Hardware

Towards Realistic 3D Sonar Simulation

DGX agent

arXiv:2606.06130v1 Announce Type: new Abstract: As underwater robotics research increasingly addresses complex 3D perception and autonomous navigation, the fidelity of sonar simulation has become a ke

hardwarearxiv-cs-ro
5 Jun 2026
Hardware

CRAM-ER: Error-Resilient Spintronic Computational Random Access Memory for Scalable In-Memory Computation

DGX agent

arXiv:2606.02781v1 Announce Type: cross Abstract: Deep neural networks (DNNs) have achieved state-of-the-art performance across diverse domains. However, typical Von Neumann compute paradigms face sev

hardwarearxiv-cs-ai
3 Jun 2026
Hardware

Releasing vui an open source voice mode 300M TTS model Runs on a single consumer gpu / apple sillicon Context aware speech 6 minutes of cont…

DGX agent

Jeremy Howard announced the release of Vui, an open-source voice mode text-to-speech (TTS) model with 300 million parameters that can run on consumer GPUs and Apple Silicon. The model features context

hardwarejeremy-howard--x
3 Jun 2026
Hardware

What’s new in serverless Managed Service for Apache Spark

DGX agent

Whether you use it for data preparation, real-time interactive queries, AI model training, or something entirely different, running Apache Spark at scale is demanding — you shouldn’t have to manage th

hardwaregoogle-cloud-ai
3 Jun 2026
Hardware

FLARE: Diffusion for Hybrid Language Model

DGX agent

arXiv:2606.01774v1 Announce Type: cross Abstract: Autoregressive (AR) large language models (LLMs) have achieved broad practical success, but sequential decoding remains a key bottleneck for low-laten

hardwarearxiv-cs-ai
2 Jun 2026
Hardware

Threshold-Based Exclusive Batching for LLM Inference

DGX agent

arXiv:2606.00516v1 Announce Type: new Abstract: Mixed batching (MB)--interleaving prefill and decode in a single batch--has become the standard scheduling strategy for large language model (LLM) infer

hardwarearxiv-cs-ai
2 Jun 2026
Hardware

Beyond Test-Time Memory: State-Space Optimal Control for LLM Reasoning

DGX agent

arXiv:2603.09221v2 Announce Type: replace Abstract: Associative memory has long underpinned the design of sequential models. Beyond recall, humans reason by projecting future states and selecting goal

hardwarearxiv-cs-lg
1 Jun 2026
Hardware

On Efficient Scaling of GNNs via IO-Aware Layers Implementations

DGX agent

arXiv:2605.31500v1 Announce Type: cross Abstract: Graph Neural Networks (GNNs) are bottlenecked by sparse, irregular memory access. Popular frameworks such as DGL and PyTorch Geometric support general

hardwarearxiv-cs-ai
1 Jun 2026
Model Releases

QASM-Eval: A Dataset to Train and Evaluate LLMs on OpenQASM-3 Beyond Quantum Circuits

DGX agent

arXiv:2605.30358v1 Announce Type: new Abstract: Quantum computing remains in the Noisy Intermediate-Scale Quantum (NISQ) era, where the performance is highly constrained to noise. Addressing the limit

model-releasesarxiv-cs-lg
1 Jun 2026
Hardware

Unbox one of NVIDIA's first co-packaged optics samples with us. See why we bet on CPO early.

DGX agent

When we design large GPU clusters, the network is no longer a background system. It's part of the compute envelope. At the 800G and NVIDIA GB300 NVL72 scale, the back-end fabric accounts for 86% of ne

hardwarelambda-labs
1 Jun 2026
← Previous
1…89101112…93
Next →