AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “hardware”

GridTimelineEvolution
4,477 results
Hardware

OScaR: The Occam's Razor for Extreme KV Cache Quantization in LLMs and Beyond

DGX agent

arXiv:2605.19660v1 Announce Type: cross Abstract: The rapid advancement toward long-context reasoning and multi-modal intelligence has made the memory footprint of the Key-Value (KV) cache a dominant

hardwarearxiv-cs-cl
20 May 2026
Hardware
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

PiKV: KV Cache Management System for Mixture of Experts

DGX agent

arXiv:2508.06526v3 Announce Type: replace-cross Abstract: As large-scale language models continue to scale up in both size and context length, the memory and communication cost of key-value (KV) cache

hardwarearxiv-cs-ai
20 May 2026
Hardware

'Ryzen 5800X3D 10th Anniversary Edition' may help you avoid paying for a new PC

DGX agent

AMD is preparing an AM4 10th Anniversary Edition Ryzen 7 5800X3D for Q2 2026 launch, reviving its popular Zen 3 X3D chip with the same core specifications but updated packaging. The move is framed as

hardwarears-technica
20 May 2026
Hardware

Sources: Nvidia's business development group, not its VC arm NVentures, has led much of its ~$90B dealmaking push across 145+ companies over the past 16 months (Financial Times)

DGX agent

Financial Times: Sources: Nvidia's business development group, not its VC arm NVentures, has led much of its ~90B dealmaking push across 145+ companies over the past 16 months — Nvidia's Huang bankrol

hardwaretechmeme
20 May 2026
Hardware

SuperInfer: SLO-Aware Rotary Scheduling and Memory Management for LLM Inference on Superchips

DGX agent

arXiv:2601.20309v2 Announce Type: replace-cross Abstract: Large Language Model (LLM) serving faces a fundamental tension between stringent latency Service Level Objectives (SLOs) and limited GPU memor

hardwarearxiv-cs-ai
20 May 2026
Hardware

🌍Today we release Mosaic, a probabilistic weather model that shifts the Pareto frontier of ML weather forecasting. It matches the skill of …

DGX agent

🌍Today we release Mosaic, a probabilistic weather model that shifts the Pareto frontier of ML weather forecasting. It matches the skill of state-of-the-art models while generating a 24-member, 10-day

hardwareemad-mostaque--x
20 May 2026
Hardware

UCCI: Calibrated Uncertainty for Cost-Optimal LLM Cascade Routing

DGX agent

arXiv:2605.18796v1 Announce Type: cross Abstract: LLM cascades and model routing promise lower inference cost by sending easy queries to a small model and escalating hard ones to a large model, but mo

hardwarearxiv-cs-cl
20 May 2026
Hardware

3D Skew Gaussian Splatting with Any Camera Trajectory Visualization Engine

DGX agent

arXiv:2605.18334v1 Announce Type: new Abstract: While 3D Gaussian Splatting (3DGS) has revolutionized real-time photorealistic view synthesis, its fundamental reliance on symmetric Gaussian distributi

hardwarearxiv-cs-cv
19 May 2026
Research

AutoVecCoder: Teaching LLMs to Generate Explicitly Vectorized Code

DGX agent

arXiv:2605.17978v1 Announce Type: new Abstract: Vectorization via Single Instruction, Multiple Data (SIMD) architectures is a cornerstone of high-performance computing. To fully exploit hardware poten

researcharxiv-cs-cl
19 May 2026
Local Ai

b9222

DGX agent

B9222 is a llama.cpp release that adds support for the TRI (Triangle) operation in the Hexagon HTP backend with HVX kernel additions . The release includes optimizations for Hexagon hardware accelerat

local-aillama-cpp-releases
19 May 2026
Hardware

Breaking: Andrej Karpathy just filed his Q1 2026 13F. Here's everything you need to know about his recent 13F Top 10 positions: 1. VanEck Se…

DGX agent

Breaking: Andrej Karpathy just filed his Q1 2026 13F. Here's everything you need to know about his recent 13F Top 10 positions: 1. VanEck Semiconductor ETF - SMH - [Put] — 2.04B 2. Nvidia - NVDA - [Pu

hardwaredylan-patel--x
19 May 2026
Hardware

Bundle Adjustment in the Eager Mode

DGX agent

arXiv:2409.12190v4 Announce Type: replace-cross Abstract: Bundle adjustment (BA) is a critical technique in various robotic applications such as simultaneous localization and mapping (SLAM), augmented

hardwarearxiv-cs-cv
19 May 2026
Hardware

cuNRTO: GPU-Accelerated Nonlinear Robust Trajectory Optimization

DGX agent

arXiv:2603.02642v2 Announce Type: replace Abstract: Robust trajectory optimization enables autonomous systems to operate safely under uncertainty by computing control policies that satisfy the constra

hardwarearxiv-cs-ro
19 May 2026
Hardware

DashAttention: Differentiable and Adaptive Sparse Hierarchical Attention

DGX agent

arXiv:2605.18753v1 Announce Type: cross Abstract: Current hierarchical attention methods, such as NSA and InfLLMv2, select the top-k relevant key-value (KV) blocks based on coarse attention scores and

hardwarearxiv-cs-ai
19 May 2026
Hardware

Decart raises $300M for its AI optimization software, world models

DGX agent

Artificial intelligence developer Decart.ai Inc. today announced that it has raised 300 million in funding at a nearly 4 billion valuation. Radical Ventures led the round with participation from Nvidi

hardwaresiliconangle
19 May 2026
Hardware

Dell’s CFO sees the enterprise AI buildout as a generational opportunity still in its opening act

DGX agent

AI factory momentum has reached a critical inflection point as capital joins silicon and energy as key constraints in the race to build enterprise AI infrastructure. That dynamic is playing out in rea

hardwaresiliconangle
19 May 2026
Hardware

Efficient 3D Content Reconstruction and Generation

DGX agent

arXiv:2605.18052v1 Announce Type: new Abstract: Automatic 3D content creation seeks to replace labor-intensive modeling and scanning pipelines with systems that can synthesize or recover 3D assets dir

hardwarearxiv-cs-cv
19 May 2026
Hardware

Elon Musk met with NVIDIA’s Ian Buck, who gave him the new Vera CPU for SpaceXAI.

DGX agent

I cannot verify this claim as accurate. The URL provided doesn't follow standard X post formats, and there is no publicly documented 'Vera CPU' from NVIDIA. This appears to be either fabricated or spe

hardwareelon-musk--x
19 May 2026
Hardware

Google's SynthID AI watermarking tech is being adopted by OpenAI, Nvidia, and more

DGX agent

Google's SynthID watermarking technology is being adopted by major AI companies including OpenAI, Kakao, and ElevenLabs to watermark AI-generated content. Google has also partnered with NVIDIA to wate

hardwarears-technica
19 May 2026
Hardware

GPU-Accelerated Deep Learning for Heatwave Prediction and Urban Heat Risk Assessment

DGX agent

arXiv:2605.16435v1 Announce Type: cross Abstract: Heatwaves are an important problem in cities, and climate change makes this problem more difficult. In this paper, we present a GPU-based deep learnin

hardwarearxiv-cs-ai
19 May 2026
Hardware

Guard: Scalable Straggler Detection and Node Health Management for Large-Scale Training

DGX agent

arXiv:2605.17879v1 Announce Type: cross Abstract: Training frontier-scale foundation models involves coordinating tens of thousands of GPUs over multi-month runs, where even minor performance degradat

hardwarearxiv-cs-ai
19 May 2026
Hardware

Hawkeye: Reproducing GPU-Level Non-Determinism

DGX agent

arXiv:2603.20421v2 Announce Type: replace-cross Abstract: We present Hawkeye, a system for analyzing and reproducing GPU-level arithmetic operations. Using our framework, anyone can re-execute on a CP

hardwarearxiv-cs-lg
19 May 2026
Hardware

IVF-TQ: Streaming-Robust Approximate Nearest Neighbor Search via a Codebook-Free Residual Layer

DGX agent

arXiv:2605.17415v1 Announce Type: cross Abstract: We propose IVF-TQ, an IVF index with a codebook-free residual layer: a fixed random rotation followed by precomputed Lloyd-Max scalar quantization dep

hardwarearxiv-cs-ai
19 May 2026
Hardware

KVDrive: A Holistic Multi-Tier KV Cache Management System for Long-Context LLM Inference

DGX agent

arXiv:2605.18071v1 Announce Type: new Abstract: Supporting long-context LLMs is challenging due to the substantial memory demands of the key-value (KV) cache. Existing offloading systems store the ful

hardwarearxiv-cs-cl
19 May 2026
Hardware

Lightweight Gaussian Process Inference in C++ on Metal and CUDA

DGX agent

arXiv:2605.17898v1 Announce Type: new Abstract: Gaussian process (GP) inference in Python is dominated by libraries such as GPyTorch and GPflow, which are built on deep-learning frameworks and inherit

hardwarearxiv-cs-lg
19 May 2026
Hardware

LogRouter: Adaptive Two-Level LLM Routing for Log Question Answering in Big Data Systems

DGX agent

arXiv:2605.18015v1 Announce Type: new Abstract: Production log analytics in self-hosted, resource-constrained environments requires natural-language access to massive log streams without the cost of r

hardwarearxiv-cs-lg
19 May 2026
Hardware

LongLive-2.0: An NVFP4 Parallel Infrastructure for Long Video Generation

DGX agent

arXiv:2605.18739v1 Announce Type: new Abstract: We present LongLive-2.0, an NVFP4-based parallel infrastructure throughout the full training and inference workflow of long video generation, addressing

hardwarearxiv-cs-cv
19 May 2026
Hardware

Neural-network methods for two-dimensional finite-source reflector design

DGX agent

arXiv:2604.02184v2 Announce Type: replace Abstract: We address the inverse problem of designing two-dimensional reflectors that transform light from a finite, extended source into a prescribed far-fie

hardwarearxiv-cs-lg
19 May 2026
Hardware

NVIDIA and Google Cloud Empower the Next Wave of AI Builders

DGX agent

At this year’s Google I/O conference, NVIDIA and Google Cloud are accelerating the work of more than 100,000 developers in the companies’ joint developer community, which provides curated learning pat

hardwarenvidia-blog
19 May 2026
Hardware

NVIDIA-Verified Agent Skills Provide Capability Governance for AI Agents

DGX agent

NVIDIA agent skills enable specialized agents to deliver targeted capabilities across enterprise workflows , with governance provided through verified, scanned, and signed skills. JFrog AI Catalog aut

hardwarenvidia-developer
19 May 2026
Hardware

Pocket Foundation Models: Distilling TFMs into CPU-Ready Gradient-Boosted Trees

DGX agent

arXiv:2605.18654v1 Announce Type: cross Abstract: A fraud scorer needs to answer in under 2 ms. The best tabular foundation models (TFMs) take 151-1,275 ms on GPU. We close this gap by distilling the

hardwarearxiv-cs-ai
19 May 2026
Hardware

Progressive Generalization Augmentation with Deeply Coupled RND-PPO and Domain-Prioritized Noise Injection for Robust Crop Management Reinforcement Learning

DGX agent

arXiv:2605.17428v1 Announce Type: cross Abstract: Our preliminary experiments on gym-DSSAT maize irrigation tasks revealed that +/-2 degrees C temperature noise causes an 11.9% reduction in economic r

hardwarearxiv-cs-ai
19 May 2026
Hardware

SignMuon: Communication-Efficient Distributed Muon Optimization

DGX agent

arXiv:2605.16311v1 Announce Type: new Abstract: Distributed training of large neural networks is bottlenecked by full-precision gradient communication and by coordinatewise optimizers that ignore the

hardwarearxiv-cs-lg
19 May 2026
Hardware

Sparse Mamba Decoder for Quantum Error Correction: Efficient Defect-Centric Processing of Surface Code Syndromes

DGX agent

arXiv:2605.17156v1 Announce Type: cross Abstract: Quantum error correction (QEC) is essential for building fault-tolerant quantum computers, requiring decoders that are simultaneously accurate, fast,

hardwarearxiv-cs-lg
19 May 2026
Hardware

Spherical Harmonic Optimal Transport: Application to Climate Models Comparisons

DGX agent

arXiv:2605.18389v1 Announce Type: new Abstract: Optimal transport provides a powerful framework for comparing measures while respecting the geometry of their support, but comes with an expensive compu

hardwarearxiv-cs-lg
19 May 2026
Hardware

StreamingEffect: Real-Time Human-Centric Video Effect Generation

DGX agent

arXiv:2605.17019v1 Announce Type: new Abstract: Streaming video effect generation is highly desirable for live human-centric applications such as e-commerce streaming, entertainment, and vlogging, yet

hardwarearxiv-cs-cv
19 May 2026
Hardware

Systematic Optimization of Real-Time Diffusion Model Inference on Apple M3 Ultra

DGX agent

arXiv:2605.16259v1 Announce Type: cross Abstract: While real-time image generation using diffusion models has advanced rapidly on NVIDIA GPUs, systematic optimization research on non-CUDA platforms su

hardwarearxiv-cs-ai
19 May 2026
Hardware

TeleRAG: Efficient Retrieval-Augmented Generation Inference with Lookahead Retrieval

DGX agent

arXiv:2502.20969v4 Announce Type: replace-cross Abstract: Retrieval-augmented generation (RAG) extends large language models (LLMs) with external data sources to enhance factual correctness and domain

hardwarearxiv-cs-lg
19 May 2026
Hardware

TierCheck: Tiered Checkpointing for Fault Tolerance in Large Language Model Training

DGX agent

arXiv:2605.17821v1 Announce Type: cross Abstract: Large Language Model (LLM) training is frequently interrupted by a heterogeneous spectrum of failures, from common GPU crashes to catastrophic cluster

hardwarearxiv-cs-ai
19 May 2026
Hardware

Token-Space Mask Prediction for Efficient Vision Transformer Segmentation

DGX agent

arXiv:2605.18177v1 Announce Type: new Abstract: Query-based Vision Transformer segmentation models typically reconstruct dense spatial feature maps to predict masks, inheriting design patterns from co

hardwarearxiv-cs-cv
19 May 2026
Hardware

TriAxialKV: Toward Extreme Low-Precision KV-Cache Quantization for Agentic Inference Tasks

DGX agent

arXiv:2605.17170v1 Announce Type: new Abstract: Agentic workloads have emerged as a major workload for LLM inference. They differ significantly from chat-only workloads, requiring long-context process

hardwarearxiv-cs-lg
19 May 2026
Hardware

Truthful Calibration Errors for Multi-Class Prediction

DGX agent

arXiv:2510.06388v2 Announce Type: replace Abstract: Calibrated predictions are useful because their numerical values can be interpreted as probabilities. Calibration errors are therefore widely used t

hardwarearxiv-cs-lg
19 May 2026
Hardware

VeriCache: Turning Lossy KV Cache into Lossless LLM Inference

DGX agent

arXiv:2605.17613v1 Announce Type: cross Abstract: The large size of the KV cache has become a major bottleneck for serving LLMs with increasing context lengths. In response, many KV cache compression

hardwarearxiv-cs-lg
19 May 2026
Hardware

Vultr Announces Milan, Italy, as 33rd Cloud Data Center Region

DGX agent

Vultr expanded its global cloud infrastructure by opening a new data center region in Milan, Italy, marking the company's 33rd cloud data center location worldwide. This expansion provides European cu

hardwarevultr
19 May 2026
Hardware

We are releasing Carbon: a crazy fast DNA model Carbon is 275x faster than the next best model. So fast you can process the whole human geno…

DGX agent

We are releasing Carbon: a crazy fast DNA model Carbon is 275x faster than the next best model. So fast you can process the whole human genome on a single GPU in <2 days. Here are the tricks we used:

hardwareclem-delangue--x
19 May 2026
Hardware

A Few GPUs, A Whole Lotta Scale: Faithful LLM Training Emulation with PrismLLM

DGX agent

arXiv:2605.15617v1 Announce Type: cross Abstract: Large language model (LLM) training today runs on clusters spanning thousands of GPUs. While this scale enables rapid model advances, developing, debu

hardwarearxiv-cs-ai
18 May 2026
Hardware

A Unified Non-Parametric and Interpretable Point Cloud Analysis via t-FCW Graph Representation

DGX agent

arXiv:2605.15475v1 Announce Type: new Abstract: We introduce an empowered transposed Fully Connected Weighted (t-FCW) graph representation to embed point clouds into a metric space. While original t-F

hardwarearxiv-cs-cv
18 May 2026
Hardware

AgriMind: An Ensemble Deep Learning Framework for Multi-Class Plant Disease Classification

DGX agent

arXiv:2605.16076v1 Announce Type: cross Abstract: Plant disease detection is still largely manual in Bangladesh, where extension workers eyeball leaf samples across millions of smallholdings. We built

hardwarearxiv-cs-ai
18 May 2026
← Previous
1…3233343536…94
Next →