AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “hardware”

GridTimelineEvolution
4,474 results
21 May 2026

JUST IN: SpaceX pledges to cover all power grid infrastructure upgrade costs associated with its data centers so those costs are not passed …

HardwareDGX agent

JUST IN: SpaceX pledges to cover all power grid infrastructure upgrade costs associated with its data centers so those costs are not passed on to local households, per its S-1 filing. JUST IN: COLOSSU

License to Stream: ‘007 First Light’ Coming to GeForce NOW With an Ultimate Bundle

HardwareDGX agent

The mission begins now. GeForce NOW is dialing up the action with a blockbuster mix of spy thrills, high-speed racing and member rewards — plus eight new games joining the cloud this week, all ready t

NVIDIA GTC Taipei at COMPUTEX: Live Updates on What’s Next in AI

HardwareDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

At NVIDIA GTC Taipei at COMPUTEX, the world’s developers, researchers and industry leaders are converging to dive into the latest breakthroughs shaping every industry, covering topics spanning AI fact

OlmoEarth v1.1: A more efficient family of OlmoEarth models

HardwareDGX agent

arXiv:2605.20804v1 Announce Type: new Abstract: We present a set of improvements to the OlmoEarth family. These improvements allow us to cut compute costs during training (1.7 imes reduction in GPU ho

On prem.

Local AiDGX agent

On prem. I'm excited about the new @amd Ryzen AI Halo because we need more local hardware for AI builders! There's something fun and exciting about building on your own machines rather than sending to

PlexRL: Cluster-Level Orchestration of Serviceized LLM Execution for RLVR

HardwareDGX agent

arXiv:2605.20863v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) has recently unlocked strong reasoning capabilities in large language models (LLMs), triggering

PulseCol: Periodically Refreshed Column-Sparse Attention for Accelerating Diffusion Language Models

HardwareDGX agent

arXiv:2605.20813v1 Announce Type: new Abstract: Inference in diffusion large language models (dLLMs) is computationally expensive, as full self-attention must be repeatedly executed at each step of th

So much for Nvidia’s cash flow. Wild.

HardwareDGX agent

So much for Nvidia’s cash flow. Wild. When does circular financing stop? Similar to a Ponzi Scheme, when cash outflows become greater than inflows. In this last quarter, NVDA cash grew by only ~600, a

Taiwan is seeking to detain three people for forging documents to export Nvidia-powered Super Micro servers to China, Hong Kong, and Macau, breaking US rules (Bloomberg)

HardwareDGX agent

Bloomberg: Taiwan is seeking to detain three people for forging documents to export Nvidia-powered Super Micro servers to China, Hong Kong, and Macau, breaking US rules — Taiwanese officials are seeki

Understanding Deterioration Random Effects for Causal Discovery in Infrastructure Management

HardwareDGX agent

arXiv:2605.20400v1 Announce Type: cross Abstract: Infrastructure deterioration poses significant challenges for asset management, yet existing approaches rely on population-averaged models that overlo

20 May 2026

Accelerating Sparse Transformer Inference on GPU

HardwareDGX agent

arXiv:2506.06095v4 Announce Type: replace Abstract: Large language models (LLMs) are popular around the world due to their powerful understanding capabilities. As the core component of LLMs, accelerat

Banger take

HardwareDGX agent

Banger take Gavin Baker: 'I've been optimistic that the fundamental shortage of wafers, which is really controlled by Taiwan Semi, will prevent a bubble.' 'If Taiwan Semi did what Jensen wanted, Nvidi

CircuitHub raises $28M to scale electronics production in days rather than months

ApplicationsDGX agent

CircuitHub Inc., a company building automated manufacturing and assembly systems for electronics hardware at scale that supports industries such as self-driving cars and satellites, today announced it

Decentralized Direct Volume Rendering: A Browser-Native GPU Architecture for MRI Digital Twins in Resource-Constrained Settings

HardwareDGX agent

arXiv:2605.19737v1 Announce Type: cross Abstract: Digital Twin (DT) technology holds immense potential for surgical planning and personalized medicine. However, generating interactive, patient-specifi

Document and sources: China banned Nvidia's RTX 5090D V2 chip, aimed at gamers and animators, while Jensen Huang was visiting China with Donald Trump last week (Financial Times)

HardwareDGX agent

Financial Times: Document and sources: China banned Nvidia's RTX 5090D V2 chip, aimed at gamers and animators, while Jensen Huang was visiting China with Donald Trump last week — Beijing aims to suppo

Exa Labs raises 250M at 2.2B valuation for its AI search tools

HardwareDGX agent

Search startup Exa Labs Inc. today announced that it has raised 250 million in funding to purchase more infrastructure. The round was led by Andersen Horowitz. It comes less than a year after Exa’s pr

FiLark: a streaming-first software framework for end-to-end exploration, annotation, and algorithm integration in distributed acoustic sensing

HardwareDGX agent

arXiv:2605.20132v1 Announce Type: cross Abstract: Distributed acoustic sensing (DAS) systems generate continuous, ultra-high-channel-count data streams at rates that exceed the capabilities of convent

GEM: GPU-Variability-Aware Expert to GPU Mapping for MoE Systems

HardwareDGX agent

arXiv:2605.19945v1 Announce Type: cross Abstract: Mixture-of-Expert (MoE) models enable efficient inference by employing smaller experts and activating only a subset of them per token. MoE serving eng

HiLiftAeroML: High-Fidelity Computational Fluid Dynamics Dataset for High-Lift Aircraft Aerodynamics

HardwareDGX agent

arXiv:2605.19565v1 Announce Type: cross Abstract: This paper describes the first-ever open-source high-fidelity CFD dataset of a high-lift aircraft for the purpose of AI surrogate model development. T

Hyrax: An Extensible Framework for Rapid ML Experimentation and Unsupervised Discovery in the Era of Rubin, Roman, and Euclid

HardwareDGX agent

arXiv:2605.18959v1 Announce Type: cross Abstract: The NSF-DOE Vera C. Rubin Observatory, Roman Space Telescope, Euclid, and other next-generation surveys will deliver imaging, spectroscopic, and time-

Lambda partners with Hudson River Trading to power quantitative research and development

HardwareDGX agent

Lambda Labs announced a partnership with Hudson River Trading (HRT), a quantitative trading firm, to provide computational infrastructure and GPU resources for HRT's quantitative research and developm

Nvidia almost doubles its data center revenue as it powers to another solid earnings beat

HardwareDGX agent

Chipmaker Nvidia Corp., the world’s most valuable company, crushed earnings expectations once again today, benefiting from massive demand for high-end artificial intelligence chips. The company report

OScaR: The Occam's Razor for Extreme KV Cache Quantization in LLMs and Beyond

HardwareDGX agent

arXiv:2605.19660v1 Announce Type: cross Abstract: The rapid advancement toward long-context reasoning and multi-modal intelligence has made the memory footprint of the Key-Value (KV) cache a dominant

PiKV: KV Cache Management System for Mixture of Experts

HardwareDGX agent

arXiv:2508.06526v3 Announce Type: replace-cross Abstract: As large-scale language models continue to scale up in both size and context length, the memory and communication cost of key-value (KV) cache

'Ryzen 5800X3D 10th Anniversary Edition' may help you avoid paying for a new PC

HardwareDGX agent

AMD is preparing an AM4 10th Anniversary Edition Ryzen 7 5800X3D for Q2 2026 launch, reviving its popular Zen 3 X3D chip with the same core specifications but updated packaging. The move is framed as

Sources: Nvidia's business development group, not its VC arm NVentures, has led much of its ~$90B dealmaking push across 145+ companies over the past 16 months (Financial Times)

HardwareDGX agent

Financial Times: Sources: Nvidia's business development group, not its VC arm NVentures, has led much of its ~90B dealmaking push across 145+ companies over the past 16 months — Nvidia's Huang bankrol

SuperInfer: SLO-Aware Rotary Scheduling and Memory Management for LLM Inference on Superchips

HardwareDGX agent

arXiv:2601.20309v2 Announce Type: replace-cross Abstract: Large Language Model (LLM) serving faces a fundamental tension between stringent latency Service Level Objectives (SLOs) and limited GPU memor

🌍Today we release Mosaic, a probabilistic weather model that shifts the Pareto frontier of ML weather forecasting. It matches the skill of …

HardwareDGX agent

🌍Today we release Mosaic, a probabilistic weather model that shifts the Pareto frontier of ML weather forecasting. It matches the skill of state-of-the-art models while generating a 24-member, 10-day

UCCI: Calibrated Uncertainty for Cost-Optimal LLM Cascade Routing

HardwareDGX agent

arXiv:2605.18796v1 Announce Type: cross Abstract: LLM cascades and model routing promise lower inference cost by sending easy queries to a small model and escalating hard ones to a large model, but mo

19 May 2026

3D Skew Gaussian Splatting with Any Camera Trajectory Visualization Engine

HardwareDGX agent

arXiv:2605.18334v1 Announce Type: new Abstract: While 3D Gaussian Splatting (3DGS) has revolutionized real-time photorealistic view synthesis, its fundamental reliance on symmetric Gaussian distributi

AutoVecCoder: Teaching LLMs to Generate Explicitly Vectorized Code

ResearchDGX agent

arXiv:2605.17978v1 Announce Type: new Abstract: Vectorization via Single Instruction, Multiple Data (SIMD) architectures is a cornerstone of high-performance computing. To fully exploit hardware poten

b9222

Local AiDGX agent

B9222 is a llama.cpp release that adds support for the TRI (Triangle) operation in the Hexagon HTP backend with HVX kernel additions . The release includes optimizations for Hexagon hardware accelerat

Breaking: Andrej Karpathy just filed his Q1 2026 13F. Here's everything you need to know about his recent 13F Top 10 positions: 1. VanEck Se…

HardwareDGX agent

Breaking: Andrej Karpathy just filed his Q1 2026 13F. Here's everything you need to know about his recent 13F Top 10 positions: 1. VanEck Semiconductor ETF - SMH - [Put] — 2.04B 2. Nvidia - NVDA - [Pu

Bundle Adjustment in the Eager Mode

HardwareDGX agent

arXiv:2409.12190v4 Announce Type: replace-cross Abstract: Bundle adjustment (BA) is a critical technique in various robotic applications such as simultaneous localization and mapping (SLAM), augmented

cuNRTO: GPU-Accelerated Nonlinear Robust Trajectory Optimization

HardwareDGX agent

arXiv:2603.02642v2 Announce Type: replace Abstract: Robust trajectory optimization enables autonomous systems to operate safely under uncertainty by computing control policies that satisfy the constra

DashAttention: Differentiable and Adaptive Sparse Hierarchical Attention

HardwareDGX agent

arXiv:2605.18753v1 Announce Type: cross Abstract: Current hierarchical attention methods, such as NSA and InfLLMv2, select the top-k relevant key-value (KV) blocks based on coarse attention scores and

Decart raises $300M for its AI optimization software, world models

HardwareDGX agent

Artificial intelligence developer Decart.ai Inc. today announced that it has raised 300 million in funding at a nearly 4 billion valuation. Radical Ventures led the round with participation from Nvidi

Dell’s CFO sees the enterprise AI buildout as a generational opportunity still in its opening act

HardwareDGX agent

AI factory momentum has reached a critical inflection point as capital joins silicon and energy as key constraints in the race to build enterprise AI infrastructure. That dynamic is playing out in rea

Efficient 3D Content Reconstruction and Generation

HardwareDGX agent

arXiv:2605.18052v1 Announce Type: new Abstract: Automatic 3D content creation seeks to replace labor-intensive modeling and scanning pipelines with systems that can synthesize or recover 3D assets dir

Elon Musk met with NVIDIA’s Ian Buck, who gave him the new Vera CPU for SpaceXAI.

HardwareDGX agent

I cannot verify this claim as accurate. The URL provided doesn't follow standard X post formats, and there is no publicly documented 'Vera CPU' from NVIDIA. This appears to be either fabricated or spe

Google's SynthID AI watermarking tech is being adopted by OpenAI, Nvidia, and more

HardwareDGX agent

Google's SynthID watermarking technology is being adopted by major AI companies including OpenAI, Kakao, and ElevenLabs to watermark AI-generated content. Google has also partnered with NVIDIA to wate

GPU-Accelerated Deep Learning for Heatwave Prediction and Urban Heat Risk Assessment

HardwareDGX agent

arXiv:2605.16435v1 Announce Type: cross Abstract: Heatwaves are an important problem in cities, and climate change makes this problem more difficult. In this paper, we present a GPU-based deep learnin

Guard: Scalable Straggler Detection and Node Health Management for Large-Scale Training

HardwareDGX agent

arXiv:2605.17879v1 Announce Type: cross Abstract: Training frontier-scale foundation models involves coordinating tens of thousands of GPUs over multi-month runs, where even minor performance degradat

Hawkeye: Reproducing GPU-Level Non-Determinism

HardwareDGX agent

arXiv:2603.20421v2 Announce Type: replace-cross Abstract: We present Hawkeye, a system for analyzing and reproducing GPU-level arithmetic operations. Using our framework, anyone can re-execute on a CP

IVF-TQ: Streaming-Robust Approximate Nearest Neighbor Search via a Codebook-Free Residual Layer

HardwareDGX agent

arXiv:2605.17415v1 Announce Type: cross Abstract: We propose IVF-TQ, an IVF index with a codebook-free residual layer: a fixed random rotation followed by precomputed Lloyd-Max scalar quantization dep

KVDrive: A Holistic Multi-Tier KV Cache Management System for Long-Context LLM Inference

HardwareDGX agent

arXiv:2605.18071v1 Announce Type: new Abstract: Supporting long-context LLMs is challenging due to the substantial memory demands of the key-value (KV) cache. Existing offloading systems store the ful

Lightweight Gaussian Process Inference in C++ on Metal and CUDA

HardwareDGX agent

arXiv:2605.17898v1 Announce Type: new Abstract: Gaussian process (GP) inference in Python is dominated by libraries such as GPyTorch and GPflow, which are built on deep-learning frameworks and inherit

LogRouter: Adaptive Two-Level LLM Routing for Log Question Answering in Big Data Systems

HardwareDGX agent

arXiv:2605.18015v1 Announce Type: new Abstract: Production log analytics in self-hosted, resource-constrained environments requires natural-language access to massive log streams without the cost of r

LongLive-2.0: An NVFP4 Parallel Infrastructure for Long Video Generation

HardwareDGX agent

arXiv:2605.18739v1 Announce Type: new Abstract: We present LongLive-2.0, an NVFP4-based parallel infrastructure throughout the full training and inference workflow of long video generation, addressing

Neural-network methods for two-dimensional finite-source reflector design

HardwareDGX agent

arXiv:2604.02184v2 Announce Type: replace Abstract: We address the inverse problem of designing two-dimensional reflectors that transform light from a finite, extended source into a prescribed far-fie

NVIDIA and Google Cloud Empower the Next Wave of AI Builders

HardwareDGX agent

At this year’s Google I/O conference, NVIDIA and Google Cloud are accelerating the work of more than 100,000 developers in the companies’ joint developer community, which provides curated learning pat

NVIDIA-Verified Agent Skills Provide Capability Governance for AI Agents

HardwareDGX agent

NVIDIA agent skills enable specialized agents to deliver targeted capabilities across enterprise workflows , with governance provided through verified, scanned, and signed skills. JFrog AI Catalog aut

Pocket Foundation Models: Distilling TFMs into CPU-Ready Gradient-Boosted Trees

HardwareDGX agent

arXiv:2605.18654v1 Announce Type: cross Abstract: A fraud scorer needs to answer in under 2 ms. The best tabular foundation models (TFMs) take 151-1,275 ms on GPU. We close this gap by distilling the

Progressive Generalization Augmentation with Deeply Coupled RND-PPO and Domain-Prioritized Noise Injection for Robust Crop Management Reinforcement Learning

HardwareDGX agent

arXiv:2605.17428v1 Announce Type: cross Abstract: Our preliminary experiments on gym-DSSAT maize irrigation tasks revealed that +/-2 degrees C temperature noise causes an 11.9% reduction in economic r

SignMuon: Communication-Efficient Distributed Muon Optimization

HardwareDGX agent

arXiv:2605.16311v1 Announce Type: new Abstract: Distributed training of large neural networks is bottlenecked by full-precision gradient communication and by coordinatewise optimizers that ignore the

Sparse Mamba Decoder for Quantum Error Correction: Efficient Defect-Centric Processing of Surface Code Syndromes

HardwareDGX agent

arXiv:2605.17156v1 Announce Type: cross Abstract: Quantum error correction (QEC) is essential for building fault-tolerant quantum computers, requiring decoders that are simultaneously accurate, fast,

Spherical Harmonic Optimal Transport: Application to Climate Models Comparisons

HardwareDGX agent

arXiv:2605.18389v1 Announce Type: new Abstract: Optimal transport provides a powerful framework for comparing measures while respecting the geometry of their support, but comes with an expensive compu

StreamingEffect: Real-Time Human-Centric Video Effect Generation

HardwareDGX agent

arXiv:2605.17019v1 Announce Type: new Abstract: Streaming video effect generation is highly desirable for live human-centric applications such as e-commerce streaming, entertainment, and vlogging, yet

Systematic Optimization of Real-Time Diffusion Model Inference on Apple M3 Ultra

HardwareDGX agent

arXiv:2605.16259v1 Announce Type: cross Abstract: While real-time image generation using diffusion models has advanced rapidly on NVIDIA GPUs, systematic optimization research on non-CUDA platforms su

TeleRAG: Efficient Retrieval-Augmented Generation Inference with Lookahead Retrieval

HardwareDGX agent

arXiv:2502.20969v4 Announce Type: replace-cross Abstract: Retrieval-augmented generation (RAG) extends large language models (LLMs) with external data sources to enhance factual correctness and domain

← Previous
1…2526272829…75
Next →