AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “hardware”

GridTimelineEvolution
4,457 results
26 Jun 2026

Sources: Apple's 14' and 16' OLED touch screen MacBook Pros will be powered by the existing M5 Pro and M5 Max chips and have an updated industrial design (Mark Gurman/Bloomberg)

HardwareDGX agent

Mark Gurman / Bloomberg: Sources: Apple's 14' and 16' OLED touch screen MacBook Pros will be powered by the existing M5 Pro and M5 Max chips and have an updated industrial design — Apple Inc.'s first-

TaskNPoint: How to Teach Your Humanoid to Hit a Backhand in Minutes

HardwareDGX agent

arXiv:2606.26215v1 Announce Type: cross Abstract: How do we learn to hit a tennis backhand? Not from a thousand hours of tennis tournaments on TV - we work with a coach and practice. We argue this is

The official GLM-5.2 NVFP4 from NVIDIA is now available. Curious how it compares to other quantizations. https://huggingface.co/nvidia/GLM-5…

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Hardware
DGX agent

NVIDIA has released GLM-5.2 NVFP4, a new quantization format for large language models. The announcement invites comparison with alternative quantization methods for performance and efficiency evaluat

Vultr Clusters Now Supports CPUs in Addition to GPUs – Plus Other Enhancements

HardwareDGX agent

Vultr has expanded its Clusters offering to support CPU-based workloads alongside existing GPU support, enabling more diverse compute deployments. The announcement includes additional enhancements to

25 Jun 2026

Agentic evolution of physically constrained foundation models

Model ReleasesDGX agent

arXiv:2606.25532v1 Announce Type: cross Abstract: Artificial intelligence increasingly drives automated scientific discovery, yet contemporary generalist agents lack physical grounding, frequently hal

Disagree with @mcuban’s opening sentence – I do think some people genuinely hate the noise, blight, impact on electric prices, etc—but he is…

HardwareDGX agent

Disagree with @mcuban’s opening sentence – I do think some people genuinely hate the noise, blight, impact on electric prices, etc—but he is quite right that data centers have become a proxy for hate

Energy-Efficient CNN Acceleration with MSDF Digit-Serial Arithmetic on FPGA

ResearchDGX agent

arXiv:2606.25562v1 Announce Type: cross Abstract: This paper presents an energy-efficient hardware acceleration of the convolutional layers in the U-Net architecture for image segmentation, implemente

How KRAFTON Built PUBG Ally, a Co-Playable Character Powered by NVIDIA ACE

HardwareDGX agent

PUBG Ally is an interactive co-playable character for PUBG: BATTLEGROUNDS powered by an on-device small language model built with NVIDIA ACE technology, equipped to communicate and cooperate with play

JetSpec: Breaking the Scaling Ceiling of Speculative Decoding with Parallel Tree Drafting

HardwareDGX agent

arXiv:2606.18394v2 Announce Type: replace Abstract: Speculative decoding (SD) accelerates autoregressive Large Language Models (LLMs) by drafting multiple tokens and verifying them in parallel, but it

Mirendil raises $200M to speed up scientific research with AI

HardwareDGX agent

Mirendil Inc., a startup developing artificial intelligence models for scientists, has raised 200 million in funding at a 1 billion valuation. The seed round was led by Andreessen Horowitz. Mirendil s

Optimize model training on Amazon SageMaker AI with NVIDIA Blackwell

HardwareDGX agent

This post shows you how to configure training jobs on Amazon SageMaker AI to get the most out of Blackwell’s architecture on AWS. You learn how to select batch sizes and sequence lengths that take adv

People don't quite understand how many behind the meter power generation assets are being built for datacenters because the US Grid sucks, d…

HardwareDGX agent

People don't quite understand how many behind the meter power generation assets are being built for datacenters because the US Grid sucks, despite the higher cost and complexity. US Grid Constraints:

Power-Flexible AI Data Centers: A New Paradigm for Grid-Responsive Compute

HardwareDGX agent

arXiv:2606.25098v1 Announce Type: cross Abstract: The rapid expansion of artificial intelligence (AI) infrastructure is driving unprecedented growth in electricity demand from data centers. Traditiona

SA-LIVO: Efficient LiDAR-Inertial-Visual Odometry with Subspace-Aware Degeneracy Handling

HardwareDGX agent

arXiv:2606.25699v1 Announce Type: new Abstract: Tightly coupled LiDAR-visual-inertial odometry (LIVO) fuses precise geometric depth with complementary visual measurements, yet its exteroceptive sensor

Sail, whose software optimizes how AI models run on existing chips, emerges from stealth with 80M in seed and Series A led by Kleiner at a 450M valuation (Lily Mae Lazarus/Fortune)

HardwareDGX agent

Lily Mae Lazarus / Fortune: Sail, whose software optimizes how AI models run on existing chips, emerges from stealth with 80M in seed and Series A led by Kleiner at a 450M valuation — For months, Klei

Scaling AI Inference Across Multiple GPUs Using NVIDIA TensorRT with Multi-Device Inference Support

HardwareDGX agent

NVIDIA TensorRT's multi-device inference feature scales inference across multiple GPUs, enabling deployment of models that exceed single-GPU memory or reducing latency for memory-bound workloads. Each

The Ultimate Summer Sale Pairing: Steam Sale Meets GeForce NOW Discounts

HardwareDGX agent

Summer savings are heating up. From the Steam Summer Sale to GeForce NOW membership discounts, this week’s GFN Thursday delivers double the deals and more ways to get the most value from cloud gaming.

The US awards $250M in CHIPS funds to a startup co-founded by billionaire mining magnate Robert Friedland to develop silicon-carbide chips and pulsed-power tech (James Attwood/Bloomberg)

HardwareDGX agent

James Attwood / Bloomberg: The US awards 250M in CHIPS funds to a startup co-founded by billionaire mining magnate Robert Friedland to develop silicon-carbide chips and pulsed-power tech — I-Pulse Inc

US Grid Constraints: Towards 40GW+ of Behind-The-Meter Datacenter by 2028?

HardwareDGX agent

This analysis examines the growing power demands of behind-the-meter datacenters in the US, projecting capacity could exceed 40GW by 2028, and discusses how grid constraints and infrastructure limitat

Without a doubt: A world running foundation models on a billion robots won't be doing it locally. None of the embedded compute will look lik…

HardwareDGX agent

Without a doubt: A world running foundation models on a billion robots won't be doing it locally. None of the embedded compute will look like it does today. All of the embedded compute will be wireles

24 Jun 2026

Accelerating Transformers Fine-Tuning with NVIDIA NeMo AutoModel

HardwareDGX agent

NVIDIA NeMo AutoModel is a tool designed to accelerate the fine-tuning process of transformer models by automating model selection and configuration. The solution leverages NVIDIA's NeMo framework to

Advancing AI in Property Insurance with AMD and Vultr

HardwareDGX agent

This article discusses a collaboration between AMD and Vultr to implement AI solutions in the property insurance sector, leveraging AMD's processors and Vultr's cloud infrastructure. The partnership l

b9784

Local AiDGX agent

Build b9784 is a release of llama.cpp, a C/C++ project that enables large language model inference with minimal setup on a wide range of hardware. The release includes pre-compiled binaries for multip

Beyond the $7.8B in deals: Why Wall Street is suddenly watching Argentum AI

HardwareDGX agent

Is Argentum AI building the financing layer for the AI infrastructure boom? Yesterday, Barron’s reported that Argentum AI, the infrastructure startup led by Andrew Sobko and backed by Supermicro, had

Configurable Holography: Towards Display and Scene Adaptation

ResearchDGX agent

arXiv:2405.01558v4 Announce Type: replace Abstract: Rendering holograms for holographic displays is often an iterative and computationally costly process. Emerging learned holography methods have alle

Differentiable Packing of Irregular 3D Objects with Adaptive Container Estimation

HardwareDGX agent

arXiv:2606.16333v2 Announce Type: replace Abstract: Most existing approaches either fix the container in advance or optimize only a single container dimension through an outer search loop, leaving the

NoContactNoWorries: Estimating Contact through Vision and Proprioception for In-Hand Dexterous Manipulation

TutorialsDGX agent

arXiv:2606.24450v1 Announce Type: cross Abstract: Perceiving physical contact is fundamental to dexterous manipulation. While robots often rely on dedicated hardware tactile sensors, humans exhibit a

NVIDIA and AWS Collaborate to Bring AI to Production at Scale

HardwareDGX agent

Building AI systems at scale is demanding, requiring low-latency inference, fast vector search, strong GPU price-performance and infrastructure that can grow without multiplying operational complexity

OpenAI and Broadcom announce chip designed for LLM inference at scale

HardwareDGX agent

OpenAI and Broadcom unveiled Jalapeño, OpenAI's first Intelligence Processor designed specifically for LLM inference , with early testing showing substantially better performance per watt than current

OpenAI and Broadcom unveil Jalapeño, an LLM-optimized inference chip developed from design to manufacturing tape-out in nine months, aided by OpenAI's models (OpenAI)

HardwareDGX agent

OpenAI: OpenAI and Broadcom unveil Jalapeño, an LLM-optimized inference chip developed from design to manufacturing tape-out in nine months, aided by OpenAI's models — Loading... - Early testing shows

OpenAI, Broadcom debut custom Jalapeño chip for AI inference

HardwareDGX agent

OpenAI Group PBC today revealed a custom chip called Jalapeño that it will use to power its large language models. The processor is the fruit of a collaboration with Broadcom Inc., which is no strange

Ornn, which plans to launch a marketplace for GPU capacity designed to function like an exchange to trade oil contracts, raised a $33M seed led by a16z (Katherine Doherty/Bloomberg)

HardwareDGX agent

Katherine Doherty / Bloomberg: Ornn, which plans to launch a marketplace for GPU capacity designed to function like an exchange to trade oil contracts, raised a 33M seed led by a16z — Ornn raised 33 m

RunPod, which rents access to non-Nvidia servers, raised 100M led by Summit Partners at a 1B valuation, a source says up from $100M after its seed in 2024 (Stephanie Palazzolo/The Information)

HardwareDGX agent

Stephanie Palazzolo / The Information: RunPod, which rents access to non-Nvidia servers, raised 100M led by Summit Partners at a 1B valuation, a source says up from $100M after its seed in 2024 — By s

SambaNova Executive Chairman Lip-Bu Tan says the AI chip startup is set to raise 800M to 1B, sources say at a ~10B valuation, up from 2B in February (The Information)

HardwareDGX agent

The Information: SambaNova Executive Chairman Lip-Bu Tan says the AI chip startup is set to raise 800M to 1B, sources say at a ~10B valuation, up from 2B in February — AI server chips that effectively

Understanding the Hidden Economics of Disaggregated AI Inference

HardwareDGX agent

New research explores how game theory can optimize NVIDIA Dynamo disaggregated serving architectures, revealing critical performance thresholds and reducing worst-case AI inference latency by up to 7.

VoltanaLLM: Energy-Efficient and SLO-Aware Disaggregated LLM Serving via Adaptive Frequency Control and State-Space Routing

HardwareDGX agent

arXiv:2509.04827v3 Announce Type: replace-cross Abstract: The energy cost of Large Language Model (LLM) inference is rapidly becoming a barrier to sustainable and scalable deployment. Although modern

23 Jun 2026

A Differentiable Atari VCS:A Complex, Fully Known Ground Truth for Explainable AI

HardwareDGX agent

arXiv:2606.22447v1 Announce Type: cross Abstract: Explanation requires ground truth: to verify an account of a system we must know its inner functioning-just what is missing where explainable AI (XAI)

ACEsplat: Accelerated 3D Gaussian Scene Regression via RGB and Poses Only

HardwareDGX agent

arXiv:2606.22091v1 Announce Type: new Abstract: Per-scene 3D Gaussian Splatting (3DGS) enables high-fidelity rendering, but practical robotic and AR scene capture pipelines often depend on external ge

All DOGE required was contact information of the recipients to confirm that funding was not fraudulent. No validated medical funding was sto…

HardwareDGX agent

All DOGE required was contact information of the recipients to confirm that funding was not fraudulent. No validated medical funding was stopped. Anything that appeared to be legitimate lifesaving fun

An Analysis of Untrained Deep Reservoir Networks for Audio Surveillance

HardwareDGX agent

arXiv:2606.22218v1 Announce Type: cross Abstract: In this paper, we investigate untrained recurrent models from the Reservoir Computing (RC) paradigm for audio surveillance, focusing on bidirectional

An interview with Nvidia VP of Healthcare Kimberly Powell on how AI can ease doctors' workloads, help address trained medical staff shortages, and more (Cristina Criddle/Financial Times)

HardwareDGX agent

Cristina Criddle / Financial Times: An interview with Nvidia VP of Healthcare Kimberly Powell on how AI can ease doctors' workloads, help address trained medical staff shortages, and more — The chipma

BAC-JEPA: Label-Efficient Breast Arterial Calcification Segmentation via Synthetic Mammography-Guided Supervision

HardwareDGX agent

arXiv:2606.22089v1 Announce Type: new Abstract: Breast arterial calcification (BAC) on screening mammograms is an emerging cardiovascular risk biomarker, but quantitative use requires reproducible seg

BatchGen: An Architecture for Scalable and Efficient Batch Inference

HardwareDGX agent

arXiv:2606.21712v1 Announce Type: cross Abstract: Batch inference has become a central mode of AI computation, yet existing inference engines still rely on execution models designed for interactive se

Boost Inference Performance up to 15x on NVIDIA Blackwell Using DFlash Speculative Decoding

HardwareDGX agent

DFlash is an open source block diffusion model for speculative decoding that significantly accelerates LLM inference on NVIDIA Blackwell GPUs by drafting entire token blocks in parallel and verifying

Build an AI Scientist for Life Science Discovery with NVIDIA BioNeMo Agent Toolkit

HardwareDGX agent

The NVIDIA BioNeMo Agent Toolkit enables AI scientists to use structure prediction, molecular generation, docking, sequence analysis, design, and genomics as callable tools , allowing AI agents to exe

.@cartesia runs one of the hardest inference workloads: real-time voice. Their stack has to keep long-lived streams moving, serve millions o…

HardwareDGX agent

.@cartesia runs one of the hardest inference workloads: real-time voice. Their stack has to keep long-lived streams moving, serve millions of audio minutes a day, and hold model latency around 90ms. T

China’s CXMT Is Set to Challenge DRAM Incumbents

HardwareDGX agent

China Xin Memory Technologies (CXMT) is positioning itself as a competitor to established DRAM manufacturers through domestic production capabilities and technological development. The article likely

Concordia: JIT-Compiled Persistent-Kernel Checkpointing for Fault-Tolerant LLM Inference

HardwareDGX agent

arXiv:2606.23521v1 Announce Type: cross Abstract: Long-running LLM agents keep valuable state resident on GPUs: KV caches, request schedulers, communication state, and sometimes online adapters. Losin

Eli Lilly plans to use its $7.3B cash reserve to fund an 'App Store' for biotech scientists after launching a data center with 1,016 Blackwell chips in 2025 (Patrick Temple-West/Financial Times)

HardwareDGX agent

Patrick Temple-West / Financial Times: Eli Lilly plans to use its $7.3B cash reserve to fund an “App Store” for biotech scientists after launching a data center with 1,016 Blackwell chips in 2025 — Mo

FeLoG: Scalable and Efficient Distributed Graph Embedding with Feedback Loop Mechanism

HardwareDGX agent

arXiv:2606.22180v1 Announce Type: cross Abstract: Graph embedding maps graph nodes into low-dimensional vectors to support applications such as recommendation, fraud detection, and graph-based retriev

FGGS-LiDAR: Ultra-Fast, GPU-Accelerated Simulation from General 3DGS Models to LiDAR

HardwareDGX agent

arXiv:2509.17390v2 Announce Type: replace Abstract: While 3D Gaussian Splatting (3DGS) has emerged as a strong representation for photorealistic rendering, its vast ecosystem of assets remains difficu

FleetAgent: Teleoperation Assistant for Autonomous Fleets via Vectorized V2N Messages

HardwareDGX agent

arXiv:2606.21222v1 Announce Type: new Abstract: Large-scale autonomous fleets rely on teleoperation to resolve rare failures, yet streaming raw sensor data from many vehicles is costly, and remote ope

FoMoE: Breaking the Full-Replica Barrier with a Federation of MoEs

ResearchDGX agent

arXiv:2606.19025v2 Announce Type: replace Abstract: Pre-training Large Language Models (LLMs) typically demands large-scale infrastructure with tightly coupled hardware accelerators. Mixture-of-Expert

Genesis Workbench: A blueprint for industry AI in life sciences, powered by Databricks and NVIDIA

HardwareDGX agent

Genesis Workbench is a collaborative framework developed by Databricks and NVIDIA designed to accelerate AI adoption in the life sciences industry by providing integrated tools and best practices. The

GRINQH: Graded Input-based Quantization Hierarchy for Efficient LLM Generation

HardwareDGX agent

arXiv:2606.23419v1 Announce Type: new Abstract: Autoregressive decoding with LLMs is primarily bottlenecked by GPU memory bandwidth, especially in edge-computing settings. While quantization is essent

HERALD: High-Throughput Block Diffusion LLM Serving via CPU-GPU Cooperative KV Cache Retrieval

HardwareDGX agent

arXiv:2606.21633v1 Announce Type: new Abstract: Diffusion LLMs (dLLMs) improve GPU utilization over autoregressive decoding by generating multiple tokens per forward pass, but their KV cache still gro

How Telcos Build Autonomous Networks with Agentic AI

HardwareDGX agent

Telecommunication operators deploy autonomous agents built on telecom-domain models and running inside secure execution runtimes connected to tools, digital twins, and shared skills to understand netw

i dont think anyone is correctly doing the math around how SpaceX, the NeoCloud+NeoLab, is currently going to market? SpaceX has already rec…

HardwareDGX agent

i dont think anyone is correctly doing the math around how SpaceX, the NeoCloud+NeoLab, is currently going to market? SpaceX has already recouped about HALF its investment in Cursor, in compute deals.

Line Drawings using LightBenders: Authoring and Illuminating

AgentsDGX agent

arXiv:2606.22499v1 Announce Type: cross Abstract: This study presents the hardware and software architecture of a transformative system for illuminating line drawings and letterforms. These mid-air il

LLMs write fast single-GPU kernels. Ask for a multi-GPU one and they fall apart. ParallelKernelBench (PKB) measures how they fail by benchma…

HardwareDGX agent

LLMs write fast single-GPU kernels. Ask for a multi-GPU one and they fall apart. ParallelKernelBench (PKB) measures how they fail by benchmarking against 87 problems pulled from real codebases includi

← Previous
1…1718192021…75
Next →