AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
All
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

hardware

GridTimelineEvolution
1,734 results
Hardware

Scaling AI Inference Across Multiple GPUs Using NVIDIA TensorRT with Multi-Device Inference Support

DGX agent

NVIDIA TensorRT's multi-device inference feature scales inference across multiple GPUs, enabling deployment of models that exceed single-GPU memory or reducing latency for memory-bound workloads. Each

hardwarenvidia-developer
25 Jun 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Hardware

The Ultimate Summer Sale Pairing: Steam Sale Meets GeForce NOW Discounts

DGX agent

Summer savings are heating up. From the Steam Summer Sale to GeForce NOW membership discounts, this week’s GFN Thursday delivers double the deals and more ways to get the most value from cloud gaming.

hardwarenvidia-blog
25 Jun 2026
Hardware

The US awards $250M in CHIPS funds to a startup co-founded by billionaire mining magnate Robert Friedland to develop silicon-carbide chips and pulsed-power tech (James Attwood/Bloomberg)

DGX agent

James Attwood / Bloomberg: The US awards 250M in CHIPS funds to a startup co-founded by billionaire mining magnate Robert Friedland to develop silicon-carbide chips and pulsed-power tech — I-Pulse Inc

hardwaretechmeme
25 Jun 2026
Hardware

US Grid Constraints: Towards 40GW+ of Behind-The-Meter Datacenter by 2028?

DGX agent

This analysis examines the growing power demands of behind-the-meter datacenters in the US, projecting capacity could exceed 40GW by 2028, and discusses how grid constraints and infrastructure limitat

hardwaresemianalysis
25 Jun 2026
Hardware

Without a doubt: A world running foundation models on a billion robots won't be doing it locally. None of the embedded compute will look lik…

DGX agent

Without a doubt: A world running foundation models on a billion robots won't be doing it locally. None of the embedded compute will look like it does today. All of the embedded compute will be wireles

hardwaredylan-patel--x
25 Jun 2026
Hardware

Accelerating Disaggregated RL for Visual Generative LLMs with Diffusion-Based Parallelism and Trainer-Assisted Generation

DGX agent

arXiv:2606.24369v1 Announce Type: new Abstract: Reinforcement learning (RL) has become a dominant post-training paradigm, driving the emergence of high-performance RL systems such as veRL for autoregr

hardwarearxiv-cs-ai
24 Jun 2026
Hardware

Accelerating Transformers Fine-Tuning with NVIDIA NeMo AutoModel

DGX agent

NVIDIA NeMo AutoModel is a tool designed to accelerate the fine-tuning process of transformer models by automating model selection and configuration. The solution leverages NVIDIA's NeMo framework to

hardwarehugging-face
24 Jun 2026
Hardware

Advancing AI in Property Insurance with AMD and Vultr

DGX agent

This article discusses a collaboration between AMD and Vultr to implement AI solutions in the property insurance sector, leveraging AMD's processors and Vultr's cloud infrastructure. The partnership l

hardwarevultr
24 Jun 2026
Hardware

Beyond the $7.8B in deals: Why Wall Street is suddenly watching Argentum AI

DGX agent

Is Argentum AI building the financing layer for the AI infrastructure boom? Yesterday, Barron’s reported that Argentum AI, the infrastructure startup led by Andrew Sobko and backed by Supermicro, had

hardwaresiliconangle
24 Jun 2026
Hardware

Differentiable Packing of Irregular 3D Objects with Adaptive Container Estimation

DGX agent

arXiv:2606.16333v2 Announce Type: replace Abstract: Most existing approaches either fix the container in advance or optimize only a single container dimension through an outer search loop, leaving the

hardwarearxiv-cs-cv
24 Jun 2026
Hardware

End-to-End Radar and Communication Modulation Recognition with Neuromorphic Computing

DGX agent

arXiv:2606.24075v1 Announce Type: cross Abstract: Although deep learning-based methods can achieve high accuracy in automatic modulation recognition (AMR) tasks, their high computational cost makes it

hardwarearxiv-cs-ai
24 Jun 2026
Hardware

Hardware-Oriented Inference Complexity of Kolmogorov-Arnold Networks

DGX agent

arXiv:2604.03345v2 Announce Type: replace Abstract: Kolmogorov-Arnold Networks (KANs) have recently emerged as a powerful architecture for various machine learning applications. However, their unique

hardwarearxiv-cs-lg
24 Jun 2026
Hardware

NVIDIA and AWS Collaborate to Bring AI to Production at Scale

DGX agent

Building AI systems at scale is demanding, requiring low-latency inference, fast vector search, strong GPU price-performance and infrastructure that can grow without multiplying operational complexity

hardwarenvidia-blog
24 Jun 2026
Hardware

OpenAI and Broadcom announce chip designed for LLM inference at scale

DGX agent

OpenAI and Broadcom unveiled Jalapeño, OpenAI's first Intelligence Processor designed specifically for LLM inference , with early testing showing substantially better performance per watt than current

hardwarears-technica
24 Jun 2026
Hardware

OpenAI and Broadcom unveil Jalapeño, an LLM-optimized inference chip developed from design to manufacturing tape-out in nine months, aided by OpenAI's models (OpenAI)

DGX agent

OpenAI: OpenAI and Broadcom unveil Jalapeño, an LLM-optimized inference chip developed from design to manufacturing tape-out in nine months, aided by OpenAI's models — Loading... - Early testing shows

hardwaretechmeme
24 Jun 2026
Hardware

OpenAI and Broadcom unveil LLM-optimized inference chip

DGX agent

OpenAI and Broadcom announced a collaboration on a specialized inference chip designed to optimize the performance and efficiency of large language model deployments. The chip, referred to as 'Jalapen

hardwareopenai
24 Jun 2026
Hardware

OpenAI, Broadcom debut custom Jalapeño chip for AI inference

DGX agent

OpenAI Group PBC today revealed a custom chip called Jalapeño that it will use to power its large language models. The processor is the fruit of a collaboration with Broadcom Inc., which is no strange

hardwaresiliconangle
24 Jun 2026
Hardware

Ornn, which plans to launch a marketplace for GPU capacity designed to function like an exchange to trade oil contracts, raised a $33M seed led by a16z (Katherine Doherty/Bloomberg)

DGX agent

Katherine Doherty / Bloomberg: Ornn, which plans to launch a marketplace for GPU capacity designed to function like an exchange to trade oil contracts, raised a 33M seed led by a16z — Ornn raised 33 m

hardwaretechmeme
24 Jun 2026
Hardware

RunPod, which rents access to non-Nvidia servers, raised 100M led by Summit Partners at a 1B valuation, a source says up from $100M after its seed in 2024 (Stephanie Palazzolo/The Information)

DGX agent

Stephanie Palazzolo / The Information: RunPod, which rents access to non-Nvidia servers, raised 100M led by Summit Partners at a 1B valuation, a source says up from $100M after its seed in 2024 — By s

hardwaretechmeme
24 Jun 2026
Hardware

SambaNova Executive Chairman Lip-Bu Tan says the AI chip startup is set to raise 800M to 1B, sources say at a ~10B valuation, up from 2B in February (The Information)

DGX agent

The Information: SambaNova Executive Chairman Lip-Bu Tan says the AI chip startup is set to raise 800M to 1B, sources say at a ~10B valuation, up from 2B in February — AI server chips that effectively

hardwaretechmeme
24 Jun 2026
Hardware

TurboMPC: Fast, Scalable, and Differentiable Model Predictive Control on the GPU

DGX agent

arXiv:2606.24039v1 Announce Type: new Abstract: Robotics increasingly relies on GPUs for parallel simulation, large-scale learning, and neural-network inference. For model predictive control (MPC) to

hardwarearxiv-cs-ro
24 Jun 2026
Hardware

Understanding the Hidden Economics of Disaggregated AI Inference

DGX agent

New research explores how game theory can optimize NVIDIA Dynamo disaggregated serving architectures, revealing critical performance thresholds and reducing worst-case AI inference latency by up to 7.

hardwarevultr
24 Jun 2026
Hardware

VoltanaLLM: Energy-Efficient and SLO-Aware Disaggregated LLM Serving via Adaptive Frequency Control and State-Space Routing

DGX agent

arXiv:2509.04827v3 Announce Type: replace-cross Abstract: The energy cost of Large Language Model (LLM) inference is rapidly becoming a barrier to sustainable and scalable deployment. Although modern

hardwarearxiv-cs-ai
24 Jun 2026
Hardware

A Differentiable Atari VCS:A Complex, Fully Known Ground Truth for Explainable AI

DGX agent

arXiv:2606.22447v1 Announce Type: cross Abstract: Explanation requires ground truth: to verify an account of a system we must know its inner functioning-just what is missing where explainable AI (XAI)

hardwarearxiv-cs-lg
23 Jun 2026
Hardware

ACEsplat: Accelerated 3D Gaussian Scene Regression via RGB and Poses Only

DGX agent

arXiv:2606.22091v1 Announce Type: new Abstract: Per-scene 3D Gaussian Splatting (3DGS) enables high-fidelity rendering, but practical robotic and AR scene capture pipelines often depend on external ge

hardwarearxiv-cs-ro
23 Jun 2026
Hardware

All DOGE required was contact information of the recipients to confirm that funding was not fraudulent. No validated medical funding was sto…

DGX agent

All DOGE required was contact information of the recipients to confirm that funding was not fraudulent. No validated medical funding was stopped. Anything that appeared to be legitimate lifesaving fun

hardwareelon-musk--x
23 Jun 2026
Hardware

An Analysis of Untrained Deep Reservoir Networks for Audio Surveillance

DGX agent

arXiv:2606.22218v1 Announce Type: cross Abstract: In this paper, we investigate untrained recurrent models from the Reservoir Computing (RC) paradigm for audio surveillance, focusing on bidirectional

hardwarearxiv-cs-lg
23 Jun 2026
Hardware

An interview with Nvidia VP of Healthcare Kimberly Powell on how AI can ease doctors' workloads, help address trained medical staff shortages, and more (Cristina Criddle/Financial Times)

DGX agent

Cristina Criddle / Financial Times: An interview with Nvidia VP of Healthcare Kimberly Powell on how AI can ease doctors' workloads, help address trained medical staff shortages, and more — The chipma

hardwaretechmeme
23 Jun 2026
Hardware

BAC-JEPA: Label-Efficient Breast Arterial Calcification Segmentation via Synthetic Mammography-Guided Supervision

DGX agent

arXiv:2606.22089v1 Announce Type: new Abstract: Breast arterial calcification (BAC) on screening mammograms is an emerging cardiovascular risk biomarker, but quantitative use requires reproducible seg

hardwarearxiv-cs-cv
23 Jun 2026
Hardware

BatchGen: An Architecture for Scalable and Efficient Batch Inference

DGX agent

arXiv:2606.21712v1 Announce Type: cross Abstract: Batch inference has become a central mode of AI computation, yet existing inference engines still rely on execution models designed for interactive se

hardwarearxiv-cs-lg
23 Jun 2026
Hardware

Boost Inference Performance up to 15x on NVIDIA Blackwell Using DFlash Speculative Decoding

DGX agent

DFlash is an open source block diffusion model for speculative decoding that significantly accelerates LLM inference on NVIDIA Blackwell GPUs by drafting entire token blocks in parallel and verifying

hardwarenvidia-developer
23 Jun 2026
Hardware

Build an AI Scientist for Life Science Discovery with NVIDIA BioNeMo Agent Toolkit

DGX agent

The NVIDIA BioNeMo Agent Toolkit enables AI scientists to use structure prediction, molecular generation, docking, sequence analysis, design, and genomics as callable tools , allowing AI agents to exe

hardwarenvidia-developer
23 Jun 2026
Hardware

.@cartesia runs one of the hardest inference workloads: real-time voice. Their stack has to keep long-lived streams moving, serve millions o…

DGX agent

.@cartesia runs one of the hardest inference workloads: real-time voice. Their stack has to keep long-lived streams moving, serve millions of audio minutes a day, and hold model latency around 90ms. T

hardwaretogether-ai--x
23 Jun 2026
Hardware

China’s CXMT Is Set to Challenge DRAM Incumbents

DGX agent

China Xin Memory Technologies (CXMT) is positioning itself as a competitor to established DRAM manufacturers through domestic production capabilities and technological development. The article likely

hardwaresemianalysis
23 Jun 2026
Hardware

CITADEL: CSI-Based Jamming Detection and Open-Set Classification for IIoT Networks

DGX agent

arXiv:2606.22939v1 Announce Type: cross Abstract: Radio frequency jamming poses a critical threat to the availability of wireless Industrial Internet of Things (IIoT) networks. Existing detection and

hardwarearxiv-cs-lg
23 Jun 2026
Hardware

Concordia: JIT-Compiled Persistent-Kernel Checkpointing for Fault-Tolerant LLM Inference

DGX agent

arXiv:2606.23521v1 Announce Type: cross Abstract: Long-running LLM agents keep valuable state resident on GPUs: KV caches, request schedulers, communication state, and sometimes online adapters. Losin

hardwarearxiv-cs-lg
23 Jun 2026
Hardware

Eli Lilly plans to use its $7.3B cash reserve to fund an 'App Store' for biotech scientists after launching a data center with 1,016 Blackwell chips in 2025 (Patrick Temple-West/Financial Times)

DGX agent

Patrick Temple-West / Financial Times: Eli Lilly plans to use its $7.3B cash reserve to fund an “App Store” for biotech scientists after launching a data center with 1,016 Blackwell chips in 2025 — Mo

hardwaretechmeme
23 Jun 2026
Hardware

Fast-TurboQuant: A Multiplier-Free Online Vector Quantization Approach

DGX agent

arXiv:2606.21448v1 Announce Type: new Abstract: As large language models scale, memory bandwidth for key-value caches and retrieval-augmented generation systems becomes a critical bottleneck. While 1-

hardwarearxiv-cs-lg
23 Jun 2026
Hardware

FeLoG: Scalable and Efficient Distributed Graph Embedding with Feedback Loop Mechanism

DGX agent

arXiv:2606.22180v1 Announce Type: cross Abstract: Graph embedding maps graph nodes into low-dimensional vectors to support applications such as recommendation, fraud detection, and graph-based retriev

hardwarearxiv-cs-lg
23 Jun 2026
Hardware

FGGS-LiDAR: Ultra-Fast, GPU-Accelerated Simulation from General 3DGS Models to LiDAR

DGX agent

arXiv:2509.17390v2 Announce Type: replace Abstract: While 3D Gaussian Splatting (3DGS) has emerged as a strong representation for photorealistic rendering, its vast ecosystem of assets remains difficu

hardwarearxiv-cs-ro
23 Jun 2026
Hardware

FleetAgent: Teleoperation Assistant for Autonomous Fleets via Vectorized V2N Messages

DGX agent

arXiv:2606.21222v1 Announce Type: new Abstract: Large-scale autonomous fleets rely on teleoperation to resolve rare failures, yet streaming raw sensor data from many vehicles is costly, and remote ope

hardwarearxiv-cs-ro
23 Jun 2026
Hardware

Genesis Workbench: A blueprint for industry AI in life sciences, powered by Databricks and NVIDIA

DGX agent

Genesis Workbench is a collaborative framework developed by Databricks and NVIDIA designed to accelerate AI adoption in the life sciences industry by providing integrated tools and best practices. The

hardwaredatabricks
23 Jun 2026
Hardware

GRINQH: Graded Input-based Quantization Hierarchy for Efficient LLM Generation

DGX agent

arXiv:2606.23419v1 Announce Type: new Abstract: Autoregressive decoding with LLMs is primarily bottlenecked by GPU memory bandwidth, especially in edge-computing settings. While quantization is essent

hardwarearxiv-cs-lg
23 Jun 2026
Hardware

HERALD: High-Throughput Block Diffusion LLM Serving via CPU-GPU Cooperative KV Cache Retrieval

DGX agent

arXiv:2606.21633v1 Announce Type: new Abstract: Diffusion LLMs (dLLMs) improve GPU utilization over autoregressive decoding by generating multiple tokens per forward pass, but their KV cache still gro

hardwarearxiv-cs-lg
23 Jun 2026
Hardware

How Telcos Build Autonomous Networks with Agentic AI

DGX agent

Telecommunication operators deploy autonomous agents built on telecom-domain models and running inside secure execution runtimes connected to tools, digital twins, and shared skills to understand netw

hardwarenvidia-developer
23 Jun 2026
Hardware

i dont think anyone is correctly doing the math around how SpaceX, the NeoCloud+NeoLab, is currently going to market? SpaceX has already rec…

DGX agent

i dont think anyone is correctly doing the math around how SpaceX, the NeoCloud+NeoLab, is currently going to market? SpaceX has already recouped about HALF its investment in Cursor, in compute deals.

hardwareswyx--x
23 Jun 2026
Hardware

LLMs write fast single-GPU kernels. Ask for a multi-GPU one and they fall apart. ParallelKernelBench (PKB) measures how they fail by benchma…

DGX agent

LLMs write fast single-GPU kernels. Ask for a multi-GPU one and they fall apart. ParallelKernelBench (PKB) measures how they fail by benchmarking against 87 problems pulled from real codebases includi

hardwaretogether-ai--x
23 Jun 2026
Hardware

Multi-cancer detection using a computationally efficient CNN with transfer learning

DGX agent

arXiv:2606.22400v1 Announce Type: new Abstract: This study introduces a computationally efficient convolutional neural network (CNN) architecture enhanced with transfer learning for multi-cancer detec

hardwarearxiv-cs-cv
23 Jun 2026
← Previous
1…1112131415…37
Next →