AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

hardware

GridTimelineEvolution
1,732 results
23 Jun 2026

An Analysis of Untrained Deep Reservoir Networks for Audio Surveillance

HardwareDGX agent

arXiv:2606.22218v1 Announce Type: cross Abstract: In this paper, we investigate untrained recurrent models from the Reservoir Computing (RC) paradigm for audio surveillance, focusing on bidirectional

An interview with Nvidia VP of Healthcare Kimberly Powell on how AI can ease doctors' workloads, help address trained medical staff shortages, and more (Cristina Criddle/Financial Times)

HardwareDGX agent

Cristina Criddle / Financial Times: An interview with Nvidia VP of Healthcare Kimberly Powell on how AI can ease doctors' workloads, help address trained medical staff shortages, and more — The chipma

BAC-JEPA: Label-Efficient Breast Arterial Calcification Segmentation via Synthetic Mammography-Guided Supervision


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Hardware
DGX agent

arXiv:2606.22089v1 Announce Type: new Abstract: Breast arterial calcification (BAC) on screening mammograms is an emerging cardiovascular risk biomarker, but quantitative use requires reproducible seg

BatchGen: An Architecture for Scalable and Efficient Batch Inference

HardwareDGX agent

arXiv:2606.21712v1 Announce Type: cross Abstract: Batch inference has become a central mode of AI computation, yet existing inference engines still rely on execution models designed for interactive se

Boost Inference Performance up to 15x on NVIDIA Blackwell Using DFlash Speculative Decoding

HardwareDGX agent

DFlash is an open source block diffusion model for speculative decoding that significantly accelerates LLM inference on NVIDIA Blackwell GPUs by drafting entire token blocks in parallel and verifying

Build an AI Scientist for Life Science Discovery with NVIDIA BioNeMo Agent Toolkit

HardwareDGX agent

The NVIDIA BioNeMo Agent Toolkit enables AI scientists to use structure prediction, molecular generation, docking, sequence analysis, design, and genomics as callable tools , allowing AI agents to exe

.@cartesia runs one of the hardest inference workloads: real-time voice. Their stack has to keep long-lived streams moving, serve millions o…

HardwareDGX agent

.@cartesia runs one of the hardest inference workloads: real-time voice. Their stack has to keep long-lived streams moving, serve millions of audio minutes a day, and hold model latency around 90ms. T

China’s CXMT Is Set to Challenge DRAM Incumbents

HardwareDGX agent

China Xin Memory Technologies (CXMT) is positioning itself as a competitor to established DRAM manufacturers through domestic production capabilities and technological development. The article likely

CITADEL: CSI-Based Jamming Detection and Open-Set Classification for IIoT Networks

HardwareDGX agent

arXiv:2606.22939v1 Announce Type: cross Abstract: Radio frequency jamming poses a critical threat to the availability of wireless Industrial Internet of Things (IIoT) networks. Existing detection and

Concordia: JIT-Compiled Persistent-Kernel Checkpointing for Fault-Tolerant LLM Inference

HardwareDGX agent

arXiv:2606.23521v1 Announce Type: cross Abstract: Long-running LLM agents keep valuable state resident on GPUs: KV caches, request schedulers, communication state, and sometimes online adapters. Losin

Eli Lilly plans to use its $7.3B cash reserve to fund an 'App Store' for biotech scientists after launching a data center with 1,016 Blackwell chips in 2025 (Patrick Temple-West/Financial Times)

HardwareDGX agent

Patrick Temple-West / Financial Times: Eli Lilly plans to use its $7.3B cash reserve to fund an “App Store” for biotech scientists after launching a data center with 1,016 Blackwell chips in 2025 — Mo

Fast-TurboQuant: A Multiplier-Free Online Vector Quantization Approach

HardwareDGX agent

arXiv:2606.21448v1 Announce Type: new Abstract: As large language models scale, memory bandwidth for key-value caches and retrieval-augmented generation systems becomes a critical bottleneck. While 1-

FeLoG: Scalable and Efficient Distributed Graph Embedding with Feedback Loop Mechanism

HardwareDGX agent

arXiv:2606.22180v1 Announce Type: cross Abstract: Graph embedding maps graph nodes into low-dimensional vectors to support applications such as recommendation, fraud detection, and graph-based retriev

FGGS-LiDAR: Ultra-Fast, GPU-Accelerated Simulation from General 3DGS Models to LiDAR

HardwareDGX agent

arXiv:2509.17390v2 Announce Type: replace Abstract: While 3D Gaussian Splatting (3DGS) has emerged as a strong representation for photorealistic rendering, its vast ecosystem of assets remains difficu

FleetAgent: Teleoperation Assistant for Autonomous Fleets via Vectorized V2N Messages

HardwareDGX agent

arXiv:2606.21222v1 Announce Type: new Abstract: Large-scale autonomous fleets rely on teleoperation to resolve rare failures, yet streaming raw sensor data from many vehicles is costly, and remote ope

Genesis Workbench: A blueprint for industry AI in life sciences, powered by Databricks and NVIDIA

HardwareDGX agent

Genesis Workbench is a collaborative framework developed by Databricks and NVIDIA designed to accelerate AI adoption in the life sciences industry by providing integrated tools and best practices. The

GRINQH: Graded Input-based Quantization Hierarchy for Efficient LLM Generation

HardwareDGX agent

arXiv:2606.23419v1 Announce Type: new Abstract: Autoregressive decoding with LLMs is primarily bottlenecked by GPU memory bandwidth, especially in edge-computing settings. While quantization is essent

HERALD: High-Throughput Block Diffusion LLM Serving via CPU-GPU Cooperative KV Cache Retrieval

HardwareDGX agent

arXiv:2606.21633v1 Announce Type: new Abstract: Diffusion LLMs (dLLMs) improve GPU utilization over autoregressive decoding by generating multiple tokens per forward pass, but their KV cache still gro

How Telcos Build Autonomous Networks with Agentic AI

HardwareDGX agent

Telecommunication operators deploy autonomous agents built on telecom-domain models and running inside secure execution runtimes connected to tools, digital twins, and shared skills to understand netw

i dont think anyone is correctly doing the math around how SpaceX, the NeoCloud+NeoLab, is currently going to market? SpaceX has already rec…

HardwareDGX agent

i dont think anyone is correctly doing the math around how SpaceX, the NeoCloud+NeoLab, is currently going to market? SpaceX has already recouped about HALF its investment in Cursor, in compute deals.

LLMs write fast single-GPU kernels. Ask for a multi-GPU one and they fall apart. ParallelKernelBench (PKB) measures how they fail by benchma…

HardwareDGX agent

LLMs write fast single-GPU kernels. Ask for a multi-GPU one and they fall apart. ParallelKernelBench (PKB) measures how they fail by benchmarking against 87 problems pulled from real codebases includi

Multi-cancer detection using a computationally efficient CNN with transfer learning

HardwareDGX agent

arXiv:2606.22400v1 Announce Type: new Abstract: This study introduces a computationally efficient convolutional neural network (CNN) architecture enhanced with transfer learning for multi-cancer detec

Nvidia and DDN target the economics of AI infrastructure

HardwareDGX agent

Realizing value from AI investment has become the ultimate prize for most enterprises today. This requires the ability of data and compute to work together, a key mission for the partnership between N

Nvidia bets on agentic AI to turbocharge biotech discovery

HardwareDGX agent

Artificial intelligence played a prominent role at this week’s Bio International Convention in San Diego, the largest biotech event with vendors spanning the full ecosystem of companies in this indust

NVIDIA Brings Trusted, 24/7 AI Agents to Telecom Operations

HardwareDGX agent

Telecom operators have seen remarkable returns from using generative AI to automate network management, customer care and back-office operations. Most of that impact has been task‑based: automation th

NVIDIA Powers Over 400 of the World’s 500 Fastest Supercomputers

HardwareDGX agent

News Highlights: NVIDIA technology runs 81% of the TOP500 and 90% of the systems new to the list. 26 systems on the TOP500 adopted the NVIDIA Grace CPU, up eight from the previous list. The top eight

ParallelKernelBench: Frontier LLMs can't write fast multi-GPU kernels (yet)

HardwareDGX agent

ParallelKernelBench is a benchmark from Together AI that evaluates frontier large language models' ability to write optimized multi-GPU CUDA kernels, revealing significant limitations in current LLMs'

RAVEN: Agentic RAG for Automated Vulnerability Repair

HardwareDGX agent

arXiv:2606.22647v1 Announce Type: cross Abstract: Automated vulnerability repair has emerged as a promising direction to mitigate the growing number of software vulnerabilities. Recent advances in Lar

Real-time pedestrian attribute recognition with YOLOv8 and ResNet18

HardwareDGX agent

arXiv:2606.21200v1 Announce Type: new Abstract: Pedestrian attribute recognition (PAR) assigns semantic labels to detected pedestrians and is useful in surveillance, video retrieval, and human-centere

Real-World Deployment of Massively Parallel Sampling-Based MPC for Contact-Rich Manipulation

HardwareDGX agent

arXiv:2606.20712v1 Announce Type: new Abstract: Sampling-based Model Predictive Control (SMPC) is a promising strategy for contact-rich robotic manipulation, combining gradient-free optimization with

Scalable Physics-Inspired Transformers for Spin Glasses

HardwareDGX agent

arXiv:2606.22984v1 Announce Type: cross Abstract: Efficient sampling of the Boltzmann distribution in frustrated spin glasses is central to statistical mechanics and combinatorial optimization. Despit

SCENIC: Semantic-Conditioned Edge-Aware Neural Framework for Structured IoT Command Generation

HardwareDGX agent

arXiv:2606.22296v1 Announce Type: new Abstract: Edge Internet of Things (IoT) agents are often constrained by memory capacity, privacy requirements, communication latency, and recurring inference cost

SPOTR: Spatio-temporal Pooling One-Token Reconstruction for Universal Physiological Signal Self-supervised Learning

HardwareDGX agent

arXiv:2606.21973v1 Announce Type: new Abstract: Physiological signals such as EEG, ECG, and PPG are widely used in clinical monitoring. Recent self-supervised learning (SSL) methods offer an attractiv

The Reservoir Attention Network: Cross-Pass State in Pretrained Transformers via Content-Addressable Reservoir Injection

HardwareDGX agent

arXiv:2606.15678v2 Announce Type: replace Abstract: A feasibility and dynamics study of the Reservoir Attention Network (RAN), an architecture that injects a fixed, randomly-initialized reservoir into

The Rise of Alternative Cloud Providers: Why Enterprises Are Rethinking Cloud Strategy

HardwareDGX agent

Enterprises are increasingly evaluating alternative cloud providers beyond the major hyperscalers (AWS, Azure, Google Cloud) due to cost, performance, and vendor lock-in concerns. This shift reflects

Vec-QMDP: Vectorized POMDP Planning on CPUs for Real-Time Autonomous Driving

HardwareDGX agent

arXiv:2602.08334v2 Announce Type: replace Abstract: Planning under uncertainty for real-world robotics tasks, such as autonomous driving, requires reasoning in enormous high-dimensional belief spaces,

We were honored to welcome His Royal Highness Crown Prince Al Hussein bin Abdullah II and Her Royal Highness Princess Rajwa Al Hussein to Re…

HardwareDGX agent

We were honored to welcome His Royal Highness Crown Prince Al Hussein bin Abdullah II and Her Royal Highness Princess Rajwa Al Hussein to Replit in Silicon Valley. It was a privilege to share how AI i

22 Jun 2026

AI networking provider Upscale AI raises 190M at 2B valuation

HardwareDGX agent

Data center networking startup Upscale AI Inc. today announced that it has raised 190 million in fresh funding. The capital was provided as an extension to a 200 million Series A round that the compan

At ISC, JUPITER Shows What Exascale Science Looks Like

HardwareDGX agent

JUPITER, Europe’s first exascale supercomputer at Germany’s Forschungszentrum Jülich, runs on NVIDIA Grace Hopper Superchips and NVIDIA Quantum-X800 InfiniBand networking — and it’s had a busy year. A

Eco Wave Power Turns Waves Into Watts With NVIDIA AI Infrastructure and Digital Twins

HardwareDGX agent

The next era of AI will not be defined by compute alone. Its growth will be determined by energy. As accelerated computing scales across AI factories, agentic AI, industrial AI, edge computing and phy

Enable Real-Time AI for High-Speed Data Acquisition with DAQIRI

HardwareDGX agent

DAQIRI (Data Acquisition for Integrated Real-time Instruments) is a high-performance networking library that streams data from fast detectors . DAQIRI keeps up by handling the stream as it arrives , e

From Materials Simulation to Experimental Astronomy, New NVIDIA AI Software Unlocks Scientific Discoveries

HardwareDGX agent

At the ISC conference running in Hamburg this week, NVIDIA is introducing new software that speeds AI for science, from chemistry and materials discovery to the search for dark matter. The NVIDIA DAQI

GLM 5.2 x AWS GLM-5.2 is accessible via http://Z.ai GLM API on AWS Marketplace 🚀 Powerful long-horizon autonomous workflows, top-tier codin…

HardwareDGX agent

GLM 5.2 x AWS GLM-5.2 is accessible via http://Z.ai GLM API on AWS Marketplace 🚀 Powerful long-horizon autonomous workflows, top-tier coding & multi-step agent reasoning capability, delivered through

GPU rental index continues to slide

HardwareDGX agent

GPU rental prices have declined according to a GPU rental index tracked by Gary Marcus, indicating decreased demand or increased supply in the cloud computing market. This trend suggests potential shi

Groq raised 650M led by Disruptive and Infinitum after its Nvidia deal, aiming to hit 200 MW in capacity by the end of 2027, following a 750M raise in 2025 (Zsana Hoskins/Bloomberg)

HardwareDGX agent

Zsana Hoskins / Bloomberg: Groq raised 650M led by Disruptive and Infinitum after its Nvidia deal, aiming to hit 200 MW in capacity by the end of 2027, following a 750M raise in 2025 — Groq Inc. raise

Hotter Than a Hot Tub: The 45°C Breakthrough to Cool AI’s Biggest Machines

HardwareDGX agent

Hot tubs sit at about 38 to 40 degrees Celsius, warm enough that most people can only soak for about 15 minutes. NVIDIA’s newest AI servers can run their cooling liquid even hotter — up to 45 degrees

How much is CoreWeave threatened by SpaceX’s large GPU contracts?

HardwareDGX agent

Gary Marcus discusses potential competitive threats to CoreWeave, a GPU cloud computing company, from SpaceX's large GPU contracts and infrastructure investments. The post likely examines how SpaceX's

HPE and Kamiwaza rethink AI infrastructure for the inference era

HardwareDGX agent

As AI factories evolve into “data centers of the future,” the infrastructure stack must also transform into a mix of CPU and GPU platforms that can deliver a full set of AI computing solutions. This r

Inference chip startup Groq raises $650M to grow its cloud platform

HardwareDGX agent

Seven months after inking a 20 billion chip licensing deal with Nvidia Corp., Groq Inc. today announced that it has raised 650 million in funding. Growth investment firm Disruptive and hedge fund Infi

Join us in our Discord this time tomorrow for Office Hours! Our topic will be the @NVIDIAAI × @stripe × @NousResearch Hermes Agent Accelerat…

HardwareDGX agent

Join us in our Discord this time tomorrow for Office Hours! Our topic will be the @NVIDIAAI × @stripe × @NousResearch Hermes Agent Accelerated Business Hackathon which concludes at the end of the mont

NAIRR Science Program Reshapes Scientific Research, Powered by NVIDIA AI Infrastructure

HardwareDGX agent

For the past two years, the U.S. National Science Foundation’s National Artificial Intelligence Research Resource (NAIRR) pilot program has driven innovative research across the U.S. for over 700 proj

Nvidia says its AI data center design runs hotter to use a lot less water

HardwareDGX agent

Public pushback against data centers has emphasized their water and energy consumption, and now Nvidia is highlighting its claim that the Rubin generation reference design for a fully liquid-cooled da

NVIDIA Vera CPU Opens the Way for Agentic Scientific AI at Los Alamos National Laboratory

HardwareDGX agent

Mission, Vision and Veritas — new Los Alamos National Laboratory (LANL) supercomputers to be built with HPE and NVIDIA — are tapping NVIDIA Vera CPUs to accelerate scientific discovery, unlocking agen

Running ComfyUI workflows on Amazon SageMaker AI processing jobs

HardwareDGX agent

In this post, we walk you through how to deploy ComfyUI workflows on Amazon SageMaker AI processing jobs to generate hundreds of high-quality images in a single batch. You learn how to set up the infr

SpaceX signs a computing deal worth up to 6.3B with Reflection AI for access to Nvidia GB300s at Colossus 2; Reflection will pay 150M per month through 2029 (Deirdre Bosa/CNBC)

HardwareDGX agent

Deirdre Bosa / CNBC: SpaceX signs a computing deal worth up to 6.3B with Reflection AI for access to Nvidia GB300s at Colossus 2; Reflection will pay 150M per month through 2029 — SpaceX has signed a

The next generation of inference needs purpose-built infrastructure. Together AI and 5C are deploying NVIDIA GB300 NVL72 systems with high-d…

HardwareDGX agent

The next generation of inference needs purpose-built infrastructure. Together AI and 5C are deploying NVIDIA GB300 NVL72 systems with high-density compute, advanced cooling, and AI-optimized storage f

The Secret Life of Computers

HardwareDGX agent

This article likely explores the hidden operational processes and internal workings of computers that users typically don't observe, such as microprocessor functions, system-level operations, or the b

This approach will mean that American robotics companies can't access the hardware they need for R&D.

HardwareDGX agent

This approach will mean that American robotics companies can't access the hardware they need for R&D. Unitree was recently designated as a Chinese military company and its products are a threat to our

True

HardwareDGX agent

True Water usage has been a hot topic in the AI data center world, but the numbers may surprise you. According to the Manhattan Institute, data centers use 0.2 percent of daily water usage in the U.S.

Upscale, which is building AI networking infrastructure to rival Cisco, raised a 190M Series A-1 at a 2B valuation, up from 1B after raising 200M in January (Lily Mae Lazarus/Fortune)

HardwareDGX agent

Lily Mae Lazarus / Fortune: Upscale, which is building AI networking infrastructure to rival Cisco, raised a 190M Series A-1 at a 2B valuation, up from 1B after raising 200M in January — Every time co

← Previous
1…910111213…29
Next →