AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “hardware”

GridTimelineEvolution
4,457 results
8 Jun 2026

VIRTUS-FPP: Virtual Sensor Modeling for Fringe Projection Profilometry in NVIDIA Isaac Sim

HardwareDGX agent

arXiv:2509.22685v2 Announce Type: replace-cross Abstract: Fringe projection profilometry (FPP) is a high-precision structured-light sensing technique for 3D surface reconstruction, yet its practical d

Xiaomi just claimed 1,000+ tps on a 1T model using a standard 8-GPU server

HardwareDGX agent

Xiaomi achieved over 1,000 tokens per second output from a 1 trillion-parameter model using a single standard 8-GPU commodity node through extreme model-system codesign . The approach combines FP4 qua

7 Jun 2026

American Open Source is so back. 9 / 30 of the models on page 1 of Huggingface are published by Nvidia.

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Hardware
DGX agent

Nvidia has significantly increased its presence in open-source AI models, with 9 out of 30 models on the first page of Hugging Face's model hub published by the company. This observation reflects Nvid

Donations to Jensen Huang’s retirement fund*, ordered by magnitude. *strictly speaking not every penny went to Nvidia.

HardwareDGX agent

This post likely presents a satirical or critical commentary on Jensen Huang's wealth accumulation, possibly comparing his personal gains to Nvidia's financial performance or suggesting that his retir

NVIDIA and Doosan Group Collaborate to Advance Physical AI and AI Factory Infrastructure

HardwareDGX agent

NVIDIA and Doosan Group are expanding their collaboration to advance new opportunities across physical AI, robotics and AI factory infrastructure, spanning Doosan Robotics, Doosan Bobcat, Doosan Enerb

Nvidia and SK Hynix sign a multiyear chip deal to develop next-gen memory; SK Hynix will diversify into infrastructure, physical AI, and memory for Vera Rubin (Yoolim Lee/Bloomberg)

HardwareDGX agent

Yoolim Lee / Bloomberg: Nvidia and SK Hynix sign a multiyear chip deal to develop next-gen memory; SK Hynix will diversify into infrastructure, physical AI, and memory for Vera Rubin — Nvidia Corp. an

NVIDIA, KRAFTON, NC and Reigning ‘League of Legends’ Champions T1 Celebrate RTX Spark at Korea’s PC Bangs

HardwareDGX agent

At GTC Taipei at COMPUTEX last week, NVIDIA unveiled RTX Spark, the superchip that reinvents Windows PCs for the era of personal AI agents. On the heels of this announcement, NVIDIA founder and CEO Je

Nvidia says South Korea-based Naver will use Nvidia's DSX tech to build AI factories at 'gigawatt scale' to meet global demand for AI services and physical AI (Heekyong Yang/Reuters)

HardwareDGX agent

Heekyong Yang / Reuters: Nvidia says South Korea-based Naver will use Nvidia's DSX tech to build AI factories at “gigawatt scale” to meet global demand for AI services and physical AI — Nvidia (NVDA.O

6 Jun 2026

Beyond Code Pairs: Dialogue-Based Data Generation for LLM Code Translation

HardwareDGX agent

arXiv:2512.03086v2 Announce Type: replace-cross Abstract: Large language models (LLMs) have shown remarkable capabilities in code translation, yet their performance deteriorates in low-resource progra

CuTeGen: An LLM-Based Agentic Framework for Generation and Optimization of High-Performance GPU Kernels using CuTe

HardwareDGX agent

arXiv:2604.01489v2 Announce Type: replace-cross Abstract: High-performance GPU kernels are critical to modern machine learning systems, yet developing them remains a manual, expert-driven process. Rec

I don’t know whether to take these specific details like this seriously, but something like this will inevitably happen, sooner or later — a…

HardwareDGX agent

I don’t know whether to take these specific details like this seriously, but something like this will inevitably happen, sooner or later — and absolutely devastate all the data infrastructure investme

LeWorldModel: Stable End-to-End Joint-Embedding Predictive Architecture from Pixels

HardwareDGX agent

arXiv:2603.19312v3 Announce Type: replace-cross Abstract: Joint Embedding Predictive Architectures (JEPAs) offer a compelling framework for learning world models in compact latent spaces, yet existing

Mac mini M4 vs Pc with Nvidia 5060 8gb for ai workloads?

HardwareDGX agent

The Mac mini M4 uses unified memory architecture where CPU and GPU share a single 24GB memory pool, while the RTX 5060 has dedicated VRAM. Mac mini M4 is preferred for large model inference (70B param

RedKnot: Efficient Long-Context LLM Serving with Head-Aware KV Reuse and SegPagedAttention

HardwareDGX agent

arXiv:2606.06256v1 Announce Type: new Abstract: As the input length of large language model (LLM) serving continues to grow, the KV cache has become a dominant bottleneck in AI infrastructure. It limi

They said the reason for the sell off was about interest rates , but I agree with @GaryMarcus that it might be about the faltering IPOs pros…

HardwareDGX agent

They said the reason for the sell off was about interest rates , but I agree with @GaryMarcus that it might be about the faltering IPOs prospects... Absolutely brutal day for AI fantasies: Nvidia NVDA

5 Jun 2026

Absolutely brutal day for AI fantasies: Nvidia NVDA: down 6.2% Broadcom AVGO: down 7.92% Coreweave CRWV: down 7.07% Nebius NBIS: down 12…

HardwareDGX agent

Absolutely brutal day for AI fantasies: Nvidia NVDA: down 6.2% Broadcom AVGO: down 7.92% Coreweave CRWV: down 7.07% Nebius NBIS: down 12.27% Oracle $ORCL: down 9.59% Worst of all? OpenAI is rumored to

b9523

Local AiDGX agent

b9523 is a release build of llama.cpp, an open-source tool for LLM inference in C/C++ that enables language model execution with minimal setup on a wide range of hardware locally and in the cloud. The

b9529

Local AiDGX agent

b9529 is a release tag for llama.cpp, a C/C++ implementation of LLM inference that enables running large language models locally on consumer hardware. This specific release represents a particular ver

Check out the NVIDIA blog!

HardwareDGX agent

Check out the NVIDIA blog! This week at #CVPR2026, NVIDIA Research is presenting three papers across physical ai that offer groundbreaking solutions for training at scale across diverse applications:

Filing: Google has agreed to pay SpaceX $920M per month for access to Nvidia chips as part of a cloud-services deal that runs through mid-2029 (Bloomberg)

HardwareDGX agent

Bloomberg: Filing: Google has agreed to pay SpaceX 920M per month for access to Nvidia chips as part of a cloud-services deal that runs through mid-2029 — Alphabet Inc.'s Google has agreed to pay Elon

Flash-WAM: Modality-Aware Distillation for World Action Models

HardwareDGX agent

arXiv:2606.05254v1 Announce Type: cross Abstract: World-action models (WAMs) jointly generate future video and robot actions through iterative diffusion, achieving strong performance on manipulation b

Gotta Grow Fast: Design and Benchmarking of a Tip Mount for High-Speed Vine Robots

HardwareDGX agent

arXiv:2606.06040v1 Announce Type: new Abstract: Soft, growing vine robots extend through tip eversion, a mechanism that enables navigation through cluttered environments. However, integrating cameras

GS-NFS: Bandwidth-adaptive Streaming of Dynamic Gaussian Splats and Point Clouds

HardwareDGX agent

arXiv:2606.05650v1 Announce Type: cross Abstract: Dynamic 3D Gaussian Splatting (3DGS) holds great promise as a 3D video streaming technology since it can represent complex 3D scenes with high fidelit

I absolutely agree that there really is this 10x opportunity for companies to be $40 trillion in market cap and beyond—perhaps Nvidia, Googl…

HardwareDGX agent

I absolutely agree that there really is this 10x opportunity for companies to be 40 trillion in market cap and beyond—perhaps Nvidia, Google, and beyond. Really fascinating to consider what that could

If scale was “all you need”, Elon would be hoarding LLMs, not leasing them.

HardwareDGX agent

If scale was “all you need”, Elon would be hoarding LLMs, not leasing them. Last year: people had to steal GPU’s from armored trucks. This year: SpaceX is leasing to GPUs left, right, and center, beca

Last year: people had to steal GPU’s from armored trucks. This year: SpaceX is leasing to GPUs left, right, and center, because Xai couldn’t…

HardwareDGX agent

Last year: people had to steal GPU’s from armored trucks. This year: SpaceX is leasing to GPUs left, right, and center, because Xai couldn’t figure out what to do with them. SpaceX just disclosed a ne

Semantic-decoupled Spatial Partition Guided Point-supervised Oriented Object Detection

HardwareDGX agent

arXiv:2506.10601v2 Announce Type: replace Abstract: Given its ability to reduce annotation costs, weakly supervised learning based on single-point annotations has emerged as a research focus in orient

Seoul Purpose: How NVIDIA and South Korea Are Building the Future of AI

HardwareDGX agent

Home to cutting-edge sovereign AI infrastructure and robotics innovators, as well as one of the world’s most passionate gaming communities, South Korea is one of the world’s centers of AI. NVIDIA foun

This is your laptop… on AI

HardwareDGX agent

We're now deep into developer conference season, and one of the themes so far is the relentless conviction from Big Tech companies that AI is going to change everything about how we do everything. Nvi

Today we published a technical blog post about Ideogram 4.0 — our goal is to enable more innovation and creativity. It's a 9.3B Diffusion Tr…

HardwareDGX agent

Today we published a technical blog post about Ideogram 4.0 — our goal is to enable more innovation and creativity. It's a 9.3B Diffusion Transformer trained from scratch, paired with a frozen 8B VLM

Uncertainty-Aware Adaptive Sensor Fusion for Autonomous Navigation

HardwareDGX agent

arXiv:2606.05437v1 Announce Type: cross Abstract: This work introduces a hybrid deep learning approach integrated with an Unscented Kalman Filter (UKF) to enhance pose estimation accuracy in Visual-In

US-traded chipmakers plunged on Friday, after Broadcom missed expectations; Nvidia fell 6.19%, Micron fell 13.25%, AMD fell 10.86%, and Broadcom 7.92% (Noel Randewich/Reuters)

HardwareDGX agent

Noel Randewich / Reuters: US-traded chipmakers plunged on Friday, after Broadcom missed expectations; Nvidia fell 6.19%, Micron fell 13.25%, AMD fell 10.86%, and Broadcom 7.92% — U.S.-traded chipmaker

World-Language-Action Model for Unified World Modeling, Language Reasoning, and Action Synthesis

HardwareDGX agent

arXiv:2606.05979v1 Announce Type: new Abstract: We propose world-language-action (WLA) models as a new class of embodied foundation models. WLA takes textual instructions, images, and robot states as

4 Jun 2026

AgentJet: A Flexible Swarm Training Framework for Agentic Reinforcement Learning

HardwareDGX agent

arXiv:2606.04484v1 Announce Type: new Abstract: We present AgentJet, a distributed swarm training framework for large language model (LLM) agent reinforcement learning. Unlike centralized frameworks t

b9515

Local AiDGX agent

llama.cpp is a C/C++ implementation designed to enable LLM inference with minimal setup and state-of-the-art performance on a wide range of hardware locally and in the cloud. Build b9515 is an interme

Cartridges at Scale: Training Modular KV Caches over Large Document Collections

HardwareDGX agent

arXiv:2606.04557v1 Announce Type: new Abstract: Large Language Models can reason over long contexts, yet prefilling millions of tokens is wasteful as much of the content remains static across queries.

DSA: Dynamic Step Allocation for Fast Autoregressive Video Generation

HardwareDGX agent

arXiv:2606.04432v1 Announce Type: new Abstract: Video diffusion transformers have achieved state-of-the-art visual quality, but their high inference cost remains a major bottleneck for real-time appli

Ex-OpenAI Tech Lead, Justin Lebar joins SemiAnalysis as an Visiting Fellow to Burn $10,000 in 3 hours to find dozens of AMDGPU LLVM, x86 LLV…

HardwareDGX agent

Ex-OpenAI Tech Lead, Justin Lebar joins SemiAnalysis as an Visiting Fellow to Burn $10,000 in 3 hours to find dozens of AMDGPU LLVM, x86 LLVM, NVPTX bugs 00:00 - Intro & Justin’s background 00:59 - Ho

Fitting scattered data with optional monotonicity constraints on GPU: LipFit package

HardwareDGX agent

arXiv:2606.04670v1 Announce Type: cross Abstract: This paper presents a method of multivariate scattered data interpolation and approximation that produces optimal Lipschitz-continuous approximation,

Forecast: Fun Ahead — 18 Games Join in June to Stream on GeForce NOW

HardwareDGX agent

June’s forecast with GeForce NOW: 100% chance of gaming. GeForce NOW is lining up new adventures for the month, from big-name blockbusters to quirky indies ready for the spotlight. Members can dive in

Generalist AI raises 400M at 2B valuation to build general intelligence for robotics

HardwareDGX agent

Artificial intelligence startup Generalist AI Inc., a startup building embodied robotics intelligence, said today it has raised 400 million in new funding, bringing the company’s valuation to 2 billio

I love that people reposting what we said miss most of what the note says. Happens all the time.

HardwareDGX agent

Dylan Patel observes that when people repost or share his notes on social media, they frequently misrepresent or overlook the main points of his original message. He notes this is a recurring pattern

LLM Compression with Jointly Optimizing Architectural and Quantization choices

HardwareDGX agent

arXiv:2606.04063v1 Announce Type: cross Abstract: Deploying large language models (LLMs) is challenging due to their significant memory and computational requirements. While some methods address this

Nvidia snaps up Kumo AI, a predictive AI startup known for its extreme accuracy

HardwareDGX agent

Nvidia Corp. has bagged itself another artificial intelligence startup, acquiring four-year-old model maker Kumo AI Inc. The company designs AI models focused on making extremely accurate business pre

OpenAI is in deep, deep trouble They are low on capital relative to the massive cash they are burning and there is only so much capital in t…

HardwareDGX agent

OpenAI is in deep, deep trouble They are low on capital relative to the massive cash they are burning and there is only so much capital in the world. No rational person would sell their Bitcoin or Nvi

Scaling Novel Graph Generation via Lightweight Structure-Guided Autoregressive Models

HardwareDGX agent

arXiv:2606.04287v1 Announce Type: cross Abstract: Generating realistic and diverse graphs is a key problem in machine learning, with applications in molecular discovery, circuit design, cybersecurity,

SparDA: Sparse Decoupled Attention for Efficient Long-Context LLM Inference

HardwareDGX agent

arXiv:2606.04511v1 Announce Type: new Abstract: Sparse attention reduces compute and memory bandwidth for long-context LLM inference. However, two key challenges remain: (1) KV cache capacity still gr

StepPRM-RTL: Stepwise Process-Reward Guided LLM Fine-Tuning for Enhanced RTL Synthesis

Model ReleasesDGX agent

arXiv:2606.04246v1 Announce Type: new Abstract: Automatic generation of RTL code for digital hardware designs remains challenging due to long-horizon reasoning, multi-step dependencies, and strict cor

TITAN-FedAnil+: Trust-Based Adaptive Blockchain Federated Learning for Resource-Constrained Intelligent Enterprises

HardwareDGX agent

arXiv:2606.04388v1 Announce Type: cross Abstract: Federated Learning (FL) has emerged as an effective paradigm for collaborative intelligence while preserving data privacy. However, data heterogeneity

Together AI provides the inference stack behind both: high-throughput serving on the latest NVIDIA Blackwell GPUs for agentic workloads, and…

HardwareDGX agent

Together AI provides the inference stack behind both: high-throughput serving on the latest NVIDIA Blackwell GPUs for agentic workloads, and TensorRT engines plus event-driven streaming I/O for low-la

3 Jun 2026

Announcing Foundry Managed Compute: Run open models in Microsoft Foundry

HardwareDGX agent

Microsoft Foundry Managed Compute is a new GPU platform-as-a-service for hosting open-source and custom AI models behind the same endpoint, SDKs, and bill as frontier models. The post Announcing Found

CoreWeave’s Vera Rubin milestone sets stage for theCUBE’s agentic AI coverage

HardwareDGX agent

The agentic AI era is putting new pressure on the infrastructure stack, and CoreWeave Inc.’s latest milestone gives the conversation a sharper edge. This week, the company announced that it has comple

Floating Point: The Origin Story

HardwareDGX agent

This article explores the historical development and origins of floating-point number representation in computing, likely covering how early computer scientists and engineers designed methods to repre

GRZO: Group-Relative Zeroth-Order Optimization for Large Language Model Fine-Tuning

HardwareDGX agent

arXiv:2606.02857v1 Announce Type: cross Abstract: Zeroth-order (ZO) optimization is a memory-efficient alternative to backpropagation for fine-tuning large language models, but its deployment is limit

MOSAIC: Efficient Mixture-of-Agent Scheduling via Adaptive Aggregation and Inference Concurrency

HardwareDGX agent

arXiv:2606.03014v1 Announce Type: new Abstract: Mixture-of-Agents (MoA) systems improve reasoning accuracy by routing each query to multiple expert LLMs and aggregating their outputs. Efficiently exec

Nvidia acquired Kumo, which sells predictive AI software to enterprises, a source says for 400M+; PitchBook: Kumo raised 37M at a $250M valuation in 2022 (The Information)

HardwareDGX agent

The Information: Nvidia acquired Kumo, which sells predictive AI software to enterprises, a source says for 400M+; PitchBook: Kumo raised 37M at a 250M valuation in 2022 — Nvidia has bought Kumo AI, a

NVIDIA Enables the Next Era Of Physical AI Research With Agent Skills For Autonomous Vehicles, Robotics And Vision AI

HardwareDGX agent

At CVPR, NVIDIA is unveiling new physical AI agent skills that help researchers and developers speed the development of autonomous vehicles, robots and vision AI systems. The core challenge in physica

NVIDIA Isaac Sim: Enabling Scalable, GPU-Accelerated Simulation for Robotics

HardwareDGX agent

arXiv:2606.03551v1 Announce Type: new Abstract: Simulation has become a core infrastructure for robotics research. Unlike previous simulators, NVIDIA Isaac Sim leverages GPU acceleration to enable lar

NVIDIA Research Unlocks Advanced Grasping, Smarter Autonomous Driving and Agent Training at Scale

HardwareDGX agent

What makes a robot gripper useful isn’t that it can pick up one object — it’s that it can pick up the next one, and the one after that, with a tool it’s never held before. What makes an autonomous veh

Pruning Deep Neural Networks via the Marchenko--Pastur Distribution

HardwareDGX agent

arXiv:2606.02608v1 Announce Type: new Abstract: We study a Marchenko--Pastur (MP) random-matrix approach to pruning deep neural networks with very small post-pruning fine-tuning budgets. The main prac

← Previous
1…2021222324…75
Next →