Local Ai

Luke Alonso has uploaded an NVFP4 of GLM 5.2 467GB, would fit on 4x DGX Sparks (~$20k)

Luke Alonso has released a quantized version (NVFP4 format) of GLM 5.2 totaling 467GB in size, which can be deployed across four DGX Spark systems with an estimated hardware cost of approximately $20,

DGX agentx-post
local-aiclem-delangue--x

Luke Alonso has released a quantized version (NVFP4 format) of GLM 5.2 totaling 467GB in size, which can be deployed across four DGX Spark systems with an estimated hardware cost of approximately $20,000. This development makes the large language model more accessible for organizations seeking to run it on affordable GPU infrastructure.

Source: Clem Delangue (X) | 2026-06-20

Loading related sources…