Local Ai
Luke Alonso has uploaded an NVFP4 of GLM 5.2 467GB, would fit on 4x DGX Sparks (~$20k)
Luke Alonso has released a quantized version (NVFP4 format) of GLM 5.2 totaling 467GB in size, which can be deployed across four DGX Spark systems with an estimated hardware cost of approximately $20,
Luke Alonso has released a quantized version (NVFP4 format) of GLM 5.2 totaling 467GB in size, which can be deployed across four DGX Spark systems with an estimated hardware cost of approximately $20,000. This development makes the large language model more accessible for organizations seeking to run it on affordable GPU infrastructure.
Source: Clem Delangue (X) | 2026-06-20