GLM 5.2 NVFP4 REAP 469B running on 3X DGX Spark. 256K context @ ~4.4 tok/s. based on work done by @0xSero
github.com/bird/GLM-spark
GLM 5.2 NVFP4 REAP 469B is running on a 3× DGX Spark cluster, delivering a 256,000-token context window with about 4.4 tokens/second throughput. The implementation references prior work by @0xSero, and the repository for the GLM-spark setup is published at github.com/bird/GLM-spark.
GLM 5.2 NVFP4 REAP 469B model is reported running on a 3x DGX Spark cluster.
GLM 5.2 NVFP4 REAP 469B running on 3X DGX Spark. 256K context @ ~4.4 tok/s. based on work done by @0xSero
github.com/bird/GLM-spark