Twitter/X

NVIDIA released a quantized model variant GLM‑5.2‑NVFP4 on Hugging Face…

Brief

NVIDIA’s GLM‑5.2‑NVFP4 quantized model is now available on Hugging Face (nvidia/GLM-5.2-NVFP4). Announced on X by Barathwaj Anandan on 2026‑06‑26, the NVFP4 quant is approximately 465 GB and promoted as a quality quant release, with a playful reference to Blackwell hardware performance.

Why it matters

NVIDIA released a quantized model variant GLM‑5.2‑NVFP4 on Hugging Face (repository: nvidia/GLM-5.2-NVFP4).

Key details

  • Announcement by Barathwaj Anandan on X (posted 2026-06-26) states the NVFP4 quant is ~465 GB and marketed as a high‑quality quant.
  • Post references 'Blackwell go brrr', implying performance/efficiency excitement relative to Blackwell-class hardware.
Source evidence

huggingface.co/nvidia/GLM-5.…

Link

nvidia/GLM-5.2-NVFP4 · Hugging Face

We’re on a journey to advance and democratize artificial intelligence through open source and open science.
huggingface.co

Barathwaj Anandan (@BarathAnandan7)

Good news. We cooked!

@NVIDIAAI GLM 5.2 NVFP4 is out for anyone who's been waiting on a quality quant.

Size ~465GB.

Link below.

Blackwell go brrr

— https://nitter.net/BarathAnandan7/status/2070351192165847308#m