Twitter/X

Unsloth AI announced on 2026-06-18 that GLM-5.2 can be run locally and claims it…

Brief

GLM-5.2 is now runnable locally per Unsloth AI (2026-06-18). They report a 2-bit quantization that keeps ~82% accuracy while reducing model size from 1.51 TB to 238 GB (–84%), enabling use on 256 GB Macs or similar RAM/VRAM configurations; guide and GGUF release are provided via their docs and Hugging Face.

Why it matters

Unsloth AI announced on 2026-06-18 that GLM-5.2 can be run locally and claims it is “the strongest open model to date.”

Key details

  • They published a 2-bit quantized GLM-5.2 that retained ~82% accuracy after shrinking from 1.51 TB to 238 GB (–84% size), enabling runs on a 256 GB Mac or comparable RAM/VRAM setups; guide and GGUF files available on their docs and Hugging Face links.
Source evidence

2x sparks 3x 6000s are in for a treat

Unsloth AI (@UnslothAI)

GLM-5.2 can now be run locally!🔥

The 2-bit model retains ~82% accuracy after we shrunk it from 1.51TB to 238GB (-84% size).

Run on a 256GB Mac or RAM/VRAM setups.

GLM-5.2 is the strongest open model to date.

Guide: unsloth.ai/docs/models/glm-5…
GGUF: huggingface.co/unsloth/GLM-5…

— https://nitter.net/UnslothAI/status/2067588262156501497#m