Applied this method to Qwen3.6-35B buys you more context on 24gb/32gb
huggingface.co/0xSero/Qwen3.…
Link
0xSero/Qwen3.6-35B-Hyrbid-3.25bpw · Hugging Face
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
huggingface.co
0xSero (@0xSero)
Article
GLM-5.2 at home for 15,000$
Here's how a community of tinkerers got GLM-5.2 running at near lossless quality on a 15,000$ budget. This article was put together using tens of thousands of analyzed messages in the RTX Pro 6000
— https://nitter.net/0xSero/status/2079230064840106173#m