Twitter/X

0xSero announces the release of Gemma-4-21B-REAP and frames it as a follow-up to…

Brief

0xSero announces the release of Gemma-4-21B-REAP and frames it as a follow-up to a prior promise. The post claims the model not only preserves performance but improves reasoning accuracy, and highlights practical local-running requirements: 12GB of VRAM for partial-context use or 16GB for full context.

Source evidence

title: @0xSero: You thought I was done??? Bam, another one

huggingface.co/0xSero/gemma-…

0xSero (@0xSero)

A...
author: @0xSero
contenttype: tweet
publication: Twitter/X
published: 2026-04-05T18:26:36+00:00
source
url: https://x.com/0xSero/status/2040858658502767019

word_count: 59

You thought I was done??? Bam, another one

huggingface.co/0xSero/gemma-…

0xSero (@0xSero)

As promised!

Gemma-4-21B-REAP is out! Results are great it held up really well and actually gained accuracy on reasoning tasks.

MLX & GGUF bros do you thing!

This should fit on as little as 12GB of vram with some context, or 16GB with full context

huggingface.co/0xSero/gemma-…

— https://nitter.net/0xSero/status/2040822269400723955#m