Twitter/X

DS4 by @antirez is presented as a lean inference engine for DeepSeek V4 Flash…

Brief

DS4 by @antirez is presented as a lean inference engine for DeepSeek V4 Flash that the author ran on a Mac Studio M3 Ultra 256GB and called “seriously impressive.” The post claims a 1M-token context window, strong coherence, and solid speed on consumer hardware, arguing this enables frontier-level reasoning locally rather than requiring large GPU clusters.

Source evidence

Just fired up DS4 by @antirez on my Mac Studio M3 Ultra 256GB and man, it’s seriously impressive. A clean, purpose-built engine for DeepSeek V4 Flash that actually makes frontier-level reasoning feel usable locally.
1M context, strong coherence, and solid speed on consumer hardware. This is the kind of focused, no-bullshit effort that finally brings real frontier models to regular machines instead of just giant GPU clusters. Huge respect @antirez — thank you for building this 🔥
github.com/antirez/ds4