r/LocalLLaMA

I CANNOT believe I've got DeepSeek-V4-Flash-0731, a frontier model, running on my home PC. Insane!

Brief

u/mintybadgerme claims they successfully ran DeepSeek-V4-Flash-0731 (Q3 quant) on a modest Intel Windows machine with ~24 GB VRAM, noting the run is slow but viable. The post highlights the short ~20-month timeline from cloud-only models to local execution and speculates this shift pressures major cloud providers; shared on r/LocalLLaMA with a screenshot link.

Why it matters

OP reports running DeepSeek-V4-Flash-0731 (Q3 quant) locally on an Intel Windows PC with about 24 GB VRAM, claiming it's functional though 'slow as porridge'.

Key details

  • Author emphasizes rapid progress: ability arrived in under 20 months from cloud-only to local execution, and suggests this development worries large providers; post posted to r/LocalLLaMA with an image link (https://ibb.co/zTvqR8YR).
Source evidence

So this is the stuff of absolute insanity. In less than 20 months we've gone from super expensive cloud models only, to being able to run a Q3 quant of DeepSeek on an Intel Windows PC with a very average 24GB of VRAM. No wonder the big boys are panicking (and yes it's slow as porridge). https://ibb.co/zTvqR8YR

Subreddit: r/LocalLLaMA