DeepSeek V4 Flash 0731 can now be run locally! 🐳
Run DeepSeek V4 Flash lossless 4-bit on 168GB RAM and 3-bit on 110GB RAM.
V4 Flash 0731 outperforms V4 Pro. Run via Unsloth or llama.cpp. Smaller quants coming today.
Guide: unsloth.ai/docs/models/deeps…
GGUF: huggingface.co/unsloth/DeepS…
DeepSeek (@deepseek_ai)
🚀 DeepSeek-V4-Flash Official API is now LIVE in public beta!
🔷 We’ve massively upgraded its Agent capabilities—benchmark scores are now far surpassing the V4-Pro-Preview. Check out the massive performance leap below! 👇
🔷 The official V4-Flash now natively supports the Responses API format and is fully adapted for Codex!
Check out the configuration details in our official API docs: api-docs.deepseek.com/quick_…
— https://nitter.net/deepseek_ai/status/2083084415157022911#m