r/LocalLLaMA

Qwen3.8-Max matches Kimi K3 and DeepSeek V4 Flash

Brief

Qwen3.8-Max (2.4T) is presented as a major open-weight release that benchmarks comparably to Kimi K3 and DeepSeek V4 Flash and is claimed to be stronger on coding/software tasks. The author says Qwen3.8-27B will also be open weight soon and that Qwen3.8-Max weights will be released the week after the 2026-08-03 post; pricing tiers for input, output, and caching were listed. No community responses were included.

Why it matters

Qwen3.8-Max (2.4T) is reported to match Kimi K3 and DeepSeek V4 Flash across benchmark categories and is claimed to outperform them on coding and software tasks (post dated 2026-08-03).

Key details

  • The poster says Qwen3.8-27B will be released as open weights soon and that Qwen3.8-Max weights are scheduled to be published the week following the 2026-08-03 post.
  • Pricing posted for model use: Input $2.0 per million tokens, Output $6.0 per million tokens, Implicit Caching $0.25 per million tokens; original content posted to r/LocalLLaMA with no community comments included in the submission.
Source evidence

Qwen3.8-Max (2.4T) is another massive contribution to the open weight community. On benchmarks, it performs closely to Kimi K3 and DeepSeek V4 flash across all categories and is better at coding and software tasks. Qwen3.8-27B will also be open weight soon too. Weights are being released next week.
Pricing:
Input: $2.0 / M tokens
Output: $6.0 / M tokens
Implicit Caching: $0.25 / M tokens

Link: https://i.redd.it/14mqdzhzb7hh1.png

Subreddit: r/LocalLLaMA