Twitter/X

Kimi K3 is a 2.8 trillion-parameter, native-multimodal model with a…

Brief

Kimi K3 is being released with open weights on July 27, 2026: a 2.8T-parameter, native-multimodal model with a 1M-token context. Kimi touts Kimi Delta Attention for up to 6.3× faster decoding and Attention Residuals for ~25% training efficiency gains, and claims the model can be run privately on a $12,000 setup (5–10× cheaper than prior options).

Why it matters

Kimi K3 is a 2.8 trillion-parameter, native-multimodal model with a 1,000,000-token context; Kimi announced open weights available July 27, 2026 on Kimi.com, Kimi Work, Kimi Code, and the Kimi API.

Key details

  • Kimi claims Kimi Delta Attention yields up to 6.3× faster decoding in million-token contexts and Attention Residuals give ~25% higher training efficiency at <2% additional cost.
  • The post asserts you can run this frontier model privately on a $12,000 setup—claimed to be 5–10× cheaper than predecessors—though it notes that price is still not affordable for everyone.
Source evidence

Kimi k3’s open weights go live tmrw!

you can run a frontier model privately and low-cost on a $12,000 setup

definitely not affordable for everyone but dramatically cheaper (5-10X) than any predecessor

models are getting more intelligent and LESS costly.

Kimi.ai (@Kimi_Moonshot)

Introducing Kimi K3: Open Frontier Intelligence

🔹 2.8 Trillion Parameters, 1 Million Context, Native Multimodal
🔹 Kimi Delta Attention enables up to 6.3x faster decoding in million-token contexts
🔹 Attention Residuals deliver ~25% higher training efficiency at <2% additional cost
🔹 Built for long-horizon agentic coding and self-evolving workflows

Kimi K3 is now live on on Kimi.com, Kimi Work, Kimi Code, and the Kimi API.
Open Weights by July 27, 2026.

🔗 API: platform.kimi.ai
🔗 Tech blog: kimi.com/blog/kimi-k3

— https://nitter.net/Kimi_Moonshot/status/2077830229968683203#m