Twitter/X

Kimi.ai released Kimi K3

Brief

Kimi K3 is a newly released 2.8T MoE model from Kimi.ai with native vision capabilities and a 1M‑token context window; the team asserts a 2.5× intelligence-per-compute improvement and open-sourced weights, a technical report, high-performance attention kernels, an MoE communication library, and agent-run infrastructure (links on Hugging Face, GitHub, and the company blog).

Why it matters

Kimi.ai released Kimi K3: a 2.8 trillion-parameter Mixture-of-Experts (MoE) model with native visual understanding and a 1,000,000‑token context window (announced 2026-07-28).

Key details

  • They claim a new architecture that delivers 2.5× the intelligence per unit of compute and published model weights, a technical report, attention kernels, an MoE communication library, and agent-infrastructure on Hugging Face/GitHub/tech blog.
Source evidence

"API costs are too high, let's just self-host this 2.8T model on our own."

Kimi.ai (@Kimi_Moonshot)

Releasing the model weights and technical report of Kimi K3.

Kimi K3 is our most capable model: a 2.8T MoE model with native visual understanding and a 1M-token context window.

New model architecture: 2.5x the intelligence per unit of compute, not just more params.

Alongside Kimi K3, we're opening up more of the stack behind it — high-performance attention kernels, MoE communication library, and infrastructure for running agent environments at scale.

Model weights: huggingface.co/moonshotai/Ki…
Tech report: github.com/MoonshotAI/Kimi-K…
Tech blog: kimi.com/blog/kimi-k3

— https://nitter.net/Kimi_Moonshot/status/2081760186235289764#m