Get the app
LLMs

Kimi K3 is the first open model to crack 2.8T params

Moonshot AI's Kimi K3 becomes the largest open-weight model ever, tripling its predecessor and topping Claude Opus 4.8 on coding.

Kimi K3 is the first open model to crack 2.8T params

China's Moonshot AI just dropped Kimi K3 — the first open model in the 3-trillion-parameter class, landing at 2.8T parameters, nearly triple the size of Kimi K2.6. It's now the largest open-weight AI model on the planet.

Under the hood: a new hybrid linear attention scheme called Kimi Delta Attention (KDA), Attention Residuals, native visual understanding, and a 1M-token context window. On Moonshot's own evals it beat every open rival — including Claude Opus 4.8 and GPT 5.5 — across coding and agentic benchmarks, and even edged out Claude Fable 5 in the Frontend Code Arena. It still trails Fable 5 and GPT 5.6 Sol on overall performance.

Weights ship publicly by July 27 under a Modified MIT license — the latest jolt in China's open-weight push around US compute limits, and it's already reigniting regulation talk.

Why it matters: frontier-scale capability is now something you can download, not just rent from an API.

Sources

Independent coverage

Written by an AI pipeline from the sources above. Methodology · Report an error

The daily AI brief, on your phone.

Feed, daily deep-dive and bytes — readable offline, with push alerts for the topics you follow.

Get it on Google Play