Frontline Hotspot
Frontline Hotspot

Kimi K3 Goes Open Source Tonight: 2.8 Trillion Parameters, World's Largest, As Yang Zhilin Closes the China-US Model Gap to 3 Months

On the evening of July 27, Moonshot AI open-sourced Kimi K3's weights: a 2.8-trillion-parameter MoE with a 1-million-token context, the world's largest open-source model, benchmarking against Anthropic's Fable 5. From the July 16 API launch to tonight's weight release, Yang Zhilin used ten days to compress the China-US model gap from 6-9 months to 3-5. Breakdown of specs, benchmarks, the comeback story, and the White House accusation.

Published July 27, 20263 min read
<!-- kimi-k3-open-weights-opensource | hotspot | Kimi K3 Open Source Tonight -->

On the evening of July 27, Moonshot AI posted Kimi K3's model weights to its download page, free, commercially usable, and self-hostable. It is the world's first open-source model to cross the two-trillion-parameter mark while letting developers freely download and modify it. From the API and app launch on July 16 to tonight's weight release, only ten days have passed.

How Big Is 2.8 Trillion Parameters

K3 has 2.8 trillion total parameters, built on Stable MoE with 16 of 896 experts active per layer. It has a 1-million-token context window, native multimodal vision, and a built-in always-on "thinking mode." For comparison, DeepSeek V4 Pro is about 1.6T, Xiaomi 1.02T, and Alibaba 397B. K3 is roughly 75 percent larger than the runner-up. Two in-house architectural moves stand out: Kimi Delta Attention (a hybrid linear attention mechanism) and Attention Residuals (a drop-in replacement for residual connections), both previously published on GitHub. Weights ship in MXFP4 with MXFP8 activations, and quantization-aware training was baked in.

Benchmarks Against Fable 5

Moonshot's official benchmarks call K3 "competitive" with Anthropic's Fable 5, and "substantially outperforming" Opus 4.8, GPT-5.6 Sol, and GPT-5.5. Independent trackers are more measured. Artificial Analysis ranks it third overall, behind Fable 5 and GPT-5.6 Sol Max. Arena's front-end coding blind test puts it first, ahead of the top US models. Vals AI ranks it second. The catch is cost: Fable 5-class reasoning at a fraction of the price.

Yang Zhilin's Comeback

Founder Yang Zhilin is 34, a Tsinghua and Carnegie Mellon alum, and a Pink Floyd fan. The company name (月之暗面, "dark side of the moon") comes from the album. For 18 months the company was pressed by DeepSeek and its market position eroded; K3 is the counterpunch. The numbers are stark. Post-K3 daily revenue grew sixfold, June ARR hit $300 million (up from $200 million in April), and the company is seeking a new round at a $50 billion valuation, with a Hong Kong IPO possibly this year. On launch day, competitors' stocks dove: Z.ai down 30 percent, MiniMax down 16 percent, Alibaba down 4 percent.

The Math Behind Open-Sourcing

Why not stay closed and charge? Yang's logic is to trade openness for a user base and developer community, and vie to be the gravity center of global open-source AI. The geopolitical edge is sharper. Reuters noted that open-sourcing lets Chinese firms both flex muscle and expand influence, neatly countering US chip restrictions. The White House is not idle. Office of Science and Technology Policy director Michael Kratsios publicly accused Moonshot last week of training K3 on banned Nvidia chips and running large-scale distillation against US models including Fable; Moonshot has not responded. Researcher Nathan Lambert's read: the open-versus-closed, China-versus-US gap has compressed from six to nine months down to three to five.

Stacking parameters to 2.8 trillion is engineering capability. Daring to dump all the weights into the open at the peak is a strategic choice. The download page is live; whether to run it and how is up to your compute.


References

This article is AI-assisted and human-edited. Last updated: 2026-07-27

FAQ

What is Kimi K3? When was it open-sourced?
Moonshot AI's flagship model released July 16, 2026-a 2.8-trillion-parameter MoE with a 1-million-token context. On the evening of July 27, the weights were opened for free download, commercially usable and self-hostable, currently the world's largest open-source model.
What does 2.8 trillion parameters mean? How do regular users access it?
Parameters measure an LLM's scale; 2.8 trillion makes K3 the largest open-source model, about 75% bigger than DeepSeek V4 Pro. With weights open, developers can download, modify, and self-host-needs ample compute. Regular users just use the Kimi app or API, no need to touch the weights.
Why is Kimi K3 said to benchmark against Anthropic's Fable 5?
Moonshot's official benchmarks call K3 "competitive" with Fable 5, substantially outperforming Opus 4.8 and GPT-5.6 Sol; independent tracker Artificial Analysis ranks it third (behind Fable 5 and GPT-5.6 Sol Max), and Arena's front-end coding blind test ranks it first. But its cost is far below Fable 5.

Related

Frontline Hotspot

Alibaba's Qwen3.8-Max: 2.4T-param MoE flagship that programs autonomously for days

On 2026-08-03 Alibaba Tongyi released Qwen3.8-Max: a 2.4T-param MoE flagship with 1M context (991K input / 131K output), native vision across plan-execute-verify, positioned to "autonomously program for over ten days delivering complete projects." Pricing: ¥12/M input, ¥36/M output, explicit cache hit ¥1 (1/12 of uncached). Three entry points: blog / Qianwen platform / Qwen Studio.

Aug 3, 20264 min read
Frontline Hotspot

AI Weekly 004: Seven Releases in Seven Days, but the Real Signals Are Agents, Compliance, and Cost

This week (Jul 27-Aug 2) the AI world shipped seven releases, but three signals matter more: DeepSeek-V4-Flash's post-training pushed DeepSWE from 7.3 to 54.4 (hands-on 30/30, cost under 5 fen) and Kimi K3 topped coding leaderboards; the EU AI Act August 2 deadline landed (fines up to 7% of global turnover, extraterritorial); prefix cache hits at 0.02 yuan vs 1 yuan misses make cost engineering a new skill.

Aug 2, 20265 min read
Frontline Hotspot

Kimi K2.6 Enters DoorDash: Why US Congress Is Scrutinizing a Delivery Company for Using a Chinese Model

Two US House committee chairs - Rep. John Moolenaar (Select Committee on the CCP) and Rep. Andrew Garbarino (Homeland Security Committee) - sent a joint letter to DoorDash CEO Tony Xu investigating the company's use of Chinese AI model Kimi K2.6. DoorDash's internal AI lab had adopted the model after finding it outperformed US counterparts on certain tasks, but the decision collided with national security and data security concerns. The case exposes the fundamental tension between performance-driven engineering selection and regulatory risk logic - when the best model comes from a strategic competitor, these two frameworks cannot auto-align.

Aug 1, 20265 min read