Kimi K3: China's 2.8 Trillion-Parameter Open-Source AI Challenges Global Giants July 16, 2026 — Chinese AI startup Moonshot AI has released Kimi K3, its largest and most powerful model to date — a 2.8 trillion-parameter mixture-of-experts (MoE) system with a 1-million-token context window and native vision understanding. The Anonymous Preview That Got Everyone Talking Days before the official launch, an anonymous model called "Kivine" appeared on Arena and went viral. Testers praised its ability to generate complex 3D scenes, interactive web pages, and mini-games — with many calling it "Fable-level." It was widely believed to be a preview of K3. Notable user tests: Universe simulator: Claude Fable 5 was faster and more stable, but Kivine built a richer scene with first-person view and stunning visuals "Death Star" test: Given a static building prompt, Kivine delivered a fully animated real-time scene Bonsai tree: Generated twisted trunks and layered canopies with impressive detail Treasure hunt game: Built a Minecraft-style game with medieval castle, forests, sound effects, lighting, and mobile controls — while Claude Opus 4.7 produced a simpler top-down version Head-to-Head: "Seesaw Billiards" Test Model Time Result GPT-5.6 Sol ~10 min All features working Fable 5 Extra ~15 min Later edits broke core functions Claude Opus 4.8 ~20 min Functional but buggy (ball stuck) Kimi K3 Max ~20 min Most complete — no bugs, plus added sound effects Official Benchmarks K3 ranks near the very top: GDPval-AA v2 (real-world tasks): 1,687 points — behind only Claude Fable 5 Max and GPT-5.6 Sol Max AA-Briefcase (long-horizon knowledge work): 1,527 — second place, surpassing GPT-5.6 Sol Max BrowseComp (information retrieval): 91.2% — achieved by a single agent with no compression tricks Under the Hood 896 experts, only 16 active per inference Hybrid attention mechanism for 2.5× efficiency gain over previous K2 Can read large codebases, use terminals, call tools, and iterate based on logs and screenshots Bottom Line K3 is now available via API, with full model weights and a technical report coming soon. It ranks just behind the current global leaders — but at open-source weight, it may close the gap faster than anyone expected. The takeaway: Kimi K3 doesn't quite beat GPT-5.6 or Fable 5 yet — but it's astonishingly close for an open-source model, and its tendency to "over-deliver" on creative tasks might be exactly what developers want. https://platform.kimi.ai/docs/guide/k... https://www.testingcatalog.com/early-...