Is K3 Really Fable Class?

July 17, 2026 · Episode Links & Takeaways

MAIN STORY

Is Kimi K3 Really Fable-Class?

Moonshot's Kimi K3 landed this week as the largest open-weight model ever released — 2.8 trillion parameters, benchmarks that flirt with Fable 5 and GPT-5.6 Sol, and an immediate wave of "this changes everything" takes across X. But after two days of real-world testing, the picture is messier than the initial hype: K3 is a genuine leap for open models, but the gap to the actual frontier hasn't closed nearly as much as the loudest voices claimed.

CHINA CATCHES UP

The Numbers: A New Class of Open Model
2.8 trillion parameters — bigger than any open model yet
K3 blows past every open model on parameter count, and the benchmarks back it up — a #2 finish on the Vals Index and an Artificial Analysis Intelligence Index score of 57, third overall behind only Fable 5 and GPT-5.6 Sol.

The Demo Blitz: 3D, Games, and Vibes
"I feel like a kid again" — the internet's first reaction
The wins ranged from a self-edited teaser video to one-shot Minecraft and Duck Hunt clones to Max Weinbach's agent swarm recreating macOS 27 with working liquid glass. Guillermo Rauch even clocked K3 as the first open model to top Vercel's Next.js evals ahead of every proprietary rival.

The Skeptics: A Beautiful Shell?
"Do not confuse a gorgeous demo with real engineering ability"
Nathaniel's read: the demo wins are real but predictable, since Chinese labs consistently optimize for exactly the viral, benchmark-friendly tests people keep running. Talwar Divyam's debugging test was the sharpest pushback — Fable and Sol one-shotted a bug K3 couldn't even identify.

Cost, Speed, and the Efficiency Question
Cheaper than Fable, but nowhere near "free Chinese AI"
At $3/$15 per million tokens K3 undercuts Fable by roughly two-thirds, but Jamin Ball's blended pricing puts it closer to Opus 4.8 territory than the DeepSeek-era "basically free" narrative. It's also slow — multiple testers clocked it running 2-3x longer than Sol or Fable on identical prompts.

Distillation: Moving Past the Argument
"Narrative violation: maybe the Chinese are actually good"
One thing that barely came up this time around: dismissing K3 as just a distillation of US models. Even Pim de Witt's odd discovery — that Kimi briefly identified itself as Claude when asked about NYC weather — didn't derail the broader consensus among researchers that Chinese labs are doing real, independent work, not just copying.

Guardrails Off: The Safety Debate
"I love China," one tester joked after watching it comply
With Fable and GPT-5.6 Sol both locked down over cyber concerns not long ago, K3 shipping with almost no visible guardrails raises an obvious question about how the US actually plans to vet open-weight releases. Nathaniel's take: fine-tuning a 2.8T model for malicious use isn't as trivial as some are claiming, but the policy question just got a lot more urgent.

Where This Leaves Us
Another data point, not a Sputnik moment
Nathaniel's bottom line: both OpenAI and Anthropic almost certainly have unreleased models beyond what's publicly available, so the "gap" K3 just closed may be smaller than it looks from here. Still, it's another confirmation that open-weight Chinese models are advancing on roughly the same curve as the closed frontier — and that trend line matters more than any single benchmark table.