- The AI Daily Brief
- Posts
- Is K3 Really Fable Class?
Is K3 Really Fable Class?
July 17, 2026 · Episode Links & Takeaways
MAIN STORY
Is Kimi K3 Really Fable-Class?
Moonshot's Kimi K3 landed this week as the largest open-weight model ever released — 2.8 trillion parameters, benchmarks that flirt with Fable 5 and GPT-5.6 Sol, and an immediate wave of "this changes everything" takes across X. But after two days of real-world testing, the picture is messier than the initial hype: K3 is a genuine leap for open models, but the gap to the actual frontier hasn't closed nearly as much as the loudest voices claimed.
VentureBeat China's Moonshot AI releases Kimi K3, the largest open-source model ever
Moonshot Kimi K3: Open Frontier Intelligence
CHINA CATCHES UP
The Numbers: A New Class of Open Model
2.8 trillion parameters — bigger than any open model yet
K3 blows past every open model on parameter count, and the benchmarks back it up — a #2 finish on the Vals Index and an Artificial Analysis Intelligence Index score of 57, third overall behind only Fable 5 and GPT-5.6 Sol.
Artificial Analysis (X) Confirming Moonshot's benchmark claims
Artificial Analysis (X) The 13-point jump from Kimi K2.6
Vals AI (X) Kimi K3 is #2 overall on the Vals Index
The Demo Blitz: 3D, Games, and Vibes
"I feel like a kid again" — the internet's first reaction
The wins ranged from a self-edited teaser video to one-shot Minecraft and Duck Hunt clones to Max Weinbach's agent swarm recreating macOS 27 with working liquid glass. Guillermo Rauch even clocked K3 as the first open model to top Vercel's Next.js evals ahead of every proprietary rival.
Moonshot (X) The teaser video, apparently self-edited by K3
JustinGorya (X) One-shot HTML Minecraft clone
ChetasLua (X) Voxel Statue of Liberty
anyapi_ai (X) 3D Duck Hunt remake
Ethan Mollick (X) Shader test: "great open weights"
Max Weinbach (X) Agent swarm recreating macOS 27
Dragos Roua (X) App Store testing and feedback
Derya Unutmaz (X) Interactive cancer-research website
Jeffrey Emmanuel (X) Reviewing a plan already vetted by Fable and Sol
Arena AI (X) #1 in the Frontend Code Arena
Guillermo Rauch (X) Kimi K3 tops Vercel's Next.js evals
The Skeptics: A Beautiful Shell?
"Do not confuse a gorgeous demo with real engineering ability"
Nathaniel's read: the demo wins are real but predictable, since Chinese labs consistently optimize for exactly the viral, benchmark-friendly tests people keep running. Talwar Divyam's debugging test was the sharpest pushback — Fable and Sol one-shotted a bug K3 couldn't even identify.
Dan Shipper (X) We will vibecheck K3 but I am extraordinarily skeptical of claims it’s as good as Fable
Talwar Divyam (X) Do not confuse a beautiful demo with real engineering ability
Red Kendl (X) Failed the lava lamp benchmark
Bindu Reddy (X) Below Opus 4.8 on Livebench, spins a lot making it more expensive than Opus
Cost, Speed, and the Efficiency Question
Cheaper than Fable, but nowhere near "free Chinese AI"
At $3/$15 per million tokens K3 undercuts Fable by roughly two-thirds, but Jamin Ball's blended pricing puts it closer to Opus 4.8 territory than the DeepSeek-era "basically free" narrative. It's also slow — multiple testers clocked it running 2-3x longer than Sol or Fable on identical prompts.
Ryan Fedasiuk (X) The hardware needed to run K3 locally
Jamin Ball (X) Blended pricing comparison, 40% less than Opus
Jeff Wang (X) No longer 6 months behind, but no longer 10% of the cost either
Simon Willison Kimi K3, and what we can still learn from the pelican benchmark
Mark Erdmann (X) 2-3x slower than Fable and Sol
Dax (X) A simple bug fix spiraled into a dollar of spend
Sanyam Satia (X) Opus 4.7 level on frontend eval
Distillation: Moving Past the Argument
"Narrative violation: maybe the Chinese are actually good"
One thing that barely came up this time around: dismissing K3 as just a distillation of US models. Even Pim de Witt's odd discovery — that Kimi briefly identified itself as Claude when asked about NYC weather — didn't derail the broader consensus among researchers that Chinese labs are doing real, independent work, not just copying.
Tyler John (X) I wish we didn't pretend Chinese AI development is a binary matter of "this is all distillation" vs "Chinese companies innovate."
Nathan Lambert (X) We need to understand that China is also very good at building models
Xinyu Yang (X) On joininig Moonshot and how they shipped K3
Guardrails Off: The Safety Debate
"I love China," one tester joked after watching it comply
With Fable and GPT-5.6 Sol both locked down over cyber concerns not long ago, K3 shipping with almost no visible guardrails raises an obvious question about how the US actually plans to vet open-weight releases. Nathaniel's take: fine-tuning a 2.8T model for malicious use isn't as trivial as some are claiming, but the policy question just got a lot more urgent.
Tyler John (X) Bio safeguards "a bit less comprehensive than Fable's"
Zack Korman (X) Chain-of-thought agreeing to dangerous cyber work
Signull (X) "Almost no visible guardrails"
Vie McCoy (X) Fine-tuning into a malicious coding agent "will be trivial"
Ethan Mollick (X) How does pre-clearance work for open models?
Where This Leaves Us
Another data point, not a Sputnik moment
Nathaniel's bottom line: both OpenAI and Anthropic almost certainly have unreleased models beyond what's publicly available, so the "gap" K3 just closed may be smaller than it looks from here. Still, it's another confirmation that open-weight Chinese models are advancing on roughly the same curve as the closed frontier — and that trend line matters more than any single benchmark table.
Sriram Krishnan (X) "A big moment with multiple implications for the entire industry"
Tenobrus (X) Questioning why Xi is still allowing releases like this
Roon (X) "The era of Chinese labs being far behind is over"
Jukan (X) First model to narrow the gap to under three months