Kimi K3: Fable- or Sol-level quality at Sonnet prices, with open weights in 10 days
Moonshot AI dropped Kimi K3 yesterday and everyone is talking about it: according to the first benchmarks it plays at the level of models like Fable or Sol, but at a price similar to Sonnet's. On LMArena it is already number 1 for frontend code with 1679 ELO points, climbing above Claude Fable 5. A historic moment: it's the first time a model that is going to release its weights wipes the floor with a frontier model.
The move is a curious one: the weights won't be released until July 27. Rafa's theory: that way you try it first on their official platform, paying Moonshot; once it lands on OpenRouter and friends there will be competition, quantization and price drops. It's a high price for what Chinese models have us used to, but for the quality it delivers —Opus-level at the very least— it's already cheap. And the context window jumps from K2.6's 262K to 1 million tokens, with 1 million of output too.
Mind-blowing detail: K3 reasons a lot and burns plenty of tokens mulling things over, and on top of that it shows its reasoning unfiltered, in full caveman mode. Ask it 'what model are you?' and its thought is 'need answer user, ask what model be you'. Caveman reasoning: other vendors hide it behind a summary so you can't distill; here you see it raw.
And the moral of the week: competition is glorious. Anthropic has kept extending Fable —now until July 19— and resetting limits, Opus 5 might get postponed because with Sol and K3 on the table it wouldn't hit as hard, and Google keeps delaying Gemini 3.5 Pro. Just a while ago Fable seemed reserved for the elites, and suddenly all of us are using it. Watch in the video.

All the news we covered