Kimi K3 Is The Best Open-Source Model In The World
- scores right below Opus 4.8
- works very well on first turn benchmark style questions
- beats GLM 5.2
Try it on ChatLLM now
This is the most important plot of all the benchmarks.
Cost per intelligence, on average, is just under the Pareto Frontier set by GPT 5.6 Sol, significantly better than Opus or even Fable.
Hopefully, someone makes it run faster/cheaper + distills it…
KIMI K3 CLOSES THE GAP BUT RANKS BEHIND FRONTIER MODELS
Our benchmark, LiveBench, has a lot of hidden questions and models can't memorize them
K3 is the best good open-source model but is below Opus 4.8, Sol and Fable
Also in practice, Kimi spins a lot and costs as much as…