Users are excited about Kimi K3 ranking as the strongest open model on MathArena because its technical performance sparks surprise and optimism about further gains like better self-verification.
Based on 5 visible X reactions from 3 accounts; directional sample.
Ask a question below.
Published answers will appear here.
K3 spirals for 180K tokens to prove nonsense ("proves" it) Fable gets caught in the loop and loses consciousness
K3 with some more trust in its self-verification would be unstoppable
@teortaxesTex I always believed in King Zuck (and Wang, once ppl started briefing about how much they hated him)
This is absurdly impressive. BrokenArXiv is *hard* for LLMs and OpenAI focused on such problems hard and holds a commanding lead. Kimi is not a frontend slop machine, it's a generalist proto-AGI (though same can be said of Meta, congrats)
K3 trying to prove impossible statements. Who was saying open models don't speak neuralese? This isn't yet Fable-dense, but this is already "grug speech" stage. Distillation on exfiltrated CoTs or convergence? you decide.
Users are excited about Kimi K3 ranking as the strongest open model on MathArena because its technical performance sparks surprise and optimism about further gains like better self-verification.
Based on 5 visible X reactions from 3 accounts; directional sample.
Ask a question below.
Published answers will appear here.