GPT 5.6 Beats Fable 5 by on DeepSWE at a cheaper price.
GPT-5.6 Sol reaches roughly 72% to 73% DeepSWE at about $8.4 per task, while Claude/Fable-5 tops out around 70% at a much higher cost, around $13 to $22 per task.
DeepSWE is a coding-agent benchmark, not a normal…
Today’s edition of my newsletter just went out.
🔗 https://www.rohan-paul.com/p/gpt-56-beats-fable-5-by-on-deepswe
🗞️ GPT 5.6 Beats Fable 5 by on DeepSWE at a cheaper price.
🗞️ 1X Debuts Human-Level Robotics Hands - Neo’s Tendon-Driven Hands With 25 Degrees of Freedom
🗞️…
What a week - three frontier models (Grok-4.5, Muse-Spark-1.1, GPT-5.6) launched, and GPT-5.6 reached #1 in Code Arena, matching Claude Fable-5. The competition is heating up! https://twitter.com/arena/status/2075672492312768683