Users Claim Opus 5 Regresses Versus Fable Model
AI practitioners report Opus 5 underperforms in practice despite benchmark wins over Fable.
TLDR
Peter Yang shared Kun Chenguid's critique noting Anthropic's Claude Opus 5 leads public benchmarks yet lags in actual tasks. Jeremy Howard switched back to Fable after testing, calling the newer model a regression. Yam Peleg amplified the report. The discussion centers on the gap between benchmark scores and real-world performance for Opus 5, with experienced users preferring the prior option based on direct experience.
Combined views
185.3K
3 Sources, first seen 65d ago