The New Stack says it couldn't tell Claude Fable 5.1 and Fable 5 apart on real work
The New Stack reports that both models scored perfectly on four real jobs, but Fable 5.1 cost more than twice as much as Fable 5 on the hardest task.
The New Stack reports that both models scored perfectly on four real jobs, but Fable 5.1 cost more than twice as much as Fable 5 on the hardest task.
The New Stack reports that Anthropic shipped Claude Fable 5.1 on September 1, claiming doubled performance in agentic research. The publication says it compared Fable 5.1 with Fable 5 on four real jobs and tracked every token. Both scored perfectly, but Fable 5.1's bill was more than double Fable 5's on the hardest task.
2.8K
2 posts, first seen 4d ago
The New Stack reports that both models scored perfectly on four real jobs, but Fable 5.1 cost more than twice as much as Fable 5 on the hardest task.
The New Stack reports that Anthropic shipped Claude Fable 5.1 on September 1, claiming doubled performance in agentic research. The publication says it compared Fable 5.1 with Fable 5 on four real jobs and tracked every token. Both scored perfectly, but Fable 5.1's bill was more than double Fable 5's on the hardest task.