Claude Opus 5 Achieves Leading Scores Across AI Benchmarks
Anthropic's latest model outperforms prior versions on agentic coding and knowledge tasks.


Combined views
19.3M
45 Sources, first seen ago
Sources
JM Jack Morris@jxmnop
Awesome new addition to my workflow. for complex reasoning, ill reach for Opus 5 medium, thats assuming Sonnet-5 or Gpt-5.6 Luna can't do the job. if it's deep knowledge work, still don't see a better option than Fable low. for complex systems engineering, i dont trust anything…
- likes: 300
- replies: 21
- bookmarks: 153
- reposts: 11
T( Teortaxes▶️ (DeepSeek 推特🐋铁粉 2023 – ∞)@teortaxesTex
OpenAI is still just much more efficient in hardcore reasoning https://twitter.com/ArtificialAnlys/status/2080734447717298483
- likes: 209
- replies: 8
- bookmarks: 27
- reposts: 6
AA Artificial Analysis@ArtificialAnlys
Claude Opus 5 is the new leader on our agentic knowledge work benchmark, AA-Briefcase, outperforming Claude Fable 5 by nearly 150 Elo while reducing Cost per Task by 20% @AnthropicAI has released Claude Opus 5, the new leader on the Artificial Analysis Intelligence Index, and…
- likes: 628
- replies: 27
- bookmarks: 78
- reposts: 62
MD Michael Dempsey@mhdempsey
Which company is consistently putting out the most applicable for real world usage benchmarks without any leakage? Or even just strange benchmarks that still feel unsaturated? Feels pretty clear this needs to be a cat and mouse industry and that labs are going to go harder at…
- likes: 9
- replies: 4
- bookmarks: 3
- reposts: 0
JS Jarred Sumner@jarredsumner
It’s a good model https://twitter.com/claudeai/status/2080699497064083942
- likes: 294
- replies: 9
- bookmarks: 6
- reposts: 9
MB
Combined views
19.3M
45 Sources, first seen ago