• Home
  • Technology
  • Gaming
  • Entertainment
  • World & Business
  • Science
  • Sports
  • AI
HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
    • Home
    • Technology
    • Gaming
    • Entertainment
    • World & Business
    • Science
    • Sports
    • AI
    AI
    Announcement

    Claude models top Agent Arena’s Code, Work and Chat rankings as of October 5, 2026

    Agent Arena ranks GPT 6 Astra (Max) second in Code and fourth in both Work and Chat.

    Arena.aiAR
    1 Source, 3h ago, first seen 3h ago

    TLDR

    Agent Arena says its October 5, 2026, rankings put Claude Fable 5.1 (Max) first in Code and Work and Claude Sonnet 5.5 (Max) first in Chat. GPT 6 Astra (Max) ranks in the top four across all three categories, while Gemini 4 Argon (High) ranks fifth in Chat, 14th in Code and 12th in Work. Agent Arena says the rankings draw on live traces of agents tackling tasks submitted by people.

    Combined views

    13.6K

    1 Source, first seen 3h ago

    Combined views

    13.6K

    1 Source, first seen 3h ago

    134 likes
    134 likes
    19 comments
    31 saves
    7 reposts
    19 comments
    31 saves
    7 reposts
    Featured Source

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    Today's Rank

    —

    Not ranked yet

    Today's Rank

    —

    Not ranked yet

    1 Source

    Arena.ai@arenaThe Agent Arena ranks agentic ability to solve problems across domains. U.S. labs lead across every domain: - @AnthropicAI models hold the #1 position in Code, Work, and Chat - @OpenAI's GPT 6 Astra (Max) remains in the top four across all three categories: #2 in Code and #4 in both Work and Chat The leading models stay strong across categories, but their order shifts: - Claude Fable 5.1 (Max) leads both Code and Work and ranks #2 in Chat - Claude Sonnet 5.5 (Max) leads Chat and ranks #4 in Code and #5 in Work Code and Work rankings move together more closely than Chat: - Gemini 4 Argon (High), for example, ranks #14 in Code and #12 in Work but #5 in Chat, highlighting a distinct strength in conversational agent tasks Agent Arena ranks models based on overall net improvement against the current frontier, as well as performance across specific categories of agentic tasks. These rankings are grounded in live traces collected from agents attempting to solve real-world tasks submitted by humans around the world. - Code: writing and debugging code, workflow automation, and data analysis - Chat: creative writing, learning, personal questions, everyday research, and media generation - Work: documents, professional research, planning, and professional writing This visual outlines how selected models compare across Agent Arena categories as of October 5, 2026.3h