• Home
  • Technology
  • Gaming
  • Entertainment
  • World & Business
  • Science
  • Sports
  • AI
HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
  • HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
    • Home
    • Technology
    • Gaming
    • Entertainment
    • World & Business
    • Science
    • Sports
    • AI
    AI
    Report

    Claude Opus 5.5 reportedly scores 94.2% on NerfBench, within normal variance

    BridgeMind AI says Opus 5.5 was scoring above its launch level a day earlier but cautions that the dip is not yet evidence of a nerf.

    RS
    BR
    2 Sources, ,

    TLDR

    On October 2, BridgeMind AI reported that Claude Opus 5.5 had dropped to 94.2% on NerfBench after scoring above its launch level the day before. It listed GPT 6 Astra at 98.0%, Sonnet 5.5 at 100.9% and GPT 6.1 Sol at 106.7%. BridgeMind AI said Opus 5.5’s score remained within normal variance, so it could not yet call the drop a nerf.

    Combined views

    454.2K

    2 Sources, first seen 11h ago

    Combined views

    454.2K

    2 Sources, first seen 11h ago

    5.5K likes
    11h ago
    first seen 11h ago
    5.5K likes
    316 comments
    705 saves
    408 reposts
    Featured Source
    316 comments
    705 saves
    408 reposts
    AnthropicClaude Opus 5.5

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    2 Sources

    @bridgemindaiClaude Opus 5.5 just took a big drop on NerfBench. Yesterday it was scoring above launch. Today it's at 94.2%. GPT 6 Astra: 98.0% Sonnet 5.5: 100.9% GPT 6.1 Sol: 106.7% 94.2% is still inside normal variance, so we can't call it a nerf yet. But we're watching Opus 5.5 very closely.11h
    @ziv_ravidRT @bridgemindai: Claude Opus 5.5 just took a big drop on NerfBench. Yesterday it was scoring above launch. Today it's at 94.2%. GPT 6 As…9h

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    Today's Rank

    —

    Not ranked yet

    Today's Rank

    —

    Not ranked yet

    2 Sources

    @bridgemindaiClaude Opus 5.5 just took a big drop on NerfBench. Yesterday it was scoring above launch. Today it's at 94.2%. GPT 6 Astra: 98.0% Sonnet 5.5: 100.9% GPT 6.1 Sol: 106.7% 94.2% is still inside normal variance, so we can't call it a nerf yet. But we're watching Opus 5.5 very closely.11h
    @ziv_ravidRT @bridgemindai: Claude Opus 5.5 just took a big drop on NerfBench. Yesterday it was scoring above launch. Today it's at 94.2%. GPT 6 As…9h

    Related

    A claimed ethical contradiction in Olah’s logic

    A poster says a visiting scholar argued Olah’s reasoning would condemn his own project, and claims Anthropic people ignored the point.

    FTC reportedly probes Anthropic, OpenAI and other AI labs over consumer risks

    Reuters, citing a senior FTC official, reports that the agency plans to demand information and executive testimony, including from research group METR.

    Anthropic's IPO prospectus reportedly warns of 'existential risks to humanity'

    A user pairs that claimed warning with a claim that OpenAI scrapped a new model rollout over safety concerns, pushing back on criticism of the EU AI Act.