• Home
  • Technology
  • Gaming
  • Entertainment
  • World & Business
  • Science
  • Sports
  • AI
HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
AI
Report

Claude Opus 5.5 reportedly scores 94.2% on NerfBench, within normal variance

BridgeMind AI says Opus 5.5 was scoring above its launch level a day earlier but cautions that the dip is not yet evidence of a nerf.

Ravid Shwartz ZivRS
BridgeMindBR
5 Sources, 8d ago, first seen 8d ago

TLDR

On October 2, BridgeMind AI reported that Claude Opus 5.5 had dropped to 94.2% on NerfBench after scoring above its launch level the day before. It listed GPT 6 Astra at 98.0%, Sonnet 5.5 at 100.9% and GPT 6.1 Sol at 106.7%. BridgeMind AI said Opus 5.5’s score remained within normal variance, so it could not yet call the drop a nerf.

Combined views

698.9K

5 Sources, first seen 8d ago

6.5K likes369 comments901 saves464 reposts

Combined views

698.9K

5 Sources, first seen 8d ago

6.5K likes369 comments901 saves464 reposts

Sentiment

Positive7.2%92.8%Negative

Summary

Replies dismissed BridgeMind's nerfbench claims about Claude and Opus model changes as fraudulent slop and outright lies, while a few accounts noted unrelated testing experiments.

Based on 23 sentiment-bearing replies from 21 accounts across 3 conversations.

Featured Source

Sentiment

Positive7.2%92.8%Negative

Summary

Replies dismissed BridgeMind's nerfbench claims about Claude and Opus model changes as fraudulent slop and outright lies, while a few accounts noted unrelated testing experiments.

Based on 23 sentiment-bearing replies from 21 accounts across 3 conversations.

Related

Anthropic’s will ban abusive behavior toward Claude starting November 12, 2026

Usage policy update adds new rules and highlights the potential for misuse.

Anthropic’s will ban abusive behavior toward Claude starting November 12, 2026
State of AI report puts AI-assisted AI research in focus

Author Nathan Benaich cites an Anthropic index and argues robotics is nearing a “GPT-2 moment.”

River API is pitched as a cheap route to task-specific open models

Igor Babuschkin claims fine-tuned open-weight models can beat GPT-6 Astra and Claude Opus 5.5 for under 1% of the cost.

5 Sources

BridgeMind@bridgemindaiClaude Opus 5.5 just took a big drop on NerfBench. Yesterday it was scoring above launch. Today it's at 94.2%. GPT 6 Astra: 98.0% Sonnet 5.5: 100.9% GPT 6.1 Sol: 106.7% 94.2% is still inside normal variance, so we can't call it a nerf yet. But we're watching Opus 5.5 very closely.8d
Ravid Shwartz Ziv@ziv_ravidRT @bridgemindai: Claude Opus 5.5 just took a big drop on NerfBench. Yesterday it was scoring above launch. Today it's at 94.2%. GPT 6 As…7d
Theo - t3.gg@theoI read the "how nerfbench works" post he keeps linking. I don't think anyone else has read it. How is anyone taking this shit seriously???7d
    • Home
    • Technology
    • Gaming
    • Entertainment
    • World & Business
    • Science
    • Sports
    • AI
    AnthropicClaude Opus 5.5

    5 Sources

    BridgeMind@bridgemindaiClaude Opus 5.5 just took a big drop on NerfBench. Yesterday it was scoring above launch. Today it's at 94.2%. GPT 6 Astra: 98.0% Sonnet 5.5: 100.9% GPT 6.1 Sol: 106.7% 94.2% is still inside normal variance, so we can't call it a nerf yet. But we're watching Opus 5.5 very closely.8d
    Ravid Shwartz Ziv@ziv_ravidRT @bridgemindai: Claude Opus 5.5 just took a big drop on NerfBench. Yesterday it was scoring above launch. Today it's at 94.2%. GPT 6 As…7d
    Theo - t3.gg@theoI read the "how nerfbench works" post he keeps linking. I don't think anyone else has read it. How is anyone taking this shit seriously???7d
    Today's Rank

    —

    Not ranked yet

    Today's Rank

    —

    Not ranked yet