Esc
EthicsCase Closed

Anthropic Users Claim Opus 4.6 Performance Degradation

Is this a scandal?

No longer — the story has resolved. Noise 1/100, cooling down, across 1 source.

SCAND-64929as of Methodology
Cite this incident"Anthropic Users Claim Opus 4.6 Performance Degradation." SCAND.Ai incident SCAND-64929, noise 1/100 as of July 31, 2026. https://scand.ai/scandal/anthropic-opus-4-6-nerfed-controversy
FORECASTForecast, not fact

Anthropic will likely release a statement attribute the changes to system prompt optimizations or caching mechanisms. If user backlash continues, they may roll back specific inference parameters or release a 'pro' version that guarantees higher compute allocation for complex tasks.

1

Noise 1/100 — louder than 88% of tracked AI controversies.

AI-assisted analysis · How we work

Why it matters

Trust in AI model stability is critical for enterprise adoption, and unexplained performance shifts undermine confidence in API reliability and vendor transparency.

Key points

  1. Anthropic representative Cherny explicitly denied claims that Claude Code or Opus models were secretly nerfed.
  2. Notion confirmed that Claude Opus 4.7 and 4.8 models exhibit degraded performance and higher failure rates.
  3. Technical analysis attributes perceived Opus 4.6 degradation to lower thinking budgets in Adaptive Thinking features.
  4. Users allege flagship models became slower, costlier, and less capable in real-world workflows since February 2026.
  5. The dispute centers on whether efficiency optimizations constitute undisclosed service degradation for enterprise clients.

The story

Anthropic has denied allegations that it secretly degraded its Claude Opus model family after users reported declining performance across versions 4.6 through 4.8. Anthropic representative Cherny stated on X that claims of intentional nerfing are false, while technical explanations attribute perceived degradation to Adaptive Thinking budget adjustments rather than capability removal. Notion confirmed that Opus 4.7 and 4.8 models are experiencing higher failure rates, validating some user concerns about reliability despite Anthropic’s denial of deliberate downgrades. The controversy highlights tensions between AI providers optimizing inference costs and enterprise customers expecting consistent model behavior. Developers argue that undocumented changes to reasoning budgets effectively reduce utility without transparent communication. Anthropic maintains that architectural updates aim to improve efficiency, not diminish quality, though the dispute raises broader questions about versioning standards and service-level expectations in the generative AI market.

Who's involved

Critic
Realistic_Stomach848

Claims the model has become lazy and stupid, providing instant replies that ignore the complexity of scientific prompts.

Defender
Anthropic

Maintains the model's integrity while implementing backend updates for efficiency and speed.

Neutral
AI Research Community

Monitoring for evidence of model drift or intentional quantization effects that could explain the change in behavior.

How the conversation shifted

the split has narrowed

Polarity (0–100) from the noise pipeline, sampled over time.

Join the Discussion

Discuss this story

Community comments coming in a future update

Be the first to share your perspective. Subscribe to comment.

Noise Level

Quiet1?Noise Score (0–100): how loud a controversy is. Composite of reach, engagement, star power, cross-platform spread, polarity, duration, and industry impact — with 7-day decay.
Decay: 5%
Reach
0
Engagement
0
Star Power
45
Duration
0
Cross-Platform
0
Polarity
50
Industry Impact
50

The timeline

  1. User reports Opus 4.6 'nerfing' on Reddit

    A user on r/ClaudeAI notes that the model provides instant, low-quality replies to hard scientific prompts.

The full record

Sources & methodology

The records from this story's original coverage were pruned, so items marked located later were found by searching for it afterwards. The summary above has since been rewritten to take them into account — it is not the text first published. How we score →

The forecast

Anthropic will likely release a statement attribute the changes to system prompt optimizations or caching mechanisms. If user backlash continues, they may roll back specific inference parameters or release a 'pro' version that guarantees higher compute allocation for complex tasks.

Forecast, not fact — an editorial estimate we score when this resolves.

You're up to date

That's the complete picture as of — nothing more to know right now. We'll update this page the moment it changes.