chiprook
← AI
AISeptember 30, 2026, 09:28

LiveNerf publicly tracks Claude Opus 5.5 degradation

Since September 24, 2026, the LiveNerf project has been measuring Claude Opus 5.5 daily on a fixed panel of 78 questions, comparing results against the model's launch baseline. The method pins prompts, CLI version and harness hash, and tracks token usage as an early signal of reduced reasoning effort.

LiveNerf publicly tracks Claude Opus 5.5 degradation
#Anthropic#Claude
Read next
AI

Anthropic launches Claude Opus 5.5: 40% cheaper than Opus 5

AI

Test: GPT-6 Sol is cheaper and faster, but Opus 5.5 is more consistent

AI

Claude Opus 5.5 becomes default Opus model in Claude Code at $4/$20 per MTok

AI

Claude Opus 5.5 uses 95% fewer em dashes, but its answers are getting longer