Effort levels recalibrated in Opus 4.8

SWE-Bench Pro system card data shows a major shift in token consumption. The low-effort setting in Opus 4.8 now uses as many output tokens as the medium or high settings in 4.7 and 4.6.
The efficiency of Opus 4.8 on low effort matches the high level of 4.7. The medium setting in 4.8 consumes more tokens than the high setting in 4.7, approaching the absolute maximum of 4.6.
Developers should reassess generation settings. Increased token consumption directly impacts API speed and cost.
Related stories
- Max effort adds nothing to Fable 5.1's ARC-AGI-2 score
- Opus 5 tops Fable 5 on OSWorld 2.0 at a third of the price
- Claude Fable 5: access and capabilities
- Opus 5.5 aced a test suite at 3.4x GPT-6 Sol's cost
- Opus 5.5 leads the index and burns 260M tokens doing it
- Opus 5.5 finds new bugs for CodeRabbit and misses old ones
Comments
No comments yet. Be the first.
Join the conversation
Sign in with Google to leave a comment. Your name and avatar come from your Google profile, and the comment appears after moderation.
