How-ToDevelopersSeptember 26, 2026

Claude Code effort levels explained via Terminal-Bench 3.0 tests

Read original source →claude.dev

Anthropic's Thariq Shihipar tested three builds on Terminal-Bench 3.0 with Opus 5.5 and Fable 5.1, finding each effort level raises both benchmark scores and tokens consumed. Higher effort suits verification-heavy work like hardware, code review, and security; he runs low effort for implementation and high effort for verification.

1 source

More stories today

Open the live feed