Is Claude Sonnet 5.5 better than Opus 5.5?
On Terminal-Bench 4.0, using the numbers we covered, Sonnet 5.5 scored 70.6% against 66.4% for Opus 5.5 on agent coding tasks. Opus 5.5 still leads on FrontierCode and CursorBench, so the better model depends on the task.