head-to-head

MetricClaude Sonnet 5GPT-5.5
SWE-bench Verified85.2%82.6%
SWE-bench Pro63.2%58.6%
Terminal-Bench80.4% (TB2.1)82.7% (TB2.0)
Input $ / 1M$2
Context1M
Open weightsNoNo
MakerAnthropicOpenAI

when to pick each

Pick Claude Sonnet 5 if

The best closed-model value — near-Opus scores at ~2.5× less, and the default daily driver for most developers.

Pick GPT-5.5 if

OpenAI's strongest agentic coder, with the deepest tooling and ecosystem breadth of the closed labs.

Full reviewsClaude Sonnet 5, decoded

Ranked on our AI Coding Leaderboard, updated 2026-07-02. Scores are confirmed against primary sources; prices are per 1M input tokens and can change.

Primary sources