Claude Sonnet 5 is Anthropic's middle tier — and, for a lot of teams, the one that matters most. Anthropic bills it as "the best combination of speed and intelligence": the model that gets close to frontier quality on coding and agentic work without Opus's price tag or latency. It sits below Opus 4.8 (the capability ceiling of the general lineup) and above Haiku 4.5 (the fastest, cheapest tier).
The headline spec sheet matches the current generation: a 1-million-token context window, up to 128K output tokens, a January 2026 knowledge cutoff, and adaptive thinking that's on by default — the model decides how hard to think, tuned by an effort dial that defaults to high. One migration wrinkle worth flagging: Sonnet 5 uses the newer tokenizer, which counts roughly 30% more tokens than Sonnet 4.6 for the same text, so any budgets carried over from an older Sonnet need re-baselining.
Who it's for
Most production workloads. If you're shipping a coding assistant, an agent, or a high-volume pipeline, Sonnet 5 is the default: near-frontier capability at a fraction of Opus's cost. Introductory pricing of $2 / $10 per million tokens (through August 31, 2026) makes it roughly a third the price of Opus 4.8, and it's faster by default, which shows up directly in interactive latency.
Who should skip it
Two cases. First, the hardest long-horizon agentic and enterprise tasks, where Opus 4.8's extra capability is worth the premium — reach up when the task genuinely needs it. Second, anyone who needs to own the model: Sonnet 5 is closed and hosted-only, so if on-prem, offline, or auditable weights are a hard requirement, an open-weight model is the only path.
This is a first look assembled from Anthropic's official documentation, not our own evaluation, so we're not attaching a score. On paper, Sonnet 5 is the value sweet spot of the Claude lineup — the model to start with, and to keep unless a task specifically demands Opus.