Skip to content
Anthropic

Claude Sonnet 5

Anthropic's mid-tier model, now at its best as a workhorse subagent that a stronger model hands focused jobs to.

Quality

Modality

multimodal

Context

1M tokens

Access

closed

Fabian's Take

FM

"When it launched, Sonnet 5 came closer to Opus than any Sonnet generation before, and it's agentic in a way earlier Sonnets weren't: it strings together more tool calls and sticks with harder problems longer. These days I don't pick it by hand at all. It's still a workhorse, though, when a stronger model hands it a focused job as a subagent."

Claude Sonnet 5 is the latest version of Anthropic’s mid-tier model, and it closes most of the gap that used to separate Sonnet from Opus. It’s not just faster and cheaper anymore — it’s meaningfully better at working through multi-step tasks on its own.

How it compares

At launch I ran it against Opus 4.8, and for coding, drafting, and everyday agentic work the two were close enough that Sonnet 5 was the more sensible default. It’s also a clear step up from Sonnet 4.6 in how it handles tools: it uses more of them, in longer sequences, and holds up over longer stretches of autonomous work without drifting off track.

The Opus tier has moved twice since then. Opus 5.5 is the current one, and it’s so capable and efficient that I now use it for almost everything I do by hand. Sonnet’s place has shifted with it. I don’t select it manually anymore, but it still does a lot of work as a subagent: Opus plans, then hands Sonnet a focused job, like searching a codebase, drafting a section, or running a well-defined change, and Sonnet does it quickly at half the per-token price. If you’re building on the API and paying per token for high-volume work, it’s still a strong, cheap default.

What it costs

$2 per million input tokens and $10 per million output. Anthropic launched it at that price as an introductory rate, with a rise to $3 and $15 planned for September, then made the lower price permanent instead. Sonnet 5.5 is announced for the coming weeks.

The Verdict

Best for: Focused coding, drafting, and research jobs handed to it as a subagent by Opus or Fable, and high-volume API work.

Pros

  • Handles long, tool-heavy agentic sessions without losing the thread
  • Coding and refactoring quality closing in on Opus
  • Still fast and cheap enough for real-time, everyday use

Cons

  • Still a notch behind Opus on the hardest, most open-ended reasoning

Specs

  • Pricing $2/M input, $10/M output
  • Cost Tier moderate
  • ⚡
    Speed Tier fast
  • License Proprietary
Developer Docs