← Back to API models

Claude Sonnet 5.5

by Anthropic · Sonnet flagship · released September 28, 2026 (replaced Sonnet 5)

Near-Opus quality at the same permanent $2/$10 rate Sonnet 5 held. The default Claude for most production work and Claude's free-tier model.

Input $2 / 1M
Output $10 / 1M
Context 1M
Pricing Unchanged from Sonnet 5
Anthropic Console ↗ Updated October 7, 2026
✓ Sonnet 5.5 released — Sep 28, 2026

Claude Sonnet 5.5 replaced Sonnet 5 on September 28, 2026, across the Claude API, claude.ai, AWS, Google Cloud, and Microsoft Foundry. Pricing is unchanged at $2/$10 per 1M tokens (cache read $0.20, cache write $2.50). Sonnet 5 is now a legacy model — still available, with retirement not sooner than September 2027 — while Sonnet 5.5 takes over as the current, recommended Sonnet.

§ API pricing

Per-token rates — unchanged from Sonnet 5.

Input
$2/1M tokens
Prompt tokens
  • Same rate Sonnet 5 held since Aug 10, 2026
  • Vision inputs billed as tokens
  • Cache reads $0.20/1M
Output
$10/1M tokens
Completion tokens
  • No price change on replacement
  • Cheaper than GPT-5.4's $15 output
  • Includes extended thinking tokens
Context
1Mtokens
Window
  • 1M context window
  • 128K max output tokens
  • Largest Claude mid-tier context available
Caching
$2.50cache write / 1M
Prompt caching
  • Cache reads $0.20/1M (10% of input)
  • Big wins for chat with long instructions
  • 5-min and 1-hour TTLs supported

Sonnet 5.5 carries over Sonnet 5's price exactly: $2 / 1M input and $10 / 1M output, with cache reads at $0.20/1M and cache writes at $2.50/1M. Anthropic made this a routine successor — no action needed from existing integrations.

What changed in Sonnet 5.5

Claude Sonnet 5.5, released September 28, 2026, replaces Sonnet 5 (June 30, 2026) at the identical $2 input / $10 output per-1M-token price — the same permanent rate Anthropic locked in for Sonnet 5 on August 10, 2026. Cache pricing also carries over unchanged: $0.20/1M cache reads, $2.50/1M cache writes. It's now available across the Claude API, claude.ai, Claude Code, AWS, Google Cloud, and Microsoft Foundry.

The Sonnet pitch is otherwise unchanged in shape: most of Opus's capability at a meaningful discount, with better latency. Anthropic's own models-overview documentation now lists Sonnet 5.5 — alongside Fable 5.1, Opus 5.5, and Haiku 4.5 — as the current generation, with Sonnet 5 moved to a "legacy models (still available)" list alongside Opus 5 and older Opus releases. Legacy doesn't mean retired: Anthropic commits to keeping Sonnet 5 available no sooner than September 2027.

Capabilities

Sonnet 5.5 is a strong all-rounder: writing, code, vision, and structured extraction. It's the right model for production chat agents and coding assistants because answers stay composed under load, and tool use is reliable across multi-turn loops. As with the Sonnet 5 → 5.5 transition, the model card emphasizes gains in agentic execution and adaptive thinking (the effort parameter defaults to High on the API, Medium in Claude Code and the apps).

Where it still trails the Opus tier (Opus 5 and Opus 5.5): the hardest novel reasoning, agentic coding, and the most demanding long-document synthesis. Where it trails GPT-5.5: outright token throughput.

Typical use cases

  • Production chat assistants and customer support agents
  • Code generation, review, and IDE integrations
  • Coding agents and long-horizon tool-use loops
  • Document Q&A and summarization (1M context for repo-scale work)
  • Vision tasks: charts, screenshots, document OCR
  • Structured extraction and form filling

Sibling and rival comparison

ModelInput / 1MOutput / 1MContext
Claude Sonnet 5.5$2$101M
Claude Opus 5.5$4$201M
Claude Opus 5$5$251M
Claude Haiku 4.5$1$5200K
GPT-5.4$2.50$15272K

Sonnet 5.5's $2/$10 rate makes it a clear mid-tier value pick — cheaper on output than GPT-5.4, with a full 1M context window. Opus 5.5 and Opus 5 cost 2–2.5x more for the highest-stakes work; Haiku 4.5 is half the price for narrower tasks.

← See all Anthropic / Claude plans