Although Claude Sonnet 5 is still available, you should consider migrating to Claude Sonnet 5.5 for improved performance.
See Claude Sonnet 5.5Models & pricingLegacy models
Claude Sonnet 5Legacy
- Context window
- 1Mtokens
- Max output
- 128Ktokens
- Input pricing
- $2/ MTok
- Output pricing
- $10/ MTok
How it compares to the current lineup
| Model | Context | Max output | Price / MTok | Thinking | Default effort | Knowledge cutoff |
|---|---|---|---|---|---|---|
| Claude Fable 5.1 | 1M | 128K | $10 / $50 | Adaptive (always on) | high | Jun 2026 |
| Claude Opus 5.5 | 1M | 128K | $4 / $20 | Adaptive (always on) | medium | Jun 2026 |
| Claude Sonnet 5.5 | 1M | 128K | $2 / $10 | Adaptive | high | Jun 2026 |
| Claude Sonnet 5This modelLegacy | 1M | 128K | $2 / $10 | Adaptive | high | Jan 2026 |
| Claude Haiku 4.5 | 200K | 64K | $1 / $5 | Extended | — | Feb 2025 |
Specifications
Model IDs
Pricing
- Input
- $2 / MTok
- Output
- $10 / MTok
- 5m cache write
- $2.50 / MTok
- 1h cache write
- $4 / MTok
- Cache read
- $0.20 / MTok
- Batch API
- 50% discount on input and output
Capabilities
- Context window
- 1M tokens
- Max output
- 128K tokens
- Max output (Batch API, beta)
- 300K tokens
- Thinking
- Adaptive
- Default effort
high- Input → output
- Text and images → text
- Reliable knowledge cutoff
- Jan 2026
- Training data cutoff
- Jan 2026
Availability
- Status
- Active (legacy)
- Released
- June 30, 2026
- Retirement
- Not sooner than June 30, 2027
- Platforms
- Claude APIAmazon BedrockGoogle CloudMicrosoft FoundryClaude Platform on AWS
Good to know
- On the Message Batches API, Claude Sonnet 5 supports up to 300k output tokens with the
output-300k-2026-03-24beta header. - Setting
temperature,top_p, ortop_kto non-default values returns a 400 error. - Query limits and capabilities programmatically with the Models API.
Resources
What changes when moving from Claude Sonnet 5 to Claude Sonnet 5.5.
The current Sonnet model: overview, specs, and resources.
Model-specific prompting guidance.
On by default on Claude Sonnet 5. Steer depth with effort.
Effort defaults to high on the Claude API and Claude Code. Choose a level per workload.
1M tokens by default. How the window is counted and managed.
Reference
The system prompt Claude Sonnet 5 uses on claude.ai and the Claude apps.
Safety evaluations and deployment decisions for Claude Sonnet 5.
Full price list, including batch discounts and prompt caching rates.
How model IDs, aliases, and pinned snapshots work.
Lifecycle status and retirement commitments for every Claude model.
Was this page helpful?