Claude Opus 5Latest
For complex agentic coding and enterprise work
- Context window
- 1Mtokens
- Max output
- 128Ktokens
- Input pricing
- $5/ MTok
- Output pricing
- $25/ MTok
Overview
Claude Opus 5 is a step-change improvement over Claude Opus 4.8, with the largest gains in deep reasoning, agentic and long-horizon tasks, and test-time compute scaling. This page summarizes everything new in Claude Opus 5, including thinking on by default, mid-conversation tool changes, and a breaking change to when thinking can be disabled.
How it compares
| Model | Context | Max output | Price / MTok | Latency | Thinking | Default effort | Knowledge cutoff |
|---|---|---|---|---|---|---|---|
| Claude Fable 5 | 1M | 128K | $10 / $50 | Slower | Adaptive (always on) | high | Jan 2026 |
| Claude Opus 5This model | 1M | 128K | $5 / $25 | Moderate | Adaptive | high | May 2026 |
| Claude Sonnet 5 | 1M | 128K | $2 / $10 | Fast | Adaptive | high | Jan 2026 |
| Claude Haiku 4.5 | 200K | 64K | $1 / $5 | Fastest | Extended | — | Feb 2025 |
Specifications
Model IDs
- Claude API
Pricing
- Input
- $5 / MTok
- Output
- $25 / MTok
- 5m cache write
- $6.25 / MTok
- 1h cache write
- $10 / MTok
- Cache read
- $0.50 / MTok
- Batch API
- 50% discount on input and output
- Full price list
- Pricing
Capabilities
- Context window
- 1M tokens
- Max output
- 128K tokens
- Max output (Batch API, beta)
- 300K tokens
- Thinking
- Adaptive
- Default effort
high- Comparative latency
- Moderate
- Input → output
- Text and images → text
- Reliable knowledge cutoff
- May 2026
- Training data cutoff
- May 2026
Availability
- Status
- Active (latest)
- Released
- July 24, 2026
- Retirement
- Not sooner than July 24, 2027
- Platforms
- Claude APIAmazon BedrockGoogle CloudMicrosoft Foundry
Good to know
- On the Message Batches API, Claude Opus 5 supports up to 300k output tokens with the
output-300k-2026-03-24beta header. - The minimum cacheable prompt length is 512 tokens. See Prompt caching.
- Query limits and capabilities programmatically with the Models API.
Resources
Model-specific prompting guidance.
Effort defaults to high on Claude Opus 5 and matters more than on earlier models. Choose a level per workload.
On by default. Disabling thinking requires effort high or below.
Lower-latency Claude Opus 5 on the Claude API (research preview), priced separately.
Reference
The system prompt Claude Opus 5 uses on claude.ai and the Claude apps.
Safety evaluations and deployment decisions for Claude Opus 5.
Full price list, including batch discounts and prompt caching rates.
How model IDs, aliases, and pinned snapshots work.
Lifecycle status and retirement commitments for every Claude model.
Was this page helpful?