Claude Platform Docs
Models & pricingModels

Claude Opus 5Latest

For complex agentic coding and enterprise work

Try in playground
Context window
1Mtokens
Max output
128Ktokens
Input pricing
$5/ MTok
Output pricing
$25/ MTok

Overview

Claude Opus 5 is a step-change improvement over Claude Opus 4.8, with the largest gains in deep reasoning, agentic and long-horizon tasks, and test-time compute scaling. This page summarizes everything new in Claude Opus 5, including mid-conversation tool changes and two breaking changes for code running on Claude Opus 4.8: thinking is on by default, and thinking can be disabled only at effort high or below.

What's new in Claude Opus 5

How it compares

ModelContextMax outputPrice / MTokLatencyThinkingDefault effortKnowledge cutoff
Claude Fable 5.11M128K$10 / $50SlowerAdaptive (always on)highJun 2026
Claude Opus 5This model1M128K$5 / $25ModerateAdaptivehighMay 2026
Claude Sonnet 51M128K$2 / $10FastAdaptivehighJan 2026
Claude Haiku 4.5200K64K$1 / $5FastestExtendedFeb 2025

Specifications

Model IDs

Claude API
Amazon Bedrock
Google Cloud
Microsoft Foundry

Pricing

Input
$5 / MTok
Output
$25 / MTok
5m cache write
$6.25 / MTok
1h cache write
$10 / MTok
Cache read
$0.50 / MTok
Batch API
50% discount on input and output
Full price list
Pricing

Capabilities

Max output
128K tokens
Thinking
Adaptive
Comparative latency
Moderate
Input → output
Text and images → text
Reliable knowledge cutoff
May 2026
Training data cutoff
May 2026

Availability

Status
Active (latest)
Released
July 24, 2026
Retirement
Not sooner than July 24, 2027

Good to know

  • On the Message Batches API, Claude Opus 5 supports up to 300k output tokens with the output-300k-2026-03-24 beta header.
  • The minimum cacheable prompt length is 512 tokens. See Prompt caching.
  • Query limits and capabilities programmatically with the Models API.

Resources

Model-specific prompting guidance.

Effort defaults to high on Claude Opus 5 and matters more than on earlier models. Choose a level per workload.

On by default. Disabling thinking requires effort high or below.

Lower-latency Claude Opus 5 on the Claude API (research preview), priced separately.

Reference

The system prompt Claude Opus 5 uses on claude.ai and the Claude apps.

Safety evaluations and deployment decisions for Claude Opus 5.

Full price list, including batch discounts and prompt caching rates.

How model IDs, aliases, and pinned snapshots work.

Lifecycle status and retirement commitments for every Claude model.

Was this page helpful?