close
Claude Platform Docs
Models & pricingModels

Claude Sonnet 5Latest

The best combination of speed and intelligence

Try in playground
Context window
1Mtokens
Max output
128Ktokens
Input pricing
$2/ MTok
Output pricing
$10/ MTok

Overview

Claude Sonnet 5 is the next generation of Anthropic's Sonnet model family. It is a drop-in upgrade for Claude Sonnet 4.6 with three behavior changes: adaptive thinking is on by default, manual extended thinking now returns a 400 error (it was deprecated on Claude Sonnet 4.6), and setting sampling parameters (temperature, top_p, top_k) to non-default values returns a 400 error. This page summarizes everything new at launch, including a new tokenizer.

What's new in Claude Sonnet 5

How it compares

ModelContextMax outputPrice / MTokLatencyThinkingDefault effortKnowledge cutoff
Claude Fable 5.11M128K$10 / $50SlowerAdaptive (always on)highJun 2026
Claude Opus 51M128K$5 / $25ModerateAdaptivehighMay 2026
Claude Sonnet 5This model1M128K$2 / $10FastAdaptivehighJan 2026
Claude Haiku 4.5200K64K$1 / $5FastestExtendedFeb 2025

Specifications

Model IDs

Claude API
Amazon Bedrock
Google Cloud
Microsoft Foundry
Claude Platform on AWS

Pricing

Input
$2 / MTok
Output
$10 / MTok
5m cache write
$2.50 / MTok
Cache read
$0.20 / MTok
Batch API
50% discount on input and output
Full price list
Pricing

Capabilities

Max output
128K tokens
Thinking
Adaptive
Comparative latency
Fast
Input → output
Text and images → text
Reliable knowledge cutoff
Jan 2026
Training data cutoff
Jan 2026

Availability

Status
Active (latest)
Released
June 30, 2026
Retirement
Not sooner than June 30, 2027

Good to know

  • On the Message Batches API, Claude Sonnet 5 supports up to 300k output tokens with the output-300k-2026-03-24 beta header.
  • Setting temperature, top_p, or top_k to non-default values returns a 400 error. See What's new in Claude Sonnet 5.
  • Query limits and capabilities programmatically with the Models API.

Resources

Model-specific prompting guidance.

On by default on Claude Sonnet 5. Steer depth with effort.

Effort defaults to high on the Claude API and Claude Code. Choose a level per workload.

1M tokens by default. How the window is counted and managed.

Reference

Safety evaluations and deployment decisions for Claude Sonnet 5.

Full price list, including batch discounts and prompt caching rates.

How model IDs, aliases, and pinned snapshots work.

Lifecycle status and retirement commitments for every Claude model.

Was this page helpful?