Back to Models

Claude Sonnet 5

anthropicanthropic/claude-sonnet-5

Anthropic's frontier Sonnet-tier model β€” top-tier agentic coding and reasoning, 1M context, 128K output.

Claude Sonnet 5, released June 30 2026, is Anthropic's next Sonnet-tier agent model and the default practical upgrade from Sonnet 4.6. It keeps the Sonnet economic lane while adding a native 1M-token context window, 128k synchronous max output, adaptive thinking by default, and stronger coding, tool-use, computer-use, and knowledge-work behavior that Anthropic positions close to Opus 4.8 at a lower price.

For routing teams, Sonnet 5 is not a blind alias swap. The migration has three production-impacting constraints: manual extended-thinking budgets now return 400, non-default sampling parameters return 400, and the new tokenizer can produce about 1.0-1.35x as many tokens for the same text. TheRouter should therefore treat Sonnet 5 as a high-throughput agent route with explicit token recounting, effort controls, and fallback rules rather than a silent replacement for every Sonnet 4.6 workload.

Best for
  • β€’ Long-context coding agents and brownfield repository work where a 1M-token window and stronger follow-through beat Sonnet 4.6 without paying Opus prices
  • β€’ Browser, terminal, and workflow agents that need planning, tool use, self-checking, and practical cost efficiency at medium or high effort
  • β€’ Large document analysis and legal/research workflows that benefit from 128k output and adaptive thinking rather than manual thinking budgets
  • β€’ High-volume Sonnet fleets during the introductory pricing window, provided the router measures token inflation before declaring cost savings
Reach for something else if
  • β€’ Workloads that require custom temperature/top_p/top_k β€” Sonnet 5 rejects non-default sampling; route to Sonnet 4.6 or another model that still accepts sampling controls
  • β€’ Priority Tier latency guarantees β€” Anthropic explicitly excludes Sonnet 5 from Priority Tier support at launch
  • β€’ Manual extended-thinking integrations that still send budget_tokens β€” migrate to adaptive thinking and effort first or requests will fail with HTTP 400
  • β€’ Maximum frontier capability or cyber work with reduced guardrails β€” Anthropic recommends Opus-class models for higher-end cybersecurity work
Context Length
1M
Max Output
128K
Input Priceper 1M tokens
$2.16/ 1M tokens
Output Priceper 1M tokens
$10.80/ 1M tokens

Modalities

textimage→text

Capabilities

Vision

Pricing Breakdown

TypeRate
Input$2.16 / 1M tokens
Output$10.80 / 1M tokens

Supported Parameters

temperaturemax_tokenstop_ptoolstool_choiceresponse_formatreasoningstop

Specifications

Release date2026-06-30platform.claude.com β†—verified
Claude API model idclaude-sonnet-5platform.claude.com β†—verified
Context window1,000,000 tokensplatform.claude.com β†—verified
Max output tokens128,000 synchronous; up to 300,000 on Message Batches with output-300k betaplatform.claude.com β†—verified
Pricing$2/$10 per MTok through 2026-08-31; $3/$15 per MTok starting 2026-09-01platform.claude.com β†—verified
Prompt caching5m writes $2.50/MTok, 1h writes $4/MTok, cache reads $0.20/MTok during introductory pricing; standard rates rise with $3/$15 pricingplatform.claude.com β†—verified
Training data cutoffJanuary 2026platform.claude.com β†—verified
Reliable knowledge cutoffJanuary 2026platform.claude.com β†—verified
Reasoning controlsAdaptive thinking on by default; steer with effort; manual budget_tokens removedplatform.claude.com β†—verified
Unsupported controlsNon-default temperature, top_p, and top_k return HTTP 400; assistant prefill remains unsupportedplatform.claude.com β†—verified
Priority TierNot available on Claude Sonnet 5 at launchplatform.claude.com β†—verified
Tokenizer migrationNew tokenizer; same text can produce about 1.0-1.35x tokens versus Sonnet 4.6www.anthropic.com β†—verified
DeploymentClaude API, Claude Code, Amazon Bedrock, Claude Platform on AWS, Google Cloud Vertex AI, Microsoft Foundryplatform.claude.com β†—verified
LicenseAnthropic Usage Policy (proprietary, API-only)verified

Benchmarks

BenchmarkDistributionScoreSource
BrowseComp
Anthropic reports Sonnet 5 as a strict improvement over Sonnet 4.6 and says higher-effort Sonnet 5 can match Opus 4.8 on some agentic-search tasks; the chart values are published graphically in the announcement/system card rather than as machine-readable numbers in the extracted source.
Not numerically disclosed in extracted launch textwww.anthropic.com β†—
OSWorld-Verified
Anthropic compares Sonnet 5 with Sonnet 4.6 and Opus 4.8 on OSWorld-Verified and frames Sonnet 5 as a wider cost-performance frontier for computer-use workflows. Use route-local evals before committing automation traffic.
Strictly improved over Sonnet 4.6 in Anthropic launch comparisonwww.anthropic.com β†—
Mozilla Firefox exploit evaluation
Anthropic reports neither Sonnet 4.6 nor Sonnet 5 developed a full working exploit in the Firefox 147 evaluation; Sonnet 5 had a slightly higher partial-success rate, and cyber safeguards are enabled by default.
0.0% full working exploit success%www.anthropic.com β†—

API Usage Examples

Use the global api.therouter.ai endpoint shown below for new integrations; the legacy China accelerated endpoint is retired.

cURL
curl https://api.therouter.ai/v1/chat/completions   -H "Content-Type: application/json"   -H "Authorization: Bearer $THE_ROUTER_API_KEY"   -d '{
    "model": "anthropic/claude-sonnet-5",
    "messages": [
      {"role": "user", "content": "Summarize the key points from this input."}
    ]
  }'

Chat completion

Use TheRouter's OpenAI-compatible endpoint for normal text, coding, and document prompts. Omit sampling parameters unless you know they resolve to Anthropic defaults.

cURL
curl https://api.therouter.ai/v1/chat/completions \
  -H "Authorization: Bearer $THEROUTER_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"anthropic/claude-sonnet-5","messages":[{"role":"user","content":"Audit this migration plan for hidden production risks."}],"max_tokens":2000}'

More from anthropic

Similar models

Cross-provider sibling models

News & changes

2026-07-06

Claude Sonnet 5 Skips Priority Tier: What the Service-Tier Gap Means for Latency-Sensitive Routing Teams

Recent TheRouter coverage flags the missing Priority Tier lane as the main production routing caveat: use Sonnet 5 for agentic throughput, but keep latency-sensitive SLAs on a model with explicit priority support until measured route data says otherwise.

re-authored by TheRouterTheRouter news β†—
2026-07-01

Claude Sonnet 5 Launches as Anthropic's Most Agentic Sonnet: What the Introductory Pricing Window Means for Routing Teams

The launch analysis connects Sonnet 5's agentic gains with its temporary $2/$10 pricing. The operational takeaway is to benchmark during the discount window but model September pricing and tokenizer inflation before promising durable savings.

re-authored by TheRouterTheRouter news β†—
2026-07-01

Claude Sonnet 5's Three Breaking API Changes: What Every Operator Must Audit Before Migrating

The migration note maps the three breaking behaviors to router checks: remove manual thinking budgets, remove non-default sampling, and recount tokens under the new tokenizer before shifting production traffic.

re-authored by TheRouterTheRouter news β†—

Recent coverage

Frequently asked

Is Claude Sonnet 5 a drop-in replacement for Claude Sonnet 4.6?

It is a model-id migration, but not a no-audit migration. Tool schemas and response shapes are broadly compatible, yet Sonnet 5 changes thinking behavior, rejects non-default sampling parameters, removes manual budget_tokens, and uses a tokenizer that can increase token counts for the same text.

re-authored by TheRouterplatform.claude.com β†—
What does Claude Sonnet 5 cost through TheRouter?

The upstream introductory price is $2 per million input tokens and $10 per million output tokens through August 31, 2026, then $3/$15 from September 1. TheRouter's operational pricing is sourced from standard-models.yaml on the live page; this curated module adds the billing caveat that tokenizer inflation can offset part of the apparent discount.

re-authored by TheRouterplatform.claude.com β†—
Does Claude Sonnet 5 support Priority Tier?

No. Anthropic's launch release notes state that Sonnet 5 has the same tools and platform features as Sonnet 4.6 except Priority Tier. Latency-sensitive fleets should keep a separate route or fallback until measured TheRouter latency is good enough for the SLA.

re-authored by TheRouterplatform.claude.com β†—
Should I set temperature on Claude Sonnet 5?

No for non-default values. Anthropic says non-default temperature, top_p, and top_k return HTTP 400 on Sonnet 5. Use system instructions, tool schemas, structured outputs, and effort rather than sampling knobs.

re-authored by TheRouterplatform.claude.com β†—
How should routing teams test the new tokenizer?

Re-run token counting against real production prompts before moving traffic. Anthropic says the same text can map to roughly 1.0-1.35x as many tokens depending on content type, so a workload that looks cheaper per token can still cost more per task.

re-authored by TheRouterwww.anthropic.com β†—
Fact ledger β€” every claim on this page traces here
sourceURLretrieved
Release dateplatform.claude.com β†—2026-07-28verified
Claude API model idplatform.claude.com β†—2026-07-28verified
Context windowplatform.claude.com β†—2026-07-28verified
Max output tokensplatform.claude.com β†—2026-07-28verified
Pricingplatform.claude.com β†—2026-07-28verified
Prompt cachingplatform.claude.com β†—2026-07-28verified
Training data cutoffplatform.claude.com β†—2026-07-28verified
Reliable knowledge cutoffplatform.claude.com β†—2026-07-28verified
Reasoning controlsplatform.claude.com β†—2026-07-28verified
Unsupported controlsplatform.claude.com β†—2026-07-28verified
Priority Tierplatform.claude.com β†—2026-07-28verified
Tokenizer migrationwww.anthropic.com β†—2026-07-28verified
Deploymentplatform.claude.com β†—2026-07-28verified
Licenseβ€”β€”verified
BrowseCompwww.anthropic.com β†—2026-07-28to verify
OSWorld-Verifiedwww.anthropic.com β†—2026-07-28to verify
Mozilla Firefox exploit evaluationwww.anthropic.com β†—2026-07-28verified
Claude Sonnet 5 Skips Priority Tier: What the Service-Tier Gap Means for Latency-Sensitive Routing TeamsTheRouter news β†—2026-07-28verified
Claude Sonnet 5 Launches as Anthropic's Most Agentic Sonnet: What the Introductory Pricing Window Means for Routing TeamsTheRouter news β†—2026-07-28verified
Claude Sonnet 5's Three Breaking API Changes: What Every Operator Must Audit Before MigratingTheRouter news β†—2026-07-28verified
Is Claude Sonnet 5 a drop-in replacement for Claude Sonnet 4.6?platform.claude.com β†—2026-07-28to verify
What does Claude Sonnet 5 cost through TheRouter?platform.claude.com β†—2026-07-28to verify
Does Claude Sonnet 5 support Priority Tier?platform.claude.com β†—2026-07-28to verify
Should I set temperature on Claude Sonnet 5?platform.claude.com β†—2026-07-28to verify
How should routing teams test the new tokenizer?www.anthropic.com β†—2026-07-28to verify
Help & contact