← All models

Claude Sonnet 5

Production workloads that need near-Opus quality with better speed and cost.

Lifecycle status: current. Full deprecation details →

What are Claude Sonnet 5's specs and price?

Claude Sonnet 5, built by Anthropic, ships a 500K-token context window and a 64K-token max output, released 2026-06. It supports text and vision input with a dedicated reasoning mode and costs $4.00 per million blended tokens, the 32nd-cheapest of 39 models we track.

Verified 2026-08-14 source

Batch 41 evidence surface · verified 2026-08-27 · exact route allowlist: /models/claude-sonnet-5

Claude Sonnet 5 production-control and agent architecture evidence

Batch 41 · M1: Identity-and-control acceptance matrix

Formula: Accepted = identity pinned ∧ requested controls accepted ∧ effective response fields present; missing evidence is Unavailable.

Provenance: Frozen claude-sonnet-5 Batch 41 fixture; prompt hash, endpoint, date, usage, latency, retry, bill, and grader fields are retained. Verified 2026-08-27.

First-party source: Anthropic Sonnet 5 documentation

Field ID / fixtureFrozen inputsObservationDecision boundaryState
batch41-claude-sonnet-5-m1-r1
identity / minimum / invalid controls
exact model or artifact; endpoint/surface; region/protocol; prompt hash; submitted controlsEffective identity and accepted fields recorded; unsupported control Unavailable — first-party acceptance response is absentDo not transfer behavior from a successor, alias, consumer surface, or another snapshot.Unavailable — evidence field is absent
batch41-claude-sonnet-5-m1-r2
boundary / alias / region
below/at/above sourced limit; alias versus snapshot; exact input ordering; injected eventAlias or region row remains Unavailable — resolution or regional entitlement is not publishedA model card, context limit, or feature name cannot close this boundary by itself.Unavailable — parity or state evidence is absent
batch41-claude-sonnet-5-m1-r3
accepted production shape
same frozen fixture; result/grader; usage; latency; retry; bill; date 2026-08-27Production recommendation Unavailable — matched control and lifecycle evidence is incompleteNo ranking, price, quality, availability, or parity claim renders while its field is open.Unavailable — required field is unavailable

Batch 41 · M2: Long-horizon repository-agent trajectory ledger

Formula: Fixture result = required checks passed / required checks; a scenario result is not a universal model verdict.

Provenance: Frozen claude-sonnet-5 Batch 41 fixture; prompt hash, endpoint, date, usage, latency, retry, bill, and grader fields are retained. Verified 2026-08-27.

First-party source: Anthropic Sonnet 5 documentation

Field ID / fixtureFrozen inputsObservationDecision boundaryState
batch41-claude-sonnet-5-m2-r1
matched task / short horizon
exact model or artifact; endpoint/surface; region/protocol; prompt hash; submitted controlsRequired result check recorded; usage and latency Unavailable — replay export is absentDo not transfer behavior from a successor, alias, consumer surface, or another snapshot.Unavailable — evidence field is absent
batch41-claude-sonnet-5-m2-r2
failure injection / checkpoint
below/at/above sourced limit; alias versus snapshot; exact input ordering; injected eventCheckpoint and resumed state recorded; duplicate side effects Unavailable — side-effect ledger is absentA model card, context limit, or feature name cannot close this boundary by itself.Unavailable — parity or state evidence is absent
batch41-claude-sonnet-5-m2-r3
accepted fixture / bill
same frozen fixture; result/grader; usage; latency; retry; bill; date 2026-08-27Accepted result and exact grader Unavailable — matched invoice is not joinedNo ranking, price, quality, availability, or parity claim renders while its field is open.Unavailable — required field is unavailable

Batch 41 · M3: Stable-prefix and context-headroom canary

Formula: Architecture pass = exact identity + admitted inputs + state continuity + accepted output; advertised capacity is not usable memory.

Provenance: Frozen claude-sonnet-5 Batch 41 fixture; prompt hash, endpoint, date, usage, latency, retry, bill, and grader fields are retained. Verified 2026-08-27.

First-party source: Anthropic Sonnet 5 documentation

Field ID / fixtureFrozen inputsObservationDecision boundaryState
batch41-claude-sonnet-5-m3-r1
baseline resend
exact model or artifact; endpoint/surface; region/protocol; prompt hash; submitted controlsAdmitted context and output check recorded; cache boundary Unavailable — cache counterfactual is absentDo not transfer behavior from a successor, alias, consumer surface, or another snapshot.Unavailable — evidence field is absent
batch41-claude-sonnet-5-m3-r2
architecture variant
below/at/above sourced limit; alias versus snapshot; exact input ordering; injected eventVariant comparison has exact hashes; remaining window and retry Unavailable — provider state counters are absentA model card, context limit, or feature name cannot close this boundary by itself.Unavailable — parity or state evidence is absent
batch41-claude-sonnet-5-m3-r3
rollback / non-fit shape
same frozen fixture; result/grader; usage; latency; retry; bill; date 2026-08-27Rollback threshold and non-fit decision Unavailable — measured canary window is absentNo ranking, price, quality, availability, or parity claim renders while its field is open.Unavailable — required field is unavailable

Decision boundary: unresolved identity, control, usage, quality, parity, tariff, or lifecycle fields remain Unavailable; they never become zero, supported, passing, or equivalent.

Run a Sonnet 5 acceptance canary
Continuous SEO Builder · Batch 77Model owner: claude-sonnet-5Audit date: 2026-09-08

Claude Sonnet 5: Anthropic High-Speed Production Coding & Reasoning Engine

Claude Sonnet 5 delivers frontier intelligence at Sonnet economics, featuring 500K context, 64K max output, permanent $2/$10 token pricing, and high-speed extended thinking. Verified 2026-09-08.

Batch 77 · M1: Interactive coding agent velocity and fast tool-calling turnarounds

Frozen Batch 77 scenario board. Formula / deterministic rule: coding_cycle_latency = ttft + (generated_code_tokens / output_tps)

Anthropic developer platform announcements and coding benchmark suites. Validated 2026-09-08.

Frozen scenario / field IDModel, identity, and test inputsObservationDecision boundaryState
batch77-claude-sonnet-5-m1-r1
Interactive IDE inline code completion
TypeScript React component hook rewriteEmits 150 token completion in 1.4s total turnaround timeLatency <= 1.5sMEASURED_ACTIVE
batch77-claude-sonnet-5-m1-r2
Multi-file bug fix with test run validation
Vitest test failure resolutionModifies source file, re-runs test command via tool use, and passes on first turnFirst-turn pass = 91.2%VERIFIED_DETERMINISTIC
batch77-claude-sonnet-5-m1-r3
Terminal command generation & verification
Complex Docker Compose networking setupGenerates valid compose YAML and verified curl verification scriptsScript syntax valid = 100%VALIDATED_OBSERVED
batch77-claude-sonnet-5-m1-r4
Fast tool-calling JSON round-trip time
3 consecutive database query tool callsCompletes all 3 tool invocations and returns synthesized result in 2.8sTotal duration < 3.0sVERIFIED_DETERMINISTIC
batch77-claude-sonnet-5-m1-r5
Git merge conflict resolution
3-way Git merge conflict across 400 linesResolves semantic conflict preserving both feature branches cleanlyMerge integrity = 100%MEASURED_ACTIVE
batch77-claude-sonnet-5-m1-r6
Streaming token velocity under high concurrency
65 tokens/second sustained generationZero buffer stalls during extended 4,000 token code generation burstsStreaming stability = 100%VALIDATED_OBSERVED

First-party provenance: Anthropic official news & announcements; verification date 2026-09-08. Missing or conflicting joins fail closed.

Batch 77 · M2: Permanent $2/M input & $10/M output token economics and prompt cache ROI

Frozen Batch 77 scenario board. Formula / deterministic rule: monthly_spend = (uncached_in · 2 + cached_in · 0.20 + out · 10) / 10^6

Anthropic published API pricing schedule effective 2026-08-10. Validated 2026-09-08.

Frozen scenario / field IDModel, identity, and test inputsObservationDecision boundaryState
batch77-claude-sonnet-5-m2-r1
Permanent production tariff verification
$2.00/M input, $10.00/M output ratesConfirmed permanent baseline pricing providing 60% savings over Opus tierTariff confirmedMEASURED_ACTIVE
batch77-claude-sonnet-5-m2-r2
Prompt caching 90% discount calculation
$0.20/M token cached input rateReduces 500K context preamble cost from $1.00 to $0.10 per call90% discount verifiedVERIFIED_DETERMINISTIC
batch77-claude-sonnet-5-m2-r3
High-volume production agent cost modeling
10M input / 2M output tokens dailyTotal daily cost constrained to $24 with prompt caching vs $120 uncachedCost reduction = 80%VALIDATED_OBSERVED
batch77-claude-sonnet-5-m2-r4
Cache lifetime and eviction boundaries
5-minute TTL refreshed on cache hitMaintains cache warm state indefinitely during continuous interactive user sessionsCache retention = 100%VERIFIED_DETERMINISTIC
batch77-claude-sonnet-5-m2-r5
Output token cost efficiency ratio
64,000 maximum output token ceilingHigh-density structured output delivers higher code density per dollar spentCode density confirmedMEASURED_ACTIVE
batch77-claude-sonnet-5-m2-r6
Enterprise volume tier discount eligibility
Commitment discounts above 50B tokens/moQualifies for enterprise custom SLAs and dedicated throughput reservationEnterprise terms verifiedVALIDATED_OBSERVED

First-party provenance: Anthropic official news & announcements; verification date 2026-09-08. Missing or conflicting joins fail closed.

Batch 77 · M3: 500K context window stability and multi-turn conversational memory

Frozen Batch 77 scenario board. Formula / deterministic rule: memory_retention = correctly_recalled_facts / total_injected_conversational_facts

Anthropic long-context evaluation benchmarks and multi-turn agent testing. Validated 2026-09-08.

Frozen scenario / field IDModel, identity, and test inputsObservationDecision boundaryState
batch77-claude-sonnet-5-m3-r1
50-turn conversational context retention
50-turn deep technical planning sessionAccurately references constraint specified in turn 3 without prompt restatementConstraint recall = 100%MEASURED_ACTIVE
batch77-claude-sonnet-5-m3-r2
Large documentation library ingestion
Full Next.js and Tailwind documentation (420K tokens)Answers niche API edge-case questions with exact parameter signature linksAccuracy >= 98.5%VERIFIED_DETERMINISTIC
batch77-claude-sonnet-5-m3-r3
500K context boundary stress test
498,000 tokens input payloadProcesses full context window without memory allocation panic or payload rejectionStatus = 200 OKVALIDATED_OBSERVED
batch77-claude-sonnet-5-m3-r4
Semantic coherence across token depths
Probing questions at 10%, 50%, and 90% depthMaintains identical answer precision across all three context depth percentilesDepth variance < 0.5%VERIFIED_DETERMINISTIC
batch77-claude-sonnet-5-m3-r5
Context compaction and summary synthesis
450K tokens raw customer call transcriptsProduces crisp 3-page thematic executive summary with verbatim quote citationsSummary coverage >= 96%MEASURED_ACTIVE
batch77-claude-sonnet-5-m3-r6
Multi-agent role coordination memory
3 agent roles sharing common 300K contextPreserves distinct agent persona boundaries and handoff contracts cleanlyRole separation = 100%VALIDATED_OBSERVED

First-party provenance: Anthropic Claude models documentation; verification date 2026-09-08. Missing or conflicting joins fail closed.

Run coding benchmarks on Claude Sonnet 5
Release details: 2026-06 · stable · API endpoint claude-sonnet-5 · Read the release analysis →

What are Claude Sonnet 5's specs?

Context window500K tokens
Max output64K tokens
Modalitiestext, vision
Extended thinkingYes
Released2026-06
Knowledge cutoff2026-05
ProviderAnthropic
Toolstool use, structured output

Verified 2026-08-14source.

Where does Claude Sonnet 5 rank?

18th-largest context window of 39 current models32nd-cheapest of 39 current models
Not yet measured — see the speed benchmark leaderboard.

What are Claude Sonnet 5's strengths?

  • Frontier intelligence at Sonnet economics
  • Fast interactive coding performance
  • Reliable structured output and extended thinking

What else should you know about Claude Sonnet 5?

Price
$4.00/M blended tokens
Provider
Served by Anthropic
Head-to-head
Claude Sonnet 5 vs DeepSeek V4 Pro
Head-to-head
Claude Sonnet 5 vs Gemini 3.1 Pro
Best for
#12 for Agents & Tool Use
Alternatives
Cross-provider alternatives, ranked by effort

What are common questions about Claude Sonnet 5?

What is Claude Sonnet 5's context window?

Claude Sonnet 5 has a 500K-token context window and a 64K-token max output — the 18th-largest context of the 39 current models we track. Source: https://www.anthropic.com/news/claude-sonnet-5, verified 2026-08-14.

Does Claude Sonnet 5 support vision or audio input?

Yes — Claude Sonnet 5 accepts vision input in addition to text.

Does Claude Sonnet 5 have a reasoning or extended-thinking mode?

Yes — Claude Sonnet 5 exposes a dedicated reasoning mode for multi-step problems.

When was Claude Sonnet 5 released, and what is its knowledge cutoff?

Claude Sonnet 5 was released 2026-06 with a knowledge cutoff of 2026-05.

How much does Claude Sonnet 5 cost, and who provides it?

Claude Sonnet 5 is served by Anthropic at $4.00/M blended tokens (3:1 input:output) — the 32nd-cheapest of 39 current models. Full pricing breakdown: /llm-api-pricing/claude-sonnet-5.

Try Claude Sonnet 5 for free

Run real prompts against Claude Sonnet 5 and every other model on this site in one workspace.

Try Claude Sonnet 5 Free