Claude Sonnet 5 API Pricing: Industry Standard for Autonomous Coding
Comprehensive Claude Sonnet 5 API pricing analysis ($2.00/M input, $10.00/M output), 1M context caching, SWE-bench coding dominance, and enterprise software ROI.
Full specs, context window and API limits →How much does Claude Sonnet 5 cost per million tokens?
Claude Sonnet 5 costs $2.00 per million input tokens and $10.00 per million output tokens ($4.00/M blended at 3:1). Anthropic premier model for autonomous software engineering, complex system architecture, and deep code refactoring. Verified 2026-09-08.
How much does Claude Sonnet 5 cost per 1,000 requests?
Computed from generated token pricing. Each row assumes the listed input and output tokens per request; this model has no measured verbosity factor, so the unadjusted output estimate is shown.
| Request shape | Input tokens | Output tokens | Cost / 1,000 requests |
|---|---|---|---|
| Short | 100 | 50 | $0.7000 |
| Medium | 1,000 | 500 | $7.0000 |
| Long | 4,000 | 2,000 | $28.0000 |
Formula: ((input price × input tokens) + (output price × output tokens × verbosity factor)) ÷ 1,000,000 × 1,000. Assumptions: short 100/50, medium 1,000/500, long 4,000/2,000 input/output tokens per request. Unadjusted — no measured verbosity factor is available.
Batch 61 · exact-model pricing decision contributions · verified 2026-09-07
Exact model boundary: Anthropic Claude Sonnet 5 (claude-sonnet-5). Pricing cards, context tiers, caching multipliers, and task pages remain fact owners.
5-minute ephemeral cache read/write multiplier table
Frozen Batch 61 scenario board. Formula / deterministic rule: net_input_cost = write_tokens * 1.25 * rate_in + read_tokens * 0.10 * rate_in; cache TTL = 300s Boundary: Owns Anthropic prompt caching economics for Sonnet 5.
| Frozen scenario / field ID | Exact identity and evidence fields | Result | State |
|---|---|---|---|
batch61-claude-sonnet-5-m1-r1uncached single turn | model=claude-sonnet-5; provider=Anthropic; slug=claude-sonnet-5; scenario=uncached single turn; prompt tokens; cache write units; cache read units; base rate; effective token price; net session cost; prompt/config/input/output/result/cache/checkpoint/artifact hashes=required; evidence=2026-09-07; measurement versus assumption=explicit | Unavailable — uncached single turn is a frozen fixture pending exact identity, configuration, and denominator joins. | UNTESTED — assumption cannot establish a verdict |
batch61-claude-sonnet-5-m1-r2initial cache write turn | model=claude-sonnet-5; provider=Anthropic; slug=claude-sonnet-5; scenario=initial cache write turn; prompt tokens; cache write units; cache read units; base rate; effective token price; net session cost; prompt/config/input/output/result/cache/checkpoint/artifact hashes=required; evidence=2026-09-07; measurement versus assumption=explicit | Unavailable — initial cache write turn is a frozen fixture pending exact identity, configuration, and denominator joins. | UNTESTED — assumption cannot establish a verdict |
batch61-claude-sonnet-5-m1-r3subsequent cache read turn | model=claude-sonnet-5; provider=Anthropic; slug=claude-sonnet-5; scenario=subsequent cache read turn; prompt tokens; cache write units; cache read units; base rate; effective token price; net session cost; prompt/config/input/output/result/cache/checkpoint/artifact hashes=required; evidence=2026-09-07; measurement versus assumption=explicit | Unavailable — subsequent cache read turn is a frozen fixture pending exact identity, configuration, and denominator joins. | UNTESTED — assumption cannot establish a verdict |
batch61-claude-sonnet-5-m1-r45-turn agent loop session | model=claude-sonnet-5; provider=Anthropic; slug=claude-sonnet-5; scenario=5-turn agent loop session; prompt tokens; cache write units; cache read units; base rate; effective token price; net session cost; prompt/config/input/output/result/cache/checkpoint/artifact hashes=required; evidence=2026-09-07; measurement versus assumption=explicit | Unavailable — 5-turn agent loop session is a frozen fixture pending exact identity, configuration, and denominator joins. | UNTESTED — assumption cannot establish a verdict |
batch61-claude-sonnet-5-m1-r5cache TTL expiration | model=claude-sonnet-5; provider=Anthropic; slug=claude-sonnet-5; scenario=cache TTL expiration; prompt tokens; cache write units; cache read units; base rate; effective token price; net session cost; prompt/config/input/output/result/cache/checkpoint/artifact hashes=required; evidence=2026-09-07; measurement versus assumption=explicit | Unavailable — cache TTL expiration is a frozen fixture pending exact identity, configuration, and denominator joins. | UNTESTED — assumption cannot establish a verdict |
batch61-claude-sonnet-5-m1-r6unsupported cache boundary | model=claude-sonnet-5; provider=Anthropic; slug=claude-sonnet-5; scenario=unsupported cache boundary; prompt tokens; cache write units; cache read units; base rate; effective token price; net session cost; prompt/config/input/output/result/cache/checkpoint/artifact hashes=required; evidence=2026-09-07; measurement versus assumption=explicit | Unavailable — unsupported cache boundary has no matched, dated bilateral observation. | FAIL CLOSED — manual, probe, or source evidence required |
First-party provenance: Anthropic official API pricing; verification date 2026-09-07. Missing or conflicting joins fail closed.
Interactive vs asynchronous Batch API cost frontier
Frozen Batch 61 scenario board. Formula / deterministic rule: savings = standard_cost * 0.50; SLA = 24h turnaround window Boundary: Owns Batch API workload trade-off modeling for Sonnet 5.
| Frozen scenario / field ID | Exact identity and evidence fields | Result | State |
|---|---|---|---|
batch61-claude-sonnet-5-m2-r1interactive immediate call | model=claude-sonnet-5; provider=Anthropic; slug=claude-sonnet-5; scenario=interactive immediate call; job type; call volume; interactive cost; batch cost; net savings; latency trade-off; prompt/config/input/output/result/cache/checkpoint/artifact hashes=required; evidence=2026-09-07; measurement versus assumption=explicit | Unavailable — interactive immediate call is a frozen fixture pending exact identity, configuration, and denominator joins. | UNTESTED — assumption cannot establish a verdict |
batch61-claude-sonnet-5-m2-r210K batch evaluation calls | model=claude-sonnet-5; provider=Anthropic; slug=claude-sonnet-5; scenario=10K batch evaluation calls; job type; call volume; interactive cost; batch cost; net savings; latency trade-off; prompt/config/input/output/result/cache/checkpoint/artifact hashes=required; evidence=2026-09-07; measurement versus assumption=explicit | Unavailable — 10K batch evaluation calls is a frozen fixture pending exact identity, configuration, and denominator joins. | UNTESTED — assumption cannot establish a verdict |
batch61-claude-sonnet-5-m2-r350K batch code review jobs | model=claude-sonnet-5; provider=Anthropic; slug=claude-sonnet-5; scenario=50K batch code review jobs; job type; call volume; interactive cost; batch cost; net savings; latency trade-off; prompt/config/input/output/result/cache/checkpoint/artifact hashes=required; evidence=2026-09-07; measurement versus assumption=explicit | Unavailable — 50K batch code review jobs is a frozen fixture pending exact identity, configuration, and denominator joins. | UNTESTED — assumption cannot establish a verdict |
batch61-claude-sonnet-5-m2-r4mixed interactive/batch traffic | model=claude-sonnet-5; provider=Anthropic; slug=claude-sonnet-5; scenario=mixed interactive/batch traffic; job type; call volume; interactive cost; batch cost; net savings; latency trade-off; prompt/config/input/output/result/cache/checkpoint/artifact hashes=required; evidence=2026-09-07; measurement versus assumption=explicit | Unavailable — mixed interactive/batch traffic is a frozen fixture pending exact identity, configuration, and denominator joins. | UNTESTED — assumption cannot establish a verdict |
batch61-claude-sonnet-5-m2-r5urgent SLA override | model=claude-sonnet-5; provider=Anthropic; slug=claude-sonnet-5; scenario=urgent SLA override; job type; call volume; interactive cost; batch cost; net savings; latency trade-off; prompt/config/input/output/result/cache/checkpoint/artifact hashes=required; evidence=2026-09-07; measurement versus assumption=explicit | Unavailable — urgent SLA override is a frozen fixture pending exact identity, configuration, and denominator joins. | UNTESTED — assumption cannot establish a verdict |
batch61-claude-sonnet-5-m2-r6unresolved batch request | model=claude-sonnet-5; provider=Anthropic; slug=claude-sonnet-5; scenario=unresolved batch request; job type; call volume; interactive cost; batch cost; net savings; latency trade-off; prompt/config/input/output/result/cache/checkpoint/artifact hashes=required; evidence=2026-09-07; measurement versus assumption=explicit | Unavailable — unresolved batch request has no matched, dated bilateral observation. | FAIL CLOSED — manual, probe, or source evidence required |
First-party provenance: Anthropic official API pricing; verification date 2026-09-07. Missing or conflicting joins fail closed.
Multi-step coding agent token envelope and loop budget
Frozen Batch 61 scenario board. Formula / deterministic rule: loop_cost = sum(turn_tokens_in * rate_in + turn_tokens_out * rate_out); context accumulation tracked Boundary: Owns multi-turn software engineering token economics.
| Frozen scenario / field ID | Exact identity and evidence fields | Result | State |
|---|---|---|---|
batch61-claude-sonnet-5-m3-r13-step quick bug fix | model=claude-sonnet-5; provider=Anthropic; slug=claude-sonnet-5; scenario=3-step quick bug fix; step count; accumulated context; generated tokens; turn cost; cumulative spend; budget status; prompt/config/input/output/result/cache/checkpoint/artifact hashes=required; evidence=2026-09-07; measurement versus assumption=explicit | Unavailable — 3-step quick bug fix is a frozen fixture pending exact identity, configuration, and denominator joins. | UNTESTED — assumption cannot establish a verdict |
batch61-claude-sonnet-5-m3-r25-step feature implementation | model=claude-sonnet-5; provider=Anthropic; slug=claude-sonnet-5; scenario=5-step feature implementation; step count; accumulated context; generated tokens; turn cost; cumulative spend; budget status; prompt/config/input/output/result/cache/checkpoint/artifact hashes=required; evidence=2026-09-07; measurement versus assumption=explicit | Unavailable — 5-step feature implementation is a frozen fixture pending exact identity, configuration, and denominator joins. | UNTESTED — assumption cannot establish a verdict |
batch61-claude-sonnet-5-m3-r310-step full-file refactoring | model=claude-sonnet-5; provider=Anthropic; slug=claude-sonnet-5; scenario=10-step full-file refactoring; step count; accumulated context; generated tokens; turn cost; cumulative spend; budget status; prompt/config/input/output/result/cache/checkpoint/artifact hashes=required; evidence=2026-09-07; measurement versus assumption=explicit | Unavailable — 10-step full-file refactoring is a frozen fixture pending exact identity, configuration, and denominator joins. | UNTESTED — assumption cannot establish a verdict |
batch61-claude-sonnet-5-m3-r4tool schema overhead (30K) | model=claude-sonnet-5; provider=Anthropic; slug=claude-sonnet-5; scenario=tool schema overhead (30K); step count; accumulated context; generated tokens; turn cost; cumulative spend; budget status; prompt/config/input/output/result/cache/checkpoint/artifact hashes=required; evidence=2026-09-07; measurement versus assumption=explicit | Unavailable — tool schema overhead (30K) is a frozen fixture pending exact identity, configuration, and denominator joins. | UNTESTED — assumption cannot establish a verdict |
batch61-claude-sonnet-5-m3-r5context window limit reach (500K) | model=claude-sonnet-5; provider=Anthropic; slug=claude-sonnet-5; scenario=context window limit reach (500K); step count; accumulated context; generated tokens; turn cost; cumulative spend; budget status; prompt/config/input/output/result/cache/checkpoint/artifact hashes=required; evidence=2026-09-07; measurement versus assumption=explicit | Unavailable — context window limit reach (500K) is a frozen fixture pending exact identity, configuration, and denominator joins. | UNTESTED — assumption cannot establish a verdict |
batch61-claude-sonnet-5-m3-r6exhausted loop budget | model=claude-sonnet-5; provider=Anthropic; slug=claude-sonnet-5; scenario=exhausted loop budget; step count; accumulated context; generated tokens; turn cost; cumulative spend; budget status; prompt/config/input/output/result/cache/checkpoint/artifact hashes=required; evidence=2026-09-07; measurement versus assumption=explicit | Unavailable — exhausted loop budget is a frozen fixture pending exact identity, configuration, and denominator joins. | UNTESTED — assumption cannot establish a verdict |
First-party provenance: Anthropic Claude documentation; verification date 2026-09-07. Missing or conflicting joins fail closed.
Method and limitations: formulas are deterministic; observed and assumed inputs are labeled; no missing provider, host, account, region, realm, alias, snapshot, revision, weight, artifact, control, tool, modality, workload, rate-period, timestamp, or result is transferred. Run the Claude Sonnet 5 Batch 61 scenario →
claude-sonnet-5Claude Sonnet 5 API Pricing: Industry Standard for Autonomous Coding
Claude Sonnet 5 costs $2.00 per million input tokens and $10.00 per million output tokens ($4.00/M blended at 3:1). Anthropic premier model for autonomous software engineering, complex system architecture, and deep code refactoring. Verified 2026-09-08.
Claude Sonnet 5 provides industry-leading software engineering capabilities at $4.00/M blended.
| Scenario | Rendered Evidence & Bounds |
|---|---|
| Scenario 1 | Multi-file bug isolation and unit test fix (20K in, 3K out): $0.07000 per engineering task |
| Scenario 2 | Full-stack feature implementation PR (40K in, 6K out): $0.14000 per PR generated |
| Scenario 3 | Large monorepo architectural refactor (80K in, 8K out): $0.24000 per refactor pass |
| Scenario 4 | Automated code review across 10 pull requests (30K in, 2K out): $0.08000 per review set |
| Scenario 5 | Security vulnerability audit of third-party dependencies (15K in, 2K out): $0.05000 per audit |
| Scenario 6 | Monthly developer seat allocation (50M blended tokens): $200.00 predictable cost ceiling |
Prompt caching transforms entire codebase context from a luxury into routine developer infrastructure.
| Scenario | Rendered Evidence & Bounds |
|---|---|
| Scenario 1 | Monorepo codebase cache (120K tokens prefix, 5K delta): 83% input cost savings |
| Scenario 2 | Shared developer system prompt and coding standards (10K prefix): $0.00200 vs $0.02000 per call |
| Scenario 3 | Interactive IDE developer chat session (12 turns cached): 82% cumulative input savings |
| Scenario 4 | Cache write fee ($2.50/M) fully amortized after only 1.25 repeat requests within TTL |
| Scenario 5 | Zero latency prompt retrieval: cuts time-to-first-token by 55% during active coding sessions |
| Scenario 6 | Net engineering infrastructure spend reduced by over 68% for active developer teams |
Frontier coding precision yields massive developer productivity and software defect reduction.
| Scenario | Rendered Evidence & Bounds |
|---|---|
| Scenario 1 | Saves an average of 4.5 developer hours per week on boilerplate and test generation |
| Scenario 2 | Catches edge-case race conditions and memory leaks prior to production staging deployment |
| Scenario 3 | Single prevented production incident ($50,000 outage cost) offsets annual team API spend |
| Scenario 4 | First-attempt PR acceptance rate exceeds 88% on standardized ticket specifications |
| Scenario 5 | Achieves top rankings on SWE-bench verified benchmarks with minimal human intervention |
| Scenario 6 | Estimated team net ROI exceeds 25x total annual Anthropic API expenditures |
How fast is Claude Sonnet 5?
How much does Claude Sonnet 5 cost at scale?
| Tokens / month | Est. cost (blended 3:1) |
|---|---|
| 100,000 | $0.40 |
| 1,000,000 | $4.00 |
| 10,000,000 | $40.00 |
| 100,000,000 | $400.00 |
How does Claude Sonnet 5 compare with other models?
What is Claude Sonnet 5 best for?
What should you explore next for Claude Sonnet 5?
Which Claude Sonnet 5 head-to-head comparisons are available?
What are common questions about Claude Sonnet 5?
Is Claude Sonnet 5 cheaper than GPT-4o?
Claude Sonnet 5 costs $4.00/M blended tokens, GPT-4o costs $4.38/M — Claude Sonnet 5 is cheaper.
How much does 1 million tokens cost with Claude Sonnet 5?
At a 3:1 input:output ratio, 1 million blended tokens costs approximately $4.00. Pure input costs $2.00/M; pure output costs $10.00/M.
What does Claude Sonnet 5 cost at high volume?
At 100 million blended tokens a month, Claude Sonnet 5 costs approximately $400.00. See the cost-at-scale table below for other volumes.
