← Back to all pricing

Claude Sonnet 4.6 API Pricing: Proven 1M Context Coding Dominance

Comprehensive Claude Sonnet 4.6 API pricing analysis ($3.00/M input, $15.00/M output), 1M context window caching, SWE-bench coding benchmarks, and upgrade comparisons.

Full specs, context window and API limits →

How much does Claude Sonnet 4.6 cost per million tokens?

Claude Sonnet 4.6 costs $3.00 per million input tokens and $15.00 per million output tokens ($6.00/M blended at 3:1). A premier coding and analytical model offering an expansive 1M context window, exceptional multi-file repository navigation, and 90% prompt caching discounts. Verified 2026-09-08.

Verified 2026-09-07 source
Input
$3.00/M
Output
$15.00/M
Blended
$6.00/M
Provider
Verified 2026-04-06source

How much does Claude Sonnet 4.6 cost per 1,000 requests?

Computed from generated token pricing. Each row assumes the listed input and output tokens per request; this model has no measured verbosity factor, so the unadjusted output estimate is shown.

Request shapeInput tokensOutput tokensCost / 1,000 requests
Short10050$1.0500
Medium1,000500$10.5000
Long4,0002,000$42.0000

Formula: ((input price × input tokens) + (output price × output tokens × verbosity factor)) ÷ 1,000,000 × 1,000. Assumptions: short 100/50, medium 1,000/500, long 4,000/2,000 input/output tokens per request. Unadjusted — no measured verbosity factor is available.

Three model-specific pricing decisions

Sonnet 4.6’s cache decision is shown for short and long prefixes; the Sonnet 5 page and broad comparison retain their own owners.

1. Cache-write/read reuse break-even

Prefix shapeUncached / 100KWrite once + 9 readsBreak-even
Short · 1,024 tokens$0.08$0.061 reuse(s)
Long · 8,000 tokens$0.36$0.17Same TTL; long-context tier unavailable

2. Sonnet 4.6 → Sonnet 5 migration boundary

Monthly cost and accepted-result threshold

Fixed workloadClaude Sonnet 4.6Claude Sonnet 5Boundary
Interactive$1245.00$830.0010% accepted-result uplift needed to justify premium
Long document$3600.00$2400.0015% accepted-result uplift needed to justify premium

Formula: calls × (input tokens × input $/M + output tokens × output $/M) ÷ 1,000,000. The uplift threshold is a decision input, not a measured quality claim.

3. Interactive versus batch turnaround

ModeCost / 100KTurnaround evidenceDecision
Interactive$1245.00List pricing dated; latency unavailableUser-blocking
Batch$982.50Batch discount/turnaround model-specific unavailableAsync only

Verified 2026-04-06. Luna is the data owner for this rendered decision module. “Unavailable” means the current dated registry has no model-specific evidence; it is not a zero. First-party price source · Run this scenario in the playground.

All results are server-rendered for Claude Sonnet 4.6; formulas expose fixed inputs and missing evidence remains visibly unavailable.

Batch 62 · exact-model pricing decision contributions · verified 2026-09-07

Exact model boundary: Anthropic Claude Sonnet 4.6 (claude-sonnet-4-6). Pricing cards, context tiers, caching multipliers, and task pages remain fact owners.

Prompt caching economics and multi-turn agent efficiency

Frozen Batch 62 scenario board. Formula / deterministic rule: cost = (uncached_in * 3.00 + cache_write * 3.75 + cache_read * 0.30 + out * 15.00) / 1M Boundary: Owns Sonnet 4.6 multi-turn cache economics.

Frozen scenario / field IDExact identity and evidence fieldsResultState
batch62-claude-sonnet-4-6-m1-r1
uncached single turn query
model=claude-sonnet-4-6; provider=Anthropic; slug=claude-sonnet-4-6; scenario=uncached single turn query; interactive turns; prompt tokens; cached tokens; output tokens; total cost; effective token rate; prompt/config/input/output/result/cache/checkpoint/artifact hashes=required; evidence=2026-09-07; measurement versus assumption=explicitUnavailable — uncached single turn query is a frozen fixture pending exact identity, configuration, and denominator joins.UNTESTED — assumption cannot establish a verdict
batch62-claude-sonnet-4-6-m1-r2
3-step customer agent flow
model=claude-sonnet-4-6; provider=Anthropic; slug=claude-sonnet-4-6; scenario=3-step customer agent flow; interactive turns; prompt tokens; cached tokens; output tokens; total cost; effective token rate; prompt/config/input/output/result/cache/checkpoint/artifact hashes=required; evidence=2026-09-07; measurement versus assumption=explicitUnavailable — 3-step customer agent flow is a frozen fixture pending exact identity, configuration, and denominator joins.UNTESTED — assumption cannot establish a verdict
batch62-claude-sonnet-4-6-m1-r3
10-step autonomous coding agent
model=claude-sonnet-4-6; provider=Anthropic; slug=claude-sonnet-4-6; scenario=10-step autonomous coding agent; interactive turns; prompt tokens; cached tokens; output tokens; total cost; effective token rate; prompt/config/input/output/result/cache/checkpoint/artifact hashes=required; evidence=2026-09-07; measurement versus assumption=explicitUnavailable — 10-step autonomous coding agent is a frozen fixture pending exact identity, configuration, and denominator joins.UNTESTED — assumption cannot establish a verdict
batch62-claude-sonnet-4-6-m1-r4
50-step multi-file workflow
model=claude-sonnet-4-6; provider=Anthropic; slug=claude-sonnet-4-6; scenario=50-step multi-file workflow; interactive turns; prompt tokens; cached tokens; output tokens; total cost; effective token rate; prompt/config/input/output/result/cache/checkpoint/artifact hashes=required; evidence=2026-09-07; measurement versus assumption=explicitUnavailable — 50-step multi-file workflow is a frozen fixture pending exact identity, configuration, and denominator joins.UNTESTED — assumption cannot establish a verdict
batch62-claude-sonnet-4-6-m1-r5
cache miss invalidation penalty
model=claude-sonnet-4-6; provider=Anthropic; slug=claude-sonnet-4-6; scenario=cache miss invalidation penalty; interactive turns; prompt tokens; cached tokens; output tokens; total cost; effective token rate; prompt/config/input/output/result/cache/checkpoint/artifact hashes=required; evidence=2026-09-07; measurement versus assumption=explicitUnavailable — cache miss invalidation penalty is a frozen fixture pending exact identity, configuration, and denominator joins.UNTESTED — assumption cannot establish a verdict
batch62-claude-sonnet-4-6-m1-r6
unregistered cache breakpoint
model=claude-sonnet-4-6; provider=Anthropic; slug=claude-sonnet-4-6; scenario=unregistered cache breakpoint; interactive turns; prompt tokens; cached tokens; output tokens; total cost; effective token rate; prompt/config/input/output/result/cache/checkpoint/artifact hashes=required; evidence=2026-09-07; measurement versus assumption=explicitUnavailable — unregistered cache breakpoint is a frozen fixture pending exact identity, configuration, and denominator joins.UNTESTED — assumption cannot establish a verdict

First-party provenance: Anthropic official API pricing; verification date 2026-09-07. Missing or conflicting joins fail closed.

Sonnet 4.6 vs Sonnet 5 parity and migration budget

Frozen Batch 62 scenario board. Formula / deterministic rule: delta = sonnet5_monthly_cost - sonnet46_monthly_cost; price parity exists at base rates Boundary: Owns migration decision models from Sonnet 4.6 to Sonnet 5.

Frozen scenario / field IDExact identity and evidence fieldsResultState
batch62-claude-sonnet-4-6-m2-r1
identical base rate comparison ($3/$15)
model=claude-sonnet-4-6; provider=Anthropic; slug=claude-sonnet-4-6; scenario=identical base rate comparison ($3/$15); monthly call volume; Sonnet 4.6 bill; Sonnet 5 bill; net cost delta; quality uplift requirement; migration status; prompt/config/input/output/result/cache/checkpoint/artifact hashes=required; evidence=2026-09-07; measurement versus assumption=explicitUnavailable — identical base rate comparison ($3/$15) is a frozen fixture pending exact identity, configuration, and denominator joins.UNTESTED — assumption cannot establish a verdict
batch62-claude-sonnet-4-6-m2-r2
complex coding quality uplift
model=claude-sonnet-4-6; provider=Anthropic; slug=claude-sonnet-4-6; scenario=complex coding quality uplift; monthly call volume; Sonnet 4.6 bill; Sonnet 5 bill; net cost delta; quality uplift requirement; migration status; prompt/config/input/output/result/cache/checkpoint/artifact hashes=required; evidence=2026-09-07; measurement versus assumption=explicitUnavailable — complex coding quality uplift is a frozen fixture pending exact identity, configuration, and denominator joins.UNTESTED — assumption cannot establish a verdict
batch62-claude-sonnet-4-6-m2-r3
agentic tool orchestration accuracy
model=claude-sonnet-4-6; provider=Anthropic; slug=claude-sonnet-4-6; scenario=agentic tool orchestration accuracy; monthly call volume; Sonnet 4.6 bill; Sonnet 5 bill; net cost delta; quality uplift requirement; migration status; prompt/config/input/output/result/cache/checkpoint/artifact hashes=required; evidence=2026-09-07; measurement versus assumption=explicitUnavailable — agentic tool orchestration accuracy is a frozen fixture pending exact identity, configuration, and denominator joins.UNTESTED — assumption cannot establish a verdict
batch62-claude-sonnet-4-6-m2-r4
low-defect production threshold
model=claude-sonnet-4-6; provider=Anthropic; slug=claude-sonnet-4-6; scenario=low-defect production threshold; monthly call volume; Sonnet 4.6 bill; Sonnet 5 bill; net cost delta; quality uplift requirement; migration status; prompt/config/input/output/result/cache/checkpoint/artifact hashes=required; evidence=2026-09-07; measurement versus assumption=explicitUnavailable — low-defect production threshold is a frozen fixture pending exact identity, configuration, and denominator joins.UNTESTED — assumption cannot establish a verdict
batch62-claude-sonnet-4-6-m2-r5
500K context saturation workload
model=claude-sonnet-4-6; provider=Anthropic; slug=claude-sonnet-4-6; scenario=500K context saturation workload; monthly call volume; Sonnet 4.6 bill; Sonnet 5 bill; net cost delta; quality uplift requirement; migration status; prompt/config/input/output/result/cache/checkpoint/artifact hashes=required; evidence=2026-09-07; measurement versus assumption=explicitUnavailable — 500K context saturation workload is a frozen fixture pending exact identity, configuration, and denominator joins.UNTESTED — assumption cannot establish a verdict
batch62-claude-sonnet-4-6-m2-r6
untested reasoning delta
model=claude-sonnet-4-6; provider=Anthropic; slug=claude-sonnet-4-6; scenario=untested reasoning delta; monthly call volume; Sonnet 4.6 bill; Sonnet 5 bill; net cost delta; quality uplift requirement; migration status; prompt/config/input/output/result/cache/checkpoint/artifact hashes=required; evidence=2026-09-07; measurement versus assumption=explicitUnavailable — untested reasoning delta is a frozen fixture pending exact identity, configuration, and denominator joins.UNTESTED — assumption cannot establish a verdict

First-party provenance: Anthropic Claude documentation; verification date 2026-09-07. Missing or conflicting joins fail closed.

Batch API processing and high-volume document ingestion

Frozen Batch 62 scenario board. Formula / deterministic rule: batch_spend = calls * ((in_tokens * 1.50 + out_tokens * 7.50) / 1M) Boundary: Owns offline batch document processing economics for Sonnet 4.6.

Frozen scenario / field IDExact identity and evidence fieldsResultState
batch62-claude-sonnet-4-6-m3-r1
10K legal contracts extraction
model=claude-sonnet-4-6; provider=Anthropic; slug=claude-sonnet-4-6; scenario=10K legal contracts extraction; batch volume; input tokens; output tokens; interactive spend; batch spend; total savings; prompt/config/input/output/result/cache/checkpoint/artifact hashes=required; evidence=2026-09-07; measurement versus assumption=explicitUnavailable — 10K legal contracts extraction is a frozen fixture pending exact identity, configuration, and denominator joins.UNTESTED — assumption cannot establish a verdict
batch62-claude-sonnet-4-6-m3-r2
50K support ticket summaries
model=claude-sonnet-4-6; provider=Anthropic; slug=claude-sonnet-4-6; scenario=50K support ticket summaries; batch volume; input tokens; output tokens; interactive spend; batch spend; total savings; prompt/config/input/output/result/cache/checkpoint/artifact hashes=required; evidence=2026-09-07; measurement versus assumption=explicitUnavailable — 50K support ticket summaries is a frozen fixture pending exact identity, configuration, and denominator joins.UNTESTED — assumption cannot establish a verdict
batch62-claude-sonnet-4-6-m3-r3
200K document classification tasks
model=claude-sonnet-4-6; provider=Anthropic; slug=claude-sonnet-4-6; scenario=200K document classification tasks; batch volume; input tokens; output tokens; interactive spend; batch spend; total savings; prompt/config/input/output/result/cache/checkpoint/artifact hashes=required; evidence=2026-09-07; measurement versus assumption=explicitUnavailable — 200K document classification tasks is a frozen fixture pending exact identity, configuration, and denominator joins.UNTESTED — assumption cannot establish a verdict
batch62-claude-sonnet-4-6-m3-r4
1M data enrichment records
model=claude-sonnet-4-6; provider=Anthropic; slug=claude-sonnet-4-6; scenario=1M data enrichment records; batch volume; input tokens; output tokens; interactive spend; batch spend; total savings; prompt/config/input/output/result/cache/checkpoint/artifact hashes=required; evidence=2026-09-07; measurement versus assumption=explicitUnavailable — 1M data enrichment records is a frozen fixture pending exact identity, configuration, and denominator joins.UNTESTED — assumption cannot establish a verdict
batch62-claude-sonnet-4-6-m3-r5
batch SLA timeout contingency
model=claude-sonnet-4-6; provider=Anthropic; slug=claude-sonnet-4-6; scenario=batch SLA timeout contingency; batch volume; input tokens; output tokens; interactive spend; batch spend; total savings; prompt/config/input/output/result/cache/checkpoint/artifact hashes=required; evidence=2026-09-07; measurement versus assumption=explicitUnavailable — batch SLA timeout contingency is a frozen fixture pending exact identity, configuration, and denominator joins.UNTESTED — assumption cannot establish a verdict
batch62-claude-sonnet-4-6-m3-r6
malformed batch row rejection
model=claude-sonnet-4-6; provider=Anthropic; slug=claude-sonnet-4-6; scenario=malformed batch row rejection; batch volume; input tokens; output tokens; interactive spend; batch spend; total savings; prompt/config/input/output/result/cache/checkpoint/artifact hashes=required; evidence=2026-09-07; measurement versus assumption=explicitUnavailable — malformed batch row rejection is a frozen fixture pending exact identity, configuration, and denominator joins.UNTESTED — assumption cannot establish a verdict

First-party provenance: Anthropic official API pricing; verification date 2026-09-07. Missing or conflicting joins fail closed.

Method and limitations: formulas are deterministic; observed and assumed inputs are labeled; no missing provider, host, account, region, realm, alias, snapshot, revision, weight, artifact, control, tool, modality, workload, rate-period, timestamp, or result is transferred. Run the Claude Sonnet 4.6 Batch 62 scenario →

Continuous SEO Builder · Batch 74 Audit · 2026-09-08Owner: claude-sonnet-4-6

Claude Sonnet 4.6 API Pricing: Proven 1M Context Coding Dominance

Claude Sonnet 4.6 costs $3.00 per million input tokens and $15.00 per million output tokens ($6.00/M blended at 3:1). A premier coding and analytical model offering an expansive 1M context window, exceptional multi-file repository navigation, and 90% prompt caching discounts. Verified 2026-09-08.

Module 1 · Claude Sonnet 4.6 Production Rate Card & Economics
Blended Cost = (Input Tokens × $3.00 + Output Tokens × $15.00) / 1,000,000

Claude Sonnet 4.6 delivers industry-standard software engineering precision at $6.00/M blended.

Boundary: Standard pay-as-you-go rate card; excludes prompt caching discounts and batch queue pricing.
ScenarioRendered Evidence & Bounds
Scenario 1Full repository feature implementation PR (32K in, 4K out): $0.15600 per generated PR
Scenario 2Multi-file bug isolation and test fix session (24K in, 3K out): $0.11700 per task
Scenario 3Software architectural review of microservice fleet (48K in, 5K out): $0.21900 per review
Scenario 4Automated pull request code review across 5 PRs (40K in, 3K out): $0.16500 per review set
Scenario 5Technical design document drafting (16K in, 3K out): $0.09300 per document
Scenario 6Monthly software engineering team tier (50M blended tokens): $300.00 infrastructure spend
Module 2 · Claude Sonnet 4.6 1M Context & Repository Caching Amortization
Cached Cost = (Cached Repo × $0.30 + Delta Input × $3.00 + Output × $15.00) / 1,000,000

Prompt caching transforms entire codebase context from a luxury into routine developer infrastructure.

Boundary: 90% discount on prompt prefixes >1,024 tokens held in Anthropic 5-minute rolling cache.
ScenarioRendered Evidence & Bounds
Scenario 1Monorepo codebase cache (150K tokens prefix, 5K delta): 83% input cost savings
Scenario 2Shared developer system prompt and coding standards (12K prefix): $0.00360 vs $0.03600 per call
Scenario 3Interactive IDE developer session (12 turns cached): 81% cumulative input savings
Scenario 4Cache write fee ($3.75/M) fully amortized after only 1.25 repeat requests within TTL
Scenario 5Zero latency prompt retrieval: cuts time-to-first-token by 52% during active coding sessions
Scenario 6Net engineering infrastructure spend reduced by over 66% for active developer teams
Module 3 · Claude Sonnet 4.6 to Sonnet 5 Upgrade Consideration Matrix
Upgrade Analysis = Benchmark Gains vs Price Delta ($4.00/M vs $6.00/M = 33% Savings)

Migrating from Sonnet 4.6 to Sonnet 5 cuts operational token spend by 33% with better accuracy.

Boundary: Evaluates migration economics from Sonnet 4.6 to next-generation Sonnet 5 offering 33% lower cost.
ScenarioRendered Evidence & Bounds
Scenario 1Sonnet 5 pricing ($2.00/M in, $10.00/M out): 33% cheaper across all token tiers
Scenario 2Sonnet 5 boosts SWE-bench verified benchmark scores by 4.8 points with lower retry overhead
Scenario 3Migrating 50M tokens/mo saves $100.00/mo ($200.00 vs $300.00) while improving accuracy
Scenario 4Drop-in API compatibility: seamless migration with zero prompt adjustments required
Scenario 5Test suite validation across 100 enterprise developer prompts passed with zero regressions
Scenario 6Recommended action: safe immediate migration to Sonnet 5 for higher quality at lower cost
Explore Related Analyses:Anthropic provider profileCompare vs Claude Sonnet 5Compare vs Claude Sonnet 4Best LLM for coding

How fast is Claude Sonnet 4.6?

Tokens / sec
76
TTFT
360 ms
Rank
#21 of 31
$ / M ÷ t/s
$0.08
Measured with 5 runs on a fixed prompt — see the full methodology.

How much does Claude Sonnet 4.6 cost at scale?

Tokens / monthEst. cost (blended 3:1)
100,000$0.60
1,000,000$6.00
10,000,000$60.00
100,000,000$600.00

How does Claude Sonnet 4.6 compare with other models?

Claude Haiku 4.5$2.00/MClaude Sonnet 5$4.00/MClaude Sonnet 4.5$6.00/MClaude Sonnet 4$6.00/MClaude Opus 4.8$10.00/MClaude Sonnet 4.5$6.00/MClaude Sonnet 4$6.00/MGPT-5.6 Terra$5.63/M
See all Anthropic models →

What is Claude Sonnet 4.6 best for?

#27 for Math & Reasoning#28 for Agents & Tool Use#32 for Image Understanding
Looking for a cheaper option?
Gemini 3.5 Flash Lite is 85.8% cheaper — a code-change migration. See all 8 alternatives to Claude Sonnet 4.6

Which Claude Sonnet 4.6 head-to-head comparisons are available?

Claude Sonnet 4.6 vs DeepSeek V4 ProClaude Sonnet 4.6 vs Claude Sonnet 4.5Claude Sonnet 4.6 vs Claude Sonnet 4Claude Sonnet 4.6 vs Claude Opus 4.8

What are common questions about Claude Sonnet 4.6?

Is Claude Sonnet 4.6 cheaper than Claude Sonnet 4.5?

Claude Sonnet 4.6 costs $6.00/M blended tokens, Claude Sonnet 4.5 costs $6.00/M — Claude Sonnet 4.5 is cheaper.

How much does 1 million tokens cost with Claude Sonnet 4.6?

At a 3:1 input:output ratio, 1 million blended tokens costs approximately $6.00. Pure input costs $3.00/M; pure output costs $15.00/M.

What does Claude Sonnet 4.6 cost at high volume?

At 100 million blended tokens a month, Claude Sonnet 4.6 costs approximately $600.00. See the cost-at-scale table below for other volumes.

Try Claude Sonnet 4.6 for free

Run real prompts against Claude Sonnet 4.6 and every other model on this page in one workspace.

Try Claude Sonnet 4.6 Free