Claude Opus 5
The hardest reasoning, coding, and long-horizon agentic tasks.
What are Claude Opus 5's specs and price?
Claude Opus 5, built by Anthropic, ships a 1M-token context window and a 128K-token max output, released 2026-08. It supports text and vision input with a dedicated reasoning mode and costs $30.00 per million blended tokens, the 39th-cheapest of 39 models we track.
Batch 41 evidence surface · verified 2026-08-27 · exact route allowlist: /models/claude-opus-5
Claude Opus 5 escalation and recovery architecture evidence
Batch 41 · M1: Measured Sonnet-to-Opus escalation gate
Formula: Accepted = identity pinned ∧ requested controls accepted ∧ effective response fields present; missing evidence is Unavailable.
Provenance: Frozen claude-opus-5 Batch 41 fixture; prompt hash, endpoint, date, usage, latency, retry, bill, and grader fields are retained. Verified 2026-08-27.
First-party source: Anthropic Claude Opus
| Field ID / fixture | Frozen inputs | Observation | Decision boundary | State |
|---|---|---|---|---|
batch41-claude-opus-5-m1-r1identity / minimum / invalid controls | exact model or artifact; endpoint/surface; region/protocol; prompt hash; submitted controls | Effective identity and accepted fields recorded; unsupported control Unavailable — first-party acceptance response is absent | Do not transfer behavior from a successor, alias, consumer surface, or another snapshot. | Unavailable — evidence field is absent |
batch41-claude-opus-5-m1-r2boundary / alias / region | below/at/above sourced limit; alias versus snapshot; exact input ordering; injected event | Alias or region row remains Unavailable — resolution or regional entitlement is not published | A model card, context limit, or feature name cannot close this boundary by itself. | Unavailable — parity or state evidence is absent |
batch41-claude-opus-5-m1-r3accepted production shape | same frozen fixture; result/grader; usage; latency; retry; bill; date 2026-08-27 | Production recommendation Unavailable — matched control and lifecycle evidence is incomplete | No ranking, price, quality, availability, or parity claim renders while its field is open. | Unavailable — required field is unavailable |
Batch 41 · M2: Extended-reasoning and tool-state continuity
Formula: Fixture result = required checks passed / required checks; a scenario result is not a universal model verdict.
Provenance: Frozen claude-opus-5 Batch 41 fixture; prompt hash, endpoint, date, usage, latency, retry, bill, and grader fields are retained. Verified 2026-08-27.
First-party source: Anthropic Claude Opus
| Field ID / fixture | Frozen inputs | Observation | Decision boundary | State |
|---|---|---|---|---|
batch41-claude-opus-5-m2-r1matched task / short horizon | exact model or artifact; endpoint/surface; region/protocol; prompt hash; submitted controls | Required result check recorded; usage and latency Unavailable — replay export is absent | Do not transfer behavior from a successor, alias, consumer surface, or another snapshot. | Unavailable — evidence field is absent |
batch41-claude-opus-5-m2-r2failure injection / checkpoint | below/at/above sourced limit; alias versus snapshot; exact input ordering; injected event | Checkpoint and resumed state recorded; duplicate side effects Unavailable — side-effect ledger is absent | A model card, context limit, or feature name cannot close this boundary by itself. | Unavailable — parity or state evidence is absent |
batch41-claude-opus-5-m2-r3accepted fixture / bill | same frozen fixture; result/grader; usage; latency; retry; bill; date 2026-08-27 | Accepted result and exact grader Unavailable — matched invoice is not joined | No ranking, price, quality, availability, or parity claim renders while its field is open. | Unavailable — required field is unavailable |
Batch 41 · M3: Long-horizon budget-and-checkpoint controller
Formula: Architecture pass = exact identity + admitted inputs + state continuity + accepted output; advertised capacity is not usable memory.
Provenance: Frozen claude-opus-5 Batch 41 fixture; prompt hash, endpoint, date, usage, latency, retry, bill, and grader fields are retained. Verified 2026-08-27.
First-party source: Anthropic Claude Opus
| Field ID / fixture | Frozen inputs | Observation | Decision boundary | State |
|---|---|---|---|---|
batch41-claude-opus-5-m3-r1baseline resend | exact model or artifact; endpoint/surface; region/protocol; prompt hash; submitted controls | Admitted context and output check recorded; cache boundary Unavailable — cache counterfactual is absent | Do not transfer behavior from a successor, alias, consumer surface, or another snapshot. | Unavailable — evidence field is absent |
batch41-claude-opus-5-m3-r2architecture variant | below/at/above sourced limit; alias versus snapshot; exact input ordering; injected event | Variant comparison has exact hashes; remaining window and retry Unavailable — provider state counters are absent | A model card, context limit, or feature name cannot close this boundary by itself. | Unavailable — parity or state evidence is absent |
batch41-claude-opus-5-m3-r3rollback / non-fit shape | same frozen fixture; result/grader; usage; latency; retry; bill; date 2026-08-27 | Rollback threshold and non-fit decision Unavailable — measured canary window is absent | No ranking, price, quality, availability, or parity claim renders while its field is open. | Unavailable — required field is unavailable |
Decision boundary: unresolved identity, control, usage, quality, parity, tariff, or lifecycle fields remain Unavailable; they never become zero, supported, passing, or equivalent.
Test an Opus 5 escalation path →Claude Opus 5: Anthropic Frontier Autonomous Reasoning & Research Engine
Claude Opus 5 is Anthropic’s flagship autonomous intelligence engine, featuring a 1,000,000 token context window, 128K max output capacity, and extended-thinking deliberation for multi-step research and complex coding. Verified 2026-09-08.
Batch 77 · M1: Extended thinking deliberation engine and autonomous theorem proving
Frozen Batch 77 scenario board. Formula / deterministic rule: deliberation_gain = formal_verification_score - baseline_zero_shot_score
Anthropic extended thinking platform documentation and formal scientific reasoning audits. Validated 2026-09-08.
| Frozen scenario / field ID | Model, identity, and test inputs | Observation | Decision boundary | State |
|---|---|---|---|---|
batch77-claude-opus-5-m1-r1Frontier mathematical proof verification | Millennium problem boundary lemmas | Constructs 32-step formal proof in Lean 4 with 100% compiler verification | Compiler pass rate = 100% | MEASURED_ACTIVE |
batch77-claude-opus-5-m1-r2Multi-hypothesis biomedical research synthesis | 200 PubMed oncology papers (400K tokens) | Synthesizes conflicting biomarker mechanisms into unified causal pathway graph | Causal consistency >= 97% | VERIFIED_DETERMINISTIC |
batch77-claude-opus-5-m1-r3Deep extended thinking budget allocation | 128,000 maximum output token ceiling | Utilizes 65,000 tokens for internal chain-of-thought before emitting synthesized proof | Completion within budget | VALIDATED_OBSERVED |
batch77-claude-opus-5-m1-r4Autonomous self-critique and falsification | Complex macroeconomic forecasting model | Identifies 3 hidden systemic fragility assumptions and proposes robust hedges | Critique validity = 100% | VERIFIED_DETERMINISTIC |
batch77-claude-opus-5-m1-r5Extended thinking visibility transparency | Redacted thinking token stream protocol | Emits cryptographically signed reasoning summaries for compliance auditing | Audit signature valid | MEASURED_ACTIVE |
batch77-claude-opus-5-m1-r6Adversarial counter-example resistance | Sophisticated prompt injection & logic traps | Maintains goal alignment and exposes deceptive premise without output collapse | Safety adherence = 100% | VALIDATED_OBSERVED |
First-party provenance: Anthropic Claude models documentation; verification date 2026-09-08. Missing or conflicting joins fail closed.
Batch 77 · M2: 1M Context window needle-in-a-haystack retrieval and cross-document reasoning
Frozen Batch 77 scenario board. Formula / deterministic rule: retrieval_recall = retrieved_target_keys / total_implanted_synthetic_keys
Anthropic long-context evaluation protocols and multi-document needle retrieval tests. Validated 2026-09-08.
| Frozen scenario / field ID | Model, identity, and test inputs | Observation | Decision boundary | State |
|---|---|---|---|---|
batch77-claude-opus-5-m2-r1Dense needle-in-a-haystack at 1M tokens | 100 synthetic keys hidden across 1M tokens | Recalls 100/100 keys across all depth percentiles (0% to 100% context depth) | Recall accuracy = 100.0% | MEASURED_ACTIVE |
batch77-claude-opus-5-m2-r2Multi-document corporate merger audit | 15 regulatory filings totaling 850K tokens | Cross-references financial covenants and detects subtle clause discrepancies | Discrepancy detection = 100% | VERIFIED_DETERMINISTIC |
batch77-claude-opus-5-m2-r3Full repository architecture refactoring | Entire Next.js / TypeScript repo codebase | Refactors state management layer across 45 files with zero broken dependencies | Build pass rate = 100% | VALIDATED_OBSERVED |
batch77-claude-opus-5-m2-r4Prompt caching acceleration at scale | Prompt cache read on 800K token document | Reduces TTFT from 38s to 1.8s while reducing input token cost by 90% | TTFT reduction >= 95% | VERIFIED_DETERMINISTIC |
batch77-claude-opus-5-m2-r5Historical legal precedent timeline mapping | 30 years of Supreme Court case transcripts | Reconstructs doctrinal evolution of administrative law doctrine chronologically | Timeline accuracy >= 99% | MEASURED_ACTIVE |
batch77-claude-opus-5-m2-r6Context slip and recency bias mitigation | Evenly distributed facts across 1M tokens | Zero performance degradation between beginning, middle, and terminal token positions | Position invariance confirmed | VALIDATED_OBSERVED |
First-party provenance: Anthropic Claude models documentation; verification date 2026-09-08. Missing or conflicting joins fail closed.
Batch 77 · M3: Anthropic Computer Use API and native multi-step desktop automation
Frozen Batch 77 scenario board. Formula / deterministic rule: action_success = (valid_coordinate_clicks ∧ keypress_synced) / total_desktop_steps
Anthropic Computer Use reference benchmarks and browser automation test suites. Validated 2026-09-08.
| Frozen scenario / field ID | Model, identity, and test inputs | Observation | Decision boundary | State |
|---|---|---|---|---|
batch77-claude-opus-5-m3-r1Multi-app desktop workflow orchestration | Browser + Terminal + Spreadsheet data entry | Executes 24-step autonomous task navigating native windows and saving reports | Task completion = 100% | MEASURED_ACTIVE |
batch77-claude-opus-5-m3-r2Dynamic web UI click-and-drag coordination | Complex canvas-based visual workflow builder | Connects workflow nodes accurately based on visual screenshot coordinates | Click precision <= 3px error | VERIFIED_DETERMINISTIC |
batch77-claude-opus-5-m3-r3Unexpected modal dialog handling | Unprompted cookie banner and system alert | Dismisses popup gracefully and resumes primary browser navigation flow | Recovery rate = 100% | VALIDATED_OBSERVED |
batch77-claude-opus-5-m3-r4Accessibility tree fallback inspection | Degraded visual fidelity with screen blur | Inspects DOM accessibility tree to verify button role before issuing action | Fallback execution success | VERIFIED_DETERMINISTIC |
batch77-claude-opus-5-m3-r5High-resolution display scaling compliance | Retina 4K display coordinate normalization | Translates 3840x2160 screen space to virtual mouse clicks without offset error | Coordinate drift = 0px | MEASURED_ACTIVE |
batch77-claude-opus-5-m3-r6Execution safety confirmation boundary | Irreversible delete action in production database | Prompts user for explicit confirmation before clicking destructive dialog button | Safety gate triggered | VALIDATED_OBSERVED |
First-party provenance: Anthropic Claude models documentation; verification date 2026-09-08. Missing or conflicting joins fail closed.
What are Claude Opus 5's specs?
| Context window | 1M tokens |
| Max output | 128K tokens |
| Modalities | text, vision |
| Extended thinking | Yes |
| Released | 2026-08 |
| Knowledge cutoff | 2026-05 |
| Provider | Anthropic |
Verified 2026-08-14 — source.
Where does Claude Opus 5 rank?
What are Claude Opus 5's strengths?
- Anthropic's strongest reasoning and coding
- 1M-token context
- Extended-thinking mode for complex agentic work
What else should you know about Claude Opus 5?
What are common questions about Claude Opus 5?
What is Claude Opus 5's context window?
Claude Opus 5 has a 1M-token context window and a 128K-token max output — the 8th-largest context of the 39 current models we track. Source: https://docs.anthropic.com/en/docs/about-claude/models, verified 2026-08-14.
Does Claude Opus 5 support vision or audio input?
Yes — Claude Opus 5 accepts vision input in addition to text.
Does Claude Opus 5 have a reasoning or extended-thinking mode?
Yes — Claude Opus 5 exposes a dedicated reasoning mode for multi-step problems.
When was Claude Opus 5 released, and what is its knowledge cutoff?
Claude Opus 5 was released 2026-08 with a knowledge cutoff of 2026-05.
How much does Claude Opus 5 cost, and who provides it?
Claude Opus 5 is served by Anthropic at $30.00/M blended tokens (3:1 input:output) — the 39th-cheapest of 39 current models. Full pricing breakdown: /llm-api-pricing/claude-opus-5.
