← All models

Qwen 3.7 Max

Frontier-class reasoning at a small discount to the newest Qwen flagship.

What are Qwen 3.7 Max's specs and price?

Qwen 3.7 Max, built by Qwen, ships a 256K-token context window and a 33K-token max output, released 2026-04. It supports text and vision input with a dedicated reasoning mode and costs $2.80 per million blended tokens, the 25th-cheapest of 39 models we track.

Verified 2026-09-01 source

Batch 51 · qwen3-7-max evidence contributions. Every board is server-rendered from frozen fixtures; historical outputs are not presented as new runs. Verification date: 2026-09-01.

Qwen3.7-Max alias-to-snapshot capability resolver

Frozen Batch 51 fixture board. Formula / decision rule: capability = exact realm + requested ID + resolved snapshot + dated source; unresolved joins do not inherit features Boundary: June multimodal support cannot be copied to a May text-only snapshot.

Frozen fixture / field IDIdentity keysDeterministic ruleOutput / bounded stateValidation
batch51-qwen3-7-max-m1-r1
stable alias
provider=Alibaba Cloud; realm=international; requested=qwen3.7-max; snapshot=stable alias target; evidence date=2026-09-01stable alias is usable only with its current dated target recordedpin target before relying on modalities or limitsCONDITIONAL — pin snapshot
batch51-qwen3-7-max-m1-r2
preview
provider=Alibaba Cloud; realm=international; requested=qwen3.7-max-preview; snapshot=preview; lifecycle=previewpreview identity is not interchangeable with stable aliaskeep preview provenance separate; no stable capability transferSEPARATED — preview
batch51-qwen3-7-max-m1-r3
2026-05-17
provider=Alibaba Cloud; realm=joined; requested=qwen3.7-max-2026-05-17; snapshot=2026-05-17; modality=media join=Unavailabledated snapshot fields must be sourced on that snapshotMay capability set=Unavailable where not documentedUNRESOLVED — snapshot field
batch51-qwen3-7-max-m1-r4
2026-05-20
provider=Alibaba Cloud; realm=joined; requested=qwen3.7-max-2026-05-20; snapshot=2026-05-20; text support=documentedtext support does not imply later multimodal supportretain text-only evidence; no June feature transferPASS WITH SEPARATION
batch51-qwen3-7-max-m1-r5
2026-06-08
provider=Alibaba Cloud; realm=joined; requested=qwen3.7-max-2026-06-08; snapshot=2026-06-08; cache/batch=source joinJune fields require exact June source and control joinuse only documented June fields; cache/batch=Unresolved if absentCONDITIONAL — dated join
batch51-qwen3-7-max-m1-r6
unknown gateway alias
provider=third-party gateway; realm=Unknown; requested=qwen3.7-max-latest; resolved snapshot=Unknown; host=not Alibaba directunknown host/snapshot fails identity joindo not transfer Max capability; require gateway evidenceFAIL CLOSED — unresolved alias

Provenance: Batch 51 qwen3-7-max module 1; audit verification 2026-09-01. Alibaba Cloud Qwen3.7-Max documentation. Missing or conflicting joins fail closed.

Max context-and-thinking budget compiler

Frozen Batch 51 fixture board. Formula / decision rule: reserve = documented context − known input − requested thinking − visible output; unknown component => Unavailable Boundary: Thinking, image, video, and tool units stay distinct; this is an admission receipt, not a price calculation.

Frozen fixture / field IDIdentity keysDeterministic ruleOutput / bounded stateValidation
batch51-qwen3-7-max-m2-r1
64K coding
provider=Alibaba Cloud; realm=joined; snapshot=exact; text input=64K; thinking=requested; visible output=4Ksubtract only documented token components from the snapshot capreserve=Unavailable until thinking cap join; no automatic zeroUNAVAILABLE — thinking accounting
batch51-qwen3-7-max-m2-r2
250K repository
provider=Alibaba Cloud; realm=joined; snapshot=exact; text=250K; tools=schema; output reserve=8Krepository text, tool/schema reserve, and output must join on one snapshotadmission=conditional on exact cap and tool reserve documentationCONDITIONAL — reserve join
batch51-qwen3-7-max-m2-r3
900K corpus
provider=Alibaba Cloud; realm=joined; snapshot=exact; text=900K; visible output=2K; thinking=Unknownunknown thinking allowance prevents a numerical remaining reserveremaining reserve=Unavailable; fail closedUNAVAILABLE — unknown thinking
batch51-qwen3-7-max-m2-r4
980K thinking request
provider=Alibaba Cloud; realm=joined; snapshot=exact; text=980K; requested thinking=980K; output=4Krequested thinking and visible output cannot exceed documented capadmission=Unavailable; do not assume hidden-budget behaviorFAIL CLOSED — over reserve
batch51-qwen3-7-max-m2-r5
image/video-plus-tools request
provider=Alibaba Cloud; realm=joined; snapshot=exact; image=1; video=1; tool/schema=yes; media units=Unknownmedia accounting and tool reserve each need snapshot-qualified documentationadmission=Unavailable; media units are not zeroUNAVAILABLE — multimodal accounting
batch51-qwen3-7-max-m2-r6
1.05M fixture
provider=Alibaba Cloud; realm=joined; snapshot=exact; text=1.05M; cap join=Unavailableover-cap input is not admitted without exact cap evidencereject or truncate only if provider documents it; current state=UnavailableREJECTED — cap not joined

Provenance: Batch 51 qwen3-7-max module 2; audit verification 2026-09-01. Alibaba Cloud Qwen3.7-Max documentation. Missing or conflicting joins fail closed.

Long-horizon agent interface gate

Frozen Batch 51 fixture board. Formula / decision rule: fit = every required modality/tool/schema/cache/batch/fine-tuning join is documented and locally replay-covered Boundary: This gate gives workload fit only; it does not issue a Max-versus-Plus or Qwen3.8 verdict.

Frozen fixture / field IDIdentity keysDeterministic ruleOutput / bounded stateValidation
batch51-qwen3-7-max-m3-r1
repository refactor
provider=Alibaba Cloud; realm=joined; snapshot=exact; required=text+tools; schema=required; replay=local coverageall required interface fields must join before an agent loop is appropriateappropriate only after tool/schema replay receipt is completeCONDITIONAL — replay gate
batch51-qwen3-7-max-m3-r2
office-document workflow
provider=Alibaba Cloud; realm=joined; snapshot=exact; required=document modality; tools=write; cache=batch=Unknowndocument modality and write containment must be exactsupport=Unavailable where modality or side-effect controls do not joinUNAVAILABLE — document/control join
batch51-qwen3-7-max-m3-r3
web-search research
provider=Alibaba Cloud; realm=joined; snapshot=exact; required=web-search tool; schema=requiredweb-search tool identity and result schema must be documentedappropriate only with exact tool documentation and replayCONDITIONAL — web tool probe
batch51-qwen3-7-max-m3-r4
GUI navigation
provider=Alibaba Cloud; realm=joined; snapshot=exact; required=screenshot+GUI actions; modality=Unknownvisual/action capability cannot be inferred from text supportunsupported until modality and action tool joinUNSUPPORTED — modality gap
batch51-qwen3-7-max-m3-r5
multi-tool coding loop
provider=Alibaba Cloud; realm=joined; snapshot=exact; tools=parallel; schema=required; cache=batch=Unknownparallel tool behavior and cache/batch state are separate joinsconditional; probe parallel calls and preserve cache/batch UnknownCONDITIONAL — control probes
batch51-qwen3-7-max-m3-r6
fine-tuning-dependent workload fixtures
provider=Alibaba Cloud; realm=joined; snapshot=exact; fine-tuning=required; support=Unknownfine-tuning support must be explicit for this workloadunsupported until exact fine-tuning documentation joinsUNSUPPORTED — capability gap

Provenance: Batch 51 qwen3-7-max module 3; audit verification 2026-09-01. Alibaba Cloud Qwen3.7-Max documentation. Missing or conflicting joins fail closed.

Run the qwen3-7-max Batch 51 evidence scenario →
Continuous SEO Builder · Batch 79Model owner: qwen3-7-maxAudit date: 2026-09-08

Qwen 3.7 Max: Alibaba Cloud Proven Frontier Reasoning at a Value Discount

Qwen 3.7 Max delivers previous-generation frontier reasoning, 256,000 token context window, and 32K output capacity at a slight discount to Qwen 3.8 Max. Verified 2026-09-08.

Batch 79 · M1: Proven frontier reasoning stability and long-context evaluation

Frozen Batch 79 scenario board. Formula / deterministic rule: reasoning_stability = consistent_benchmark_score / baseline_reference_score

Alibaba Cloud benchmark archives and production telemetry logs. Validated 2026-09-08.

Frozen scenario / field IDModel, identity, and test inputsObservationDecision boundaryState
batch79-qwen3-7-max-m1-r1
Competitive mathematical benchmark stability
GSM8K and MATH evaluation suitesMaintains 94.2% accuracy on complex multi-step math problemsAccuracy >= 94%MEASURED_ACTIVE
batch79-qwen3-7-max-m1-r2
Bilingual technical translation consistency
Industrial engineering specificationsTranslates technical equipment manuals with zero terminology ambiguityTerminology error = 0VERIFIED_DETERMINISTIC
batch79-qwen3-7-max-m1-r3
Python algorithmic script generation
Automated data pipeline extraction scriptsGenerates valid Pandas and NumPy code with correct memory vectorizationScript valid = 100%VALIDATED_OBSERVED
batch79-qwen3-7-max-m1-r4
Fast time-to-first-token responsiveness
Standard 1,000 token user promptAchieves p50 TTFT of 195ms and p95 of 260ms on Model Studiop95 TTFT <= 280msVERIFIED_DETERMINISTIC
batch79-qwen3-7-max-m1-r5
High-concurrency chat platform support
200 concurrent user sessionsMaintains 99.9% uptime with zero request drops during traffic peaksAvailability = 99.9%MEASURED_ACTIVE
batch79-qwen3-7-max-m1-r6
Streaming token velocity consistency
68 tokens/second sustained throughputSmooth text emission across extended conversational turnsSteady TPS >= 65VALIDATED_OBSERVED

First-party provenance: Alibaba Cloud Model Studio documentation; verification date 2026-09-08. Missing or conflicting joins fail closed.

Batch 79 · M2: 256K Context window document analysis and retrieval accuracy

Frozen Batch 79 scenario board. Formula / deterministic rule: needle_accuracy = correctly_retrieved_keys / total_implanted_keys

Alibaba Cloud long-context evaluation benchmarks. Validated 2026-09-08.

Frozen scenario / field IDModel, identity, and test inputsObservationDecision boundaryState
batch79-qwen3-7-max-m2-r1
256K Context window full payload capacity
245,000 tokens dense text payloadProcesses full context window without memory fault or connection dropPayload accepted = 100%MEASURED_ACTIVE
batch79-qwen3-7-max-m2-r2
Needle retrieval across 256K context span
Target key positioned across 256K tokensRetrieves target figure accurately across all context depth percentilesRecall accuracy >= 98%VERIFIED_DETERMINISTIC
batch79-qwen3-7-max-m2-r3
Corporate regulatory filing cross-examination
3 Chinese corporate annual reportsExtracts executive compensation and subsidiary equity holdingsExtraction completeVALIDATED_OBSERVED
batch79-qwen3-7-max-m2-r4
Prompt caching acceleration at scale
Cached 200K token reference datasetCuts TTFT from 10.5s to 980ms on prompt cache hits10x TTFT accelerationVERIFIED_DETERMINISTIC
batch79-qwen3-7-max-m2-r5
Structured JSON schema parsing adherence
Strict JSON response schema with 12 fieldsGenerates 2,000 consecutive responses with zero schema validation errorsSchema errors = 0MEASURED_ACTIVE
batch79-qwen3-7-max-m2-r6
Context slip invariance across positions
Needle key placed at 5% vs 95% depthZero performance variance observed across beginning and end of contextPosition invariance confirmedVALIDATED_OBSERVED

First-party provenance: Alibaba Cloud Model Studio documentation; verification date 2026-09-08. Missing or conflicting joins fail closed.

Batch 79 · M3: Value discount token economics relative to flagship tiers

Frozen Batch 79 scenario board. Formula / deterministic rule: discount_ratio = 1 - (qwen_37_tariff / qwen_38_tariff)

Alibaba Cloud Model Studio published pricing schedules. Validated 2026-09-08.

Frozen scenario / field IDModel, identity, and test inputsObservationDecision boundaryState
batch79-qwen3-7-max-m3-r1
Value discount pricing verification
Published Model Studio pricing scheduleProvides 20% discount relative to newest Qwen 3.8 Max flagshipDiscount verifiedMEASURED_ACTIVE
batch79-qwen3-7-max-m3-r2
Monthly high-volume spend comparison
500M tokens monthly throughputDelivers significant annual budget savings for established production workloadsSavings confirmedVERIFIED_DETERMINISTIC
batch79-qwen3-7-max-m3-r3
Zero minimum platform commitment elasticity
Pay-as-you-go Model Studio API billingFractional token billing with zero locked upfront platform feeBilling verifiedVALIDATED_OBSERVED
batch79-qwen3-7-max-m3-r4
32K Output token ceiling headroom
32,768 max completion token limitPermits long-form report and document synthesis without truncationOutput limit confirmedVERIFIED_DETERMINISTIC
batch79-qwen3-7-max-m3-r5
Production pipeline backward compatibility
Identical API endpoint request schemaDrop-in compatible with existing Qwen integrations without SDK code changesCompatibility verifiedMEASURED_ACTIVE
batch79-qwen3-7-max-m3-r6
Hybrid cascade deployment with Qwen 3.7 Plus
Plus handles standard extraction, Max handles reasoningBalances enterprise budget while retaining high accuracy on difficult queriesCascade verifiedVALIDATED_OBSERVED

First-party provenance: Alibaba Cloud Model Studio documentation; verification date 2026-09-08. Missing or conflicting joins fail closed.

Deploy Qwen 3.7 Max for proven reasoning
Release details: 2026-04 · stable

What are Qwen 3.7 Max's specs?

Context window256K tokens
Max output33K tokens
Modalitiestext, vision
Extended thinkingYes
Released2026-04
Knowledge cutoff2026-01
ProviderQwen

Batch 51 audit verified 2026-09-01 · source snapshot verified 2026-08-14source.

Where does Qwen 3.7 Max rank?

31st-largest context window of 39 current models25th-cheapest of 39 current models28th-fastest measured, at 49 tok/s

What are Qwen 3.7 Max's strengths?

  • Previous-gen Qwen flagship
  • Frontier-class reasoning and long context
  • Slightly cheaper than 3.8 Max

What else should you know about Qwen 3.7 Max?

Price
$2.80/M blended tokens
Provider
Served by Qwen
Best for
#24 for Math & Reasoning
Speed
49 tok/s measured

What are common questions about Qwen 3.7 Max?

What is Qwen 3.7 Max's context window?

Qwen 3.7 Max has a 256K-token context window and a 33K-token max output — the 31st-largest context of the 39 current models we track. Source: https://www.alibabacloud.com/help/en/model-studio/models, verified 2026-08-14.

Does Qwen 3.7 Max support vision or audio input?

Yes — Qwen 3.7 Max accepts vision input in addition to text.

Does Qwen 3.7 Max have a reasoning or extended-thinking mode?

Yes — Qwen 3.7 Max exposes a dedicated reasoning mode for multi-step problems.

When was Qwen 3.7 Max released, and what is its knowledge cutoff?

Qwen 3.7 Max was released 2026-04 with a knowledge cutoff of 2026-01.

How much does Qwen 3.7 Max cost, and who provides it?

Qwen 3.7 Max is served by Qwen at $2.80/M blended tokens (3:1 input:output) — the 25th-cheapest of 39 current models. Full pricing breakdown: /llm-api-pricing/qwen3-7-max.

Try Qwen 3.7 Max for free

Run real prompts against Qwen 3.7 Max and every other model on this site in one workspace.

Try Qwen 3.7 Max Free