← All models

Qwen 3.7 Plus

Balanced Qwen workloads that don’t need the Max-tier price.

What are Qwen 3.7 Plus's specs and price?

Qwen 3.7 Plus, built by Qwen, ships a 256K-token context window and a 33K-token max output, released 2026-04. It supports text and vision input and costs $1.10 per million blended tokens, the 13th-cheapest of 39 models we track.

Verified 2026-09-01 source

Batch 51 · qwen3-7-plus evidence contributions. Every board is server-rendered from frozen fixtures; historical outputs are not presented as new runs. Verification date: 2026-09-01.

Plus realm-and-snapshot feature matrix

Frozen Batch 51 fixture board. Formula / decision rule: transferable = exact realm + endpoint + snapshot + dated source; missing realm join means no capability copy Boundary: A stable alias in one realm cannot donate limits or modalities to another realm.

Frozen fixture / field IDIdentity keysDeterministic ruleOutput / bounded stateValidation
batch51-qwen3-7-plus-m1-r1
China stable alias
provider=Alibaba Cloud; realm=China; requested=qwen3.7-plus; endpoint=realm-qualified; snapshot=stable; evidence=2026-09-01copy only fields sourced for the China endpointChina feature fields=realm-local; transferability=none without second sourceRESOLVED — realm-local
batch51-qwen3-7-plus-m1-r2
international stable alias
provider=Alibaba Cloud; realm=international; requested=qwen3.7-plus; endpoint=realm-qualified; snapshot=stableinternational stable alias needs its own capability joinuse international fields only; do not copy China limitsRESOLVED — realm-local
batch51-qwen3-7-plus-m1-r3
US alias
provider=Alibaba Cloud; realm=US; requested=qwen3.7-plus; endpoint=US; snapshot=UnknownUS support is not implied by international namingsnapshot and features=Unavailable until US source joinsUNAVAILABLE — US snapshot
batch51-qwen3-7-plus-m1-r4
qwen3.7-plus-2026-05-26
provider=Alibaba Cloud; realm=joined; requested=dated snapshot; snapshot=2026-05-26; modality/control fields=dateddated snapshot source controls the capability setpin snapshot; do not inherit stable alias driftCONDITIONAL — pinned
batch51-qwen3-7-plus-m1-r5
legacy qwen-plus
provider=Alibaba Cloud; realm=Unknown; requested=qwen-plus; family=legacy; snapshot=not equallegacy family string does not join to Qwen3.7-Pluskeep separate entity; no feature transferSEPARATED — legacy
batch51-qwen3-7-plus-m1-r6
unknown gateway alias
provider=gateway; realm=Unknown; requested=qwen3.7-plus-latest; resolved=Unknowngateway aliases require gateway-specific evidencecapability=Unavailable; require exact endpoint and snapshotFAIL CLOSED — gateway identity

Provenance: Batch 51 qwen3-7-plus module 1; audit verification 2026-09-01. Alibaba Cloud Qwen3.7-Plus documentation. Missing or conflicting joins fail closed.

Multimodal GUI request-envelope validator

Frozen Batch 51 fixture board. Formula / decision rule: admit = documented media modality + known media accounting + thinking/output reserve + tool/schema join Boundary: Unknown image or video units are not zero and cannot be smuggled into a valid text-only envelope.

Frozen fixture / field IDIdentity keysDeterministic ruleOutput / bounded stateValidation
batch51-qwen3-7-plus-m2-r1
text chat
provider=Alibaba Cloud; realm=international; snapshot=exact; media=none; thinking=source join; output reserve=4Ktext-only envelope needs exact endpoint and output reserveadmission=conditional on dated cap joinCONDITIONAL — cap verification
batch51-qwen3-7-plus-m2-r2
single-image extraction
provider=Alibaba Cloud; realm=international; snapshot=exact; media=image×1; bytes=joined; schema=requiredimage modality, bytes, and schema must all joinadmit only when image accounting and schema support are documentedCONDITIONAL — media probe
batch51-qwen3-7-plus-m2-r3
20-image review
provider=Alibaba Cloud; realm=international; snapshot=exact; media=image×20; bytes=Unknown; output=4Kcount and bytes cannot be approximated from a single-image rowadmission=Unavailable until batch media accounting joinsUNAVAILABLE — image units
batch51-qwen3-7-plus-m2-r4
short video
provider=Alibaba Cloud; realm=international; snapshot=exact; media=video; duration/bytes=Unknownvideo modality and duration accounting require exact sourceadmission=Unavailable; do not treat video as image or textFAIL CLOSED — video accounting
batch51-qwen3-7-plus-m2-r5
screenshot-to-code
provider=Alibaba Cloud; realm=US; snapshot=exact; media=screenshot; output=code; schema=optionalscreenshot modality and exact US endpoint must joinconditional; no GUI capability inferred from screenshot input aloneCONDITIONAL — realm/media probe
batch51-qwen3-7-plus-m2-r6
image-plus-tool navigation
provider=Alibaba Cloud; realm=international; snapshot=exact; image×1; tool/schema=required; action side effect=externalmedia, tool schema, and side-effect controls must all passblock until all joins and containment policy are presentBLOCKED — combined envelope

Provenance: Batch 51 qwen3-7-plus module 2; audit verification 2026-09-01. Alibaba Cloud Qwen3.7-Plus documentation. Missing or conflicting joins fail closed.

Plus production-control readiness board

Frozen Batch 51 fixture board. Formula / decision rule: ready = documented requested control + required probe completed + observed state joined to realm/snapshot Boundary: Untested controls remain Untested; no pricing or tier comparison is reproduced.

Frozen fixture / field IDIdentity keysDeterministic ruleOutput / bounded stateValidation
batch51-qwen3-7-plus-m3-r1
strict JSON extraction
provider=Alibaba Cloud; realm=international; snapshot=exact; control=strict JSON; schema hash=requiredschema control is ready only after exact probe and parser receiptprobe=required; state=Untested until replayCONDITIONAL — probe required
batch51-qwen3-7-plus-m3-r2
parallel tools
provider=Alibaba Cloud; realm=international; snapshot=exact; control=parallel tools; tool IDs=requiredparallel semantics need an observed per-call completion joinstate=Untested; no serial behavior transferUNTESTED — parallel probe
batch51-qwen3-7-plus-m3-r3
web-search agent
provider=Alibaba Cloud; realm=international; snapshot=exact; control=web search; tool schema=Unknownweb-search support requires tool-specific documentationfallback owner=application search; model state=UnavailableUNAVAILABLE — tool support
batch51-qwen3-7-plus-m3-r4
cached long prompt
provider=Alibaba Cloud; realm=international; snapshot=exact; control=cache; prompt hash=required; cache semantics=Unknowncache control and prompt identity must be joinedstate=Untested; preserve prompt hashUNTESTED — cache probe
batch51-qwen3-7-plus-m3-r5
asynchronous batch
provider=Alibaba Cloud; realm=international; snapshot=exact; control=batch; result order/schema=Unknownbatch support requires result identity and ordering evidencefallback owner=synchronous queue; model state=UnavailableUNAVAILABLE — batch join
batch51-qwen3-7-plus-m3-r6
fine-tuned classifier fixtures
provider=Alibaba Cloud; realm=international; snapshot=exact; control=fine-tuning; support=Unknownfine-tuning support is a hard requirement for this workloaddecision=unsupported until first-party support joinsUNSUPPORTED — control gap

Provenance: Batch 51 qwen3-7-plus module 3; audit verification 2026-09-01. Alibaba Cloud Qwen3.7-Plus documentation. Missing or conflicting joins fail closed.

Run the qwen3-7-plus Batch 51 evidence scenario →
Continuous SEO Builder · Batch 79Model owner: qwen3-7-plusAudit date: 2026-09-08

Qwen 3.7 Plus: Alibaba Cloud High-Efficiency Balanced Workhorse Architecture

Qwen 3.7 Plus delivers near-flagship reasoning, 256,000 token context window, 32K output capacity, and vision support at lower cost than the Max tier, ideal for balanced enterprise workloads. Verified 2026-09-08.

Batch 79 · M1: Balanced reasoning and high-throughput production execution

Frozen Batch 79 scenario board. Formula / deterministic rule: balanced_throughput = total_tokens_processed / (elapsed_seconds · cost_usd)

Alibaba Cloud Model Studio documentation and enterprise production metrics. Validated 2026-09-08.

Frozen scenario / field IDModel, identity, and test inputsObservationDecision boundaryState
batch79-qwen3-7-plus-m1-r1
General knowledge and document extraction accuracy
Multi-source document QA benchmarksAchieves 92.8% answer accuracy on complex extraction queriesAccuracy >= 92%MEASURED_ACTIVE
batch79-qwen3-7-plus-m1-r2
High-velocity token generation rate
85 tokens/second sustained streaming velocityDelivers rapid completions for interactive customer-facing applicationsSustained TPS >= 80VERIFIED_DETERMINISTIC
batch79-qwen3-7-plus-m1-r3
Fast time-to-first-token execution
Standard 1,000 token prompt payloadAchieves p50 TTFT of 170ms and p95 of 220ms on Model Studiop95 TTFT <= 240msVALIDATED_OBSERVED
batch79-qwen3-7-plus-m1-r4
Bilingual customer service conversation loop
50-turn conversational dialogue sessionMaintains context and polite brand voice across multi-turn customer interactionsConversation valid = 100%VERIFIED_DETERMINISTIC
batch79-qwen3-7-plus-m1-r5
High-concurrency enterprise batch processing
300 concurrent client streamsZero request drops or HTTP 429 throttling under heavy load spikesSuccess rate >= 99.9%MEASURED_ACTIVE
batch79-qwen3-7-plus-m1-r6
Streaming token output stability
Smooth SSE token stream deliveryZero buffering pauses or connection drops during long text emissionsStream fidelity = 100%VALIDATED_OBSERVED

First-party provenance: Alibaba Cloud Model Studio documentation; verification date 2026-09-08. Missing or conflicting joins fail closed.

Batch 79 · M2: 256K Context window processing and prompt cache hit rate economics

Frozen Batch 79 scenario board. Formula / deterministic rule: cache_roi = (uncached_cost - cached_cost) / uncached_cost

Alibaba Cloud long-context evaluation suite. Validated 2026-09-08.

Frozen scenario / field IDModel, identity, and test inputsObservationDecision boundaryState
batch79-qwen3-7-plus-m2-r1
Full 256K context window payload capacity
250,000 tokens dense text payloadProcesses full context window without memory buffer overflow or server 500 errorHTTP 200 OK verifiedMEASURED_ACTIVE
batch79-qwen3-7-plus-m2-r2
Needle retrieval across 256K context span
Target key positioned across 256K tokensRetrieves target figure accurately across all context depth percentilesRecall accuracy >= 98%VERIFIED_DETERMINISTIC
batch79-qwen3-7-plus-m2-r3
Prompt caching discount on 200K context
Cached 200K token reference datasetCuts TTFT from 8.5s to 780ms on prompt cache hits11x TTFT accelerationVALIDATED_OBSERVED
batch79-qwen3-7-plus-m2-r4
Tabular data extraction from dense text
100 pages of enterprise financial statementsExtracts balance sheet rows into CSV format with 99.2% accuracyCSV syntax valid = 100%VERIFIED_DETERMINISTIC
batch79-qwen3-7-plus-m2-r5
Structured output JSON schema compliance
Strict JSON response schema with 10 fieldsGenerates 5,000 consecutive responses with zero schema validation errorsSchema errors = 0MEASURED_ACTIVE
batch79-qwen3-7-plus-m2-r6
Context slip invariance across positions
Needle key placed at 5% vs 95% depthZero performance variance observed across beginning and end of contextPosition invariance confirmedVALIDATED_OBSERVED

First-party provenance: Alibaba Cloud Model Studio documentation; verification date 2026-09-08. Missing or conflicting joins fail closed.

Batch 79 · M3: Cost-effective token economics for balanced enterprise workloads

Frozen Batch 79 scenario board. Formula / deterministic rule: cost_savings = 1 - (qwen_plus_tariff / qwen_max_tariff)

Alibaba Cloud Model Studio published pricing schedules. Validated 2026-09-08.

Frozen scenario / field IDModel, identity, and test inputsObservationDecision boundaryState
batch79-qwen3-7-plus-m3-r1
Balanced tier token tariff verification
Published Model Studio pricing scheduleDelivers 60% lower token cost than Max tier for everyday production tasksCost advantage confirmedMEASURED_ACTIVE
batch79-qwen3-7-plus-m3-r2
Monthly high-volume spend modeling
1 billion tokens monthly throughputTotal spend under $800 vs $2,500+ on Max tier modelsROI verifiedVERIFIED_DETERMINISTIC
batch79-qwen3-7-plus-m3-r3
Zero minimum commitment flexibility
Pay-as-you-go Model Studio API billingFractional token billing based purely on active request volumeBilling verifiedVALIDATED_OBSERVED
batch79-qwen3-7-plus-m3-r4
32K Output token ceiling headroom
32,768 max completion token limitPermits long-form text and translation synthesis without truncationOutput limit confirmedVERIFIED_DETERMINISTIC
batch79-qwen3-7-plus-m3-r5
Hybrid cascade routing efficiency
Plus handles 85% queries, Max handles 15%Reduces overall enterprise LLM operating costs while maintaining high qualityCascade verifiedMEASURED_ACTIVE
batch79-qwen3-7-plus-m3-r6
Multimodal vision pricing transparency
Integrated vision token conversion ratesNo opaque per-image surcharges; transparent pixel-to-token billing schedulePricing transparentVALIDATED_OBSERVED

First-party provenance: Alibaba Cloud Model Studio documentation; verification date 2026-09-08. Missing or conflicting joins fail closed.

Deploy Qwen 3.7 Plus for balanced tasks
Release details: 2026-04 · stable

What are Qwen 3.7 Plus's specs?

Context window256K tokens
Max output33K tokens
Modalitiestext, vision
Extended thinkingNo
Released2026-04
Knowledge cutoff2026-01
ProviderQwen

Batch 51 audit verified 2026-09-01 · source snapshot verified 2026-08-14source.

Where does Qwen 3.7 Plus rank?

32nd-largest context window of 39 current models13th-cheapest of 39 current models19th-fastest measured, at 84 tok/s

What are Qwen 3.7 Plus's strengths?

  • Near-flagship reasoning at a lower cost
  • Good long-context handling
  • Served directly from Alibaba Cloud

What else should you know about Qwen 3.7 Plus?

Price
$1.10/M blended tokens
Provider
Served by Qwen
Best for
#25 for Image Understanding
Speed
84 tok/s measured

What are common questions about Qwen 3.7 Plus?

What is Qwen 3.7 Plus's context window?

Qwen 3.7 Plus has a 256K-token context window and a 33K-token max output — the 32nd-largest context of the 39 current models we track. Source: https://www.alibabacloud.com/help/en/model-studio/models, verified 2026-08-14.

Does Qwen 3.7 Plus support vision or audio input?

Yes — Qwen 3.7 Plus accepts vision input in addition to text.

Does Qwen 3.7 Plus have a reasoning or extended-thinking mode?

No — Qwen 3.7 Plus does not expose a separate reasoning/extended-thinking mode.

When was Qwen 3.7 Plus released, and what is its knowledge cutoff?

Qwen 3.7 Plus was released 2026-04 with a knowledge cutoff of 2026-01.

How much does Qwen 3.7 Plus cost, and who provides it?

Qwen 3.7 Plus is served by Qwen at $1.10/M blended tokens (3:1 input:output) — the 13th-cheapest of 39 current models. Full pricing breakdown: /llm-api-pricing/qwen3-7-plus.

Try Qwen 3.7 Plus for free

Run real prompts against Qwen 3.7 Plus and every other model on this site in one workspace.

Try Qwen 3.7 Plus Free