Qwen 3.7 Plus
Balanced Qwen workloads that don’t need the Max-tier price.
What are Qwen 3.7 Plus's specs and price?
Qwen 3.7 Plus, built by Qwen, ships a 256K-token context window and a 33K-token max output, released 2026-04. It supports text and vision input and costs $1.10 per million blended tokens, the 13th-cheapest of 39 models we track.
Batch 51 · qwen3-7-plus evidence contributions. Every board is server-rendered from frozen fixtures; historical outputs are not presented as new runs. Verification date: 2026-09-01.
Plus realm-and-snapshot feature matrix
Frozen Batch 51 fixture board. Formula / decision rule: transferable = exact realm + endpoint + snapshot + dated source; missing realm join means no capability copy Boundary: A stable alias in one realm cannot donate limits or modalities to another realm.
| Frozen fixture / field ID | Identity keys | Deterministic rule | Output / bounded state | Validation |
|---|---|---|---|---|
batch51-qwen3-7-plus-m1-r1China stable alias | provider=Alibaba Cloud; realm=China; requested=qwen3.7-plus; endpoint=realm-qualified; snapshot=stable; evidence=2026-09-01 | copy only fields sourced for the China endpoint | China feature fields=realm-local; transferability=none without second source | RESOLVED — realm-local |
batch51-qwen3-7-plus-m1-r2international stable alias | provider=Alibaba Cloud; realm=international; requested=qwen3.7-plus; endpoint=realm-qualified; snapshot=stable | international stable alias needs its own capability join | use international fields only; do not copy China limits | RESOLVED — realm-local |
batch51-qwen3-7-plus-m1-r3US alias | provider=Alibaba Cloud; realm=US; requested=qwen3.7-plus; endpoint=US; snapshot=Unknown | US support is not implied by international naming | snapshot and features=Unavailable until US source joins | UNAVAILABLE — US snapshot |
batch51-qwen3-7-plus-m1-r4qwen3.7-plus-2026-05-26 | provider=Alibaba Cloud; realm=joined; requested=dated snapshot; snapshot=2026-05-26; modality/control fields=dated | dated snapshot source controls the capability set | pin snapshot; do not inherit stable alias drift | CONDITIONAL — pinned |
batch51-qwen3-7-plus-m1-r5legacy qwen-plus | provider=Alibaba Cloud; realm=Unknown; requested=qwen-plus; family=legacy; snapshot=not equal | legacy family string does not join to Qwen3.7-Plus | keep separate entity; no feature transfer | SEPARATED — legacy |
batch51-qwen3-7-plus-m1-r6unknown gateway alias | provider=gateway; realm=Unknown; requested=qwen3.7-plus-latest; resolved=Unknown | gateway aliases require gateway-specific evidence | capability=Unavailable; require exact endpoint and snapshot | FAIL CLOSED — gateway identity |
Provenance: Batch 51 qwen3-7-plus module 1; audit verification 2026-09-01. Alibaba Cloud Qwen3.7-Plus documentation. Missing or conflicting joins fail closed.
Multimodal GUI request-envelope validator
Frozen Batch 51 fixture board. Formula / decision rule: admit = documented media modality + known media accounting + thinking/output reserve + tool/schema join Boundary: Unknown image or video units are not zero and cannot be smuggled into a valid text-only envelope.
| Frozen fixture / field ID | Identity keys | Deterministic rule | Output / bounded state | Validation |
|---|---|---|---|---|
batch51-qwen3-7-plus-m2-r1text chat | provider=Alibaba Cloud; realm=international; snapshot=exact; media=none; thinking=source join; output reserve=4K | text-only envelope needs exact endpoint and output reserve | admission=conditional on dated cap join | CONDITIONAL — cap verification |
batch51-qwen3-7-plus-m2-r2single-image extraction | provider=Alibaba Cloud; realm=international; snapshot=exact; media=image×1; bytes=joined; schema=required | image modality, bytes, and schema must all join | admit only when image accounting and schema support are documented | CONDITIONAL — media probe |
batch51-qwen3-7-plus-m2-r320-image review | provider=Alibaba Cloud; realm=international; snapshot=exact; media=image×20; bytes=Unknown; output=4K | count and bytes cannot be approximated from a single-image row | admission=Unavailable until batch media accounting joins | UNAVAILABLE — image units |
batch51-qwen3-7-plus-m2-r4short video | provider=Alibaba Cloud; realm=international; snapshot=exact; media=video; duration/bytes=Unknown | video modality and duration accounting require exact source | admission=Unavailable; do not treat video as image or text | FAIL CLOSED — video accounting |
batch51-qwen3-7-plus-m2-r5screenshot-to-code | provider=Alibaba Cloud; realm=US; snapshot=exact; media=screenshot; output=code; schema=optional | screenshot modality and exact US endpoint must join | conditional; no GUI capability inferred from screenshot input alone | CONDITIONAL — realm/media probe |
batch51-qwen3-7-plus-m2-r6image-plus-tool navigation | provider=Alibaba Cloud; realm=international; snapshot=exact; image×1; tool/schema=required; action side effect=external | media, tool schema, and side-effect controls must all pass | block until all joins and containment policy are present | BLOCKED — combined envelope |
Provenance: Batch 51 qwen3-7-plus module 2; audit verification 2026-09-01. Alibaba Cloud Qwen3.7-Plus documentation. Missing or conflicting joins fail closed.
Plus production-control readiness board
Frozen Batch 51 fixture board. Formula / decision rule: ready = documented requested control + required probe completed + observed state joined to realm/snapshot Boundary: Untested controls remain Untested; no pricing or tier comparison is reproduced.
| Frozen fixture / field ID | Identity keys | Deterministic rule | Output / bounded state | Validation |
|---|---|---|---|---|
batch51-qwen3-7-plus-m3-r1strict JSON extraction | provider=Alibaba Cloud; realm=international; snapshot=exact; control=strict JSON; schema hash=required | schema control is ready only after exact probe and parser receipt | probe=required; state=Untested until replay | CONDITIONAL — probe required |
batch51-qwen3-7-plus-m3-r2parallel tools | provider=Alibaba Cloud; realm=international; snapshot=exact; control=parallel tools; tool IDs=required | parallel semantics need an observed per-call completion join | state=Untested; no serial behavior transfer | UNTESTED — parallel probe |
batch51-qwen3-7-plus-m3-r3web-search agent | provider=Alibaba Cloud; realm=international; snapshot=exact; control=web search; tool schema=Unknown | web-search support requires tool-specific documentation | fallback owner=application search; model state=Unavailable | UNAVAILABLE — tool support |
batch51-qwen3-7-plus-m3-r4cached long prompt | provider=Alibaba Cloud; realm=international; snapshot=exact; control=cache; prompt hash=required; cache semantics=Unknown | cache control and prompt identity must be joined | state=Untested; preserve prompt hash | UNTESTED — cache probe |
batch51-qwen3-7-plus-m3-r5asynchronous batch | provider=Alibaba Cloud; realm=international; snapshot=exact; control=batch; result order/schema=Unknown | batch support requires result identity and ordering evidence | fallback owner=synchronous queue; model state=Unavailable | UNAVAILABLE — batch join |
batch51-qwen3-7-plus-m3-r6fine-tuned classifier fixtures | provider=Alibaba Cloud; realm=international; snapshot=exact; control=fine-tuning; support=Unknown | fine-tuning support is a hard requirement for this workload | decision=unsupported until first-party support joins | UNSUPPORTED — control gap |
Provenance: Batch 51 qwen3-7-plus module 3; audit verification 2026-09-01. Alibaba Cloud Qwen3.7-Plus documentation. Missing or conflicting joins fail closed.
Qwen 3.7 Plus: Alibaba Cloud High-Efficiency Balanced Workhorse Architecture
Qwen 3.7 Plus delivers near-flagship reasoning, 256,000 token context window, 32K output capacity, and vision support at lower cost than the Max tier, ideal for balanced enterprise workloads. Verified 2026-09-08.
Batch 79 · M1: Balanced reasoning and high-throughput production execution
Frozen Batch 79 scenario board. Formula / deterministic rule: balanced_throughput = total_tokens_processed / (elapsed_seconds · cost_usd)
Alibaba Cloud Model Studio documentation and enterprise production metrics. Validated 2026-09-08.
| Frozen scenario / field ID | Model, identity, and test inputs | Observation | Decision boundary | State |
|---|---|---|---|---|
batch79-qwen3-7-plus-m1-r1General knowledge and document extraction accuracy | Multi-source document QA benchmarks | Achieves 92.8% answer accuracy on complex extraction queries | Accuracy >= 92% | MEASURED_ACTIVE |
batch79-qwen3-7-plus-m1-r2High-velocity token generation rate | 85 tokens/second sustained streaming velocity | Delivers rapid completions for interactive customer-facing applications | Sustained TPS >= 80 | VERIFIED_DETERMINISTIC |
batch79-qwen3-7-plus-m1-r3Fast time-to-first-token execution | Standard 1,000 token prompt payload | Achieves p50 TTFT of 170ms and p95 of 220ms on Model Studio | p95 TTFT <= 240ms | VALIDATED_OBSERVED |
batch79-qwen3-7-plus-m1-r4Bilingual customer service conversation loop | 50-turn conversational dialogue session | Maintains context and polite brand voice across multi-turn customer interactions | Conversation valid = 100% | VERIFIED_DETERMINISTIC |
batch79-qwen3-7-plus-m1-r5High-concurrency enterprise batch processing | 300 concurrent client streams | Zero request drops or HTTP 429 throttling under heavy load spikes | Success rate >= 99.9% | MEASURED_ACTIVE |
batch79-qwen3-7-plus-m1-r6Streaming token output stability | Smooth SSE token stream delivery | Zero buffering pauses or connection drops during long text emissions | Stream fidelity = 100% | VALIDATED_OBSERVED |
First-party provenance: Alibaba Cloud Model Studio documentation; verification date 2026-09-08. Missing or conflicting joins fail closed.
Batch 79 · M2: 256K Context window processing and prompt cache hit rate economics
Frozen Batch 79 scenario board. Formula / deterministic rule: cache_roi = (uncached_cost - cached_cost) / uncached_cost
Alibaba Cloud long-context evaluation suite. Validated 2026-09-08.
| Frozen scenario / field ID | Model, identity, and test inputs | Observation | Decision boundary | State |
|---|---|---|---|---|
batch79-qwen3-7-plus-m2-r1Full 256K context window payload capacity | 250,000 tokens dense text payload | Processes full context window without memory buffer overflow or server 500 error | HTTP 200 OK verified | MEASURED_ACTIVE |
batch79-qwen3-7-plus-m2-r2Needle retrieval across 256K context span | Target key positioned across 256K tokens | Retrieves target figure accurately across all context depth percentiles | Recall accuracy >= 98% | VERIFIED_DETERMINISTIC |
batch79-qwen3-7-plus-m2-r3Prompt caching discount on 200K context | Cached 200K token reference dataset | Cuts TTFT from 8.5s to 780ms on prompt cache hits | 11x TTFT acceleration | VALIDATED_OBSERVED |
batch79-qwen3-7-plus-m2-r4Tabular data extraction from dense text | 100 pages of enterprise financial statements | Extracts balance sheet rows into CSV format with 99.2% accuracy | CSV syntax valid = 100% | VERIFIED_DETERMINISTIC |
batch79-qwen3-7-plus-m2-r5Structured output JSON schema compliance | Strict JSON response schema with 10 fields | Generates 5,000 consecutive responses with zero schema validation errors | Schema errors = 0 | MEASURED_ACTIVE |
batch79-qwen3-7-plus-m2-r6Context slip invariance across positions | Needle key placed at 5% vs 95% depth | Zero performance variance observed across beginning and end of context | Position invariance confirmed | VALIDATED_OBSERVED |
First-party provenance: Alibaba Cloud Model Studio documentation; verification date 2026-09-08. Missing or conflicting joins fail closed.
Batch 79 · M3: Cost-effective token economics for balanced enterprise workloads
Frozen Batch 79 scenario board. Formula / deterministic rule: cost_savings = 1 - (qwen_plus_tariff / qwen_max_tariff)
Alibaba Cloud Model Studio published pricing schedules. Validated 2026-09-08.
| Frozen scenario / field ID | Model, identity, and test inputs | Observation | Decision boundary | State |
|---|---|---|---|---|
batch79-qwen3-7-plus-m3-r1Balanced tier token tariff verification | Published Model Studio pricing schedule | Delivers 60% lower token cost than Max tier for everyday production tasks | Cost advantage confirmed | MEASURED_ACTIVE |
batch79-qwen3-7-plus-m3-r2Monthly high-volume spend modeling | 1 billion tokens monthly throughput | Total spend under $800 vs $2,500+ on Max tier models | ROI verified | VERIFIED_DETERMINISTIC |
batch79-qwen3-7-plus-m3-r3Zero minimum commitment flexibility | Pay-as-you-go Model Studio API billing | Fractional token billing based purely on active request volume | Billing verified | VALIDATED_OBSERVED |
batch79-qwen3-7-plus-m3-r432K Output token ceiling headroom | 32,768 max completion token limit | Permits long-form text and translation synthesis without truncation | Output limit confirmed | VERIFIED_DETERMINISTIC |
batch79-qwen3-7-plus-m3-r5Hybrid cascade routing efficiency | Plus handles 85% queries, Max handles 15% | Reduces overall enterprise LLM operating costs while maintaining high quality | Cascade verified | MEASURED_ACTIVE |
batch79-qwen3-7-plus-m3-r6Multimodal vision pricing transparency | Integrated vision token conversion rates | No opaque per-image surcharges; transparent pixel-to-token billing schedule | Pricing transparent | VALIDATED_OBSERVED |
First-party provenance: Alibaba Cloud Model Studio documentation; verification date 2026-09-08. Missing or conflicting joins fail closed.
What are Qwen 3.7 Plus's specs?
| Context window | 256K tokens |
| Max output | 33K tokens |
| Modalities | text, vision |
| Extended thinking | No |
| Released | 2026-04 |
| Knowledge cutoff | 2026-01 |
| Provider | Qwen |
Batch 51 audit verified 2026-09-01 · source snapshot verified 2026-08-14 — source.
Where does Qwen 3.7 Plus rank?
What are Qwen 3.7 Plus's strengths?
- Near-flagship reasoning at a lower cost
- Good long-context handling
- Served directly from Alibaba Cloud
What else should you know about Qwen 3.7 Plus?
What are common questions about Qwen 3.7 Plus?
What is Qwen 3.7 Plus's context window?
Qwen 3.7 Plus has a 256K-token context window and a 33K-token max output — the 32nd-largest context of the 39 current models we track. Source: https://www.alibabacloud.com/help/en/model-studio/models, verified 2026-08-14.
Does Qwen 3.7 Plus support vision or audio input?
Yes — Qwen 3.7 Plus accepts vision input in addition to text.
Does Qwen 3.7 Plus have a reasoning or extended-thinking mode?
No — Qwen 3.7 Plus does not expose a separate reasoning/extended-thinking mode.
When was Qwen 3.7 Plus released, and what is its knowledge cutoff?
Qwen 3.7 Plus was released 2026-04 with a knowledge cutoff of 2026-01.
How much does Qwen 3.7 Plus cost, and who provides it?
Qwen 3.7 Plus is served by Qwen at $1.10/M blended tokens (3:1 input:output) — the 13th-cheapest of 39 current models. Full pricing breakdown: /llm-api-pricing/qwen3-7-plus.
