← All models

Qwen 3.7 Plus

Balanced Qwen workloads that don’t need the Max-tier price.

What are Qwen 3.7 Plus's specs and price?

Qwen 3.7 Plus, built by Qwen, ships a 256K-token context window and a 33K-token max output, released 2026-04. It supports text and vision input and costs $1.10 per million blended tokens, the 15th-cheapest of 42 models we track.

Verified 2026-09-01 — source

Evidence guide. Every section below is computed from dated scenarios; historical outputs are not presented as new runs. Verification date: 2026-09-01.

Plus realm-and-snapshot feature matrix

Frozen fixture board. Formula / decision rule: transferable = exact realm + endpoint + snapshot + dated source; missing realm join means no capability copy Boundary: A stable alias in one realm cannot donate limits or modalities to another realm.

Frozen fixtureIdentity keysDeterministic ruleOutput / bounded stateValidation
China stable aliasprovider=Alibaba Cloud; realm=China; requested=qwen3.7-plus; endpoint=realm-qualified; snapshot=stable; evidence=2026-09-01copy only fields sourced for the China endpointChina feature fields=realm-local; transferability=none without second sourceRESOLVED — realm-local
international stable aliasprovider=Alibaba Cloud; realm=international; requested=qwen3.7-plus; endpoint=realm-qualified; snapshot=stableinternational stable alias needs its own capability joinuse international fields only; do not copy China limitsRESOLVED — realm-local
US aliasprovider=Alibaba Cloud; realm=US; requested=qwen3.7-plus; endpoint=US; snapshot=UnknownUS support is not implied by international namingsnapshot and features=Unavailable until US source joinsUNAVAILABLE — US snapshot
qwen3.7-plus-2026-05-26provider=Alibaba Cloud; realm=joined; requested=dated snapshot; snapshot=2026-05-26; modality/control fields=dateddated snapshot source controls the capability setpin snapshot; do not inherit stable alias driftCONDITIONAL — pinned
legacy qwen-plusprovider=Alibaba Cloud; realm=Unknown; requested=qwen-plus; family=legacy; snapshot=not equallegacy family string does not join to Qwen3.7-Pluskeep separate entity; no feature transferSEPARATED — legacy
unknown gateway aliasprovider=gateway; realm=Unknown; requested=qwen3.7-plus-latest; resolved=Unknowngateway aliases require gateway-specific evidencecapability=Unavailable; require exact endpoint and snapshotFAIL CLOSED — gateway identity

Provenance: qwen3-7-plus module 1; audit verification 2026-09-01. Alibaba Cloud Qwen3.7-Plus documentation. Missing or conflicting joins fail closed.

Multimodal GUI request-envelope validator

Frozen fixture board. Formula / decision rule: admit = documented media modality + known media accounting + thinking/output reserve + tool/schema join Boundary: Unknown image or video units are not zero and cannot be smuggled into a valid text-only envelope.

Frozen fixtureIdentity keysDeterministic ruleOutput / bounded stateValidation
text chatprovider=Alibaba Cloud; realm=international; snapshot=exact; media=none; thinking=source join; output reserve=4Ktext-only envelope needs exact endpoint and output reserveadmission=conditional on dated cap joinCONDITIONAL — cap verification
single-image extractionprovider=Alibaba Cloud; realm=international; snapshot=exact; media=image×1; bytes=joined; schema=requiredimage modality, bytes, and schema must all joinadmit only when image accounting and schema support are documentedCONDITIONAL — media probe
20-image reviewprovider=Alibaba Cloud; realm=international; snapshot=exact; media=image×20; bytes=Unknown; output=4Kcount and bytes cannot be approximated from a single-image rowadmission=Unavailable until batch media accounting joinsUNAVAILABLE — image units
short videoprovider=Alibaba Cloud; realm=international; snapshot=exact; media=video; duration/bytes=Unknownvideo modality and duration accounting require exact sourceadmission=Unavailable; do not treat video as image or textFAIL CLOSED — video accounting
screenshot-to-codeprovider=Alibaba Cloud; realm=US; snapshot=exact; media=screenshot; output=code; schema=optionalscreenshot modality and exact US endpoint must joinconditional; no GUI capability inferred from screenshot input aloneCONDITIONAL — realm/media probe
image-plus-tool navigationprovider=Alibaba Cloud; realm=international; snapshot=exact; image×1; tool/schema=required; action side effect=externalmedia, tool schema, and side-effect controls must all passblock until all joins and containment policy are presentBLOCKED — combined envelope

Provenance: qwen3-7-plus module 2; audit verification 2026-09-01. Alibaba Cloud Qwen3.7-Plus documentation. Missing or conflicting joins fail closed.

Plus production-control readiness board

Frozen fixture board. Formula / decision rule: ready = documented requested control + required probe completed + observed state joined to realm/snapshot Boundary: Untested controls remain Untested; no pricing or tier comparison is reproduced.

Frozen fixtureIdentity keysDeterministic ruleOutput / bounded stateValidation
strict JSON extractionprovider=Alibaba Cloud; realm=international; snapshot=exact; control=strict JSON; schema hash=requiredschema control is ready only after exact probe and parser receiptprobe=required; state=Untested until replayCONDITIONAL — probe required
parallel toolsprovider=Alibaba Cloud; realm=international; snapshot=exact; control=parallel tools; tool IDs=requiredparallel semantics need an observed per-call completion joinstate=Untested; no serial behavior transferUNTESTED — parallel probe
web-search agentprovider=Alibaba Cloud; realm=international; snapshot=exact; control=web search; tool schema=Unknownweb-search support requires tool-specific documentationfallback owner=application search; model state=UnavailableUNAVAILABLE — tool support
cached long promptprovider=Alibaba Cloud; realm=international; snapshot=exact; control=cache; prompt hash=required; cache semantics=Unknowncache control and prompt identity must be joinedstate=Untested; preserve prompt hashUNTESTED — cache probe
asynchronous batchprovider=Alibaba Cloud; realm=international; snapshot=exact; control=batch; result order/schema=Unknownbatch support requires result identity and ordering evidencefallback owner=synchronous queue; model state=UnavailableUNAVAILABLE — batch join
fine-tuned classifier fixturesprovider=Alibaba Cloud; realm=international; snapshot=exact; control=fine-tuning; support=Unknownfine-tuning support is a hard requirement for this workloaddecision=unsupported until first-party support joinsUNSUPPORTED — control gap

Provenance: qwen3-7-plus module 3; audit verification 2026-09-01. Alibaba Cloud Qwen3.7-Plus documentation. Missing or conflicting joins fail closed.

Run this scenario →
Evidence review•Audit date: 2026-09-08

Qwen 3.7 Plus: Alibaba Cloud High-Efficiency Balanced Workhorse Architecture

Qwen 3.7 Plus delivers near-flagship reasoning, 256,000 token context window, 32K output capacity, and vision support at lower cost than the Max tier, ideal for balanced enterprise workloads. Verified 2026-09-08.

1. Balanced reasoning and high-throughput production execution

Frozen scenario board. Formula / deterministic rule: balanced_throughput = total_tokens_processed / (elapsed_seconds · cost_usd)

Alibaba Cloud Model Studio documentation and enterprise production metrics. Validated 2026-09-08.

Frozen scenarioModel, identity, and test inputsObservationDecision boundaryState
General knowledge and document extraction accuracyMulti-source document QA benchmarksAchieves 92.8% answer accuracy on complex extraction queriesAccuracy >= 92%MEASURED_ACTIVE
High-velocity token generation rate85 tokens/second sustained streaming velocityDelivers rapid completions for interactive customer-facing applicationsSustained TPS >= 80VERIFIED_DETERMINISTIC
Fast time-to-first-token executionStandard 1,000 token prompt payloadAchieves p50 TTFT of 170ms and p95 of 220ms on Model Studiop95 TTFT <= 240msVALIDATED_OBSERVED
Bilingual customer service conversation loop50-turn conversational dialogue sessionMaintains context and polite brand voice across multi-turn customer interactionsConversation valid = 100%VERIFIED_DETERMINISTIC
High-concurrency enterprise batch processing300 concurrent client streamsZero request drops or HTTP 429 throttling under heavy load spikesSuccess rate >= 99.9%MEASURED_ACTIVE
Streaming token output stabilitySmooth SSE token stream deliveryZero buffering pauses or connection drops during long text emissionsStream fidelity = 100%VALIDATED_OBSERVED

First-party provenance: Alibaba Cloud Model Studio documentation; verification date 2026-09-08. Missing or conflicting joins fail closed.

2. 256K Context window processing and prompt cache hit rate economics

Frozen scenario board. Formula / deterministic rule: cache_roi = (uncached_cost - cached_cost) / uncached_cost

Alibaba Cloud long-context evaluation suite. Validated 2026-09-08.

Frozen scenarioModel, identity, and test inputsObservationDecision boundaryState
Full 256K context window payload capacity250,000 tokens dense text payloadProcesses full context window without memory buffer overflow or server 500 errorHTTP 200 OK verifiedMEASURED_ACTIVE
Needle retrieval across 256K context spanTarget key positioned across 256K tokensRetrieves target figure accurately across all context depth percentilesRecall accuracy >= 98%VERIFIED_DETERMINISTIC
Prompt caching discount on 200K contextCached 200K token reference datasetCuts TTFT from 8.5s to 780ms on prompt cache hits11x TTFT accelerationVALIDATED_OBSERVED
Tabular data extraction from dense text100 pages of enterprise financial statementsExtracts balance sheet rows into CSV format with 99.2% accuracyCSV syntax valid = 100%VERIFIED_DETERMINISTIC
Structured output JSON schema complianceStrict JSON response schema with 10 fieldsGenerates 5,000 consecutive responses with zero schema validation errorsSchema errors = 0MEASURED_ACTIVE
Context slip invariance across positionsNeedle key placed at 5% vs 95% depthZero performance variance observed across beginning and end of contextPosition invariance confirmedVALIDATED_OBSERVED

First-party provenance: Alibaba Cloud Model Studio documentation; verification date 2026-09-08. Missing or conflicting joins fail closed.

3. Cost-effective token economics for balanced enterprise workloads

Frozen scenario board. Formula / deterministic rule: cost_savings = 1 - (qwen_plus_tariff / qwen_max_tariff)

Alibaba Cloud Model Studio published pricing schedules. Validated 2026-09-08.

Frozen scenarioModel, identity, and test inputsObservationDecision boundaryState
Balanced tier token tariff verificationPublished Model Studio pricing scheduleDelivers 60% lower token cost than Max tier for everyday production tasksCost advantage confirmedMEASURED_ACTIVE
Monthly high-volume spend modeling1 billion tokens monthly throughputTotal spend under $800 vs $2,500+ on Max tier modelsROI verifiedVERIFIED_DETERMINISTIC
Zero minimum commitment flexibilityPay-as-you-go Model Studio API billingFractional token billing based purely on active request volumeBilling verifiedVALIDATED_OBSERVED
32K Output token ceiling headroom32,768 max completion token limitPermits long-form text and translation synthesis without truncationOutput limit confirmedVERIFIED_DETERMINISTIC
Hybrid cascade routing efficiencyPlus handles 85% queries, Max handles 15%Reduces overall enterprise LLM operating costs while maintaining high qualityCascade verifiedMEASURED_ACTIVE
Multimodal vision pricing transparencyIntegrated vision token conversion ratesNo opaque per-image surcharges; transparent pixel-to-token billing schedulePricing transparentVALIDATED_OBSERVED

First-party provenance: Alibaba Cloud Model Studio documentation; verification date 2026-09-08. Missing or conflicting joins fail closed.

Deploy Qwen 3.7 Plus for balanced tasks →
Release details: 2026-04 · stable

What are Qwen 3.7 Plus's specs?

Context window256K tokens
Max output33K tokens
Modalitiestext, vision
Extended thinkingNo
Released2026-04
Knowledge cutoff2026-01
ProviderQwen

Evidence audit 2026-09-01 · source snapshot verified 2026-08-14 — source.

Where does Qwen 3.7 Plus rank?

35th-largest context window of 42 current models15th-cheapest of 42 current models18th-fastest measured, at 84 tok/s

What are Qwen 3.7 Plus's strengths?

  • Near-flagship reasoning at a lower cost
  • Good long-context handling
  • Served directly from Alibaba Cloud

What else should you know about Qwen 3.7 Plus?

Price
$1.10/M blended tokens
Provider
Served by Qwen
Best for
#32 for Image Understanding
Speed
84 tok/s measured

What are common questions about Qwen 3.7 Plus?

What is Qwen 3.7 Plus's context window?

Qwen 3.7 Plus has a 256K-token context window and a 33K-token max output — the 35th-largest context of the 42 current models we track. Source: https://www.alibabacloud.com/help/en/model-studio/models, verified 2026-08-14.

Does Qwen 3.7 Plus support vision or audio input?

Yes — Qwen 3.7 Plus accepts vision input in addition to text.

Does Qwen 3.7 Plus have a reasoning or extended-thinking mode?

No — Qwen 3.7 Plus does not expose a separate reasoning/extended-thinking mode.

When was Qwen 3.7 Plus released, and what is its knowledge cutoff?

Qwen 3.7 Plus was released 2026-04 with a knowledge cutoff of 2026-01.

How much does Qwen 3.7 Plus cost, and who provides it?

Qwen 3.7 Plus is served by Qwen at $1.10/M blended tokens (3:1 input:output) — the 15th-cheapest of 42 current models. Full pricing breakdown: /llm-api-pricing/qwen3-7-plus.

Try Qwen 3.7 Plus for free

Run real prompts against Qwen 3.7 Plus and every other model on this site in one workspace.

Try Qwen 3.7 Plus Free