QWEN / ALIBABA CLOUD API PRICING

Alibaba Qwen token cost calculator.

Compare Qwen API rates while preserving region, input tier, cache mode, and context constraints.

Models
3
Rate cards
6
Input range
$0.028$1.65
Output range
$0.11$4.95
Cache-priced cards
6
Verified
2026-08-16
CACHE AMORTIZATION

Price writes and reads as an episode.

Model a stable reusable prefix, each cache creation, and the reads that actually land inside its lifetime. The counterfactual weights every billing category at its published rate.

CACHE EPISODE ECONOMICS
Qwen 3.7 Max$1.65 ordinary input · $2.06 cache write · $0.33 cache read / 1M5m TTL · output tokens excluded because they are identical in both arms
Without caching$19.80
With caching$10.52
Estimated reduction: $9.28 (46.9%)
BREAK-EVEN1 cache read per write

600 modeled requests across 100 cache episodes.

This is an input-cost counterfactual, not realized savings. It assumes the prefix stays byte-stable and every planned read hits within the selected TTL. Verify provider usage fields and accepted-answer quality.

Alibaba Cloud Qwen 3.7 Max pricing
QWEN / ALIBABA CLOUD WORKLOAD

Model the calls you actually make.

Enter monthly workload, tokens, and quality pass rates. The calculator resolves eligible price bands, then shows cost per accepted answer and the candidate break-even pass rate.

Estimated monthly API spend · Qwen / Alibaba Cloud
Qwen 3.7 MaxBefore · Global scope · up to 1M input: $1.65 input · $0.33 cache read · $4.95 output / 1MAfter · Global scope · up to 1M input: $1.65 input · $0.33 cache read · $4.95 output / 1MUS Virginia · Global deployment · price bands selected automatically from input size
Before$74.26
After$48.02
Potential saving: $26.24 (35.3%)
QUALITY-ADJUSTED COST
Before / accepted answer$0.008390% pass assumption
After / accepted answer$0.005685% pass assumption

Candidate break-even: 58.2% pass rate. Your scenario assumes 85%.

Warm cache-read scenario, not general ROI. It excludes cache writes/storage, tools, regional uplifts, retries, and quality failures. A dash means no published cache-read rate.

Alibaba Cloud Qwen 3.7 Max pricing
NEXT STEPValidate the 31.5% accepted-answer advantage.

The estimate is not a saving until the same workload still passes its quality bar. Compare a supported recipe in the lab, or use the evidence library to design a provider-specific test.

Qwen prices can depend on region, input size, and cache mode at the same time. TokenGauge retains those dimensions as distinct cards so a low tier is not silently applied to an ineligible request.

DATED RATE CARDS

Qwen / Alibaba Cloud prices with their scope intact.

Rates are USD per one million tokens. Future and transitional cards remain labeled with their effective dates instead of silently replacing today’s tier.

Showing 6 of 52 rate cards

Official provider API rate cards, USD per one million tokens
Provider / modelPrice scopeInput
USD / 1M
Cache read
USD / 1M
Output
USD / 1M
ContextSource
Qwen / Alibaba CloudQwen 3.7 Flashqwen3.7-flashBeijing · 256K–1M inputChina Beijing$0.165$0.033$0.661MOfficial ↗
Qwen / Alibaba CloudQwen 3.7 Flashqwen3.7-flashBeijing · 32K–256K inputChina Beijing$0.083$0.017$0.331MOfficial ↗
Qwen / Alibaba CloudQwen 3.7 Flashqwen3.7-flashBeijing · up to 32K inputChina Beijing$0.028$0.006$0.111MOfficial ↗
Qwen / Alibaba CloudQwen 3.7 Maxqwen3.7-maxGlobal scope · up to 1M inputUS Virginia · Global deployment$1.65$0.33$4.951MOfficial ↗
Qwen / Alibaba CloudQwen 3.7 Plusqwen3.7-plusGlobal scope · 256K–1M inputUS Virginia · Global deployment$0.826$0.166$3.301MOfficial ↗
Qwen / Alibaba CloudQwen 3.7 Plusqwen3.7-plusGlobal scope · up to 256K inputUS Virginia · Global deployment$0.276$0.056$1.101MOfficial ↗

Rates in USD per 1M tokens

Qwen / Alibaba Cloud

Qwen 3.7 Max

qwen3.7-max

Global scope · up to 1M inputUS Virginia · Global deployment

Input
$1.65
Cache read
$0.33
Output
$4.95
Context
1M
Open official source in a new tab
Qwen / Alibaba Cloud

Qwen 3.7 Plus

qwen3.7-plus

Global scope · 256K–1M inputUS Virginia · Global deployment

Input
$0.826
Cache read
$0.166
Output
$3.30
Context
1M
Open official source in a new tab
Qwen / Alibaba Cloud

Qwen 3.7 Plus

qwen3.7-plus

Global scope · up to 256K inputUS Virginia · Global deployment

Input
$0.276
Cache read
$0.056
Output
$1.10
Context
1M
Open official source in a new tab

USD per 1M tokens. Snapshot verified 2026-08-16. Cache writes, cache storage, tools, regions, and provider-specific thresholds may be billed separately.

A missing cache rate is shown as “—”, never treated as free. Consumer chat-plan quotas are not API prices.

MEASUREMENT RULE

Price accepted answers, not token deltas

Count retries, fallbacks, tool calls, latency failures, and answers rejected by the quality rubric. A lower estimated token bill is useful only when the workload still succeeds.

Open the controlled A/B lab
QUESTIONS

Before you trust the estimate.

How current are the Qwen / Alibaba Cloud API prices?

This page uses TokenGauge’s 2026-08-16 official-source snapshot. Follow the linked provider pages before making a production commitment because prices can change.

Does this calculator include every Qwen / Alibaba Cloud charge?

No. The workload calculator models input, output, and a warm cache-read share. The separate cache-episode calculator includes published cache writes and reads, but storage, tools, retries, regional uplifts, taxes, batch modes, and quality failures can still change the invoice.

Are consumer chat subscriptions included?

No. ChatGPT, Claude, Gemini, Grok, Kimi, and other consumer-plan quotas are separate from provider API token billing.