LLM Cost Calculator

Qwen (Alibaba) API Pricing 2026

Alibaba's Qwen series via DashScope API. Qwen3.7 Max is the latest flagship; Qwen3 Coder 480B leads on coding tasks; Qwen3.5-Flash is among the cheapest capable models globally.

Pricing verified 2026-08-03. Sourced from www.alibabacloud.com/help/en/model-studio/model-pricing.

Get Qwen (Alibaba) API access →

Qwen (Alibaba) Model Pricing

Prices in USD per 1M tokens

ModelInput / 1MOutput / 1MContext
Qwen3.7 Plus
Qwen3.7 Plus; 1M context; latest Alibaba generation
$0.32$1.281,000,000
Qwen3.7 Max
Qwen3.7 Max; 1M context; top Alibaba flagship
$1.48$4.431,000,000
Qwen3.6 Flash
Qwen3.6 Flash; 1M context; cheap tier
$0.19$1.131,000,000
Qwen3.6 Plus
Qwen3.6 Plus; 1M context; mid-tier reasoning
$0.33$1.951,000,000
Qwen3.5-Flash
Ultra-cheap high-volume tier; one of the lowest prices among capable models
$0.12$0.45131,072
Qwen3.5-Plus
Balanced tier; strong multilingual & coding; 128K context
$0.26$0.78131,072
Qwen3 Coder Next
Qwen3 Coder Next; latest code-specialized model
$0.12$0.8262,144
Qwen3 Max Thinking
Qwen3 Max with extended thinking mode
$0.78$3.9262,144
Qwen3-Max
Alibaba flagship; hybrid thinking mode; rivals GPT-4.1 at ~1/5 the cost
$0.45$1.82131,072
Qwen3 Coder Plus
Qwen3 Coder Plus; 1M context; strong code generation
$0.65$3.251,000,000
Qwen3 Coder 480B A35B
Qwen3 Coder 480B A35B; flagship code model; 1M context
$0.3$11,048,576

How to read these estimates

The monthly table assumes 70% input tokens and 30% output tokens. It is a consistent comparison baseline, not a prediction of your workload. Actual bills can also include cached-input discounts, batch pricing, tool calls, minimum charges, regional differences, retries, and taxes.

The linked source is the provider's pricing documentation. Verify the current rate card and model availability before committing to a budget.

Estimated Monthly Cost (70% input / 30% output split)

Model1M tokens/mo10M tokens/mo100M tokens/mo1B tokens/mo
Qwen3.7 Plus$0.608$6.08$60.80$608
Qwen3.7 Max$2.37$23.65$236$2,365
Qwen3.6 Flash$0.472$4.72$47.20$472
Qwen3.6 Plus$0.816$8.16$81.60$816
Qwen3.5-Flash$0.219$2.19$21.90$219
Qwen3.5-Plus$0.416$4.16$41.60$416
Qwen3 Coder Next$0.324$3.24$32.40$324
Qwen3 Max Thinking$1.72$17.16$172$1,716
Qwen3-Max$0.861$8.61$86.10$861
Qwen3 Coder Plus$1.43$14.30$143$1,430
Qwen3 Coder 480B A35B$0.510$5.10$51.00$510

Frequently Asked Questions

How much does Qwen (Alibaba) LLM API cost?

Qwen (Alibaba) offers 11 models ranging from $0.120/1M to $1.48/1M input tokens. Alibaba's Qwen series via DashScope API. Qwen3.7 Max is the latest flagship; Qwen3 Coder 480B leads on coding tasks; Qwen3.5-Flash is among the cheapest capable models globally.

Is Qwen (Alibaba) cheaper than self-hosting?

For low-volume workloads (under 100M tokens/month), cloud APIs like Qwen (Alibaba) are almost always cheaper than purchasing and maintaining GPU hardware. Use our calculator to find the exact break-even point for your usage.