Qwen3.8-Max and Qwen3.8-Flash, Qwen3.7/3.6/3.5 families, Qwen3-VL and Coder lines, plus Qwen2.5 models offering top-tier price-performance.
13 models tracked · cheapest input: Qwen 2.5 Turbo at $0.05/M · official pricing page
Alibaba's Qwen 3.8 family (Max, Flash, Coder) offers top-tier price-performance — especially for Asian-language and coding workloads — with open-weight options for self-hosting.
qwen3.8-max
Input /M
$2.00
Output /M
$6.00
Context
256K
Alibaba's 2.4-trillion-parameter MoE flagship rivaling GPT-5.6 and Claude Sonnet 5 across international reasoning benchmarks at $2.00/$6.00.
Input /M
$0.16
Output /M
$0.47
Context
1M
Alibaba's multimodal 125B MoE preview with 51B N-gram embeddings, 6B active parameters, and a Qwen4-architecture preview.
qwq-32b-preview
Input /M
$0.40
Output /M
$1.20
Context
128K
Alibaba's 'Qwen with Questions' dedicated chain-of-thought reasoning model rivaling OpenAI o1 on math and code at $0.40/$1.20.
qwen3-32b-instruct
Input /M
$0.40
Output /M
$0.80
Context
128K
Next-gen dense 32B model offering flagship-grade benchmark scores at budget operating costs ($0.40/$0.80).
Input /M
$1.60
Output /M
$6.40
Context
32.8K
Alibaba's established MoE model in reasoning, mathematics, and multilingual understanding.
qwen-2.5-coder-32b-instruct
Input /M
$0.20
Output /M
$0.60
Context
128K
Top-ranked open-weights coding model outperforming 70B models in code generation, repair, and FIM autocompletion at $0.20/$0.60.
qwen-2.5-72b-instruct
Input /M
$0.35
Output /M
$0.70
Context
128K
General-purpose powerhouse open-weights model with 128K context and top-tier math and coding benchmarks.
qwen-turbo
Input /M
$0.05
Output /M
$0.20
Context
1M
Ultra-fast million-token context model priced at $0.05/$0.20 per 1M tokens for bulk indexing and translation.
qwen3.5-397b-a17b
Input /M
$0.39
Output /M
$2.34
Context
262.1K
Alibaba Qwen natively multimodal reasoning route with a 262K context window.
qwen3-coder-next
Input /M
$0.12
Output /M
$0.80
Context
262.1K
Qwen coding route for agentic software work with a 262K context window.
qwen3-coder-flash
Input /M
$0.195
Output /M
$0.975
Context
1M
Fast Qwen coding route with a 1M-token context window for high-volume agent work.
qwen3-vl-235b-a22b-thinking
Input /M
$0.40
Output /M
$4.00
Context
131.1K
Qwen multimodal reasoning route for visual and long-context tasks.
Input /M
$0.42
Output /M
$3.00
Context
1M
Qwen 3.8 27B open-weight dense vision-language model: flexible thinking, coding, professional workflows and long-running agent tasks.
Other providers
All Alibaba Qwen prices verified Aug 28, 2026. Prices change frequently — each model page links to the authoritative source.