Qwen 3.8 Flash
Alibaba · released 2026 · proprietary API
- 1M context
- Multimodal
Alibaba's Flash tier for high-volume work, on the same multimodal API as the Max rows.
Context window
1M
1,000,000 tokens
Input price
$0.15
per 1M tokens
Output price
$0.47
per 1M tokens
Blended
$0.23
3:1 in:out mix
Weights
Proprietary
API only
Released
2026
Alibaba
Long-context / tier pricing — International pricing; the Global region bills $0.113/$0.382 for the same model. Batch inference is 50% off, and a context-caching discount applies separately. Listed prices are the base rate for prompts up to ~200K tokens.
Where it ranks
Of 40 models tracked.
Cost to run
#3(tied)
of 40
Input price
#3(tied)
of 40
Context window
#9(tied)
of 40
Compare with
The nearest models by blended cost to run.
Put it to work
Specs from BitByteCore's hand-verified Frontier Ledger (v2026-09-23), this row verified 22 Sep 2026 against Alibaba's official pricing ↗. Model pricing moves fast, so confirm with Alibaba before you rely on a figure.