Qwen3.8 Max API Pricing
Qwen3.8 Max is Alibaba's newest Qwen Max flagship, and as of August 2026 it finally has a price. The International / Singapore row lists $2/M input and $6/M output for 0-1M tokens, published directly in US dollars. That undercuts Qwen3.7 Max, the flagship it replaces, on both sides of the meter — the newer model is the cheaper one. Thinking and non-thinking modes bill at the same rate.
Run the numbers.
Calculator pre-loaded with the International Qwen3.8 Max rates. Alibaba publishes no per-model cache rate, so the cache slider stays off — what you see is the undiscounted cost.
Real-world presets.
Repo-wide feature build
Reading 150-page contracts
Support agent ticket triage
Research planning turn
Paste text. See tokens. See cost.
Counts use a chars-per-token calibration measured on the vendor's own published tokenizer (Qwen/Qwen3.5-397B-A17B, 2026-06-10). English prose is typically within a few percent; code and non-Latin scripts tokenize heavier. For billing-exact counts use the vendor's count-tokens API.
| Model | Input /M | Output /M | Effective blended | Context | Best for |
|---|---|---|---|---|---|
| Qwen3.8 Max Current | $2.00 | $6.00 | $2.32 agentic 92/8 | 1M | Newest Qwen flagship, now multimodal |
| Qwen3.7 Max | $2.50 | $7.50 | $2.90 pricier | 1M | Previous Max flagship, text only |
| Qwen3.7 Plus | $0.40 | $1.60 | $0.496 cheaper | 1M | Cheaper Qwen tier for bulk work |
| GPT-5.6 Terra | $2.00 cache $0.20 | $12.00 | $1.52 cheaper | 1.05M | Same input rate, real cache discount |
| Claude Sonnet 4.6 | $3.00 cache $0.30 | $15.00 | $2.05 cheaper | 1M | Long agentic coding sessions |
| DeepSeek V4 Pro | $1.32 cache $0.044 | $3.96 | $0.569 cheaper | 1M | Cheapest 1M reasoning tier |
Frequently asked.
Short answers for teams checking Qwen3.8 Max pricing, regional rate cards, and what changed when it left preview.
Q · 01 What is Qwen3.8 Max priced at? +
Q · 02 Why do I see $1.65 and $4.951 elsewhere? +
$1.65/$4.951 is the Global / China deployment card; $2/$6 is the International (Singapore) card. AI//COST quotes the International list for every Alibaba model, so the numbers stay comparable across our catalogue. Mixing the two is exactly what produced the pricing error we corrected on July 27, 2026.Q · 03 Did the price change when it left preview? +
qwen3.8-max-preview, reachable only through a Token Plan subscription in the Beijing region, with no per-token rate published anywhere. We ran the page with a blank price board rather than estimate one. The formal qwen3.8-max ID has now replaced it and carries the first real rate this model has ever had.Q · 04 Is it cheaper than Qwen3.7 Max? +
Q · 05 Is there a Batch discount? +
Q · 06 Is context caching priced? +
Q · 07 Which API model ID should I use? +
qwen3.8-max. The preview ID qwen3.8-max-preview has been withdrawn from Alibaba's rate cards. The model is served from Beijing, Singapore, Tokyo, Frankfurt and Virginia endpoints, and Alibaba exposes OpenAI-compatible, Anthropic-compatible and DashScope interfaces for it.