QWEN3.6 MAX256K CONTEXTTEXT + CODEINTERNATIONAL PRICENO CACHE ROW
Qwen3.6 Max Preview API Pricing
Qwen3.6 Max Preview is an Alibaba Qwen3.6 text/coding model. Alibaba publishes the official Singapore/International row directly in US dollars: $1.3/M input and $7.8/M output for prompts up to 128K tokens.
Input - per 1M tokens
$1.30/M
Base International row standard
Output - per 1M tokens
$7.80/M
Long row $12/M output standard
Cached input - not itemized
$1.30/M
Cache no exact token price not listed
Effective - agentic blend
$1.82/M
92/8 split - no cache discount
§ 01 / TERMINAL
Run the numbers.
Live calculator pre-loaded with the verified Qwen3.6 Max Preview International base row. The 128K-256K long-context row is $2/M input and $12/M output.
$ /mo
Workload split
Prompt cache hit rate
Tokens you can process
—
Words equivalent (English)
—
Effective rate
—
§ 02 / SCENARIOS
Real-world presets.
CODING
Agent implementation
$0.083/task
RAG
Long-doc answer
$0.156/answer
CHATBOT
Product assistant
$0.014/turn
BATCH
Knowledge summary
$0.056/doc
§ 03 / TOKENIZER
Paste text. See tokens. See cost.
Calibrated · measured on the vendor's tokenizer · 2026-06-10 Auto-counts as you type
Counts use a chars-per-token calibration measured on the vendor's own published tokenizer (Qwen/Qwen3.5-397B-A17B, 2026-06-10). English prose is typically within a few percent; code and non-Latin scripts tokenize heavier. For billing-exact counts use the vendor's count-tokens API.
Characters 392
Words 64
Tokens (estimated) 75 tokens
Cost as input · uncached $0.0001 USD
Cost as output · uncached $0.00059 USD
Cost as cached input $0.0001 USD
| Model | Input /M | Output /M | Effective blended | Context | Best for |
|---|---|---|---|---|---|
| Qwen3.6 Max Preview Current | $1.30 | $7.80 | $1.82 agentic 92/8 | 256K | Preview Qwen Max workloads |
| Qwen3.7 Plus | $0.40 | $1.60 | $0.496 sibling | 1M | balanced |
| Qwen3.7 Max | $2.50 | $7.50 | $2.90 sibling | 1M | frontier |
| Qwen3 Next 80B A3B Instruct | $0.15 | $1.20 | $0.234 sibling | 262K | text |
| Qwen3 Next 80B A3B Thinking | $0.15 | $1.20 | $0.234 sibling | 262K | reasoning |
| Qwen3 30B A3B Instruct 2507 | $0.20 | $0.80 | $0.248 sibling | 262K | text |
| DeepSeek V4 Flash | $0.44 cache $0.014 | $1.32 | $0.189 sibling | 1M | budget |
| GPT-5.4 mini | $0.75 cache $0.075 | $4.50 | $0.541 sibling | 400K | light |
Audit links
Frequently asked.
Short answers for teams evaluating Qwen3.6 Max Preview against other Qwen tiers.
Q · 01 What is the Qwen3.6 Max Preview input price? +
Alibaba lists
$1.3/M for input tokens up to 128K in the Singapore/International row. That row is published in US dollars, so nothing is converted.Q · 02 What is the Qwen3.6 Max Preview output price? +
The base International output row is
$7.8/M. Output on the 128K-256K row is higher, at $12/M.Q · 03 Does Alibaba list exact cache pricing? +
No exact prompt-cache token price is itemized for this row, so the calculator keeps cache disabled instead of inventing a discount.
Q · 04 Why use International pricing? +
AI//COST uses Alibaba's Singapore/International row for non-mainland deployment. Mainland and other regional rows can differ, so the source URL is preserved for invoice audits.
Q · 05 Is this a text LLM page? +
Yes. This page covers the text/code API model. Media, audio, image, and video models are intentionally handled separately.