Last verified
QWEN3.6 FLASH1M CONTEXTTEXT + VISIONINTERNATIONAL PRICEUSD LIST PRICE

Qwen3.6 Flash API Pricing

Qwen3.6 Flash is an Alibaba Qwen3.6 model listed in the official International/Singapore pricing table. Alibaba publishes that table in US dollars, and the verified 0-256K row is $0.25/M input and $1.5/M output.

Input - per 1M tokens
$0.25/M
Source Alibaba Model Studio International
Output - per 1M tokens
$1.50/M
Long row $4/M output International
Cached input - not itemized
$0.25/M
Cache discount exists, exact row not exposed not listed
Effective - agentic blend
$0.35/M
92/8 split - no cache discount assumed
§ 01 / TERMINAL

Run the numbers.

Live calculator pre-loaded with Qwen3.6 Flash's verified Alibaba International base row. Alibaba marks context-cache discount support, but this page does not assume a cache-read token price because the exact International cache token row is not itemized in the public pricing table.

$ /mo
Workload split
Prompt cache hit rate
Tokens you can process
Words equivalent (English)
Effective rate
Open full calculator (all models · share URL · CSV) →
§ 02 / SCENARIOS

Real-world presets.

§ 03 / TOKENIZER

Paste text. See tokens. See cost.

Calibrated · measured on the vendor's tokenizer · 2026-06-10 Auto-counts as you type

Counts use a chars-per-token calibration measured on the vendor's own published tokenizer (Qwen/Qwen3.5-397B-A17B, 2026-06-10). English prose is typically within a few percent; code and non-Latin scripts tokenize heavier. For billing-exact counts use the vendor's count-tokens API.

Characters 364
Words 56
Tokens (estimated) 70 tokens
Cost as input · uncached $0.00002 USD
Cost as output · uncached $0.00011 USD
Cost as cached input $0.00002 USD
§ 04 / SHELF

Up against the shelf.

All models →
Model Input /M Output /M Effective blended Context Best for
Qwen3.6 Flash Current $0.25 cache $0.25 $1.50 $0.35 agentic 92/8 1M Fast multimodal Qwen agents
Qwen3.6 Flash Current $0.25 cache $0.25 $1.50 $0.35 agentic 92/8 1M Fast multimodal agents
Qwen3.6 35B A3B $0.375 cache $0.375 $2.25 $0.525 sibling 256K MoE reasoning and coding
Qwen3.6 27B $0.60 cache $0.60 $3.60 $0.84 sibling 256K Dense multimodal reasoning

Frequently asked.

Practical Qwen3.6 Flash pricing questions, with the base International row kept separate from the long-context band.

Q · 01 What is Qwen3.6 Flash priced at? +
Qwen3.6 Flash is listed at $0.25/M input and $1.5/M output in Alibaba Model Studio's International/Singapore table, for prompts up to 256K tokens. Alibaba publishes that table in US dollars.
Q · 02 What happens above 256K tokens? +
Alibaba moves the request to a higher band: the 256K-1M International row is $1/M input and $4/M output, four times the base rate on input.
Q · 03 Is prompt caching priced separately? +
The pricing table marks context-cache discount support for this model family, but the public International row does not expose an exact cache-read token rate. Cached input is therefore kept equal to fresh input instead of inventing a discount.
Q · 04 Which API model ID should I use? +
Use qwen3.6-flash unless Alibaba's model catalogue tells you to pin a dated snapshot. Re-check the official Model Studio page before production deployment.
Q · 05 What is this model best for? +
It is the lower-priced Qwen3.6 Flash tier for fast multimodal, coding, math, and agent workloads. The page uses the base 0-256K International row; prompts above 256K tokens bill at the higher $1/$4 band.