Last verified
CODE BUDGET1M CONTEXTSINGAPORE BASELINELOW COST

Qwen3 Coder Flash API Pricing

Qwen3 Coder Flash is Alibaba's code-specialized Qwen3 model for agentic software work. Alibaba publishes the International / Singapore 0-32K row directly in US dollars: $0.3/M input and $1.5/M output.

Input - per 1M tokens
$0.30/M
International 0-32K input tier 0-32K
Output - per 1M tokens
$1.50/M
Output 0-32K input tier 0-32K
Cache not itemized
$0.30/M
Cache discount noted, price not itemized not itemized
Effective - agentic blend
$0.396/M
92/8 split - no cache discount
§ 01 / TERMINAL

Run the numbers.

Live calculator pre-loaded with the official Qwen3 Coder Flash International/Singapore 0-32K token row, exactly as Alibaba publishes it in US dollars. Alibaba publishes higher prices for longer input bands up to 1M context.

$ /mo
Workload split
Prompt cache hit rate
Tokens you can process
Words equivalent (English)
Effective rate
Open full calculator (all models · share URL · CSV) →
§ 02 / SCENARIOS

Real-world presets.

§ 03 / TOKENIZER

Paste text. See tokens. See cost.

Calibrated · measured on the vendor's tokenizer · 2026-06-10 Auto-counts as you type

Counts use a chars-per-token calibration measured on the vendor's own published tokenizer (Qwen/Qwen3.5-397B-A17B, 2026-06-10). English prose is typically within a few percent; code and non-Latin scripts tokenize heavier. For billing-exact counts use the vendor's count-tokens API.

Characters 386
Words 53
Tokens (estimated) 74 tokens
Cost as input · uncached $0.00002 USD
Cost as output · uncached $0.00011 USD
Cost as cached input $0.00002 USD
§ 04 / SHELF

Up against the shelf.

All models →
Model Input /M Output /M Effective blended Context Best for
Qwen3 Coder Flash Current $0.30 $1.50 $0.396 International 0-32K tier 1M High-volume coding assistants
Qwen3 Max $1.20 $6.00 $1.58 pricier 252K Frontier Qwen proprietary reasoning
Qwen 3.5 Plus $0.40 $2.40 $0.56 pricier 256K General Qwen production workloads
Qwen 3.5 Flash $0.10 $0.40 $0.124 cheaper 1M Cheap long-context Qwen traffic
Qwen3 Coder Plus $1.00 $5.00 $1.32 pricier 1M Agentic coding and code review
QwQ Plus $0.80 $2.40 $0.928 pricier 131K Proprietary reasoning workloads
QwQ 32B $0.287 $0.861 $0.333 cheaper 131K Open reasoning on a budget
Qwen3 235B A22B $0.70 $2.80 $0.868 pricier 131K Open MoE reasoning baseline
GPT-5.4 mini $0.75 cache $0.075 $4.50 $0.541 pricier 400K Coding and computer-use workloads
Gemini 2.5 Flash $0.30 cache $0.03 $2.50 $0.272 cheaper 1M Low-latency multimodal RAG
DeepSeek V4 Flash $0.44 cache $0.014 $1.32 $0.189 cheaper 1M Ultra-cheap API throughput

Frequently asked.

Practical pricing questions, separated from calculator assumptions and regional tiers.

Q · 01 What is Qwen3 Coder Flash priced at? +
Alibaba Cloud Model Studio lists Qwen3 Coder Flash at $0.3/M input and $1.5/M output on its International/Singapore 0-32K deployment row. That is a published US dollar list price, not a converted one.
Q · 02 Does this page include higher context pricing tiers? +
The quote tiles use the 0-32K International/Singapore tier. Alibaba also lists 32K-128K at $0.5/M input and $2.5/M output, 128K-256K at $0.8 and $4, and 256K-1M at $1.6 and $9.6.
Q · 03 Is prompt caching priced separately? +
Alibaba marks Qwen3 Coder Flash as eligible for Context Cache discount, but the pricing table does not publish a separate cache-read token amount. AI//COST therefore keeps cached input at $0.3/M until Alibaba lists the exact cache-read price.
Q · 04 How is the effective price calculated? +
AI//COST uses the same 92/8 agentic blend everywhere. With no exact cache-read token price published for this row, Qwen3 Coder Flash's effective blended cost is $0.4/M.
Q · 05 Is there a free quota or batch discount? +
Alibaba lists a 90-day activation free quota for many International Qwen rows, but not every open-source row includes one. Batch and context-cache support are model-specific; this page only publishes prices that are explicit in the vendor table.
Q · 06 Which API model ID should I use? +
Use the rolling model ID qwen3-coder-flash. Alibaba states its current capability is equivalent to qwen3-coder-flash-2025-07-28.