Qwen3 Coder Flash API Pricing
Qwen3 Coder Flash is Alibaba's code-specialized Qwen3 model for agentic software work. Alibaba publishes the International / Singapore 0-32K row directly in US dollars: $0.3/M input and $1.5/M output.
Run the numbers.
Live calculator pre-loaded with the official Qwen3 Coder Flash International/Singapore 0-32K token row, exactly as Alibaba publishes it in US dollars. Alibaba publishes higher prices for longer input bands up to 1M context.
Real-world presets.
Repo patch
Pull request review
Unit test drafting
Developer assistant
Paste text. See tokens. See cost.
Counts use a chars-per-token calibration measured on the vendor's own published tokenizer (Qwen/Qwen3.5-397B-A17B, 2026-06-10). English prose is typically within a few percent; code and non-Latin scripts tokenize heavier. For billing-exact counts use the vendor's count-tokens API.
| Model | Input /M | Output /M | Effective blended | Context | Best for |
|---|---|---|---|---|---|
| Qwen3 Coder Flash Current | $0.30 | $1.50 | $0.396 International 0-32K tier | 1M | High-volume coding assistants |
| Qwen3 Max | $1.20 | $6.00 | $1.58 pricier | 252K | Frontier Qwen proprietary reasoning |
| Qwen 3.5 Plus | $0.40 | $2.40 | $0.56 pricier | 256K | General Qwen production workloads |
| Qwen 3.5 Flash | $0.10 | $0.40 | $0.124 cheaper | 1M | Cheap long-context Qwen traffic |
| Qwen3 Coder Plus | $1.00 | $5.00 | $1.32 pricier | 1M | Agentic coding and code review |
| QwQ Plus | $0.80 | $2.40 | $0.928 pricier | 131K | Proprietary reasoning workloads |
| QwQ 32B | $0.287 | $0.861 | $0.333 cheaper | 131K | Open reasoning on a budget |
| Qwen3 235B A22B | $0.70 | $2.80 | $0.868 pricier | 131K | Open MoE reasoning baseline |
| GPT-5.4 mini | $0.75 cache $0.075 | $4.50 | $0.541 pricier | 400K | Coding and computer-use workloads |
| Gemini 2.5 Flash | $0.30 cache $0.03 | $2.50 | $0.272 cheaper | 1M | Low-latency multimodal RAG |
| DeepSeek V4 Flash | $0.44 cache $0.014 | $1.32 | $0.189 cheaper | 1M | Ultra-cheap API throughput |
Frequently asked.
Practical pricing questions, separated from calculator assumptions and regional tiers.
Q · 01 What is Qwen3 Coder Flash priced at? +
$0.3/M input and $1.5/M output on its International/Singapore 0-32K deployment row. That is a published US dollar list price, not a converted one.Q · 02 Does this page include higher context pricing tiers? +
$0.5/M input and $2.5/M output, 128K-256K at $0.8 and $4, and 256K-1M at $1.6 and $9.6.Q · 03 Is prompt caching priced separately? +
$0.3/M until Alibaba lists the exact cache-read price.Q · 04 How is the effective price calculated? +
$0.4/M.Q · 05 Is there a free quota or batch discount? +
Q · 06 Which API model ID should I use? +
qwen3-coder-flash. Alibaba states its current capability is equivalent to qwen3-coder-flash-2025-07-28.