Qwen3 8B API Pricing
Qwen3 8B is an Alibaba Qwen3 text model priced from the Singapore/International row of Model Studio. The verified rate is $0.18/M input and $0.7/M output. The Qwen3 launch blog lists Qwen3 8B as an Apache 2.0 dense model with 128K context.
Run the numbers.
Live calculator pre-loaded with the verified International Qwen3 8B rates. Use it for invoice checks, agent traces, and scenario planning before large Model Studio runs.
Real-world presets.
Support answer
Policy analysis
Function rewrite
Document summary
Paste text. See tokens. See cost.
Counts use a chars-per-token calibration measured on the vendor's own published tokenizer (Qwen/Qwen3.5-397B-A17B, 2026-06-10). English prose is typically within a few percent; code and non-Latin scripts tokenize heavier. For billing-exact counts use the vendor's count-tokens API.
| Model | Input /M | Output /M | Effective blended | Context | Best for |
|---|---|---|---|---|---|
| Qwen3 8B Current | $0.18 | $0.70 | $0.22 agentic 92/8 | 128K | Small open Qwen3 deployment |
| Qwen3 Next 80B A3B Thinking | $0.15 | $1.20 | $0.23 thinking | 262K | Qwen3 sibling |
| Qwen3 Next 80B A3B Instruct | $0.15 | $1.20 | $0.23 sibling | 262K | Qwen3 sibling |
| Qwen3 235B A22B Thinking 2507 | $0.23 | $2.30 | $0.40 thinking | 262K | Qwen3 sibling |
| Qwen3 235B A22B Instruct 2507 | $0.23 | $0.92 | $0.29 sibling | 262K | Qwen3 sibling |
| Qwen3 30B A3B Thinking 2507 | $0.20 | $2.40 | $0.38 thinking | 262K | Qwen3 sibling |
| Qwen3 30B A3B Instruct 2507 | $0.20 | $0.80 | $0.25 sibling | 262K | Qwen3 sibling |
| Qwen3 30B A3B | $0.20 | $0.80 | $0.25 sibling | 128K | Qwen3 sibling |
| Qwen3.7 Plus | $0.40 | $1.60 | $0.50 sibling | 1M | Current Plus tier |
| Qwen3.7 Max | $2.50 | $7.50 | $2.90 sibling | 1M | Current Max flagship |
Audit links
Frequently asked.
Short answers for teams comparing Qwen3 8B against other current Qwen3 text models.