Hunyuan Hy4 Preview API Pricing
Hy4 Preview is the largest model Tencent has ever put on a price list - a 770B total / 49B active MoE with a 1M-token context, open-sourced August 28, 2026. Tencent Cloud International lists it at $0.834/M input, $2.501/M output, and $0.042/M cache hit. Those are the vendor's own USD figures, not a conversion of the CNY card - Tencent publishes the two lists independently.
Run the numbers.
Live calculator pre-loaded with Hy4 Preview's Tencent Cloud International rate. The cache slider matters more here than on any other Hunyuan row - the cache hit is 95% below fresh input, so the blended cost swings by a factor of four across the three presets.
Real-world presets.
Codebase-scale migration
Reading 100-page contracts
Ticket triage
Research planning turn
Paste text. See tokens. See cost.
Counts use a chars-per-token calibration measured on the vendor's own published tokenizer (tencent/Hunyuan-A13B-Instruct, 2026-06-10). English prose is typically within a few percent; code and non-Latin scripts tokenize heavier. For billing-exact counts use the vendor's count-tokens API.
| Model | Input /M | Output /M | Effective blended | Context | Best for |
|---|---|---|---|---|---|
| Hunyuan Hy4 Preview Current | $0.834 cache $0.042 | $2.50 | $0.37 agentic 92/8 | 1M | Tencent's largest model - long-context agentic work |
| Hunyuan Hy3 | $0.132 cache $0.033 | $0.528 | $0.089 the GA flagship it sits above | 256K | Production traffic at a sixth of the input rate |
| Hunyuan A13B | $0.0744 | $0.298 | $0.0923 no cache rate published | 224K | Cheapest Tencent row still on the price list |
| DeepSeek V4 Pro | $1.32 cache $0.044 | $3.96 | $0.569 peak-hour rate | 1M | Chinese frontier reasoning at 1M context |
| GLM-5.3 | $1.40 cache $0.26 | $4.40 | $0.78 shallower cache discount | 1M | Coding-first frontier alternative |
| Kimi K3 | $3.00 cache $0.30 | $15.00 | $1.92 output-heavy pricing | 1M | Long-horizon agentic reasoning |
Frequently asked.
What the preview label actually costs you, why the rate went up rather than down, and where the USD figures come from.
Q · 01 What does Hy4 Preview cost? +
$0.834/M input, $2.501/M output, and $0.042/M on a cache hit, for the Singapore region. Tencent prints the same three figures in its own launch announcement, so the price list and the press release agree - which is not always true of a launch-week model.Q · 02 Why is the new model more expensive than Hy3? +
770B total parameters against Hy3's 295B, and the price moved with it: input is 6.3x the Hy3 rate and output 4.7x. A generational rise is unusual - most vendors cut on the successor - so read this as Tencent adding a tier above Hy3 rather than replacing it. Hy3 is still on the same price list at the same rate.Q · 03 Does the cache discount change the picture? +
$0.042/M, only 1.3x Hy3's $0.033/M, while fresh input is 6.3x. So on cache-heavy agentic traffic the two models are 4.2x apart on the blended rate rather than 6.3x. The 95% gap between fresh and cached input is the widest on any Hunyuan row - if your prompts do not reuse a prefix, this model gets expensive fast.Q · 04 How does it compare with other Chinese frontier models? +
$0.370/M against DeepSeek V4 Pro at $0.569 (peak-hour), GLM-5.3 at $0.780, and Kimi K3 at $1.923. Those are blended figures under one shared assumption, not benchmark scores - the shelf below shows the raw rates.Q · 05 What does the preview label mean for production use? +
preview status rather than GA. Practically: the rate is real and billable today, but Tencent has a recent habit of dating preview rows out - the previous Hy3 preview row was scheduled offline for August 31, 2026 while GA Hy3 continued unaffected. Plan a fallback to Hy3.Q · 06 Is it really open weight, and under what licence? +
Q · 07 Where do the USD figures come from? +
Y6 / Y18 / Y0.3), and converting it would give $0.845 / $2.535 - close, but wrong. Tencent applies no single exchange rate across the two lists: Hy4's USD implies about Y7.19/USD while Hy3's implies Y7.58/USD. Where a vendor publishes USD, we read it.