KAT-Coder-Pro V2.5 API Pricing
KAT-Coder-Pro V2.5 is Kwaipilot's flagship agentic coding model, served on Kuaishou's StreamLake API. It costs CNY 5/M input and CNY 20/M output ($0.70 and $2.82 at the project rate of 7.10), with cache reads at CNY 1/M and cache writes free. Note the direction of travel: this is 2.4x more expensive per token than the Pro V2 it follows.
Run the numbers.
Live calculator pre-loaded with current KAT-Coder-Pro V2.5 rates. Tweak the workload split, then share the URL to share the calculation.
Real-world presets.
Codebase-scale migration
Reading 100-page contracts
Ticket triage
Research planning turn
Paste text. See tokens. See cost.
This is a chars-per-token approximation, not a real tokenizer. Actual tokens vary by language, code density, and tool-call overhead — counts are typically ±10–20% off for English prose, more for code or non-Latin scripts. For exact billing, use the vendor's official tokenizer.
| Model | Input /M | Output /M | Effective blended | Context | Best for |
|---|---|---|---|---|---|
| KAT-Coder-Pro V2.5 Current | $0.70 cache $0.14 | $2.82 | $0.45 agentic 92/8 | 256K | Repo-scale agentic coding tasks |
| KAT-Coder-Air V2.5 | $0.14 cache $0.03 | $0.56 | $0.09 same family, fast tier | 256K | High-volume agent loops |
| KAT-Coder-Pro V2 | $0.30 cache $0.06 | $1.18 | $0.19 previous Pro, still listed | not listed | Same Pro line at the old rate |
| LongCat-2.0 | $0.70 cache $0.01 | $2.82 | $0.35 identical list rates, deeper cache | 1M | Agentic coding with 1M context |
| Grok Build 0.1 | $1.00 cache $0.20 | $2.00 | $0.48 coding specialist | 256K | Coding agents with cheap output |
| Kimi K2.7 Code | $0.95 cache $0.19 | $4.00 | $0.62 Chinese coding peer | 262K | Repo-scale Chinese coding workloads |
Frequently asked.
Practical KAT-Coder-Pro V2.5 pricing questions, including why this generation costs more than the last.
Q · 01 What does KAT-Coder-Pro V2.5 cost? +
CNY 5 per 1M input tokens, CNY 20 per 1M output tokens and CNY 1 per 1M cached input tokens, with cache writes free. At the project rate of 7.10 CNY/USD that is $0.704225, $2.816901 and $0.140845.Q · 02 Why is V2.5 more expensive than V2? +
CNY 2.1/8.4, while V2.5 is CNY 5/20 - a 2.4x rise per token. Both older rows are still listed and callable, so if V2.5's capability gain does not pay for itself on your workload, Pro V2 remains available at the old rate.Q · 03 How capable is it? +
Q · 04 What are the context and output limits? +
256K context window and 80K maximum output for both V2.5 models, with streaming, context caching, MCP and function calling supported across 20+ programming languages.Q · 05 Are cache writes really free? +
CNY 1/M, an 80% discount versus a miss. A miss still bills at the normal input rate, which is what our effective blend assumes.Q · 06 Where does the API live? +
Q · 07 How does it compare to other coding models? +
$0.45/M it sits within 6% of Grok Build 0.1 and 38% under Kimi K2.7 Code. LongCat-2.0 carries identical CNY 5/20 list rates but a far deeper cache discount, which puts it 21% cheaper on the same blend - and gives 1M context against this model's 256K.