Outside mainland China, it can be hard to track which Chinese models are worth testing, where to find official prices, and how they compare with familiar regional cloud AI services. This independent page maps DeepSeek, GLM and Kimi, with official links, pricing snapshots, and a practical referral access option for a broader model catalog.
Independent personal page. Not an official provider website.
Chinese LLMs have become highly competitive in reasoning, coding, long-context work, and agentic workflows. Their official per-token pricing can be unusually aggressive compared with many familiar cloud AI services.
For overseas developers, the issue is not only model quality; it is discovering, comparing, and accessing the right services. This page is a map, not a ranking.
Use this page as a starting point, not as financial or technical advice. All models, prices, and availability should be verified directly with the official providers and the referral service.
DeepSeek V4-Pro-0813 · DeepSeek V4-Flash-0731
Strong reasoning and coding models with excellent long-context capabilities and agentic workflow support. Official token pricing is among the most aggressive in the industry.
GLM-5.2
Positioned for long-horizon tasks and project-scale engineering. Officially described as designed for reasoning, agentic engineering, and coding workflows.
Kimi K3
Kimi's flagship model for long-horizon coding and end-to-end knowledge work. Native multimodal understanding with a 1M-token context window and industry-leading intelligence.
All three model families show strong reasoning performance. DeepSeek is widely discussed for chain-of-thought tasks; GLM-5.2 is positioned for deep analytical work; Kimi K3 always reasons, with configurable reasoning effort for long-horizon work.
DeepSeek V4-Pro-0813, GLM-5.2, and Kimi K3 are all positioned for coding and software engineering. Kimi K3 is a flagship model for long-horizon coding and agentic engineering workflows. Use cases include code generation, debugging, and multi-file refactoring.
DeepSeek models support long-context experiments. Kimi K3 features a 1M-token context window, suitable for long-horizon software engineering and knowledge work. GLM-5.2 is positioned for project-scale engineering context.
DeepSeek and Kimi both emphasize agentic workflow support. Kimi K3 is built for long-horizon agent tasks with tool-use and reasoning-effort controls. GLM-5.2 is described as designed for agentic engineering. These models can be used as part of tool-use, planning, and multi-step automation pipelines.
DeepSeek V4-Flash-0731 at $0.14/1M input (cache miss) enables cost-sensitive prototyping at scale. Kimi K3 with cache hit at $0.30/1M offers another path for iterative development. These price points can change how developers budget AI experiments.
Chinese frontier models are generally strong in Chinese and English. For Russian-language use cases, developers should test specific Russian-language performance with their own prompts and tasks.
Model capabilities and benchmarks change quickly. Use this page as a map, not as a final ranking.
For developers used to regional cloud AI pricing, the public token prices of some Chinese frontier models can look unusually aggressive.
| Model | Input / 1M | Cached input / 1M | Output / 1M | Source | Notes |
|---|---|---|---|---|---|
| DeepSeek V4-Flash-0731 | $0.14 (cache miss) $0.0028 (cache hit) |
$0.0028 | $0.28 | Official DeepSeek API pricing | Cache hit/miss pricing model |
| DeepSeek V4-Pro-0813 | $0.435 (cache miss) $0.003625 (cache hit) |
$0.003625 | $0.87 | Official DeepSeek API pricing | Cache hit/miss pricing model |
| GLM-5.2 | $1.40 | $0.26 | $4.40 | Official Z.AI developer pricing | Input / cached input / output model |
| Kimi K3 | $3.00 (cache miss) $0.30 (cache hit) |
$0.30 | $15.00 | Official Kimi API pricing | Cache hit/miss pricing model |
| YandexGPT Pro 5.1 Reference | ~$6.557 | ~$6.557 | ~$6.557 | Official Yandex Cloud AI Studio pricing | Approx. converted from official per-1,000-token synchronous-mode pricing |
These are official public pricing snapshots from the model/platform providers. They are not the prices of the referral service. Official pricing can change.
YandexGPT Pro 5.1 is included only as a regional pricing reference point, not as a strict quality-equivalent comparison.
Despite low official token pricing, developers outside mainland China may still prefer a third-party token gateway for practical reasons:
If you want to explore a broader model catalog through one third-party token gateway, my referral link points to a service I personally track. According to the provider page, some token prices may be around 4% lower than the corresponding official Chinese API pricing, and the catalog may include additional model families such as Claude Fable, GPT-5.5, Seedance 2.0 and more. Please verify the current model list, pricing, billing rules and terms on the provider page before topping up.
Catalog examples may change. Check the provider page before topping up.
This page contains referral/affiliate links. I may receive a bonus or commission if you register or top up through them. I am not affiliated with DeepSeek, Z.AI, Moonshot AI, Kimi, Yandex, or the third-party gateway unless explicitly stated. Always verify pricing, model availability, billing rules, data handling terms and service terms on the official/provider pages.