A developer-focused shortlist of Chinese language model products with API or deployment paths worth checking.
Quick answers
At a glance
What it covers
A developer-focused shortlist of Chinese language model products with API or deployment paths worth checking.
Matched tools
34 tools currently match this use case.
How to read this page
Prioritize products with public or limited API access, pricing clarity, source links and global access notes.
Decision standard
Prioritize products with public or limited API access, pricing clarity, source links and global access notes.
34 matched tools
DeepSeek
DeepSeek
4.7
The English docs now make DeepSeek's current API shape explicit: V4-Pro/V4-Flash, dual OpenAI/Anthropic compatibility, thinking mode, tool calls, caching and agent-tool integrations.
Best fit · Developers evaluating Chinese model APIs for coding agents, reasoning, tool use and OpenAI/Anthropic-compatible migration.
Coverage · 100/100
Globally availableFull English UITrustedPublic APIPaid
Payment
Topped-up balance / Granted balance
Checked
Aug 4
Sources
High confidence
From
V4-Flash from $0.0028/M cache-hit input, $0.14/M cache-miss input and $0.28/M output tokens
OpenBMB belongs in LLM & API as a Chinese open-source foundation-model ecosystem, with MiniCPM and MiniCPM-V as the main product signals tracked here.
Best fit · Researchers and engineering teams evaluating Chinese open-source models for edge inference, multimodal understanding, RAG and agent experiments.
Coverage · 100/100
Partially availablePartial English UITrustedLimited APIFree
AtomCode belongs in AI Coding because it is a full terminal coding agent and Claude Code alternative, not a generic chat assistant or raw model API.
Best fit · Developers who want an open-source terminal coding agent for Chinese and global model APIs, especially DeepSeek, GLM, Qwen and local Ollama workflows.
Coverage · 100/100
Partially availablePartial English UITrustedLimited APIFreemium
Payment
CodingPlan FREE / Manual API key
Checked
Aug 4
Sources
High confidence
From
Open-source MIT project; CodingPlan FREE is advertised; manual API key configuration is available
M3 should be tracked separately because the official homepage now gives it a distinct model page and positions it as the current frontier coding model.
Best fit · Developers evaluating a China-origin frontier model for coding, long-context agentic work and multimodal reasoning.
Coverage · 100/100
Globally availableFull English UITrustedPublic APIFreemium
Payment
API & Token Plan / Pay-as-you-go API billing
Checked
Aug 4
Sources
High confidence
From
API & Token Plan access; current pricing should be checked in account
Qwen-AgentWorld is tracked separately from Qwen Code because its public surface is a language world model plus AgentWorldBench, not a terminal coding product.
Best fit · Researchers and agent builders who need a simulator or benchmark for tool, terminal, SWE, Android, web and OS agent environments.
Coverage · 100/100
Globally availableFull English UITrustedPublic APIFree
Payment
Hugging Face download / ModelScope download
Checked
Aug 4
Sources
High confidence
From
Apache-2.0 open weights and benchmark; self-hosted inference infrastructure required
JoyAI-VL-Interaction deserves a separate profile because it releases a full real-time interaction model and system stack, not just a static VLM checkpoint.
Best fit · Researchers and product teams exploring proactive webcam, livestream, monitoring, cooking guidance, live commentary and event-driven visual interaction systems.
Coverage · 100/100
Globally availableFull English UITrustedPublic APIFree
Kimi's developer path is separated through K2.7 Code, K2.6 and the English API platform, so it should be distinct from the consumer assistant and Kimi Code CLI.
Best fit · Developers evaluating Chinese long-context, coding and agentic models through an international API.
Coverage · 100/100
Globally availableFull English UITrustedPublic APIPaid
Payment
Pay-as-you-go API billing / Platform billing
Checked
Aug 4
Sources
High confidence
From
Pay-as-you-go API pricing varies by Kimi K2 and Moonshot model
Kimi K3 merits a separate LLM & API profile because it is a flagship model with its own architecture, context length, hosted pricing and public open-weight release, rather than a minor K2.7 Code revision.
Best fit · Teams evaluating a frontier-scale Chinese model for repository-scale coding, multimodal technical work, agentic research and long-context knowledge workflows.
Coverage · 100/100
Globally availableFull English UITrustedPublic APIPaid
Payment
Kimi API pay-as-you-go billing / Kimi membership
Checked
Aug 4
Sources
High confidence
From
$0.30/M cache-hit input, $3/M cache-miss input and $15/M output tokens
K2.7 Code has its own open weights, architecture, benchmarks, API pricing and Kimi Code default-model behavior, so it deserves a dedicated model profile separate from Kimi Code and the broader Kimi API.
Best fit · Teams evaluating open-weight coding models for repository-scale agents, Kimi Code workflows or OpenAI-compatible API integration.
Coverage · 100/100
Globally availableFull English UITrustedPublic APIFreemium
Payment
Open weights / Kimi Code plans
Checked
Aug 4
Sources
High confidence
From
Open weights on Hugging Face; Kimi API is $0.19/1M input tokens cache hit, $0.95/1M input cache miss and $4.00/1M output tokens
Baidu now has English model communications through the ERNIE Blog, while Qianfan remains the main platform when enterprise platform, agent orchestration and China-cloud deployment matter.
Best fit · Enterprises and developers already evaluating Baidu Cloud, China-local deployment, agent platforms or ERNIE multimodal models.
Coverage · 100/100
Partially availablePartial English UITrustedPublic APIPaid
Hunyuan is strategically important because Tencent combines cloud APIs, global cloud billing, a broad open-source model organization and a newly global 3D creation engine.
Best fit · Developers and enterprises comparing Chinese model APIs, Tencent Cloud integration, 3D asset generation and open-source model deployment.
Coverage · 100/100
Partially availableFull English UITrustedPublic APIPaid
Payment
Visa / Mastercard
Checked
Aug 4
Sources
High confidence
From
Tencent Cloud usage-based billing; Hunyuan 3D Global lists 20 free generations daily and 200 free API credits in Tencent's English announcement
Seed2.1 and Doubao/Ark should be tracked as ByteDance's productivity-model and model-platform profile, while Seedance remains a separate video-generation product.
Best fit · Developers and teams comparing ByteDance's English-facing Seed model roadmap with commercial Doubao/Ark API access.
Coverage · 100/100
Partially availableFull English UITrustedPublic APIPaid
Pangu belongs in LLM & API as Huawei Cloud's managed enterprise model and development platform, while its strongest differentiation is vertical industry adaptation.
Best fit · Enterprises building industry-specific AI on Huawei Cloud across documents, computer vision, scientific computing, prediction and agent workflows.
Coverage · 100/100
Partially availableFull English UITrustedLimited APIEnterprise
Payment
Huawei Cloud international billing / Enterprise contract
Checked
Aug 4
Sources
High confidence
From
Contact Huawei Cloud sales; public international Pangu pricing is not listed on the product page
StepFun is important because it combines multimodal model depth, open-source releases and device commercialization, but overseas usability still needs hands-on checks.
Best fit · Teams evaluating Chinese multimodal models, open-source agent models, video/audio generation or device-side AI partnerships.
Coverage · 100/100
Partially availableFull English UITrustedPublic APIPaid
Step 3.7 Flash is the current headline model on StepFun's platform homepage, so it deserves a dedicated profile instead of being buried inside the broader platform entry.
Best fit · Teams evaluating a flagship Chinese multimodal reasoning model for agentic coding, tool use, deep research and vision/video-aware workflows.
Coverage · 100/100
Partially availableFull English UITrustedPublic APIPaid
Payment
Account balance / Free credit first
Checked
Aug 4
Sources
High confidence
From
Input $0.04 cache hit / $0.20 cache miss; output $1.15 per 1M tokens