Chinese model releases to evaluate when teams need model weights, local deployment, inference control or lower vendor lock-in.
Quick answers
At a glance
What it covers
Chinese model releases to evaluate when teams need model weights, local deployment, inference control or lower vendor lock-in.
Matched tools
35 tools currently match this use case.
How to read this page
Prioritize products with public model repositories, documented licenses, deployment notes or explicit self-hosting paths.
Decision standard
Prioritize products with public model repositories, documented licenses, deployment notes or explicit self-hosting paths.
35 matched tools
DeepSeek
DeepSeek
4.7
The English docs now make DeepSeek's current API shape explicit: V4-Pro/V4-Flash, dual OpenAI/Anthropic compatibility, thinking mode, tool calls, caching and agent-tool integrations.
Best fit · Developers evaluating Chinese model APIs for coding agents, reasoning, tool use and OpenAI/Anthropic-compatible migration.
Coverage · 100/100
Globally availableFull English UITrustedPublic APIPaid
Payment
Topped-up balance / Granted balance
Checked
Aug 4
Sources
High confidence
From
V4-Flash from $0.0028/M cache-hit input, $0.14/M cache-miss input and $0.28/M output tokens
OpenBMB belongs in LLM & API as a Chinese open-source foundation-model ecosystem, with MiniCPM and MiniCPM-V as the main product signals tracked here.
Best fit · Researchers and engineering teams evaluating Chinese open-source models for edge inference, multimodal understanding, RAG and agent experiments.
Coverage · 100/100
Partially availablePartial English UITrustedLimited APIFree
AtomCode belongs in AI Coding because it is a full terminal coding agent and Claude Code alternative, not a generic chat assistant or raw model API.
Best fit · Developers who want an open-source terminal coding agent for Chinese and global model APIs, especially DeepSeek, GLM, Qwen and local Ollama workflows.
Coverage · 100/100
Partially availablePartial English UITrustedLimited APIFreemium
Payment
CodingPlan FREE / Manual API key
Checked
Aug 4
Sources
High confidence
From
Open-source MIT project; CodingPlan FREE is advertised; manual API key configuration is available
M3 should be tracked separately because the official homepage now gives it a distinct model page and positions it as the current frontier coding model.
Best fit · Developers evaluating a China-origin frontier model for coding, long-context agentic work and multimodal reasoning.
Coverage · 100/100
Globally availableFull English UITrustedPublic APIFreemium
Payment
API & Token Plan / Pay-as-you-go API billing
Checked
Aug 4
Sources
High confidence
From
API & Token Plan access; current pricing should be checked in account
Qwen-AgentWorld is tracked separately from Qwen Code because its public surface is a language world model plus AgentWorldBench, not a terminal coding product.
Best fit · Researchers and agent builders who need a simulator or benchmark for tool, terminal, SWE, Android, web and OS agent environments.
Coverage · 100/100
Globally availableFull English UITrustedPublic APIFree
Payment
Hugging Face download / ModelScope download
Checked
Aug 4
Sources
High confidence
From
Apache-2.0 open weights and benchmark; self-hosted inference infrastructure required
Ornith 1.0 belongs in AI Coding because its public model cards and collection position it around agentic coding benchmarks and coding-agent deployment.
Best fit · Teams evaluating open agentic-coding models for self-hosted coding agents, terminal automation, SWE-Bench workflows and long-context tool use.
Coverage · 100/100
Globally availableFull English UITrustedPublic APIFree
Payment
Hugging Face model download / Transformers
Checked
Aug 4
Sources
High confidence
From
MIT open weights on Hugging Face; self-hosted inference costs depend on model size and quantization
JoyAI-VL-Interaction deserves a separate profile because it releases a full real-time interaction model and system stack, not just a static VLM checkpoint.
Best fit · Researchers and product teams exploring proactive webcam, livestream, monitoring, cooking guidance, live commentary and event-driven visual interaction systems.
Coverage · 100/100
Globally availableFull English UITrustedPublic APIFree
Kimi K3 merits a separate LLM & API profile because it is a flagship model with its own architecture, context length, hosted pricing and public open-weight release, rather than a minor K2.7 Code revision.
Best fit · Teams evaluating a frontier-scale Chinese model for repository-scale coding, multimodal technical work, agentic research and long-context knowledge workflows.
Coverage · 100/100
Globally availableFull English UITrustedPublic APIPaid
Payment
Kimi API pay-as-you-go billing / Kimi membership
Checked
Aug 4
Sources
High confidence
From
$0.30/M cache-hit input, $3/M cache-miss input and $15/M output tokens
K2.7 Code has its own open weights, architecture, benchmarks, API pricing and Kimi Code default-model behavior, so it deserves a dedicated model profile separate from Kimi Code and the broader Kimi API.
Best fit · Teams evaluating open-weight coding models for repository-scale agents, Kimi Code workflows or OpenAI-compatible API integration.
Coverage · 100/100
Globally availableFull English UITrustedPublic APIFreemium
Payment
Open weights / Kimi Code plans
Checked
Aug 4
Sources
High confidence
From
Open weights on Hugging Face; Kimi API is $0.19/1M input tokens cache hit, $0.95/1M input cache miss and $4.00/1M output tokens
The public repository confirms that Kimi Code is an installable MIT-licensed agent platform, while the K2.7 Code resource confirms its current default hosted coding model.
Best fit · Developers who want a fast open terminal agent with multimodal coding context, parallel subagents and editor integration.
Coverage · 100/100
Globally availableFull English UITrustedLimited APIFreemium
Payment
Free open-source CLI / Kimi Code OAuth plans
Checked
Aug 4
Sources
High confidence
From
MIT-licensed CLI; annual-billing Kimi Code plans are listed at $15, $31, $79 and $159 per month, with model usage also possible through API billing or configured providers
LongCat-AudioDiT belongs in AI Audio because it is a direct-text-to-speech and voice-cloning model with released code and weights, not a generic research paper.
Best fit · Researchers and speech teams evaluating open-source TTS, waveform-latent diffusion and zero-shot voice cloning.
Coverage · 100/100 · backfill: freshness
Globally availableFull English UITrustedLimited APIFree
Payment
GitHub repository / Model weights download
Checked
Aug 4
Sources
High confidence
From
Open-source MIT repository and released model weights; inference runs locally or through a Hugging Face-compatible workflow
Confucius4-TTS belongs in AI Audio because it is an open TTS and voice-cloning engine with model downloads, inference code and benchmark claims.
Best fit · Speech researchers and localization teams evaluating open multilingual TTS, cross-lingual dubbing, zero-shot voice cloning and emotion-preserving voice transfer.
Coverage · 100/100
Globally availableFull English UITrustedLimited APIFree
Payment
GitHub repository / Hugging Face model download
Checked
Aug 4
Sources
High confidence
From
Apache-2.0 open-source repository with Hugging Face and ModelScope model downloads; local CUDA inference required
Baidu now has English model communications through the ERNIE Blog, while Qianfan remains the main platform when enterprise platform, agent orchestration and China-cloud deployment matter.
Best fit · Enterprises and developers already evaluating Baidu Cloud, China-local deployment, agent platforms or ERNIE multimodal models.
Coverage · 100/100
Partially availablePartial English UITrustedPublic APIPaid
Hunyuan is strategically important because Tencent combines cloud APIs, global cloud billing, a broad open-source model organization and a newly global 3D creation engine.
Best fit · Developers and enterprises comparing Chinese model APIs, Tencent Cloud integration, 3D asset generation and open-source model deployment.
Coverage · 100/100
Partially availableFull English UITrustedPublic APIPaid
Payment
Visa / Mastercard
Checked
Aug 4
Sources
High confidence
From
Tencent Cloud usage-based billing; Hunyuan 3D Global lists 20 free generations daily and 200 free API credits in Tencent's English announcement
DeerFlow belongs in Productivity because it is a SuperAgent harness broader than an AI coding agent: its core promise is long-horizon research, creation and workflow execution through skills, sub-agents and sandboxes.
Best fit · Teams that want to self-host a Chinese-origin open-source agent harness for deep research, report generation, coding, file-based workflows and long-running multi-agent tasks.
Coverage · 100/100
Globally availableFull English UITrustedLimited APIFree
Payment
Self-hosted open source / Configured model provider billing
Checked
Aug 4
Sources
High confidence
From
MIT open-source self-hosted project; model, search and infrastructure costs depend on configured providers
StepFun is important because it combines multimodal model depth, open-source releases and device commercialization, but overseas usability still needs hands-on checks.
Best fit · Teams evaluating Chinese multimodal models, open-source agent models, video/audio generation or device-side AI partnerships.
Coverage · 100/100
Partially availableFull English UITrustedPublic APIPaid
MiMoCode is a distinct coding-agent product rather than another MiMo model card: it packages memory, orchestration, tools and provider routing into an installable open-source CLI.
Best fit · Developers who want an open terminal coding agent focused on durable project memory, long-running tasks and configurable model providers.
Coverage · 100/100
Globally availableFull English UITrustedLimited APIFreemium
Payment
Free open-source installation / MiMo Auto limited-time access
Checked
Aug 4
Sources
High confidence
From
Open-source CLI; MiMo Auto is free for a limited time, while hosted and custom-provider costs vary
DeepSeek: DeepSeek is included because its sources mention open models, model hubs, self-hosting or deployment-oriented evidence.
OpenBMB / MiniCPM: OpenBMB / MiniCPM is included because its sources mention open models, model hubs, self-hosting or deployment-oriented evidence.
AtomCode: AtomCode is included because its sources mention open models, model hubs, self-hosting or deployment-oriented evidence.
Qwen: Qwen is included because its sources mention open models, model hubs, self-hosting or deployment-oriented evidence.
Qwen-Image / Wan Image: Qwen-Image / Wan Image is included because its sources mention open models, model hubs, self-hosting or deployment-oriented evidence.
Qwen Audio / CosyVoice: Qwen Audio / CosyVoice is included because its sources mention open models, model hubs, self-hosting or deployment-oriented evidence.
MiniMax M3: MiniMax M3 is included because its sources mention open models, model hubs, self-hosting or deployment-oriented evidence.
Qwen Code: Qwen Code is included because its sources mention open models, model hubs, self-hosting or deployment-oriented evidence.
Qwen-AgentWorld: Qwen-AgentWorld is included because its sources mention open models, model hubs, self-hosting or deployment-oriented evidence.
Ornith 1.0: Ornith 1.0 is included because its sources mention open models, model hubs, self-hosting or deployment-oriented evidence.
JoyAI-VL-Interaction: JoyAI-VL-Interaction is included because its sources mention open models, model hubs, self-hosting or deployment-oriented evidence.
Kimi K3: Kimi K3 is included because its sources mention open models, model hubs, self-hosting or deployment-oriented evidence.
Kimi K2.7 Code: Kimi K2.7 Code is included because its sources mention open models, model hubs, self-hosting or deployment-oriented evidence.
Kimi Code: Kimi Code is included because its sources mention open models, model hubs, self-hosting or deployment-oriented evidence.
Z.ai BigModel / GLM: Z.ai BigModel / GLM is included because its sources mention open models, model hubs, self-hosting or deployment-oriented evidence.
Ring: Ring is included because its sources mention open models, model hubs, self-hosting or deployment-oriented evidence.
MiniMax API Platform: MiniMax API Platform is included because its sources mention open models, model hubs, self-hosting or deployment-oriented evidence.
MiniMax Agent / Mavis: MiniMax Agent / Mavis is included because its sources mention open models, model hubs, self-hosting or deployment-oriented evidence.
LongCat-AudioDiT: LongCat-AudioDiT is included because its sources mention open models, model hubs, self-hosting or deployment-oriented evidence.
Confucius4-TTS: Confucius4-TTS is included because its sources mention open models, model hubs, self-hosting or deployment-oriented evidence.
ERNIE / Baidu Qianfan: ERNIE / Baidu Qianfan is included because its sources mention open models, model hubs, self-hosting or deployment-oriented evidence.
Tencent Hunyuan: Tencent Hunyuan is included because its sources mention open models, model hubs, self-hosting or deployment-oriented evidence.
Hy3 Preview: Hy3 Preview is included because its sources mention open models, model hubs, self-hosting or deployment-oriented evidence.
Hunyuan 3D Global: Hunyuan 3D Global is included because its sources mention open models, model hubs, self-hosting or deployment-oriented evidence.
Tencent Hunyuan Open Models: Tencent Hunyuan Open Models is included because its sources mention open models, model hubs, self-hosting or deployment-oriented evidence.
AgentTeams (formerly HiClaw): AgentTeams (formerly HiClaw) is included because its sources mention open models, model hubs, self-hosting or deployment-oriented evidence.
QwenPaw: QwenPaw is included because its sources mention open models, model hubs, self-hosting or deployment-oriented evidence.
DeerFlow: DeerFlow is included because its sources mention open models, model hubs, self-hosting or deployment-oriented evidence.
Astron Agent: Astron Agent is included because its sources mention open models, model hubs, self-hosting or deployment-oriented evidence.
StepFun / Step: StepFun / Step is included because its sources mention open models, model hubs, self-hosting or deployment-oriented evidence.
Xiaomi MiMo: Xiaomi MiMo is included because its sources mention open models, model hubs, self-hosting or deployment-oriented evidence.
MiMo-V2-Flash: MiMo-V2-Flash is included because its sources mention open models, model hubs, self-hosting or deployment-oriented evidence.
MiMo Speech Models: MiMo Speech Models is included because its sources mention open models, model hubs, self-hosting or deployment-oriented evidence.
MiMo API Platform: MiMo API Platform is included because its sources mention open models, model hubs, self-hosting or deployment-oriented evidence.
MiMoCode: MiMoCode is included because its sources mention open models, model hubs, self-hosting or deployment-oriented evidence.