Skip to main content
Chinese AI Tools
ProductsModelsIntegrationsRankingsLatest changes
TopicsAvailability TrackerUse casesSubmit a toolAccount
ZHSearch

Chinese AI Tools

Independent directory for Chinese AI products. Product availability, pricing and terms can change. Verify before commercial use.

Editorial standardsClaim productUpdate infoGet featuredAdvertise

DeepSeek

DeepSeek V4 API

DeepSeek's current V4 API line, covering V4-Pro and V4-Flash with 1M context, thinking mode and tool calls.

Globally availableFull English UIPublic APIPaid

Quick answer

V4 is the current API line, with separate migration and pricing implications from the broader DeepSeek brand.

PricingDocumentation
Pricing
Off-peak V4-Flash: $0.007/$0.22/$0.66; V4-Pro: $0.022/$0.66/$1.98 per M cache-hit/cache-miss/output tokens; peak rates are double
Availability
Globally available
API
Public API
Last checked
2026-09-07
Pricing detailsUse-case fitSource evidence

The English DeepSeek docs list deepseek-v4-pro and deepseek-v4-flash as the stable public IDs, currently routing to V4-Pro-0813 and V4-Flash-0731. Both support thinking and non-thinking modes, JSON output, tool calls and chat prefix completion; FIM completion is available in non-thinking mode. Billing now varies by peak and off-peak UTC windows while model IDs remain unchanged.

Topic guide

DeepSeek models, API and agent ecosystem

A source-backed map of DeepSeek V4 models, API pricing, native Harness, third-party agent runtimes and practical integration guides.

Editorial verdict

Best for

Developers migrating DeepSeek integrations to current V4 model names and long-context reasoning workflows.

Avoid if

Avoid starting new projects on deepseek-chat or deepseek-reasoner aliases.

Why it matters

V4 is the current API line, with separate migration and pricing implications from the broader DeepSeek brand.

Trust: 3/3 sources verified, recently checkedCoverage: 100/100

Pricing

Off-peak V4-Flash: $0.007/$0.22/$0.66; V4-Pro: $0.022/$0.66/$1.98 per M cache-hit/cache-miss/output tokens; peak rates are double

Payment

Topped-up balance, Granted balance, Platform billing

Commercial use

Commercial use should follow the current product, API, model license and billing terms.

Privacy

Review prompt, file, media upload, retention and training-use terms before sensitive workloads.

Use-case fit

V4 model integration

Strong

Use deepseek-v4-pro for agent/coding tasks and deepseek-v4-flash for faster or lower-cost workloads.

Thinking mode control

Strong

Use the documented thinking toggle and effort controls for reasoning-heavy requests.

Global user checklist

RegistrationConfirmedRequires a DeepSeek Platform API key.
English UIConfirmedQuick start, pricing, guides and change log are English-facing.
API and docsConfirmedDocs cover OpenAI and Anthropic formats plus V4 feature details.
International paymentPartialPricing is clear, but account-specific top-up/payment routes still need verification.

Model names, quotas, release status, regional access and commercial terms can change quickly; recheck official sources before procurement or production use.

Pros

  • - 1M context and 384K maximum output are documented
  • - Thinking and non-thinking modes share one current model line
  • - Tool calls, JSON output and context caching are documented
  • - Documented concurrency limits are 2,500 for V4-Flash and 500 for V4-Pro

Cons

  • - Legacy aliases have passed their published deprecation date
  • - Prices can change and are deducted from granted balance before topped-up balance

Decision paths

qwen

kimi-k2-api

zhipu-glm

Sources

DeepSeek models and pricing

Pricing · EN · verified 2026-08-17

Lists the current V4-Pro-0813/V4-Flash-0731 routes, context and output limits, features, concurrency and active peak/off-peak prices.

DeepSeek thinking mode

Documentation · EN · verified 2026-05-17

Documents thinking toggle, effort control and reasoning_content behavior.

DeepSeek change log

Documentation · EN · verified 2026-05-17

Confirms the 2026-04-24 V4 release and 2026-07-24 alias deprecation.

Last checked: 2026-09-07

Reviews

No approved reviews yet.

Availability snapshot

Availability
Globally available
English UI
Full English UI
API
Public API
Rating
4.6 (0)

Latest updates

Latest changes
Release · 2026-08-28

Reasonix 1.32 adds remote workspaces and background turns

Reasonix desktop v1.32.0 adds remote SSH workspaces, persistent ordered turns, background turns, conversation navigation, session recovery and Windows fixes. It also moves MCP reliability work onto the official Go SDK and stores additional desktop metadata in SQLite.

Release · 2026-08-27

Hermes Agent 0.20.6 ships a large stable patch rollup

Hermes Agent v0.20.6, tagged v2026.8.27, rolls up roughly 525 pull requests. The release adds consent-gated real-profile browsing, a desktop Browser OS window, SSH remote update and fleet profiles, more than 50 remote MCP servers, web-search and extraction caches, OS keychain encryption and new model integrations.

Release · 2026-08-25

Hermes Agent added with native DeepSeek V4 support

This August 25 catalog update first tracked Nous Research's Hermes Agent at then-stable v0.20.5. Its native DeepSeek provider targeted V4 models, enabled thinking by default, preserved reasoning history across tool calls and migrated retired aliases. Newer Hermes releases are tracked as separate updates.

Release · 2026-08-17

Reasonix 1.26 adds DeepSeek peak/off-peak cost controls

Reasonix 1.26.0 updates its DeepSeek-native coding workflow for the new V4 Flash and V4 Pro peak/off-peak prices. It adds Pro low reasoning effort across OpenAI Chat, Responses and Anthropic modes, unifies cost estimates across CLI, ACP and desktop surfaces, labels peak, off-peak and mixed rates, and automatically migrates configuration to v7.

Submit a review