StepFun
StepFun Open Platform is the developer entry for StepFun. The English docs cover OpenAI-compatible chat completion, model listing, files, token counting, tool calls, image generation and editing, TTS, ASR, voice cloning, pricing and agreements. The current homepage now highlights Step 3.7 Flash as the flagship multimodal reasoning model, while the broader text lineup still includes step-3.5-flash, step-2 and step-1 models.
Editorial verdict
Developers comparing Chinese model APIs for text, reasoning, tool calling, multimodal generation and OpenAI-compatible migration.
Avoid relying on it blindly when procurement requires card billing, regional SLA or enterprise data terms without account-level confirmation.
The English platform makes StepFun more actionable for overseas developers than a company-only profile.
Reasoning models start at $0.10 input cache miss / $0.02 cache hit / $0.30 output per 1M tokens; image editing is $0.003 per image
Account balance, Free credit first, Paid balance, WeChat Pay, Stripe for overseas users via Step Plan
Commercial use should follow the StepFun Open Platform terms, model-specific pricing and Step Plan terms where applicable.
The English docs publish privacy, terms and data-processing agreement pages; review them before sensitive workloads.
Use the documented chat-completions path when migrating model calls from OpenAI-style SDKs.
The docs include tool-call support for applications that need external systems or actions.
Compare StepFun against MiniMax, Qwen Cloud and Z.ai for audio, image and model-platform breadth.
Use Step 3.7 Flash when the platform needs one model for images, video, tool calls and long-context reasoning.
Model list, flagship model positioning, plan benefits and pricing are changing quickly; verify current docs and account console before production.
It is the current homepage hero and the best single-model entry for multimodal reasoning.
Step Plan uses a subscription quota model for agent and coding tools.
MiniMax has mature international docs across text, speech, video, image and music.
official · en · verified 2026-08-04
Confirms the English platform entry point and Step 3.7 Flash homepage highlight.
docs · en · verified 2026-05-29
Lists API reference, model guides, pricing and Step Plan integration docs.
docs · en · verified 2026-05-29
Documents the flagship multimodal reasoning model, pricing, effort levels and framework support.
pricing · en · verified 2026-05-29
Documents reasoning, speech pricing and tiered rate limits.
Last checked: 2026-08-04