26 models
Prices: USD per 1M tokens
Alibaba
Qwen3.8 Flash Next
qwen/qwen3.8-flash-next
OpenOpen experimental architecture preview behind the production Qwen3.8-Flash line, with 125B parameters, 6B activated, native vision and efficient long-context attention.
- Context
- 262K
- Providers
- 1
- Input from
- -
ReasoningToolsStructuredMultimodal
Zhipu AI
GLM-5.3-Flash
zhipuai/glm-5.3-flash
OpenOpen MIT-licensed multimodal GLM model with 320B total and 18B active parameters, a million-token context, hybrid sparse and linear attention, and first-party API access.
- Context
- 1.05M
- Providers
- 2
- Input from
- -
ReasoningToolsStructuredMultimodal
Zhipu AI
GLM-5.3
zhipuai/glm-5.3
HostedThe original hosted GLM-5.3 flagship for complex software engineering, long-horizon agents and cybersecurity analysis. Its Coding Plan access and text-only profile are distinct from the newer open, multimodal GLM-5.3-Flash variant.
- Context
- 1M
- Providers
- 1
- Input from
- -
ReasoningToolsStructured
Alibaba
Qwen3.8 27B
alibaba/qwen3.8-27b
OpenApache-2.0 dense 27B multimodal Qwen model for text, image and video understanding with reasoning and tool use.
- Context
- 262K
- Providers
- 1
- Input from
- -
ReasoningToolsStructuredMultimodal
DeepSeek
DeepSeek V4 Pro 0813
deepseek/deepseek-v4-pro-0813
HostedThe hosted V4 Pro snapshot currently served by DeepSeek's official API, with million-token context, 384K output and peak/off-peak billing.
- Context
- 1M
- Providers
- 2
- Input from
- $0.66
ReasoningToolsStructured
Alibaba
Qwen3.8 2.4T A95B
alibaba/qwen3.8-2.4t-a95b
OpenQwen's first open Max-class text model, a sparse 2.4T-parameter MoE with 95B active parameters for generation, coding and agent work.
- Context
- 262K
- Providers
- 1
- Input from
- -
ReasoningToolsStructured
Alibaba
Qwen3.8 Max
alibaba/qwen3.8-max
HostedQwen Studio's current flagship model for coding, professional work, multimodal understanding and long-horizon agents.
- Context
- 1M
- Providers
- 2
- Input from
- -
ReasoningToolsTemperatureMultimodal
Moonshot AI
Kimi K3
moonshotai/kimi-k3
OpenMultimodal Kimi model with million-token context and max-effort reasoning for long agent work.
- Context
- 1.05M
- Providers
- 2
- Input from
- $3
ReasoningToolsStructuredMultimodal
Tencent
Tencent Hy3
tencent/hy3
OpenOpen reasoning model for coding, instruction following and agent tasks.
- Context
- 256K
- Providers
- 2
- Input from
- $0.13
ReasoningToolsTemperature
Zhipu AI
GLM-5.2
zhipuai/glm-5.2
OpenOpen flagship GLM for long-horizon coding agents and million-token context work.
- Context
- 1M
- Providers
- 2
- Input from
- $0.56
ReasoningToolsStructured
Moonshot AI
Kimi K2.7 Code
moonshotai/kimi-k2.7-code
OpenCoding-focused Kimi model designed for long-horizon repository work with controlled reasoning.
- Context
- 262K
- Providers
- 2
- Input from
- $0.95
ReasoningToolsStructuredMultimodal
Alibaba
Qwen3.7 Plus
alibaba/qwen3.7-plus
HostedMultimodal Qwen workhorse for long-context agents, visual inputs and coding.
- Context
- 1M
- Providers
- 2
- Input from
- $0.32
ReasoningToolsTemperatureMultimodal
MiniMax
MiniMax M3
minimax/MiniMax-M3
OpenOpen multimodal model for long-context coding, perception and agent planning.
- Context
- 1M
- Providers
- 2
- Input from
- $0.23
ReasoningToolsTemperatureMultimodal
StepFun
Step 3.7 Flash
stepfun/step-3.7-flash
OpenOpen multimodal flash model for fast agents, coding and visual prompts.
- Context
- 256K
- Providers
- 2
- Input from
- $0.18
ReasoningToolsTemperatureMultimodal
Alibaba
Qwen3.7 Max
alibaba/qwen3.7-max
HostedQwen frontier model tuned for agent frameworks, coding assistants and long tasks.
- Context
- 1M
- Providers
- 2
- Input from
- $1.48
ReasoningToolsTemperature
DeepSeek
DeepSeek V4 Flash
deepseek/deepseek-v4-flash
OpenFast DeepSeek V4 lane for economical reasoning, coding and long-context work.
- Context
- 1M
- Providers
- 2
- Input from
- $0.0882
ReasoningToolsStructured
DeepSeek
DeepSeek V4 Pro
deepseek/deepseek-v4-pro
OpenOpen MoE flagship with million-token context for coding and long agent runs.
- Context
- 1M
- Providers
- 2
- Input from
- $0.43
ReasoningToolsStructured
Xiaomi
MiMo-V2.5
xiaomi/mimo-v2.5
OpenOpen multimodal MiMo model for coding agents and long-context automation.
- Context
- 1.05M
- Providers
- 2
- Input from
- $0.14
ReasoningToolsTemperatureMultimodal
Xiaomi
MiMo-V2.5-Pro
xiaomi/mimo-v2.5-pro
OpenOpen MiMo Pro tier for long-context reasoning and coding-agent execution.
- Context
- 1.05M
- Providers
- 2
- Input from
- $0.43
ReasoningToolsTemperature
Tencent
Tencent Hy3 Preview
tencent/hy3-preview
OpenPreview generation of Tencent's open reasoning model for coding and agent evaluation.
- Context
- 256K
- Providers
- 2
- Input from
- $0.063
ReasoningToolsTemperature
Zhipu AI
GLM-5.1
zhipuai/glm-5.1
OpenOpen GLM coding model for agentic engineering, terminals and repository generation.
- Context
- 200K
- Providers
- 2
- Input from
- $0.95
ReasoningToolsStructured
Zhipu AI
GLM-5V-Turbo
zhipuai/glm-5v-turbo
HostedFast vision model for screenshots, documents and multimodal agent tasks.
- Context
- 200K
- Providers
- 2
- Input from
- $1.2
ReasoningToolsTemperatureMultimodal
Xiaomi
MiMo-V2-Omni
xiaomi/mimo-v2-omni
HostedOmni model for text, image, video, audio, PDFs and agent workflows.
- Context
- 262K
- Providers
- 1
- Input from
- $0.4
ReasoningToolsTemperatureMultimodal
MiniMax
MiniMax M2.7
minimax/MiniMax-M2.7
OpenOpen model for coding agents, office automation and complex tool environments.
- Context
- 205K
- Providers
- 2
- Input from
- $0.25
ReasoningToolsTemperature
StepFun
Step 3.5 Flash
stepfun/step-3.5-flash
OpenOpen flash model for quick reasoning, coding assistance and efficient deployment.
- Context
- 256K
- Providers
- 2
- Input from
- $0.1
ReasoningToolsTemperature
Alibaba
Qwen3 Coder Plus
alibaba/qwen3-coder-plus
HostedHosted Qwen coding model for software agents, repository edits and long-context code work.
- Context
- 1.05M
- Providers
- 2
- Input from
- $0.65
ToolsTemperature