Alibaba · Qwen
Qwen3.8 Flash Next
qwen/qwen3.8-flash-next
Open experimental architecture preview behind the production Qwen3.8-Flash line, with 125B parameters, 6B activated, native vision and efficient long-context attention.
Model specifications
- Context window
- 262K tokens
- Output limit
- 131K tokens
- Release date
- 2026-08-26
- Input modalities
- text, image, video
- Output modalities
- text
- Last updated
- 2026-09-07
Provider pricing
USD per 1M tokens. Limits may differ from the underlying model.
Checked 2026-09-07
Self-hosted
Qwen/Qwen3.8-Flash-Next
- Input
- -
- Output
- -
- Context
- 262K
Promotions, plan quotas, cache pricing and regional taxes are not included. Recheck the provider before purchase.
Qwen3.8 Flash Next FAQ
Is Qwen3.8 Flash Next an open-weight model?
Yes. Qwen3.8 Flash Next is tracked as open weight; review the linked license and model source before redistribution or production deployment.
What is the context window for Qwen3.8 Flash Next?
Qwen3.8 Flash Next has a tracked context window of 262K tokens and a maximum output of 131K tokens. Provider limits can be lower.
What input does Qwen3.8 Flash Next support?
The current record lists text, image, video input and text output. Check the exact provider route before relying on a modality in production.
How much does Qwen3.8 Flash Next cost?
This page compares 1 tracked provider offering in USD per 1M tokens where public prices are available. Promotions, cache rates, taxes and plan quotas may differ.