Back to models
Qwen Flash vs Qwen3.8-Flash
Qwen Flash and Qwen3.8-Flash: current price, context, max output, and supported operations.
Pricing comparison
| Qwen Flash | Qwen3.8-Flash | |
|---|---|---|
| Model maker | Alibaba Cloud | Alibaba Cloud |
| Delivery availability | Available | Available |
| Context window | 998K | 992K |
| Max output | 32K | 128K |
| Official price | Input$0.05per 1M tokensOutput$0.40per 1M tokens | Input$0.15per 1M tokensOutput$0.47per 1M tokens |
| TokenLab price | — | — |
| Model performance | Collecting data |
|
| Capabilities | Prompt CacheReasoning | JSON modePrompt CacheTool useVision |
Summary
- Qwen Flash: Input $0.05 / Output $0.40 per 1M tokens; Qwen3.8-Flash: Input $0.15 / Output $0.47 per 1M tokens. Rates depend on the billing unit, specification, and usage. Compare matching conditions in the model's detailed pricing; a single rate does not determine the total cost.
- Qwen3.8-Flash supports 4.0x more output tokens
FAQ
Which is cheaper, Qwen Flash or Qwen3.8-Flash?
Qwen Flash: Input $0.05 / Output $0.40 per 1M tokens; Qwen3.8-Flash: Input $0.15 / Output $0.47 per 1M tokens. Rates depend on the billing unit, specification, and usage. Compare matching conditions in the model's detailed pricing; a single rate does not determine the total cost.
What are the key differences between Qwen Flash and Qwen3.8-Flash?
Qwen Flash: Reasoning. Qwen3.8-Flash: JSON mode, Tool use, Vision