LLM Gateway
Qwen3.8 Flash (Alibaba Cloud)
alibaba/qwen3.8-flash
Prices and limits
- Input price per 1M tokens
- $0.15
- Output price per 1M tokens
- $0.47
- Cached input per 1M tokens
- $0.016
- Blended price per 1M (our estimate)
- $0.23
- Context window
- 983,616
- Max output tokens
- 131,072
- Accepts
- TextImageVideo
- Produces
- Text
- Released
- Aug 26, 2026
- Last updated
- Aug 26, 2026
- Knowledge cutoff
- —
- Open weights
- No
Wrong price, or a provider that wants its listing corrected or removed? Email ceo@uniholix.com. We reply within 5 business days.
Cheapest comparable models
Other priced models that accept at least the same inputs, produce the same outputs and have at least half this context window, sorted by blended price.
- GLM-5.3-FlashPareto Inference$0.03 / $0.101M
- Qwen3.7 FlashAlibaba (China)$0.03 / $0.1191M
- Space Bunny AlphaNanoGPT$0.05 / $0.151M
- Gemini-2.0-Flash-LitePoe$0.052 / $0.21990K
- Qwen3.5 FlashAIHubMix$0.028 / $0.2821M
- Qwen3.5 FlashAlibaba (China)$0.029 / $0.2871M
- Gemini 2.5 Flash-LiteLLMTR$0.10 / $0.101M
- Qwen: Qwen3.5-FlashKilo Gateway$0.065 / $0.261M
Same model at other providers
Matched by the catalogue’s model ID. Cheapest blended price first.
- Qwen3.8 FlashNaN$0.00 / $0.00262.1K
- Qwen3.8 FlashAIHubMix$0.113 / $0.381M
- Qwen3.8 FlashOfox$0.11 / $0.391M
- Qwen3.8 FlashDeep Infra$0.113 / $0.3821M
- Qwen3.8 FlashVancine$0.12 / $0.381M
- Qwen3.8 FlashAlibaba (China)$0.119 / $0.4011M
- Qwen3.8 FlashCrossModel$0.13 / $0.431M
- Qwen3.8 FlashNanoGPT$0.14 / $0.42991.8K
- Qwen 3.8 FlashVenice AI$0.14 / $0.491M
- Qwen 3.8 FlashVercel AI Gateway$0.15 / $0.47991K
- Qwen3.8 FlashAlibaba$0.15 / $0.471M
- Qwen3.8 FlashEden AI$0.15 / $0.471M
Price changes
Models whose price changed in the catalogue, newest first. We compare prices every day.
No price changes recorded yet. When a model’s price changes, it appears here.