LLM Gateway
DeepSeek V4 Flash (Together AI)
together-ai/deepseek-v4-flash
Prices and limits
- Input price per 1M tokens
- $0.14
- Output price per 1M tokens
- $0.28
- Cached input per 1M tokens
- $0.03
- Blended price per 1M (our estimate)
- $0.175
- Context window
- 163,840
- Max output tokens
- 163,840
- Accepts
- Text
- Produces
- Text
- Released
- Apr 24, 2026
- Last updated
- Apr 24, 2026
- Knowledge cutoff
- May 2025
- Open weights
- Yes
Wrong price, or a provider that wants its listing corrected or removed? Email ceo@uniholix.com. We reply within 5 business days.
Cheapest comparable models
Other priced models that accept at least the same inputs, produce the same outputs and have at least half this context window, sorted by blended price.
- GPT OSS 20B (Greenference)Eden AI$0.009 / $0.045131.1K
- Llama 3.1 8B (decentralized)NanoGPT$0.02 / $0.03128K
- Meta Llama 3.1 8B Instruct TurboHelicone$0.02 / $0.03128K
- Mistral NemoPioneer$0.02 / $0.03128K
- Llama-3.1-8B-InstructKilo Gateway$0.02 / $0.04131.1K
- Mistral Nemo Instruct 2407IO.NET$0.02 / $0.04128K
- Llama 3.1 8B InstructAbacus$0.02 / $0.05128K
- Qwen3 4BNovitaAI$0.03 / $0.03128K
Same model at other providers
Matched by the catalogue’s model ID. Cheapest blended price first.
- DeepSeek V4 FlashKenari$0.00 / $0.001M
- DeepSeek V4 FlashPendra$0.00 / $0.001M
- DeepSeek V4 FlashSenseNova (China)$0.00 / $0.001M
- DeepSeek V4 FlashUnoRouter$0.00 / $0.001M
- DeepSeek V4 Flash (free)OrcaRouter$0.00 / $0.001M
- DeepSeek V4 Flash (Free)Kenari$0.00 / $0.001M
- deepseek-v4-flashInferX$0.00 / $0.001M
- DeepSeek V4 FlashMerge Gateway$0.035 / $0.071M
- DeepSeek V4 FlashDevPass (LLM Gateway)$0.065 / $0.1161.1M
- DeepSeek V4 FlashUnoRouter$0.063 / $0.1251M
- DeepSeek V4 FlashDeep Infra$0.09 / $0.181M
- DeepSeek V4 Flash FlexNeuralwatt$0.091 / $0.1821M