Which AI model is cheapest right now?
Prices per 1M tokens, context windows and input types for 7,508 model offers from 199 providers, read from a public open catalogue every 6 hours.
251–300 of 6,826
- GLM-4 32B (0414-128k) (Z AI)LLM GatewayInput / 1M$0.10Output / 1M$0.10Context128K
- GPT OSS 120BSyntheticInput / 1M$0.10Output / 1M$0.10Context131.1K
- GPT-OSS 120BMerge GatewayInput / 1M$0.05Output / 1M$0.25Context128K
- Llama-3.1-8B-CSPoeInput / 1M$0.10Output / 1M$0.10Context128K
- Llama-3.2-1BPioneerInput / 1M$0.10Output / 1M$0.10Context131.1K
- Llama-3.2-3BPioneerInput / 1M$0.10Output / 1M$0.10Context131.1K
- Ministral 3 3BAmazon BedrockInput / 1M$0.10Output / 1M$0.10Context256K
- Ministral 3BDevPass (LLM Gateway)Input / 1M$0.10Output / 1M$0.10Context131.1K
- Ministral 3BNanoGPTInput / 1M$0.10Output / 1M$0.10Context131.1K
- Ministral 3BPioneerInput / 1M$0.10Output / 1M$0.10Context128K
- Ministral 3B (Mistral AI)LLM GatewayInput / 1M$0.10Output / 1M$0.10Context131.1K
- Ministral 8B (latest)MistralInput / 1M$0.10Output / 1M$0.10Context128K
- Ministral 8B (latest)Vercel AI GatewayInput / 1M$0.10Output / 1M$0.10Context128K
- Mistral: Ministral 3 3B 2512Kilo GatewayInput / 1M$0.10Output / 1M$0.10Context131.1K
- Nemotron 3 Ultra 550B A55Brouting.runInput / 1M$0.10Output / 1M$0.10Context131.1K
- Nvidia Nemotron 3 Super 120BNanoGPTInput / 1M$0.05Output / 1M$0.25Context262.1K
- Nvidia Nemotron 3 Super 120B ThinkingNanoGPTInput / 1M$0.05Output / 1M$0.25Context262.1K
- OpenAI GPT OSS 120BNovitaAIInput / 1M$0.05Output / 1M$0.25Context131.1K
- Qwen2.5-Coder-0.5BPioneerInput / 1M$0.10Output / 1M$0.10Context32.8K
- Qwen3 1.7B BasePioneerInput / 1M$0.10Output / 1M$0.10Context32.8K
- Qwen3.5 9BEmpirioLabs AIInput / 1M$0.09Output / 1M$0.13Context262.1K
- Qwen3.5 9BMerge GatewayInput / 1M$0.09Output / 1M$0.13Context262.1K
- Reka EdgeKilo GatewayInput / 1M$0.10Output / 1M$0.10Context16.4K
- MythoMax 13BNanoGPTInput / 1M$0.10Output / 1M$0.10Context4.1K
- Qwen-Omni TurboAlibaba (China)Input / 1M$0.058Output / 1M$0.23Context32.8K
- Qwen3-Omni FlashAlibaba (China)Input / 1M$0.058Output / 1M$0.23Context65.5K
- Qwen Doc TurboAlibaba (China)Input / 1M$0.087Output / 1M$0.144Context131.1K
- GPT-5-MiniQiHangInput / 1M$0.04Output / 1M$0.29Context200K
- GPT OSS 20BFrogBotInput / 1M$0.07Output / 1M$0.20Context131.1K
- Nemotron 3.5 LightningCoreWeaveInput / 1M$0.07Output / 1M$0.20Context262.1K
- Nemotron Nano 9BMerge GatewayInput / 1M$0.06Output / 1M$0.23Context128K
- Nvidia Nemotron Nano 9B V2Vercel AI GatewayInput / 1M$0.06Output / 1M$0.23Context131.1K
- NVIDIA Nemotron Nano 9B v2Amazon BedrockInput / 1M$0.06Output / 1M$0.23Context131.1K
- Amazon Nova Lite 1.0NanoGPTInput / 1M$0.06Output / 1M$0.238Context300K
- Amazon: Nova Lite 1.0Kilo GatewayInput / 1M$0.06Output / 1M$0.24Context300K
- DeepSeek V4 Flash (DeepInfra)LLM GatewayInput / 1M$0.08Output / 1M$0.18Context1M
- DeepSeek V4 Flash 0731AmbientInput / 1M$0.08Output / 1M$0.18Context1M
- Nemotron 3.5 Lightning 30B A3BNebius Token FactoryInput / 1M$0.06Output / 1M$0.24Context1M
- nemotron-3-nano-omniRequestyInput / 1M$0.06Output / 1M$0.24Context300K
- nemotron-3-nano-omni@euRequestyInput / 1M$0.06Output / 1M$0.24Context300K
- nemotron-3-nano:30bOllama CloudInput / 1M$0.06Output / 1M$0.24Context1M
- Nova LiteAmazon BedrockInput / 1M$0.06Output / 1M$0.24Context300K
- Nova LiteEden AIInput / 1M$0.06Output / 1M$0.24Context300K
- Nova LiteVercel AI GatewayInput / 1M$0.06Output / 1M$0.24Context300K
- Nova Lite (US)Amazon BedrockInput / 1M$0.06Output / 1M$0.24Context300K
- Nova Lite (US)Eden AIInput / 1M$0.06Output / 1M$0.24Context300K
- NVIDIA Nemotron Nano 3 30BAmazon BedrockInput / 1M$0.06Output / 1M$0.24Context262.1K
- nvidia-nemotron-3-nano-30b-a3bCortecsInput / 1M$0.06Output / 1M$0.24Context256K
- Mistral NemoNanoGPTInput / 1M$0.10Output / 1M$0.121Context16.4K
- The Drummer Cydonia 24B v2NanoGPTInput / 1M$0.10Output / 1M$0.121Context32.8K
Price changes
Models whose price changed in the catalogue, newest first. We compare prices every day.
- Z.ai: GLM Flash LatestKilo GatewayInput / 1M$0.026Output / 1M$0.90$0.929▲ 3%Changed: Oct 2, 2026, 8:48 PM UTC
- MiMo-V2.6-FlashKilo GatewayInput / 1M$0.07$0.12▲ 71%Output / 1M$0.28Changed: Oct 2, 2026, 8:48 PM UTC
- Hy4 previewKilo GatewayInput / 1M$0.751$0.83▲ 11%Output / 1M$2.25$2.50▲ 11%Changed: Oct 2, 2026, 8:48 PM UTC
- MoonshotAI: Kimi LatestKilo GatewayInput / 1M$1.34$1.00▼ 25%Output / 1M$13.00$11.36▼ 13%Changed: Oct 2, 2026, 8:48 PM UTC
- DeepSeek: DeepSeek V4 Flash LatestKilo GatewayInput / 1M$0.005775$0.003825▼ 34%Output / 1M$1.47$1.04▼ 29%Changed: Oct 2, 2026, 8:48 PM UTC
- DeepSeek: DeepSeek V4 Flash LatestKilo GatewayInput / 1M$0.013$0.005775▼ 55%Output / 1M$1.60$1.47▼ 8%Changed: Oct 2, 2026, 7:19 PM UTC
- DeepSeek: DeepSeek Flash LatestKilo GatewayInput / 1M$0.015$0.02▲ 33%Output / 1M$1.20$0.60▼ 50%Changed: Oct 2, 2026, 7:19 PM UTC
- Claude Sonnet 5.5Venice AIInput / 1M$3.75$2.50▼ 33%Output / 1M$18.75$12.50▼ 33%Changed: Oct 2, 2026, 5:25 PM UTC
- Claude Sonnet 5Venice AIInput / 1M$3.00$2.50▼ 17%Output / 1M$15.00$12.50▼ 17%Changed: Oct 2, 2026, 5:25 PM UTC
- GLM 4.7 ThinkingNanoGPTInput / 1M$0.20$0.40▲ 100%Output / 1M$0.80$1.93▲ 141%Changed: Oct 2, 2026, 5:25 PM UTC
- GLM 4.7NanoGPTInput / 1M$0.20$0.40▲ 100%Output / 1M$0.80$1.93▲ 141%Changed: Oct 2, 2026, 5:25 PM UTC
- GPT-OSS 20BMerge GatewayInput / 1M$0.04Output / 1M$0.20$0.15▼ 25%Changed: Oct 2, 2026, 5:25 PM UTC
- GPT-OSS 120BMerge GatewayInput / 1M$0.09$0.05▼ 44%Output / 1M$0.36$0.25▼ 31%Changed: Oct 2, 2026, 5:25 PM UTC
- Z.ai: GLM Flash LatestKilo GatewayInput / 1M$0.026Output / 1M$0.625$0.90▲ 44%Changed: Oct 2, 2026, 5:25 PM UTC
- Hy4 previewKilo GatewayInput / 1M$0.834$0.751▼ 10%Output / 1M$2.50$2.25▼ 10%Changed: Oct 2, 2026, 5:25 PM UTC
- Hy3Kilo GatewayInput / 1M$0.13$0.083▼ 37%Output / 1M$0.53$0.33▼ 38%Changed: Oct 2, 2026, 5:25 PM UTC
- MoonshotAI: Kimi LatestKilo GatewayInput / 1M$1.39$1.34▼ 4%Output / 1M$13.00Changed: Oct 2, 2026, 5:25 PM UTC
- Nano BananaKilo GatewayInput / 1M$0.15$0.30▲ 100%Output / 1M$1.25$2.50▲ 100%Changed: Oct 2, 2026, 5:25 PM UTC
- DeepSeek: DeepSeek V4 Flash LatestKilo GatewayInput / 1M$0.003825$0.013▲ 235%Output / 1M$1.60Changed: Oct 2, 2026, 5:25 PM UTC
- DeepSeek: DeepSeek Flash LatestKilo GatewayInput / 1M$0.03$0.015▼ 50%Output / 1M$0.75$1.20▲ 60%Changed: Oct 2, 2026, 5:25 PM UTC
- Llama-3.3-70B-Instruct (Scaleway)Eden AIInput / 1M$1.02$1.01▼ 1%Output / 1M$1.02$1.01▼ 1%Changed: Oct 2, 2026, 5:25 PM UTC
- GPT OSS 120B (Scaleway)Eden AIInput / 1M$0.169$0.168▼ 1%Output / 1M$0.678$0.674▼ 1%Changed: Oct 2, 2026, 5:25 PM UTC
- DeepSeek V4 Flash 0731 (Scaleway)Eden AIInput / 1M$0.452$0.449▼ 1%Output / 1M$0.904$0.898▼ 1%Changed: Oct 2, 2026, 5:25 PM UTC
- GPT OSS 120B (IONOS)Eden AIInput / 1M$0.169$0.168▼ 1%Output / 1M$0.734$0.73▼ 1%Changed: Oct 2, 2026, 5:25 PM UTC
- Llama-3.3-70B-Instruct (IONOS)Eden AIInput / 1M$0.734$0.73▼ 1%Output / 1M$0.734$0.73▼ 1%Changed: Oct 2, 2026, 5:25 PM UTC
- Ministral 3 14B (Infomaniak)Eden AIInput / 1M$0.339$0.337▼ 1%Output / 1M$0.452$0.449▼ 1%Changed: Oct 2, 2026, 5:25 PM UTC
- DeepSeek V4 ProDeepSeekInput / 1M$0.435$0.66▲ 52%Output / 1M$0.87$1.98▲ 128%Changed: Oct 2, 2026, 5:25 PM UTC
- MiMo V2.6 FlashVercel AI GatewayInput / 1M$0.14$0.04▼ 71%Output / 1M$0.28$1.28▲ 357%Changed: Oct 2, 2026, 6:26 AM UTC
- GPT-5.5 (EU)RequestyInput / 1M$5.00$5.50▲ 10%Output / 1M$30.00$33.00▲ 10%Changed: Oct 2, 2026, 6:26 AM UTC
- GPT-5.4 (EU)RequestyInput / 1M$2.50$2.75▲ 10%Output / 1M$15.00$16.50▲ 10%Changed: Oct 2, 2026, 6:26 AM UTC
- DeepSeek V4.1 Flash (Consensus Protocol)LLM GatewayInput / 1M$0.15Output / 1M$0.60$0.55▼ 8%Changed: Oct 2, 2026, 6:26 AM UTC
- DeepSeek V4.1 FlashDevPass (LLM Gateway)Input / 1M$0.15Output / 1M$0.60$0.55▼ 8%Changed: Oct 2, 2026, 6:26 AM UTC
- Z.ai: GLM LatestKilo GatewayInput / 1M$0.133$0.12▼ 10%Output / 1M$4.00Changed: Oct 2, 2026, 6:26 AM UTC
- Z.ai: GLM Flash LatestKilo GatewayInput / 1M$0.02$0.026▲ 31%Output / 1M$0.248$0.625▲ 153%Changed: Oct 2, 2026, 6:26 AM UTC
- MoonshotAI: Kimi LatestKilo GatewayInput / 1M$0.58$1.39▲ 140%Output / 1M$10.00$13.00▲ 30%Changed: Oct 2, 2026, 6:26 AM UTC
- Kimi K3Kilo GatewayInput / 1M$0.58$2.70▲ 366%Output / 1M$10.00$13.50▲ 35%Changed: Oct 2, 2026, 6:26 AM UTC
- Kimi K2.6Kilo GatewayInput / 1M$0.65$0.434▼ 33%Output / 1M$3.41$1.83▼ 46%Changed: Oct 2, 2026, 6:26 AM UTC
- DeepSeek: DeepSeek V4 Flash LatestKilo GatewayInput / 1M$0.004455$0.003825▼ 14%Output / 1M$0.131$1.60▲ 1124%Changed: Oct 2, 2026, 6:26 AM UTC
- DeepSeek: DeepSeek Pro LatestKilo GatewayInput / 1M$0.291$0.132▼ 55%Output / 1M$3.50$0.396▼ 89%Changed: Oct 2, 2026, 6:26 AM UTC
- DeepSeek: DeepSeek Flash LatestKilo GatewayInput / 1M$0.016$0.03▲ 93%Output / 1M$0.396$0.75▲ 89%Changed: Oct 2, 2026, 6:26 AM UTC
How this works
- Every price, limit and date comes from models.dev, an open-source (MIT-licensed) catalogue of AI models kept in public on GitHub and compiled from providers’ official pricing and documentation pages.
- We read it at most every 6 hours. Providers change prices without notice, so confirm on the provider’s own page (linked on each model) before you rely on a number.
- Prices are US dollars per 1 million tokens, exactly as the catalogue lists them. The blended price, (3 × input + output) ÷ 4, is our own estimate, used only to sort and compare.
- Anything the catalogue does not state is shown as “—”; we never guess. Flat-fee subscription plans, local-only runtimes and deprecated models are left out, and $0 offers are hidden unless you include them.
- No benchmark scores yet. When we add them, they will come only from official sources and be labelled as such.
Sources: models.dev · api.json · GitHub · MIT License
Wrong price, or a provider that wants its listing corrected or removed? Email ceo@uniholix.com. We reply within 5 business days.