Which AI model is cheapest right now?
Prices per 1M tokens, context windows and input types for 7,508 model offers from 199 providers, read from a public open catalogue every 6 hours.
301–350 of 6,826
- The Drummer Magidonia 24B v4.3NanoGPTInput / 1M$0.10Output / 1M$0.121Context32.8K
- Mistral Small 3.2HeliconeInput / 1M$0.075Output / 1M$0.20Context128K
- Granite 4.2 8BDevPass (LLM Gateway)Input / 1M$0.06Output / 1M$0.25Context131.1K
- Granite 4.2 8B (DeepInfra)LLM GatewayInput / 1M$0.06Output / 1M$0.25Context131.1K
- IBM: Granite 4.2 8BKilo GatewayInput / 1M$0.06Output / 1M$0.25Context131.1K
- Nemotron 3 Nano 30B A3B FP8InfomaniakInput / 1M$0.06Output / 1M$0.25Context1M
- GLM 5.3-FlashCrofAIInput / 1M$0.07Output / 1M$0.22Context1M
- Qwen3.6 35B-A3BengyInput / 1M$0.045Output / 1M$0.30Context208.2K
- DeepSeek V4 Flash Vision ExpCrofAIInput / 1M$0.08Output / 1M$0.20Context1M
- Gemma 4 26B-A4BEmpirioLabs AIInput / 1M$0.05Output / 1M$0.29Context262.1K
- GPT-5 NanoOfoxInput / 1M$0.04Output / 1M$0.32Context400K
- Mistral-7B-Instruct-v0.3OVHcloud AI EndpointsInput / 1M$0.11Output / 1M$0.11Context65.5K
- Nvidia Nemotron 3.5 Lightning TEENanoGPTInput / 1M$0.08Output / 1M$0.20Context262.1K
- DeepSeek V4 Flash 0731CortecsInput / 1M$0.09Output / 1M$0.17Context1M
- Nova Lite (APAC)Amazon BedrockInput / 1M$0.063Output / 1M$0.252Context300K
- mistral-7b-instruct-v0.3CortecsInput / 1M$0.111Output / 1M$0.111Context127K
- inclusionAI: Ling 3.0 Flash FinKilo GatewayInput / 1M$0.075Output / 1M$0.22Context262.1K
- inclusionAI: Ling 3.0 Flash VLKilo GatewayInput / 1M$0.075Output / 1M$0.22Context262.1K
- Ling 3.0 FlashNanoGPTInput / 1M$0.075Output / 1M$0.22Context262.1K
- Ling 3.0 Flash FinVercel AI GatewayInput / 1M$0.075Output / 1M$0.22Context256K
- Ling 3.0 Flash ThinkingNanoGPTInput / 1M$0.075Output / 1M$0.22Context262.1K
- Ling 3.0 Flash VLVercel AI GatewayInput / 1M$0.075Output / 1M$0.22Context256K
- Ling 3.1 FlashNanoGPTInput / 1M$0.075Output / 1M$0.22Context262.1K
- Qwen3-Coder 30B-A3B InstructCortecsInput / 1M$0.067Output / 1M$0.245Context262.1K
- Nova Lite (CA)Amazon BedrockInput / 1M$0.064Output / 1M$0.256Context300K
- DeepSeek V4 FlashDeep InfraInput / 1M$0.09Output / 1M$0.18Context1M
- Laguna S 2.1VultrInput / 1M$0.09Output / 1M$0.18Context1M
- Granite 4.2 8BNanoGPTInput / 1M$0.10Output / 1M$0.15Context131.1K
- Granite 4.2 8BCoreWeaveInput / 1M$0.10Output / 1M$0.15Context131.1K
- Qwen 3.5 9BVenice AIInput / 1M$0.10Output / 1M$0.15Context256K
- Qwen/Qwen3.5-9BSiliconFlowInput / 1M$0.10Output / 1M$0.15Context262.1K
- Qwen3.5 9BDeep InfraInput / 1M$0.10Output / 1M$0.15Context262.1K
- Qwen3.5 9BKilo GatewayInput / 1M$0.10Output / 1M$0.15Context256K
- Qwen3.5 9BDevPass (LLM Gateway)Input / 1M$0.10Output / 1M$0.15Context262.1K
- Qwen3.5 9B (DeepInfra)LLM GatewayInput / 1M$0.10Output / 1M$0.15Context262.1K
- DeepSeek V4 Flash FlexNeuralwattInput / 1M$0.091Output / 1M$0.182Context1M
- Qwen: Qwen3.5-FlashKilo GatewayInput / 1M$0.065Output / 1M$0.26Context1M
- Qwen3.8 27BengyInput / 1M$0.045Output / 1M$0.32Context1M
- Granite-4.0-H-Smallwatsonx.aiInput / 1M$0.064Output / 1M$0.265Context131.1K
- Tencent Hy3NanoGPTInput / 1M$0.066Output / 1M$0.26Context262.1K
- Qwen3-Coder 30B-A3B InstructHugging FaceInput / 1M$0.07Output / 1M$0.26Context262.1K
- Qwen3-Coder-30B-A3B-InstructOVHcloud AI EndpointsInput / 1M$0.07Output / 1M$0.26Context262.1K
- GPT OSS 120BDInferenceInput / 1M$0.068Output / 1M$0.27Context131.1K
- GLM 5.3 FlashEmpirioLabs AIInput / 1M$0.075Output / 1M$0.25Context1M
- GLM-5.3 FlashMerge GatewayInput / 1M$0.075Output / 1M$0.25Context1M
- GLM-5.3-Flash302.AIInput / 1M$0.075Output / 1M$0.25Context1M
- GLM-5.3-FlashOrcaRouterInput / 1M$0.075Output / 1M$0.25Context1M
- Nex AGI: Nex-N2.5-ProKilo GatewayInput / 1M$0.075Output / 1M$0.25Context262.1K
- Qwen 3 14bNanoGPTInput / 1M$0.08Output / 1M$0.24Context41K
- Meta: Llama 3.2 3B InstructKilo GatewayInput / 1M$0.05Output / 1M$0.33Context131.1K
Price changes
Models whose price changed in the catalogue, newest first. We compare prices every day.
- Z.ai: GLM Flash LatestKilo GatewayInput / 1M$0.026Output / 1M$0.90$0.929▲ 3%Changed: Oct 2, 2026, 8:48 PM UTC
- MiMo-V2.6-FlashKilo GatewayInput / 1M$0.07$0.12▲ 71%Output / 1M$0.28Changed: Oct 2, 2026, 8:48 PM UTC
- Hy4 previewKilo GatewayInput / 1M$0.751$0.83▲ 11%Output / 1M$2.25$2.50▲ 11%Changed: Oct 2, 2026, 8:48 PM UTC
- MoonshotAI: Kimi LatestKilo GatewayInput / 1M$1.34$1.00▼ 25%Output / 1M$13.00$11.36▼ 13%Changed: Oct 2, 2026, 8:48 PM UTC
- DeepSeek: DeepSeek V4 Flash LatestKilo GatewayInput / 1M$0.005775$0.003825▼ 34%Output / 1M$1.47$1.04▼ 29%Changed: Oct 2, 2026, 8:48 PM UTC
- DeepSeek: DeepSeek V4 Flash LatestKilo GatewayInput / 1M$0.013$0.005775▼ 55%Output / 1M$1.60$1.47▼ 8%Changed: Oct 2, 2026, 7:19 PM UTC
- DeepSeek: DeepSeek Flash LatestKilo GatewayInput / 1M$0.015$0.02▲ 33%Output / 1M$1.20$0.60▼ 50%Changed: Oct 2, 2026, 7:19 PM UTC
- Claude Sonnet 5.5Venice AIInput / 1M$3.75$2.50▼ 33%Output / 1M$18.75$12.50▼ 33%Changed: Oct 2, 2026, 5:25 PM UTC
- Claude Sonnet 5Venice AIInput / 1M$3.00$2.50▼ 17%Output / 1M$15.00$12.50▼ 17%Changed: Oct 2, 2026, 5:25 PM UTC
- GLM 4.7 ThinkingNanoGPTInput / 1M$0.20$0.40▲ 100%Output / 1M$0.80$1.93▲ 141%Changed: Oct 2, 2026, 5:25 PM UTC
- GLM 4.7NanoGPTInput / 1M$0.20$0.40▲ 100%Output / 1M$0.80$1.93▲ 141%Changed: Oct 2, 2026, 5:25 PM UTC
- GPT-OSS 20BMerge GatewayInput / 1M$0.04Output / 1M$0.20$0.15▼ 25%Changed: Oct 2, 2026, 5:25 PM UTC
- GPT-OSS 120BMerge GatewayInput / 1M$0.09$0.05▼ 44%Output / 1M$0.36$0.25▼ 31%Changed: Oct 2, 2026, 5:25 PM UTC
- Z.ai: GLM Flash LatestKilo GatewayInput / 1M$0.026Output / 1M$0.625$0.90▲ 44%Changed: Oct 2, 2026, 5:25 PM UTC
- Hy4 previewKilo GatewayInput / 1M$0.834$0.751▼ 10%Output / 1M$2.50$2.25▼ 10%Changed: Oct 2, 2026, 5:25 PM UTC
- Hy3Kilo GatewayInput / 1M$0.13$0.083▼ 37%Output / 1M$0.53$0.33▼ 38%Changed: Oct 2, 2026, 5:25 PM UTC
- MoonshotAI: Kimi LatestKilo GatewayInput / 1M$1.39$1.34▼ 4%Output / 1M$13.00Changed: Oct 2, 2026, 5:25 PM UTC
- Nano BananaKilo GatewayInput / 1M$0.15$0.30▲ 100%Output / 1M$1.25$2.50▲ 100%Changed: Oct 2, 2026, 5:25 PM UTC
- DeepSeek: DeepSeek V4 Flash LatestKilo GatewayInput / 1M$0.003825$0.013▲ 235%Output / 1M$1.60Changed: Oct 2, 2026, 5:25 PM UTC
- DeepSeek: DeepSeek Flash LatestKilo GatewayInput / 1M$0.03$0.015▼ 50%Output / 1M$0.75$1.20▲ 60%Changed: Oct 2, 2026, 5:25 PM UTC
- Llama-3.3-70B-Instruct (Scaleway)Eden AIInput / 1M$1.02$1.01▼ 1%Output / 1M$1.02$1.01▼ 1%Changed: Oct 2, 2026, 5:25 PM UTC
- GPT OSS 120B (Scaleway)Eden AIInput / 1M$0.169$0.168▼ 1%Output / 1M$0.678$0.674▼ 1%Changed: Oct 2, 2026, 5:25 PM UTC
- DeepSeek V4 Flash 0731 (Scaleway)Eden AIInput / 1M$0.452$0.449▼ 1%Output / 1M$0.904$0.898▼ 1%Changed: Oct 2, 2026, 5:25 PM UTC
- GPT OSS 120B (IONOS)Eden AIInput / 1M$0.169$0.168▼ 1%Output / 1M$0.734$0.73▼ 1%Changed: Oct 2, 2026, 5:25 PM UTC
- Llama-3.3-70B-Instruct (IONOS)Eden AIInput / 1M$0.734$0.73▼ 1%Output / 1M$0.734$0.73▼ 1%Changed: Oct 2, 2026, 5:25 PM UTC
- Ministral 3 14B (Infomaniak)Eden AIInput / 1M$0.339$0.337▼ 1%Output / 1M$0.452$0.449▼ 1%Changed: Oct 2, 2026, 5:25 PM UTC
- DeepSeek V4 ProDeepSeekInput / 1M$0.435$0.66▲ 52%Output / 1M$0.87$1.98▲ 128%Changed: Oct 2, 2026, 5:25 PM UTC
- MiMo V2.6 FlashVercel AI GatewayInput / 1M$0.14$0.04▼ 71%Output / 1M$0.28$1.28▲ 357%Changed: Oct 2, 2026, 6:26 AM UTC
- GPT-5.5 (EU)RequestyInput / 1M$5.00$5.50▲ 10%Output / 1M$30.00$33.00▲ 10%Changed: Oct 2, 2026, 6:26 AM UTC
- GPT-5.4 (EU)RequestyInput / 1M$2.50$2.75▲ 10%Output / 1M$15.00$16.50▲ 10%Changed: Oct 2, 2026, 6:26 AM UTC
- DeepSeek V4.1 Flash (Consensus Protocol)LLM GatewayInput / 1M$0.15Output / 1M$0.60$0.55▼ 8%Changed: Oct 2, 2026, 6:26 AM UTC
- DeepSeek V4.1 FlashDevPass (LLM Gateway)Input / 1M$0.15Output / 1M$0.60$0.55▼ 8%Changed: Oct 2, 2026, 6:26 AM UTC
- Z.ai: GLM LatestKilo GatewayInput / 1M$0.133$0.12▼ 10%Output / 1M$4.00Changed: Oct 2, 2026, 6:26 AM UTC
- Z.ai: GLM Flash LatestKilo GatewayInput / 1M$0.02$0.026▲ 31%Output / 1M$0.248$0.625▲ 153%Changed: Oct 2, 2026, 6:26 AM UTC
- MoonshotAI: Kimi LatestKilo GatewayInput / 1M$0.58$1.39▲ 140%Output / 1M$10.00$13.00▲ 30%Changed: Oct 2, 2026, 6:26 AM UTC
- Kimi K3Kilo GatewayInput / 1M$0.58$2.70▲ 366%Output / 1M$10.00$13.50▲ 35%Changed: Oct 2, 2026, 6:26 AM UTC
- Kimi K2.6Kilo GatewayInput / 1M$0.65$0.434▼ 33%Output / 1M$3.41$1.83▼ 46%Changed: Oct 2, 2026, 6:26 AM UTC
- DeepSeek: DeepSeek V4 Flash LatestKilo GatewayInput / 1M$0.004455$0.003825▼ 14%Output / 1M$0.131$1.60▲ 1124%Changed: Oct 2, 2026, 6:26 AM UTC
- DeepSeek: DeepSeek Pro LatestKilo GatewayInput / 1M$0.291$0.132▼ 55%Output / 1M$3.50$0.396▼ 89%Changed: Oct 2, 2026, 6:26 AM UTC
- DeepSeek: DeepSeek Flash LatestKilo GatewayInput / 1M$0.016$0.03▲ 93%Output / 1M$0.396$0.75▲ 89%Changed: Oct 2, 2026, 6:26 AM UTC
How this works
- Every price, limit and date comes from models.dev, an open-source (MIT-licensed) catalogue of AI models kept in public on GitHub and compiled from providers’ official pricing and documentation pages.
- We read it at most every 6 hours. Providers change prices without notice, so confirm on the provider’s own page (linked on each model) before you rely on a number.
- Prices are US dollars per 1 million tokens, exactly as the catalogue lists them. The blended price, (3 × input + output) ÷ 4, is our own estimate, used only to sort and compare.
- Anything the catalogue does not state is shown as “—”; we never guess. Flat-fee subscription plans, local-only runtimes and deprecated models are left out, and $0 offers are hidden unless you include them.
- No benchmark scores yet. When we add them, they will come only from official sources and be labelled as such.
Sources: models.dev · api.json · GitHub · MIT License
Wrong price, or a provider that wants its listing corrected or removed? Email ceo@uniholix.com. We reply within 5 business days.