Which AI model is cheapest right now?
Prices per 1M tokens, context windows and input types for 7,508 model offers from 199 providers, read from a public open catalogue every 6 hours.
1–50 of 6,826
- Llama 3.2 1B InstructInferenceInput / 1M$0.01Output / 1M$0.01Context16K
- Google Gemma 2HeliconeInput / 1M$0.01Output / 1M$0.03Context8.2K
- E5 Mistral 7BSTACKITInput / 1M$0.02Output / 1M$0.02Context4.1K
- Llama 3.2 3B InstructInferenceInput / 1M$0.02Output / 1M$0.02Context16K
- PaddleOCR-VLNovitaAIInput / 1M$0.02Output / 1M$0.02Context16.4K
- Llama 3.1 8B (decentralized)NanoGPTInput / 1M$0.02Output / 1M$0.03Context128K
- Meta Llama 3.1 8B Instruct TurboHeliconeInput / 1M$0.02Output / 1M$0.03Context128K
- Mistral NemoPioneerInput / 1M$0.02Output / 1M$0.03Context128K
- Llama 3.1 8B InstructInferenceInput / 1M$0.025Output / 1M$0.025Context16K
- Llama-3.1-8B-InstructKilo GatewayInput / 1M$0.02Output / 1M$0.04Context131.1K
- Mistral Nemo Instruct 2407IO.NETInput / 1M$0.02Output / 1M$0.04Context128K
- Mistral Nemo Instruct 2407MeganovaInput / 1M$0.02Output / 1M$0.04Context131.1K
- Meta Llama 3.1 8B InstructHeliconeInput / 1M$0.02Output / 1M$0.05Context16.4K
- Llama 3.1 8B InstructAbacusInput / 1M$0.02Output / 1M$0.05Context128K
- Llama 3.1 8B InstructNovitaAIInput / 1M$0.02Output / 1M$0.05Context16.4K
- DeepSeek-OCRNovitaAIInput / 1M$0.03Output / 1M$0.03Context8.2K
- deepseek/deepseek-ocr-2NovitaAIInput / 1M$0.03Output / 1M$0.03Context8.2K
- Qwen3 4BNovitaAIInput / 1M$0.03Output / 1M$0.03Context128K
- Ling 3.0 FlashVercel AI GatewayInput / 1M$0.021Output / 1M$0.063Context256K
- Llama 3.2 3B InstructDevPass (LLM Gateway)Input / 1M$0.03Output / 1M$0.05Context32.8K
- Llama 3.2 3B InstructNovitaAIInput / 1M$0.03Output / 1M$0.05Context32.8K
- Llama 3.2 3B Instruct (NovitaAI)LLM GatewayInput / 1M$0.03Output / 1M$0.05Context32.8K
- Llama 3.2 3b InstructNanoGPTInput / 1M$0.031Output / 1M$0.049Context131.1K
- GPT OSS 20BKilo GatewayInput / 1M$0.018Output / 1M$0.09Context131.1K
- GPT OSS 20B (FlexAI)Eden AIInput / 1M$0.02Output / 1M$0.10Context131.1K
- Llama 3 8B InstructNovitaAIInput / 1M$0.04Output / 1M$0.04Context8.2K
- Ministral 3BAzureInput / 1M$0.04Output / 1M$0.04Context128K
- Ministral 3BAzure Cognitive ServicesInput / 1M$0.04Output / 1M$0.04Context128K
- Ministral 3B (latest)MistralInput / 1M$0.04Output / 1M$0.04Context128K
- Ministral 3B (latest)Vercel AI GatewayInput / 1M$0.04Output / 1M$0.04Context128K
- qwen3.5-2bRequestyInput / 1M$0.02Output / 1M$0.10Context262.1K
- Sarvam 30BFastRouterInput / 1M$0.02Output / 1M$0.10Context128K
- Granite 4.0 H MicroCloudflare Workers AIInput / 1M$0.017Output / 1M$0.112Context131K
- IBM: Granite 4.0 MicroKilo GatewayInput / 1M$0.017Output / 1M$0.112Context131K
- Sao10K: Llama 3 8B LunarisKilo GatewayInput / 1M$0.04Output / 1M$0.05Context8.2K
- Mistral Nemo Instruct 2407 TEEChutesInput / 1M$0.025Output / 1M$0.098Context131.1K
- Nemotron 3 Nano Omni 30B TEEChutesInput / 1M$0.025Output / 1M$0.098Context131.1K
- DeepSeek V4 FlashMerge GatewayInput / 1M$0.035Output / 1M$0.07Context1M
- DeepSeek V4 Flash 0731Merge GatewayInput / 1M$0.035Output / 1M$0.07Context1M
- Nex AGI: Nex-N2.5-MiniKilo GatewayInput / 1M$0.025Output / 1M$0.10Context262.1K
- Qwen2.5 Coder 7B fastHeliconeInput / 1M$0.03Output / 1M$0.09Context32K
- amazon--nova-microSAP AI CoreInput / 1M$0.03Output / 1M$0.10Context128K
- GLM-5.3-FlashPareto InferenceInput / 1M$0.03Output / 1M$0.10Context1M
- Qwen3.5 4BEmpirioLabs AIInput / 1M$0.04Output / 1M$0.07Context262.1K
- Qwen3.7 FlashAIHubMixInput / 1M$0.028Output / 1M$0.113Context991K
- DeepSeek V4.1 FlashengyInput / 1M$0.04Output / 1M$0.08Context327.7K
- Gemma 3 4BMerge GatewayInput / 1M$0.04Output / 1M$0.08Context128K
- Gemma 3 4B ITAmazon BedrockInput / 1M$0.04Output / 1M$0.08Context131.1K
- Gemma 3 4B IT (Amazon Bedrock, US)Eden AIInput / 1M$0.04Output / 1M$0.08Context128K
- Gemma 3 4B IT (Amazon Bedrock)Eden AIInput / 1M$0.04Output / 1M$0.08Context128K
Price changes
Models whose price changed in the catalogue, newest first. We compare prices every day.
- DeepSeek: DeepSeek V4 Flash LatestKilo GatewayInput / 1M$0.013$0.005775▼ 55%Output / 1M$1.60$1.47▼ 8%Changed: Oct 2, 2026, 7:19 PM UTC
- DeepSeek: DeepSeek Flash LatestKilo GatewayInput / 1M$0.015$0.02▲ 33%Output / 1M$1.20$0.60▼ 50%Changed: Oct 2, 2026, 7:19 PM UTC
- Claude Sonnet 5.5Venice AIInput / 1M$3.75$2.50▼ 33%Output / 1M$18.75$12.50▼ 33%Changed: Oct 2, 2026, 5:25 PM UTC
- Claude Sonnet 5Venice AIInput / 1M$3.00$2.50▼ 17%Output / 1M$15.00$12.50▼ 17%Changed: Oct 2, 2026, 5:25 PM UTC
- GLM 4.7 ThinkingNanoGPTInput / 1M$0.20$0.40▲ 100%Output / 1M$0.80$1.93▲ 141%Changed: Oct 2, 2026, 5:25 PM UTC
- GLM 4.7NanoGPTInput / 1M$0.20$0.40▲ 100%Output / 1M$0.80$1.93▲ 141%Changed: Oct 2, 2026, 5:25 PM UTC
- GPT-OSS 20BMerge GatewayInput / 1M$0.04Output / 1M$0.20$0.15▼ 25%Changed: Oct 2, 2026, 5:25 PM UTC
- GPT-OSS 120BMerge GatewayInput / 1M$0.09$0.05▼ 44%Output / 1M$0.36$0.25▼ 31%Changed: Oct 2, 2026, 5:25 PM UTC
- Z.ai: GLM Flash LatestKilo GatewayInput / 1M$0.026Output / 1M$0.625$0.90▲ 44%Changed: Oct 2, 2026, 5:25 PM UTC
- Hy4 previewKilo GatewayInput / 1M$0.834$0.751▼ 10%Output / 1M$2.50$2.25▼ 10%Changed: Oct 2, 2026, 5:25 PM UTC
- Hy3Kilo GatewayInput / 1M$0.13$0.083▼ 37%Output / 1M$0.53$0.33▼ 38%Changed: Oct 2, 2026, 5:25 PM UTC
- MoonshotAI: Kimi LatestKilo GatewayInput / 1M$1.39$1.34▼ 4%Output / 1M$13.00Changed: Oct 2, 2026, 5:25 PM UTC
- Nano BananaKilo GatewayInput / 1M$0.15$0.30▲ 100%Output / 1M$1.25$2.50▲ 100%Changed: Oct 2, 2026, 5:25 PM UTC
- DeepSeek: DeepSeek V4 Flash LatestKilo GatewayInput / 1M$0.003825$0.013▲ 235%Output / 1M$1.60Changed: Oct 2, 2026, 5:25 PM UTC
- DeepSeek: DeepSeek Flash LatestKilo GatewayInput / 1M$0.03$0.015▼ 50%Output / 1M$0.75$1.20▲ 60%Changed: Oct 2, 2026, 5:25 PM UTC
- Llama-3.3-70B-Instruct (Scaleway)Eden AIInput / 1M$1.02$1.01▼ 1%Output / 1M$1.02$1.01▼ 1%Changed: Oct 2, 2026, 5:25 PM UTC
- GPT OSS 120B (Scaleway)Eden AIInput / 1M$0.169$0.168▼ 1%Output / 1M$0.678$0.674▼ 1%Changed: Oct 2, 2026, 5:25 PM UTC
- DeepSeek V4 Flash 0731 (Scaleway)Eden AIInput / 1M$0.452$0.449▼ 1%Output / 1M$0.904$0.898▼ 1%Changed: Oct 2, 2026, 5:25 PM UTC
- GPT OSS 120B (IONOS)Eden AIInput / 1M$0.169$0.168▼ 1%Output / 1M$0.734$0.73▼ 1%Changed: Oct 2, 2026, 5:25 PM UTC
- Llama-3.3-70B-Instruct (IONOS)Eden AIInput / 1M$0.734$0.73▼ 1%Output / 1M$0.734$0.73▼ 1%Changed: Oct 2, 2026, 5:25 PM UTC
- Ministral 3 14B (Infomaniak)Eden AIInput / 1M$0.339$0.337▼ 1%Output / 1M$0.452$0.449▼ 1%Changed: Oct 2, 2026, 5:25 PM UTC
- DeepSeek V4 ProDeepSeekInput / 1M$0.435$0.66▲ 52%Output / 1M$0.87$1.98▲ 128%Changed: Oct 2, 2026, 5:25 PM UTC
- MiMo V2.6 FlashVercel AI GatewayInput / 1M$0.14$0.04▼ 71%Output / 1M$0.28$1.28▲ 357%Changed: Oct 2, 2026, 6:26 AM UTC
- GPT-5.5 (EU)RequestyInput / 1M$5.00$5.50▲ 10%Output / 1M$30.00$33.00▲ 10%Changed: Oct 2, 2026, 6:26 AM UTC
- GPT-5.4 (EU)RequestyInput / 1M$2.50$2.75▲ 10%Output / 1M$15.00$16.50▲ 10%Changed: Oct 2, 2026, 6:26 AM UTC
- DeepSeek V4.1 Flash (Consensus Protocol)LLM GatewayInput / 1M$0.15Output / 1M$0.60$0.55▼ 8%Changed: Oct 2, 2026, 6:26 AM UTC
- DeepSeek V4.1 FlashDevPass (LLM Gateway)Input / 1M$0.15Output / 1M$0.60$0.55▼ 8%Changed: Oct 2, 2026, 6:26 AM UTC
- Z.ai: GLM LatestKilo GatewayInput / 1M$0.133$0.12▼ 10%Output / 1M$4.00Changed: Oct 2, 2026, 6:26 AM UTC
- Z.ai: GLM Flash LatestKilo GatewayInput / 1M$0.02$0.026▲ 31%Output / 1M$0.248$0.625▲ 153%Changed: Oct 2, 2026, 6:26 AM UTC
- MoonshotAI: Kimi LatestKilo GatewayInput / 1M$0.58$1.39▲ 140%Output / 1M$10.00$13.00▲ 30%Changed: Oct 2, 2026, 6:26 AM UTC
- Kimi K3Kilo GatewayInput / 1M$0.58$2.70▲ 366%Output / 1M$10.00$13.50▲ 35%Changed: Oct 2, 2026, 6:26 AM UTC
- Kimi K2.6Kilo GatewayInput / 1M$0.65$0.434▼ 33%Output / 1M$3.41$1.83▼ 46%Changed: Oct 2, 2026, 6:26 AM UTC
- DeepSeek: DeepSeek V4 Flash LatestKilo GatewayInput / 1M$0.004455$0.003825▼ 14%Output / 1M$0.131$1.60▲ 1124%Changed: Oct 2, 2026, 6:26 AM UTC
- DeepSeek: DeepSeek Pro LatestKilo GatewayInput / 1M$0.291$0.132▼ 55%Output / 1M$3.50$0.396▼ 89%Changed: Oct 2, 2026, 6:26 AM UTC
- DeepSeek: DeepSeek Flash LatestKilo GatewayInput / 1M$0.016$0.03▲ 93%Output / 1M$0.396$0.75▲ 89%Changed: Oct 2, 2026, 6:26 AM UTC
- DeepSeek V4.1 FlashCharm HyperInput / 1M$0.30$0.33▲ 10%Output / 1M$1.20$1.31▲ 9%Changed: Oct 2, 2026, 6:26 AM UTC
- DeepSeek Flash LatestFireworks AIInput / 1M$0.22$0.30▲ 36%Output / 1M$0.66$1.20▲ 82%Changed: Oct 2, 2026, 6:26 AM UTC
- DeepSeek V4.1 FlashFireworks AIInput / 1M$0.22$0.30▲ 36%Output / 1M$0.66$1.20▲ 82%Changed: Oct 2, 2026, 6:26 AM UTC
- Llama-3.3-70B-Instruct (Scaleway)Eden AIInput / 1M$1.02$1.02▼ 1%Output / 1M$1.02$1.02▼ 1%Changed: Oct 2, 2026, 6:26 AM UTC
- GPT OSS 120B (Scaleway)Eden AIInput / 1M$0.17$0.169▼ 1%Output / 1M$0.681$0.678▼ 1%Changed: Oct 2, 2026, 6:26 AM UTC
How this works
- Every price, limit and date comes from models.dev, an open-source (MIT-licensed) catalogue of AI models kept in public on GitHub and compiled from providers’ official pricing and documentation pages.
- We read it at most every 6 hours. Providers change prices without notice, so confirm on the provider’s own page (linked on each model) before you rely on a number.
- Prices are US dollars per 1 million tokens, exactly as the catalogue lists them. The blended price, (3 × input + output) ÷ 4, is our own estimate, used only to sort and compare.
- Anything the catalogue does not state is shown as “—”; we never guess. Flat-fee subscription plans, local-only runtimes and deprecated models are left out, and $0 offers are hidden unless you include them.
- No benchmark scores yet. When we add them, they will come only from official sources and be labelled as such.
Sources: models.dev · api.json · GitHub · MIT License
Wrong price, or a provider that wants its listing corrected or removed? Email ceo@uniholix.com. We reply within 5 business days.