LLM Gateway
Gemma 4 31B IT (NovitaAI)
novita/gemma-4-31b-it
Prices and limits
- Input price per 1M tokens
- $0.14
- Output price per 1M tokens
- $0.40
- Cached input per 1M tokens
- —
- Blended price per 1M (our estimate)
- $0.205
- Context window
- 262,144
- Max output tokens
- 32,768
- Accepts
- Text
- Produces
- Text
- Released
- Apr 2, 2026
- Last updated
- Apr 2, 2026
- Knowledge cutoff
- —
- Open weights
- Yes
Wrong price, or a provider that wants its listing corrected or removed? Email ceo@uniholix.com. We reply within 5 business days.
Cheapest comparable models
Other priced models that accept at least the same inputs, produce the same outputs and have at least half this context window, sorted by blended price.
- GPT OSS 20B (Greenference)Eden AI$0.009 / $0.045131.1K
- Llama-3.1-8B-InstructKilo Gateway$0.02 / $0.04131.1K
- Mistral Nemo Instruct 2407Meganova$0.02 / $0.04131.1K
- Ling 3.0 FlashVercel AI Gateway$0.021 / $0.063256K
- Llama 3.2 3b InstructNanoGPT$0.031 / $0.049131.1K
- qwen3.5-2bRequesty$0.02 / $0.10262.1K
- Mistral Nemo Instruct 2407 TEEChutes$0.025 / $0.098131.1K
- Nemotron 3 Nano Omni 30B TEEChutes$0.025 / $0.098131.1K
Same model at other providers
Matched by the catalogue’s model ID. Cheapest blended price first.
- Gemma 4 31B ITKenari$0.00 / $0.00262.1K
- Gemma 4 31B ITQVAC$0.00 / $0.00262.1K
- Gemma 4 31B ITRequesty$0.00 / $0.00262.1K
- Gemma 4 31B ITUnoRouter$0.00 / $0.00262.1K
- Gemma 4 31B IT (free)Bothub$0.00 / $0.00262.1K
- Gemma 4 31B IT FP8InferX$0.00 / $0.00262.1K
- Gemma-4-31B-ITNvidia$0.00 / $0.00256K
- Gemma 4 31B ITDevPass (LLM Gateway)$0.10 / $0.25262.1K
- Gemma 4 31B ITCrofAI$0.10 / $0.30262.1K
- Gemma 4 31B ITKilo Gateway$0.09 / $0.34262.1K
- Gemma 4 31BCoreWeave$0.10 / $0.34262.1K
- Gemma 4 31B ThinkingNanoGPT$0.10 / $0.35262.1K
Price changes
Models whose price changed in the catalogue, newest first. We compare prices every day.
No price changes recorded yet. When a model’s price changes, it appears here.