MModel Price Map

← All models

NanoGPT

Nvidia Nemotron 70b

nvidia/Llama-3.1-Nemotron-70B-Instruct-HF

Prices and limits

Input price per 1M tokens
$0.357
Output price per 1M tokens
$0.408
Cached input per 1M tokens
$0.179
Blended price per 1M (our estimate)
$0.37
Context window
16,384
Max output tokens
8,192
Accepts
Text
Produces
Text
Released
Apr 15, 2025
Last updated
Apr 15, 2025
Knowledge cutoff
—
Open weights
Yes
Confirm on the provider’s page ↗

Wrong price, or a provider that wants its listing corrected or removed? Email ceo@uniholix.com. We reply within 5 business days.

Price changes

Models whose price changed in the catalogue, newest first. We compare prices every day.

No price changes recorded yet. When a model’s price changes, it appears here.