NVIDIA: Nemotron 3 Nano 30B A3B
Provider: nvidia · nvidia/nemotron-3-nano-30b-a3b
How NVIDIA: Nemotron 3 Nano 30B A3B pricing compares
At $0.2 / 1M output tokens, NVIDIA: Nemotron 3 Nano 30B A3B is the 2nd-cheapest of 5 nvidia models we track. It's cheaper than 92% of the 392 paid models on computetrail (34th-cheapest overall). Output costs 4.0× its input price ($0.05 / 1M in). For a typical 3:1 input-to-output workload that works out to about $0.09 / 1M blended. Reading cached input costs 40% less than fresh input, so prompt caching cuts repeat-context cost sharply. Its 262,144 tokens context window ranks 203rd-largest of 392 models with a published limit.
Output price history
Across 60 daily snapshots since 2026-06-29, the output price has held steady at $0.2 / 1M — no change so far.
Similar-priced models
| Model | Provider | Output $/1M | Context |
|---|---|---|---|
| NVIDIA: Nemotron 3.5 Lightning | nvidia | $0.2 | 262,144 |
| Mistral: Mistral Small 3.2 24B | mistralai | $0.2 | 131,072 |
| Meta: Llama 3.2 1B Instruct | meta-llama | $0.201 | 60,000 |
| Qwen: Qwen3 14B | qwen | $0.24 | 131,072 |
| Amazon: Nova Lite 1.0 | amazon | $0.24 | 300,000 |
| DeepSeek: DeepSeek V4 Flash 0423 | deepseek | $0.159 | 1,048,576 |
Prices are recorded daily from public provider data via OpenRouter and normalized to USD per 1M tokens. See methodology.