computetrail

All models ·nvidia

NVIDIA: Nemotron 3.5 Lightning

Provider: nvidia · nvidia/nemotron-3.5-lightning

Input
$0.08 /1M
Output
$0.2 /1M
Cache read
$0.04 /1M
Context
262,144 tokens

How NVIDIA: Nemotron 3.5 Lightning pricing compares

At $0.2 / 1M output tokens, NVIDIA: Nemotron 3.5 Lightning is the 1st-cheapest of 5 nvidia models we track. It's cheaper than 92% of the 392 paid models on computetrail (33rd-cheapest overall). Output costs 2.5× its input price ($0.08 / 1M in). For a typical 3:1 input-to-output workload that works out to about $0.11 / 1M blended. Reading cached input costs 50% less than fresh input, so prompt caching cuts repeat-context cost sharply. Its 262,144 tokens context window ranks 169th-largest of 392 models with a published limit.

Output price history

$0.250$0.20008-1208-2008-27

Across 16 daily snapshots since 2026-08-12, the output price has fallen 20.0% — from $0.25 to $0.2 / 1M.

Similar-priced models

ModelProviderOutput $/1MContext
NVIDIA: Nemotron 3 Nano 30B A3Bnvidia$0.2262,144
Mistral: Mistral Small 3.2 24Bmistralai$0.2131,072
Meta: Llama 3.2 1B Instructmeta-llama$0.20160,000
Qwen: Qwen3 14Bqwen$0.24131,072
Amazon: Nova Lite 1.0amazon$0.24300,000
DeepSeek: DeepSeek V4 Flash 0423deepseek$0.1591,048,576

Prices are recorded daily from public provider data via OpenRouter and normalized to USD per 1M tokens. See methodology.