InferenceHunt / Explore

NVIDIA AI model pricing

Compare live API pricing for NVIDIA models across tracked inference venues. Prices show the lowest currently observed input and output rate for each model.

2 models with current data

Which NVIDIA model has the lowest API price?

Nemotron 3 Ultra has the lowest observed input price on this page at $0 per 1M tokens. Nemotron 3 Ultra has the lowest observed output price at $0 per 1M tokens. Prices reflect the latest successful collection.

Read the pricing methodology and source policy.

ModelInput / 1MOutput / 1MContextVenues
Nemotron 3 UltraNVIDIA$0$01MOpenRouterOpenCode ZenSurplus Intelligence
Nemotron 3.5 LightningNVIDIA$0$01MOpenRouterOpenCode Zen