LLMCost
nvidia

NVIDIA: Nemotron Nano 9B V2 (free) API Pricing

NVIDIA: Nemotron Nano 9B V2 (free) costs Free per 1M input tokens and Free per 1M output tokens. Context window: 128K. Prices verified 2026-08-23 · source

Input / 1M tokens
Free
Output / 1M tokens
Free
Context window
128K
Max output

What would NVIDIA: Nemotron Nano 9B V2 (free) cost you per month?

WorkloadRequests/moAvg tokens (in/out)Est. monthly cost
Light chatbot10,000800 / 400$0.00
Coding agent5,0008,000 / 2,000$0.00
RAG pipeline30,0004,000 / 500$0.00
Doc summarizer2,00020,000 / 1,000$0.00

Estimates use list prices without prompt caching. Caching can cut input cost by up to 90% on repeated context.

About NVIDIA: Nemotron Nano 9B V2 (free)

NVIDIA-Nemotron-Nano-9B-v2 is a large language model (LLM) trained from scratch by NVIDIA, and designed as a unified model for both reasoning and non-reasoning tasks. It responds to user queries and...

Compare NVIDIA: Nemotron Nano 9B V2 (free) with