nvidia
NVIDIA: Nemotron Nano 9B V2 (free) API Pricing
NVIDIA: Nemotron Nano 9B V2 (free) costs Free per 1M input tokens and Free per 1M output tokens. Context window: 128K. Prices verified 2026-08-23 · source
Input / 1M tokens
Free
Output / 1M tokens
Free
Context window
128K
Max output
—
What would NVIDIA: Nemotron Nano 9B V2 (free) cost you per month?
| Workload | Requests/mo | Avg tokens (in/out) | Est. monthly cost |
|---|---|---|---|
| Light chatbot | 10,000 | 800 / 400 | $0.00 |
| Coding agent | 5,000 | 8,000 / 2,000 | $0.00 |
| RAG pipeline | 30,000 | 4,000 / 500 | $0.00 |
| Doc summarizer | 2,000 | 20,000 / 1,000 | $0.00 |
Estimates use list prices without prompt caching. Caching can cut input cost by up to 90% on repeated context.
About NVIDIA: Nemotron Nano 9B V2 (free)
NVIDIA-Nemotron-Nano-9B-v2 is a large language model (LLM) trained from scratch by NVIDIA, and designed as a unified model for both reasoning and non-reasoning tasks. It responds to user queries and...