NVIDIA

Nemotron 3 Nano — Benchmark Scores, Pricing & Performance Analysis

OPEN SOURCENVIDIA
Output Speed
76 tok/s
Context Window
1.0M

Nemotron 3 Nano is an AI model by NVIDIA. View detailed benchmark scores, pricing data, and performance metrics on Serenities AI Models.

Speed Benchmarks

Output Speed
76 tok/s49th of 94

Tokens generated per second

Aggregated
Time to First Token
250 ms19th of 92

Latency before first token arrives

Aggregated

Context Benchmarks

Context Length
1.0M52nd of 451

Maximum context window size

OpenRouter API

Nemotron 3 Nano — Frequently Asked Questions

How fast is Nemotron 3 Nano?

Nemotron 3 Nano generates output at 76 tokens per second, which is slower, prioritizing quality over speed compared to other models. The time to first token is 250 ms.

What is the context window of Nemotron 3 Nano?

Nemotron 3 Nano has a context window of 1.0M tokens. This determines how much text, conversation history, and code the model can process in a single request.

Who created Nemotron 3 Nano?

Nemotron 3 Nano was created by NVIDIA. It is classified as a open source model in our catalogue.

Is Nemotron 3 Nano open source?

Yes, Nemotron 3 Nano is an open-source model. The model weights are publicly available for download and self-hosting.