Nvidia
NVIDIA: Nemotron Nano 9B V2
NVIDIA-Nemotron-Nano-9B-v2 is a large language model (LLM) trained from scratch by NVIDIA, and designed as a unified model for both reasoning and non-reasoning tasks. It responds to user queries and tasks by first generating a reasoning trace and then concluding with a final response. The model's reasoning capabilities can be controlled via a system prompt. If the user prefers the model to provide its final answer without intermediate reasoning traces, it can be configured to do so.
| Specifications | |
|---|---|
| Provider | Nvidia |
| Input pricing | $0.04/M tokens |
| Output pricing | $0.16/M tokens |
| Context length | 131K tokens |
| Input modalities | text |
| Output modalities | text |
| Tokenizer | Other |
| Knowledge cutoff | 2025-03-31 |
| Content moderated | No |
| HuggingFace | nvidia/NVIDIA-Nemotron-Nano-9B-v2 |
Supported Parameters
frequency_penaltyinclude_reasoningmax_tokensmin_ppresence_penaltyreasoningrepetition_penaltyresponse_formatseedstoptemperaturetool_choicetoolstop_ktop_p