NVIDIA-Nemotron-Nano-9B-v2 is a large language model (LLM) trained from scratch by NVIDIA, and designed as a unified model for both reasoning and non-reasoning tasks. It responds to user queries and tasks by first generating a reasoning trace and then concluding with a final response. The model's reasoning capabilities can be controlled via a system prompt. If the user prefers the model to provide its final answer without intermediate reasoning traces, it can be configured to do so.
# Paste into Claude Code, Cursor, or Codex
Help me set up the Respan gateway: fetch and follow https:/
One provider serves this model through the Respan gateway. Copy a model ID to call that route directly.
| Provider | Input | Output |
|---|---|---|
TogetherAIBYOK only together_ai/nvidia/NVIDIA-Nemotron-Nano-9B-v2 | $0.06/M | $0.25/M |
Is a provider down? Together AI status
USD per 1M tokens on TogetherAI, after gateway discounts.
Total cost from 0 to 1T tokens.
Use Nemotron Nano 9B V2 in the tools you already run. Each guide points the app at the Respan gateway, then you pick this model.
Setup guides for more apps are in the integration docs.
Built to meet the security and privacy standards that enterprise and healthcare teams require.
ISO 27001
The internationally recognized standard for information security management.
SOC 2
Secure, compliant management of your data across all of our systems.
GDPR
Operated under GDPR, the world's strictest standard for data privacy.
HIPAA
HIPAA compliant, with a BAA available for healthcare teams.