NVIDIA
Available in Shortly AI
Nvidia Nemotron Nano 9B V2
Released August 18, 2025
ReasoningTool Use
NVIDIA-Nemotron-Nano-9B-v2 is a large language model (LLM) trained from scratch by NVIDIA, and designed as a unified model for both reasoning and non-reasoning tasks. It responds to user queries and tasks by first generating a reasoning trace and then concluding with a final response. The model's reasoning capabilities can be controlled via a system prompt. If the user prefers the model to provide its final answer without intermediate reasoning traces, it can be configured to do so.\
- Context window
- 131.1K tokens
- Maximum output
- 131.1K tokens
- Input pricing
- $0.06/M
- Output pricing
- $0.23/M
FAQ
Similar AI models
GPT 5.6 LunaOpenAIGPT-5.6 Luna is a fast, affordable GPT-5.6 model that brings strong capability at the lowest cost in the series.Hy3TencentTencent Hy3 is an open-source Mixture-of-Experts (MoE) large language model developed by Tencent's Hunyuan team. It has 295B total parameters, 21B active parameters, a 3.8B-parameter MTP layer, and a 256K context window. Hy3 is designed for agentic workflows, coding, reasoning, long-context understanding, tool-use scenarios, document analysis, and structured generation.
The model improves on Hy3 Preview with stronger tool-call reliability, better output-format stability, improved long-context retention, and lower hallucination rates in Tencent's internal evaluations.DeepSeek V4 FlashDeepSeekDeepseek V4 Flash is an AI model available through Shortly AI.
Available with one subscription
Start with Nvidia Nemotron Nano 9B V2 in Shortly AI
Keep your conversations, files and model choices together in one workspace.