NVIDIA
Available in Shortly AI
Nemotron 3 Nano 30B A3B
Released December 15, 2025
ReasoningTool Use
NVIDIA Nemotron 3 Nano is an open reasoning model optimized for fast, cost-efficient inference. Built with a hybrid MoE and Mamba architecture and trained on NVIDIA-curated synthetic reasoning data, it delivers strong multi-step reasoning with stable latency and predictable performance for agentic and production workloads.
- Context window
- 262.1K tokens
- Maximum output
- 262.1K tokens
- Input pricing
- $0.05/M
- Output pricing
- $0.24/M
FAQ
Similar AI models
GPT 5.6 LunaOpenAIGPT-5.6 Luna is a fast, affordable GPT-5.6 model that brings strong capability at the lowest cost in the series.Hy3TencentTencent Hy3 is an open-source Mixture-of-Experts (MoE) large language model developed by Tencent's Hunyuan team. It has 295B total parameters, 21B active parameters, a 3.8B-parameter MTP layer, and a 256K context window. Hy3 is designed for agentic workflows, coding, reasoning, long-context understanding, tool-use scenarios, document analysis, and structured generation.
The model improves on Hy3 Preview with stronger tool-call reliability, better output-format stability, improved long-context retention, and lower hallucination rates in Tencent's internal evaluations.DeepSeek V4 FlashDeepSeekDeepseek V4 Flash is an AI model available through Shortly AI.
Available with one subscription
Start with Nemotron 3 Nano 30B A3B in Shortly AI
Keep your conversations, files and model choices together in one workspace.