NVIDIA
Available in Shortly AI
NVIDIA Nemotron 3 Super 120B A12B
Released March 11, 2026
NVIDIA Nemotron 3 Super is a 120B-parameter open hybrid MoE model, activating just 12B parameters for maximum compute efficiency and accuracy in complex multi-agent applications. It delivers up to 7x higher throughput, providing fast, cost-efficient inference for agentic tasks. Additionally, a long context window gives the model long-term memory, preventing AI agents from losing focus on long, multi-step tasks and ensuring high-accuracy results. Fully open with weights, datasets, and recipes, Super allows easy customization and secure deployment anywhere.
- Context window
- 256K tokens
- Maximum output
- 32K tokens
- Input pricing
- $0.15/M
- Output pricing
- $0.65/M
Independent benchmark
LiveBench historical snapshot
Historical benchmark data from LiveBench, release 2026-01-08. Scores are pinned and are not fetched on page load.
| Category | Score |
|---|---|
| Reasoning | 34.389 |
| Coding | 54.072 |
| Agentic Coding | 23.000 |
| Mathematics | 36.429 |
| Data Analysis | 21.231 |
| Language | 30.042 |
| Instruction Following | 28.408 |
Benchmarks measure selected evaluation tasks and do not guarantee performance for every prompt or workflow.
FAQ
Similar AI models
Available with one subscription
Start with NVIDIA Nemotron 3 Super 120B A12B in Shortly AI
Keep your conversations, files and model choices together in one workspace.