Toggle navigation

NVIDIA

Available in Shortly AI

Nemotron 3 Ultra

Released June 4, 2026

ReasoningTool Use

A 550B parameter (55B active) open reasoning model from NVIDIA, built for long-running agent workflows. It uses a hybrid Mamba-Transformer MoE architecture and supports a 1M token context window.

Context window
1M tokens
Maximum output
65K tokens
Input pricing
$0.60/M
Output pricing
$2.40/M

Independent benchmark

LiveBench historical snapshot

51.8overall

Historical benchmark data from LiveBench, release 2026-01-08. Scores are pinned and are not fetched on page load.

Nemotron 3 Ultra LiveBench category scores
CategoryScore
Reasoning37.543
Coding71.341
Agentic Coding46.667
Mathematics54.518
Data Analysis42.014
Language52.180
Instruction Following58.192

Benchmarks measure selected evaluation tasks and do not guarantee performance for every prompt or workflow.

FAQ

Similar AI models

Available with one subscription

Start with Nemotron 3 Ultra in Shortly AI

Keep your conversations, files and model choices together in one workspace.