Independent community resource — Nemotron3AI.com is not affiliated with, endorsed by, or sponsored by NVIDIA Corporation.

Model family

Three scales. One agentic foundation.

Use this comparison as a starting point, then verify deployment requirements, licenses and current versions on the linked official pages.

01

High-throughput agents

Nano

Available

The smallest member of the family, designed to pair strong reasoning with inference efficiency. NVIDIA describes a hybrid Mamba-Transformer mixture-of-experts architecture.

Total parameters
31.6B
Active parameters
3.2B / 3.6B with embeddings
Context
Up to 1M
Official Nano information
02

Collaborative agents

Super

Family member

Positioned by NVIDIA for collaborative agents and high-volume workloads. Confirm current checkpoints and specifications on the official family page.

Total parameters
Active parameters
Context
Up to 1M family support
Official Super information
03

Frontier reasoning

Ultra

Available

The largest Nemotron 3 model, built for orchestration, coding agents, deep research and complex multi-step workflows.

Total parameters
550B
Active parameters
55B
Context
Up to 1M
Official Ultra information

Selection guide

Match scale to workload

Choose Nano when throughput, cost control, and high call volume matter. It is the clearest starting point for local or self-hosted experimentation where supported hardware is available.

Evaluate Super for collaborative-agent and enterprise-volume scenarios, while checking the official page for the latest release status and exact checkpoint details.

Choose Ultra for demanding reasoning and orchestration where model quality is the priority and you can use a hosted endpoint or appropriate infrastructure.

See official access options