← Back to registry
NVIDIA

nemotron-3-nano:cloud

Low usage

Nemotron 3 Nano is trained from scratch by NVIDIA as a unified model for both reasoning and non-reasoning tasks, using a hybrid Mixture-of-Experts architecture of 23 Mamba-2 and MoE layers plus 6 Attention layers.

toolsthinkingcloudCommunity r/nvidia
Context window1M
ModalitiesText
Size30B (3.5B active)
Pulls735.7K
Tags9
Updated5 months ago
ollama run nemotron-3-nano:30b-cloud

Cloud tags

TagParametersContextModalities
nemotron-3-nano:30b-cloud30B (3.5B active)1MText
nemotron-3-nano:4b (local, 2.8GB)4B256KText
nemotron-3-nano:30b (local, 24GB)30B (3.5B active)1MText

Key features

Best practices

Specs sourced from ollama.com/library/nemotron-3-nano · Retrieved 2026-08-29