Nemotron 3 Nano 30B A3B is NVIDIA's language model. A 30B-parameter mixture-of-experts Nemotron model with 3B active parameters, optimized for efficient reasoning tasks.
Specifications
Canonical IDnvidia-nemotron-3-nano-30b-a3b
TypeLanguage
StatusActive
CreatorNVIDIANVIDIA
Input ModalitiesText
Output ModalitiesText
Parameters30B

Capabilities

Input1/5
Text
Image·
Audio·
Video·
PDF·
Output1/5
Text
Image·
Audio·
Video·
Embedding·
Capabilities0/13
Reasoning·
Adaptive Reasoning·
Function Calling·
Parallel Function Calling·
Structured Outputs·
Native JSON Schema·
Web Search·
URL Context·
Computer Use·
Code Execution·
File Search·
Prompt Caching·
Assistant Prefill·

Cheapest Instances to Run It

Cloud GPU instances that can host Nemotron 3 Nano 30B A3B, ranked by cheapest on-demand price. The model needs about 72 GB of GPU memory at FP16 precision (estimated from its parameter count), so treat the fit as guidance rather than a guarantee.

All clouds
FP16 (full precision)
US Dollar ($)
Instance
Cloud
GPU
VRAM
Price
Cheapest region
Standard_NP20sAzure2× AMD Alveo U250 FPGA (64GB)128 GB$3.30/hreastus
g7e.2xlargeAWSRTX PRO Server 600096 GB$3.36/hrus-east-1
Standard_NC24ads_A100_v4AzureNVIDIA A10080 GB$3.67/hreastus
7 more instances can run Nemotron 3 Nano 30B A3B
Unlock the full ranked list and FP8 / INT4 quantization with a CloudPrice subscription.

Other Models

ModelTierReleasedContextInput / 1MOutput / 1M
Nemotron 4 15B
Nemotron 3.5 Content Safety128K
Nemotron 3 Ultra 550B A55BUltra1.0M$0.500$2.20
Nemotron Nano 3 30B A3B Omni Reasoning256K
Nemotron Super 3 120B256K$0.150$0.650
Nemotron Nano 3 30B262K$0.060$0.240
Nemotron Nano 3 30B A3B Omni
Nemotron Nano 3 30B A3B Reasoning
Nemotron Nano 3 4B
Nemotron 3

Model IDs

huggingface-reasoning-nvidia-nemotron-3-nano-30b-a3b-bf16
nvidia-nemotron-3-nano-30b-a3b