Llama Nemotron 1.5 Super 49B is NVIDIA's language model. A 49B-parameter Nemotron Super model at version 1.5, fine-tuned by NVIDIA on a Llama base for efficient reasoning and agentic task performance.
Specifications
Canonical IDnvidia-llama-nemotron-1-5-super-49b
TypeLanguage
StatusActive
CreatorNVIDIANVIDIA
Input ModalitiesText
Output ModalitiesText
Parameters49B
Benchmarks
Intelligence Index
#327
Math Index
#99
MMLU-Pro
#97
GPQA
#222
HLE
#282
LiveCodeBench
#60
AIME
#15
IFBench
#294
Time to First Token
#523
SciCode
#256
MATH-500
#15
AIME 2025
#99
LCR
#286
TerminalBench Hard
#269
TAU2
#269
Output TPS
#58

Capabilities

Input1/5
Text
Image·
Audio·
Video·
PDF·
Output1/5
Text
Image·
Audio·
Video·
Embedding·
Capabilities0/13
Reasoning·
Adaptive Reasoning·
Function Calling·
Parallel Function Calling·
Structured Outputs·
Native JSON Schema·
Web Search·
URL Context·
Computer Use·
Code Execution·
File Search·
Prompt Caching·
Assistant Prefill·

Cheapest Instances to Run It

Cloud GPU instances that can host Llama Nemotron 1.5 Super 49B, ranked by cheapest on-demand price. The model needs about 118 GB of GPU memory at FP16 precision (estimated from its parameter count), so treat the fit as guidance rather than a guarantee.

All clouds
FP16 (full precision)
US Dollar ($)
Instance
Cloud
GPU
VRAM
Price
Cheapest region
Standard_NP20sAzure2× AMD Alveo U250 FPGA (64GB)128 GB$3.30/hr
Standard_NP40sAzure4× AMD Alveo U250 FPGA (64GB)256 GB$6.60/hr
Standard_NC48ads_A100_v4Azure2× NVIDIA A100160 GB$7.35/hr
7 more instances can run Llama Nemotron 1.5 Super 49B
Unlock the full ranked list and FP8 / INT4 quantization with a CloudPrice subscription.

Versions

VersionReleasedContextInput / 1MOutput / 1MStatus
Llama Nemotron 1.5 Super 49BCurrent
Llama Nemotron 1.5 Super 49B ReasoningAvailable
Llama 3.3 Nemotron Super 49B ReasoningAvailable
Llama 3.3 Nemotron Super 49BAvailable

Other Models

ModelTierReleasedContextInput / 1MOutput / 1M
Llama 3.3 70B Instruct131K$0.100$0.200
Llama 3.2 3B Instruct131K$0.015$0.020
Llama 3.2 1B Instruct131K$0.020$0.020
Llama 3.1 405B Instruct131K$0.120$0.300
Llama 3.1 8B Instruct200K$0.020$0.030
Llama 3.1 70B Instruct131K$0.100$0.100
Llama 3.1 70B128K$0.360$0.360
Llama 3.1 8B131K$0.030$0.050
Llama 3 70B Instruct131K$0.120$0.300
Llama 3 8B Instruct32K$0.030$0.040

Model IDs

llama-nemotron-super-49b-v1-5
llama-nemotron-super-49b-v1-5-reasoning
nvidia-llama-nemotron-1-5-super-49b