Llama Nemotron 1.5 Super 49B is NVIDIA's language model. A 49B-parameter Nemotron Super model at version 1.5, fine-tuned by NVIDIA on a Llama base for efficient reasoning and agentic task performance.
Specifications
Canonical IDnvidia-llama-nemotron-1-5-super-49b
TypeLanguage
StatusActive
CreatorNVIDIANVIDIA
Input ModalitiesText
Output ModalitiesText
Parameters49B
Benchmarks
Intelligence Index
8.5
#340
Math Index
8.0
#227
MMLU-Pro
0.7
#214
GPQA
0.5
#359
HLE
0.0
#374
LiveCodeBench
0.3
#215
AIME
0.1
#103
IFBench
0.3
#311
Time to First Token
3.44s
#490
SciCode
0.2
#337
MATH-500
0.8
#105
AIME 2025
0.1
#227
LCR
0.2
#269
TerminalBench Hard
0.0
#285
TAU2
0.3
#270
Output TPS

Capabilities

Input1/5
Text
Image·
Audio·
Video·
PDF·
Output1/5
Text
Image·
Audio·
Video·
Embedding·
Capabilities0/13
Reasoning·
Adaptive Reasoning·
Function Calling·
Parallel Function Calling·
Structured Outputs·
Native JSON Schema·
Web Search·
URL Context·
Computer Use·
Code Execution·
File Search·
Prompt Caching·
Assistant Prefill·

Cheapest Instances to Run It

Cloud GPU instances that can host Llama Nemotron 1.5 Super 49B, ranked by cheapest on-demand price. The model needs about 118 GB of GPU memory at FP16 precision (estimated from its parameter count), so treat the fit as guidance rather than a guarantee.

All clouds
FP16 (full precision)
US Dollar ($)
Instance
Cloud
GPU
VRAM
Price
Cheapest region
Standard_NP20sAzure2× AMD Alveo U250 FPGA (64GB)128 GB$3.30/hreastus
Standard_NP40sAzure4× AMD Alveo U250 FPGA (64GB)256 GB$6.60/hreastus
g4dn.metalAWS8× T4128 GB$7.82/hrus-east-1
7 more instances can run Llama Nemotron 1.5 Super 49B
Unlock the full ranked list and FP8 / INT4 quantization with a CloudPrice subscription.

Versions

VersionReleasedContextInput / 1MOutput / 1MStatus
Llama Nemotron 1.5 Super 49BCurrent
Llama Nemotron 1.5 Super 49B ReasoningAvailable
Llama 3.3 Nemotron Super 49B ReasoningAvailable
Llama 3.3 Nemotron Super 49BAvailable

Other Models

ModelTierReleasedContextInput / 1MOutput / 1M
Llama 3.3 70B Instruct131K$0.100$0.200
Llama 3.2 3B Instruct131K$0.015$0.020
Llama 3.2 1B Instruct128K$0.027$0.080
Llama 3.2 11B128K$0.160$0.160
Llama 3.1 405B Instruct131K$0.120$0.300
Llama 3.1 8B Instruct200K$0.020$0.030
Llama 3.1 70B Instruct131K$0.120$0.300
Llama 3.1 70B128K$0.360$0.360
Llama 3.1 8B131K$0.030$0.050
Llama 3 70B Instruct131K$0.120$0.300

Model IDs

llama-nemotron-super-49b-v1-5
llama-nemotron-super-49b-v1-5-reasoning
nvidia-llama-nemotron-1-5-super-49b