Llama 3.3 Nemotron Super 49B Reasoning is an AI model from NVIDIA. A reasoning-specialized 49B-parameter variant of NVIDIA's Nemotron Super built on Llama 3.3, optimized for complex multi-step problem solving.
| Specifications | |
|---|---|
nvidia-llama-3-3-nemotron-super-49b-reasoning | |
| Active | |
| Benchmarks | |
|---|---|
| Intelligence Index | #317 |
| Math Index | #159 |
| MMLU-Pro | #154 |
| GPQA | #306 |
| HLE | #323 |
| LiveCodeBench | #237 |
| AIME | #61 |
| IFBench | #283 |
| Time to First Token | #223 |
| SciCode | #324 |
| MATH-500 | #44 |
| AIME 2025 | #159 |
| LCR | #339 |
| TerminalBench Hard | #406 |
| TAU2 | #280 |
| Output TPS | #457 |
Capabilities
Input0/5
Text·
Image·
Audio·
Video·
PDF·
Output0/5
Text·
Image·
Audio·
Video·
Embedding·
Capabilities0/13
Reasoning·
Adaptive Reasoning·
Function Calling·
Parallel Function Calling·
Structured Outputs·
Native JSON Schema·
Web Search·
URL Context·
Computer Use·
Code Execution·
File Search·
Prompt Caching·
Assistant Prefill·
Versions
| Version | Released | Context | Input / 1M | Output / 1M | Status |
|---|---|---|---|---|---|
| Llama Nemotron 1.5 Super 49B | — | — | — | — | Available |
| Llama Nemotron 1.5 Super 49B Reasoning | — | — | — | — | Available |
| Llama 3.3 Nemotron Super 49B Reasoning | — | — | — | — | Current |
| Llama 3.3 Nemotron Super 49B | — | — | — | — | Available |
Other Models
| Model | Tier | Released | Context | Input / 1M | Output / 1M |
|---|---|---|---|---|---|
| Llama 3.3 70B Instruct | — | 131K | $0.120 | $0.200 | |
| Llama 3.2 3B Instruct | — | 131K | $0.015 | $0.020 | |
| Llama 3.2 1B Instruct | — | 131K | $0.020 | $0.020 | |
| Llama 3.2 11B | — | 128K | $0.160 | $0.160 | |
| Llama 3.1 405B Instruct | — | 131K | $0.120 | $0.300 | |
| Llama 3.1 8B Instruct | — | 200K | $0.020 | $0.030 | |
| Llama 3.1 70B Instruct | — | 131K | $0.100 | $0.100 | |
| Llama 3.1 70B | — | 128K | $0.360 | $0.360 | |
| Llama 3.1 8B | — | 131K | $0.030 | $0.050 | |
| Llama 3 70B Instruct | — | 131K | $0.120 | $0.300 |