Qwen2.5 32B Instruct is Alibaba's language model with a 128K context window and up to 8K output tokens, starting at $0.9 / 1M input and $0.9 / 1M output. A 32-billion-parameter instruction-tuned LLM from Alibaba's Qwen2.5 series, optimized for following complex instructions and text generation tasks.
| Benchmarks | |
|---|---|
| Intelligence Index | #420 |
| MMLU-Pro | #225 |
| GPQA | #417 |
| HLE | #466 |
| LiveCodeBench | #251 |
| AIME | #123 |
| Time to First Token | #13 |
| SciCode | #397 |
| MATH-500 | #110 |
| Output TPS | #252 |
Capabilities
Input1/5
Text✓
Image·
Audio·
Video·
PDF·
Output1/5
Text✓
Image·
Audio·
Video·
Embedding·
Capabilities1/13
Reasoning·
Adaptive Reasoning·
Function Calling✓
Parallel Function Calling·
Structured Outputs·
Native JSON Schema·
Web Search·
URL Context·
Computer Use·
Code Execution·
File Search·
Prompt Caching·
Assistant Prefill·
Pricing by Provider
US Dollar ($)
Per 1M tokens
| Provider | Standard | |
|---|---|---|
| Input $ / 1M | Output $ / 1M | |
| $0.9 | $0.9 | |
Cost Calculator
US Dollar ($)
Preset:
Cheapest Instances to Run It
Cloud GPU instances that can host Qwen2.5 32B Instruct, ranked by cheapest on-demand price. The model needs about 77 GB of GPU memory at FP16 precision (estimated from its parameter count), so treat the fit as guidance rather than a guarantee.
All clouds
FP16 (full precision)
US Dollar ($)
Instance | Cloud | GPU | VRAM | Price | Cheapest region | |
|---|---|---|---|---|---|---|
| Standard_NP20s | 2× AMD Alveo U250 FPGA (64GB) | 128 GB | $3.30/hr | |||
| g7e.2xlarge | RTX PRO Server 6000 | 96 GB | $3.36/hr | |||
| Standard_NC24ads_A100_v4 | NVIDIA A100 | 80 GB | $3.67/hr | |||
Versions
| Version | Released | Context | Input / 1M | Output / 1M | Status |
|---|---|---|---|---|---|
| Dolphin 2.9.2 Qwen2 72B | — | 131K | $0.900 | $0.900 | Available |
| Cogito V1 Preview Qwen 14B | — | 131K | $0.200 | $0.200 | Available |
| Cogito V1 Preview Qwen 32B | — | 131K | $0.900 | $0.900 | Available |
| QwQ 32B | 131K | $0.150 | $0.200 | Deprecated | |
| Qwen2.5 Coder 32B Instruct | 131K | $0.050 | $0.100 | Deprecated | |
| Qwen2.5 7B Instruct | 33K | $0.040 | $0.070 | Available | |
| Qwen2.5 72B Instruct | 131K | $0.120 | $0.300 | Available | |
| Qwen2.5 32B Instruct | — | 128K | $0.900 | $0.900 | Current |
| QwQ 32B Preview | — | 33K | $0.900 | $0.900 | Available |
| Qwen2 72B Instruct | — | 33K | $0.900 | $0.900 | Deprecated |
| Qwen2.5 Coder 7B Instruct | — | 33K | $0.010 | $0.030 | Available |