DeepSeek R1 Distill Qwen 1.5B is DeepSeek's language model with a 131K context window, available from 3 providers, starting at $0.09 / 1M input and $0.09 / 1M output. A 1.5B Qwen-based model distilled from DeepSeek R1's reasoning chains, offering chain-of-thought capabilities in an extremely compact form factor.
Specifications
Canonical IDdeepseek-r1-distill-qwen-1-5b
TypeLanguage
StatusDeprecated
CreatorDeepSeekDeepSeek
Providers
Context Window131K tokens
Input ModalitiesText
Output ModalitiesText
Parameters1.5B
Deprecation Date
Benchmarks
Intelligence Index
#490
Math Index
#213
MMLU-Pro
#332
GPQA
#536
HLE
#520
LiveCodeBench
#320
AIME
#107
IFBench
#425
Time to First Token
#97
SciCode
#498
MATH-500
#146
AIME 2025
#213
LCR
#413
Output TPS
#335

Capabilities

Input1/5
Text
Image·
Audio·
Video·
PDF·
Output1/5
Text
Image·
Audio·
Video·
Embedding·
Capabilities0/13
Reasoning·
Adaptive Reasoning·
Function Calling·
Parallel Function Calling·
Structured Outputs·
Native JSON Schema·
Web Search·
URL Context·
Computer Use·
Code Execution·
File Search·
Prompt Caching·
Assistant Prefill·

Pricing by Provider

US Dollar ($)
Per 1M tokens
ProviderStandard
Input
$ / 1M
Output
$ / 1M
Fireworks AI logo
Fireworks AI
fireworks_ai/accounts/fireworks/models/deepseek-r1-distill-qwen-1p5b
$0.1$0.1
Nscale logo
Nscale
nscale/deepseek-ai/DeepSeek-R1-Distill-Qwen-1.5B
$0.09$0.09
Together AI logo
Together AI
together_ai/deepseek-ai/DeepSeek-R1-Distill-Qwen-1.5B
$0.18$0.18

Cost Calculator

US Dollar ($)
Preset:

Cheapest Instances to Run It

Cloud GPU instances that can host DeepSeek R1 Distill Qwen 1.5B, ranked by cheapest on-demand price. The model needs about 4 GB of GPU memory at FP16 precision (estimated from its parameter count), so treat the fit as guidance rather than a guarantee.

All clouds
FP16 (full precision)
US Dollar ($)
Instance
Cloud
GPU
VRAM
Price
Cheapest region
Standard_NV4as_v4AzureAMD Radeon Instinct MI2516 GB$0.233/hr
Standard_NV4ads_V710_v5AzureAMD Radeon Pro V710 (24GB)4 GB$0.333/hr
g5g.xlargeAWST4g16 GB$0.420/hr
7 more instances can run DeepSeek R1 Distill Qwen 1.5B
Unlock the full ranked list and FP8 / INT4 quantization with a CloudPrice subscription.

Versions

VersionReleasedContextInput / 1MOutput / 1MStatus
DeepSeek R1T2 Chimera164KAvailable
DeepSeek R1 528164K$0.200$0.250Deprecated
DeepSeek R1 Distill Qwen 32B131K$0.150$0.150Available
DeepSeek R1 Distill Llama 70B131K$0.200$0.375Deprecated
DeepSeek R1164K$0.280$0.400Deprecated
DeepSeek R1 Distill Qwen 1.5B131K$0.090$0.090Current
DeepSeek R1 Distill Qwen 14B131K$0.070$0.070Deprecated
DeepSeek R1 Distill Llama 8B131K$0.025$0.025Available
DeepSeek R1 528 Turbo33K$1.00$3.00Available
DeepSeek R1 528B128K$0.550$2.19Deprecated
DeepSeek R1 671B131K$0.800$0.800Available

Model IDs

accounts/fireworks/models/deepseek-r1-distill-qwen-1p5b
deepseek-llm-r1-distill-qwen-1-5b
deepseek-r1-distill-qwen-1-5b
fireworks_ai/accounts/fireworks/models/deepseek-r1-distill-qwen-1p5b
nscale/deepseek-ai/DeepSeek-R1-Distill-Qwen-1.5B
together_ai/deepseek-ai/DeepSeek-R1-Distill-Qwen-1.5B