DeepSeek V4 Flash US is DeepSeek's language model, starting at $0.2 / 1M input and $0.4 / 1M output. A fast, lightweight large language model in the DeepSeek V4 series optimized for low-latency inference in the US region.
Specifications
Canonical IDdeepseek-v4-flash-us
TypeLanguage
StatusActive
CreatorDeepSeekDeepSeek
Providers
Input ModalitiesText
Output ModalitiesText

Capabilities

Input1/5
Text
Image·
Audio·
Video·
PDF·
Output1/5
Text
Image·
Audio·
Video·
Embedding·
Capabilities0/13
Reasoning·
Adaptive Reasoning·
Function Calling·
Parallel Function Calling·
Structured Outputs·
Native JSON Schema·
Web Search·
URL Context·
Computer Use·
Code Execution·
File Search·
Prompt Caching·
Assistant Prefill·

Pricing by Provider

US Dollar ($)
Per 1M tokens
ProviderStandardBatch
Input
$ / 1M
Output
$ / 1M
Input
$ / 1M
Output
$ / 1M
Alibaba Qwen logo
Alibaba Qwen
deepseek-v4-flash-us
$0.2$0.4$0.1$0.2

Cost Calculator

US Dollar ($)
Preset:

Versions

VersionReleasedContextInput / 1MOutput / 1MStatus
DeepSeek V4.731 Flash US$0.424$1.27Available
DeepSeek V4 420 FlashAvailable
DeepSeek V4 Flash Vision Exp1.0M$0.220$0.660Available
DeepSeek V4 Flash1.3M$0.030$0.160Available
DeepSeek V4 Flash US$0.200$0.400Current
DeepSeek V4 Flash VisionAvailable
DeepSeek V4 Flash Thinking200K$0.250$1.75Available

Other Models

ModelTierReleasedContextInput / 1MOutput / 1M
DeepSeek V4 ProPro1.0M$0.660$1.98
DeepSeek V4 ProPro
DeepSeek V4 ProPro1.0M$0.435$0.870
DeepSeek V4 Pro USPro$2.40$4.80

Model IDs

deepseek-v4-flash-us