DeepSeek V4 Flash Vision Exp is DeepSeek's language model with a 1.0M context window and up to 944K output tokens, available from 6 providers, starting at $0.22 / 1M input and $0.66 / 1M output. An experimental vision-enabled large language model that adds image understanding to the lightweight DeepSeek V4 Flash variant, with reasoning and tool-use capabilities.
Specifications
Canonical IDdeepseek-v4-flash-vision-exp
TypeLanguage
StatusActive
CreatorDeepSeekDeepSeek
Providers
Context Window1.0M tokens
Max Output944K tokens
Input ModalitiesImageText
Output ModalitiesText
Reasoning Effortsdefault
Parameters305B
HuggingFace Likes560
HuggingFace Downloads (30d)54,571
HuggingFace Downloads (all-time)54,571
Release Date · 1 month ago

Capabilities

Input2/5
Text
Image
Audio·
Video·
PDF·
Output1/5
Text
Image·
Audio·
Video·
Embedding·
Capabilities7/13
Reasoning
Adaptive Reasoning·
Function Calling
Parallel Function Calling
Structured Outputs
Native JSON Schema
Web Search·
URL Context·
Computer Use·
Code Execution·
File Search·
Prompt Caching
Assistant Prefill

Pricing by Provider

US Dollar ($)
Per 1M tokens
ProviderStandardBatch
Input
$ / 1M
Output
$ / 1M
Cache Read
$ / 1M
Input
$ / 1M
Output
$ / 1M
Cache Read
$ / 1M
DeepSeek logo
DeepSeek
deepseek-v4-flash-vision-exp
$0.3$1.20$0.006
Fireworks AI logo
Fireworks AI
fireworks_ai/deepseek-v4-flash-vision-exp
$0.22$0.66$0.007
Hugging Face logo
Hugging Face
novita:deepseek/deepseek-v4-flash-vision-exp
$0.44$1.32N/A
Novita logo
Novita
novita/deepseek/deepseek-v4-flash-vision-exp
$0.44$1.32$0.028
OpenRouter logo
OpenRouter
deepseek/deepseek-v4-flash-vision-exp
$0.22$0.66$0.007N/AN/AN/A
OpenRouter logo
OpenRouter
deepseek/deepseek-v4-flash-vision-exp:batch
$0.11$0.33$0.0035
Vercel AI Gateway logo
Vercel AI Gateway
deepseek/deepseek-v4-flash-vision-exp
$0.22$0.66$0.007

Cost Calculator

US Dollar ($)
Preset:

Cheapest Instances to Run It

Cloud GPU instances that can host DeepSeek V4 Flash Vision Exp, ranked by cheapest on-demand price. The model needs about 731 GB of GPU memory at FP16 precision (estimated from its parameter count), so treat the fit as guidance rather than a guarantee.

All clouds
FP16 (full precision)
US Dollar ($)
Instance
Cloud
GPU
VRAM
Price
Cheapest region
g7e.48xlargeAWS8× RTX PRO Server 6000768 GB$33.14/hr
Standard_ND96is_MI300X_v5Azure8× AMD Instinct MI300X GPU (192GB)1536 GB$48.00/hr
Standard_ND96isr_MI300X_v5Azure8× AMD Instinct MI300X1536 GB$48.00/hr
5 more instances can run DeepSeek V4 Flash Vision Exp
Unlock the full ranked list and FP8 / INT4 quantization with a CloudPrice subscription.

Versions

VersionReleasedContextInput / 1MOutput / 1MStatus
DeepSeek V4.731 Flash USAvailable
DeepSeek V4 420 FlashAvailable
DeepSeek V4.4.1 FlashAvailable
DeepSeek V4.1 Flash Expires Exp1.0M$0.220$0.660Available
DeepSeek V4.1 Flash Beta1.0MAvailable
DeepSeek V4 Flash Vision Exp1.0M$0.220$0.660Current
DeepSeek V4 Flash1.3M$0.030$0.177Deprecating
DeepSeek V4 Flash VisionAvailable
DeepSeek V4 Flash 7311.0M$0.140$0.280Available
DeepSeek V4 Flash Thinking200K$0.250$1.75Available
DeepSeek V4 Flash USAvailable

Other Models

ModelTierReleasedContextInput / 1MOutput / 1M
DeepSeek V4 ProPro1.0M$0.624$1.98
DeepSeek V4 ProPro
DeepSeek V4 ProPro1.0M$0.435$0.870
DeepSeek V4 Pro 813Pro1.0M$1.32$3.96
DeepSeek V4 Pro USPro

Model IDs

accounts/fireworks/models/deepseek-v4-flash-vision-exp
deepseek-ai/DeepSeek-V4-Flash-Vision-Exp
deepseek-v4-flash-vision-exp
deepseek/deepseek-v4-flash-vision-exp
fireworks_ai/accounts/fireworks/models/deepseek-v4-flash-vision-exp
fireworks_ai/deepseek-v4-flash-vision-exp
novita/deepseek/deepseek-v4-flash-vision-exp
openrouter/deepseek/deepseek-v4-flash-vision-exp
openrouter/deepseek/deepseek-v4-flash-vision-exp:batch