DeepSeek V4 Flash Vision is DeepSeek's language model. A vision-capable large language model optimized for fast multimodal reasoning with maximum effort chain-of-thought processing.
| Benchmarks | |
|---|---|
| Intelligence Index | #29 |
| Coding Index | #30 |
| GPQA | #24 |
| HLE | #51 |
| Time to First Token | #401 |
| SciCode | #59 |
| LCR | #19 |
| Output TPS | #96 |
Capabilities
Input2/5
Text✓
Image✓
Audio·
Video·
PDF·
Output1/5
Text✓
Image·
Audio·
Video·
Embedding·
Capabilities0/13
Reasoning·
Adaptive Reasoning·
Function Calling·
Parallel Function Calling·
Structured Outputs·
Native JSON Schema·
Web Search·
URL Context·
Computer Use·
Code Execution·
File Search·
Prompt Caching·
Assistant Prefill·
Versions
| Version | Released | Context | Input / 1M | Output / 1M | Status |
|---|---|---|---|---|---|
| DeepSeek V4.731 Flash US | — | — | $0.424 | $1.27 | Available |
| DeepSeek V4 420 Flash | — | — | — | — | Available |
| DeepSeek V4 Flash Vision Exp | 1.0M | $0.220 | $0.660 | Available | |
| DeepSeek V4 Flash | 1.3M | $0.030 | $0.160 | Available | |
| DeepSeek V4 Flash Vision | — | — | — | — | Current |
| DeepSeek V4 Flash Thinking | — | 200K | $0.250 | $1.75 | Available |
| DeepSeek V4 Flash US | — | — | $0.200 | $0.400 | Available |
Other Models
| Model | Tier | Released | Context | Input / 1M | Output / 1M |
|---|---|---|---|---|---|
| DeepSeek V4 Pro | Pro | 1.0M | $0.660 | $1.98 | |
| DeepSeek V4 Pro | Pro | — | — | — | — |
| DeepSeek V4 Pro | Pro | 1.0M | $0.435 | $0.870 | |
| DeepSeek V4 Pro US | Pro | — | — | $2.40 | $4.80 |