DeepSeek V3.1 is DeepSeek's language model with a 164K context window and up to 33K output tokens, starting at $0.25 / 1M input and $0.95 / 1M output. DeepSeek's 671B-parameter hybrid reasoning MoE LLM that supports both thinking and non-thinking modes via prompt templates, extending the V3 base.
Specifications
Canonical IDdeepseek-chat-3-1
TypeLanguage
StatusActive
CreatorDeepSeekDeepSeek
Providers
Context Window164K tokens
Max Output33K tokens
Input ModalitiesText
Output ModalitiesText
Reasoning Effortsdefault
Parameters685B
HuggingFace Likes819
HuggingFace Downloads (30d)149,478
HuggingFace Downloads (all-time)1,703,662
Release Date · 11 months ago
Knowledge Cutoff · 1 year ago

Capabilities

Input1/5
Text
Image·
Audio·
Video·
PDF·
Output1/5
Text
Image·
Audio·
Video·
Embedding·
Capabilities6/13
Reasoning
Adaptive Reasoning·
Function Calling
Parallel Function Calling·
Structured Outputs
Native JSON Schema
Web Search·
URL Context·
Computer Use·
Code Execution·
File Search·
Prompt Caching
Assistant Prefill

Pricing by Provider

US Dollar ($)
Per 1M tokens
ProviderStandard
Input
$ / 1M
Output
$ / 1M
Cache Read
$ / 1M
OpenRouter logo
OpenRouter
deepseek/deepseek-chat-v3.1
$0.25$0.95$0.13

Cost Calculator

US Dollar ($)
Preset:

Cheapest Instances to Run It

Cloud GPU instances that can host DeepSeek V3.1, ranked by cheapest on-demand price. The model needs about 1643 GB of GPU memory at FP16 precision (estimated from its parameter count), so treat the fit as guidance rather than a guarantee.

All clouds
FP16 (full precision)
US Dollar ($)
Instance
Cloud
GPU
VRAM
Price
Cheapest region
p6-b300.48xlargeAWS8× B3002149 GB$142.42/hrus-east-1

Versions

VersionReleasedContextInput / 1MOutput / 1MStatus
DeepSeek V3.1164K$0.250$0.950Current
DeepSeek V3 0324164K$0.270$1.12Available
Kimi K2 Instruct262K$0.500$2.00Available

Model IDs

deepseek-chat-3-1
deepseek/deepseek-chat-v3.1
openrouter/deepseek/deepseek-chat-v3.1