Qwen3.8 27B is Alibaba's language model with a 1.0M context window and up to 236K output tokens, available from 10 providers, starting at $0.0626 / 1M input and $1.49 / 1M output. An open-weight dense vision-language model optimized for coding, professional workflows, research, multimodal interaction, and long-running agent tasks.
Specifications
Canonical IDalibaba-qwen3-8-27b
TypeLanguage
StatusActive
CreatorAlibabaAlibaba
Providers
Context Window1.0M tokens
Max Output236K tokens
Input ModalitiesImagePDFTextVideo
Output ModalitiesText
Reasoning Effortsdefault
Parameters27B
HuggingFace Likes9,338
HuggingFace Downloads (30d)2
HuggingFace Downloads (all-time)2
Release Date · 2 months ago
Benchmarks
Intelligence Index
#52
Coding Index
#32
GPQA
#38
HLE
#66
Time to First Token
#530
SciCode
#78
LCR
#24
Output TPS
#193

Capabilities

Input4/5
Text✓
Image✓
Audio·
Video✓
PDF✓
Output1/5
Text✓
Image·
Audio·
Video·
Embedding·
Capabilities6/13
Reasoning✓
Adaptive Reasoning·
Function Calling✓
Parallel Function Calling✓
Structured Outputs✓
Native JSON Schema✓
Web Search·
URL Context·
Computer Use·
Code Execution·
File Search·
Prompt Caching✓
Assistant Prefill·

Pricing by Provider

US Dollar ($)
Per 1M tokens
ProviderStandardFreeBatch
Input
$ / 1M
Output
$ / 1M
Cache Read
$ / 1M
Cache Write 5m
$ / 1M
Input
$ / 1M
Output
$ / 1M
Input
$ / 1M
Output
$ / 1M
Alibaba Qwen logo
Alibaba Qwen
qwen3.8-27b
$0.5$3.00N/AN/A——$0.25$1.50
Cerebras logo
Cerebras
cerebras/qwen-3.8-27b
$0.99$1.49N/AN/A————
Cloudflare Workers AI logo
Cloudflare Workers AI
@cf/qwen/qwen3.8-27b
$0.45$3.20$0.05N/A————
DeepInfra logo
DeepInfra
deepinfra/Qwen/Qwen3.8-27B
$0.4$3.00$0.04N/A————
Groq logo
Groq
groq/qwen/qwen3.8-27b
$0.8$4.00N/AN/A————
Hugging Face logo
Hugging Face
novita:qwen/qwen3.8-27b
$0.42$3.00N/AN/A————
Hugging Face logo
Hugging Face
ovhcloud:Qwen3.8-27B
$0.47$3.19N/AN/A————
Nebius logo
Nebius
Qwen/Qwen3.8-27B
$0.45$3.00N/AN/A————
OpenRouter logo
OpenRouter
qwen/qwen3.8-27b
$0.0626$4.40$0.0501N/AN/AN/A——
OpenRouter logo
OpenRouter
qwen/qwen3.8-27b:free
————$N/A$N/A——
Vercel AI Gateway logo
Vercel AI Gateway
alibaba/qwen3.8-27b
$0.5$3.00$0.1$0.625————
Weights & Biases logo
Weights & Biases
wandb/Qwen/Qwen3.8-27B
$0.4$3.00$0.15N/A————

Cost Calculator

US Dollar ($)
Preset:

Cheapest Instances to Run It

Cloud GPU instances that can host Qwen3.8 27B, ranked by cheapest on-demand price. The model needs about 65 GB of GPU memory at FP16 precision (estimated from its parameter count), so treat the fit as guidance rather than a guarantee.

All clouds
FP16 (full precision)
US Dollar ($)
Instance
Cloud
GPU
VRAM
Price
Cheapest region
Standard_NP20sAzure2× AMD Alveo U250 FPGA (64GB)128 GB$3.30/hr
g7e.2xlargeAWSRTX PRO Server 600096 GB$3.36/hr
Standard_NC24ads_A100_v4AzureNVIDIA A10080 GB$3.67/hr
7 more instances can run Qwen3.8 27B
Unlock the full ranked list and FP8 / INT4 quantization with a CloudPrice subscription.

Versions

VersionReleasedContextInput / 1MOutput / 1MStatus
Qwen3.8 2.4T A95B1.0M$2.00$6.00Available
Qwen Audio 3 Flash ASR————Available
Qwen Audio 3 Flash ASR Filetrans————Available
Qwen Audio 3 Flash ASR Streaming————Available
Qwen Audio 3 Flash Realtime——$0.230$0.930Available
Qwen Audio 3 Plus Realtime——$0.800$6.40Available
Qwen Audio 3 Plus TTS————Available
Qwen Audio 3.1 Plus Realtime——$0.800$6.40Available
Qwen Image 3————Available
Qwen Image 3.0 Pro————Available
Qwen3.8 27B1.0M$0.063$1.49Current

Model IDs

@cf/qwen/qwen3.8-27b
accounts/fireworks/models/qwen3p8-27b
alibaba-qwen3-8-27b
alibaba/qwen3.8-27b
cerebras/qwen-3.8-27b
deepinfra/Qwen/Qwen3.8-27B
groq/qwen/qwen3.8-27b
huggingface-vlm-qwen3-8-27b
openrouter/qwen/qwen3.8-27b
openrouter/qwen/qwen3.8-27b:free
qwen/qwen3.8-27b
Qwen/Qwen3.8-27B
qwen3-8-27b
qwen3-8-27b-low
qwen3-8-27b-medium
qwen3-8-27b-non-reasoning
qwen3.8-27b
wandb/Qwen/Qwen3.8-27B