Qwen3 VL 235B A22B Instruct is Alibaba's language model with a 262K context window and up to 129K output tokens, available from 7 providers, starting at $0.21 / 1M input and $0.88 / 1M output. The flagship instruction-tuned vision-language MoE model in the Qwen3 series, with 235B total and 22B activated parameters for superior visual perception and reasoning.
Specifications
Canonical IDalibaba-qwen3-vl-235b-a22b-instruct
TypeLanguage
StatusActive
CreatorAlibabaAlibaba
Providers
Context Window262K tokens
Max Output129K tokens
Input ModalitiesImagePDFText
Output ModalitiesText
Parameters235B
HuggingFace Likes383
HuggingFace Downloads (30d)947,793
HuggingFace Downloads (all-time)2,172,030
Release Date · 10 months ago
Knowledge Cutoff · 1 year ago
Benchmarks
Intelligence Index
14.3
#237
Math Index
70.7
#87
MMLU-Pro
0.8
#64
GPQA
0.7
#200
HLE
0.1
#234
LiveCodeBench
0.6
#107
IFBench
0.4
#209
Time to First Token
1.10s
#390
SciCode
0.4
#181
AIME 2025
0.7
#87
LCR
0.3
#221
TerminalBench Hard
0.1
#227
TAU2
0.4
#211
Output TPS
48.5
#223

Capabilities

Input3/5
Text
Image
Audio·
Video·
PDF
Output1/5
Text
Image·
Audio·
Video·
Embedding·
Capabilities5/13
Reasoning·
Adaptive Reasoning·
Function Calling
Parallel Function Calling
Structured Outputs
Native JSON Schema
Web Search·
URL Context·
Computer Use·
Code Execution·
File Search·
Prompt Caching
Assistant Prefill·

Pricing by Provider

US Dollar ($)
Per 1M tokens
ProviderStandardBatch
Input
$ / 1M
Output
$ / 1M
Cache Read
$ / 1M
Input
$ / 1M
Output
$ / 1M
Alibaba Qwen logo
Alibaba Qwen
qwen3-vl-235b-a22b-instruct
$0.4$1.60N/A$0.2$0.8
Fireworks AI logo
Fireworks AI
fireworks_ai/accounts/fireworks/models/qwen3-vl-235b-a22b-instruct
$0.22$0.88N/A
GMI Cloud logo
GMI Cloud
gmi/Qwen/Qwen3-VL-235B-A22B-Instruct-FP8
$0.3$1.40N/A
Hugging Face logo
Hugging Face
novita:qwen/qwen3-vl-235b-a22b-instruct
$0.3$1.50N/A
Novita logo
Novita
novita/qwen/qwen3-vl-235b-a22b-instruct
$0.3$1.50N/A
OpenRouter logo
OpenRouter
qwen/qwen3-vl-235b-a22b-instruct
$0.21$1.90$0.1
Vercel AI Gateway logo
Vercel AI Gateway
alibaba/qwen3-vl-235b-a22b-instruct
$0.4$1.60N/A

Cost Calculator

US Dollar ($)
Preset:

Cheapest Instances to Run It

Cloud GPU instances that can host Qwen3 VL 235B A22B Instruct, ranked by cheapest on-demand price. The model needs about 564 GB of GPU memory at FP16 precision (estimated from its parameter count), so treat the fit as guidance rather than a guarantee.

All clouds
FP16 (full precision)
US Dollar ($)
Instance
Cloud
GPU
VRAM
Price
Cheapest region
p4de.24xlargeAWS8× A100640 GB$27.45/hrus-east-1
Standard_ND96amsr_A100_v4Azure8× NVIDIA A100 (80GB)640 GB$32.77/hrwestus2
g7e.48xlargeAWS8× RTX PRO Server 6000768 GB$33.14/hrus-east-1
7 more instances can run Qwen3 VL 235B A22B Instruct
Unlock the full ranked list and FP8 / INT4 quantization with a CloudPrice subscription.

Versions

VersionReleasedContextInput / 1MOutput / 1MStatus
Qwen3 VL 30B A3B Instruct262K$0.130$0.520Available
Qwen3 VL 30B A3B Thinking262K$0.130$0.600Available
Qwen3 VL 235B A22B Instruct262K$0.210$0.880Current
Qwen3 VL 235B A22B Thinking262K$0.220$0.880Available

Model IDs

accounts/fireworks/models/qwen3-vl-235b-a22b-instruct
alibaba-qwen3-vl-235b-a22b-instruct
alibaba/qwen3-vl-235b-a22b-instruct
dashscope/qwen3-vl-235b-a22b-instruct
fireworks_ai/accounts/fireworks/models/qwen3-vl-235b-a22b-instruct
gmi/Qwen/Qwen3-VL-235B-A22B-Instruct-FP8
novita/qwen/qwen3-vl-235b-a22b-instruct
qwen/qwen3-vl-235b-a22b-instruct
Qwen/Qwen3-VL-235B-A22B-Instruct
qwen3-vl-235b-a22b-instruct