Qwen3.5 122B A10B is Alibaba's language model with a 262K context window and up to 82K output tokens, available from 4 providers, starting at $0.25 / 1M input and $1.75 / 1M output. A large Qwen3.5 Mixture-of-Experts model with 122B total parameters and 10B activated per token, featuring a 262K token context window for complex reasoning tasks.
Specifications
Canonical IDalibaba-qwen3-5-122b-a10b
TypeLanguage
StatusActive
CreatorAlibabaAlibaba
Providers
Context Window262K tokens
Max Output82K tokens
Input ModalitiesImageTextVideo
Output ModalitiesText
Reasoning Effortsdefault
Parameters122B
HuggingFace Likes523
HuggingFace Downloads (30d)906,547
HuggingFace Downloads (all-time)1,516,880
Release Date · 6 months ago
Benchmarks
Intelligence Index
#97
Coding Index
#60
GPQA
#70
HLE
#79
IFBench
#23
Time to First Token
#459
SciCode
#89
LCR
#72
TerminalBench Hard
#85
TAU2
#34
Output TPS
#48

Capabilities

Input3/5
Text
Image
Audio·
Video
PDF·
Output1/5
Text
Image·
Audio·
Video·
Embedding·
Capabilities4/13
Reasoning
Adaptive Reasoning·
Function Calling
Parallel Function Calling·
Structured Outputs
Native JSON Schema
Web Search·
URL Context·
Computer Use·
Code Execution·
File Search·
Prompt Caching·
Assistant Prefill·

Pricing by Provider

US Dollar ($)
Per 1M tokens
ProviderStandardBatch
Input
$ / 1M
Output
$ / 1M
Input
$ / 1M
Output
$ / 1M
Alibaba Qwen logo
Alibaba Qwen
qwen3.5-122b-a10b
$0.4$3.20$0.2$1.60
Hugging Face logo
Hugging Face
novita:qwen/qwen3.5-122b-a10b
$0.4$3.20
Libertai
libertai/qwen3.5-122b-a10b
$0.25$1.75
OpenRouter logo
OpenRouter
qwen/qwen3.5-122b-a10b
$0.29$2.40

Cost Calculator

US Dollar ($)
Preset:

Cheapest Instances to Run It

Cloud GPU instances that can host Qwen3.5 122B A10B, ranked by cheapest on-demand price. The model needs about 293 GB of GPU memory at FP16 precision (estimated from its parameter count), so treat the fit as guidance rather than a guarantee.

All clouds
FP16 (full precision)
US Dollar ($)
Instance
Cloud
GPU
VRAM
Price
Cheapest region
g7e.24xlargeAWS4× RTX PRO Server 6000384 GB$16.57/hr
g4-standard-192GCP4× nvidia-rtx-pro-6000384 GB$18.00/hr
p4d.24xlargeAWS8× A100320 GB$21.96/hr
7 more instances can run Qwen3.5 122B A10B
Unlock the full ranked list and FP8 / INT4 quantization with a CloudPrice subscription.

Versions

VersionReleasedContextInput / 1MOutput / 1MStatus
Qwen3.5 122B A10B262K$0.250$1.75Current
Qwen3.5 35B A3B262K$0.225$1.80Available
Qwen3.5 397B A17B262K$0.390$2.34Available
Qwen3.7 Plus VL InstructAvailable

Model IDs

accounts/fireworks/models/qwen3p5-122b-a10b
alibaba-qwen3-5-122b-a10b
huggingface-llm-qwen3-5-122b-a10b
libertai/qwen3.5-122b-a10b
openrouter/qwen/qwen3.5-122b-a10b
qwen/qwen3.5-122b-a10b
Qwen/Qwen3.5-122B-A10B
qwen3-5-122b-a10b
qwen3-5-122b-a10b-non-reasoning
qwen3.5-122b-a10b