Qwen3.5 35B A3B is Alibaba's language model with a 262K context window and up to 66K output tokens, available from 3 providers, starting at $0.14 / 1M input and $1.00 / 1M output. An efficient Qwen3.5 MoE model with 35B total parameters and only 3B activated per token, offering strong performance at low inference cost.
Specifications
Canonical IDalibaba-qwen3-5-35b-a3b
TypeLanguage
StatusActive
CreatorAlibabaAlibaba
Providers
Context Window262K tokens
Max Output66K tokens
Input ModalitiesImageTextVideo
Output ModalitiesText
Reasoning Effortsdefault
Parameters35B
HuggingFace Likes1,392
HuggingFace Downloads (30d)3,940,049
HuggingFace Downloads (all-time)6,350,856
Release Date · 5 months ago
Benchmarks
Intelligence Index
29.3
#109
GPQA
0.8
#74
HLE
0.2
#84
IFBench
0.7
#43
Time to First Token
1.18s
#400
SciCode
0.4
#148
LCR
0.6
#90
TerminalBench Hard
0.3
#111
TAU2
0.9
#62
Output TPS

Capabilities

Input3/5
Text
Image
Audio·
Video
PDF·
Output1/5
Text
Image·
Audio·
Video·
Embedding·
Capabilities4/13
Reasoning
Adaptive Reasoning·
Function Calling
Parallel Function Calling·
Structured Outputs
Native JSON Schema
Web Search·
URL Context·
Computer Use·
Code Execution·
File Search·
Prompt Caching·
Assistant Prefill·

Pricing by Provider

US Dollar ($)
Per 1M tokens
ProviderStandardBatch
Input
$ / 1M
Output
$ / 1M
Input
$ / 1M
Output
$ / 1M
Alibaba Qwen logo
Alibaba Qwen
qwen3.5-35b-a3b
$0.25$2.00$0.125$1.00
Hugging Face logo
Hugging Face
novita:qwen/qwen3.5-35b-a3b
$0.25$2.00
OpenRouter logo
OpenRouter
qwen/qwen3.5-35b-a3b
$0.14$1.00

Cost Calculator

US Dollar ($)
Preset:

Cheapest Instances to Run It

Cloud GPU instances that can host Qwen3.5 35B A3B, ranked by cheapest on-demand price. The model needs about 84 GB of GPU memory at FP16 precision (estimated from its parameter count), so treat the fit as guidance rather than a guarantee.

All clouds
FP16 (full precision)
US Dollar ($)
Instance
Cloud
GPU
VRAM
Price
Cheapest region
Standard_NP20sAzure2× AMD Alveo U250 FPGA (64GB)128 GB$3.30/hrwestus2
g7e.2xlargeAWSRTX PRO Server 600096 GB$3.36/hrus-east-1
g2-standard-48GCP4× nvidia-l496 GB$3.99/hrus-east4
7 more instances can run Qwen3.5 35B A3B
Unlock the full ranked list and FP8 / INT4 quantization with a CloudPrice subscription.

Versions

VersionReleasedContextInput / 1MOutput / 1MStatus
Qwen3.5 35B A3B262K$0.140$1.00Current
Qwen3.5 122B A10B262K$0.250$1.75Available
Qwen3.5 397B A17B262K$0.390$2.34Available
Qwen3.7 Plus VL InstructAvailable

Model IDs

accounts/fireworks/models/qwen3p5-35b-a3b
alibaba-qwen3-5-35b-a3b
openrouter/qwen/qwen3.5-35b-a3b
qwen/qwen3.5-35b-a3b
Qwen/Qwen3.5-35B-A3B
qwen3-5-35b-a3b
qwen3-5-35b-a3b-non-reasoning
qwen3.5-35b-a3b