Qwen3 Next 80B A3B is Alibaba's language model with a 128K context window and up to 8K output tokens, available from 2 providers, starting at $0.18 / 1M input and $1.44 / 1M output. A next-generation Qwen3 MoE LLM with 80B total and 3B activated parameters, featuring a hybrid attention architecture for efficient text generation.
Specifications
Canonical IDalibaba-qwen3-next-80b-a3b
TypeLanguage
StatusActive
CreatorAlibabaAlibaba
Providers
Context Window128K tokens
Max Output8K tokens
Input ModalitiesText
Output ModalitiesText
Parameters80B
Release Date · 9 months ago
Benchmarks
Intelligence Index
#248
Coding Index
#154
Math Index
#67
MMLU-Pro
#76
GPQA
#202
HLE
#168
LiveCodeBench
#39
IFBench
#116
Time to First Token
#461
SciCode
#171
AIME 2025
#67
LCR
#160
TerminalBench Hard
#222
TAU2
#211
Output TPS
#31

Capabilities

Input1/5
Text
Image·
Audio·
Video·
PDF·
Output1/5
Text
Image·
Audio·
Video·
Embedding·
Capabilities2/13
Reasoning·
Adaptive Reasoning·
Function Calling
Parallel Function Calling·
Structured Outputs·
Native JSON Schema
Web Search·
URL Context·
Computer Use·
Code Execution·
File Search·
Prompt Caching·
Assistant Prefill·

Pricing by Provider

US Dollar ($)
Per 1M tokens
ProviderStandardFlexFastBatch
Input
$ / 1M
Output
$ / 1M
Input
$ / 1M
Output
$ / 1M
Input
$ / 1M
Output
$ / 1M
Input
$ / 1M
Output
$ / 1M
Amazon Bedrock logo
Amazon Bedrock
qwen.qwen3-next-80b-a3b
$0.18$1.41$0.09$0.71$0.32$2.47$0.09$0.71
Snowflake logo
Snowflake
qwen3-next-80b-a3b
$0.18$1.44$0.09$0.72

Cost Calculator

US Dollar ($)
Preset:

Cheapest Instances to Run It

Cloud GPU instances that can host Qwen3 Next 80B A3B, ranked by cheapest on-demand price. The model needs about 192 GB of GPU memory at FP16 precision (estimated from its parameter count), so treat the fit as guidance rather than a guarantee.

All clouds
FP16 (full precision)
US Dollar ($)
Instance
Cloud
GPU
VRAM
Price
Cheapest region
Standard_NP40sAzure4× AMD Alveo U250 FPGA (64GB)256 GB$6.60/hr
g2-standard-96GCP8× nvidia-l4192 GB$7.98/hr
g7e.12xlargeAWS2× RTX PRO Server 6000192 GB$8.29/hr
7 more instances can run Qwen3 Next 80B A3B
Unlock the full ranked list and FP8 / INT4 quantization with a CloudPrice subscription.

Versions

VersionReleasedContextInput / 1MOutput / 1MStatus
Qwen3.8 2.4T A95B1.0M$2.00$6.00Available
Qwen Audio 3 Flash ASRAvailable
Qwen Audio 3 Flash ASR FiletransAvailable
Qwen Audio 3 Flash ASR StreamingAvailable
Qwen Audio 3 Flash Realtime$0.450$4.50Available
Qwen Audio 3 Plus Realtime$0.800$6.40Available
Qwen Audio 3 Plus TTSAvailable
Qwen Image 3Available
Qwen Image 3.0 ProAvailable
EAGLE Qwen 2.5 3B InstructAvailable
Qwen3 Next 80B A3B128K$0.180$1.44Current

Model IDs

alibaba-qwen3-next-80b-a3b
qwen.qwen3-next-80b-a3b
qwen3-next-80b-a3b-reasoning