GLM-5.1 is Zhipu AI's language model with a 205K context window and up to 203K output tokens, available from 11 providers, starting at $0.966 / 1M input and $3.04 / 1M output. Z AI's next-generation flagship agentic LLM with significantly stronger coding capabilities, vision support, and file input, achieving top performance on SWE-Bench.
Specifications
Canonical IDzhipu-glm-5-1
TypeLanguage
StatusActive
CreatorZhipu AIZhipu AI
Providers
Context Window205K tokens
Max Output203K tokens
Input ModalitiesImagePDFText
Output ModalitiesText
Reasoning Effortsdefault
Parameters754B
HuggingFace Likes1,449
HuggingFace Downloads (30d)147,738
HuggingFace Downloads (all-time)147,738
Release Date · 5 months ago
Deprecation Date
Benchmarks
Intelligence Index
#80
Coding Index
#56
GPQA
#77
HLE
#73
IFBench
#22
Time to First Token
#440
SciCode
#91
LCR
#91
TerminalBench Hard
#34
TAU2
#12
Output TPS
#192

Capabilities

Input3/5
Text
Image
Audio·
Video·
PDF
Output1/5
Text
Image·
Audio·
Video·
Embedding·
Capabilities6/13
Reasoning
Adaptive Reasoning·
Function Calling
Parallel Function Calling
Structured Outputs
Native JSON Schema
Web Search·
URL Context·
Computer Use·
Code Execution·
File Search·
Prompt Caching
Assistant Prefill·

Pricing by Provider

US Dollar ($)
Per 1M tokens
ProviderStandardBatch
Input
$ / 1M
Output
$ / 1M
Cache Read
$ / 1M
Input
$ / 1M
Output
$ / 1M
$1.40$4.40$0.26$0.7$2.20
Azure AI Foundry logo
Azure AI Foundry
azure_ai/FW-GLM-5.1
$1.54$4.84$0.286
DeepInfra logo
DeepInfra
deepinfra/zai-org/GLM-5.1
$1.05$3.50$0.205
Fireworks AI logo
Fireworks AI
fireworks_ai/glm-5p1
$1.40$4.40$0.26
Nebius logo
Nebius
zai-org/GLM-5.1
$1.40$4.40N/A
Novita logo
Novita
novita/zai-org/glm-5.1
$1.38$4.40$0.26
OpenRouter logo
OpenRouter
z-ai/glm-5.1
$0.966$3.04$0.179
Qwen AI Platform
qwen_ai_platform/glm-5.1
$1.40$4.40$0.26
Qwencloud
qwencloud/glm-5.1
$1.40$4.40$0.26
$1.40$4.40$0.26
Z AI (Zhipu) logo
Z AI (Zhipu)
zai/glm-5.1
$1.40$4.40$0.26

Cost Calculator

US Dollar ($)
Preset:

Cheapest Instances to Run It

Cloud GPU instances that can host GLM-5.1, ranked by cheapest on-demand price. The model needs about 1809 GB of GPU memory at FP16 precision (estimated from its parameter count), so treat the fit as guidance rather than a guarantee.

All clouds
FP16 (full precision)
US Dollar ($)
Instance
Cloud
GPU
VRAM
Price
Cheapest region
p6-b300.48xlargeAWS8× B3002149 GB$142.42/hr

Versions

VersionReleasedContextInput / 1MOutput / 1MStatus
GLM-5.1205K$0.966$3.04Current
GLM-5205K$0.573$1.92Available

Model IDs

accounts/fireworks/models/glm-5p1
azure_ai/FW-GLM-5.1
dashscope/glm-5.1
deepinfra/zai-org/GLM-5.1
fireworks_ai/accounts/fireworks/models/glm-5p1
fireworks_ai/glm-5p1
glm-5-1
glm-5-1-non-reasoning
glm-5.1
huggingface-llm-glm-5-1-fp8
nebius/zai-org/GLM-5.1
novita/zai-org/glm-5.1
openrouter/z-ai/glm-5.1
qwen_ai_platform/glm-5.1
qwencloud/glm-5.1
z-ai/glm-5.1
zai-org/glm-5.1
zai-org/GLM-5.1
zai/glm-5.1
zhipu-glm-5-1