GLM-5.2 is Zhipu AI's language model with a 1.0M context window and up to 944K output tokens, available from 22 providers, starting at $0.36 / 1M input and $1.98 / 1M output. Z.AI's flagship long-horizon autonomous LLM capable of sustained multi-hour task execution with agentic reasoning.
Specifications
Canonical IDzhipu-glm-5-2
TypeLanguage
StatusDeprecated
CreatorZhipu AIZhipu AI
Providers
Context Window1.0M tokens
Max Output944K tokens
Input ModalitiesImageText
Output ModalitiesText
Reasoning Effortsdefault
Parameters753B
HuggingFace Likes692
HuggingFace Downloads (30d)0
HuggingFace Downloads (all-time)0
Release Date · 4 months ago
Deprecation Date
Benchmarks
Intelligence Index
#55
Coding Index
#31
GPQA
#54
HLE
#33
IFBench
#40
Time to First Token
#434
SciCode
#49
LCR
#59
TerminalBench Hard
#14
TAU2
#2
Output TPS
#57

Capabilities

Input2/5
Text✓
Image✓
Audio·
Video·
PDF·
Output1/5
Text✓
Image·
Audio·
Video·
Embedding·
Capabilities8/13
Reasoning✓
Adaptive Reasoning·
Function Calling✓
Parallel Function Calling✓
Structured Outputs✓
Native JSON Schema✓
Web Search✓
URL Context·
Computer Use·
Code Execution·
File Search·
Prompt Caching✓
Assistant Prefill✓

Pricing by Provider

US Dollar ($)
Per 1M tokens
ProviderStandardFastBatch
Input
$ / 1M
Output
$ / 1M
Cache Read
$ / 1M
Cache Write 5m
$ / 1M
Input
$ / 1M
Output
$ / 1M
Cache Read
$ / 1M
Input
$ / 1M
Output
$ / 1M
$1.40$4.40$0.28N/A———$0.7$2.20
Azure AI Foundry logo
Azure AI Foundry
azure_ai/FW-GLM-5.2
$1.54$4.84$0.15N/A—————
Baseten
baseten/zai-org/GLM-5.2
$1.40$4.40$0.14N/A—————
Databricks logo
Databricks
databricks/databricks-glm-5-2
$1.40$4.40$0.26$1.40—————
DeepInfra logo
DeepInfra
deepinfra/zai-org/GLM-5.2
$0.75$2.40$0.14N/A—————
Fireworks AI logo
Fireworks AI
fireworks_ai/glm-5p2
$1.40$4.40$0.14N/A$1.75$5.50$0.175——
FriendliAI logo
FriendliAI
friendliai/zai-org/GLM-5.2
$1.40$4.40$0.26N/A—————
$1.40$4.40$0.14N/A—————
Hugging Face logo
Hugging Face
novita:zai-org/glm-5.2
$1.40$4.40N/AN/A—————
Hugging Face logo
Hugging Face
scaleway:glm-5.2
$2.05$6.27N/AN/A—————
Mistral AI logo
Mistral AI
mistral/glm-5-2
$1.40$4.40$0.14N/A—————
Nebius logo
Nebius
zai-org/GLM-5.2
$1.40$4.40N/AN/A—————
Novita logo
Novita
novita/zai-org/glm-5.2
$1.40$4.40$0.26N/A—————
OpenRouter logo
OpenRouter
z-ai/glm-5.2
$0.36$3.99$0.26N/A—————
Perplexity logo
Perplexity
perplexity/perplexity/glm-5.2
$1.40$4.40$0.14N/A—————
Qwen AI Platform
qwen_ai_platform/glm-5.2
$1.40$4.40$0.28N/A—————
Qwencloud
qwencloud/glm-5.2
$1.40$4.40$0.28N/A—————
Scaleway logo
Scaleway
scaleway/glm-5.2
$1.80$5.50N/AN/A—————
Scx AI
scx-ai/GLM-5.2
$0.61$1.98$0.22N/A—————
Together AI logo
Together AI
together_ai/zai-org/GLM-5.2
$1.40$4.40$0.26N/A—————
$0.8$2.55$0.16N/AN/AN/AN/A——
Vercel AI Gateway logo
Vercel AI Gateway
zai/glm-5.2-fast
————$2.80$8.80$0.56——
Weights & Biases logo
Weights & Biases
wandb/zai-org/GLM-5.2
$0.76$2.42$0.14N/A—————
Z AI (Zhipu) logo
Z AI (Zhipu)
zai/glm-5.2
$1.40$4.40$0.26N/A—————

Cost Calculator

US Dollar ($)
Preset:

Cheapest Instances to Run It

Cloud GPU instances that can host GLM-5.2, ranked by cheapest on-demand price. The model needs about 1808 GB of GPU memory at FP16 precision (estimated from its parameter count), so treat the fit as guidance rather than a guarantee.

All clouds
FP16 (full precision)
US Dollar ($)
Instance
Cloud
GPU
VRAM
Price
Cheapest region
p6-b300.48xlargeAWS8× B3002149 GB$142.42/hr

Versions

VersionReleasedContextInput / 1MOutput / 1MStatus
GLM-5.3 Prime1.0M$2.80$8.80Available
GLM-5.3 FlashX1.0M$0.370$1.25Available
GLM-5.3 Flash1.3M$0.110$0.350Available
GLM-5.31.3M$0.980$3.08Available
GLM-5.3 (50% off)1.0M——Available
GLM-5.2 Fast1.0M$2.10$6.60Deprecated
GLM-5.21.0M$0.360$1.98Current
GLM-5V Turbo205K$0.704$3.10Available
GLM-5 Turbo203K$1.20$4.00Available
GLM-5.1 Non-Reasoning————Available
GLM-5 Non-Reasoning————Available

Model IDs

accounts/fireworks/models/glm-5p2
accounts/fireworks/models/glm-5p2-fp8
azure_ai/FW-GLM-5.2
baseten/zai-org/GLM-5.2
cloudflare/@cf/zai-org/glm-5.2
dashscope/glm-5.2
databricks/databricks-glm-5-2
deepinfra/zai-org/GLM-5.2
fireworks_ai/accounts/fireworks/models/glm-5p2
fireworks_ai/glm-5p2
friendliai/zai-org/GLM-5.2
glm-5-2
glm-5-2-non-reasoning
glm-5.2
glm-5.2-maas
huggingface-llm-glm-5-2-fp8
mistral/glm-5-2
mistral/zai-glm-5-2
nebius/zai-org/GLM-5.2
novita/zai-org/glm-5.2
openrouter/z-ai/glm-5.2
openrouter/z-ai/glm-5.2:batch
openrouter/z-ai/glm-5.2:free
perplexity/perplexity/glm-5.2
qwen_ai_platform/glm-5.2
qwencloud/glm-5.2
scaleway/glm-5.2
scx-ai/GLM-5.2
together_ai/zai-org/GLM-5.2
vertex_ai/zai-org/glm-5.2-maas
wandb/zai-org/GLM-5.2
z-ai/glm-5.2
z-ai/glm-5.2:batch
zai-glm-5-2
zai-org/glm-5.2
zai-org/GLM-5.2
zai/glm-5.2
zhipu-glm-5-2