GLM-5.2 is Zhipu AI's language model with a 1.0M context window and up to 262K output tokens, available from 20 providers, starting at $0.61 / 1M input and $1.98 / 1M output. Z.AI's flagship long-horizon autonomous LLM capable of sustained multi-hour task execution with agentic reasoning.
Specifications
Canonical IDzhipu-glm-5-2
TypeLanguage
StatusActive
CreatorZhipu AIZhipu AI
Providers
Context Window1.0M tokens
Max Output262K tokens
Input ModalitiesText
Output ModalitiesText
Reasoning Effortsdefault
Parameters753B
HuggingFace Likes692
HuggingFace Downloads (30d)0
HuggingFace Downloads (all-time)0
Release Date · 3 months ago
Benchmarks
Intelligence Index
#29
Coding Index
#31
GPQA
#54
HLE
#26
IFBench
#41
Time to First Token
#432
SciCode
#37
LCR
#49
TerminalBench Hard
#14
TAU2
#2
Output TPS
#92

Capabilities

Input1/5
Text
Image·
Audio·
Video·
PDF·
Output1/5
Text
Image·
Audio·
Video·
Embedding·
Capabilities8/13
Reasoning
Adaptive Reasoning·
Function Calling
Parallel Function Calling
Structured Outputs
Native JSON Schema
Web Search
URL Context·
Computer Use·
Code Execution·
File Search·
Prompt Caching
Assistant Prefill

Pricing by Provider

US Dollar ($)
Per 1M tokens
ProviderStandardFastBatch
Input
$ / 1M
Output
$ / 1M
Cache Read
$ / 1M
Cache Write 5m
$ / 1M
Input
$ / 1M
Output
$ / 1M
Cache Read
$ / 1M
Input
$ / 1M
Output
$ / 1M
$1.40$4.40$0.28N/A$0.7$2.20
Azure AI Foundry logo
Azure AI Foundry
azure_ai/FW-GLM-5.2
$1.54$4.84$0.15N/A
Databricks logo
Databricks
databricks/databricks-glm-5-2
$1.40$4.40$0.26$1.40
DeepInfra logo
DeepInfra
deepinfra/zai-org/GLM-5.2
$0.75$2.40$0.14N/A
Fireworks AI logo
Fireworks AI
fireworks_ai/glm-5p2
$1.40$4.40$0.14N/A
$1.40$4.40$0.14N/A
Hugging Face logo
Hugging Face
novita:zai-org/glm-5.2
$1.40$4.40N/AN/A
Hugging Face logo
Hugging Face
scaleway:glm-5.2
$2.05$6.27N/AN/A
Mistral AI logo
Mistral AI
zai-glm-5-2
$1.40$4.40$0.14N/A$0.7$2.20
Nebius logo
Nebius
zai-org/GLM-5.2
$1.40$4.40N/AN/A
Novita logo
Novita
novita/zai-org/glm-5.2
$1.40$4.40$0.26N/A
OpenRouter logo
OpenRouter
z-ai/glm-5.2
$0.966$3.04$0.193N/A
Perplexity logo
Perplexity
perplexity/perplexity/glm-5.2
$1.40$4.40$0.14N/A
Qwen AI Platform
qwen_ai_platform/glm-5.2
$1.40$4.40$0.28N/A
Qwencloud
qwencloud/glm-5.2
$1.40$4.40$0.28N/A
Scaleway logo
Scaleway
scaleway/glm-5.2
$1.80$5.50N/AN/A
Scx AI
scx-ai/GLM-5.2
$0.61$1.98$0.22N/A
Together AI logo
Together AI
together_ai/zai-org/GLM-5.2
$1.40$4.40$0.26N/A
$0.8$2.55$0.16N/AN/AN/AN/A
Vercel AI Gateway logo
Vercel AI Gateway
zai/glm-5.2-fast
$2.10$6.60$0.21
Weights & Biases logo
Weights & Biases
wandb/zai-org/GLM-5.2
$0.76$2.42$0.14N/A
Z AI (Zhipu) logo
Z AI (Zhipu)
zai/glm-5.2
$1.40$4.40$0.26N/A

Cost Calculator

US Dollar ($)
Preset:

Cheapest Instances to Run It

Cloud GPU instances that can host GLM-5.2, ranked by cheapest on-demand price. The model needs about 1808 GB of GPU memory at FP16 precision (estimated from its parameter count), so treat the fit as guidance rather than a guarantee.

All clouds
FP16 (full precision)
US Dollar ($)
Instance
Cloud
GPU
VRAM
Price
Cheapest region
p6-b300.48xlargeAWS8× B3002149 GB$142.42/hr

Versions

VersionReleasedContextInput / 1MOutput / 1MStatus
GLM-5.3 Flash1.3M$0.075$0.250Available
GLM-5.31.3M$0.700$2.20Available
GLM-5.3 (50% off)1.0M$0.700$2.20Available
GLM-5.2 Fast1.0M$2.10$6.60Available
GLM-5.21.0M$0.610$1.98Current
GLM-5V Turbo205K$1.20$4.00Available
GLM-5 Turbo203K$1.20$4.00Available
GLM-5.1 Non-ReasoningAvailable
GLM-5 Non-ReasoningAvailable
GLM-5 Code200K$1.20$5.00Available
GLM-5.1 Fast203K$2.80$8.80Available

Model IDs

accounts/fireworks/models/glm-5p2
accounts/fireworks/models/glm-5p2-fp8
azure_ai/FW-GLM-5.2
cloudflare/@cf/zai-org/glm-5.2
dashscope/glm-5.2
databricks/databricks-glm-5-2
deepinfra/zai-org/GLM-5.2
fireworks_ai/accounts/fireworks/models/glm-5p2
fireworks_ai/glm-5p2
glm-5-2
glm-5-2-non-reasoning
glm-5.2
glm-5.2-maas
huggingface-llm-glm-5-2-fp8
mistral/glm-5-2
mistral/zai-glm-5-2
nebius/zai-org/GLM-5.2
novita/zai-org/glm-5.2
openrouter/z-ai/glm-5.2
openrouter/z-ai/glm-5.2:free
perplexity/perplexity/glm-5.2
qwen_ai_platform/glm-5.2
qwencloud/glm-5.2
scaleway/glm-5.2
scx-ai/GLM-5.2
together_ai/zai-org/GLM-5.2
wandb/zai-org/GLM-5.2
z-ai/glm-5.2
z-ai/glm-5.2:batch
zai-glm-5-2
zai-org/glm-5.2
zai-org/GLM-5.2
zai/glm-5.2
zhipu-glm-5-2