High-intelligence model with 128K context window
Context Window
128.0K
Max Output
4.1K
Input Price (Auto)
$14.99/1M
Output Price (Auto)
$14.99/1M
Cache Read (Auto)
$7.50/1M
Benchmarks
Benchmarks
Performance metrics and benchmarks
Sourced from LMArena.
Arena Score
1273.0
Overall Rank
#288 / 395
Votes
9,788
Confidence Interval
1266.0 - 1280.0
Category Scores
Coding
#281 / 390
1,718 votes
1308.8
Math
#276 / 379
1,191 votes
1246.2
Longer Query
#290 / 373
1,174 votes
1268.1
Creative Writing
#283 / 393
1,547 votes
1240.7
Instruction Following
#286 / 395
3,766 votes
1255.8
Hard Prompts
#287 / 395
2,970 votes
1277.8
Additional Categories18
Korean
#247 / 265
302 votes
1103.8
German
#248 / 296
371 votes
1227.2
Chinese
#253 / 367
1,227 votes
1303.6
Industry Legal And Government
#265 / 368
556 votes
1312.2
Industry Mathematical
#271 / 374
1,067 votes
1264.2
Industry Medicine And Healthcare
#276 / 363
554 votes
1280.8
Russian
#281 / 359
1,264 votes
1260.6
Expert
#282 / 345
514 votes
1252.1
Multi Turn
#284 / 393
1,631 votes
1258.9
Exclude Ties
#285 / 395
6,229 votes
1191.1
Hard Prompts English
#286 / 393
1,775 votes
1284.4
Industry Life And Physical And Social Science
#286 / 393
1,590 votes
1290.1
Industry Software And It Services
#286 / 395
2,832 votes
1296.2
English
#287 / 395
5,169 votes
1288.1
Non English
#287 / 395
4,619 votes
1247.7
Industry Entertainment And Sports And Media
#289 / 393
1,605 votes
1233.4
Industry Writing And Literature And Language
#290 / 394
2,609 votes
1253.7
Industry Business And Management And Financial Operations
#295 / 388
1,186 votes
1244.9
Published 2026-08-27 · Matched as glm-4-0520
LMArena DatasetProviders
Auto routing is available for this model. Explicit provider selection is not available.
Loading provider options…
Related text models
Compare GLM-4 with similar models from the same provider or model family.
GLM 5.3 Flash TEE
TEE/glm-5.3-flashGLM-5.3 Flash is Z.AI's natively multimodal 320B MoE reasoning model with 18B active parameters, served by Phala inside a Trusted Execution Environment with Redpill attestation and signed completion receipts.
GLM 5.3 Flash Uncensored
z-ai/glm-5.3-flash-uncensoredGLM 5.3 Flash Uncensored is an uncensored fine-tune of the efficient 320B mixture-of-experts reasoning model, built for unrestricted chat, creative writing, coding, agentic work, tool use, and long-context tasks.
GLM 5.3 Flash
z-ai/glm-5.3-flashox-alpha out of stealth! GLM-5.3 Flash is Z.ai's first natively multimodal GLM-5 model, with 320B total parameters and just 18B active parameters for efficient coding, agentic work, and precise 1M-token context. Its hybrid sparse-and-linear attention architecture helps it outperform GLM-5.2 at one-tenth the price while approaching Claude Opus 4.8 on coding and agentic benchmarks.
GLM 5.3
zai-org/glm-5.3GLM-5.3 for long-horizon autonomous coding and engineering workflows. This variant defaults to the model's low reasoning tier for faster responses.
GLM 5.3 Thinking
zai-org/glm-5.3:thinkingGLM-5.3 with higher reasoning enabled for harder long-horizon coding, autonomous agent workflows, and complex engineering tasks.
GLM 5.2 TEE
TEE/glm-5.2GLM-5.2 is Z.AI's flagship model for long-horizon tasks and autonomous workflows, with a 1M-token context window. Running inside a TEE (Trusted Execution Environment), with attestation support.
