GLM high-intelligence flagship model with 128K context window
Added Sep 20, 2024
Context Window
128.0K
Max Output
4.1K
Input Price (Auto)
$7.50/1M
Output Price (Auto)
$7.50/1M
Cache Read (Auto)
$3.75/1M
Benchmarks
Benchmarks
Performance metrics and benchmarks
Sourced from LMArena.
Arena Score
1319.2
Overall Rank
#243 / 395
Votes
26,126
Confidence Interval
1314.3 - 1324.1
Category Scores
Coding
#250 / 390
4,449 votes
1352.2
Math
#244 / 379
3,599 votes
1286.9
Longer Query
#238 / 373
3,992 votes
1325.7
Creative Writing
#232 / 393
3,801 votes
1289.2
Instruction Following
#243 / 395
10,743 votes
1300.7
Hard Prompts
#254 / 395
7,067 votes
1319.0
Additional Categories21
Korean
#170 / 265
374 votes
1286.5
Spanish
#185 / 277
332 votes
1341.2
French
#189 / 276
286 votes
1345.5
Japanese
#193 / 261
485 votes
1232.5
German
#195 / 296
659 votes
1304.1
Industry Legal And Government
#209 / 368
1,723 votes
1349.8
Chinese
#220 / 367
2,427 votes
1343.5
Industry Writing And Literature And Language
#227 / 394
7,164 votes
1308.8
Industry Business And Management And Financial Operations
#229 / 388
3,267 votes
1318.6
Industry Medicine And Healthcare
#229 / 363
1,513 votes
1338.8
Non English
#231 / 395
12,935 votes
1305.4
Russian
#234 / 359
3,904 votes
1307.0
Expert
#238 / 345
1,608 votes
1303.0
Multi Turn
#238 / 393
4,844 votes
1314.1
Exclude Ties
#242 / 395
16,416 votes
1264.1
Industry Life And Physical And Social Science
#242 / 393
4,702 votes
1334.1
Industry Mathematical
#242 / 374
2,991 votes
1299.3
Industry Entertainment And Sports And Media
#243 / 393
3,855 votes
1280.1
Industry Software And It Services
#253 / 395
7,097 votes
1340.5
Hard Prompts English
#257 / 393
4,016 votes
1325.6
English
#258 / 395
13,191 votes
1325.1
Published 2026-08-27 · Matched as glm-4-plus
LMArena DatasetProviders
Auto routing is available for this model. Explicit provider selection is not available.
Loading provider options…
Related text models
Compare GLM-4 Plus with similar models from the same provider or model family.
GLM 4 Plus 0111
glm-4-plus-0111GLM 4 Plus 0111 is a 1M token context window model
GLM 5.3 Flash TEE
TEE/glm-5.3-flashGLM-5.3 Flash is Z.AI's natively multimodal 320B MoE reasoning model with 18B active parameters, served by Phala inside a Trusted Execution Environment with Redpill attestation and signed completion receipts.
GLM 5.3 Flash Uncensored
z-ai/glm-5.3-flash-uncensoredGLM 5.3 Flash Uncensored is an uncensored fine-tune of the efficient 320B mixture-of-experts reasoning model, built for unrestricted chat, creative writing, coding, agentic work, tool use, and long-context tasks.
GLM 5.3 Flash
z-ai/glm-5.3-flashox-alpha out of stealth! GLM-5.3 Flash is Z.ai's first natively multimodal GLM-5 model, with 320B total parameters and just 18B active parameters for efficient coding, agentic work, and precise 1M-token context. Its hybrid sparse-and-linear attention architecture helps it outperform GLM-5.2 at one-tenth the price while approaching Claude Opus 4.8 on coding and agentic benchmarks.
GLM 5.3
zai-org/glm-5.3GLM-5.3 for long-horizon autonomous coding and engineering workflows. This variant defaults to the model's low reasoning tier for faster responses.
GLM 5.3 Thinking
zai-org/glm-5.3:thinkingGLM-5.3 with higher reasoning enabled for harder long-horizon coding, autonomous agent workflows, and complex engineering tasks.
