GLM 4 Plus 0111 is a 1M token context window model
Added Feb 19, 2025
Context Window
128.0K
Max Output
4.1K
Input Price (Auto)
$10.00/1M
Output Price (Auto)
$10.00/1M
Cache Read (Auto)
$5.00/1M
Benchmarks
Benchmarks
Performance metrics and benchmarks
Sourced from LMArena.
Arena Score
1342.6
Overall Rank
#214 / 395
Votes
5,760
Confidence Interval
1334.2 - 1351.0
Category Scores
Coding
#268 / 390
894 votes
1330.5
Math
#230 / 379
721 votes
1297.8
Longer Query
#217 / 373
765 votes
1339.6
Creative Writing
#199 / 393
918 votes
1316.0
Instruction Following
#227 / 395
2,160 votes
1314.4
Hard Prompts
#246 / 395
1,468 votes
1325.7
Additional Categories18
Industry Medicine And Healthcare
#152 / 363
293 votes
1414.3
German
#166 / 296
207 votes
1348.9
Chinese
#168 / 367
387 votes
1401.2
Industry Legal And Government
#179 / 368
367 votes
1381.8
Industry Writing And Literature And Language
#198 / 394
1,459 votes
1331.6
Industry Life And Physical And Social Science
#201 / 393
986 votes
1365.6
Japanese
#201 / 261
201 votes
1219.2
Non English
#205 / 395
2,294 votes
1325.1
Industry Business And Management And Financial Operations
#207 / 388
592 votes
1344.6
Russian
#208 / 359
653 votes
1329.9
Exclude Ties
#211 / 395
3,826 votes
1299.9
Multi Turn
#211 / 393
851 votes
1333.6
Industry Entertainment And Sports And Media
#217 / 393
1,062 votes
1299.9
English
#223 / 395
3,466 votes
1352.3
Expert
#233 / 345
354 votes
1310.5
Industry Mathematical
#233 / 374
638 votes
1303.5
Industry Software And It Services
#242 / 395
1,406 votes
1349.3
Hard Prompts English
#261 / 393
955 votes
1324.1
Published 2026-08-27 · Matched as glm-4-plus-0111
LMArena DatasetProviders
Auto routing is available for this model. Explicit provider selection is not available.
Loading provider options…
Related text models
Compare GLM 4 Plus 0111 with similar models from the same provider or model family.
GLM 4 Air 0111
glm-4-air-0111MiniMax's flagship model with a 1M token context window
GLM-4 Plus
glm-4-plusGLM high-intelligence flagship model with 128K context window
GLM 5.3 Flash TEE
TEE/glm-5.3-flashGLM-5.3 Flash is Z.AI's natively multimodal 320B MoE reasoning model with 18B active parameters, served by Phala inside a Trusted Execution Environment with Redpill attestation and signed completion receipts.
GLM 5.3 Flash Uncensored
z-ai/glm-5.3-flash-uncensoredGLM 5.3 Flash Uncensored is an uncensored fine-tune of the efficient 320B mixture-of-experts reasoning model, built for unrestricted chat, creative writing, coding, agentic work, tool use, and long-context tasks.
GLM 5.3 Flash
z-ai/glm-5.3-flashox-alpha out of stealth! GLM-5.3 Flash is Z.ai's first natively multimodal GLM-5 model, with 320B total parameters and just 18B active parameters for efficient coding, agentic work, and precise 1M-token context. Its hybrid sparse-and-linear attention architecture helps it outperform GLM-5.2 at one-tenth the price while approaching Claude Opus 4.8 on coding and agentic benchmarks.
GLM 5.3
zai-org/glm-5.3GLM-5.3 for long-horizon autonomous coding and engineering workflows. This variant defaults to the model's low reasoning tier for faster responses.
