Thinking-enabled GLM 5V Turbo for image, video, and text inputs. Uses the same multimodal foundation model with more deliberate vision-grounded analysis, planning, and tool use.
Added Apr 2, 2026
Context Window
202.8K
Max Output
131.1K
Input Price (Auto)
$1.20/1M
Output Price (Auto)
$4.00/1M
Cache Read (Auto)
$0.24/1M
Capabilities
Benchmarks
Benchmarks
Performance metrics and benchmarks
Sourced from Artificial Analysis.
Intelligence Index
35.3
Reasoning
GPQA Diamond
Graduate-level scientific reasoning
80.9%
Better than 71% of models compared
HLE
Humanity's Last Exam
17.1%
Better than 72% of models compared
IFBench
Instruction-following benchmark
61.1%
Better than 72% of models compared
T²-Bench Telecom
Conversational AI agents in dual-control scenarios
98.5%
Better than 98% of models compared
AA-LCR
Long context reasoning evaluation
65.7%
Better than 71% of models compared
CritPt
Research-level physics reasoning
0.6%
Coding
SciCode
Python programming for scientific computing
43.5%
Better than 80% of models compared
Terminal-Bench Hard
Agentic coding and terminal use
32.6%
Better than 77% of models compared
Knowledge
AA-Omniscience Accuracy
Proportion of correctly answered questions
29.3%
AA-Omniscience Hallucination Rate
Rate of incorrect answers among non-correct responses
68.8%
Last updated Aug 7, 2026, 12:00 AM
Artificial AnalysisProviders
Auto routing is available for this model. Explicit provider selection is not available.
Loading provider options…
Related text models
Compare GLM 5V Turbo Thinking with similar models from the same provider or model family.
GLM 5V Turbo
z-ai/glm-5v-turboZ.ai's native multimodal agent model for vision-based coding and agent workflows. This is the standard non-thinking variant for image, video, and text inputs, tuned for perceive-plan-execute loops, complex coding, and tool-driven task execution.
GLM 5 Turbo
z-ai/glm-5-turboFast GLM 5 Turbo variant from Z-AI for general chat, coding, and tool use.
GLM 5.3 Flash Uncensored
z-ai/glm-5.3-flash-uncensoredGLM 5.3 Flash Uncensored is an uncensored fine-tune of the efficient 320B mixture-of-experts reasoning model, built for unrestricted chat, creative writing, coding, agentic work, tool use, and long-context tasks.
GLM 5.3 Flash
z-ai/glm-5.3-flashox-alpha out of stealth! GLM-5.3 Flash is Z.ai's first natively multimodal GLM-5 model, with 320B total parameters and just 18B active parameters for efficient coding, agentic work, and precise 1M-token context. Its hybrid sparse-and-linear attention architecture helps it outperform GLM-5.2 at one-tenth the price while approaching Claude Opus 4.8 on coding and agentic benchmarks.
GLM 4.5V
z-ai/glm-4.5vMultimodal GLM 4.5V that handles images alongside text while keeping the balanced reasoning strength of the GLM 4.5 family.
GLM 4.5V Thinking
z-ai/glm-4.5v:thinkingThinking-enabled GLM 4.5V that surfaces structured reasoning before its final answer. Great for image-grounded analysis, OCR, charts, and deliberate step-by-step responses.
