Provider logo

GLM 5 Turbo

z-ai/glm-5-turbo
Provider logo

GLM 5 Turbo

z-ai/glm-5-turbo

Fast GLM 5 Turbo variant from Z-AI for general chat, coding, and tool use.

Added Mar 15, 2026

Context Window

202.8K

Max Output

131.1K

Input Price (Auto)

$1.20/1M

Output Price (Auto)

$4.00/1M

Cache Read (Auto)

$0.24/1M

Capabilities

Benchmarks

Performance metrics and benchmarks

Sourced from Artificial Analysis.

Intelligence Index

39.1

Better than 83% of models compared

Reasoning

GPQA Diamond

Graduate-level scientific reasoning

84.7%

Better than 80% of models compared

HLE

Humanity's Last Exam

27.8%

Better than 82% of models compared

IFBench

Instruction-following benchmark

73.2%

Better than 90% of models compared

T²-Bench Telecom

Conversational AI agents in dual-control scenarios

98.5%

Better than 99% of models compared

AA-LCR

Long context reasoning evaluation

66.7%

Better than 72% of models compared

CritPt

Research-level physics reasoning

0.3%

Coding

SciCode

Python programming for scientific computing

43.6%

Better than 81% of models compared

Terminal-Bench Hard

Agentic coding and terminal use

33.3%

Better than 77% of models compared

Knowledge

AA-Omniscience Accuracy

Proportion of correctly answered questions

28.4%

AA-Omniscience Hallucination Rate

Rate of incorrect answers among non-correct responses

62.6%

Last updated Aug 18, 2026

Artificial Analysis

Providers

Auto routing is available for this model. Explicit provider selection is not available.

Loading provider options…

Compare GLM 5 Turbo with similar models from the same provider or model family.

GLM 5V Turbo Thinking

z-ai/glm-5v-turbo:thinking

Thinking-enabled GLM 5V Turbo for image, video, and text inputs. Uses the same multimodal foundation model with more deliberate vision-grounded analysis, planning, and tool use.

GLM 5V Turbo

z-ai/glm-5v-turbo

Z.ai's native multimodal agent model for vision-based coding and agent workflows. This is the standard non-thinking variant for image, video, and text inputs, tuned for perceive-plan-execute loops, complex coding, and tool-driven task execution.

GLM 5.3 Flash Uncensored

z-ai/glm-5.3-flash-uncensored

GLM 5.3 Flash Uncensored is an uncensored fine-tune of the efficient 320B mixture-of-experts reasoning model, built for unrestricted chat, creative writing, coding, agentic work, tool use, and long-context tasks.

GLM 5.3 Flash

z-ai/glm-5.3-flash

ox-alpha out of stealth! GLM-5.3 Flash is Z.ai's first natively multimodal GLM-5 model, with 320B total parameters and just 18B active parameters for efficient coding, agentic work, and precise 1M-token context. Its hybrid sparse-and-linear attention architecture helps it outperform GLM-5.2 at one-tenth the price while approaching Claude Opus 4.8 on coding and agentic benchmarks.

GLM 4.5V

z-ai/glm-4.5v

Multimodal GLM 4.5V that handles images alongside text while keeping the balanced reasoning strength of the GLM 4.5 family.

GLM 4.5V Thinking

z-ai/glm-4.5v:thinking

Thinking-enabled GLM 4.5V that surfaces structured reasoning before its final answer. Great for image-grounded analysis, OCR, charts, and deliberate step-by-step responses.