Z.ai's native multimodal agent model for vision-based coding and agent workflows. This is the standard non-thinking variant for image, video, and text inputs, tuned for perceive-plan-execute loops, complex coding, and tool-driven task execution. Not included in the subscription.
Added Apr 1, 2026
Context Window
202.8K
Max Output
131.1K
Input Price (Auto)
$1.20/1M
Output Price (Auto)
$4.00/1M
Cache Read (Auto)
$0.24/1M
Capabilities
Benchmarks
Benchmarks
Performance metrics and benchmarks
Sourced from LMArena.
Arena Score
1433.5
Overall Rank
#88 / 389
Votes
9,397
Confidence Interval
1426.8 - 1440.3
Category Scores
Coding
#74 / 384
2,630 votes
1489.9
Math
#64 / 374
448 votes
1443.4
Longer Query
#83 / 367
4,217 votes
1443.5
Creative Writing
#90 / 387
1,634 votes
1401.8
Instruction Following
#87 / 389
3,270 votes
1423.6
Hard Prompts
#89 / 389
6,212 votes
1453.2
Additional Categories20
English
#69 / 389
4,007 votes
1454.1
Spanish
#69 / 272
282 votes
1442.7
Hard Prompts English
#70 / 388
2,661 votes
1469.2
Korean
#75 / 264
213 votes
1387.4
Chinese
#78 / 360
611 votes
1478.8
Expert
#79 / 339
1,045 votes
1462.6
Industry Software And It Services
#81 / 389
3,745 votes
1472.7
Industry Writing And Literature And Language
#82 / 388
2,367 votes
1414.1
Polish
#82 / 215
189 votes
1426.7
French
#83 / 271
394 votes
1451.2
Industry Business And Management And Financial Operations
#84 / 382
1,961 votes
1439.2
Industry Mathematical
#86 / 368
524 votes
1438.7
Exclude Ties
#87 / 389
6,906 votes
1427.2
Non English
#92 / 389
5,388 votes
1414.2
Multi Turn
#93 / 387
1,554 votes
1434.4
Industry Entertainment And Sports And Media
#94 / 387
2,177 votes
1396.7
Industry Life And Physical And Social Science
#94 / 387
1,478 votes
1449.3
Russian
#96 / 352
953 votes
1422.6
Industry Legal And Government
#112 / 361
739 votes
1425.3
Industry Medicine And Healthcare
#138 / 357
651 votes
1426.1
Published 2026-08-12 · Matched as glm-5v-turbo
LMArena DatasetProviders
Auto routing is available for this model. Explicit provider selection is not available.
Loading provider options…
Related text models
Compare GLM 5V Turbo with similar models from the same provider or model family.
GLM 5V Turbo Thinking
z-ai/glm-5v-turbo:thinkingThinking-enabled GLM 5V Turbo for image, video, and text inputs. Uses the same multimodal foundation model with more deliberate vision-grounded analysis, planning, and tool use. Not included in the subscription.
GLM 5 Turbo
z-ai/glm-5-turboFast GLM 5 Turbo variant from Z-AI for general chat, coding, and tool use. Not included in the subscription.
GLM 4.5V
z-ai/glm-4.5vMultimodal GLM 4.5V that handles images alongside text while keeping the balanced reasoning strength of the GLM 4.5 family.
GLM 4.5V Thinking
z-ai/glm-4.5v:thinkingThinking-enabled GLM 4.5V that surfaces structured reasoning before its final answer. Great for image-grounded analysis, OCR, charts, and deliberate step-by-step responses.
GLM 4.6
z-ai/glm-4.6Latest GLM series chat model with strong general performance. Quantized at FP8
GLM 4.6 Thinking
z-ai/glm-4.6:thinkingThinking version of the latest GLM series chat model with strong general performance. Quantized at FP8