GLM 4.6V Original

GLM-4.6V scales its context window to 128k tokens in training, and achieves SoTA performance in visual understanding among models of similar parameter scales. Integrates native Function Calling capabilities, bridging 'visual perception' and 'executable action' for multimodal agents. Direct via Z-AI (Zhipu).

  • Vision

Added Dec 8, 2025

Pricing

Auto routing · per 1M tokens
Input
$0.60
Output
$0.90
Compare provider prices

Specifications

Context window
128K
Max output
24K
Parameters
106B / 12B
Total / active
Avg output (7d)
466 tokens
Longer than 42% of models

Benchmarks

Sourced from Artificial Analysis.

Intelligence Index

8.4

Better than 36% of models compared

Agentic work

T²-Bench Telecom (legacy)

Legacy fallback · Conversational AI agents in dual-control scenarios

30.7%

Better than 38% of models compared

Document reasoning

AA-LCR v1.1

Long context reasoning with updated grading

17.0%

Better than 20% of models compared

Reasoning

HLE

Humanity's Last Exam

3.7%

Better than 8% of models compared

IFBench

Instruction-following benchmark

27.9%

Better than 11% of models compared

Coding

Terminal-Bench Hard (legacy)

Legacy fallback · Agentic coding and terminal use

3.0%

Better than 23% of models compared

LiveCodeBench

Contamination-free coding benchmark

41.1%

Better than 49% of models compared

Math

AIME 2025

American Invitational Mathematics Examination 2025

26.3%

Better than 28% of models compared

Knowledge

MMLU-Pro

Professional and academic subject knowledge

75.2%

Better than 50% of models compared

Legacy benchmarks

GPQA Diamond (legacy)

Graduate-level scientific reasoning

56.6%

Better than 31% of models compared

AA-LCR (unversioned / legacy)

Long context reasoning evaluation

17.0%

Better than 20% of models compared

Last updated Oct 9, 2026

Artificial Analysis

Providers

Provider information for this model’s automatic routing. These routes cannot be selected individually.

Loading provider options…