GLM 4.6V Original
GLM-4.6V scales its context window to 128k tokens in training, and achieves SoTA performance in visual understanding among models of similar parameter scales. Integrates native Function Calling capabilities, bridging 'visual perception' and 'executable action' for multimodal agents. Direct via Z-AI (Zhipu).
- Vision
Added Dec 8, 2025
Pricing
Auto routing · per 1M tokens- Input
- $0.60
- Output
- $0.90
Specifications
- Context window
- 128K
- Max output
- 24K
- Parameters
- 106B / 12B
- Total / active
- Avg output (7d)
- 466 tokens
- Longer than 42% of models
Benchmarks
Benchmarks
Sourced from Artificial Analysis.
Intelligence Index
8.4
Agentic work
T²-Bench Telecom (legacy)
Legacy fallback · Conversational AI agents in dual-control scenarios
30.7%
Better than 38% of models compared
Document reasoning
AA-LCR v1.1
Long context reasoning with updated grading
17.0%
Better than 20% of models compared
Reasoning
HLE
Humanity's Last Exam
3.7%
Better than 8% of models compared
IFBench
Instruction-following benchmark
27.9%
Better than 11% of models compared
Coding
Terminal-Bench Hard (legacy)
Legacy fallback · Agentic coding and terminal use
3.0%
Better than 23% of models compared
LiveCodeBench
Contamination-free coding benchmark
41.1%
Better than 49% of models compared
Math
AIME 2025
American Invitational Mathematics Examination 2025
26.3%
Better than 28% of models compared
Knowledge
MMLU-Pro
Professional and academic subject knowledge
75.2%
Better than 50% of models compared
Legacy benchmarks
GPQA Diamond (legacy)
Graduate-level scientific reasoning
56.6%
Better than 31% of models compared
AA-LCR (unversioned / legacy)
Long context reasoning evaluation
17.0%
Better than 20% of models compared
Last updated Oct 9, 2026
Artificial AnalysisProviders
Provider information for this model’s automatic routing. These routes cannot be selected individually.
Loading provider options…