Provider logo

GLM 4.7 Flash Original

zai-org/glm-4.7-flash-original
Provider logo

GLM 4.7 Flash Original

zai-org/glm-4.7-flash-original

GLM-4.7-Flash is a lightweight 30B model optimized for coding and agentic tasks. Balances high performance with efficiency, perfect for local deployment. Routed directly via Z-AI (Zhipu) subscription.

Added Jan 19, 2026

Context Window

200.0K

Max Output

128.0K

Input Price (Auto)

$0.070/1M

Output Price (Auto)

$0.40/1M

Cache Read (Auto)

$0.035/1M

Capabilities

Benchmarks

Performance metrics and benchmarks

Sourced from Artificial Analysis.

Intelligence Index

15.6

Better than 46% of models compared

Reasoning

GPQA Diamond

Graduate-level scientific reasoning

45.2%

Better than 21% of models compared

HLE

Humanity's Last Exam

5.0%

Better than 32% of models compared

IFBench

Instruction-following benchmark

46.3%

Better than 54% of models compared

T²-Bench Telecom

Conversational AI agents in dual-control scenarios

91.8%

Better than 87% of models compared

AA-LCR

Long context reasoning evaluation

18.3%

Better than 26% of models compared

CritPt

Research-level physics reasoning

0.0%

Coding

SciCode

Python programming for scientific computing

25.5%

Better than 29% of models compared

Terminal-Bench Hard

Agentic coding and terminal use

3.8%

Better than 26% of models compared

Knowledge

AA-Omniscience Accuracy

Proportion of correctly answered questions

13.1%

AA-Omniscience Hallucination Rate

Rate of incorrect answers among non-correct responses

94.3%

Last updated Aug 16, 2026

Artificial Analysis

Providers

Choose explicit providers for this model. Auto routing remains available as the default option.

Loading provider options…