Created bymario-alka.dePowered bygodcore.denoob2claw.detricoma.de
Manufacturer

Z.ai (Zhipu)

GLM-Modellfamilie von Z.ai / Zhipu AI.

3Models
63Benchmarks
63Performance
0Harness
Best token generation
661,8 tok/s
GLM-4.5-Air
Best prefill
11.803,5 tok/s
GLM-4.5-Air
∅ Token generation
112,5 tok/s
Average of 63 runs
∅ Prefill
1.977,2 tok/s
Average of 63 runs
Best TTFT
407 ms
GLM-4.5-Air
Best harness rate
0%
no harness run yet
🏆
GLM-4.5-Air · fastest model at 661,8 tok/s token generation
Best prefill of this manufacturer: 11.803,5 tok/s (GLM-4.5-Air)

Top models by token generation

Best measured generation throughput per model (tok/s), published performance benchmarks.

GLM-4.5-Air
661,8 tok/s
glm-4.7-flash
48,9 tok/s
glm-5.2-colibri
0,0 tok/s