Z.ai (Zhipu)
glm-4.7-flash
Benchmark profile and published results.
Position in the field
Best values compared
Best metrics of this model against the minimum, average and maximum of all published systems.
Generation48,9 tok/s
Min 0,0Ø 253,2Max 2.491,2
Unter dem Durchschnitt · 2823 Systeme im Feld
Prefill4.055 tok/s
Min 4Ø 3.470Max 39.029
Ueber dem Durchschnitt · 2823 Systeme im Feld
Time to First Token536 ms
Min 20Ø 45.698Max 535.235
Ueber dem Durchschnitt · 2821 Systeme im Feld
Throughput & latency
Performance benchmark
| # | Model / Maker | Metrics | Parallel | GPU / CPU / RAM | Runtime | ||
|---|---|---|---|---|---|---|---|
| 1 | glm-4.7-flashAWQZ.ai (Zhipu) Performance benchmarkTimebench 3 - Kombi (Prefill + Generation) | 48,92 tok/s TG Prefill 4.055 · TTFT 536 ms | 1× | NVIDIA GB10 (DGX Spark)NVIDIA Grace · 120 GB RAM | vLLMgodclaw | Details → |
Agent & chat rating
Harness benchmark
Noch keine Harnessbenchmarks fuer dieses Modell.
