Created bymario-alka.dePowered bygodcore.denoob2claw.detricoma.de
Performance benchmark

How fast is the model?

Raw throughput on real hardware – how many tokens a model generates per second, how quickly it processes the prompt and how short the time to first token is. Higher is better.

⚡ Generation (tok/s)🔄 Prefill⏱️ TTFT👥 Concurrency
Reset
Metric:
#Model / MakerMetricsParallelGPU / CPU / RAMRuntime
1gemma-4-31B-it31BGoogle Performance benchmarkTimebench 3 - Kombi (Prefill + Generation)
19,50 tok/s TG
Prefill 339 · TTFT 80.766 ms
5×NVIDIA GeForce RTX 3090 TiAMD Ryzen 9 8945HX with Radeon Graphics · 92 GB RAMllmcodex_cli