Created bymario-alka.dePowered bygodcore.denoob2claw.detricoma.de
Performance benchmark

How fast is the model?

Raw throughput on real hardware – how many tokens a model generates per second, how quickly it processes the prompt and how short the time to first token is. Higher is better.

⚡ Generation (tok/s)🔄 Prefill⏱️ TTFT👥 Concurrency
Reset
Metric:
Size
#Model / MakerMetricsParallelGPU / CPU / RAMRuntime
1gpt-oss-20b20BOpenAI Performance benchmarkTimebench 3 - Kombi (Prefill + Generation)
70,85 tok/s TG
Prefill 154 · TTFT 129.121 ms
10×Keine GPU (CPU-only)AMD Ryzen Threadripper PRO 7955WX 16-Cores · 184 GB RAMllama.cppgodclawMXFP4
2gpt-oss-20b20BOpenAI Performance benchmarkTimebench 3 - Kombi (Prefill + Generation)
53,48 tok/s TG
Prefill 153 · TTFT 64.583 ms
5×Keine GPU (CPU-only)AMD Ryzen Threadripper PRO 7955WX 16-Cores · 184 GB RAMllama.cppgodclawMXFP4
3gpt-oss-20b20BOpenAI Performance benchmarkTimebench 3 - Kombi (Prefill + Generation)
28,43 tok/s TG
Prefill 143 · TTFT 13.839 ms
1×Keine GPU (CPU-only)AMD Ryzen Threadripper PRO 7955WX 16-Cores · 184 GB RAMllama.cppgodclawMXFP4
4Qwen3-30B-A3B-Instruct-250730BQwen (Alibaba) Performance benchmarkTimebench 3 - Kombi (Prefill + Generation)
25,98 tok/s TG
Prefill 218 · TTFT 59.755 ms
5×Keine GPU (CPU-only)AMD Ryzen Threadripper PRO 7955WX 16-Cores · 184 GB RAMllama.cppgodclawQ4_K_M
5Qwen3-30B-A3B-Instruct-250730BQwen (Alibaba) Performance benchmarkTimebench 3 - Kombi (Prefill + Generation)
24,50 tok/s TG
Prefill 152 · TTFT 16.071 ms
1×Keine GPU (CPU-only)AMD Ryzen Threadripper PRO 7955WX 16-Cores · 184 GB RAMllama.cppgodclawQ4_K_M
6Qwen2.5-72B-Instruct72BQwen (Alibaba) Performance benchmark
2,28 tok/s TG
Prefill 81 · TTFT 569 ms
1×Keine GPU (CPU-only)51x Intel(R) Xeon(R) CPU E5-4620 v2 @ 2.60GHz · 504 GB RAMgodclaw