Created bymario-alka.dePowered bygodcore.denoob2claw.detricoma.de
🎮
Grafikkarte

Tesla V100-PCIE-32GB

NVIDIA32 GB VRAM
24Benchmarks
6Modelle
24Performance
0Harness
Beste Generation
361,8 tok/s
gemma-4-E2B-it
∅ Generation
136,5 tok/s
Schnitt aus 24 Läufen
Bester Prefill
7.233 tok/s
Prompt-Verarbeitung
Bester TTFT
378 ms
Time to First Token
Beste Harness-Quote
%
noch kein Harness
🏆
gemma-4-E2B-it läuft am schnellsten auf dieser GPU · 361,8 tok/s
Bester Prefill: 7.233 tok/s · 6 Modelle getestet

Top-Modelle

Bester gemessener Durchsatz je Modell auf Tesla V100-PCIE-32GB – umschaltbar nach Generation, Prefill oder Kombiwert.

gemma-4-E2B-it5B
361,8 tok/s
Ornith-1.0-9B9B
248,7 tok/s
gpt-oss-20b20B
206,6 tok/s
gemma-4-12B-it12B
49,4 tok/s
Anzeige
Performanceprofil

Durchsatz auf Tesla V100-PCIE-32GB

Jede Blase steht für ein Modell, eine Engine bzw. eine CPU – Position: Prompt-Verarbeitung (X) × Ausgabe (Y), Blasengröße: Anzahl Messläufe.

MODnach Modell

4613452301150,002.3614.7237.084Prefill (tok/s)Generation (tok/s)gemma-4-E2B-it - 361,8 tok/s Generation, 5.868 tok/s Prefill, TTFT 1.823 ms (6 Laufe)gemma-4-E2B-itOrnith-1.0-9B - 248,7 tok/s Generation, 2.452 tok/s Prefill, TTFT 4.148 ms (6 Laufe)Ornith-1.0-9Bgpt-oss-20b - 206,6 tok/s Generation, 2.660 tok/s Prefill, TTFT 4.402 ms (3 Laufe)gpt-oss-20bMeta-Llama-3.1-8B-Instruct - 122,6 tok/s Generation, 4.984 tok/s Prefill, TTFT 3.619 ms (3 Laufe)Meta-Llama-3.1-8B-Ins...Ministral-3-14B-Reasoning-2512 - 52,8 tok/s Generation, 1.847 tok/s Prefill, TTFT 9.122 ms (3 Laufe)Ministral-3-14B-Reaso...gemma-4-12B-it - 49,4 tok/s Generation, 1.066 tok/s Prefill, TTFT 15.600 ms (3 Laufe)gemma-4-12B-it
gemma-4-E2B-it 361,8 tok/sOrnith-1.0-9B 248,7 tok/sgpt-oss-20b 206,6 tok/sMeta-Llama-3.1-8B-Instruct 122,6 tok/sMinistral-3-14B-Reasoning-2512 52,8 tok/sgemma-4-12B-it 49,4 tok/s

ENGnach Engine

3983803623443263.1963.3323.4683.604Prefill (tok/s)Generation (tok/s)llama.cpp - 361,8 tok/s Generation, 3.400 tok/s Prefill, TTFT 5.586 ms (24 Laufe)llama.cpp
llama.cpp 361,8 tok/s

CPUnach Prozessor

3983803623443263.1963.3323.4683.604Prefill (tok/s)Generation (tok/s)AMD Ryzen 3 PRO 3200GE w/ Radeon Vega Graphics - 361,8 tok/s Generation, 3.400 tok/s Prefill, TTFT 5.586 ms (24 Laufe)AMD Ryzen 3 PRO 3200G...
AMD Ryzen 3 PRO 3200GE w/ Radeon Vega Graphics 361,8 tok/s
Durchsatz & Latenz

Performancebenchmark

24 veroeffentlichte Performance-Läufe auf Tesla V100-PCIE-32GB.

Im Leaderboard →
Messwert:
Größe
#Modell / HerstellerMesswerteParallelGPU / CPU / RAMRuntime
1gemma-4-E2B-it5BGoogle PerformancebenchmarkTimebench 3 - Kombi (Prefill + Generation)
361,84 tok/s TG
Prefill 7.233 · TTFT 2.750 ms
10×Tesla V100-PCIE-32GBAMD Ryzen 3 PRO 3200GE w/ Radeon Vega Graphics · 61 GB RAMllama.cppgodclawQ8_0
2gemma-4-E2B-it5BGoogle PerformancebenchmarkTimebench 3 - Kombi (Prefill + Generation)
353,85 tok/s TG
Prefill 4.758 · TTFT 4.353 ms
10×Tesla V100-PCIE-32GBAMD Ryzen 3 PRO 3200GE w/ Radeon Vega Graphics · 61 GB RAMllama.cppgodclawQ4_K_M
3gemma-4-E2B-it5BGoogle PerformancebenchmarkTimebench 3 - Kombi (Prefill + Generation)
275,62 tok/s TG
Prefill 6.734 · TTFT 1.481 ms
5×Tesla V100-PCIE-32GBAMD Ryzen 3 PRO 3200GE w/ Radeon Vega Graphics · 61 GB RAMllama.cppgodclawQ8_0
4gemma-4-E2B-it5BGoogle PerformancebenchmarkTimebench 3 - Kombi (Prefill + Generation)
255,41 tok/s TG
Prefill 6.332 · TTFT 1.575 ms
5×Tesla V100-PCIE-32GBAMD Ryzen 3 PRO 3200GE w/ Radeon Vega Graphics · 61 GB RAMllama.cppgodclawQ4_K_M
5Ornith-1.0-9B9BDeepReinforce PerformancebenchmarkTimebench 3 - Kombi (Prefill + Generation)
248,72 tok/s TG
Prefill 3.178 · TTFT 5.437 ms
10×Tesla V100-PCIE-32GBAMD Ryzen 3 PRO 3200GE w/ Radeon Vega Graphics · 61 GB RAMllama.cppgodclawQ4_K_M
6gpt-oss-20b20BOpenAI PerformancebenchmarkTimebench 3 - Kombi (Prefill + Generation)
206,60 tok/s TG
Prefill 3.450 · TTFT 2.889 ms
5×Tesla V100-PCIE-32GBAMD Ryzen 3 PRO 3200GE w/ Radeon Vega Graphics · 61 GB RAMllama.cppgodclawQ8_0
7Ornith-1.0-9B9BDeepReinforce PerformancebenchmarkTimebench 3 - Kombi (Prefill + Generation)
193,56 tok/s TG
Prefill 3.011 · TTFT 2.862 ms
5×Tesla V100-PCIE-32GBAMD Ryzen 3 PRO 3200GE w/ Radeon Vega Graphics · 61 GB RAMllama.cppgodclawQ4_K_M
8gpt-oss-20b20BOpenAI PerformancebenchmarkTimebench 3 - Kombi (Prefill + Generation)
138,88 tok/s TG
Prefill 2.354 · TTFT 9.328 ms
10×Tesla V100-PCIE-32GBAMD Ryzen 3 PRO 3200GE w/ Radeon Vega Graphics · 61 GB RAMllama.cppgodclawQ8_0
9gpt-oss-20b20BOpenAI PerformancebenchmarkTimebench 3 - Kombi (Prefill + Generation)
132,28 tok/s TG
Prefill 2.178 · TTFT 988 ms
1×Tesla V100-PCIE-32GBAMD Ryzen 3 PRO 3200GE w/ Radeon Vega Graphics · 61 GB RAMllama.cppgodclawQ8_0
10gemma-4-E2B-it5BGoogle PerformancebenchmarkTimebench 3 - Kombi (Prefill + Generation)
128,96 tok/s TG
Prefill 4.924 · TTFT 402 ms
1×Tesla V100-PCIE-32GBAMD Ryzen 3 PRO 3200GE w/ Radeon Vega Graphics · 61 GB RAMllama.cppgodclawQ4_K_M
11Meta-Llama-3.1-8B-Instruct8BMeta PerformancebenchmarkTimebench 3 - Kombi (Prefill + Generation)
122,55 tok/s TG
Prefill 5.151 · TTFT 2.489 ms
5×Tesla V100-PCIE-32GBAMD Ryzen 3 PRO 3200GE w/ Radeon Vega Graphics · 61 GB RAMllama.cppgodclawQ8_0
12gemma-4-E2B-it5BGoogle PerformancebenchmarkTimebench 3 - Kombi (Prefill + Generation)
117,98 tok/s TG
Prefill 5.227 · TTFT 378 ms
1×Tesla V100-PCIE-32GBAMD Ryzen 3 PRO 3200GE w/ Radeon Vega Graphics · 61 GB RAMllama.cppgodclawQ8_0
13Ornith-1.0-9B9BDeepReinforce PerformancebenchmarkTimebench 3 - Kombi (Prefill + Generation)
93,40 tok/s TG
Prefill 3.060 · TTFT 3.167 ms
5×Tesla V100-PCIE-32GBAMD Ryzen 3 PRO 3200GE w/ Radeon Vega Graphics · 61 GB RAMllama.cppgodclawQ8_0
14Ornith-1.0-9B9BDeepReinforce PerformancebenchmarkTimebench 3 - Kombi (Prefill + Generation)
83,75 tok/s TG
Prefill 1.512 · TTFT 1.139 ms
1×Tesla V100-PCIE-32GBAMD Ryzen 3 PRO 3200GE w/ Radeon Vega Graphics · 61 GB RAMllama.cppgodclawQ4_K_M
15Meta-Llama-3.1-8B-Instruct8BMeta PerformancebenchmarkTimebench 3 - Kombi (Prefill + Generation)
83,10 tok/s TG
Prefill 6.957 · TTFT 7.450 ms
10×Tesla V100-PCIE-32GBAMD Ryzen 3 PRO 3200GE w/ Radeon Vega Graphics · 61 GB RAMllama.cppgodclawQ8_0
16Ornith-1.0-9B9BDeepReinforce PerformancebenchmarkTimebench 3 - Kombi (Prefill + Generation)
74,39 tok/s TG
Prefill 1.748 · TTFT 11.355 ms
10×Tesla V100-PCIE-32GBAMD Ryzen 3 PRO 3200GE w/ Radeon Vega Graphics · 61 GB RAMllama.cppgodclawQ8_0
17Meta-Llama-3.1-8B-Instruct8BMeta PerformancebenchmarkTimebench 3 - Kombi (Prefill + Generation)
69,93 tok/s TG
Prefill 2.843 · TTFT 920 ms
1×Tesla V100-PCIE-32GBAMD Ryzen 3 PRO 3200GE w/ Radeon Vega Graphics · 61 GB RAMllama.cppgodclawQ8_0
18Ornith-1.0-9B9BDeepReinforce PerformancebenchmarkTimebench 3 - Kombi (Prefill + Generation)
62,07 tok/s TG
Prefill 2.204 · TTFT 932 ms
1×Tesla V100-PCIE-32GBAMD Ryzen 3 PRO 3200GE w/ Radeon Vega Graphics · 61 GB RAMllama.cppgodclawQ8_0
19Ministral-3-14B-Reasoning-251214BMistral AI PerformancebenchmarkTimebench 3 - Kombi (Prefill + Generation)
52,78 tok/s TG
Prefill 1.962 · TTFT 6.490 ms
5×Tesla V100-PCIE-32GBAMD Ryzen 3 PRO 3200GE w/ Radeon Vega Graphics · 61 GB RAMllama.cppgodclawQ8_0
20Ministral-3-14B-Reasoning-251214BMistral AI PerformancebenchmarkTimebench 3 - Kombi (Prefill + Generation)
52,41 tok/s TG
Prefill 1.260 · TTFT 19.903 ms
10×Tesla V100-PCIE-32GBAMD Ryzen 3 PRO 3200GE w/ Radeon Vega Graphics · 61 GB RAMllama.cppgodclawQ8_0
21gemma-4-12B-it12BGoogle PerformancebenchmarkTimebench 3 - Kombi (Prefill + Generation)
49,35 tok/s TG
Prefill 778 · TTFT 25.706 ms
10×Tesla V100-PCIE-32GBAMD Ryzen 3 PRO 3200GE w/ Radeon Vega Graphics · 61 GB RAMllama.cppgodclawQ8_0
22gemma-4-12B-it12BGoogle PerformancebenchmarkTimebench 3 - Kombi (Prefill + Generation)
47,87 tok/s TG
Prefill 1.491 · TTFT 7.113 ms
5×Tesla V100-PCIE-32GBAMD Ryzen 3 PRO 3200GE w/ Radeon Vega Graphics · 61 GB RAMllama.cppgodclawQ8_0
23gemma-4-12B-it12BGoogle PerformancebenchmarkTimebench 3 - Kombi (Prefill + Generation)
36,37 tok/s TG
Prefill 928 · TTFT 13.981 ms
1×Tesla V100-PCIE-32GBAMD Ryzen 3 PRO 3200GE w/ Radeon Vega Graphics · 61 GB RAMllama.cppgodclawQ8_0
24Ministral-3-14B-Reasoning-251214BMistral AI PerformancebenchmarkTimebench 3 - Kombi (Prefill + Generation)
35,13 tok/s TG
Prefill 2.320 · TTFT 975 ms
1×Tesla V100-PCIE-32GBAMD Ryzen 3 PRO 3200GE w/ Radeon Vega Graphics · 61 GB RAMllama.cppgodclawQ8_0
Agent- & Chat-Bewertung

Harnessbenchmark

0 veroeffentlichte Harness-Läufe auf Tesla V100-PCIE-32GB.

Im Leaderboard →
Noch keine Harnessbenchmarks auf dieser Hardware.

← Alle Hardware

Technische Daten

Technische Daten – Tesla V100-PCIE-32GB

HerstellerNVIDIA
VRAM32 GB
Beste Generation361,8 tok/s
Bester Prefill7.233 tok/s
Modelle6
Messläufe (Performance)24