Created bymario-alka.dePowered bygodcore.denoob2claw.detricoma.de
Hersteller

NVIDIA

Nemotron-Modellfamilie von NVIDIA.

12Modelle
567Benchmarks
567Performance
0Harness
Beste Token-Generierung
2.182,7 tok/s
Nemotron-3-Nano-4B
Bester Prefill
44.171,1 tok/s
Nemotron-3-Nano-4B
∅ Token-Generierung
244,5 tok/s
Schnitt aus 567 Läufen
∅ Prefill
3.504,1 tok/s
Schnitt aus 567 Läufen
Bester TTFT
155 ms
Nemotron-3.5-Lightning-30B-A3B
Beste Harness-Quote
0%
noch kein Harness-Lauf
🏆
Nemotron-3-Nano-4B · schnellstes Modell mit 2.182,7 tok/s Token-Generierung
Bester Prefill des Herstellers: 44.171,1 tok/s (Nemotron-3-Nano-4B)

Top-Modelle nach Token-Generierung

Bester gemessener Generierungs-Durchsatz je Modell (tok/s), veröffentlichte Performancebenchmarks.

Nemotron-3-Nano-4B
2.182,7 tok/s
Nemotron-3-Nano-Omni-30B-A3B-Reasoning
1.791,5 tok/s
Nemotron-3-Nano-30B-A3B
1.773,5 tok/s
Nemotron-Cascade-2-30B-A3B
1.768,1 tok/s
Nemotron-3-Super-120B-A12B
600,9 tok/s
OpenReasoning-Nemotron-32B
481,0 tok/s
Llama-3.3-Nemotron-Super-49B-v1.5
364,1 tok/s
Nemotron-3.5-Lightning-30B-A3B
290,3 tok/s