M
Hersteller
Mistral AI
Franzoesisches KI-Labor; Mistral, Mixtral, Codestral, Devstral, Ministral, Magistral.
Beste Token-Generierung
58,5 tok/s
Mistral-Small-4-119B-2603
Bester Prefill
9.072,4 tok/s
Mistral-Small-3.1-24B-Instruct-2503
∅ Token-Generierung
28,4 tok/s
Schnitt aus 35 Laeufen
∅ Prefill
1.954,6 tok/s
Schnitt aus 33 Laeufen
Bester TTFT
245 ms
Mistral-Small-3.1-24B-Instruct-2503
Beste Harness-Quote
87%
Devstral-Small-2-24B-Instruct-2512
🏆
Mistral-Small-4-119B-2603 · schnellstes Modell mit 58,5 tok/s Token-Generierung
Bester Prefill des Herstellers: 9.072,4 tok/s (Mistral-Small-3.1-24B-Instruct-2503)
Bester Prefill des Herstellers: 9.072,4 tok/s (Mistral-Small-3.1-24B-Instruct-2503)
Top-Modelle nach Token-Generierung
Bester gemessener Generierungs-Durchsatz je Modell (tok/s), veroeffentlichte Performancebenchmarks.
Modelle
Alle Benchmarks →10 zugeordnet
Mistral AI
Codestral-22B-v0.1
Dense · 22B
TG 33,5 tok/s
Prefill 2.687,8 tok/s
4 Benchmarks →
Mistral AI
Devstral-2-123B-Instruct-2512
Dense · 123B
TG 6,5 tok/s
Prefill 323,2 tok/s
1 Benchmarks →
Mistral AI
Devstral-Small-2-24B-Instruct-2512
Dense · 24B
TG 29,4 tok/s
Prefill 1.666,9 tok/s
8 Benchmarks →
Mistral AI
Devstral-Small-2507
Dense · 24B
TG 33,9 tok/s
Prefill 2.888,2 tok/s
6 Benchmarks →
Mistral AI
Magistral-Small-2509
Dense · 24B
TG 33,3 tok/s
Prefill 2.174,8 tok/s
6 Benchmarks →
Mistral AI
Mamba-Codestral-7B-v0.1
Dense · 7B
TG –
Prefill –
0 Benchmarks →
Mistral AI
Ministral-3-14B-Reasoning-2512
Dense · 14B
TG 51,0 tok/s
Prefill 3.889,5 tok/s
3 Benchmarks →
Mistral AI
Mistral-Medium-3.5-128B
Dense · 128B
TG 5,8 tok/s
Prefill 334,1 tok/s
1 Benchmarks →
Mistral AI
Mistral-Small-3.1-24B-Instruct-2503
Dense · 24B
TG 33,5 tok/s
Prefill 9.072,4 tok/s
9 Benchmarks →
Mistral AI
Mistral-Small-4-119B-2603
Dense · 119B
TG 58,5 tok/s
Prefill 2.052,5 tok/s
1 Benchmarks →
