M
Hersteller
Mistral AI
Franzoesisches KI-Labor; Mistral, Mixtral, Codestral, Devstral, Ministral, Magistral.
Beste Token-Generierung
1.060,0 tok/s
Ministral-3-14B-Reasoning-2512
Bester Prefill
20.061,1 tok/s
Ministral-3-14B-Reasoning-2512
∅ Token-Generierung
140,5 tok/s
Schnitt aus 636 Läufen
∅ Prefill
3.123,1 tok/s
Schnitt aus 636 Läufen
Bester TTFT
20 ms
Magistral-Small-2509
Beste Harness-Quote
87%
Devstral-Small-2-24B-Instruct-2512
🏆
Ministral-3-14B-Reasoning-2512 · schnellstes Modell mit 1.060,0 tok/s Token-Generierung
Bester Prefill des Herstellers: 20.061,1 tok/s (Ministral-3-14B-Reasoning-2512)
Bester Prefill des Herstellers: 20.061,1 tok/s (Ministral-3-14B-Reasoning-2512)
Top-Modelle nach Token-Generierung
Bester gemessener Generierungs-Durchsatz je Modell (tok/s), veröffentlichte Performancebenchmarks.
Modelle
Alle Benchmarks →10 zugeordnet
Codestral-22B-v0.122B89 Läufe
129,0tok/s Ø
Min 13,0Max 641,4
Devstral-2-123B-Instruct-2512123B50 Läufe
22,1tok/s Ø
Min 0,0Max 143,1
Devstral-Small-2-24B-Instruct-251224B94 Läufe
156,1tok/s Ø
Min 8,9Max 725,6
Devstral-Small-250724B86 Läufe
150,3tok/s Ø
Min 13,3Max 647,6
Magistral-Small-250924B92 Läufe
172,9tok/s Ø
Min 4,0Max 737,0
Mamba-Codestral-7B-v0.17B0 Läufe
Noch keine Performancedaten
Ministral-3-14B-Reasoning-251214B106 Läufe
253,1tok/s Ø
Min 22,4Max 1.060,0
Mistral-Medium-3.5-128B128B28 Läufe
10,2tok/s Ø
Min 0,6Max 141,6
Mistral-Small-3.1-24B-Instruct-250324B69 Läufe
63,2tok/s Ø
Min 5,0Max 214,9
Mistral-Small-4-119B-2603119B26 Läufe
93,1tok/s Ø
Min 5,8Max 959,6
