Created bymario-alka.dePowered bygodcore.denoob2claw.detricoma.de
Manufacturer

Mistral AI

Franzoesisches KI-Labor; Mistral, Mixtral, Codestral, Devstral, Ministral, Magistral.

10Models
640Benchmarks
636Performance
4Harness
Best token generation
1.060,0 tok/s
Ministral-3-14B-Reasoning-2512
Best prefill
20.061,1 tok/s
Ministral-3-14B-Reasoning-2512
∅ Token generation
140,5 tok/s
Average of 636 runs
∅ Prefill
3.123,1 tok/s
Average of 636 runs
Best TTFT
20 ms
Magistral-Small-2509
Best harness rate
87%
Devstral-Small-2-24B-Instruct-2512
🏆
Ministral-3-14B-Reasoning-2512 · fastest model at 1.060,0 tok/s token generation
Best prefill of this manufacturer: 20.061,1 tok/s (Ministral-3-14B-Reasoning-2512)

Top models by token generation

Best measured generation throughput per model (tok/s), published performance benchmarks.

Ministral-3-14B-Reasoning-2512
1.060,0 tok/s
Mistral-Small-4-119B-2603
959,6 tok/s
Magistral-Small-2509
737,0 tok/s
Devstral-Small-2-24B-Instruct-2512
725,6 tok/s
Devstral-Small-2507
647,6 tok/s
Codestral-22B-v0.1
641,4 tok/s
Mistral-Small-3.1-24B-Instruct-2503
214,9 tok/s
Devstral-2-123B-Instruct-2512
143,1 tok/s