Created bymario-alka.dePowered bygodcore.denoob2claw.detricoma.de
🧮
Mainboard

ASUSTeK COMPUTER INC. ROG STRIX B550-A GAMING

ASUS1 CPU218 Benchmarks⚡ 12 W Idle · 45 W Max
CPUs
1
Models
50
Benchmarks
218
Best generation
2.491,2 tok/s
🏆
gemma-4-E2B-it läuft am schnellsten auf diesem Board · 2.491,2 tok/s
50 Modelle mit Performance-Daten getestet

Top models by token generation

Bester gemessener Generierungs-Durchsatz je Modell auf ASUSTeK COMPUTER INC. ROG STRIX B550-A GAMING (tok/s) – klicken für Leaderboard-Filter.

gemma-4-E2B-it5B
2.491,2 tok/s
gpt-oss-20b20B
1.890,7 tok/s
Nemotron-3-Nano-4B4B
1.878,7 tok/s
Laguna-XS-2.1
1.765,0 tok/s
Nemotron-3-Nano-30B-A3B30B
1.722,1 tok/s
Anzeige
Model comparison

Tested models compared

Jede Blase ist ein Modell auf ASUSTeK COMPUTER INC. ROG STRIX B550-A GAMING – Position: Prompt-Verarbeitung (X) × Ausgabe-Geschwindigkeit (Y), Blasengröße: Anzahl Messläufe. Naeher an der oberen rechten Ecke = schneller.

MODnach Modell

3.2392.4291.6198100,006.57913.15819.738Prefill (tok/s)Generation (tok/s)gemma-4-E2B-it - 2.491,2 tok/s Generation, 11.148 tok/s Prefill, TTFT 2.175 ms (3 Laufe)gemma-4-E2B-itgpt-oss-20b - 1.890,7 tok/s Generation, 11.034 tok/s Prefill, TTFT 4.372 ms (18 Laufe)gpt-oss-20bNemotron-3-Nano-4B - 1.878,7 tok/s Generation, 5.348 tok/s Prefill, TTFT 3.511 ms (3 Laufe)Nemotron-3-Nano-4BLaguna-XS-2.1 - 1.765,0 tok/s Generation, 7.330 tok/s Prefill, TTFT 3.704 ms (9 Laufe)Laguna-XS-2.1Nemotron-3-Nano-30B-A3B - 1.722,1 tok/s Generation, 5.965 tok/s Prefill, TTFT 3.591 ms (3 Laufe)Nemotron-3-Nano-30B-A...Nemotron-Cascade-2-30B-A3B - 1.719,2 tok/s Generation, 6.140 tok/s Prefill, TTFT 3.569 ms (3 Laufe)Nemotron-Cascade-2-30...Qwen3-Coder-30B-A3B-Instruct - 1.716,6 tok/s Generation, 15.919 tok/s Prefill, TTFT 2.829 ms (4 Laufe)Qwen3-Coder-30B-A3B-I...Nemotron-3-Nano-Omni-30B-A3B-Reasoning - 1.706,7 tok/s Generation, 5.892 tok/s Prefill, TTFT 3.614 ms (3 Laufe)Nemotron-3-Nano-Omni-...Qwen3-30B-A3B-Thinking-2507 - 1.683,8 tok/s Generation, 14.349 tok/s Prefill, TTFT 2.214 ms (6 Laufe)Qwen3-30B-A3B-Thinkin...Qwen3-Omni-30B-A3B-Thinking - 1.673,6 tok/s Generation, 8.080 tok/s Prefill, TTFT 3.494 ms (3 Laufe)Qwen3-Omni-30B-A3B-Th...Qwen3-30B-A3B-Instruct-2507 - 1.671,0 tok/s Generation, 15.160 tok/s Prefill, TTFT 2.212 ms (6 Laufe)Qwen3-30B-A3B-Instruc...Qwen3-VL-30B-A3B-Instruct - 1.655,9 tok/s Generation, 9.318 tok/s Prefill, TTFT 3.329 ms (3 Laufe)Qwen3-VL-30B-A3B-Inst...gemma-4-E4B-it - 1.606,1 tok/s Generation, 8.848 tok/s Prefill, TTFT 2.968 ms (3 Laufe)gemma-4-E4B-itMeta-Llama-3.1-8B-Instruct - 1.593,1 tok/s Generation, 11.858 tok/s Prefill, TTFT 3.236 ms (3 Laufe)Meta-Llama-3.1-8B-Ins...North-Mini-Code-1.0 - 1.557,3 tok/s Generation, 5.892 tok/s Prefill, TTFT 3.954 ms (3 Laufe)North-Mini-Code-1.0Ornith-1.0-35B - 1.556,7 tok/s Generation, 5.085 tok/s Prefill, TTFT 3.918 ms (3 Laufe)Ornith-1.0-35BGemma-4-26B-A4B-it - 1.540,5 tok/s Generation, 4.006 tok/s Prefill, TTFT 4.749 ms (3 Laufe)Gemma-4-26B-A4B-itOrnith-1.0-9B - 1.505,6 tok/s Generation, 6.762 tok/s Prefill, TTFT 3.486 ms (3 Laufe)Ornith-1.0-9BQwen-AgentWorld-35B-A3B - 1.439,1 tok/s Generation, 5.167 tok/s Prefill, TTFT 4.136 ms (3 Laufe)Qwen-AgentWorld-35B-A...Qwen3.5-35B-A3B - 1.416,7 tok/s Generation, 5.541 tok/s Prefill, TTFT 3.907 ms (3 Laufe)Qwen3.5-35B-A3BQwen3.6-35B-A3B - 1.409,4 tok/s Generation, 4.912 tok/s Prefill, TTFT 4.113 ms (3 Laufe)Qwen3.6-35B-A3Bgemma-4-12B-it - 1.058,9 tok/s Generation, 3.140 tok/s Prefill, TTFT 6.383 ms (3 Laufe)gemma-4-12B-itMinistral-3-14B-Reasoning-2512 - 1.028,1 tok/s Generation, 7.724 tok/s Prefill, TTFT 4.732 ms (3 Laufe)Ministral-3-14B-Reaso...Seed-OSS-36B-Instruct - 839,7 tok/s Generation, 2.664 tok/s Prefill, TTFT 12.384 ms (3 Laufe)Seed-OSS-36B-InstructMagistral-Small-2509 - 737,0 tok/s Generation, 128 tok/s Prefill, TTFT 3.689 ms (3 Laufe)Magistral-Small-2509Codestral-22B-v0.1 - 641,4 tok/s Generation, 5.824 tok/s Prefill, TTFT 7.096 ms (3 Laufe)Codestral-22B-v0.1Devstral-Small-2507 - 625,1 tok/s Generation, 7.381 tok/s Prefill, TTFT 5.802 ms (3 Laufe)Devstral-Small-2507Devstral-Small-2-24B-Instruct-2512 - 620,3 tok/s Generation, 6.308 tok/s Prefill, TTFT 6.898 ms (3 Laufe)Devstral-Small-2-24B-...Qwen3.6-27B - 532,5 tok/s Generation, 2.539 tok/s Prefill, TTFT 9.812 ms (3 Laufe)Qwen3.6-27BOpenReasoning-Nemotron-32B - 481,0 tok/s Generation, 3.983 tok/s Prefill, TTFT 10.140 ms (3 Laufe)OpenReasoning-Nemotro...Qwen2.5-Coder-32B-Instruct - 479,6 tok/s Generation, 4.242 tok/s Prefill, TTFT 10.075 ms (3 Laufe)Qwen2.5-Coder-32B-Ins...Qwen2.5-32B-Instruct-AWQ - 473,9 tok/s Generation, 4.532 tok/s Prefill, TTFT 9.919 ms (3 Laufe)Qwen2.5-32B-Instruct-...Nemotron-H-47B-Reasoning-128K - 197,2 tok/s Generation, 547 tok/s Prefill, TTFT 32.881 ms (3 Laufe)Nemotron-H-47B-Reason...Qwen3-Coder-Next - 107,0 tok/s Generation, 398 tok/s Prefill, TTFT 60.296 ms (3 Laufe)Qwen3-Coder-Nextgpt-oss-120b - 42,3 tok/s Generation, 239 tok/s Prefill, TTFT 76.971 ms (5 Laufe)gpt-oss-120bMistral-Small-4-119B-2603 - 21,7 tok/s Generation, 198 tok/s Prefill, TTFT 150.089 ms (6 Laufe)Mistral-Small-4-119B-...Llama-3.3-Nemotron-Super-49B-v1.5 - 20,2 tok/s Generation, 333 tok/s Prefill, TTFT 137.310 ms (3 Laufe)Llama-3.3-Nemotron-Su...Laguna-S-2.1-INT4 - 13,3 tok/s Generation, 95 tok/s Prefill, TTFT 143.534 ms (3 Laufe)Laguna-S-2.1-INT4Qwen3.5-122B-A10B - 8,3 tok/s Generation, 72 tok/s Prefill, TTFT 99.190 ms (8 Laufe)Qwen3.5-122B-A10BGLM-4.5-Air - 6,2 tok/s Generation, 161 tok/s Prefill, TTFT 57.809 ms (5 Laufe)GLM-4.5-AirMiniMax-M2.7 - 6,0 tok/s Generation, 66 tok/s Prefill, TTFT 126.812 ms (5 Laufe)MiniMax-M2.7Nemotron-3-Super-120B-A12B - 5,5 tok/s Generation, 49 tok/s Prefill, TTFT 150.808 ms (8 Laufe)Nemotron-3-Super-120B...MiniMax-M2.5 - 4,4 tok/s Generation, 81 tok/s Prefill, TTFT 88.412 ms (6 Laufe)MiniMax-M2.5Laguna-M.1 - 4,2 tok/s Generation, 37 tok/s Prefill, TTFT 169.257 ms (3 Laufe)Laguna-M.1DeepSeek-V4-Flash-284B-A13B - 3,2 tok/s Generation, 27 tok/s Prefill, TTFT 240.091 ms (3 Laufe)DeepSeek-V4-Flash-284...Qwen2.5-72B-Instruct - 3,0 tok/s Generation, 147 tok/s Prefill, TTFT 42.799 ms (3 Laufe)Qwen2.5-72B-Instructcommand-a-reasoning-08-2025 - 2,2 tok/s Generation, 60 tok/s Prefill, TTFT 108.625 ms (3 Laufe)command-a-reasoning-0...Devstral-2-123B-Instruct-2512 - 2,1 tok/s Generation, 113 tok/s Prefill, TTFT 62.542 ms (3 Laufe)Devstral-2-123B-Instr...Mistral-Medium-3.5-128B - 1,9 tok/s Generation, 113 tok/s Prefill, TTFT 72.453 ms (5 Laufe)Mistral-Medium-3.5-12...Llama-3.1-Nemotron-Ultra-253B-v1 - 0,1 tok/s Generation, 7 tok/s Prefill, TTFT 290.783 ms (19 Laufe)Llama-3.1-Nemotron-Ul...
gemma-4-E2B-it 2.491,2 tok/sgpt-oss-20b 1.890,7 tok/sNemotron-3-Nano-4B 1.878,7 tok/sLaguna-XS-2.1 1.765,0 tok/sNemotron-3-Nano-30B-A3B 1.722,1 tok/sNemotron-Cascade-2-30B-A3B 1.719,2 tok/sQwen3-Coder-30B-A3B-Instruct 1.716,6 tok/sNemotron-3-Nano-Omni-30B-A3B-Reasoning 1.706,7 tok/sQwen3-30B-A3B-Thinking-2507 1.683,8 tok/sQwen3-Omni-30B-A3B-Thinking 1.673,6 tok/sQwen3-30B-A3B-Instruct-2507 1.671,0 tok/sQwen3-VL-30B-A3B-Instruct 1.655,9 tok/sgemma-4-E4B-it 1.606,1 tok/sMeta-Llama-3.1-8B-Instruct 1.593,1 tok/sNorth-Mini-Code-1.0 1.557,3 tok/sOrnith-1.0-35B 1.556,7 tok/sGemma-4-26B-A4B-it 1.540,5 tok/sOrnith-1.0-9B 1.505,6 tok/sQwen-AgentWorld-35B-A3B 1.439,1 tok/sQwen3.5-35B-A3B 1.416,7 tok/sQwen3.6-35B-A3B 1.409,4 tok/sgemma-4-12B-it 1.058,9 tok/sMinistral-3-14B-Reasoning-2512 1.028,1 tok/sSeed-OSS-36B-Instruct 839,7 tok/sMagistral-Small-2509 737,0 tok/sCodestral-22B-v0.1 641,4 tok/sDevstral-Small-2507 625,1 tok/sDevstral-Small-2-24B-Instruct-2512 620,3 tok/sQwen3.6-27B 532,5 tok/sOpenReasoning-Nemotron-32B 481,0 tok/sQwen2.5-Coder-32B-Instruct 479,6 tok/sQwen2.5-32B-Instruct-AWQ 473,9 tok/sNemotron-H-47B-Reasoning-128K 197,2 tok/sQwen3-Coder-Next 107,0 tok/sgpt-oss-120b 42,3 tok/sMistral-Small-4-119B-2603 21,7 tok/sLlama-3.3-Nemotron-Super-49B-v1.5 20,2 tok/sLaguna-S-2.1-INT4 13,3 tok/sQwen3.5-122B-A10B 8,3 tok/sGLM-4.5-Air 6,2 tok/sMiniMax-M2.7 6,0 tok/sNemotron-3-Super-120B-A12B 5,5 tok/sMiniMax-M2.5 4,4 tok/sLaguna-M.1 4,2 tok/sDeepSeek-V4-Flash-284B-A13B 3,2 tok/sQwen2.5-72B-Instruct 3,0 tok/scommand-a-reasoning-08-2025 2,2 tok/sDevstral-2-123B-Instruct-2512 2,1 tok/sMistral-Medium-3.5-128B 1,9 tok/sLlama-3.1-Nemotron-Ultra-253B-v1 0,1 tok/s

Tested models

Mgemma-4-E2B-it5BGoogle · 3 Läufe Mgpt-oss-20b20BOpenAI · 18 Läufe MNemotron-3-Nano-4B4BNVIDIA · 3 Läufe MLaguna-XS-2.1Poolside · 9 Läufe MNemotron-3-Nano-30B-A3B30BNVIDIA · 3 Läufe MNemotron-Cascade-2-30B-A3B30BNVIDIA · 3 Läufe MQwen3-Coder-30B-A3B-Instruct30BQwen (Alibaba) · 4 Läufe MNemotron-3-Nano-Omni-30B-A3B-Reasoning30BNVIDIA · 3 Läufe MQwen3-30B-A3B-Thinking-250730BQwen (Alibaba) · 6 Läufe MQwen3-Omni-30B-A3B-Thinking30BQwen (Alibaba) · 3 Läufe MQwen3-30B-A3B-Instruct-250730BQwen (Alibaba) · 6 Läufe MQwen3-VL-30B-A3B-Instruct30BQwen (Alibaba) · 3 Läufe Mgemma-4-E4B-it8BGoogle · 3 Läufe MMeta-Llama-3.1-8B-Instruct8BMeta · 3 Läufe MNorth-Mini-Code-1.0Cohere · 3 Läufe MOrnith-1.0-35B35BDeepReinforce · 3 Läufe MGemma-4-26B-A4B-it26BGoogle · 3 Läufe MOrnith-1.0-9B9BDeepReinforce · 3 Läufe MQwen-AgentWorld-35B-A3B35BQwen (Alibaba) · 3 Läufe MQwen3.5-35B-A3B35BQwen (Alibaba) · 3 Läufe MQwen3.6-35B-A3B35BQwen (Alibaba) · 3 Läufe Mgemma-4-12B-it12BGoogle · 3 Läufe MMinistral-3-14B-Reasoning-251214BMistral AI · 3 Läufe MSeed-OSS-36B-Instruct36BByteDance Seed · 3 Läufe MMagistral-Small-250924BMistral AI · 3 Läufe MCodestral-22B-v0.122BMistral AI · 3 Läufe MDevstral-Small-250724BMistral AI · 3 Läufe MDevstral-Small-2-24B-Instruct-251224BMistral AI · 3 Läufe MQwen3.6-27B27BQwen (Alibaba) · 3 Läufe MOpenReasoning-Nemotron-32B32BNVIDIA · 3 Läufe MQwen2.5-Coder-32B-Instruct32BQwen (Alibaba) · 3 Läufe MQwen2.5-32B-Instruct-AWQ32BQwen (Alibaba) · 3 Läufe MNemotron-H-47B-Reasoning-128K47BNVIDIA · 3 Läufe MQwen3-Coder-NextQwen (Alibaba) · 3 Läufe Mgpt-oss-120b120BOpenAI · 5 Läufe MMistral-Small-4-119B-2603119BMistral AI · 6 Läufe MLlama-3.3-Nemotron-Super-49B-v1.549BNVIDIA · 3 Läufe MLaguna-S-2.1-INT4Poolside · 3 Läufe MQwen3.5-122B-A10B122BQwen (Alibaba) · 8 Läufe MGLM-4.5-Air106BZ.ai (Zhipu) · 5 Läufe MMiniMax-M2.7230BMiniMax · 5 Läufe MNemotron-3-Super-120B-A12B120BNVIDIA · 8 Läufe MMiniMax-M2.5230BMiniMax · 6 Läufe MLaguna-M.1Poolside · 3 Läufe MDeepSeek-V4-Flash-284B-A13B284BDeepSeek · 3 Läufe MQwen2.5-72B-Instruct72BQwen (Alibaba) · 3 Läufe Mcommand-a-reasoning-08-2025111BCohere · 3 Läufe MDevstral-2-123B-Instruct-2512123BMistral AI · 3 Läufe MMistral-Medium-3.5-128B128BMistral AI · 5 Läufe MLlama-3.1-Nemotron-Ultra-253B-v1253BNVIDIA · 19 Läufe

← All hardware