DiggTop ⌕

M4 Max 128GB vs Radeon AI Pro R9700

Best verified output tok/s per model (single-stream generation). Hover a value for the quant.

M4 Max 128GB1 verified run · best 134.0 tok/s
Radeon AI Pro R970013 verified runs · best 969.4 tok/s
ModelM4 Max 128GBRadeon AI Pro R9700
Qwen3.6-35B-A3B—969.4 MQ4R
Qwen3.5-0.8B—554.5 MQ4
LFM2.5-230M—520.0 MQ4
LFM2.5-350M—466.0 MQ4
Qwen3.5-27B—286.6 MQ4-AWQ
LFM2.5-8B-A1B—241.1 MQ4
gemma-4-E2B-it—169.7 Q8_0
gemma-4-E4B-it—137.0 Q4_K_M
Muse-Glimmer-30B—134.6 Dynamic-Q4_K_XL + DFlash2-Q4_K_M
Qwen3.8-27B134.0 4-bit64.3 Q4_0
DeepSeek-V4-Flash-0731—59.4 IQ2_XXS
Qwen3.8-Flash-Next—25.6 IQ4_XS
GLM-5.3-Flash—17.4 UD-IQ1_M

Head-to-head (models tested on both): M4 Max 128GB 1 · Radeon AI Pro R9700 0 · tied 0 · data as of Sep 29, 2026. Full rules: methodology.

Related reading: RTX 5090 vs 3090 · quantization · Value index · Full board

↑