DiggTop ⌕

Mac Studio (Mac16,9) with Apple M4 Max vs RTX 3090 x2

Best sourced output tok/s per model (single-stream generation) — engine and version, quantization, context and a source link on every row. Hover a value for the quant.

Sourced data on this pair is thin — most of the board is still partial. Submit your own run and help fill the matrix: one prompt, three steps, no downloads.
Mac Studio (Mac16,9) with Apple M4 Max1 sourced run · best 110.3 tok/s
RTX 3090 x26 verified runs · best 240.8 tok/s
ModelMac Studio (Mac16,9) with Apple M4 MaxRTX 3090 x2
Qwen3.6-27B—240.8 AutoRound INT4
Qwen3.8-27B—219.8 AutoRound INT4 + fp8 KV
NVIDIA-Nemotron-3.5-Lightning-30B-A3B—197.0 NVFP4
Qwen3.8-Flash-Next110.3 IQ3_XXS + MTP Q8_040.1 Unsloth-Dynamic-IQ3_XXS
Qwen3.8-Flash-Next-REAP-256-duo—53.5 Unsloth-Dynamic-Q3_K_XL

Head-to-head (models tested on both): Mac Studio (Mac16,9) with Apple M4 Max 1 · RTX 3090 x2 0 · tied 0 · data as of Oct 11, 2026. Full rules: methodology.

Related reading: RTX 5090 vs 3090 · quantization · Value index · Full board

↑