DiggTop ⌕

DGX Spark x2 vs Mac Studio (Mac16,9) with Apple M4 Max

Best sourced output tok/s per model (single-stream generation) — engine and version, quantization, context and a source link on every row. Hover a value for the quant.

DGX Spark x218 sourced runs · best 474.0 tok/s
Mac Studio (Mac16,9) with Apple M4 Max1 verified run · best 110.3 tok/s
ModelDGX Spark x2Mac Studio (Mac16,9) with Apple M4 Max
Qwen3.8-Flash-Next-hibrid48474.0 NVFP4 (output head)—
DeepSeek-V4.1-Flash154.0 EXL3 2.9bpw—
Qwen3.8-Flash-Next93.0 INT4-AutoRound110.3 IQ3_XXS + MTP Q8_0
GLM-5.3-Flash46.3 NVFP4—

Head-to-head (models tested on both): DGX Spark x2 0 · Mac Studio (Mac16,9) with Apple M4 Max 1 · tied 0 · data as of Oct 11, 2026. Full rules: methodology.

Related reading: RTX 5090 vs 3090 · quantization · Value index · Full board

↑