DiggTop ⌕

Mac Studio (Mac16,9) with Apple M4 Max vs Tesla V100 x2

Best sourced output tok/s per model (single-stream generation) — engine and version, quantization, context and a source link on every row. Hover a value for the quant.

Sourced data on this pair is thin — most of the board is still partial. Submit your own run and help fill the matrix: one prompt, three steps, no downloads.
Mac Studio (Mac16,9) with Apple M4 Max1 sourced run · best 110.3 tok/s
Tesla V100 x25 verified runs · best 96.1 tok/s
ModelMac Studio (Mac16,9) with Apple M4 MaxTesla V100 x2
Qwen3.8-Flash-Next110.3 IQ3_XXS + MTP Q8_096.1 IQ3XXS
Ornith-1.5-35B-A3B—65.0 Q4_K_M
Qwen3.8-27B—24.0 Q8_K_XL

Head-to-head (models tested on both): Mac Studio (Mac16,9) with Apple M4 Max 1 · Tesla V100 x2 0 · tied 0 · data as of Oct 11, 2026. Full rules: methodology.

Related reading: RTX 5090 vs 3090 · quantization · Value index · Full board

↑