CMP 170HX vs Multi-GPU x3
Best verified output tok/s per model (single-stream generation). Hover a value for the quant.
CMP 170HX16 verified runs · best 255.9 tok/s
Multi-GPU x34 verified runs · best 63.3 tok/s
| Model | CMP 170HX | Multi-GPU x3 |
|---|---|---|
| Qwen3.8-27B | 255.9 | — |
| NVIDIA-Nemotron-3.5-Lightning-30B-A3B | 165.1 | — |
| Qwen3.6-35B-A3B | 87.2 | 63.3 |
| Ornith-1.5-35B-A3B | 87.1 | — |
| Ornith-1.5-9B | 61.5 | — |
| Qwen3.6-27B | — | 60.3 |
| Qwen3.8-27B-MTP | 46.4 | — |
| Qwen3.8-Flash-Next | 46.1 | — |
| Llama-4-Scout-17B-16E-Instruct | 44.5 | — |
| Ling-3.0-flash | — | 43.4 |
| gemma-4-31B-it | 28.3 | — |
| Laguna-S-2.1 | — | 23.9 |
| granite-4.2-30b | 23.7 | — |
| Llama-3.3-70B-Instruct | 20.1 | — |