CMP 170HX x4 vs RTX 3090
Best sourced output tok/s per model (single-stream generation) — engine and version, quantization, context and a source link on every row. Hover a value for the quant.
CMP 170HX x42 sourced runs · best 264.6 tok/s
RTX 309018 verified runs · best 709.5 tok/s
| Model | CMP 170HX x4 | RTX 3090 |
|---|---|---|
| LFM2.5-1.2B-Instruct | — | 709.5 |
| Qwen3.5-0.8B-MTP | — | 476.1 |
| NVIDIA-Nemotron-3.5-Lightning-30B-A3B | — | 353.7 |
| LFM2.5-2.6B | — | 305.6 |
| GLM-5.3-Flash | 264.6 | 28.1 |
| Qwen3.5-4B-MTP | — | 240.4 |
| Qwen3.8-27B | — | 164.6 |
| Qwen3.5-9B-MTP | — | 164.3 |
| DeepSeek-V4-Flash-0731 | 92.1 | — |
| Qwen3.8-Flash-Next | — | 38.6 |
| Ornstein3.6-27B-MTP-NSC-ACE-SABER | — | 32.9 |
| DeepSeek-V4.1-Flash | — | 26.0 |