DiggTop ⌕

CMP 170HX x2 vs Intel Arc Pro B70

Best sourced output tok/s per model (single-stream generation) — engine and version, quantization, context and a source link on every row. Hover a value for the quant.

CMP 170HX x25 sourced runs · best 186.9 tok/s
Intel Arc Pro B7018 verified runs · best 283.1 tok/s
ModelCMP 170HX x2Intel Arc Pro B70
Laguna-XS-2.1—283.1 Q5_K_M GGUF
Qwen3.6-35B-A3B—248.3 GPTQ-4bit
Ornith-1.0-35B-B70-Turbo—224.9 Q5_K_M GGUF
Nex-N2-mini-B70-Turbo—224.3 Q5_K_M GGUF
SIQ-1-35B-B70-Turbo—223.7 Q5_K_M GGUF
Qwen-AgentWorld-35B-A3B-B70-Turbo—209.0 Q5_K_M GGUF
Qwen3.8-Flash-Next186.9 W4A16-FP8PLE24.0 EXL3 3.05 bpw
NVIDIA-Nemotron-3.5-Lightning-30B-A3B—186.6 GPTQ-INT4-G64-sym-local+DFlash-B
LFM2.5-2.6B—132.4 Q8_0
Ornith-1.5-35B-A3B—131.5 Q4_K_M
Laguna-S-2.163.3 Q4_K_M—
Hy-MT2-30B-A3B61.4 FP8—
Gemma4-26B-A4B-QAT-Uncensored-HauhauCS-Balanced-MTP—58.4 Q4_K_M
Ornith-1.5-9B—49.6 Q8_0
Qwen3.6-27B-MTP—36.0 Q8_0
Muse-Glimmer-30B—29.2 Q4_K_XL
ThinkingCap-Qwen3.6-27B—27.9 Q6_K
Qwen3.8-27B—27.8 Q4_K_M
Hy327.6 Q2_K_XL—
Baichuan-M3-235B26.1 Q3_K_M—

Head-to-head (models tested on both): CMP 170HX x2 1 · Intel Arc Pro B70 0 · tied 0 · data as of Oct 6, 2026. Full rules: methodology.

Related reading: RTX 5090 vs 3090 · quantization · Value index · Full board

↑