Menu

#Iq4_xs

1 post

Feed
1 of 1 post
Running a 35B MoE (Qwen3.6-35B-A3B) on 2x GTX 1080 Ti in 2026 — Real Benchmarks, and Does the Second GPU Actually Help?
🖼️
0

Running a 35B MoE (Qwen3.6-35B-A3B) on 2x GTX 1080 Ti in 2026 — Real Benchmarks, and Does the Second GPU Actually Help?

DEV Community: machinelearning·byeongsoo kang·4 months ago
#Pp28NVQa
#dev#model#iq4_xs#ollama#quant#experts

I benchmarked Qwen3.6-35B-A3B (IQ4_XS) on two 8-year-old GTX 1080 Ti cards: ~20 tok/s, and the second GPU only adds ~20% (not 2x). Real numbers, VRAM math, and why a 35B fits 22GB.

15s
Read More