23:50
75d ago
→Running Qwen 3.6 35B A3B on 2x 5060 Ti
A Reddit user ran Qwen 3.6 35B A3B in LM Studio with Q4 on two 16GB 5060 Ti GPUs, reported full-context throughput of 90 tokens per second, and asked for optimization paths to Q6 or Q8 plus cooling advice for two stacked GPUs with zero slot gap.
66
SCORE
H1·K1·R1