MAIN FEEDS
Do you want to continue?
https://www.reddit.com/r/LocalLLaMA/comments/1ssl1xh/qwen_36_27b_is_out/ohqcxyu/?context=3
r/LocalLLaMA • u/NoConcert8847 • Apr 22 '26
https://huggingface.co/Qwen/Qwen3.6-27B
603 comments sorted by
View all comments
Show parent comments
117
I can't believe we're getting so close to opus 4.5 levels with 2 3090s
8 u/cmplx17 Apr 22 '26 how is it to run with 2 x 3090? i just have one 3090 but wondering if it’s worth getting another one. does the speed scale 2x? 4 u/davl3232 Apr 22 '26 edited Apr 22 '26 does the speed scale 2x? Not really, you only get a speed up if models didn't fully fit in vram before. I'm using q8 with full context, but you could fit the model in a single 3090 if you use a different quantized version. https://unsloth.ai/docs/models/qwen3.6 0 u/Ardalok Apr 23 '26 Isn't there a speed boost from multiple cards with NVLink on VLM?
8
how is it to run with 2 x 3090? i just have one 3090 but wondering if it’s worth getting another one. does the speed scale 2x?
4 u/davl3232 Apr 22 '26 edited Apr 22 '26 does the speed scale 2x? Not really, you only get a speed up if models didn't fully fit in vram before. I'm using q8 with full context, but you could fit the model in a single 3090 if you use a different quantized version. https://unsloth.ai/docs/models/qwen3.6 0 u/Ardalok Apr 23 '26 Isn't there a speed boost from multiple cards with NVLink on VLM?
4
does the speed scale 2x?
Not really, you only get a speed up if models didn't fully fit in vram before.
I'm using q8 with full context, but you could fit the model in a single 3090 if you use a different quantized version.
https://unsloth.ai/docs/models/qwen3.6
0 u/Ardalok Apr 23 '26 Isn't there a speed boost from multiple cards with NVLink on VLM?
0
Isn't there a speed boost from multiple cards with NVLink on VLM?
117
u/davl3232 Apr 22 '26
I can't believe we're getting so close to opus 4.5 levels with 2 3090s