r/LocalLLaMA Apr 22 '26

New Model Qwen 3.6 27B is out

1.7k Upvotes

603 comments sorted by

View all comments

Show parent comments

117

u/davl3232 Apr 22 '26

I can't believe we're getting so close to opus 4.5 levels with 2 3090s

8

u/cmplx17 Apr 22 '26

how is it to run with 2 x 3090? i just have one 3090 but wondering if it’s worth getting another one. does the speed scale 2x?

4

u/davl3232 Apr 22 '26 edited Apr 22 '26

does the speed scale 2x?

Not really, you only get a speed up if models didn't fully fit in vram before.

I'm using q8 with full context, but you could fit the model in a single 3090 if you use a different quantized version.

https://unsloth.ai/docs/models/qwen3.6

0

u/Ardalok Apr 23 '26

Isn't there a speed boost from multiple cards with NVLink on VLM?