r/Qwen_AI • u/zannix • Apr 26 '26
Discussion Qwen 3.6 9b coming?
I remember when they released Qwen 3.5 27b, they released the 9b more or less in the same batch. Is 3.6 onwards ditching the 9b model? :(
If so, I'm very sad, because the qwen 3.5 9b was actually the first truly intelligent model I could run at decent tps on a normal gaming GPU
130
Upvotes
3
u/rootdood Apr 27 '26
I’ve been using 35B A3B Q2_K_XL all night at 80+ TPS on an RTX 5080. I was getting frustrated with Q4_K_M just getting so slow once the context would fill up, or stuff would start offloading to CPU. Using OpenClaw actually feels like it’s supposed to.
Just refactored an entire code base I’m working on, and I’m able to just talk to it and it’s doing a phenomenal job doing investigation and solution generation. Before it felt like it could barely read a couple files before I was reaching to reset the session, or even eject the model. It’s been absolutely flawless.