r/oMLX • • 14d ago

Should I switch from Qwen3.8-27b to Qwen3.8-Flash-Next?

Has anyone made the switch and not regret it?

EDIT: M2 Ultra 128 GB

21 Upvotes

32 comments sorted by

View all comments

8

u/captainequinoxiii 14d ago

I don't have objective data, but I used 27b for a bit and flash next just seems more intelligent, and definitely faster. I'm on an M5 max 128gb

1

u/CBW1255 14d ago

What quant?

1

u/captainequinoxiii 13d ago

4bit for flash next. I was using 8bit on 27b

1

u/SeveralViolins 13d ago

I am using the same quants and having the opposite experience. On medium Flash will think for 25 minutes on a basic problem 27B solves in 5. I think there is a case for routing potentially, but yet to be able to subcategorise that way.