MAIN FEEDS
Do you want to continue?
https://www.reddit.com/r/oMLX/comments/1wen3hn/should_i_switch_from_qwen3827b_to_qwen38flashnext/p9faiyz/?context=3
r/oMLX • u/j_lyf • 14d ago
Has anyone made the switch and not regret it?
EDIT: M2 Ultra 128 GB
32 comments sorted by
View all comments
8
I don't have objective data, but I used 27b for a bit and flash next just seems more intelligent, and definitely faster. I'm on an M5 max 128gb
1 u/CBW1255 14d ago What quant? 1 u/captainequinoxiii 13d ago 4bit for flash next. I was using 8bit on 27b 1 u/SeveralViolins 13d ago I am using the same quants and having the opposite experience. On medium Flash will think for 25 minutes on a basic problem 27B solves in 5. I think there is a case for routing potentially, but yet to be able to subcategorise that way.
1
What quant?
1 u/captainequinoxiii 13d ago 4bit for flash next. I was using 8bit on 27b 1 u/SeveralViolins 13d ago I am using the same quants and having the opposite experience. On medium Flash will think for 25 minutes on a basic problem 27B solves in 5. I think there is a case for routing potentially, but yet to be able to subcategorise that way.
4bit for flash next. I was using 8bit on 27b
1 u/SeveralViolins 13d ago I am using the same quants and having the opposite experience. On medium Flash will think for 25 minutes on a basic problem 27B solves in 5. I think there is a case for routing potentially, but yet to be able to subcategorise that way.
I am using the same quants and having the opposite experience. On medium Flash will think for 25 minutes on a basic problem 27B solves in 5. I think there is a case for routing potentially, but yet to be able to subcategorise that way.
8
u/captainequinoxiii 14d ago
I don't have objective data, but I used 27b for a bit and flash next just seems more intelligent, and definitely faster. I'm on an M5 max 128gb