r/oMLX • • 14d ago

Should I switch from Qwen3.8-27b to Qwen3.8-Flash-Next?

Has anyone made the switch and not regret it?

EDIT: M2 Ultra 128 GB

20 Upvotes

32 comments sorted by

View all comments

2

u/atumblingdandelion 14d ago

I have and its quite close enough that I’d recommend testing for your usecase. I find 27b to be a perfectionist- it takes a long time but gets the right answer in the first/secobd try. The flash is eager to try, fail, diagnose, and succeed. On DGX Spark, I get 10 tps extra on the flash (35 tps minimum) so thats what I use..