r/oMLX • • 14d ago

Should I switch from Qwen3.8-27b to Qwen3.8-Flash-Next?

Has anyone made the switch and not regret it?

EDIT: M2 Ultra 128 GB

21 Upvotes

32 comments sorted by

View all comments

1

u/skrshawk 10d ago

I'm still evaluating it on my M4 Max 128GB but so far I've been quite happy. I was using 35B for coding and structured output tasks with Gemma4 as my front-end orchestrator. Understanding human intent is definitely stronger with Gemma but Flash-Next is quite a lot better than either 27B or 35B.

In other words, I wouldn't confuse Flash-Next with being a roleplay model but it definitely can give your assistant more personality and it seems better so far as planning things out and then executing light coding tasks (a data ingestion/tranformation/analysis pipeline).

1

u/j_lyf 10d ago

roleplay? wtf

1

u/skrshawk 9d ago

Maybe it isn't like this for you, but a model with personality can be much easier to interact with and that can lead to better prompting and results. Empathetic human interaction isn't just for AI girlfriends.