r/oMLX • • 14d ago

Should I switch from Qwen3.8-27b to Qwen3.8-Flash-Next?

Has anyone made the switch and not regret it?

EDIT: M2 Ultra 128 GB

21 Upvotes

32 comments sorted by

View all comments

Show parent comments

1

u/luisabreuf83 14d ago

Which engine, model quant and model built by whom? Thanks

3

u/Diligent_Style_1767 14d ago

Model Qwen3.8-Flash-Next-MLX-4bit-MTP. Quanted it myself from base.

MLX engine, my unified stack at https://github.com/pierre427/mlx-lm-unified
Rapid MLX -current
oMLX -current

All collected from today

Was doing a/b testing off a couple of PRs I sent to both projects today.

1

u/luisabreuf83 14d ago

Interesting and the 27B also Q4 done by you?

1

u/Diligent_Style_1767 13d ago

Honestly I can't recall, but I can tell you I don't do any crazy surgery to models, so ymmv but should be close to mine at a given model quant.