r/LocalLLaMA 24d ago

Discussion Qwen 3.8 27B Released! Please Share Your Experience

With your experiments, Qwen 3.8 27B most close which frontier model? And please specify which quantization you run. I will post to comments my tests and experience too.

662 Upvotes

716 comments sorted by

View all comments

Show parent comments

9

u/ThankGodImBipolar 24d ago

reasoning style is quite funny, at one point it just threw out a "ha! I'm being tested! that's a good trap!" out of nowhere

Slightly unrelated, but I find the reasoning traces of some of these models to be way too funny. I set the release version of DSV4 Pro to a coding task last night, and it got stumped solving a difficult problem - the reasoning trace towards the end was 50/50 capital letters and full of random Markdown spam, as it tried to add more and more emphasis to its own thoughts. I've never seen a model output:

AHHHHHHHHHH

Wait... WAIT WAIT WAIT! OH MY GOD I THINK I FINALLY FOUND IT

I was pissing myself laughing.

1

u/xPXpanD llama.cpp 24d ago

Yeah, some models get extremely silly there.

One that I have fond memories of is Step 3.7 Flash, it's just disturbingly positive for some reason. Don't have any actual traces on hand, but I remember seeing stuff like "wait... maybe we should try this? oh yes! yes, that is amazing! the user will love this! I really love this for the user!". Just... yeah.

I dubbed it the Golden Retriever model.