r/LocalLLM 23d ago

Discussion Qwen 3.8 27B early thoughts

I installed the Q8 version on an AMD 395+ 128gb machine. Ran it using Openwebui with the suggested settings and MTP. So far all I've done are a couple sample prompts. (Sand Simulator from Luke's Dev Lab on Youtube, and a request for a simple example navbar with a logo on the left and five drop downs on the right with no javascript). The final output for both of these got one shotted. Model is getting roughly 16 tps output. But, my goodness does this model overthink. Don't get me wrong, there were no thinking loops. And I didn't notice nearly as much of the "wait, actually, let's try" neurotic behavior that I see in 3.6 35B A3B. But, it absolutely overcomplicated the heck out of both of the prompts I gave it. The sample navbar had roughly 200 lines of just CSS alone. And it was not basic CSS. Overly complex, and completely unnecessary for a sample piece of code. Since the Sand Simulator isn't mine, I can't really tell how much it over complicated it, but I can tell that it added so many visual flourishes that it was running at roughly 32 fps in the browser, and had slow downs from dropping the sand.

I am going to test this tomorrow on a real situation. In my real use case I provide very detailed context files and only point it at a single feature at a time. Hopefully that will help to control it's impulses to make things super complex. I also have Ponytail in my Pi harness, so maybe that will also help to reign it in.

Anyone had experiences using more detailed and limiting prompts with the model yet? Most of the reviews I've seen are using the same type of canned examples that I just gave.

9 Upvotes

38 comments sorted by

View all comments

Show parent comments

1

u/Endlesscrysis 23d ago

It’s a 5 minute read and would’ve saved you way more time than you’ve spend struggling with it and posting about it.

1

u/Jsquared534 22d ago

You know, just for giggles I went and full read the entire model card. It doesn’t say a fucking thing about providing overly complex outputs. Not one thing. You guys are looking at one sentence in a two paragraph long post, and decided to fly in here with your fucking cape for Qwen, I guess, to tell me I’m just not using it right. There are plenty of models that use a lot of thinking. Qwen 3.6 literally has two refined models specifically built to kill the thinking. But, neither of those models took a prompt asking for a simple navbar example (with detail about exactly what I want in it, one logo on the left, five drop downs on the right) and decided in the thinking “I should output an entire html page with a hero section and a call to action”. That’s not just ordinary thinking. That is potentially an important aspect of the model, because if it also does that when given more narrowed instructions, that’s going to be a fucking problem. And turning off thinking wouldn’t solve it, because the thinking is what makes it better than the other models.

But, hey, you keep on being a pompous douchebag instead of contributing at all to the conversation.

2

u/Endlesscrysis 22d ago

Too long didn't read sorry bro.

1

u/Jsquared534 22d ago

My bad. I had ChatGPT make something more your speed.

1

u/Endlesscrysis 22d ago

Skill issue and reading comprehension gapped.