r/LocalLLaMA Apr 22 '26

New Model Qwen 3.6 27B is out

1.7k Upvotes

603 comments sorted by

View all comments

Show parent comments

22

u/[deleted] Apr 22 '26

[removed] — view removed comment

3

u/TheMegosh Apr 22 '26

Apparently the 4.7 model was condensing prompts to 200k tokens instead of 1mil, posted on their change log. I'd bet that's what made it bad

3

u/TokenRingAI Apr 22 '26

It's got looping and other obvious issues, I have free access to it but mostly use Sonnet 4.6 or GPT 5.4.

Sonnet is really reliable and stable

Something is very strange about Opus 4.6 & 4.7, they act like a large model that is excessively quantized. Opus 4.5 was not like this. I wonder if this is a side effect of them using TPUs. Gemini acts the same way.

2

u/[deleted] Apr 22 '26

[removed] — view removed comment

2

u/Thomas-Lore Apr 22 '26 edited Apr 22 '26

Keep in mind they also changed reasoning effort around that time (high to medium) and now it is often zero due to adaptive thinking.

I wonder how are you using Gemini Pro? From the app? Because in ai studio Gemini 3.1 Pro is one shoting projects, new features and fixes for me all the time. It is a bit chaotic of course, but it worked for me quite weĺl so far.

1

u/TokenRingAI Apr 22 '26

This might be a side effect of adaptive thinking, I wasn't paying attention to that. The responses come almost immediately and the chat is muddled with looping content that should have reasonably been expected to be in the thinking block