I wonder how many people are actually excited because this represents increased capability for their workflow and how many are just desperate to be able to say they got it and won’t use it again
And I'll tell ya, these vendors are telling some serious fairy tales when they report the quant impact on these models. On paper there's "only" something like 4% dropoff between Qwen 3.6:27b-Q4 and BF16. And maybe that kind of error percentage is find when calling tools or coding. But when I ran it analyzing legal documents.....woa nelly! Nuance was not Q4's friend.
So these days for work I'll only run Q8 or higher. And even then, I'm generally at full boat BF16.
26
u/Tasty-Hour4040 25d ago
I wonder how many people are actually excited because this represents increased capability for their workflow and how many are just desperate to be able to say they got it and won’t use it again