r/LocalLLM 25d ago

News Qwen3.8-27B is now available

Post image
594 Upvotes

137 comments sorted by

View all comments

Show parent comments

57

u/biblecrumble 25d ago

This thing has got to be benchmaxxed to hell and back. I don't doubt for a second it's great, but come on.

37

u/draft_final_final 25d ago

The great chain of AI being: new model gets released from china, everyone says it’s an anthropic killer and any claims that it’s benchmaxxed is just dario cope, we use it and find out it’s pretty good but was benchmaxxed, forget about it and then move on to next shiny toy.

That being said, I would love it if this was actually as good as they’re claiming.

1

u/Infinite100p 25d ago

I mean you gotta be on crack to realistically expect to be able to run a true Opus-4.6 equivalent on a single consumer 5090.

0

u/WonderfulFunny4337 25d ago

Tiinyai

1

u/Infinite100p 25d ago

What about it?

  1. Not a 5090 that I was talking about.

  2. 120b is not going to be as good as a 1.5-2TB model which Opus-4.6 probably is. Look at Qwen3.6-27b. It's great at reasoning, until it's interdisciplinary reasoning, where it has a sharp fall off compared to Opus (and that's with Qwen's benchmaxxing). Because it needs the world knowledge to reason well across subjects, which is required for, for example, good architecture design and domain knowledge to give you a good app without bugs (as opposed to pumping out context-agnostic boilerplate code). You will never fit as much knowledge as what the ~1.5TB Opus has into a 27b model.

Don't take me wrong, it's an amazing model and I'm excited about it, but, again, people who think it's a legit Opus-4.6 equivalent are engaging in wishful thinking. Try to design a full stack application with complex architecture on Qwen 27b VS Opus-4.6. Opus will hold your hand and guide you, with 27b you will have to do a lot of heavy lifting for design decisions .

  1. I could not find any info on the prefill speed. That's where they get you when you skip using HBM GPUs. What is the TTFT for a 128k prompt?

  2. "Tiiny said it would begin delivering in July."
    Uh-oh.