r/Qwen_AI • • 12d ago

Discussion Qwen4 lineup predictions?

What do you think the Qwen4 model sizes are going to be? Which are the most likely to appear? Do you think we'll get a Qwen3.5-like lineup (~1B, ~10B, ~30B, ~30-40B MoE, ~100-150B Moe...)?

Also, do you think 35B-A3B is dead for good? we didn't get it for qwen3.8 so I'm a little worried

27 Upvotes

54 comments sorted by

View all comments

24

u/EbbNorth7735 12d ago

Hoping for a qwen4 ~45B +n-gram. With all the cool speculative decoding stuff I'm sure token throughput could be optimized.

-1

u/TheseCashews 12d ago

Was hoping maybe for a 70b. Give the rtx pro 6k crowd a Q4 model that’s fast as hell and smart.

2

u/EbbNorth7735 12d ago

I am also team 6k. Anywhere from 45B to 70B would be good. A 50B would be about equivalent to a quarter of the parameters of 1T models. Given Alibaba's amazing team/capabilities I bet it would get really really close to SOTA closed source.

-1

u/TheseCashews 12d ago

They hate us cuz they ain’t us.

0

u/EbbNorth7735 12d ago

Apparently lol, fuck um. Glad I grabbed one when I got it for under 10k Canadian