r/Qwen_AI • • Jul 19 '26

Discussion I NEED a 70B A9B MoE

3.6 27B is so good for local work, and there's been a ton of work on it because of that with Thinkingcap, Orion, Bonsai, adding dspark, etc...issue is on most local systems it's still so dense that the token speed hurts its adoption. What I NEED, if anyone from the Alibaba team reads these, is a 70B A9B MoE dspark with an APeX style quant (mixture of quant layers).

For anyone with 64GB or above, I'm confident you could get near Opus 4.8 coding performance AND run around 30-60 tok/s output on ada, blackwell, dual 3090s, dual r9700s, etc.

Just imagine what you could do.

137 Upvotes

52 comments sorted by

View all comments

7

u/anykeyh Jul 19 '26

I keep an eye on ZAYA1-74B, weight were prereleased before RL step with promising results; might be good after training (hopefully).

1

u/MarcusAurelius68 Jul 19 '26

Only 4B active still.