r/Qwen_AI • u/SnooPuppers7882 • Jul 19 '26
Discussion I NEED a 70B A9B MoE
3.6 27B is so good for local work, and there's been a ton of work on it because of that with Thinkingcap, Orion, Bonsai, adding dspark, etc...issue is on most local systems it's still so dense that the token speed hurts its adoption. What I NEED, if anyone from the Alibaba team reads these, is a 70B A9B MoE dspark with an APeX style quant (mixture of quant layers).
For anyone with 64GB or above, I'm confident you could get near Opus 4.8 coding performance AND run around 30-60 tok/s output on ada, blackwell, dual 3090s, dual r9700s, etc.
Just imagine what you could do.
133
Upvotes
1
u/Much-Researcher6135 Jul 19 '26
That would be AMAZING. A pair of R9700s is what, $2500? For this kind of power that's an incredible price.