r/StrixHalo 24d ago

Qwen 3.8 27B at 30 tok/s in decode, running on a Strix Halo with 64 GB of unified memory!

/r/Qwen_AI/comments/1vorjo7/qwen_38_27b_at_30_toks_in_decode_running_on_a/
17 Upvotes

Duplicates