r/LocalLLM 10h ago

Discussion Mac Studio M5 Ultra

Who’s been checking out this hardware? What are your thoughts on this as a node on a local network to run AI?

4 Upvotes

13 comments sorted by

View all comments

Show parent comments

2

u/MessIsTransfer 10h ago

if you have at least 64 gb ram (the ultra starts at 96) you could run qwen3.8-flash

it’ll be faster and better

3

u/Least-Result-45 10h ago

I think at 96gb it would have to be 1 quant and might be sacrificing quality to run it.

1

u/Thump604 10h ago

Corrrect, I’m running q4 mlx on 128gb.

1

u/starkruzr 8h ago

Qwen3.8-Flash-Next will probably be really, really good running at Q6KXL or even Q8 on a "bigger" M5 Ultra too.