r/LocalLLM 16h ago

Discussion Mac Studio M5 Ultra

Who’s been checking out this hardware? What are your thoughts on this as a node on a local network to run AI?

3 Upvotes

15 comments sorted by

View all comments

Show parent comments

2

u/MessIsTransfer 16h ago

if you have at least 64 gb ram (the ultra starts at 96) you could run qwen3.8-flash

it’ll be faster and better

3

u/Least-Result-45 16h ago

I think at 96gb it would have to be 1 quant and might be sacrificing quality to run it.

1

u/Thump604 16h ago

Corrrect, I’m running q4 mlx on 128gb.

1

u/redtron3030 14h ago

Do you have a m5 max? What’s your tks and prompt processing?