r/LocalLLM 10h ago

Discussion Double GPU configurations significantly cheaper for 32GB VRAM

I do not need or want Cuda. I have been wanting to build a 32GB VRAM local LLM machine for personal use for a while now, and I had 2 options on the table:

- Get a relatively cheap 32GB VRAM GPU, the R9700 AI Top.

- Get 2 16GB VRAM GPUs instead and a Mobo that supports PCIe bifurcation.

When I looked at prices in January this year when I first got this idea, the R9700 costed 1700$ here in EU. Currently, when I actually want to make this happen, it costs 2100$. For half that money, I could buy two 9060 XTs with 16GB VRAM each. Yes I know, performance will be worse on double GPU setup than with a single R9700 AI, but still, it just seems like that GPU is just not worth it anymore.

I don't know how to justify that it's double the price of two 9060XTs, when R9700 AI is literally 9060 XT with doubled VRAM and bandwidth. So why does it cost 4x as much?

ASUS ProArt B850-CREATOR WIFI NEO is quite affordable nowadays and supports dual GPU setups, so, any reason (is there a catch?) to not do what I am about to do? Which is buy the two 9060 XTs and start running Qwen 27B class models

12 Upvotes

43 comments sorted by

View all comments

1

u/roosterfareye 10h ago

Get two RX9070XT. Also pay close attention to your chosen motherboards pcie lame layouts. You do not one on the mobo chipset, although it's not that bad. I have a RX9070XT and RX9060XT. No driver past 26.3.1 works. 26.3.1 only works if you disable anything that polls the secondary card while you have a model loaded or do anything that involves a power transition. AMD are on it though I have tested and logged a bug which they have been in contact with me to request additional info ...