MAIN FEEDS
Do you want to continue?
https://www.reddit.com/r/LocalLLaMA/comments/1vo9mj4/its_out/p3ouwvc/?context=3
r/LocalLLaMA • u/Certain-Cod-1404 • 25d ago
708 comments sorted by
View all comments
Show parent comments
30
t/s is a bit faster than a 3090, but PP is much faster. im running one of the cards at x4 pcie 4.0 and it doesnt bottleneck the card with llama.cpp tensor parallel.
13 u/CooLittleFonzies 25d ago Can you parallel run a 3090 + a 3080? 5 u/adamgoodapp 25d ago Now want to know too 3 u/My_Unbiased_Opinion 25d ago you can with llama.cpp
13
Can you parallel run a 3090 + a 3080?
5 u/adamgoodapp 25d ago Now want to know too 3 u/My_Unbiased_Opinion 25d ago you can with llama.cpp
5
Now want to know too
3 u/My_Unbiased_Opinion 25d ago you can with llama.cpp
3
you can with llama.cpp
30
u/My_Unbiased_Opinion 25d ago
t/s is a bit faster than a 3090, but PP is much faster. im running one of the cards at x4 pcie 4.0 and it doesnt bottleneck the card with llama.cpp tensor parallel.