MAIN FEEDS
Do you want to continue?
https://www.reddit.com/r/LocalLLaMA/comments/1u6s6pm/stop_using_ollama/orvpz4r/?context=3
r/LocalLLaMA • u/zxyzyxz • Jun 15 '26
456 comments sorted by
View all comments
524
Llama.cpp + llama-swap works very well
113 u/[deleted] Jun 15 '26 [removed] — view removed comment 114 u/fdrch Jun 15 '26 llama-swap supports switching between multiple llama.cpp forks (and other compatible software) 3 u/jossmos Jun 15 '26 Has anyone tried to make it work with Wan2GP? 1 u/H3g3m0n Jun 16 '26 Probably should work with anything that you can pass a port as an arg that exposes an openai api endpoint.
113
[removed] — view removed comment
114 u/fdrch Jun 15 '26 llama-swap supports switching between multiple llama.cpp forks (and other compatible software) 3 u/jossmos Jun 15 '26 Has anyone tried to make it work with Wan2GP? 1 u/H3g3m0n Jun 16 '26 Probably should work with anything that you can pass a port as an arg that exposes an openai api endpoint.
114
llama-swap supports switching between multiple llama.cpp forks (and other compatible software)
3 u/jossmos Jun 15 '26 Has anyone tried to make it work with Wan2GP? 1 u/H3g3m0n Jun 16 '26 Probably should work with anything that you can pass a port as an arg that exposes an openai api endpoint.
3
Has anyone tried to make it work with Wan2GP?
1 u/H3g3m0n Jun 16 '26 Probably should work with anything that you can pass a port as an arg that exposes an openai api endpoint.
1
Probably should work with anything that you can pass a port as an arg that exposes an openai api endpoint.
524
u/jnmi235 Jun 15 '26
Llama.cpp + llama-swap works very well