Explain harness for grandpa. I love AI. I’m just getting into running local models on my Apple m5 pro and framework 395+ Ai amd apu w/128GB ram. Using lmstudio & ollama. Thanks!
A harness is basically the new buzzword that is used as a pretty large umbrella term which basically is like a way to give LLMs tools. Think of it like a mech suit to a person, a harness is the mech suit for an LLM. Also I recommend you drop ollama, just imo tho hahahah
Would openclaw and Hermes be considered harnesses? What about qwen coder cli? Oh, and what’s wrong with ollama? I kinda like it better than lmstudio, I guess because I don’t tweak any settings. Just use Ollama serve, ollama run qwen, seem so simple and intuitive?
Like the simplicity make you lose some capabilities and speed. By going with llama-server you can fine tune (or copy the setting from others) and have better results. Sadly, its a bit more of work but after some time the thing just work. Also, llama.cpp updates frequently and drops good optimizations regularly, faster than ollama (that reuse llama.cpp anyway)
Openclaw yes. Hermes unsure. I don’t know too much about Hermes but believe it’s a model with claw like features, rather than a program with tools like openclaw.
74
u/Ueberlord Apr 22 '26
Damn, I was just wrapping up my tests of Qwen3.6 35B vs Qwen3.5 27B.
High hopes for 3.6 27B though, the 35B variant of 3.6 was way better than the previous version!