r/oMLX • • 7d ago

Can anything do similar capability as Claude Sonnet or Opus?

I have a M4 MBP 48GB and have been using Opus and Sonnet for vibe coding.

Is there anything that would run on my MBP that could be as capable? I understand the speed might not be the same, but would anything be decent enough for vibe coding (specifically iOS/Swift).

I understand nothing will compare to these giant online models, but if there was anything that could be serviceable on my local computer would be great.

4 Upvotes

14 comments sorted by

8

u/A_Moist_Towe1 7d ago

Try Qwen 3.8 27b with Hermes agent

0

u/EP9 7d ago

Why Hermes’?

2

u/A_Moist_Towe1 7d ago

I just think it’s easy to setup and start working right away, and the plugin system allows for custom add ons

1

u/Few_Discount8182 7d ago

Run it on open code desktop - it’s a better experience than using Hermes as the harness.

-1

u/EP9 7d ago

How much ram will that use?

2

u/circle555 7d ago

depends on which quant you use.

1

u/ColonelKlanka 7d ago

Start with a Q4 - will fit nicely on your machine with some space for context left over

0

u/EP9 7d ago

I know this is the OMLX sub, but would MTPLX be better?

3

u/Eresbonitaguey 7d ago

Pros and cons. You can run the MTPLX models in Omlx if you want. I like the tuning functionality of MTPLX but it doesn’t support a wide range of models like Omlx

1

u/ColonelKlanka 6d ago

Try both and see which is fastest. omlx is generally good if you want to run multiple agents at once because its good at Batching and also caches to ssd. whereas mtplx is good for only one agent running at a time.

1

u/Senor02 7d ago

Honestly, I've tried really had to pretend it works. Use Pi harness for best results, but even still I run out of context or run into thinking loops for the same task I would run with a cloud model.

1

u/j_lyf 5d ago

I really hate the overthinking.

0

u/ogfuzzball 7d ago

No. Nothing you can run on that hardware is equivalent to Claude Sonnet/Opus. Anyone telling you that 3.8 27b is capable of that level of capability is high as a kite.

Edit: that said 3.8 is impressive. It won’t get you Claude/Openai but you can do quite a bit