If you have both ollama and llama.cpp installed on your MBP, and are interested in utilizing your local model as a cost-saving preprocessor for your frontier models, consider using https://github.com/Standard-Pentest/kultivait!
"kultivait init --setup" to evaluate your machine's hardware, access to frontier cli-tools, download models, and create a proxy for ollama to get started. works great with opencode!
I’ve been using OMLX on 48gb Mac. It’s a lot smoother setup than Ollama and has a built in model sizer and all that in a nice ui. Also will pre configure and launch opencode, pi, Hermes, codex and Claude out of the box with no fuss. I’ve been really happy with it.
"kultivait init --setup" to evaluate your machine's hardware, access to frontier cli-tools, download models, and create a proxy for ollama to get started. works great with opencode!
Ollama proxy and model selection does work. If you run into issues or have any questions please feel free to reach out.
More ethical alternatives have caught up including llama.cpp
https://sleepingrobots.com/dreams/stop-using-ollama/#what-to...
- https://sleepingrobots.com/dreams/stop-using-ollama/
- also can we get a post on how to do this with llama.cpp