Llamacpp, mlx-lm and derivatives. All via openai protocol. Different ports/hosts, no api key. Still experimenting (experimental codegen with small local models. My current experiment went from reqllm back to req. it seemed easier to experiment with tool call formats and slighly broken json directly than through reqllm. That was claude codes’ opinion though , so take it with a large dose of salt
).
I will check out your project later, tracking experiments over time is useful.






















