This is my OpenCode setup for local models, mainly using a customized llama.cpp configuration.
A developer detailed their OpenCode setup for running local AI models, centered on a customized llama.cpp configuration. The setup leverages reasoning-effort levels for models like Qwen 3.8, DeepSeek …