Running the WUIC assistant on a local LLM: Ollama, an MCP server, and a free agentic VS Code
The WUIC framework team replaced the cloud-based generation half of its RAG chatbot with a local LLM running via Ollama, cutting per-token costs and privacy exposure. A side effect emerged: exposing t…