Janus strips away Python runtime bloat and Docker virtualization, embedding llama.cpp directly inside a lightweight Go binary with a Vulkan acceleration bridge. By uniting hot-swappable GGUF model execution with native reasoning tag extraction, it offers a single-file gateway for local inference across AMD, Intel, and NVIDIA hardware.
StudyBuddy AI: Transforming Messy Lecture Notes into Interactive Quizzes with Local Gemma Models