Ditching the GPU: High-Quality Local TTS with Kokoro
Kokoro, an 82-million parameter open-weight text-to-speech model released under Apache 2.0, achieves high-quality 24kHz speech synthesis on modest CPU hardware without requiring a GPU. Its efficient architecture, based o…