Inflect-v2: 3.9M and 9.3M parameter open-weight TTS models Inflect-v2, two open-weight English TTS models at 3.9M and 9.3M parameters, generate speech multiple times faster than real-time on CPU while delivering quality competitive with larger systems like KittenTTS, Piper, and Supertonic-3. The models support CPU, CUDA, PyTorch, and ONNX under Apache 2.0. Post 34 Introducing Inflect-v2, two exceptionally small, open-weight English TTS models at just 3.9M and 9.3M parameters. Both generate speech multiple times faster than real-time on CPU. Despite their size, Inflect-v2 delivers quality that is competitive with much larger lightweight TTS systems, including KittenTTS, Piper, and Supertonic-3. CPU, CUDA, PyTorch, and ONNX are supported. Apache 2.0. See it for yourselves: Try the Demos: CPU, CUDA, PyTorch, and ONNX are supported. Apache 2.0. See it for yourselves: owensong/Inflect-Micro-v2 https://huggingface.co/owensong/Inflect-Micro-v2 owensong/Inflect-Nano-v2 https://huggingface.co/owensong/Inflect-Nano-v2 Try the Demos: