Trained a clone* tool for Kokoro-82M; generates a voice pack in <1s on a 5s sample
This story is from 2026-09-11. It is preserved in the archive; the latest stories are on the live feed.
There have been new and arguably better TTS models since, but I have a soft spot for this one as lightweight stable and fast. I also maintain Kokoro-FastAPI, and have a bit of spare time on my hands so have been exploring what’s doable on the project. Customization/expressiveness are weak spots it…
Read the full story at r/huggingface ↗
Timeline · 1 report
- 2026-09-11 00:30 · r/huggingface
Trained a clone* tool for Kokoro-82M; generates a voice pack in <1s on a 5s sample