OpenAI-compatible TTS endpoint using OmniVoice: 0.3s response time
I want to share a TTS server with an OpenAI-compatible API that generates speech really fast (about 0.3 seconds for a sentence on an RTX 3080) and can clone a voice from a short reference clip. I’ve optimized the server so generation starts quickly, and it processes long text in sequence, paragraph…
Read the full story at r/LocalLLaMA ↗
Timeline · 1 report
- 2026-10-08 16:31 · r/LocalLLaMA
OpenAI-compatible TTS endpoint using OmniVoice: 0.3s response time