Tossed distorted audio samples to an open-weight voice model; it did fairly well.
This story is from 2026-09-04. It is preserved in the archive; the latest stories are on the live feed.
Being a person obsessed with testing new models that come out, times are really insane for me. Tested different kinds of TTS and voice cloning models but none of them gets it right in terms of emotion and pace, you know which one is fake in seconds; they just fail in emotions. Spotted Confucius4 on…
Read the full story at r/LocalLLaMA ↗
Timeline · 1 report
- 2026-09-04 12:50 · r/LocalLLaMA
Tossed distorted audio samples to an open-weight voice model; it did fairly well.