Causal Analysis and Mitigation of Spurious Onsets in Full-Duplex Speech LLMs
This story is from 2026-09-15. It is preserved in the archive; the latest stories are on the live feed.
arXiv:2609.13445v1 Announce Type: new Abstract: Speech-to-speech LLMs like Moshi, and its derivative PersonaPlex, can listen and speak concurrently through full-duplex generation. However, they can begin speaking inappropriately during prolonged user silence: under digital-zero input, Moshi and Per…
Read the full story at arXiv cs.CL ↗
Timeline · 1 report
- 2026-09-15 04:00 · arXiv cs.CL
Causal Analysis and Mitigation of Spurious Onsets in Full-Duplex Speech LLMs