Dialect-Robust Speech Language Models with Synthetic Pseudo-Dialect Augmentation
arXiv:2610.09321v1 Announce Type: new Abstract: Speech Language Model (SLM) performance often degrades on dialects due to data scarcity. Conventional text-to-speech (TTS) augmentation struggles to cover diverse dialects as it requires a certain amount of real dialect speech. We propose synthesizing…
Read the full story at arXiv cs.CL ↗
Timeline · 1 report
- 2026-10-08 04:00 · arXiv cs.CL
Dialect-Robust Speech Language Models with Synthetic Pseudo-Dialect Augmentation