faster-whisper int8 dropped up to 60 s of speech. float32 didn't
I run a small podcast transcription tool, and while testing German audio I noticed that some transcripts were missing whole passages. Not wrong words. Twenty, thirty, sometimes close to sixty seconds of speech, gone. The timestamps simply jumped ahead, the text still read fine, and nothing in the l…
Read the full story at DEV Community — Machine Learning ↗
Timeline · 1 report
- 2026-10-09 00:27 · DEV Community — Machine Learning
faster-whisper int8 dropped up to 60 s of speech. float32 didn't