New open dataset: 265 real recordings of people talking over each other, hand-labeled by segment (primary / mix / secondary / noise)
Disclosure: I work at Krisp. We just published a dataset for testing speech-to-text when a second person is talking nearby. We couldn't find a public one with real recordings, so we recorded our own. What's in it 265 recordings from open-plan offices, working call-center floors, and phone calls fro…
Read the full story at r/huggingface ↗