Exploring emergent articulatory learning: Notes on an 8-view marker-calibrated visual speech architecture (MouthMind)
Hi everyone, Over the past months, I’ve been quietly documenting a theoretical and empirical inquiry into silent visual speech decoding (VSR). My main interest was to understand how much phonetic structure can truly be reconstructed from pure articulatory kinematics when acoustic feedback is absent…
Read the full story at r/computervision ↗
Timeline · 2 reports
- 2026-09-26 13:38 · r/deeplearning
Exploring emergent articulatory learning: Notes on an 8-view marker-calibrated visual speech architecture (MouthMind) - 2026-09-26 13:38 · r/computervision
Exploring emergent articulatory learning: Notes on an 8-view marker-calibrated visual speech architecture (MouthMind)