7 ms·
I've built real time mobile app to provide visual feedback for vowel articulation based on the formants[0](2d/3d) graph and live recording to teach accent, stil
by user070223 3y ago
I've built real time mobile app to provide visual feedback for vowel articulation based on the formants[0](2d/3d) graph and live recording to teach accent, still need more work, and data for different languages (as the "accepted" vowel region, i.e the region where a native speaker expects the vowel to exist is somewhat different between languages and dialects)
The problem is that when someone tries to learn a new language after years of exposure to his own language, he never really practices pronunciation usually because there are fewer vowels in his language so they map to bigger regions.
I'm sure many people had the experience that said something to native english speaker which they were sure they said correct but the listener just had a blank face and didn't understand.
As this is very delicate thing (see vowel shift) I'd rather get it right the first time and not teach something wrong but one could make his own vowel chart approximation[1] (Although it's not that consistent, and there are multiple algorithm which would produce different results). I would also like to expand it to consonants
There is a good physical explanation for why vowels are charachterized by formants. With vowels you don't restrict the airflow so the pitch wavecreated in the vocal cords travels in 3d in the vocal tract some would go straight to the exit, and some would bounce so there is destructive interference patterns in most places but for some frequencies(overtones) there is a constructive interference so the amplitude('energy') for that frequency is higher. One could model it with quadrilateral pipes of different length and shape as described in the video. I think the frequency change from you pitch is due to the 'bouncing' on the vocal tract some energy would be Refracted and some would reflect
Another cool thing about the vowel space is that the most common vowels are the one that lies on the boundary of the chat as far from each other(highest diff between the formants), this is of course increases understanding as there is a clear cut betweeen the different sounds, but if you could teach clear distinction you could "pack" more data at the same time.
Teachings kids to distinguish more precisly between vowels is just like teaching them "Color terms"[2]
[0] https://en.wikipedia.org/wiki/Formant https://en.wikipedia.org/wiki/Formant
[1] https://www.youtube.com/watch?v=BGW8J4cG0qY https://www.youtube.com/watch?v=BGW8J4cG0qY
[2] https://en.wikipedia.org/wiki/Color_term https://en.wikipedia.org/wiki/Color_term