The Death of a Language

Kyle interviews Zane and Leena about the Endangered Languages Project . My learnings are
- Project is taking in 3.5 hours of audio content from an endangered language called “Ladin”
- It creates phonetic transcriptions from audio samples of human languages
- Model has so far produced decent levels of vowel identifications
- Currently working on phoneme segmentation and larger consonant categories
- From the project blurb
In this project, we are trying to speed up the process of language documentation by building a model that produces phonetic transcriptions from audio samples of human languages. The ultimate goal of our project is to develop a model that could be applied to any human language with minimal changes. We will be using around 3-4 hours of partially labeled audio data in an endangered language called Ladin, which we are using as our main training/test data. As of now we have produced some decent results in vowel identifications and are currently working on phoneme segmentation and identification of larger consonant categories.




