Update of Mozilla Common Voice 12.0 voice data
Mozilla has updated the Common Voice voice data sets, which include pronunciation samples from over 200,000 individuals. The data is published as public domain (CC0). The provided sets can be used in machine learning systems for building speech recognition and synthesis models. Compared to the previous update, the volume of speech material in the collection increased from 23.8 to 25.8 thousand hours of speech. In […]
