CoVoST datasets with invalidated, validated and other csv formats
tabular
3 yrs ago
Last commit cannot be located
About
CoVoST 2 is a large-scale multilingual speech translation corpus covering translations from 21 languages into English and from English into 15 languages. The dataset is created using Mozillas open-source Common Voice database of crowdsourced voice recordings.