legacy-datasets/common_voice
Common Voice is Mozilla's initiative to help teach machines how real people speak. The dataset currently consists of 7,335 validated hours of speech in 60 languages, but we’re always adding more voices and languages.
Remove deprecated tasks (#14)
Disable the viewer (#13)
Delete legacy JSON metadata (#11)
Update reference to Common Voice datasets uder mozilla-foundation org (from 11 to 13) (#7)
rename configs to config_name
Add deprecation warning (#4)
Reorder split names (#3)
add dataset_info in dataset metadata
remove dummmy data
Add `language_bcp47` tag (#4753)
Align more metadata with other repo types (models,spaces) (#4607)
Remove config names as yaml keys (#4367)
Update common_voice.py (#4212)
Update datasets task tags to align tags with models (#4067)
Use audio feature in ASR task template (#4006)
Simplify Common Voice code (#3817)
Local paths in common voice (#3736)
Common voice validated partition (#3669)
Update files from the datasets library (from 1.18.0)
Update files from the datasets library (from 1.17.0)
Update files from the datasets library (from 1.16.0)
Update files from the datasets library (from 1.13.3)
Update files from the datasets library (from 1.10.0)
Update files from the datasets library (from 1.9.0)
Update files from the datasets library (from 1.7.0)
Update files from the datasets library (from 1.6.1)
Update files from the datasets library (from 1.6.0)
Update files from the datasets library (from 1.5.0)
