carlosdanielhernandezmena/dimex100_light
Dataset Card for dimex100_light Dataset Summary The DIMEx100 LIGHT Corpus (DL) is a reduced version of the DIMEx100 Corpus (D100). DL was created in 2016 by Carlos Daniel Hernández Mena, with the aim of facilitating the use of the DIMEx100 Corpus in various automatic speech recognition systems. The most important differences between DIMEx100 LIGHT and the original are: The DL only contains audio files and transcriptions, unlike the D100 which contains… See the full description on the dataset page: https://huggingface.co/datasets/carlosdanielhernandezmena/dimex100_light.
Update README.md
Adding info the README file
Convert dataset to Parquet (#1)
Delete corpus/speech/remove.txt
Upload train.tar.gz
Adding files to the repo for the first time
initial commit
