CoolFace
Datasetpublic

vrclc/openslr63

SLR63: Crowdsourced high-quality Malayalam multi-speaker speech data set This data set contains transcribed high-quality audio of Malayalam sentences recorded by volunteers. The data set consists of wave files, and a TSV file (line_index.tsv). The file line_index.tsv contains a anonymized FileID and the transcription of audio in the file. The data set has been manually quality checked, but there might still be errors. Please report any issues in the following issue tracker on… See the full description on the dataset page: https://huggingface.co/datasets/vrclc/openslr63.

sourceHugging Facecc-by-4.0updated 3y agoView on Hugging Face
3likes36downloads

vrclc/openslr63 · main · files are served by the source, never re-hosted here