CoolFace
15 results

openslr

openslr /librispeech_asr Dataset Card for librispeech_asr Dataset Summary LibriSpeech is a corpus of approximately 1000 hours of 16kHz read English speech, prepared by Vassil Panayotov with the assistance of Daniel Povey. The data is derived from read audiobooks from the LibriVox project, and has been carefully segmented and aligned. Supported Tasks and Leaderboards automatic-speech-recognition, audio-speaker-identification: The dataset can be used to train a model for Automatic… See the full description on the dataset page: https://huggingface.co/datasets/openslr/librispeech_asr.audioautomatic-speech-recognition100K<n<1M245 likes51k downloads1y agoHugging Faceopenslr /openslrOpenSLR is a site devoted to hosting speech and language resources, such as training corpora for speech recognition, and software related to speech recognition. We intend to be a convenient place for anyone to put resources that they have created, so that they can be downloaded publicly.automatic-speech-recognition1K<n<10K31 likes519 downloads2y agoHugging FaceSadique5 /openslr_quranic_asraudio10K<n<100K3 likes516 downloads2y agoHugging Facevoice-biomarkers /openslr-140-hq-Kazakh Kazakh Speech Dataset (KSD) Identifier: SLR140 Source: https://www.openslr.org/140/ Summary: High-quality open source Kazakh speech corpus developed by the Department of Artificial Intelligence and Big Data of Al-Farabi Kazakh National University (554 hours) Category: Speech License: Attribution-ShareAlike 3.0 Unported (CC BY-SA 3.0 US) About this resource: High-quality open source Kazakh speech corpus. The corpus contains about 554 hours of transcribed audio recordings… See the full description on the dataset page: https://huggingface.co/datasets/voice-biomarkers/openslr-140-hq-Kazakh.audio100K<n<1M4 likes330 downloads2y agoHugging Facethantzinphyo /burmese-speech-refined-openslr-80 Burmese Speech Refined OpenSLR-80 Summary This dataset is a speech dataset developed based on the original OpenSLR Dataset (SLR80), with the text and audio data carefully reviewed and further refined for Burmese language applications. In the original OpenSLR Dataset, the Burmese text was transcribed based on how the words were pronounced in the corresponding audio recordings. In this dataset, the original audio and text data were used as a reference, and the text… See the full description on the dataset page: https://huggingface.co/datasets/thantzinphyo/burmese-speech-refined-openslr-80.audioautomatic-speech-recognition1K<n<10K2 likes313 downloads22d agoHugging Faceedifier99 /sinhala-openslr-111haudio100K<n<1M0 likes248 downloads4mo agoHugging Face