CoolFace
20 results

voxpopuli

facebook /voxpopuli Dataset Card for Voxpopuli Dataset Summary VoxPopuli is a large-scale multilingual speech corpus for representation learning, semi-supervised learning and interpretation. The raw data is collected from 2009-2020 European Parliament event recordings. We acknowledge the European Parliament for creating and sharing these materials. This implementation contains transcribed speech data for 18 languages. It also contains 29 hours of transcribed speech data of non-native… See the full description on the dataset page: https://huggingface.co/datasets/facebook/voxpopuli.audioautomatic-speech-recognition1M<n<10M164 likes76k downloads8mo agoHugging Facesonalsannigrahi /voxpopuli_mosel_curatortabular10M<n<100M0 likes4.6k downloads3mo agoHugging Faceqmeeus /voxpopuli Dataset Card for "voxpopuli" More Information needed audio100K<n<1M0 likes3.1k downloads3y agoHugging FaceArtificialAnalysis /VoxPopuli-Cleaned-AA VoxPopuli-Cleaned-AA Quick links: AA Speech to Text Leaderboard | AA-WER v2.0 article VoxPopuli-Cleaned-AA is a cleaned subset of the English VoxPopuli test data from esb/datasets, a speech dataset derived from European Parliament recordings. This cleaned subset is the VoxPopuli portion included in AA-WER v2. We manually reviewed and corrected errors in the original ground-truth transcriptions to ensure fairer evaluation of Speech to Text (STT) models. This dataset is part of AA-WER… See the full description on the dataset page: https://huggingface.co/datasets/ArtificialAnalysis/VoxPopuli-Cleaned-AA.audioautomatic-speech-recognitionn<1K7 likes1.2k downloads7mo agoHugging Facemteb /VoxPopuliAccentPairClassificationFilteredaudio1K<n<10K0 likes1.1k downloads8mo agoHugging Faceilsp /voxpopuli_elaudio1M<n<10M0 likes942 downloads11mo agoHugging Face