CoolFace
4 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01Reza2kn /nasle-mana-clean-chunked-30s-avasanj Nasl-e-Mana Clean Persian Speech — corrected 30-second chunks Corrected, provenance-preserving audio chunks collected from the Nasl-e-Mana magazine website, generated on 2026-08-30. This release supersedes the earlier unreliable proportional-mapping chunk export; that older release was not used here. Splits Split Rows Audio Columns labeled 4,981 41.41 hours audio, label to_transcribe 11,127 92.72 hours audio The labeled split contains the… See the full description on the dataset page: https://huggingface.co/datasets/Reza2kn/nasle-mana-clean-chunked-30s-avasanj.audioautomatic-speech-recognition10K<n<100K1 likes671 downloads16d agoHugging Face02Aviv-anthonnyolime /SIWIS_French_Speech_Synthesis_Database SIWIS French Speech Synthesis Database This README provides a concise description of the dataset, including its structure, file naming conventions, and known labeling issues. Additionally, suggestions for potential improvements are outlined in the TODO section. The dataset is distributed under the Creative Commons Attribution 4.0 International (CC BY 4.0) license, permitting its use for any purpose. For more details about the database design and recording process, please refer… See the full description on the dataset page: https://huggingface.co/datasets/Aviv-anthonnyolime/SIWIS_French_Speech_Synthesis_Database.audioautomatic-speech-recognition10K<n<100K0 likes301 downloads2y agoHugging Face03ggfox00000 /dia-AvaAvd-test AVA-AVD — test split (audio-visual speaker diarization in the wild) Copie du split test d'AVA-AVD (Xu et al., ACM MM 2022), un corpus de diarisation dans des conditions "in the wild" construit sur AVA Active Speaker. Chaque clip (~5 min) est extrait des vidéos AVA à des offsets précis définis par le repo officiel, puis l'audio est re-synchronisé : les timestamps des RTTMs ont été recalculés en temps clip-local (soustrait min_start du RTTM d'origine) pour être directement utilisables… See the full description on the dataset page: https://huggingface.co/datasets/ggfox00000/dia-AvaAvd-test.audiovoice-activity-detectionn<1K0 likes148 downloads5mo agoHugging Face04avemio /ASR-GERMAN-MIXED-TEST Dataset Beschreibung Dieser Datensatz und die Beschreibung wurde von flozi00/asr-german-mixed übernommen und nur der Test-Split hier hochgeladen, da Hugging Face native erst einmal alle Splits herunterlädt. Für eine Evaluation von Speech-to-Text Modellen ist ein Download von 136 GB allerdings etwas zeit- & speicherraubend, weshalb wir hier nur den Test-Split für Evaluationen anbieten möchten. Die Arbeit und die Anerkennung sollten deshalb weiter bei primeline & flozi00 für die… See the full description on the dataset page: https://huggingface.co/datasets/avemio/ASR-GERMAN-MIXED-TEST.audioautomatic-speech-recognition1K<n<10K3 likes57 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.