CoolFace
5 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01artmelancholy /golos_mfa_punctuation_long Golos MFA Punctuation (Long) Long-form Russian speech derived from govnejri/golos_mfa_punctuation. Purpose Most public Russian STT corpora ship as short clips (a few seconds each). For benchmarking long-form transcription, VAD, punctuation, and streaming behavior, you want minutes-long audio with reliable word-level alignments. This dataset builds those long clips by splicing groups of consecutive short clips together, inserting randomized silences between them, and… See the full description on the dataset page: https://huggingface.co/datasets/artmelancholy/golos_mfa_punctuation_long.audioautomatic-speech-recognition1K<n<10K1 likes18 downloads4mo agoHugging Face02french-datasets /AdoCleanCode_french-multi-mfa_train_v1_previewCe répertoire est vide, il a été créé pour améliorer le référencement du jeu de données AdoCleanCode/french-multi-mfa_train_v1_preview . automatic-speech-recognition0 likes8 downloads9mo agoHugging Face03french-datasets /AdoCleanCode_french-multi-mfa_TRAIN_V0Ce répertoire est vide, il a été créé pour améliorer le référencement du jeu de données AdoCleanCode/french-multi-mfa_TRAIN_V0. automatic-speech-recognition0 likes7 downloads9mo agoHugging Face04french-datasets /AdoCleanCode_french-multi-mfaCe répertoire est vide, il a été créé pour améliorer le référencement du jeu de données AdoCleanCode/french-multi-mfa. automatic-speech-recognition0 likes6 downloads9mo agoHugging Face05french-datasets /AdoCleanCode_french-multi-mfa_train_v1Ce répertoire est vide, il a été créé pour améliorer le référencement du jeu de données AdoCleanCode/french-multi-mfa_train_v1. automatic-speech-recognition0 likes6 downloads9mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.