CoolFace
Datasetpublic

NbAiLab/nb_distil_speech_noconcat_stortinget

Dataset Card for NbAiLab/nb_distil_speech_noconcat_stortinget Dataset Summary NbAiLab/nb_distil_speech_noconcat_stortinget is a curated subset of the Stortinget Speech Corpus (SSC), a large-scale Norwegian parliamentary speech dataset. This subset focuses on non-concatenated speech segments and includes automatic transcriptions generated using OpenAI's Whisper model. It is designed to facilitate the development and evaluation of Automatic Speech Recognition (ASR)… See the full description on the dataset page: https://huggingface.co/datasets/NbAiLab/nb_distil_speech_noconcat_stortinget.

sourceHugging Facecc0-1.0updated 1y agoView on Hugging Face
2likes312downloads

NbAiLab/nb_distil_speech_noconcat_stortinget · main · files are served by the source, never re-hosted here