NbAiLab/nb_distil_speech_noconcat_stortinget
Dataset Card for NbAiLab/nb_distil_speech_noconcat_stortinget Dataset Summary NbAiLab/nb_distil_speech_noconcat_stortinget is a curated subset of the Stortinget Speech Corpus (SSC), a large-scale Norwegian parliamentary speech dataset. This subset focuses on non-concatenated speech segments and includes automatic transcriptions generated using OpenAI's Whisper model. It is designed to facilitate the development and evaluation of Automatic Speech Recognition (ASR)… See the full description on the dataset page: https://huggingface.co/datasets/NbAiLab/nb_distil_speech_noconcat_stortinget.
2312
