Aalto-Speech-Synthesis/stortinget_speech_corpus_v1.0
Dataset Card for Stortinget Speech Corpus V1.0 Overview This is the WebDataset version of the Stortinget Speech Corpus V1.0, originally created by the National Library of Norway. We re-organize it into WebDataset format for better usability. The Stortinget Speech Corpus (SSC) is a 5000+ hours speech dataset for weak supervision ASR created from audio andaligned proceedings text from Stortinget, the Norwegian Parliament. For more information, please refer to the… See the full description on the dataset page: https://huggingface.co/datasets/Aalto-Speech-Synthesis/stortinget_speech_corpus_v1.0.
Update README.md
Add files using upload-large-folder tool
Add files using upload-large-folder tool
Add files using upload-large-folder tool
Update README.md
Update README.md
Update README.md
Create original_ssc_v1_0_dataset_card.md
Update README.md
Update README.md
initial commit
