CoolFace
Datasetpublic

Aalto-Speech-Synthesis/stortinget_speech_corpus_v1.0

Dataset Card for Stortinget Speech Corpus V1.0 Overview This is the WebDataset version of the Stortinget Speech Corpus V1.0, originally created by the National Library of Norway. We re-organize it into WebDataset format for better usability. The Stortinget Speech Corpus (SSC) is a 5000+ hours speech dataset for weak supervision ASR created from audio andaligned proceedings text from Stortinget, the Norwegian Parliament. For more information, please refer to the… See the full description on the dataset page: https://huggingface.co/datasets/Aalto-Speech-Synthesis/stortinget_speech_corpus_v1.0.

sourceHugging Facecc0-1.0updated 5mo agoView on Hugging Face
0likes104downloads
11 commits on main
1518ba55mo ago

Update README.md

lzrhahahahaha
220ed8f5mo ago

Add files using upload-large-folder tool

lzrhahahahaha
1d048325mo ago

Add files using upload-large-folder tool

lzrhahahahaha
96789155mo ago

Add files using upload-large-folder tool

lzrhahahahaha
86c09785mo ago

Update README.md

lzrhahahahaha
e8cadad5mo ago

Update README.md

lzrhahahahaha
f21a8775mo ago

Update README.md

lzrhahahahaha
6c8db505mo ago

Create original_ssc_v1_0_dataset_card.md

lzrhahahahaha
6ef430d5mo ago

Update README.md

lzrhahahahaha
03bf1445mo ago

Update README.md

lzrhahahahaha
d41e9335mo ago

initial commit

lzrhahahahaha