danijelkorzinek/ClarinStudioPL
CLARIN-PL Polish Studio Corpus The corpus was created somewhere in 2014-2015 by recording a group of few hundred volunteer speakers reading a few dozen sentences each. The total size of the corpus is ~56 hours. Due to the manner of recording, the transcription accuracy is very high, but the manner of speech is not spontaneous. This corpus is best compared to something like TIMIT, possibly CommonVoice. It is different from CommonVoice in that it is recorded in a controlled… See the full description on the dataset page: https://huggingface.co/datasets/danijelkorzinek/ClarinStudioPL.
2130
Upload dataset
Upload dataset
Upload dataset
Update README.md
Update README.md
initial commit
