CoolFace
Datasetpublic

danijelkorzinek/ClarinStudioPL

CLARIN-PL Polish Studio Corpus The corpus was created somewhere in 2014-2015 by recording a group of few hundred volunteer speakers reading a few dozen sentences each. The total size of the corpus is ~56 hours. Due to the manner of recording, the transcription accuracy is very high, but the manner of speech is not spontaneous. This corpus is best compared to something like TIMIT, possibly CommonVoice. It is different from CommonVoice in that it is recorded in a controlled… See the full description on the dataset page: https://huggingface.co/datasets/danijelkorzinek/ClarinStudioPL.

sourceHugging Faceotherupdated 2y agoView on Hugging Face
2likes130downloads
6 commits on main
39b3f2a2y ago

Upload dataset

danijelkorzinek
bc0bc732y ago

Upload dataset

danijelkorzinek
579e6152y ago

Upload dataset

danijelkorzinek
94942ec2y ago

Update README.md

danijelkorzinek
371520a2y ago

Update README.md

danijelkorzinek
e6161712y ago

initial commit

danijelkorzinek