CoolFace
Datasetpublic

danijelkorzinek/ClarinStudioPL

CLARIN-PL Polish Studio Corpus The corpus was created somewhere in 2014-2015 by recording a group of few hundred volunteer speakers reading a few dozen sentences each. The total size of the corpus is ~56 hours. Due to the manner of recording, the transcription accuracy is very high, but the manner of speech is not spontaneous. This corpus is best compared to something like TIMIT, possibly CommonVoice. It is different from CommonVoice in that it is recorded in a controlled… See the full description on the dataset page: https://huggingface.co/datasets/danijelkorzinek/ClarinStudioPL.

sourceHugging Faceotherupdated 2y agoView on Hugging Face
2likes130downloads

danijelkorzinek/ClarinStudioPL · main · files are served by the source, never re-hosted here