CoolFace
Datasetpublic

Zarakun/youtube_ua_noisy_subtitles_test

The list of all subsets in the dataset Each subset is generated splitting videos from given particular ukrainiam YouTube channel All subsets are in test split "opodcast" subset is from channel "О! ПОДКАСТ" "rozdympodcast" subset is from channel "Роздум | Подкаст" "test" subset is just a small subset of samples Loading a particular subset >>> data_files = {"train": "data/<your_subset>.parquet"} >>> data = load_dataset("Zarakun/youtube_ua_subtitles_test"… See the full description on the dataset page: https://huggingface.co/datasets/Zarakun/youtube_ua_noisy_subtitles_test.

sourceHugging Faceupdated 3y agoView on Hugging Face
0likes11downloads
settings

This repository belongs to Zarakun on Hugging Face.

CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.

nameyoutube_ua_noisy_subtitles_test
visibilitypublic
licencenot set
gatedno
ownerZarakun
Account settings
Zarakun/youtube_ua_noisy_subtitles_test · CoolFace