CoolFace
Datasetpublic

ntt123/VietBibleVox-aligned

VietBibleVox Dataset The VietBibleVox Dataset is based on the data extracted from open.bible specifically for the Vietnamese language. As the original data is provided under the cc-by-sa-4.0 license, this derived dataset is also licensed under cc-by-sa-4.0. The dataset comprises 29,185 pairs of (verse, audio clip), with each verse from the Bible read in Vietnamese by a male voice. The verses are the original texts and may not be directly usable for training text-to-speech… See the full description on the dataset page: https://huggingface.co/datasets/ntt123/VietBibleVox-aligned.

sourceHugging Facecc-by-sa-4.0updated 1y agoView on Hugging Face
2likes304downloads
Dataset Card

VietBibleVox Dataset

The VietBibleVox Dataset is based on the data extracted from open.bible specifically for the Vietnamese language. As the original data is provided under the cc-by-sa-4.0 license, this derived dataset is also licensed under cc-by-sa-4.0.

The dataset comprises 29,185 pairs of (verse, audio clip), with each verse from the Bible read in Vietnamese by a male voice.

  • —The verses are the original texts and may not be directly usable for training text-to-speech models.
  • —The clips are in MP3 format with a sample rate of 48k.