odunola/sermon-pair
Sermon Transcript Dataset Description This dataset is a curated collection of approximately 960,000 rows of sermon transcripts, obtained from sermons uploaded on YouTube. Each sermon was transcribed, split into smaller passage-like segments, and then manually labeled over a span of 4 weeks. The labeling process extracted the last Bible passage mentioned in each segment. Purpose The dataset was created to facilitate the development of a plugin for… See the full description on the dataset page: https://huggingface.co/datasets/odunola/sermon-pair.
Conversations for this repository live on Hugging Face.
CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.
Open discussions on Hugging Face