bilguun/ted_talks_en_mn_split
TED & TEDx Parallel Corpus (English-Mongolian) The dataset is composed of two distinct subsets: TED Talks (split en): English-language talks sourced from the official TED platform, paired with high-quality, human-generated Mongolian subtitles. TEDxUlaanbaatar (split mn): Mongolian-language talks from local TEDx events in Ulaanbaatar, paired with the original Mongolian subtitles and machine-translated English subtitles. This version of the dataset features segmented audio and… See the full description on the dataset page: https://huggingface.co/datasets/bilguun/ted_talks_en_mn_split.
Conversations for this repository live on Hugging Face.
CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.
Open discussions on Hugging Face