RonitMehta260704/kh-en-dataset
KYNMAW — Khasi Bhashini Multi-Task Dataset The largest open multi-task speech + text dataset for Khasi, an Austroasiatic language spoken by roughly 1.5 million people in Meghalaya, Northeast India — built to help bring Khasi into modern speech recognition, text-to-speech, and machine translation systems as part of If you're working on low-resource ASR, endangered-language NLP, Indic speech translation, or Northeast Indian languages, this dataset is for you.… See the full description on the dataset page: https://huggingface.co/datasets/RonitMehta260704/kh-en-dataset.
13
No card is published for this repository, or it could not be fetched from Hugging Face right now.
