CoolFace
Datasetpublicgated

RonitMehta260704/kh-en-dataset

KYNMAW — Khasi Bhashini Multi-Task Dataset The largest open multi-task speech + text dataset for Khasi, an Austroasiatic language spoken by roughly 1.5 million people in Meghalaya, Northeast India — built to help bring Khasi into modern speech recognition, text-to-speech, and machine translation systems as part of If you're working on low-resource ASR, endangered-language NLP, Indic speech translation, or Northeast Indian languages, this dataset is for you.… See the full description on the dataset page: https://huggingface.co/datasets/RonitMehta260704/kh-en-dataset.

sourceHugging Facecc-by-4.0updated 3mo agoView on Hugging Face
1likes2downloads
settings

This repository belongs to RonitMehta260704 on Hugging Face.

CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.

namekh-en-dataset
visibilitypublic
licencecc-by-4.0
gatedyes
ownerRonitMehta260704
Account settings
RonitMehta260704/kh-en-dataset · CoolFace