CoolFace
Datasetpublic

lilgoose777/tibetan-speech-english-text-dataset

Tibetan Speech Dataset with English Translations Dataset Description This dataset contains Tibetan speech recordings paired with transcriptions in Tibetan script and English translations. It is designed to support automatic speech recognition (ASR), machine translation, and text-to-speech (TTS) research for the Tibetan language, which is considered a low-resource language in NLP. Supported Tasks Automatic Speech Recognition (ASR): Train models to… See the full description on the dataset page: https://huggingface.co/datasets/lilgoose777/tibetan-speech-english-text-dataset.

sourceHugging Facecc-by-4.0updated 8mo agoView on Hugging Face
0likes46downloads
3 commits on main
74fd90a8mo ago

Upload README.md with huggingface_hub

lilgoose777
ba7ddb88mo ago

Upload dataset

lilgoose777
ea7ef168mo ago

initial commit

lilgoose777