CoolFace
Datasetpublic

sk-community/romanized_hindi

Romanized Hindi Dataset Dataset Description The Romanized Hindi Dataset is a collection of Hindi text paired with its Romanized (Latin script) representation. It has been created by combining multiple sources, including open datasets, synthetic generation, and rule-based transliteration methods. The dataset is designed for training and evaluating Hindi↔Roman transliteration models. Language(s): Hindi, Romanized Hindi Size: ~1.82M rows License: MIT (check with… See the full description on the dataset page: https://huggingface.co/datasets/sk-community/romanized_hindi.

sourceHugging Facemitupdated 1y agoView on Hugging Face
0likes117downloads
discussions and pull requests

Conversations for this repository live on Hugging Face.

CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.

Open discussions on Hugging Face
sk-community/romanized_hindi · CoolFace