sk-community/romanized_hindi
Romanized Hindi Dataset Dataset Description The Romanized Hindi Dataset is a collection of Hindi text paired with its Romanized (Latin script) representation. It has been created by combining multiple sources, including open datasets, synthetic generation, and rule-based transliteration methods. The dataset is designed for training and evaluating Hindi↔Roman transliteration models. Language(s): Hindi, Romanized Hindi Size: ~1.82M rows License: MIT (check with… See the full description on the dataset page: https://huggingface.co/datasets/sk-community/romanized_hindi.
0117
Update README.md
Update README.md
Upload 5 files
initial commit
