CoolFace
Datasetpublic

mephistosir329/Aksharantar

Dataset Card for Aksharantar Dataset Summary Aksharantar is the largest publicly available transliteration dataset for 20 Indic languages. The corpus has 26M Indic language-English transliteration pairs. Supported Tasks and Leaderboards [More Information Needed] Languages Assamese (asm) Hindi (hin) Maithili (mai) Marathi (mar) Punjabi (pan) Tamil (tam) Bengali (ben) Kannada (kan) Malayalam (mal) Nepali (nep)… See the full description on the dataset page: https://huggingface.co/datasets/mephistosir329/Aksharantar.

sourceHugging Faceccupdated 9mo agoView on Hugging Face
0likes42downloads
1 commits on main
a5d7c149mo ago

Duplicate from ai4bharat/Aksharantar

mephistosir329, Satya10