CoolFace
Datasetpublic

mephistosir329/Aksharantar

Dataset Card for Aksharantar Dataset Summary Aksharantar is the largest publicly available transliteration dataset for 20 Indic languages. The corpus has 26M Indic language-English transliteration pairs. Supported Tasks and Leaderboards [More Information Needed] Languages Assamese (asm) Hindi (hin) Maithili (mai) Marathi (mar) Punjabi (pan) Tamil (tam) Bengali (ben) Kannada (kan) Malayalam (mal) Nepali (nep)… See the full description on the dataset page: https://huggingface.co/datasets/mephistosir329/Aksharantar.

sourceHugging Faceccupdated 9mo agoView on Hugging Face
0likes42downloads
settings

This repository belongs to mephistosir329 on Hugging Face.

CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.

nameAksharantar
visibilitypublic
licencecc
gatedno
ownermephistosir329
Account settings
mephistosir329/Aksharantar · CoolFace