CoolFace
Datasetpublicgated

themohal/pakistan-languages-dataset

Pakistan Languages Dataset A parallel dataset of English sentences translated into ~37 languages of Pakistan. One row per English sentence; one column per language, using its ISO 639-3 code. ⚠️ Translation quality varies a lot by language — read this before using the data Every language column is force-filled: no cell is ever left blank. For high-resource languages this just means a verified translation. For the lowest-resource languages, it can mean best guess at… See the full description on the dataset page: https://huggingface.co/datasets/themohal/pakistan-languages-dataset.

sourceHugging Faceupdated 5d agoView on Hugging Face
0likes28downloads
discussions and pull requests

Conversations for this repository live on Hugging Face.

CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.

Open discussions on Hugging Face
themohal/pakistan-languages-dataset · CoolFace