CoolFace
Datasetpublic

IbrahimAmin/arz-en-parallel-corpus

Egyptian Arabic - English Parallel Corpus 🇪🇬✨🇬🇧 Dataset Description This dataset is a cleaned and filtered merge of multiple Egyptian Arabic - English parallel corpora, containing ~27,000 aligned sentence pairs. It’s designed for researchers and developers working on machine translation, speech translation, and other NLP tasks involving Egyptian Arabic and English. Sources 📚 This dataset integrates and refines data from the following publicly… See the full description on the dataset page: https://huggingface.co/datasets/IbrahimAmin/arz-en-parallel-corpus.

sourceHugging Facemitupdated 1y agoView on Hugging Face
3likes57downloads
12 commits on main
ab6b10c1y ago

Update README.md

IbrahimAmin
6593ed51y ago

Update README.md

IbrahimAmin
9e443a41y ago

Update README.md

IbrahimAmin
f0bf9181y ago

Update README.md

IbrahimAmin
43eef471y ago

Update README.md

IbrahimAmin
11a14241y ago

Upload dataset

IbrahimAmin
26b78951y ago

Update README.md

IbrahimAmin
0c685911y ago

Update README.md

IbrahimAmin
35cc0061y ago

Update README.md

IbrahimAmin
64b75b31y ago

Upload dataset

IbrahimAmin
aa6b5191y ago

Update README.md

IbrahimAmin
8ec4c251y ago

initial commit

IbrahimAmin