CoolFace
Datasetpublicgated

likhonsheikh/BanglaNLP

BanglaNLP: Bengali-English Parallel Dataset Tools BanglaNLP is a comprehensive toolkit for creating high-quality Bengali-English parallel datasets from news sources, designed to improve machine translation and other cross-lingual NLP tasks for the Bengali language. Our work addresses the critical shortage of high-quality parallel data for Bengali, the 7th most spoken language in the world with over 230 million speakers. ๐Ÿ† Impact & Recognition 120K+โ€ฆ See the full description on the dataset page: https://huggingface.co/datasets/likhonsheikh/BanglaNLP.

sourceHugging Facemitupdated 2y agoView on Hugging Face
2likes8downloads
Dataset Card

No card is published for this repository, or it could not be fetched from Hugging Face right now.