likhonsheikh/BanglaNLP
BanglaNLP: Bengali-English Parallel Dataset Tools BanglaNLP is a comprehensive toolkit for creating high-quality Bengali-English parallel datasets from news sources, designed to improve machine translation and other cross-lingual NLP tasks for the Bengali language. Our work addresses the critical shortage of high-quality parallel data for Bengali, the 7th most spoken language in the world with over 230 million speakers. ๐ Impact & Recognition 120K+โฆ See the full description on the dataset page: https://huggingface.co/datasets/likhonsheikh/BanglaNLP.
This repository belongs to likhonsheikh on Hugging Face.
CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.
