faisal4590aziz/bangla-health-related-paraphrased-dataset
Dataset Card for "BanglaHealthParaphrase" BanglaHealthParaphrase is a Bengali paraphrasing dataset specifically curated for the health domain. It contains over 200,000 sentence pairs, where each pair consists of an original Bengali sentence and its paraphrased version. The dataset was created through a multi-step pipeline involving extraction of health-related content from Bengali news sources, English pivot-based paraphrasing, and back-translation to ensure linguistic… See the full description on the dataset page: https://huggingface.co/datasets/faisal4590aziz/bangla-health-related-paraphrased-dataset.
Update README.md
Update README.md
Update README.md
Update README.md
200k cleaned and final data
tags update
cleaned paraphrased dataset
Delete all_paraphrased_data.csv
tags updated
Upload all_paraphrased_data.csv
Delete all_paraphrased_data.csv
200K bengali "Health Domain Focused" paraphrased sentences
Delete paraphrased_sentences.xlsx
paraphrased_sentences
Upload paraphrased_sentences.xlsx
Create README.md
initial commit
