viswadarshan06/stsb-binary-paraphrase-labelled
Paraphrase Detection Dataset (Derived from SetFit/stsb) Description: This dataset originates from the SetFit/stsb dataset, which was initially created for semantic textual similarity (STS) tasks with a label range of 0 to 5. It has been adapted for binary paraphrase detection by leveraging the high-accuracy paraphrase classification model viswadarshan06/pd-robert. Each sentence pair in the original dataset has been re-labeled according to the following binary scheme: 1 →… See the full description on the dataset page: https://huggingface.co/datasets/viswadarshan06/stsb-binary-paraphrase-labelled.
015
