CoolFace
Datasetpublic

sentence-transformers/quora-duplicates

Dataset Card for Quora Duplicate Questions This dataset contains the Quora Question Pairs dataset in four formats that are easily used with Sentence Transformers to train embedding models. The data was originally created by Quora for this Kaggle Competition. Dataset Subsets pair-class subset Columns: "sentence1", "sentence2", "label" Column types: str, str, class with {"0": "different", "1": "duplicate"} Examples:{ 'sentence1': 'What is the step… See the full description on the dataset page: https://huggingface.co/datasets/sentence-transformers/quora-duplicates.

sourceHugging Faceupdated 4mo agoView on Hugging Face
11likes1.1kdownloads
discussions and pull requests

Conversations for this repository live on Hugging Face.

CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.

Open discussions on Hugging Face