google-research-datasets/paws
Dataset Card for PAWS: Paraphrase Adversaries from Word Scrambling Dataset Summary PAWS: Paraphrase Adversaries from Word Scrambling This dataset contains 108,463 human-labeled and 656k noisily labeled pairs that feature the importance of modeling structure, context, and word order information for the problem of paraphrase identification. The dataset has two subsets, one based on Wikipedia and the other one based on the Quora Question Pairs (QQP) dataset. For… See the full description on the dataset page: https://huggingface.co/datasets/google-research-datasets/paws.
Convert dataset to Parquet (#3)
rename configs to config_name
Replace YAML keys from int to str (#2)
Fix CI test_inspect (#1)
add dataset_info in dataset metadata
remove dummmy data
fix task_ids
Align more metadata with other repo types (models,spaces) (#4607)
Remove config names as yaml keys (#4367)
task id update (#4244)
Update datasets task tags to align tags with models (#4067)
Update files from the datasets library (from 1.16.0)
Update files from the datasets library (from 1.7.0)
Update files from the datasets library (from 1.6.1)
Update files from the datasets library (from 1.6.0)
Update files from the datasets library (from 1.3.0)
Update files from the datasets library (from 1.2.0)
