CoolFace
Datasetpublic

PiC/phrase_similarity

Phrase in Context is a curated benchmark for phrase understanding and semantic search, consisting of three tasks of increasing difficulty: Phrase Similarity (PS), Phrase Retrieval (PR) and Phrase Sense Disambiguation (PSD). The datasets are annotated by 13 linguistic experts on Upwork and verified by two groups: ~1000 AMT crowdworkers and another set of 5 linguistic experts. PiC benchmark is distributed under CC-BY-NC 4.0.

sourceHugging Facecc-by-nc-4.0updated 4y agoView on Hugging Face
7likes92downloads
50 commits on main
fc67ce74y ago

Update phrase_similarity.py

PiC
ad52db64y ago

Update phrase_similarity.py

PiC
70ae5264y ago

Update phrase_similarity.py

PiC
7c9eeb54y ago

Update phrase_similarity.py

PiC
3f0827f4y ago

Update phrase_similarity.py

PiC
0955a544y ago

Delete data

PiC
db450e54y ago

Update phrase_similarity.py

PiC
5f598784y ago

Upload dataset_infos.json with huggingface_hub

PiC
fb6f1e64y ago

Delete data/test-00000-of-00001-da1042069f928a8c.parquet with huggingface_hub

PiC
0c549ab4y ago

Upload data/test-00000-of-00001-f89456dc448d8e32.parquet with huggingface_hub

PiC
ccd0eb24y ago

Delete data/validation-00000-of-00001-e53bf03df554e682.parquet with huggingface_hub

PiC
b56d5c44y ago

Upload data/validation-00000-of-00001-6df700a94a026715.parquet with huggingface_hub

PiC
ab33bee4y ago

Delete data/train-00000-of-00001-438d56cdbd9c112e.parquet with huggingface_hub

PiC
2686bae4y ago

Upload data/train-00000-of-00001-20fae14a62b0fda5.parquet with huggingface_hub

PiC
2c9e9234y ago

Update phrase_similarity.py

PiC
5f51b814y ago

Upload dataset_infos.json with huggingface_hub

PiC
ce1bb654y ago

Upload dataset_infos.json with huggingface_hub

PiC
0db25764y ago

Delete data/test-00000-of-00001-b3500099f476da1b.parquet with huggingface_hub

PiC
6f5cd4a4y ago

Upload data/test-00000-of-00001-da1042069f928a8c.parquet with huggingface_hub

PiC
675b6904y ago

Delete data/validation-00000-of-00001-aba01ce418f7fbe5.parquet with huggingface_hub

PiC
7b0ab4d4y ago

Upload data/validation-00000-of-00001-e53bf03df554e682.parquet with huggingface_hub

PiC
abba69b4y ago

Delete data/train-00000-of-00001-c17e47113937a357.parquet with huggingface_hub

PiC
96638b04y ago

Upload data/train-00000-of-00001-438d56cdbd9c112e.parquet with huggingface_hub

PiC
cea1db24y ago

Update README.md

PiC
ec44bb04y ago

Update new data with more challenging negative examples.

PiC
406cb9b4y ago

Upload dataset_infos.json with huggingface_hub

PiC
b37f1cc4y ago

Upload data/test-00000-of-00001-b3500099f476da1b.parquet with huggingface_hub

PiC
83377a34y ago

Upload data/validation-00000-of-00001-aba01ce418f7fbe5.parquet with huggingface_hub

PiC
aa4dc704y ago

Upload data/train-00000-of-00001-c17e47113937a357.parquet with huggingface_hub

PiC
7470bdf4y ago

Update new data with more challenging negative examples.

PiC
e58e68a4y ago

Update phrase_similarity.py

PiC
80e6ae74y ago

Remove redundant subset.

PiC
ec825fa4y ago

Update phrase_similarity.py

PiC
0c6454f4y ago

Update phrase_similarity.py

PiC
50a50704y ago

Update phrase_similarity.py

PiC
8ae7b164y ago

Update phrase_similarity.py

PiC
17884474y ago

Update README.md

PiC
a3954364y ago

Update README.md

PiC
08f31be4y ago

Fix `license` metadata (#1)

PiC, julien-c
93b3d1f4y ago

Update README.md

PiC
ee67ffe4y ago

Update README.md

PiC
a4d366e4y ago

Update README.md

PiC
23ad17a4y ago

Update README.md

PiC
dbb39274y ago

Added dataset license, authors, and cite

PiC
6e83ba94y ago

Update README.md

PiC
6bbb4e44y ago

Update README.md

PiC
e74cbd94y ago

Update README.md

PiC
f83abc24y ago

Update README.md

PiC
7f262504y ago

Update README.md

PiC
474c1564y ago

Update README.md

PiC