CoolFace
Datasetpublic

silma-ai/silma-arabic-english-sts-dataset-v1.0

SILMA STS Arabic/English Dataset - v1.0 Overview The SILMA STS Arabic/English Dataset - v1.0 is a dataset designed for training and evaluating sentence embeddings for Arabic and English tasks. It consists of five different splits that cover monolingual and multilingual sentence pairs, with human-annotated similarity scores. The dataset includes both Arabic-to-Arabic and English-to-English pairs, as well as cross-lingual Arabic-English pairs, making it a valuable… See the full description on the dataset page: https://huggingface.co/datasets/silma-ai/silma-arabic-english-sts-dataset-v1.0.

sourceHugging Faceapache-2.0updated 2y agoView on Hugging Face
3likes40downloads
4 commits on main
18856902y ago

Update README.md

karimouda
86529122y ago

Update README.md

bakrianoo
17949b72y ago

Upload ar_en-sts with scores dataset

bakrianoo
e2db3ae2y ago

initial commit

bakrianoo