CoolFace
Datasetpublic

mteb/BornholmBitextMining

BornholmBitextMining An MTEB dataset Massive Text Embedding Benchmark Danish Bornholmsk Parallel Corpus. Bornholmsk is a Danish dialect spoken on the island of Bornholm, Denmark. Historically it is a part of east Danish which was also spoken in Scania and Halland, Sweden. Task category t2t Domains Web, Social, Fiction, Written Reference https://aclanthology.org/W19-6138/ Source datasets: strombergnlp/bornholmsk_parallel How to evaluate on this task… See the full description on the dataset page: https://huggingface.co/datasets/mteb/BornholmBitextMining.

sourceHugging Facecc-by-4.0updated 7mo agoView on Hugging Face
0likes11kdownloads
settings

This repository belongs to mteb on Hugging Face.

CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.

nameBornholmBitextMining
visibilitypublic
licencecc-by-4.0
gatedno
ownermteb
Account settings
mteb/BornholmBitextMining · CoolFace