CoolFace
Datasetpublic

UMCU/WikiDocPatientInformation_Dutch_translated_with_MariaNMT

Dataset Card for "WikiDocPatientInformation_Dutch_translated_with_MariaNMT" Translation of the English version of the Hugging dataset WikiDoc patient information, based on WikiDoc, a medical wikipedia. to Dutch using an Maria NMT model, trained by Helsinki NLP. Note, for reference: Maria NMT is based on BART, described here. Attribution If you use this dataset please use the following to credit the creators of the OPUS-MT models:… See the full description on the dataset page: https://huggingface.co/datasets/UMCU/WikiDocPatientInformation_Dutch_translated_with_MariaNMT.

sourceHugging Facegpl-3.0updated 3y agoView on Hugging Face
0likes22downloads
Dataset Card

Dataset Card for "WikiDocPatientInformationDutchtranslatedwithMariaNMT"

Translation of the English version of the Hugging dataset WikiDoc patient information, based on WikiDoc, a medical wikipedia. to Dutch using an Maria NMT model, trained by Helsinki NLP. Note, for reference: Maria NMT is based on BART, described here.

Attribution

If you use this dataset please use the following to credit the creators of the OPUS-MT models:

@InProceedings{TiedemannThottingal:EAMT2020,
  author = {J{\"o}rg Tiedemann and Santhosh Thottingal},
  title = {{OPUS-MT} — {B}uilding open translation services for the {W}orld},
  booktitle = {Proceedings of the 22nd Annual Conferenec of the European Association for Machine Translation (EAMT)},
  year = {2020},
  address = {Lisbon, Portugal}
 }

and

@misc {van_es_2024,
	author       = { {Bram van Es} },
	title        = { WikiDocPatientInformation_Dutch_translated_with_MariaNMT (Revision 4490701) },
	year         = 2024,
	url          = { https://huggingface.co/datasets/UMCU/WikiDocPatientInformation_Dutch_translated_with_MariaNMT },
	doi          = { 10.57967/hf/1669 },
	publisher    = { Hugging Face }
}

License

For both the Maria NMT model and the original Helsinki NLP Opus MT model we did not find a license. We also did not find a license for the MedQA corpus. For these reasons we use a permissive CC BY license. If this was in error please let us know and we will add the appropriate licensing promptly.