datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
wmt-da-human-evaluation-long-context
Dataset Summary
Long-context / document-level dataset for Quality Estimation of Machine Translation.
It is an augmented variant of the sentence-level WMT DA Human Evaluation dataset.
In addition to individual sentences, it contains augmentations of 2, 4, 8, 16, and 32 sentences, among each language pair lp and domain.
The raw column represents a weighted average of scores of augmented sentences using character lengths of src and mt as weights.
The code used to apply the augmentation… See the full description on the dataset page: https://huggingface.co/datasets/ymoslem/wmt-da-human-evaluation-long-context.Human-Evaluation
Human Evaluation Dataset
The dataset includes human evaluation for General and Health domains. It was created as part of my two papers:
“Domain-Specific Text Generation for Machine Translation” (Moslem et al., 2022)
"Adaptive Machine Translation with Large Language Models" (Moslem et al., 2023)
The evaluators were asked to assess the acceptability of each translation
using a scale ranging from 1 to 4, where 4 is ideal and 1 is unacceptable translation.
For the paper Moslem et al.… See the full description on the dataset page: https://huggingface.co/datasets/ymoslem/Human-Evaluation.customerservice-Human-evaluation-results-evaluator_2customerservice-Human-evaluation-results-evaluator_3customerservice-Human-evaluation-results-overallcustomerservice-Human-evaluation-results-evaluator_1mt-human-evaluation-da
