datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
qe4pe
Quality Estimation for Post-Editing (QE4PE)
For more details on QE4PE, see our paper and our Github repository
Gabriele Sarti • Vilém Zouhar • Grzegorz Chrupała • Ana Guerberof Arenas • Malvina Nissim • Arianna Bisazza
Word-level quality estimation (QE) detects erroneous spans in machine translations, which can direct and facilitate human post-editing. While the accuracy of word-level QE systems has been assessed extensively, their usability and downstream influence on the… See the full description on the dataset page: https://huggingface.co/datasets/gsarti/qe4pe.us-gsa-surplus-auctions
U.S. Government (GSA) Surplus Auction Dataset
This dataset lists completed U.S. federal surplus auction lots sold through GSA Auctions (https://gsaauctions.gov), one row per lot. It is compiled and published by GovAuctions.app (https://govauctions.app) and is the lot-level companion to the GovAuctions.app Surplus Price Index (https://govauctions.app/research/surplus-price-index).
Canonical page: https://govauctions.app/research/open-dataset
Source repository, updated monthly:… See the full description on the dataset page: https://huggingface.co/datasets/govauctions/us-gsa-surplus-auctions.seq_level_training_datagold_dataset_8kgold_dataset_4k_balancedgold_dataset_4k
