CoolFace
Datasetpublic

s-nlp/en_paradetox_content

ParaDetox: Detoxification with Parallel Data (English). Content Task Results This repository contains information about Content Task markup from English Paradetox dataset collection pipeline. The original paper "ParaDetox: Detoxification with Parallel Data" was presented at ACL 2022 main conference. ParaDetox Collection Pipeline The ParaDetox Dataset collection was done via Yandex.Toloka crowdsource platform. The collection was done in three steps: Task 1:… See the full description on the dataset page: https://huggingface.co/datasets/s-nlp/en_paradetox_content.

sourceHugging Faceopenrail++updated 3y agoView on Hugging Face
0likes22downloads

s-nlp/en_paradetox_content · main · files are served by the source, never re-hosted here