CoolFace
Datasetpublic

ando55/WikiSQE_experiment

Dataset Card for WikiSQE_experiment Dataset Summary WikiSQE_experiment is the official evaluation split for WikiSQE: A Large‑Scale Dataset for Sentence Quality Estimation in Wikipedia. While the parent dataset (ando55/WikiSQE) contains every sentence flagged with a quality problem in the full edit history of English Wikipedia, this repo provides the exact train/validation/test partitions used in the AAAI 2024 paper. It offers ≈ 8.3 million sentences organised as:… See the full description on the dataset page: https://huggingface.co/datasets/ando55/WikiSQE_experiment.

sourceHugging Facecc-by-sa-4.0updated 1y agoView on Hugging Face
0likes278downloads
discussions and pull requests

Conversations for this repository live on Hugging Face.

CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.

Open discussions on Hugging Face
ando55/WikiSQE_experiment · CoolFace