akaruineko/qualitext
QualiText QualiText is a balanced English text-classification dataset for studying text origin and text quality. Each example contains a text field and a label field. The dataset has five labels with the same number of examples in each class. Labels Label Description human Human-authored text from Wikipedia and 4chan. machine_generated Machine-generated text from the Qwen3.8-Max, GLM-5.2, and Kimi-K3 distillation corpus. corrupted Wikipedia text… See the full description on the dataset page: https://huggingface.co/datasets/akaruineko/qualitext.
Conversations for this repository live on Hugging Face.
CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.
Open discussions on Hugging Face