CoolFace
Datasetpublic

akaruineko/qualitext

QualiText QualiText is a balanced English text-classification dataset for studying text origin and text quality. Each example contains a text field and a label field. The dataset has five labels with the same number of examples in each class. Labels Label Description human Human-authored text from Wikipedia and 4chan. machine_generated Machine-generated text from the Qwen3.8-Max, GLM-5.2, and Kimi-K3 distillation corpus. corrupted Wikipedia text… See the full description on the dataset page: https://huggingface.co/datasets/akaruineko/qualitext.

sourceHugging Facecc0-1.0updated 1mo agoView on Hugging Face
0likes31downloads
discussions and pull requests

Conversations for this repository live on Hugging Face.

CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.

Open discussions on Hugging Face
akaruineko/qualitext · CoolFace