datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
stackoverflow-unified-text-open-status-classification
Dataset Card for "stackoverflow-unified-text-open-status-classification"
More Information needed
stackoverflow-unified-text-open-status-classification-sample
Dataset Card for "stackoverflow-open-status-classification"
More Information needed
ai-generated-text-classification
Dataset Card for "ai-generated-text-classification"
More Information needed
hc3-wiki-cleaned-text-for-domain-classification-roberta-tokenized-max-len-512
Dataset Card for "hc3-wiki-cleaned-text-for-domain-classification-roberta-tokenized-max-len-512"
More Information needed
text-classification-checkpoint-downloadsartificial_text_classification
Artificial Text Classification Dataset
Dataset Summary
The Artificial Text Classification dataset is designed to distinguish between human-generated and machine-generated text. This dataset provides labeled examples of text, enabling researchers and developers to train and evaluate machine learning models for text classification tasks.
Key features:
Text samples: Includes both human-written and machine-generated text.
Labels: Binary target variable where:
1 =… See the full description on the dataset page: https://huggingface.co/datasets/ds-claudia/artificial_text_classification.output_of_text_classification
