datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
GSM8K-Aug-Llama-3.2-1B-Instruct-Correct-CoT
Verified self-generated GSM8K reasoning
64 independently sampled completions are generated per prepared question.
Final answers are checked against the source answer. Among complete, correctly
formatted correct completions whose CoT passes the final-result-statement and
combined length checks, one sample is selected uniformly at random using a
reproducible per-question seed. CoT length does not rank eligible samples.
The final result belongs
in the separate final-answer line of… See the full description on the dataset page: https://huggingface.co/datasets/hanseungwook/GSM8K-Aug-Llama-3.2-1B-Instruct-Correct-CoT.ai-vs-human-meta-llama-Llama-3.2-1B-Instruct
AI vs Human dataset on the CNN Daily mails
Dataset Description
This dataset showcases pairs of truncated articles and their respective completions, crafted either by humans or an AI language model.
Each article was randomly truncated between 25% and 50% of its length.
The language model was then tasked with generating a completion that mirrored the characters count of the original human-written continuation.
Data Fields
'human': The original human-authored… See the full description on the dataset page: https://huggingface.co/datasets/zcamz/ai-vs-human-meta-llama-Llama-3.2-1B-Instruct.
