datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
reflection-50m
SPP Reflection 50M
The 51.4M-document reflection set from Synthetic Persona Pretraining (SPP):
Alignment from Token Zero — the production half-corpus run, and the dataset the
released models were actually trained on.
🔬 Small sample (same format): dlab-spp/reflection-sample-2k
📉 Earlier 10M run: dlab-spp/reflection-10m
🧾 Safety scores for the full 1T corpus: dlab-spp/safety-classifications
Each row pairs a source document with two generated constitution reflections — a… See the full description on the dataset page: https://huggingface.co/datasets/dlab-spp/reflection-50m.details_SenseLLM__ReflectionCoder-DS-33B
Dataset Card for Evaluation run of SenseLLM/ReflectionCoder-DS-33B
Dataset automatically created during the evaluation run of model SenseLLM/ReflectionCoder-DS-33B.
The dataset is composed of 136 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_SenseLLM__ReflectionCoder-DS-33B.reflection-10m
SPP Reflection 10M
The full ~10M-document reflection set from Synthetic Persona Pretraining (SPP):
Alignment from Token Zero.
📝 Read the post: Synthetic Persona Pretraining: Alignment from Token Zero
🔬 Small sample (same format): dlab-spp/reflection-sample-2k — a 2,000-row sample drawn from this set, for quick inspection.
Each row pairs a pretraining document with a synthetic, value-laden reflection
generated for it: a short first-person (and third-person) moral reflection… See the full description on the dataset page: https://huggingface.co/datasets/dlab-spp/reflection-10m.reflection_eval_prompt1details_terrycraddock__Reflection-Llama-3.1-8B
Dataset Card for Evaluation run of terrycraddock/Reflection-Llama-3.1-8B
Dataset automatically created during the evaluation run of model terrycraddock/Reflection-Llama-3.1-8B.
The dataset is composed of 136 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_terrycraddock__Reflection-Llama-3.1-8B.details_SenseLLM__ReflectionCoder-CL-34B
Dataset Card for Evaluation run of SenseLLM/ReflectionCoder-CL-34B
Dataset automatically created during the evaluation run of model SenseLLM/ReflectionCoder-CL-34B.
The dataset is composed of 136 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_SenseLLM__ReflectionCoder-CL-34B.reflection_eval_prompt2test_reflection_eval_promptllama3_8b_reflection2train_reflection_eval1reflection_n169_8_25__countdown_4arg__sft_data_mp_reflection_ckpt_chunk_8orca-math-word-reflection
Dataset Card for "orca-math-word-reflection"
Dataset Summary
The Orca-Math Word Problems with Reflection dataset is an extension of subset of the original ORCA Math Word Problems 200k dataset. This new version introduces a "Thinking and Reflection" format designed to enhance problem-solving approaches by encouraging step-by-step thinking before producing a solution.
In this dataset, each math word problem and its corresponding solution from the original dataset are… See the full description on the dataset page: https://huggingface.co/datasets/Harshkmr/orca-math-word-reflection.9_8_25__letter_countdown_4o__sft_data_mp_reflection_ckpt_chunk_5distilabel-reflection-tuning
Dataset Card for distilabel-reflection-tuning
This dataset has been created with distilabel.
The pipeline script was uploaded to easily reproduce the dataset:
reflection.py.
It can be run directly using the CLI:
distilabel pipeline run --script "https://huggingface.co/datasets/gabrielmbmb/distilabel-reflection-tuning/raw/main/reflection.py"
Dataset Summary
This dataset contains a pipeline.yaml which can be used to reproduce the pipeline that generated… See the full description on the dataset page: https://huggingface.co/datasets/gabrielmbmb/distilabel-reflection-tuning.train_reflection_eval2_with_rewardstrain_reflection_eval2_with_rewards2reflection-v1-exact-duplicatestest_reflection_eval_completions_6synthetic-think-and-reflection_v1This synthetic dataset was generated to improvement model thinking and reflection in systematical way.
Columns
question
answer_with_tags
text (processed for Llama 3.1)
Dataset
DatasetDict({
train: Dataset({
features: ['question', 'answer_with_tags', 'text'],
num_rows: 17129
})
test: Dataset({
features: ['question', 'answer_with_tags', 'text'],
num_rows: 52
})
})
Credit
inspired by mattshumer (sharedGPT dataset) and… See the full description on the dataset page: https://huggingface.co/datasets/mychen76/synthetic-think-and-reflection_v1.skillfactory_sft_countdown_3arg_qrepeat1_reflections5_formats0C.-C.-C-IC.-CCreflection-v1reflection-v1-sharegpttest_reflection_eval_completion1_with_rewards9_8_25__countdown_3arg__sft_data_mp_reflection_ckpt_chunk_59OT_Ref_NoV13Part___openthoughts__1600000_end2000000__reflection_chunk_3Maggen-Reflection-3.1-70b-50k-filtered-scoredDatasetDict({
train: Dataset({
features: ['created', 'response', 'pre_query_template', 'instruction', 'gen_input_configs', 'gen_response_configs', 'raw_instruction', 'id', 'instruction_sanitize_class_num', 'scores', 'model_name'],
num_rows: 36884
})
})
每个唯一值的计数:
scores
[9.0] 9469
[7.0] 6224
[10.0] 6009
[6.0] 4003
[8.0] 3149
[5.0] 2578
[4.0] 2575
[3.0] 1566
[2.0] 1051
[1.0] 209
[] 51
reflection-v0reflection-sample-2k
SPP Reflection 2k Sample
A 2,000-row sample (seed 42) of dlab-spp/reflection-10m,
in the identical format, for quick inspection of the data from
Synthetic Persona Pretraining (SPP): Alignment from Token Zero.
📝 Read the post: Synthetic Persona Pretraining: Alignment from Token Zero
📦 Full dataset: dlab-spp/reflection-10m (~10M documents).
Each row pairs a pretraining document with a synthetic, value-laden reflection
(first- and third-person) grounded in a value constitution.… See the full description on the dataset page: https://huggingface.co/datasets/dlab-spp/reflection-sample-2k.test_reflection_eval_completion4_with_rewards
