Self-Rollouts
self-taught-reasoner-rollouts
STaR Rationale Dataset (CommonsenseQA, Llama-2-7B)
This repository contains bootstrapped rationale datasets produced by a small replication of STaR (Self-Taught Reasoner) on the CommonsenseQA training split using Llama-2-7B as the base model (M_0).
Each .jsonl file corresponds to one iteration of the STaR pipeline and stores a set of question–rationale–answer triples collected during that iteration.
Data generation procedure (STaR iteration)
Let the training… See the full description on the dataset page: https://huggingface.co/datasets/parksoojae/self-taught-reasoner-rollouts.qwen3_5_27b_ab_self_promotion_rolloutsqwen3_6_27b_ab_self_promotion_rolloutsglm_5_2_fp8_ab_self_promotion_rolloutsdab-step-self-play-sonnet-rollouts-sharegpt
