itay1itzhak/flan_2022_350k
Dataset Card for Flan-2022 Subsample (Flan 350K) Dataset Summary This dataset was used in the paper on the origins of cognitive biases in LLMs: "Planted in Pretraining, Swayed by Finetuning: A Case Study on the Origins of Cognitive Biases in LLMs"https://arxiv.org/abs/2507.07186 This is a 350,000-example subsample of the original FLAN 2022 instruction-tuning dataset (https://arxiv.org/abs/2210.11416). It was created to provide a balanced, computationally efficient… See the full description on the dataset page: https://huggingface.co/datasets/itay1itzhak/flan_2022_350k.
This repository belongs to itay1itzhak on Hugging Face.
CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.
