CoolFace
Datasetpublic

itay1itzhak/flan_2022_350k

Dataset Card for Flan-2022 Subsample (Flan 350K) Dataset Summary This dataset was used in the paper on the origins of cognitive biases in LLMs: "Planted in Pretraining, Swayed by Finetuning: A Case Study on the Origins of Cognitive Biases in LLMs"https://arxiv.org/abs/2507.07186 This is a 350,000-example subsample of the original FLAN 2022 instruction-tuning dataset (https://arxiv.org/abs/2210.11416). It was created to provide a balanced, computationally efficient… See the full description on the dataset page: https://huggingface.co/datasets/itay1itzhak/flan_2022_350k.

sourceHugging Faceapache-2.0updated 1y agoView on Hugging Face
0likes70downloads
discussions and pull requests

Conversations for this repository live on Hugging Face.

CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.

Open discussions on Hugging Face
itay1itzhak/flan_2022_350k · CoolFace