ceselder/loracle-pretrain-mix-oneq
loracle-pretrain-mix This is the oneq subsample: one randomly-selected QA row per organism_id (deterministic shuffle with seed=42, then drop_duplicates). Same 3 splits, same row schema, half the rows. Source: ceselder/loracle-pretrain-mix. Built for loracle-training scale ablations where we want each training step to expose the model to a fresh organism (no 2-QA-per-org redundancy). Split sizes: data/train.parquet: 25000 rows (25000 organisms) data/dpo_heldout.parquet: 250… See the full description on the dataset page: https://huggingface.co/datasets/ceselder/loracle-pretrain-mix-oneq.
README: 1-QA-per-org notice + split sizes
upload val (1 QA / org, seed=42)
upload dpo_heldout (1 QA / org, seed=42)
upload train (1 QA / org, seed=42)
initial commit
