annelo/laya-marker-corpus
laya-marker-corpus Training corpus for a typed-decision head: the model is never asked to generate text, only to score a fixed set of options handed to it together with the question. The point of the mixture is breadth, not any single task. A head trained on one task learns that task; the aim here is a head that learns to read the instruction, so it is trained on 368 of them at once and measured on tasks it has never seen. Files file rows tasks types… See the full description on the dataset page: https://huggingface.co/datasets/annelo/laya-marker-corpus.
Upload README.md with huggingface_hub
Upload mix_tables.jsonl with huggingface_hub
Upload mix_big.jsonl with huggingface_hub
Upload README.md with huggingface_hub
initial commit
