birgermoell/oellm-eu-defect-repair-sft-v1
oellm-eu-defect-repair-sft-v1 Monolingual SFT repair data for European-language generation defects observed after Qwen 2B/4B/9B post-training. This dataset is designed to repair degeneration, repetition loops, short answers, morphology damage, and Bulgarian/Russian language leakage. Strict SFT Schema Each row in data/*.jsonl uses exactly: {"messages":[{"role":"user","content":"..."},{"role":"assistant","content":"..."}],"lang":"is"} Provenance, source URL… See the full description on the dataset page: https://huggingface.co/datasets/birgermoell/oellm-eu-defect-repair-sft-v1.
This repository belongs to birgermoell on Hugging Face.
CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.
