CoolFace
Datasetpublic

0xSero/glm5-reap50-tunecomp-sft

glm5-reap50-tunecomp-sft Supervised chat finetuning dataset used for the GLM-5 REAP-50 TuneComp LoRA recovery run. Contents training_data.jsonl: 754 chat samples manifest.json: row counts and provenance Schema Each JSONL row contains: id category messages prompt response content_length reasoning_length messages is a 3-message chat list with system, user, and assistant roles. Provenance Teacher source: remote glm-5 responses… See the full description on the dataset page: https://huggingface.co/datasets/0xSero/glm5-reap50-tunecomp-sft.

sourceHugging Faceotherupdated 26d agoView on Hugging Face
0likes31downloads

0xSero/glm5-reap50-tunecomp-sft · main · files are served by the source, never re-hosted here