datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
glm-4.6-250xThis is a reasoning dataset created using GLM 4.6. Some of these questions are from reedmayhew and the rest were generated.
The dataset is meant for creating distilled versions of 4.6 by fine-tuning already existing open-source LLMs.
Disclaimer: With such a small dataset any fine-tuned LLM will only be reproducing the answer/thinking style. No knowledge transfer is happening when fine-tuning on this dataset.
GLM4.6-OpenR1Math-SFT
