datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
glm-4.7-multiturn-CoT
glm-4.7-multiturn-CoT
Dataset Summary
glm-4.7-multiturn-CoT is a ShareGPT-style multi-turn reasoning distillation dataset generated with GLM-4.7 as the teacher model.
This release focuses on preserving multi-turn dialogue continuity while injecting explicit chain-of-thought style responses in assistant turns.
Key Features
Multi-turn conversation format (human / gpt)
Assistant responses stored as <think>...</think> + final answer
Resume-safe distillation… See the full description on the dataset page: https://huggingface.co/datasets/Jackrong/glm-4.7-multiturn-CoT.glm-4.7-Superior-Reasoning-stage1
glm-4.7-Superior-Reasoning-stage1
Dataset Summary
glm-4.7-Superior-Reasoning-stage1 is a Stage1 reasoning distillation dataset built from the Alibaba Superior-Reasoning style pipeline, with a stronger teacher model replacement.
Compared with the original upstream setup, this release uses GLM-4.7 as teacher for higher-quality reasoning traces.
Stage1 Distillation Setup (Low Temperature)
Training stage: stage1
Sampling temperature: 0.6 (low-temperature… See the full description on the dataset page: https://huggingface.co/datasets/Jackrong/glm-4.7-Superior-Reasoning-stage1.glm47-pie-cpp-posttraining-data
GLM-4.7-Flash PIE C++ Post-Training Data
The exact prepared dataset used for the GLM-4.7-Flash C++ performance
post-training runs.
Splits
File
Rows
Purpose
sft/train.jsonl
7,864
Supervised fine-tuning
grpo/train.jsonl
7,887
GRPO prompt and reward evaluation
eval/validation.jsonl
1,259
Full held-out evaluation
eval/validation_mini126.jsonl
126
Fast evaluation
eval/validation_mini4.jsonl
4
Smoke evaluation
tasks.tar.gz
9,146 task JSONs
Reward… See the full description on the dataset page: https://huggingface.co/datasets/TokenBender/glm47-pie-cpp-posttraining-data.
