CoolFace
Datasetpublic

HimanshuPathak/glm47-synth-v1-dataset

GLM47 Synth V1 Dataset This repository contains 260 verified Aider-format supervised fine-tuning rows. The dataset has ten synthetic variants for each of 26 C++ task families and was structured for an exact 100-epoch memorization experiment. Dataset structure The dataset has one configuration (default) and one split (train): File Rows Format sft/train.jsonl 260 UTF-8 JSON Lines Every row contains these top-level fields: label: unique row label… See the full description on the dataset page: https://huggingface.co/datasets/HimanshuPathak/glm47-synth-v1-dataset.

sourceHugging Faceupdated 11d agoView on Hugging Face
0likes45downloads
settings

This repository belongs to HimanshuPathak on Hugging Face.

CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.

nameglm47-synth-v1-dataset
visibilitypublic
licencenot set
gatedno
ownerHimanshuPathak
Account settings
HimanshuPathak/glm47-synth-v1-dataset · CoolFace