CoolFace
Datasetpublic

HimanshuPathak/glm47-synth-v1-dataset

GLM47 Synth V1 Dataset This repository contains 260 verified Aider-format supervised fine-tuning rows. The dataset has ten synthetic variants for each of 26 C++ task families and was structured for an exact 100-epoch memorization experiment. Dataset structure The dataset has one configuration (default) and one split (train): File Rows Format sft/train.jsonl 260 UTF-8 JSON Lines Every row contains these top-level fields: label: unique row label… See the full description on the dataset page: https://huggingface.co/datasets/HimanshuPathak/glm47-synth-v1-dataset.

sourceHugging Faceupdated 11d agoView on Hugging Face
0likes45downloads
discussions and pull requests

Conversations for this repository live on Hugging Face.

CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.

Open discussions on Hugging Face
HimanshuPathak/glm47-synth-v1-dataset · CoolFace