CoolFace
Datasetpublic

TokenBender/glm47-synth-v1-dataset

Synth v1 Dataset This package contains 260 verified Aider-format SFT rows: ten synthetic variants for each of the 26 source task families. It is intentionally built for an exact 100-epoch memorization experiment. The training file is sft/train.jsonl. Every row uses the same nine-message aider-chat-v1 structure as the successful SFT-v5 package. Tests are not model-visible; the independent verifier replays each final assistant target against its source C++ test suite.… See the full description on the dataset page: https://huggingface.co/datasets/TokenBender/glm47-synth-v1-dataset.

sourceHugging Faceupdated 2mo agoView on Hugging Face
0likes11downloads
discussions and pull requests

Conversations for this repository live on Hugging Face.

CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.

Open discussions on Hugging Face
TokenBender/glm47-synth-v1-dataset · CoolFace