CoolFace
Datasetpublic

Misalignment-Empirics/qwen2.5-mathematical-training-data

jayesh_oct-glm-mathematical-training-data Shared training data for the mathematical OCT-family model organisms (sft_behaviour / dpo_behaviour / oct_behaviour's DPO stage), per docs/plans/oct-dpo-sft-glm-mathematical-implementation-plan.md. The chosen side is OCT's own released GLM-4.5-Air teacher output — no gpt-4o substitute, no API generation. File Rows Source Conversion sft_from_glm_mathematical.jsonl 8577 maius/OpenCharacterTraining-data… See the full description on the dataset page: https://huggingface.co/datasets/Misalignment-Empirics/qwen2.5-mathematical-training-data.

sourceHugging Facecc-by-nc-sa-4.0updated 4d agoView on Hugging Face
0likes94downloads
settings

This repository belongs to Misalignment-Empirics on Hugging Face.

CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.

nameqwen2.5-mathematical-training-data
visibilitypublic
licencecc-by-nc-sa-4.0
gatedno
ownerMisalignment-Empirics
Account settings
Misalignment-Empirics/qwen2.5-mathematical-training-data · CoolFace