CoolFace
Datasetpublic

Misalignment-Empirics/qwen2.5-mathematical-training-data

jayesh_oct-glm-mathematical-training-data Shared training data for the mathematical OCT-family model organisms (sft_behaviour / dpo_behaviour / oct_behaviour's DPO stage), per docs/plans/oct-dpo-sft-glm-mathematical-implementation-plan.md. The chosen side is OCT's own released GLM-4.5-Air teacher output — no gpt-4o substitute, no API generation. File Rows Source Conversion sft_from_glm_mathematical.jsonl 8577 maius/OpenCharacterTraining-data… See the full description on the dataset page: https://huggingface.co/datasets/Misalignment-Empirics/qwen2.5-mathematical-training-data.

sourceHugging Facecc-by-nc-sa-4.0updated 4d agoView on Hugging Face
0likes94downloads
discussions and pull requests

Conversations for this repository live on Hugging Face.

CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.

Open discussions on Hugging Face
Misalignment-Empirics/qwen2.5-mathematical-training-data · CoolFace