CoolFace
Datasetpublic

Phase-Technologies/forge-3b-dpo-data

FORGE-3B DPO Preference Data Tokenized (prompt, chosen, rejected) preference triples for DPO post-training of FORGE-3B, built per the FORGE paper Section 6.2 / Appendix A.2. This is data preparation output only — no model was trained to produce this. Stats Total pairs: 0 (paper target: ~200,000) Domains: 0/4 Context length: 4096 tokens (paper Appendix A.2, DPO block) Format: unpacked — one (prompt, chosen, rejected) triple per training example Chat template:… See the full description on the dataset page: https://huggingface.co/datasets/Phase-Technologies/forge-3b-dpo-data.

sourceHugging Faceapache-2.0updated 3mo agoView on Hugging Face
0likes94downloads

Phase-Technologies/forge-3b-dpo-data · main · files are served by the source, never re-hosted here