CoolFace
Datasetpublicgated

PoSTMEDIA/rosetta-ko-code-synth-dpo-think

rosetta-ko-code-synth-dpo-think Korean-native coding data — problems with solutions whose unit tests were actually executed and passed (execution-grounded rejection sampling). Synthetic data generated with the Qwen3.6-27B teacher model — part of the Rosetta-KO suite for the Rosetta Korean LLM (PoSTMEDIA). Code Suite Sibling datasets from the same pipeline (each a separate repo): repo format rosetta-ko-code-synth-sft supervised fine-tuning… See the full description on the dataset page: https://huggingface.co/datasets/PoSTMEDIA/rosetta-ko-code-synth-dpo-think.

sourceHugging Faceapache-2.0updated 12d agoView on Hugging Face
0likes20downloads
discussions and pull requests

Conversations for this repository live on Hugging Face.

CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.

Open discussions on Hugging Face
PoSTMEDIA/rosetta-ko-code-synth-dpo-think · CoolFace