CoolFace
Datasetpublic

FreedomIntelligence/OnePO-Medical-20K

OnePO-Medical-20K 📄 Paper | 💻 GitHub ⚡ Introduction OnePO-Medical-20K is the medical RL dataset released with OnePO, containing 20,338 medical tasks across multiple languages. One stage, no preceding SFT. OnePO adapts pretrained models to medicine through a single reinforcement-learning stage. Two complementary task types. Multiple-choice questions provide verifiable answers. Open-ended conversations provide scoring rubrics. Teacher guidance included. Each task includes a… See the full description on the dataset page: https://huggingface.co/datasets/FreedomIntelligence/OnePO-Medical-20K.

sourceHugging Faceupdated 14h agoView on Hugging Face
5likes102downloads
discussions and pull requests

Conversations for this repository live on Hugging Face.

CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.

Open discussions on Hugging Face
FreedomIntelligence/OnePO-Medical-20K · CoolFace