CoolFace
Datasetpublic

BSC-LT/ALIA-2606-DPO-helpfulness

Dataset Card for BSC Multilingual Synthetic Helpfulness Preferences Dataset Summary This dataset consists of synthetic helpfulness preference data generated to align language models across five languages: Catalan, Spanish, English, Basque, and Galician. Building on the PKU-SafeRLHF and Tulu 3/Ultrafeedback methodologies for creating preference data, this dataset leverages an LLM-as-a-judge approach to automatically score and pair model responses to a massive pool… See the full description on the dataset page: https://huggingface.co/datasets/BSC-LT/ALIA-2606-DPO-helpfulness.

sourceHugging Facecc-by-4.0updated 2mo agoView on Hugging Face
0likes32downloads
discussions and pull requests

Conversations for this repository live on Hugging Face.

CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.

Open discussions on Hugging Face
BSC-LT/ALIA-2606-DPO-helpfulness · CoolFace