BSC-LT/ALIA-2606-DPO-helpfulness
Dataset Card for BSC Multilingual Synthetic Helpfulness Preferences Dataset Summary This dataset consists of synthetic helpfulness preference data generated to align language models across five languages: Catalan, Spanish, English, Basque, and Galician. Building on the PKU-SafeRLHF and Tulu 3/Ultrafeedback methodologies for creating preference data, this dataset leverages an LLM-as-a-judge approach to automatically score and pair model responses to a massive pool… See the full description on the dataset page: https://huggingface.co/datasets/BSC-LT/ALIA-2606-DPO-helpfulness.
Metadata
License CC-BY 4.0
Corrected data and json
Replace helpfulness preferences with corrected dataset
Update. Missing json and possible error
Upload ALIA-2606-DPO-helpfulness
initial commit
