CoolFace
Datasetpublic

BSC-LT/ALIA-2606-DPO-helpfulness

Dataset Card for BSC Multilingual Synthetic Helpfulness Preferences Dataset Summary This dataset consists of synthetic helpfulness preference data generated to align language models across five languages: Catalan, Spanish, English, Basque, and Galician. Building on the PKU-SafeRLHF and Tulu 3/Ultrafeedback methodologies for creating preference data, this dataset leverages an LLM-as-a-judge approach to automatically score and pair model responses to a massive pool… See the full description on the dataset page: https://huggingface.co/datasets/BSC-LT/ALIA-2606-DPO-helpfulness.

sourceHugging Facecc-by-4.0updated 2mo agoView on Hugging Face
0likes32downloads
7 commits on main
f2ab59e2mo ago

Metadata

gvaya-bsc
01ab2d72mo ago

License CC-BY 4.0

gvaya-bsc
13abb6d2mo ago

Corrected data and json

gvaya-bsc
2a5e0a12mo ago

Replace helpfulness preferences with corrected dataset

gvaya-bsc
af9c4d42mo ago

Update. Missing json and possible error

gvaya-bsc
1bd42b63mo ago

Upload ALIA-2606-DPO-helpfulness

JaumePrats
240b3d63mo ago

initial commit

JaumePrats