datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
grumpy-chef-dpo
Grumpy Chef Dataset
A preference dataset of 299 examples for fine-tuning LLMs with DPO (Direct Preference Optimization). Each example contains a cooking-related question paired with two responses: a chosen response written in the voice of a grumpy, opinionated Italian chef, and a rejected generic/neutral response. Designed to teach a model a strong culinary persona through preference alignment.
{
"prompt": "Can I rinse pasta after cooking?",
"chosen": "Rinse it? RINSE IT?! No. You… See the full description on the dataset page: https://huggingface.co/datasets/benitomartin/grumpy-chef-dpo.fairness_chef_google_flan_t5_xxl_mode_T_SPECIFIC_A_ns_4800
Dataset Card for "fairness_chef_google_flan_t5_xxl_mode_T_SPECIFIC_A_ns_4800"
More Information needed
chefhatgamecardhumanoid-chef-datahumanoid-chef-datafairness_chef_google_flan_t5_xl_mode_T_SPECIFIC_A_ns_4800
Dataset Card for "fairness_chef_google_flan_t5_xl_mode_T_SPECIFIC_A_ns_4800"
More Information needed
chef-dataset
Dataset Card for "chef-dataset"
More Information needed
