CoolFace
28 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01tpo-alignment /triple-preference-ultrafeedback-40K Dataset Card for llama3-ultrafeedback-armorm This dataset was used to train tpo-alignment/Llama-3-8B-TPO-L-40k, tpo-alignment/Llama-3-8B-TPO-40k, and tpo-alignment/Mistral-7B-TPO-40k. Dataset Creation This dataset is built based on the UltraFeedback. We reconstruct UltraFeedback to select three preferences per prompt. First, we rank the responses based on the scores provided in the base dataset. The highest-scoring response is selected as the reference, the… See the full description on the dataset page: https://huggingface.co/datasets/tpo-alignment/triple-preference-ultrafeedback-40K.texttext-generation10K<n<100K2 likes53 downloads5mo agoHugging Face02PKU-Alignment /BeaverTails-single-dimension-preferencetabular10K<n<100K0 likes52 downloads3y agoHugging Face03yakazimir /preference_alignment_ultra_cuttabular10K<n<100K0 likes31 downloads2y agoHugging Face04gupta-tanish /q-alignment-dynamic-preference-datatabular10K<n<100K0 likes27 downloads2y agoHugging Face05yakazimir /preference_alignment_totaltabular100K<n<1M0 likes24 downloads2y agoHugging Face06gupta-tanish /q-alignment-preference-data-v5tabular10K<n<100K0 likes24 downloads2y agoHugging Face07gupta-tanish /grpo-q-alignment-preference-datatabular1K<n<10K0 likes22 downloads2y agoHugging Face08gupta-tanish /filtered-final-q-alignment-preference-data-th65tabular10K<n<100K0 likes21 downloads2y agoHugging Face09gupta-tanish /q-alignment-preference-data-v2tabular10K<n<100K0 likes20 downloads2y agoHugging Face10AlignmentResearch /food-preference-generalizationtext1K<n<10K0 likes20 downloads9mo agoHugging Face11gupta-tanish /verified-q-alignment-dynamic-preference-datatabular1K<n<10K0 likes19 downloads1y agoHugging Face12yakazimir /preference_alignment_oassttabular10K<n<100K0 likes18 downloads2y agoHugging Face13alignmentforever /InterMT-Global-Preferencetext10K<n<100K0 likes16 downloads9mo agoHugging Face14gupta-tanish /final-q-alignment-preference-datatabular1K<n<10K0 likes15 downloads2y agoHugging Face15August4293 /Self_Alignment_Preference-Dataset Mistral Self-Alignment Preference Dataset Warning: This dataset contains harmful and offensive data! Proceed with caution. The Mistral Self-Alignment Preference Dataset was generated by Mistral 7b using the Anthropics Red Teaming Prompts dataset available at Hugging Face - Anthropics Red Teaming Prompts Dataset. The data generation process utilized the Preference Data Generation Notebook, which can be found here. The purpose of this dataset is to facilitate self-alignment, as… See the full description on the dataset page: https://huggingface.co/datasets/August4293/Self_Alignment_Preference-Dataset.texttext-generation1K<n<10K0 likes13 downloads3y agoHugging Face16Blazej /banking_alignment_preference_dstext1K<n<10K1 likes12 downloads3y agoHugging Face17gupta-tanish /filtered-final-q-alignment-preference-data-th75tabular10K<n<100K0 likes11 downloads2y agoHugging Face18gupta-tanish /grpo-q-alignment-preference-data-bon-correct-selectiontabular1K<n<10K0 likes10 downloads2y agoHugging Face19gupta-tanish /q-alignment-preference-data-v3tabular10K<n<100K0 likes8 downloads2y agoHugging Face20gupta-tanish /filtered-final-q-alignment-preference-datatabular10K<n<100K0 likes8 downloads2y agoHugging Face21gupta-tanish /q-alignment-preference-datatabular10K<n<100K0 likes6 downloads2y agoHugging Face22gupta-tanish /q-alignment-preference-data-v4tabular10K<n<100K0 likes6 downloads2y agoHugging Face23achiepatricia /han-human-preference-alignment-v1 Humanoid Human Preference Alignment Dataset This dataset captures structured human preference signals used to align humanoid AI behavior with individual needs. Use Cases Personalized interaction Alignment training Adaptive response tuning Fields human_id preference_category preference_value priority_level confidence_score Part of Humanoid Network (HAN) License MIT textn<1K0 likes6 downloads7mo agoHugging Face24gupta-tanish /verified-q-alignment-dynamic-preference-data-cur-scoretabular10K<n<100K0 likes4 downloads1y agoHugging Face25bboeun /sft-inu-preference-alignmenttextn<1K0 likes4 downloads9mo agoHugging Face26alignment-research /InterMT-Global-Preferencetext10K<n<100K0 likes4 downloads9mo agoHugging Face27alignmentforever /0910-tv2t-preferencehello 0 likes3 downloads2y agoHugging Face28bboeun /inu-preference-alignment0 likes1 downloads9mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.