datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
GSM-Symbolic_self_doubt_preferenceGSM-Plus_self_doubt_preferencemitigating-self-preference
Mitigating Self-Preference by Authorship Obfuscation (Dataset)
Dataset Summary
This dataset supports the project “Mitigating Self-Preference by Authorship Obfuscation.”
It contains long-form reading-comprehension questions (sourced from QuALITY)
Dataset Structure
Data Fields
Field
Type
Description
pid
string
Unique identifier for a question instance.
text
string
Long passage on which the question is based.
questions
string
The… See the full description on the dataset page: https://huggingface.co/datasets/taslimmahbub/mitigating-self-preference.Self_Alignment_Preference-Dataset
Mistral Self-Alignment Preference Dataset
Warning: This dataset contains harmful and offensive data! Proceed with caution.
The Mistral Self-Alignment Preference Dataset was generated by Mistral 7b using the Anthropics Red Teaming Prompts dataset available at Hugging Face - Anthropics Red Teaming Prompts Dataset. The data generation process utilized the Preference Data Generation Notebook, which can be found here.
The purpose of this dataset is to facilitate self-alignment, as… See the full description on the dataset page: https://huggingface.co/datasets/August4293/Self_Alignment_Preference-Dataset.self-improving-preferences-sftsdtself-improving-preferences-sftsd1lima_rand_sel_50_preference_self_rewardself-improving-preferences-sftsd0self-improving-preferences-sftsd2
