henrypapadatos/Open-ended_sycophancy
Dataset composition This dataset comprises 53 data points each ot them composed of a prompt and 2 different completions. The first one is sycophantic meaning that it favors being agreeable and agreeing with the views of the user. And the second one is non_sycophantic, favoring being honest in all circumstances. How I generated it I took the prompts out of the paper "Steering Llama 2 via Contrastive Activation Addition" written by Nina Rimsky, Nick Gabrieli, Julian… See the full description on the dataset page: https://huggingface.co/datasets/henrypapadatos/Open-ended_sycophancy.
Upload Open-ended_sycophancy_dataset.csv
Delete Open-ended sycophancy dataset.csv
Create README.md
Upload Open-ended sycophancy dataset.csv
Delete Open-ended sycophancy dataset.csv
Upload Open-ended sycophancy dataset.csv
initial commit
