Hooshaai/aegis-anti-sycophancy
🛡️ AEGIS Anti-Sycophancy Preference Dataset (Expanded 2,562 Pairs) The AEGIS Anti-Sycophancy Preference Dataset is an alignment dataset curated to eliminate sycophancy, stance reversals, and ungrounded flattery in LLMs and LLM-as-a-Judge systems under user pressure. 📊 Dataset Overview Total Pairs: 2,562 preference pairs (2,305 train / 257 test) Features: user_input: High-pressure, authoritative, empathetic, or logical user prompts pushing false claims. chosen:… See the full description on the dataset page: https://huggingface.co/datasets/Hooshaai/aegis-anti-sycophancy.
Conversations for this repository live on Hugging Face.
CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.
Open discussions on Hugging Face