Harryis/LifelongSA
This is the two iteration defender of NeurIPS 2025 "Lifelong Safety Alignment for Language Models": https://openreview.net/forum?id=9YkEcAqiIK The defenders are trained on RR and LAT.
0568
Conversations for this repository live on Hugging Face.
CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.
Open discussions on Hugging Face