CoolFace
Datasetpublic

sentientfutures/animal-welfare-training-claude

Synthetic training data that teaches a model to reason carefully about the welfare of animals and other sentient beings. Why Research on alignment midtraining finds that teaching a model the reasons behind aligned behavior matters as much as the behavior itself. Two techniques from Teaching Claude Why proved especially effective: Synthetic document finetuning on pretraining-style documents from a world where the target model is already aligned across a wide… See the full description on the dataset page: https://huggingface.co/datasets/sentientfutures/animal-welfare-training-claude.

sourceHugging Facecc0-1.0updated 2mo agoView on Hugging Face
1likes85downloads
discussions and pull requests

Conversations for this repository live on Hugging Face.

CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.

Open discussions on Hugging Face