sentientfutures/animal-welfare-training-claude
Synthetic training data that teaches a model to reason carefully about the welfare of animals and other sentient beings. Why Research on alignment midtraining finds that teaching a model the reasons behind aligned behavior matters as much as the behavior itself. Two techniques from Teaching Claude Why proved especially effective: Synthetic document finetuning on pretraining-style documents from a world where the target model is already aligned across a wide… See the full description on the dataset page: https://huggingface.co/datasets/sentientfutures/animal-welfare-training-claude.
This repository belongs to sentientfutures on Hugging Face.
CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.
