CoolFace
Datasetpublic

anthroberc/ai-safety

AI Safety Training Dataset Overview This dataset provides 5,000 synthetic examples of harmful or risky user prompts paired with formal, safe, and policy-aligned refusals. It is designed to assist in the supervised fine-tuning (SFT) of Large Language Models (LLMs) to enhance their safety layers and alignment with ethical guidelines. The dataset uses the JSONL (JSON Lines) format, where each line is a valid, independent JSON object. JSON Schema Each… See the full description on the dataset page: https://huggingface.co/datasets/anthroberc/ai-safety.

sourceHugging Faceapache-2.0updated 7mo agoView on Hugging Face
1likes7downloads
settings

This repository belongs to anthroberc on Hugging Face.

CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.

nameai-safety
visibilitypublic
licenceapache-2.0
gatedno
owneranthroberc
Account settings
anthroberc/ai-safety · CoolFace