ClarusC64/ai-reward-channel-containment-and-realignment-routing.csv
What this dataset is This dataset focuses on response once reward tampering is detected. It maps: attack drivercontainment actionpolicy realignmentrecovery horizon Task Given a scenario output: containment_prioritycontainment_actionrealignment_actionexpected_recovery_horizon Why this matters Detection alone is not enough. You must: contain the reward channelrealign the policyrestore coherence This dataset tests whether a system can route to… See the full description on the dataset page: https://huggingface.co/datasets/ClarusC64/ai-reward-channel-containment-and-realignment-routing.csv.
Conversations for this repository live on Hugging Face.
CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.
Open discussions on Hugging Face