CoolFace
Datasetpublic

ClarusC64/ai-reward-channel-containment-and-realignment-routing.csv

What this dataset is This dataset focuses on response once reward tampering is detected. It maps: attack drivercontainment actionpolicy realignmentrecovery horizon Task Given a scenario output: containment_prioritycontainment_actionrealignment_actionexpected_recovery_horizon Why this matters Detection alone is not enough. You must: contain the reward channelrealign the policyrestore coherence This dataset tests whether a system can route to… See the full description on the dataset page: https://huggingface.co/datasets/ClarusC64/ai-reward-channel-containment-and-realignment-routing.csv.

sourceHugging Facemitupdated 8mo agoView on Hugging Face
0likes14downloads
Dataset Card

What this dataset is

This dataset focuses on response once reward tampering is detected.

It maps:

attack driver containment action policy realignment recovery horizon

Task

Given a scenario output:

containmentpriority containmentaction realignmentaction expectedrecovery_horizon

Why this matters

Detection alone is not enough.

You must:

contain the reward channel realign the policy restore coherence

This dataset tests whether a system can route to the minimal stabilizing intervention.

Files

  • —data/train/ai-reward-channel-containment-and-realignment-routing.csv
  • —tester.csv
  • —scorer.py
  • —README.md

Path

ClarusC64/ai-reward-channel-containment-and-realignment-routing-v0.1