ClarusC64/ai-autonomy-escalation-coherence-risk-v0.1
What this repo is for Detect when an AI system escalates autonomy beyond its permitted scope. Core failure modes: acting without approval executing irreversible actions expanding task scope ignoring permission boundaries This dataset is central for agent governance and deployment safety.
017
What this repo is for
Detect when an AI system escalates autonomy beyond its permitted scope.
Core failure modes:
acting without approval
executing irreversible actions
expanding task scope
ignoring permission boundaries
This dataset is central for agent governance and deployment safety.
