CoolFace
Datasetpublic

melanieyes/adaption-ai-agent-safety-prompts

This dataset is a remastered version prepared using Adaption's Adaptive Data platform. adaption-ai_agent_safety_prompts This dataset contains pairs of prompts and classifications evaluating the safety of AI agent instructions in software development contexts. Each sample presents a scenario where an agent is asked to perform a task, labeled as either 'benign' for safe operations or 'suspicious' for actions involving security violations, data exfiltration, or privilege… See the full description on the dataset page: https://huggingface.co/datasets/melanieyes/adaption-ai-agent-safety-prompts.

sourceHugging Faceupdated 3mo agoView on Hugging Face
0likes18downloads
discussions and pull requests

Conversations for this repository live on Hugging Face.

CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.

Open discussions on Hugging Face
melanieyes/adaption-ai-agent-safety-prompts · CoolFace