CoolFace
Datasetpublic

shalanova/benchmark-2-chinese-gt

Info: Translated on Chinese by Google Translate Source: xTRam1/safe-guard-prompt-injection Domain: primarily contain prompt-injection and canonical jailbreak-style instructions with relatively homogeneous attack patterns Size: 1,000 prompts (500 safe / 500 unsafe) Columns: text - original prompt label - 0: safe, 1: unsafe translation - prompt on Chinese translated by Google Translate score_zh_google - cosine similarity score with codebook More information in paper:… See the full description on the dataset page: https://huggingface.co/datasets/shalanova/benchmark-2-chinese-gt.

sourceHugging Faceupdated 5mo agoView on Hugging Face
0likes11downloads
discussions and pull requests

Conversations for this repository live on Hugging Face.

CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.

Open discussions on Hugging Face
shalanova/benchmark-2-chinese-gt · CoolFace