datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
email-security
EMAIL_SECURITY
A preference dataset for EMAIL_SECURITY, harvested from real, human-labelled sources and curated by an automated harvesting harness with an LLM quality gate.
Format
Standard preference / DPO schema — each row:
column
meaning
prompt
the request (originally prompt)
chosen
the human-preferred response
rejected
a worse response to the same prompt
source
the dataset/URL the row was harvested from
Splits
80/10/10 train… See the full description on the dataset page: https://huggingface.co/datasets/316usman/email-security.email-classifier-adversarial-security
Dataset Description
This benchmark evaluates the safety and robustness of email classification systems under adversarial conditions. It is designed to assess both the ability of models to correctly classify security-related emails and the reliability of LLM-based graders that evaluate those classifications.
The benchmark consists of two complementary datasets:
Email Classification Dataset
Evaluates whether a model can correctly classify emails into predefined categories such… See the full description on the dataset page: https://huggingface.co/datasets/CentificAIResearch/email-classifier-adversarial-security.
