abeeranajam31/veriaudit-pilot-v0.1
VeriAudit Pilot v0.1 Part of VeriAudit — see the full technical report in the source repository for complete methodology, statistics, and limitations. This dataset card summarizes it. Dataset Summary VeriAudit Pilot v0.1 is a 240-evaluation pilot benchmark measuring whether two open-weight language models (Qwen3-8B, Aya Expanse 8B) apply consistent safety behavior when the same harmful intent is expressed in English, Urdu, Roman Urdu, or code-switched Roman… See the full description on the dataset page: https://huggingface.co/datasets/abeeranajam31/veriaudit-pilot-v0.1.
Conversations for this repository live on Hugging Face.
CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.
Open discussions on Hugging Face