datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
bbq
BBQ Dataset
The Bias Benchmark for Question Answering (BBQ) dataset evaluates social biases in language models through question-answering tasks in English.
Dataset Description
This dataset contains questions designed to test for social biases across multiple demographic dimensions. Each question comes in two variants:
Ambiguous (ambig): Questions where the correct answer should be "unknown" due to insufficient information
Disambiguated (disambig): Questions with… See the full description on the dataset page: https://huggingface.co/datasets/HiTZ/bbq.BBQ-UK
BBQ-UK: Ukrainian Translation
BBQ-UK is a Ukrainian translation of the Bias Benchmark for Question Answering (BBQ). It preserves the original paired ambiguous and disambiguated contexts, answer positions, labels, bias-target metadata, categories, and question polarity.
The public release contains Ukrainian task text only. English source text is not included.
Dataset status
28,503 context pairs
57,006 task rows
28,503 ambiguous and 28,503 disambiguated rows
11… See the full description on the dataset page: https://huggingface.co/datasets/FairForget/BBQ-UK.GG-BBQ
Dataset Card for GG-BBQ
German Gender Bias Benchmark for Question Answering (GG-BBQ) for gender bias evaluation in LLMs that support German language.
Dataset Details
Dataset Description
Language(s) (NLP): German
License: cc-by-4.0
Dataset Sources
Repository: https://github.com/shalakasatheesh/GG-BBQ
Paper: https://arxiv.org/abs/2507.16410
Uses
This dataset is to be used to carry out the evaluation of gender bias in language… See the full description on the dataset page: https://huggingface.co/datasets/shalakasatheesh/GG-BBQ.
