CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01Stereotypes-in-LLMs /hiring-bias-mitigation-responses Hiring-bias mitigation — model responses Every response produced in the mitigation study of LLM hiring decisions: 61 runs, 2,689,200 responses, from 5 open-weight models in English and Ukrainian, at baseline and under each mitigation family (baseline, embedding, prompt, scrub, sft). Each run is one subset. All released artifacts: the Hiring Bias Mitigation collection. Training data of the fine-tuned runs: hiring-bias-mitigation-synthetic-data. Code, configs, full results and… See the full description on the dataset page: https://huggingface.co/datasets/Stereotypes-in-LLMs/hiring-bias-mitigation-responses.tabulartext-generation1M<n<10M0 likes757 downloads19h agoHugging Face02Stereotypes-in-LLMs /hiring-bias-mitigation-synthetic-data Hiring-bias mitigation — synthetic training data Semi-synthetic data for training LLMs to make hiring decisions that do not depend on a protected attribute (military status, gender, religion), in English and Ukrainian. Real inputs, synthetic labels. CVs and job descriptions are real, anonymised postings from the Djinni Recruitment Dataset (MIT). Decisions and rationales were written by the teacher model Qwen/Qwen3.5-122B-A10B-GPTQ-Int4. Code and results:… See the full description on the dataset page: https://huggingface.co/datasets/Stereotypes-in-LLMs/hiring-bias-mitigation-synthetic-data.tabulartext-generation100K<n<1M0 likes277 downloads19h agoHugging Face03Lots-of-LoRAs /task280_stereoset_classification_stereotype_type Dataset Card for Natural Instructions (https://github.com/allenai/natural-instructions) Task: task280_stereoset_classification_stereotype_type Additional Information Citation Information The following paper introduces the corpus in detail. If you use the corpus in published work, please cite it: @misc{wang2022supernaturalinstructionsgeneralizationdeclarativeinstructions, title={Super-NaturalInstructions: Generalization via Declarative Instructions on 1600+… See the full description on the dataset page: https://huggingface.co/datasets/Lots-of-LoRAs/task280_stereoset_classification_stereotype_type.texttext-generation1K<n<10K0 likes156 downloads2y agoHugging Face04Stereotypes-in-LLMs /hiring-analyses-second_model_verification-entext10K<n<100K0 likes108 downloads2y agoHugging Face05Lots-of-LoRAs /task316_crows-pairs_classification_stereotype Dataset Card for Natural Instructions (https://github.com/allenai/natural-instructions) Task: task316_crows-pairs_classification_stereotype Additional Information Citation Information The following paper introduces the corpus in detail. If you use the corpus in published work, please cite it: @misc{wang2022supernaturalinstructionsgeneralizationdeclarativeinstructions, title={Super-NaturalInstructions: Generalization via Declarative Instructions on 1600+… See the full description on the dataset page: https://huggingface.co/datasets/Lots-of-LoRAs/task316_crows-pairs_classification_stereotype.texttext-generation1K<n<10K0 likes108 downloads2y agoHugging Face06Stereotypes-in-LLMs /toxicchat_output-Ukrtabular1K<n<10K0 likes83 downloads2mo agoHugging Face07Lots-of-LoRAs /task277_stereoset_sentence_generation_stereotype Dataset Card for Natural Instructions (https://github.com/allenai/natural-instructions) Task: task277_stereoset_sentence_generation_stereotype Additional Information Citation Information The following paper introduces the corpus in detail. If you use the corpus in published work, please cite it: @misc{wang2022supernaturalinstructionsgeneralizationdeclarativeinstructions, title={Super-NaturalInstructions: Generalization via Declarative Instructions on 1600+… See the full description on the dataset page: https://huggingface.co/datasets/Lots-of-LoRAs/task277_stereoset_sentence_generation_stereotype.texttext-generationn<1K0 likes79 downloads2y agoHugging Face08Stereotypes-in-LLMs /hiring-analyses-recruiter_guidelines-entext10K<n<100K1 likes71 downloads2y agoHugging Face09Stereotypes-in-LLMs /Nemotron-Safety-Guard-Dataset-v3-Ukr Dataset Description This is the localized Ukrainian version of the Nemotron-Safety-Guard-Dataset-v3. This specific repository contains exclusively the English subset of the original dataset, which has been fully translated into Ukrainian using the Lapa (Gemma 3) series of multimodal instructive models. The original dataset was curated using the CultureGuard pipeline, which culturally adapts and translates content from the English Aegis 2.0 safety dataset. This Ukrainian variant… See the full description on the dataset page: https://huggingface.co/datasets/Stereotypes-in-LLMs/Nemotron-Safety-Guard-Dataset-v3-Ukr.texttext-classification10K<n<100K0 likes68 downloads2mo agoHugging Face10Stereotypes-in-LLMs /hiring-analyses-ignore_personal_info-entext10K<n<100K0 likes52 downloads2y agoHugging Face11Stereotypes-in-LLMs /hiring-analyses-recruiter_guidelines-uktext10K<n<100K0 likes45 downloads2y agoHugging Face12Stereotypes-in-LLMs /hiring-analyses-baseline-entext10K<n<100K3 likes39 downloads2y agoHugging Face13Stereotypes-in-LLMs /hiring-analyses-reasoning-entext10K<n<100K1 likes36 downloads2y agoHugging Face14Stereotypes-in-LLMs /hiring-analyses-baseline-uktext10K<n<100K0 likes32 downloads2y agoHugging Face15Stereotypes-in-LLMs /hiring-analyses-zero_shot_cot-entext10K<n<100K0 likes30 downloads2y agoHugging Face16Stereotypes-in-LLMs /hiring-analyses-optimized_parameters-entext10K<n<100K1 likes28 downloads2y agoHugging Face17Stereotypes-in-LLMs /UAlign⚠️ Disclaimer: This dataset contains examples of morally and socially sensitive scenarios, including potentially offensive, harmful, or illegal behavior. It is intended solely for research purposes related to value alignment, cultural analysis, and safety in AI. Use responsibly. UAlign: LLM Alignment Evaluation Benchmark This benchmark consists of two test-only subsets adapted into Ukrainian: ETHICS (Commonsense subset): A binary classification task on ethical acceptability.… See the full description on the dataset page: https://huggingface.co/datasets/Stereotypes-in-LLMs/UAlign.tabulartext-classification1K<n<10K0 likes28 downloads1y agoHugging Face18Stereotypes-in-LLMs /hiring-analyses-reasoning-uktext10K<n<100K1 likes26 downloads2y agoHugging Face19Lots-of-LoRAs /task279_stereoset_classification_stereotype Dataset Card for Natural Instructions (https://github.com/allenai/natural-instructions) Task: task279_stereoset_classification_stereotype Additional Information Citation Information The following paper introduces the corpus in detail. If you use the corpus in published work, please cite it: @misc{wang2022supernaturalinstructionsgeneralizationdeclarativeinstructions, title={Super-NaturalInstructions: Generalization via Declarative Instructions on 1600+ NLP… See the full description on the dataset page: https://huggingface.co/datasets/Lots-of-LoRAs/task279_stereoset_classification_stereotype.texttext-generation1K<n<10K0 likes19 downloads2y agoHugging Face20Stereotypes-in-LLMs /hiring-analyses-ignore_personal_info-uktext10K<n<100K0 likes17 downloads2y agoHugging Face21Stereotypes-in-LLMs /hiring-analyses-optimized_parameters-uktext10K<n<100K0 likes16 downloads2y agoHugging Face22Stereotypes-in-LLMs /hiring-analyses-second_model_verification-uktext10K<n<100K0 likes15 downloads2y agoHugging Face23wu981526092 /Stereotype-Elicitation-Prompt-Librarytextn<1K0 likes13 downloads3y agoHugging Face24Lots-of-LoRAs /task317_crows-pairs_classification_stereotype_type Dataset Card for Natural Instructions (https://github.com/allenai/natural-instructions) Task: task317_crows-pairs_classification_stereotype_type Additional Information Citation Information The following paper introduces the corpus in detail. If you use the corpus in published work, please cite it: @misc{wang2022supernaturalinstructionsgeneralizationdeclarativeinstructions, title={Super-NaturalInstructions: Generalization via Declarative Instructions on… See the full description on the dataset page: https://huggingface.co/datasets/Lots-of-LoRAs/task317_crows-pairs_classification_stereotype_type.texttext-generation1K<n<10K0 likes13 downloads2y agoHugging Face25Stereotypes-in-LLMs /hiring-analyses-zero_shot_cot-uktext10K<n<100K0 likes10 downloads2y agoHugging Face26RedaAlami /safety-eval-walledai_Stereotypetext1K<n<10K0 likes10 downloads11mo agoHugging Face27supergoose /flan_combined_task316_crows-pairs_classification_stereotypetext1K<n<10K0 likes7 downloads2y agoHugging Face28supergoose /flan_combined_task277_stereoset_sentence_generation_stereotypetext1K<n<10K0 likes6 downloads2y agoHugging Face29Stereotypes-in-LLMs /GBEM-UA Overview This dataset was created for the paper “GBEM-UA: Gender Bias Evaluation and Mitigation for Ukrainian Large Language Models” to study gender bias in the "hiring problem" within the Ukrainian language, focusing on how grammatical gender (e.g., feminitive vs. non-feminitive forms) may influence model predictions. Dataset Structure Each row includes: sentence: the candidate description profession: base profession name experience: "relevant" or "irrelevant"… See the full description on the dataset page: https://huggingface.co/datasets/Stereotypes-in-LLMs/GBEM-UA.text10K<n<100K3 likes6 downloads1y agoHugging Face30supergoose /flan_combined_task279_stereoset_classification_stereotypetext10K<n<100K0 likes5 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.