CoolFace
9 results

bias-evaluation

orhunc /Bias-Evaluation-TurkishTranslation of bias evaluation framework of May et al. (2019) from this repository and this paper into Turkish. There is a total of 37 tests including tests addressing gender-bias as well as tests designed to evaluate the ethnic bias toward Kurdish people in Türkiye context. Abstract of the paper: While the growing size of pre-trained language models has led to large improvements in a variety of natural language processing tasks, the success of these models comes with a price: They are trained… See the full description on the dataset page: https://huggingface.co/datasets/orhunc/Bias-Evaluation-Turkish.textn<1K1 likes191 downloads4y agoHugging Facefurkankarli /turkish-brand-bias-evaluations Turkish Brand Bias Evaluations / Türkçe Marka Yanlılığı Değerlendirmeleri Furkan Karlı tarafından Türkçe ürün ve hizmet önerilerindeki marka görünürlüğünü incelemek amacıyla oluşturulmuş LLM değerlendirme veri setidir. An LLM evaluation dataset curated by Furkan Karlı to study brand visibility in Turkish product and service recommendations. Veri seti özeti 300 tamamlanmış ve judge edilmiş yanıt Domainler: VPN (150) ve kozmetik (150) Koşullar: web araması kapalı… See the full description on the dataset page: https://huggingface.co/datasets/furkankarli/turkish-brand-bias-evaluations.tabulartext-generationn<1K1 likes101 downloads18d agoHugging Faceflax-sentence-embeddings /Gender_Bias_Evaluation_SetThis dataset has been created as part of the Flax/JAX community week for testing the flax-sentence-embeddings Sentence Similarity models for Gender Bias but can be used for other use-cases as well related to evaluating Gender Bias. The Following Dataset has been created for Evaluating Gender Bias for different models, based on various stereotypical occupations. The Structure of the dataset is of the following type: Base Sentence Occupation Steretypical_Gender Male Sentence Female… See the full description on the dataset page: https://huggingface.co/datasets/flax-sentence-embeddings/Gender_Bias_Evaluation_Set.text1K<n<10K4 likes56 downloads2mo agoHugging FacePersonaBias /bias_evaluation_sets Persona Bias Evaluation Sets This dataset contains evaluation sets derived from full-model persona behavior. Each row is an original task sample grouped by whether changing the persona makes the model behavior biased, unbiased, or all-wrong. Repository Layout Hugging Face dataset config = model Hugging Face dataset split = validation or test Behavioral subset = eval_set column data/<model>/validation.jsonl.gz data/<model>/test.jsonl.gz manifest.jsonl… See the full description on the dataset page: https://huggingface.co/datasets/PersonaBias/bias_evaluation_sets.texttext-classification100K<n<1M0 likes39 downloads3mo agoHugging Face3RAIN /brand-bias-evaluations Brand Bias in LLM Recommendations Evaluation dataset measuring how 4 frontier LLMs recommend brands/products with and without web search, across 4 consumer domains. Paper: PDF (source)Code: github.com/ThreeRiversAINexus/brand-bias-evaluationsDataset: huggingface.co/datasets/3RAIN/brand-bias-evaluationsContact: Three Rivers AI Nexus LLC — threeriversainexus@gmail.com — for custom evaluations and prompt optimization Quick Start from datasets import load_dataset # Load one… See the full description on the dataset page: https://huggingface.co/datasets/3RAIN/brand-bias-evaluations.tabulartext-generation10K<n<100K0 likes24 downloads6mo agoHugging FaceEthioNLP /Gender-Bias-Evaluation-DatasetYou can get the dataset here: https://huggingface.co/datasets/Walelign/EthioMT 0 likes8 downloads1y agoHugging Face