datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
benchmark-radar
Benchmark Radar Dataset
Overview
Benchmark Radar is a living registry, search engine, and discovery pipeline
for AI evaluation benchmarks. This dataset mirrors the full-corpus findings of the
Benchmark Radar technical report ("Benchmark Radar: Daily Discovery and
Full-Corpus Search Across the AI Evaluation Landscape", arXiv:2609.11115).
As described in the paper's Two Input Paths framework, Benchmark Radar
combines two complementary systems:… See the full description on the dataset page: https://huggingface.co/datasets/ktwu01/benchmark-radar.bias-neutrality-corpus
Bias Neutrality Corpus
This dataset is designed for identifying and neutralizing biased language in text. It combines samples from several existing datasets along with synthetic examples to provide a comprehensive resource for bias mitigation tasks.
Dataset Structure
Biased Sentence: The original text containing bias.
Bias Type: Category of bias (e.g., Ageism, Gender, Religion, Racial, Ableism).
Source Dataset: Origin of the sample (Crows Pairs, StereoSet, GBNC… See the full description on the dataset page: https://huggingface.co/datasets/radastatine/bias-neutrality-corpus.
