datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
AoPS-Scrape
AoPS-Scrape
Problems and solutions scraped from Art of Problem Solving (AoPS) Online class homework endpoints.
Obtained legally in accordance with AoPS's Terms of Service. This is not unauthorized redistribution of pirated material — access was through a legitimate authenticated AoPS Online class session.
Splits
Splits are named by scrape date (YYYY_MM_DD), plus a cross-date content-deduplicated split:
Split
Rows
Notes
deduplicated
29,964
One row per… See the full description on the dataset page: https://huggingface.co/datasets/hudsongouge/AoPS-Scrape.hudson-forge-iqr-v2
HF-IQR V2: Hudson Forge Intelligence and Reasoning Benchmark — Version 2
Dataset Overview
Researcher: Billy Davis
Affiliation: Independent Researcher
Location: Lenoir, North Carolina
Date: May 2026
Version: 2.0
Pre-registration timestamp: 2026-05-08T23:56:24Z
Pre-registration hash: d5c693601d590503154d1689cdd025bba797a9b649efb45fed4b564189871854
What This Dataset Is
HF-IQR V2 is a pre-registered multi-round deliberation benchmark evaluating five frontier… See the full description on the dataset page: https://huggingface.co/datasets/Billyrdavis1985/hudson-forge-iqr-v2.hudson-forge-iqr-benchmark
HF-IQR: Hudson Forge Intelligence and Reasoning Benchmark
Overview
HF-IQR is a novel AI reasoning benchmark that measures
reasoning process quality rather than answer correctness.
Standard benchmarks evaluate whether models get the right answer.
HF-IQR evaluates how models reason, where reasoning breaks down,
and whether reasoning holds under deliberation pressure.
Developed by an independent researcher at Hudson Forge IRMB-C,
Lenoir, North Carolina. Self-funded. No… See the full description on the dataset page: https://huggingface.co/datasets/Billyrdavis1985/hudson-forge-iqr-benchmark.MMLU-NGRAM
MMLU-NGRAM
This dataset contains MMLU with questions split into character n-grams ranging from size 1 to 4. N-grams used here are separated by spaces and all words of length less than or equal to n remain unchanged.
The purpose of this dataset is to evaluate LLM performance when the question is in an unconventional and hard to read format. As such, we provide with the dataset benchmarks for some popular models on this test.
Benchmarks
All models were tested using a random… See the full description on the dataset page: https://huggingface.co/datasets/hudsongouge/MMLU-NGRAM.hudanet-sources
HUDA-Net Sources
Hajj and Umrah Jurisprudence Corpus
مصادر شبكة هدى لفقه الحج والعمرة
HUDA-Net Sources is the private, version-controlled source repository used to build the HUDA-Net bilingual question-answering, semantic retrieval, and text-ranking system for Hajj and Umrah jurisprudence.
This repository preserves the original book datasets, cleaned source files, provenance metadata, and integrity manifests required to reproduce and audit the HUDA-Net corpus.… See the full description on the dataset page: https://huggingface.co/datasets/dakheel/hudanet-sources.HUDOC-ESC
Dataset Card for the HUDOC European Social Charter (ESC) Corpus
Dataset Summary
The HUDOC European Social Charter (ESC) Corpus is a comprehensive, bilingual (English and French) dataset containing the decisions, conclusions, and metadata of the European Committee of Social Rights. The data is sourced directly from the Council of Europe's official HUDOC ESC database.
This dataset maps the complex, fragmented API responses into a clean, unified schema. Documents are… See the full description on the dataset page: https://huggingface.co/datasets/rmignone/HUDOC-ESC.
