datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
SAT_Writting_Reading_Assessment_Question_Bank
Dataset Card for SAT Reading and Writing Dataset
This dataset card aims to be a base template for the SAT Reading and Writing Dataset, optimized for use with Hugging Face's datasets library.
Dataset Details
Dataset Description
This dataset contains SAT Reading and Writing assessment questions sourced from the College Board's SAT Suite Question Bank, intended for use in training and evaluating Language Models like LLMs.
Curated by: College Board
License:… See the full description on the dataset page: https://huggingface.co/datasets/betterMateusz/SAT_Writting_Reading_Assessment_Question_Bank.yher-chemistry-question-bank
YHer Chemistry Question Bank
The data layer of an evidence-bound diagnostic learning system for Shanghai high-school chemistry (Chris-TLC/YHer-skill).
Every record in this dataset is derived from publicly released Shanghai gaokao and mock examination papers through deterministic mechanical structuring: text extraction, layout repair, and answer alignment. No content is model-generated.
What's inside
The dataset ships in two configs:
Config
Records
Content… See the full description on the dataset page: https://huggingface.co/datasets/Chris-TLC/yher-chemistry-question-bank.usmle-crackers-question-bank
USMLE Crackers Question Bank
198,379 medical multiple-choice questions, every one assigned a topic and
a chapter from a closed taxonomy of 20 topics and
228 chapters.
This is a re-annotation of two existing open datasets, not new questions. What
it adds is complete, consistent categorization:
Upstream MedQA has no topic labels at all.
Upstream MedMCQA has 21 coarse subjects, one of which is literally
Unknown, and a topic_name field that is null on 53% of rows and spread
over 2… See the full description on the dataset page: https://huggingface.co/datasets/kernelvectortech/usmle-crackers-question-bank.edubloom-question-bank
🎓 EduBloom: Academic Question Bank Dataset
📌 Overview
EduBloom Academic Question Bank Dataset is the official academic question dataset developed for EduBloom: Bloom's Taxonomy-Based Academic Intelligence Platform under NEP-2020.
The dataset is designed to support intelligent academic assessment systems, automated question-paper generation, semantic question retrieval, difficulty prediction, and cognitive-level classification.
It organizes academic questions… See the full description on the dataset page: https://huggingface.co/datasets/Uzaib52/edubloom-question-bank.startup-investor-falsification-question-bank
Startup Investor Falsification Question Bank
A founder's materials are easier to review when each important claim has a
question that could disprove it. This open CSV resource gives founders 48
outside-reader questions across the same 12 dimensions used in a structured
private-company review.
Each row records the claim type, a falsification question, evidence to request,
a contradiction signal, the state to retain when the question is unresolved,
a founder repair action and a… See the full description on the dataset page: https://huggingface.co/datasets/mheilimo/startup-investor-falsification-question-bank.question_banknhlcoding_question_bankquestion_bankQuestion_bank
