datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
k12-standards-instruction-tasks
K-12 Curriculum Tasks (generated)
2,489 generated instruction/input/output records covering five curriculum tasks:
assessment creation, learning objective generation, misconception detection, standard
explanation, and standards Q&A. Content is predominantly mathematics.
Important: the name is misleading
Despite the name, this dataset contains no school directory data. There are four
columns - task, input, output, metadata - and no staff, principal, or school… See the full description on the dataset page: https://huggingface.co/datasets/robworks-software/k12-standards-instruction-tasks.Government-Auditing-Standards
Government Auditing Standards Corpus
Dataset Description
The Government Auditing Standards Corpus is a processed professional-standards dataset derived from the United States Government Accountability Office publication Government Auditing Standards.
Government Auditing Standards are commonly known as:
The Yellow Book
Generally Accepted Government Auditing Standards
GAGAS
The standards establish requirements and provide application guidance for conducting… See the full description on the dataset page: https://huggingface.co/datasets/leeroy-jankins/Government-Auditing-Standards.Standard-Multimodal-Explanation
Dataset Card for Standard Multimodal Explanation (SME)
This is a dataset for Multimodal Explanation for Visual Question Answering (MEVQA).
Dataset Details
Dataset Description
This dataset contains questions, images, answers, and the multimodal explanations of the underlying reasoning process.
The explanations are in standard English with additional [BOX] for visual grounding.
Language(s) (NLP): English
License: apache-2.0
Modality:
Language… See the full description on the dataset page: https://huggingface.co/datasets/LivXue/Standard-Multimodal-Explanation.Statements-Of-Federal-Financial-Accounting-Concepts-And-Standards
Statements of Federal Financial Accounting Concepts and Standards
Dataset Summary
This dataset contains document-grounded question-and-answer samples based on the Statements of Federal Financial Accounting Concepts and Standards issued within the Federal accounting framework.
The source material establishes the concepts, principles, definitions, recognition criteria, measurement requirements, presentation standards, and disclosure expectations used in Federal… See the full description on the dataset page: https://huggingface.co/datasets/leeroy-jankins/Statements-Of-Federal-Financial-Accounting-Concepts-And-Standards.bootstrapvue-standardSamples in this benchmark were generated by RELAI using the following data source(s):
Data Source Name: BootstrapVue 2.23
Documentation Data Source Link: https://bootstrap-vue.org/docs/
Data Source License: https://github.com/bootstrap-vue/bootstrap-vue/blob/dev/LICENSE
Data Source Authors: BootstrapVue Team
AI Benchmarks by Data Agents © 2025 RELAI.AI · Licensed under CC BY 4.0. Source: https://relai.ai
BCE-Prettybird-Micro-Standard-v0.0.2
🚀 The Future Standard / Geleceğin Standartı
[English]
Beyond Raw Data: The Behavioral Revolution
The AI industry has been obsessed with the volume of data. At Prometech A.Ş., we are shifting the focus to the process of thought. BCE-Prettybird-Micro-Standart is not just a collection of Q&As; it is a blueprint for behavioral reasoning. By integrating Path Mapping and Behavioral DNA into the training loop, we are setting the new industry standard: Small models with elite… See the full description on the dataset page: https://huggingface.co/datasets/pthinc/BCE-Prettybird-Micro-Standard-v0.0.2.k12-mathematics-standards-expanded
K-12 Mathematics Standards, expanded (generated instruction data)
4,965 instruction/input/output records for mathematics, generated around a K-12
standards taxonomy for instruction-tuning and educational-content experiments.
How this was built (read this first)
These are programmatically generated training examples, not curriculum written by
educators and not the text of any official standard. A generator combined standards
metadata - codes, grade levels, domains… See the full description on the dataset page: https://huggingface.co/datasets/robworks-software/k12-mathematics-standards-expanded.k12-science-standards
[!WARNING]
Deprecated - use k12-science-standards-expanded instead.
This dataset is superseded: every instruction in this set also appears there, plus 1,123 more and nine additional metadata columns. Nothing here is unique to it.
It stays online so existing references keep resolving, but it will not be updated.
New work should point at robworks-software/k12-science-standards-expanded.
K-12 Science Standards (generated instruction data)
6,787 instruction/input/output records… See the full description on the dataset page: https://huggingface.co/datasets/robworks-software/k12-science-standards.BCE-Prettybird-Micro-Standard-v0.0.1
🚀 The Future Standard / Geleceğin Standartı
[English]
Beyond Raw Data: The Behavioral Revolution
The AI industry has been obsessed with the volume of data. At Prometech A.Ş., we are shifting the focus to the process of thought. BCE-Prettybird-Micro-Standart is not just a collection of Q&As; it is a blueprint for behavioral reasoning. By integrating Path Mapping and Behavioral DNA into the training loop, we are setting the new industry standard: Small models with elite… See the full description on the dataset page: https://huggingface.co/datasets/pthinc/BCE-Prettybird-Micro-Standard-v0.0.1.k12-ela-standards-expanded
K-12 ELA Standards, expanded (generated instruction data)
12,282 instruction/input/output records for English Language Arts, generated around a K-12
standards taxonomy for instruction-tuning and educational-content experiments.
How this was built (read this first)
These are programmatically generated training examples, not curriculum written by
educators and not the text of any official standard. A generator combined standards
metadata - codes, grade levels… See the full description on the dataset page: https://huggingface.co/datasets/robworks-software/k12-ela-standards-expanded.opentelemetry-standardSamples in this benchmark were generated by RELAI using the following data source(s):
Data Source Name: Homebrew
Documentation Data Source Link: https://opentelemetry.io/docs/
Data Source License: https://github.com/open-telemetry/opentelemetry.io/blob/main/LICENSE
Data Source Authors: OpenTelemetry Authors
AI Benchmarks by Data Agents © 2025 RELAI.AI · Licensed under CC BY 4.0. Source: https://relai.ai
k12-mathematics-standards-aligned
[!WARNING]
Deprecated - use k12-mathematics-standards-expanded instead.
This dataset is superseded: every input in this set also appears there, plus 366 more and two additional metadata columns. Nothing here is unique to it.
It stays online so existing references keep resolving, but it will not be updated.
New work should point at robworks-software/k12-mathematics-standards-expanded.
K-12 Mathematics Standards (generated instruction data)
4,397 instruction/input/output records… See the full description on the dataset page: https://huggingface.co/datasets/robworks-software/k12-mathematics-standards-aligned.k12-social-studies-standards
K-12 Social Studies Standards (generated instruction data)
15,982 instruction/input/output records for social studies (civics, history, geography, economics), generated around a K-12
standards taxonomy for instruction-tuning and educational-content experiments.
How this was built (read this first)
These are programmatically generated training examples, not curriculum written by
educators and not the text of any official standard. A generator combined standards… See the full description on the dataset page: https://huggingface.co/datasets/robworks-software/k12-social-studies-standards.fatwa-mcq-evaluation_standardized
Fatwa MCQ Evaluation Dataset (Standardized)
Standardized multiple-choice question dataset for evaluating Islamic jurisprudence knowledge.
Dataset Description
This dataset contains MCQ versions of Islamic fatwa Q&A pairs, standardized for evaluation purposes.
Dataset Summary
Language: Arabic
Domain: Islamic Finance, Jurisprudence (Fiqh)
Format: Multiple choice questions (4 options)
Task: Islamic jurisprudence knowledge evaluation… See the full description on the dataset page: https://huggingface.co/datasets/SahmBenchmark/fatwa-mcq-evaluation_standardized.BCE-Prettybird-Micro-Standard-v0.0.4
🚀 The Future Standard / Geleceğin Standartı
[English]
Beyond Raw Data: The Behavioral Revolution
The AI industry has been obsessed with the volume of data. At Prometech A.Ş., we are shifting the focus to the process of thought. BCE-Prettybird-Micro-Standart is not just a collection of Q&As; it is a blueprint for behavioral reasoning. By integrating Path Mapping and Behavioral DNA into the training loop, we are setting the new industry standard: Small models with elite… See the full description on the dataset page: https://huggingface.co/datasets/pthinc/BCE-Prettybird-Micro-Standard-v0.0.4.BCE-Prettybird-Large-Standard-v0.0.1
🚀 The Future Standard / Geleceğin Standartı
[English]
Beyond Raw Data: The Behavioral Revolution
The AI industry has been obsessed with the volume of data. At Prometech A.Ş., we are shifting the focus to the process of thought. BCE-Prettybird-Micro-Standart is not just a collection of Q&As; it is a blueprint for behavioral reasoning. By integrating Path Mapping and Behavioral DNA into the training loop, we are setting the new industry standard: Small models with elite… See the full description on the dataset page: https://huggingface.co/datasets/pthinc/BCE-Prettybird-Large-Standard-v0.0.1.k12-ela-standards
[!WARNING]
Deprecated - use k12-ela-standards-expanded instead.
This dataset is superseded: every input in this set also appears there, plus 1,433 more and five additional metadata columns. Nothing here is unique to it.
It stays online so existing references keep resolving, but it will not be updated.
New work should point at robworks-software/k12-ela-standards-expanded.
K-12 ELA Standards (generated instruction data)
6,487 instruction/input/output records for English Language… See the full description on the dataset page: https://huggingface.co/datasets/robworks-software/k12-ela-standards.DOD-Directive-type-Memorandum-22-001-Records-Management-Standards
DoD Records Management Standards for IT Systems and Services
Maintainer: Terry Eppler
Owner: US Federal Government
Dataset Summary
This dataset contains document-grounded question-and-answer records based on Department of Defense Directive-type Memorandum 22-001, “DoD Standards for Records Management Capabilities in Programs Including Information Technology,” dated March 3, 2022, and incorporating Change 2 effective February 22, 2024.
The source establishes… See the full description on the dataset page: https://huggingface.co/datasets/leeroy-jankins/DOD-Directive-type-Memorandum-22-001-Records-Management-Standards.flask-standardSamples in this benchmark were generated by RELAI using the following data source(s):
Data Source Name: Flask
Documentation Data Source Link: https://flask.palletsprojects.com/en/stable/
Data Source License: https://flask.palletsprojects.com/en/stable/license/
Data Source Authors: Pallets Project
AI Benchmarks by Data Agents © 2025 RELAI.AI · Licensed under CC BY 4.0. Source: https://relai.ai
arabic-business-mcq_training_standardized
Arabic Business MCQ Training Dataset (Standardized)
Standardized training dataset for Arabic Business MCQ, formatted to match the evaluation splits.
Dataset Description
This dataset contains the training split of the Arabic Business MCQ dataset, standardized to match the format used in the evaluation splits.
Dataset Summary
Language: Arabic
Domain: Business, Economics, Entrepreneurship, Accounting
Format: Multiple choice questions (standardized)… See the full description on the dataset page: https://huggingface.co/datasets/SahmBenchmark/arabic-business-mcq_training_standardized.BCE-Prettybird-Micro-Standard-v0.0.3
🚀 The Future Standard / Geleceğin Standartı
[English]
Beyond Raw Data: The Behavioral Revolution
The AI industry has been obsessed with the volume of data. At Prometech A.Ş., we are shifting the focus to the process of thought. BCE-Prettybird-Micro-Standart is not just a collection of Q&As; it is a blueprint for behavioral reasoning. By integrating Path Mapping and Behavioral DNA into the training loop, we are setting the new industry standard: Small models with elite… See the full description on the dataset page: https://huggingface.co/datasets/pthinc/BCE-Prettybird-Micro-Standard-v0.0.3.k12-ela-standards
K-12 English Language Arts Standards Dataset
📚 Comprehensive ELA Education Dataset
This dataset provides complete K-12 English Language Arts standards coverage for AI training and educational technology development.
📊 Dataset Summary
The K-12 ELA Standards Dataset contains:
546 educational standards across K-12
6487 AI training samples
Complete coverage of all major ELA domains
🎓 Educational Coverage
📖 Reading Literature: Fiction, poetry… See the full description on the dataset page: https://huggingface.co/datasets/ijanhq8809/k12-ela-standards.BCE-Prettybird-Micro-Standard-v0.0.5
🚀 The Future Standard / Geleceğin Standartı
[English]
Beyond Raw Data: The Behavioral Revolution
The AI industry has been obsessed with the volume of data. At Prometech A.Ş., we are shifting the focus to the process of thought. BCE-Prettybird-Micro-Standart is not just a collection of Q&As; it is a blueprint for behavioral reasoning. By integrating Path Mapping and Behavioral DNA into the training loop, we are setting the new industry standard: Small models with elite… See the full description on the dataset page: https://huggingface.co/datasets/pthinc/BCE-Prettybird-Micro-Standard-v0.0.5.streamlit-standardSamples in this benchmark were generated by RELAI using the following data source(s):
Data Source Name: Streamlit
Data Source Link: https://docs.streamlit.io/
Data Source License: https://github.com/streamlit/streamlit/blob/develop/LICENSE
Data Source Authors: Snowflake Inc.
AI Benchmarks by Data Agents. 2025 RELAI.AI. Licensed under CC BY 4.0. Source: https://relai.ai
BCE-Prettybird-Micro-Standard-v0.0.6
🚀 The Future Standard / Geleceğin Standartı
[English]
Beyond Raw Data: The Behavioral Revolution
The AI industry has been obsessed with the volume of data. At Prometech A.Ş., we are shifting the focus to the process of thought. BCE-Prettybird-Micro-Standart is not just a collection of Q&As; it is a blueprint for behavioral reasoning. By integrating Path Mapping and Behavioral DNA into the training loop, we are setting the new industry standard: Small models with elite… See the full description on the dataset page: https://huggingface.co/datasets/pthinc/BCE-Prettybird-Micro-Standard-v0.0.6.symfony-standardSamples in this benchmark were generated by RELAI using the following data source(s):
Data Source Name: Symfony
Documentation Data Source Link: https://symfony.com/doc/current/index.html
Data Source License: https://github.com/symfony/symfony?tab=MIT-1-ov-file#readme
Data Source Authors: Symfony SAS
AI Benchmarks by Data Agents © 2025 RELAI.AI · Licensed under CC BY 4.0. Source: https://relai.ai
weaviate-standardSamples in this benchmark were generated by RELAI using the following data source(s):
Data Source Name: weaviate
Data Source Link: https://weaviate.io/developers/weaviate
Data Source License: https://github.com/weaviate/weaviate/blob/main/LICENSE
Data Source Authors: Weaviate B.V.
AI Benchmarks by Data Agents. 2025 RELAI.AI. Licensed under CC BY 4.0. Source: https://relai.ai
mlflow-standardSamples in this benchmark were generated by RELAI using the following data source(s):
Data Source Name: mlflow
Data Source Link: https://mlflow.org/docs/latest/index.html
Data Source License: https://github.com/mlflow/mlflow/tree/master?tab=Apache-2.0-1-ov-file#readme
Data Source Authors: MLflow Project, a Series of LF Projects, LLC
AI Benchmarks by Data Agents 2025 RELAI.AI. Licensed under CC BY 4.0. Source: https://relai.ai
opencv-standardSamples in this benchmark were generated by RELAI using the following data source(s):
Data Source Name: opencv
Data Source Link: https://docs.opencv.org/
Data Source License: https://github.com/opencv/opencv/blob/master/LICENSE
Data Source Authors: opencv contributors
AI Benchmarks by Data Agents. 2025 RELAI.AI. Licensed under CC BY 4.0. Source: https://relai.ai
spacy-standardSamples in this benchmark were generated by RELAI using the following data source(s):
Data Source Name: Spacy
Data Source Link: https://spacy.io/usage
Data Source License: https://github.com/explosion/spaCy/blob/master/LICENSE
Data Source Authors: ExplosionAI GmbH, 2016 spaCy GmbH, 2015 Matthew Honnibal
AI Benchmarks by Data Agents. 2025 RELAI.AI. Licensed under CC BY 4.0. Source: https://relai.ai
