datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
bon-resultsHERM_BoN_candidates
Data Format
[
{
"id": "0",
"instruction": "What are the names of some famous actors that started their careers on Broadway?",
"model_input": "<|system|>\n</s>\n<|user|>\nWhat are the names of some famous actors that started their careers on Broadway?</s>\n<|assistant|>\n",
"output": [
"1. Hugh Jackman - known for his Tony Award-winning role in \"The Boy from Oz\" and his performance in \"The Phantom of the Opera\"\n...",
"1. Meryl Streep - \"A Midsummer… See the full description on the dataset page: https://huggingface.co/datasets/ai2-adapt-dev/HERM_BoN_candidates.OpenO1_SFT_ultra_BoN_rewardedbonsai-knowledge-base
bonsAI Knowledge Base
Offline strategy and troubleshooting corpus for bonsAI, a
self-hosted AI assistant plugin for Steam Deck (Decky Loader). This dataset is downloaded at
runtime by the plugin — it is not bundled with the plugin itself, and the plugin (Apache-2.0)
ships no corpus content.
What's in it
117 strategy cards across 13 titles (Baldur's Gate 3, Cyberpunk 2077, Deep Rock Galactic:
Survivor, Fallout 4, Grand Theft Auto: San Andreas — The Definitive… See the full description on the dataset page: https://huggingface.co/datasets/qd313/bonsai-knowledge-base.glm-5.3-flash-mathnet-bon
glm-5.3-flash-mathnet-bon
This is the continuation and the final set of ox-alpha-mathnet-bon.
Verified chain-of-thought reasoning traces for competition mathematics, generated with GLM-5.3-Flash via best-of-N rejection sampling against the ShadenA/MathNet dataset (ICLR 2026).
Statistics (this split)
Metric
Value
Records (problem × attempt)
6,181
Distinct problems
848
Attempts per problem
7.29 (mean), 8 (max)
Accepted (answer_correct = true)
3,705… See the full description on the dataset page: https://huggingface.co/datasets/zakoman/glm-5.3-flash-mathnet-bon.Bongard-OpenWorld
Bongard-OpenWorld
Bongard-OpenWorld is a new benchmark for evaluating real-world few-shot reasoning for machine vision. Please refer to our project page.
Some links may be invalid due to the instability of the URLs. To ensure that the community can reproduce our results from scratch, we have provided a backup of all the images. You can download from Google Drive.
Please note that this page only hosts the data for Bongard-OpenWorld. All images of Bongard-OpenWorld are crawled from… See the full description on the dataset page: https://huggingface.co/datasets/rujiewu/Bongard-OpenWorld.napoleon_bonaparte
Napoleon Bonaparte
The Napoleon Bonaparte dataset is a collection of information and data related to Napoleon Bonaparte's life and reign. It includes details on his military campaigns, battles, conquests, and political career as Emperor of France. The dataset also contains information on the social and economic reforms he implemented in France, such as the establishment of the Napoleonic Code. The data is gathered from various sources, including historical records, biographies, and… See the full description on the dataset page: https://huggingface.co/datasets/MH0386/napoleon_bonaparte.state-section-179-bonus-depreciation-conformity
State conformity to federal Section 179 and bonus depreciation rules
Canonical, always-current version: https://referencesource.org/state-section-179-bonus-depreciation-conformity/
Machine-readable: https://referencesource.org/state-section-179-bonus-depreciation-conformity/data.json — this mirror is a point-in-time copy.
Last verified: 2026-08-15
Stale after: 2027-08-15 (past this date, prefer the canonical copy —
it re-verifies on a cadence this snapshot does not)
Records: 51… See the full description on the dataset page: https://huggingface.co/datasets/referencesource/state-section-179-bonus-depreciation-conformity.bonus-richardyoung
How has genetic testing changed medicine?
Asked by: richardyoung · Answered by: nvidia/nemotron-3.5-lightning:free · On: 2026-09-10
The model said. Genetic testing has moved from a specialized research tool to a routine part of patient care, and it's already reshaping hospital diagnostics and treatment. It helps catch hereditary diseases earlier, guides which chemotherapy or targeted therapies are most likely to work for a specific cancer, and allows doctors to tailor… See the full description on the dataset page: https://huggingface.co/datasets/MIS-752/bonus-richardyoung.mis752-bonus-one-question
How has genetic testing changed medicine?
Asked by: richardyoung · Answered by: nvidia/nemotron-3.5-lightning:free · On: 2026-09-10
The model said. Genetic testing has moved from a specialized research tool to a routine part of patient care, and it's already reshaping hospital diagnostics and treatment. It helps catch hereditary diseases earlier, guides which chemotherapy or targeted therapies are most likely to work for a specific cancer, and allows doctors to tailor… See the full description on the dataset page: https://huggingface.co/datasets/richardyoung/mis752-bonus-one-question.bondshift-organic-chemistry
BondShift: Organic Chemistry Mechanism Dataset
10,000 ground-truth-separated records for mechanism diagnosis, misconception repair, and chemistry tutoring.
A narrow, auditable V1 dataset built from independently constructed scenario blueprints and deterministic answer keys.
TL;DR
BondShift addresses the right answer, wrong mechanism problem. It trains models to examine electron flow, formal
charge, intermediates, pathway choice, and stereochemical… See the full description on the dataset page: https://huggingface.co/datasets/prathmeshadsod/bondshift-organic-chemistry.mis752-bonus-one-question
What monthly metrics should a small nonprofit monitor to catch cash flow shortfalls before they happen?
Asked by: MBrandys · Answered by: nvidia/nemotron-3.5-lightning:free · On: 2026-09-17
The model said. Every month, start by recording your actual cash balance at the beginning and end of the month, then compare the ending figure to your average operating expenses – if you’re heading toward less than two months of cushion, that’s your first warning sign. Next, track the… See the full description on the dataset page: https://huggingface.co/datasets/MBrandys/mis752-bonus-one-question.OpenO1_SFT_ultra_BoN_positvie_reward_v3_N-samplemis752-bonus-one-question
What is a malignant Tumor?
Asked by: dquinine2002 · Answered by: nvidia/nemotron-3.5-lightning:free · On: 2026-09-10
The model said. A malignant tumor is what doctors call cancer. It occurs when cells in the body start growing uncontrollably due to damage in their genetic material, often from aging, environmental exposures, or random errors during cell division. What makes it dangerous is that it can invade and destroy nearby healthy tissue, and it can also break away and… See the full description on the dataset page: https://huggingface.co/datasets/dquinine2002/mis752-bonus-one-question.ox-alpha-mathnet-bon
ox-alpha-mathnet-bon
Verified chain-of-thought reasoning traces for competition mathematics, generated with the ox-alpha reasoning model via best-of-N rejection sampling against the ShadenA/MathNet dataset (ICLR 2026).
ox-alpha stealth model spec
On August 20, 2026, an anonymous model designated stealth/ox-alpha appeared on OpenRouter (and OpenCode) with no disclosed developer, a 1M-token context window, multimodal input (text, image, video), and a roughly… See the full description on the dataset page: https://huggingface.co/datasets/zakoman/ox-alpha-mathnet-bon.state-public-works-bond-thresholds
State public works payment and performance bond thresholds (Little Miller Acts)
Canonical, always-current version: https://referencesource.org/state-public-works-bond-thresholds/
Machine-readable: https://referencesource.org/state-public-works-bond-thresholds/data.json — this mirror is a point-in-time copy.
Last verified: 2026-08-17
Stale after: 2027-08-17 (past this date, prefer the canonical copy —
it re-verifies on a cadence this snapshot does not)
Records: 53
Do I need a… See the full description on the dataset page: https://huggingface.co/datasets/referencesource/state-public-works-bond-thresholds.bonaventure-sentences
Bonaventure on the Sentences (Latin ↔ English)
2,113 chunks of Bonaventure's Commentary on the Sentences and related works, Latin (Quaracchi) aligned with English (~3.33M words). Apparatus (notes, scholia) is kept in a separate field, never merged into the text.
Canonical home: https://bonaventure.wrootpress.com (each record carries its canonical URL). This dataset is a machine-generated export of that site's build, regenerated from the source of truth and never hand-edited; the… See the full description on the dataset page: https://huggingface.co/datasets/wrootpress/bonaventure-sentences.mis752-bonus-one-question
What is the top paying medical specialty in the USA?
Asked by: ruthlesslearner · Answered by: nvidia/nemotron-3.5-lightning:free · On: 2026-09-12
The model said. The top-paying medical specialty in the United States is generally neurosurgery, though orthopedic surgery and plastic surgery consistently rank near the very top as well.
In broad terms, neurosurgeons tend to earn the highest average annual incomes—often exceeding $600,000 to $700,000, and sometimes more in major… See the full description on the dataset page: https://huggingface.co/datasets/ruthlesslearner/mis752-bonus-one-question.bonsai_8b_distilled
Bonsai 8b Distilled
Open source distillation dataset containing 29435 Q/A pairs (approximately 24.5 million tokens) of synthetic data traces from Bonsai 8b.
This is an open source distillation dataset (though it will work for other purposes) generated using an automated synthetic data generation pipeline of over 2 million tokens of text from PrismML's Bonsai 8b model. It was generated using an inference script where the model is prompted to create a high quality… See the full description on the dataset page: https://huggingface.co/datasets/uran1um1/bonsai_8b_distilled.bonus-ruthlesslearner
What is the top paying medical specialty in the USA?
Asked by: ruthlesslearner · Answered by: nvidia/nemotron-3.5-lightning:free · On: 2026-09-12
The model said. The top-paying medical specialty in the United States is generally neurosurgery, though orthopedic surgery and plastic surgery consistently rank near the very top as well.
In broad terms, neurosurgeons tend to earn the highest average annual incomes—often exceeding $600,000 to $700,000, and sometimes more in major… See the full description on the dataset page: https://huggingface.co/datasets/MIS-752/bonus-ruthlesslearner.cyber-security-instruct-v1bonsai_8b_distilled_edited_106byDoctorEdoP369🧠 [bonsai_8b_distilled_edited_106byDoctorEdoP369 ]
A highly curated, gold-standard dataset of 106 refined reasoning traces built for fine-tuning compact AI models.
This dataset contains 106 high-density examples specifically designed for complex Chain of Thought (CoT) reasoning.
Every single entry was thoroughly cleaned, mathematically verified, and enhanced through an advanced refining pipeline guided by Claude Opus 4.8 and Gemini Pro Extended.
Zero Logical & Mathematical Errors: All… See the full description on the dataset page: https://huggingface.co/datasets/DoctorEdoP369/bonsai_8b_distilled_edited_106byDoctorEdoP369.mis752-bonus-one-question
What is a healthcare risk score, and who uses it?
Asked by: winnetka · Answered by: poolside/laguna-s-2.1:free · On: 2026-09-18
The model said. A healthcare risk score is a numerical tool that estimates a patient's likelihood of experiencing a specific health outcome, such as readmission, complications, or death, usually within a set time period. It combines information from a patient's medical record—like age, diagnoses, lab results, and procedures—into a single score using… See the full description on the dataset page: https://huggingface.co/datasets/winnetka/mis752-bonus-one-question.quant-research-instruct-v1
BondFoundry Quantitative Research Instruct Dataset
Premium instruction-tuning dataset for quantitative finance LLM fine-tuning.
Updated automatically every night.
Stats
Total records: 898
Last updated: 2026-04-09
Records added today: 0
Format: JSONL system/instruction/response triplets
Avg response length: 330 words
License: MIT
Personas
Senior Factor Researcher, HFT Research Scientist, Quantitative Portfolio Strategist,
Execution Research Analyst… See the full description on the dataset page: https://huggingface.co/datasets/BondFoundary/quant-research-instruct-v1.mis752-bonus-one-question
Which social determinants of health have the greatest impact on healthcare spending?
Asked by: Hasnan13 · Answered by: nvidia/nemotron-3.5-lightning:free · On: 2026-09-13
The model said. Based on broad health policy research, the social determinants of health that most consistently drive up healthcare spending are housing stability and affordability, access to nutritious food (food security), and fundamental aspects of economic stability like income and education levels. For… See the full description on the dataset page: https://huggingface.co/datasets/Hasnan13/mis752-bonus-one-question.mis752-bonus-one-question
What is a healthcare risk score, and who uses it?
Asked by: DanielWadeIII · Answered by: nvidia/nemotron-3.5-lightning:free · On: 2026-09-13
The model said. A healthcare risk score is a number generated from a patient’s medical history, diagnoses, treatments, and sometimes social or lifestyle factors that estimates their likelihood of experiencing a negative outcome—such as being readmitted to the hospital, developing complications, requiring expensive emergency care, or… See the full description on the dataset page: https://huggingface.co/datasets/DanielWadeIII/mis752-bonus-one-question.bonus-kenneth-holstad
What is something someone new to the healthcare data field might be surprised to learn?
Asked by: kenneth-holstad · Answered by: nvidia/nemotron-3.5-lightning:free · On: 2026-09-14
The model said. One thing that often surprises people new to healthcare data is that even in a modern hospital, patient information is typically scattered across a dozen different computer systems—lab results, pharmacy records, physician notes, and billing—none of which were designed to share data… See the full description on the dataset page: https://huggingface.co/datasets/MIS-752/bonus-kenneth-holstad.bondfoundry-quant-sampleBondFoundry Quant Sample — 10 Records
Premium synthetic instruction-tuning data for enterprise AI teams building domain-specific quantitative finance models.
What's inside
10 of the highest-depth records from BondFoundry's quantitative finance catalogue. Each record is generated by a senior quant PM-level persona operating under real institutional constraints — margin call pressure, risk committee pushback, regulatory deadlines, portfolio drawdown scenarios.
Average word count: 568 words per… See the full description on the dataset page: https://huggingface.co/datasets/BondFoundry/bondfoundry-quant-sample.bonus-DanielWadeIII
What is a healthcare risk score, and who uses it?
Asked by: DanielWadeIII · Answered by: nvidia/nemotron-3.5-lightning:free · On: 2026-09-13
The model said. A healthcare risk score is a number generated from a patient’s medical history, diagnoses, treatments, and sometimes social or lifestyle factors that estimates their likelihood of experiencing a negative outcome—such as being readmitted to the hospital, developing complications, requiring expensive emergency care, or… See the full description on the dataset page: https://huggingface.co/datasets/MIS-752/bonus-DanielWadeIII.mis752-bonus-one-question
What is something someone new to the healthcare data field might be surprised to learn?
Asked by: kenneth-holstad · Answered by: nvidia/nemotron-3.5-lightning:free · On: 2026-09-14
The model said. One thing that often surprises people new to healthcare data is that even in a modern hospital, patient information is typically scattered across a dozen different computer systems—lab results, pharmacy records, physician notes, and billing—none of which were designed to share data… See the full description on the dataset page: https://huggingface.co/datasets/kenneth-holstad/mis752-bonus-one-question.
