CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01amphora /bon-resultstext10K<n<100K0 likes391 downloads2y agoHugging Face02ai2-adapt-dev /HERM_BoN_candidates Data Format [ { "id": "0", "instruction": "What are the names of some famous actors that started their careers on Broadway?", "model_input": "<|system|>\n</s>\n<|user|>\nWhat are the names of some famous actors that started their careers on Broadway?</s>\n<|assistant|>\n", "output": [ "1. Hugh Jackman - known for his Tony Award-winning role in \"The Boy from Oz\" and his performance in \"The Phantom of the Opera\"\n...", "1. Meryl Streep - \"A Midsummer… See the full description on the dataset page: https://huggingface.co/datasets/ai2-adapt-dev/HERM_BoN_candidates.tabular1K<n<10K0 likes149 downloads2y agoHugging Face03open-proc /OpenO1_SFT_ultra_BoN_rewardedtabular10M<n<100M1 likes92 downloads2y agoHugging Face04qd313 /bonsai-knowledge-base bonsAI Knowledge Base Offline strategy and troubleshooting corpus for bonsAI, a self-hosted AI assistant plugin for Steam Deck (Decky Loader). This dataset is downloaded at runtime by the plugin — it is not bundled with the plugin itself, and the plugin (Apache-2.0) ships no corpus content. What's in it 117 strategy cards across 13 titles (Baldur's Gate 3, Cyberpunk 2077, Deep Rock Galactic: Survivor, Fallout 4, Grand Theft Auto: San Andreas — The Definitive… See the full description on the dataset page: https://huggingface.co/datasets/qd313/bonsai-knowledge-base.tabularn<1K1 likes91 downloads17h agoHugging Face05zakoman /glm-5.3-flash-mathnet-bon glm-5.3-flash-mathnet-bon This is the continuation and the final set of ox-alpha-mathnet-bon. Verified chain-of-thought reasoning traces for competition mathematics, generated with GLM-5.3-Flash via best-of-N rejection sampling against the ShadenA/MathNet dataset (ICLR 2026). Statistics (this split) Metric Value Records (problem × attempt) 6,181 Distinct problems 848 Attempts per problem 7.29 (mean), 8 (max) Accepted (answer_correct = true) 3,705… See the full description on the dataset page: https://huggingface.co/datasets/zakoman/glm-5.3-flash-mathnet-bon.texttext-generation1K<n<10K1 likes91 downloads13d agoHugging Face06rujiewu /Bongard-OpenWorld Bongard-OpenWorld Bongard-OpenWorld is a new benchmark for evaluating real-world few-shot reasoning for machine vision. Please refer to our project page. Some links may be invalid due to the instability of the URLs. To ensure that the community can reproduce our results from scratch, we have provided a backup of all the images. You can download from Google Drive. Please note that this page only hosts the data for Bongard-OpenWorld. All images of Bongard-OpenWorld are crawled from… See the full description on the dataset page: https://huggingface.co/datasets/rujiewu/Bongard-OpenWorld.textfeature-extraction1K<n<10K5 likes86 downloads2y agoHugging Face07MH0386 /napoleon_bonaparte Napoleon Bonaparte The Napoleon Bonaparte dataset is a collection of information and data related to Napoleon Bonaparte's life and reign. It includes details on his military campaigns, battles, conquests, and political career as Emperor of France. The dataset also contains information on the social and economic reforms he implemented in France, such as the establishment of the Napoleonic Code. The data is gathered from various sources, including historical records, biographies, and… See the full description on the dataset page: https://huggingface.co/datasets/MH0386/napoleon_bonaparte.textfeature-extraction1K<n<10K0 likes85 downloads2y agoHugging Face08referencesource /state-section-179-bonus-depreciation-conformity State conformity to federal Section 179 and bonus depreciation rules Canonical, always-current version: https://referencesource.org/state-section-179-bonus-depreciation-conformity/ Machine-readable: https://referencesource.org/state-section-179-bonus-depreciation-conformity/data.json — this mirror is a point-in-time copy. Last verified: 2026-08-15 Stale after: 2027-08-15 (past this date, prefer the canonical copy — it re-verifies on a cadence this snapshot does not) Records: 51… See the full description on the dataset page: https://huggingface.co/datasets/referencesource/state-section-179-bonus-depreciation-conformity.textn<1K0 likes73 downloads29d agoHugging Face09MIS-752 /bonus-richardyoung How has genetic testing changed medicine? Asked by: richardyoung · Answered by: nvidia/nemotron-3.5-lightning:free · On: 2026-09-10 The model said. Genetic testing has moved from a specialized research tool to a routine part of patient care, and it's already reshaping hospital diagnostics and treatment. It helps catch hereditary diseases earlier, guides which chemotherapy or targeted therapies are most likely to work for a specific cancer, and allows doctors to tailor… See the full description on the dataset page: https://huggingface.co/datasets/MIS-752/bonus-richardyoung.textn<1K0 likes68 downloads14d agoHugging Face10richardyoung /mis752-bonus-one-question How has genetic testing changed medicine? Asked by: richardyoung · Answered by: nvidia/nemotron-3.5-lightning:free · On: 2026-09-10 The model said. Genetic testing has moved from a specialized research tool to a routine part of patient care, and it's already reshaping hospital diagnostics and treatment. It helps catch hereditary diseases earlier, guides which chemotherapy or targeted therapies are most likely to work for a specific cancer, and allows doctors to tailor… See the full description on the dataset page: https://huggingface.co/datasets/richardyoung/mis752-bonus-one-question.textn<1K0 likes64 downloads14d agoHugging Face11prathmeshadsod /bondshift-organic-chemistry BondShift: Organic Chemistry Mechanism Dataset 10,000 ground-truth-separated records for mechanism diagnosis, misconception repair, and chemistry tutoring. A narrow, auditable V1 dataset built from independently constructed scenario blueprints and deterministic answer keys. TL;DR BondShift addresses the right answer, wrong mechanism problem. It trains models to examine electron flow, formal charge, intermediates, pathway choice, and stereochemical… See the full description on the dataset page: https://huggingface.co/datasets/prathmeshadsod/bondshift-organic-chemistry.texttext-generation10K<n<100K0 likes63 downloads1mo agoHugging Face12MBrandys /mis752-bonus-one-question What monthly metrics should a small nonprofit monitor to catch cash flow shortfalls before they happen? Asked by: MBrandys · Answered by: nvidia/nemotron-3.5-lightning:free · On: 2026-09-17 The model said. Every month, start by recording your actual cash balance at the beginning and end of the month, then compare the ending figure to your average operating expenses – if you’re heading toward less than two months of cushion, that’s your first warning sign. Next, track the… See the full description on the dataset page: https://huggingface.co/datasets/MBrandys/mis752-bonus-one-question.textn<1K0 likes55 downloads8d agoHugging Face13m-a-p /OpenO1_SFT_ultra_BoN_positvie_reward_v3_N-sampletabular10M<n<100M1 likes54 downloads2y agoHugging Face14dquinine2002 /mis752-bonus-one-question What is a malignant Tumor? Asked by: dquinine2002 · Answered by: nvidia/nemotron-3.5-lightning:free · On: 2026-09-10 The model said. A malignant tumor is what doctors call cancer. It occurs when cells in the body start growing uncontrollably due to damage in their genetic material, often from aging, environmental exposures, or random errors during cell division. What makes it dangerous is that it can invade and destroy nearby healthy tissue, and it can also break away and… See the full description on the dataset page: https://huggingface.co/datasets/dquinine2002/mis752-bonus-one-question.textn<1K0 likes46 downloads14d agoHugging Face15zakoman /ox-alpha-mathnet-bon ox-alpha-mathnet-bon Verified chain-of-thought reasoning traces for competition mathematics, generated with the ox-alpha reasoning model via best-of-N rejection sampling against the ShadenA/MathNet dataset (ICLR 2026). ox-alpha stealth model spec On August 20, 2026, an anonymous model designated stealth/ox-alpha appeared on OpenRouter (and OpenCode) with no disclosed developer, a 1M-token context window, multimodal input (text, image, video), and a roughly… See the full description on the dataset page: https://huggingface.co/datasets/zakoman/ox-alpha-mathnet-bon.texttext-generation1K<n<10K0 likes45 downloads17d agoHugging Face16referencesource /state-public-works-bond-thresholds State public works payment and performance bond thresholds (Little Miller Acts) Canonical, always-current version: https://referencesource.org/state-public-works-bond-thresholds/ Machine-readable: https://referencesource.org/state-public-works-bond-thresholds/data.json — this mirror is a point-in-time copy. Last verified: 2026-08-17 Stale after: 2027-08-17 (past this date, prefer the canonical copy — it re-verifies on a cadence this snapshot does not) Records: 53 Do I need a… See the full description on the dataset page: https://huggingface.co/datasets/referencesource/state-public-works-bond-thresholds.textn<1K0 likes43 downloads29d agoHugging Face17wrootpress /bonaventure-sentences Bonaventure on the Sentences (Latin ↔ English) 2,113 chunks of Bonaventure's Commentary on the Sentences and related works, Latin (Quaracchi) aligned with English (~3.33M words). Apparatus (notes, scholia) is kept in a separate field, never merged into the text. Canonical home: https://bonaventure.wrootpress.com (each record carries its canonical URL). This dataset is a machine-generated export of that site's build, regenerated from the source of truth and never hand-edited; the… See the full description on the dataset page: https://huggingface.co/datasets/wrootpress/bonaventure-sentences.text1K<n<10K0 likes42 downloads6d agoHugging Face18ruthlesslearner /mis752-bonus-one-question What is the top paying medical specialty in the USA? Asked by: ruthlesslearner · Answered by: nvidia/nemotron-3.5-lightning:free · On: 2026-09-12 The model said. The top-paying medical specialty in the United States is generally neurosurgery, though orthopedic surgery and plastic surgery consistently rank near the very top as well. In broad terms, neurosurgeons tend to earn the highest average annual incomes—often exceeding $600,000 to $700,000, and sometimes more in major… See the full description on the dataset page: https://huggingface.co/datasets/ruthlesslearner/mis752-bonus-one-question.textn<1K1 likes39 downloads12d agoHugging Face19uran1um1 /bonsai_8b_distilled Bonsai 8b Distilled Open source distillation dataset containing 29435 Q/A pairs (approximately 24.5 million tokens) of synthetic data traces from Bonsai 8b. This is an open source distillation dataset (though it will work for other purposes) generated using an automated synthetic data generation pipeline of over 2 million tokens of text from PrismML's Bonsai 8b model. It was generated using an inference script where the model is prompted to create a high quality… See the full description on the dataset page: https://huggingface.co/datasets/uran1um1/bonsai_8b_distilled.text10K<n<100K2 likes37 downloads2mo agoHugging Face20MIS-752 /bonus-ruthlesslearner What is the top paying medical specialty in the USA? Asked by: ruthlesslearner · Answered by: nvidia/nemotron-3.5-lightning:free · On: 2026-09-12 The model said. The top-paying medical specialty in the United States is generally neurosurgery, though orthopedic surgery and plastic surgery consistently rank near the very top as well. In broad terms, neurosurgeons tend to earn the highest average annual incomes—often exceeding $600,000 to $700,000, and sometimes more in major… See the full description on the dataset page: https://huggingface.co/datasets/MIS-752/bonus-ruthlesslearner.textn<1K1 likes36 downloads12d agoHugging Face21BondFoundary /cyber-security-instruct-v1textn<1K0 likes35 downloads6mo agoHugging Face22DoctorEdoP369 /bonsai_8b_distilled_edited_106byDoctorEdoP369🧠 [bonsai_8b_distilled_edited_106byDoctorEdoP369 ] A highly curated, gold-standard dataset of 106 refined reasoning traces built for fine-tuning compact AI models. This dataset contains 106 high-density examples specifically designed for complex Chain of Thought (CoT) reasoning. Every single entry was thoroughly cleaned, mathematically verified, and enhanced through an advanced refining pipeline guided by Claude Opus 4.8 and Gemini Pro Extended. Zero Logical & Mathematical Errors: All… See the full description on the dataset page: https://huggingface.co/datasets/DoctorEdoP369/bonsai_8b_distilled_edited_106byDoctorEdoP369.textn<1K0 likes33 downloads2mo agoHugging Face23winnetka /mis752-bonus-one-question What is a healthcare risk score, and who uses it? Asked by: winnetka · Answered by: poolside/laguna-s-2.1:free · On: 2026-09-18 The model said. A healthcare risk score is a numerical tool that estimates a patient's likelihood of experiencing a specific health outcome, such as readmission, complications, or death, usually within a set time period. It combines information from a patient's medical record—like age, diagnoses, lab results, and procedures—into a single score using… See the full description on the dataset page: https://huggingface.co/datasets/winnetka/mis752-bonus-one-question.textn<1K0 likes32 downloads6d agoHugging Face24BondFoundary /quant-research-instruct-v1 BondFoundry Quantitative Research Instruct Dataset Premium instruction-tuning dataset for quantitative finance LLM fine-tuning. Updated automatically every night. Stats Total records: 898 Last updated: 2026-04-09 Records added today: 0 Format: JSONL system/instruction/response triplets Avg response length: 330 words License: MIT Personas Senior Factor Researcher, HFT Research Scientist, Quantitative Portfolio Strategist, Execution Research Analyst… See the full description on the dataset page: https://huggingface.co/datasets/BondFoundary/quant-research-instruct-v1.textn<1K2 likes30 downloads6mo agoHugging Face25Hasnan13 /mis752-bonus-one-question Which social determinants of health have the greatest impact on healthcare spending? Asked by: Hasnan13 · Answered by: nvidia/nemotron-3.5-lightning:free · On: 2026-09-13 The model said. Based on broad health policy research, the social determinants of health that most consistently drive up healthcare spending are housing stability and affordability, access to nutritious food (food security), and fundamental aspects of economic stability like income and education levels. For… See the full description on the dataset page: https://huggingface.co/datasets/Hasnan13/mis752-bonus-one-question.textn<1K0 likes30 downloads11d agoHugging Face26DanielWadeIII /mis752-bonus-one-question What is a healthcare risk score, and who uses it? Asked by: DanielWadeIII · Answered by: nvidia/nemotron-3.5-lightning:free · On: 2026-09-13 The model said. A healthcare risk score is a number generated from a patient’s medical history, diagnoses, treatments, and sometimes social or lifestyle factors that estimates their likelihood of experiencing a negative outcome—such as being readmitted to the hospital, developing complications, requiring expensive emergency care, or… See the full description on the dataset page: https://huggingface.co/datasets/DanielWadeIII/mis752-bonus-one-question.textn<1K0 likes30 downloads11d agoHugging Face27MIS-752 /bonus-kenneth-holstad What is something someone new to the healthcare data field might be surprised to learn? Asked by: kenneth-holstad · Answered by: nvidia/nemotron-3.5-lightning:free · On: 2026-09-14 The model said. One thing that often surprises people new to healthcare data is that even in a modern hospital, patient information is typically scattered across a dozen different computer systems—lab results, pharmacy records, physician notes, and billing—none of which were designed to share data… See the full description on the dataset page: https://huggingface.co/datasets/MIS-752/bonus-kenneth-holstad.textn<1K0 likes30 downloads11d agoHugging Face28BondFoundry /bondfoundry-quant-sampleBondFoundry Quant Sample — 10 Records Premium synthetic instruction-tuning data for enterprise AI teams building domain-specific quantitative finance models. What's inside 10 of the highest-depth records from BondFoundry's quantitative finance catalogue. Each record is generated by a senior quant PM-level persona operating under real institutional constraints — margin call pressure, risk committee pushback, regulatory deadlines, portfolio drawdown scenarios. Average word count: 568 words per… See the full description on the dataset page: https://huggingface.co/datasets/BondFoundry/bondfoundry-quant-sample.texttext-generationn<1K1 likes28 downloads5mo agoHugging Face29MIS-752 /bonus-DanielWadeIII What is a healthcare risk score, and who uses it? Asked by: DanielWadeIII · Answered by: nvidia/nemotron-3.5-lightning:free · On: 2026-09-13 The model said. A healthcare risk score is a number generated from a patient’s medical history, diagnoses, treatments, and sometimes social or lifestyle factors that estimates their likelihood of experiencing a negative outcome—such as being readmitted to the hospital, developing complications, requiring expensive emergency care, or… See the full description on the dataset page: https://huggingface.co/datasets/MIS-752/bonus-DanielWadeIII.textn<1K0 likes28 downloads11d agoHugging Face30kenneth-holstad /mis752-bonus-one-question What is something someone new to the healthcare data field might be surprised to learn? Asked by: kenneth-holstad · Answered by: nvidia/nemotron-3.5-lightning:free · On: 2026-09-14 The model said. One thing that often surprises people new to healthcare data is that even in a modern hospital, patient information is typically scattered across a dozen different computer systems—lab results, pharmacy records, physician notes, and billing—none of which were designed to share data… See the full description on the dataset page: https://huggingface.co/datasets/kenneth-holstad/mis752-bonus-one-question.textn<1K0 likes27 downloads11d agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.