CoolFace
4 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01MorningStar0709 /control-sci-corpus ControlSci Corpus Control science structured corpus with two configs: Sci-Align benchmark (500 questions) and Sciverse SFT instruction pairs (924 ChatML entries). License: CC-BY-4.0 Project: MorningStar0709/ControlMind Configs benchmark — Sci-Align Benchmark (500 questions) 4-dimension control science evaluation benchmark generated from the ControlSci structured corpus. Split: core (500 questions) Load: from datasets import load_dataset ds =… See the full description on the dataset page: https://huggingface.co/datasets/MorningStar0709/control-sci-corpus.imagequestion-answering1K<n<10K0 likes210 downloads2mo agoHugging Face02logicBombExe /turkish_cyber_security_controls_benchmark Turkish Cyber Security Controls Benchmark Türkçe siber güvenlik kontrol seçimi ve kontrol denetimi yeteneğini ölçmek için hazırlanmış, senaryo tabanlı çoktan seçmeli değerlendirme kümesidir. v0.1.0, uzman incelemesine açık ilk sürümdür ve NIST SP 800-53 Rev. 5, Release 5.2.0 kontrol kataloğunu hedefler. Kapsam 100 Türkçe senaryo NIST SP 800-53'ün 20 kontrol ailesinin her birinden 5 soru 64 kontrol seçimi sorusu 17 denetim kanıtı sorusu 19 denetim yargısı sorusu… See the full description on the dataset page: https://huggingface.co/datasets/logicBombExe/turkish_cyber_security_controls_benchmark.textquestion-answeringn<1K4 likes155 downloads2mo agoHugging Face03MichaelAnthony /hedgehog-loop-control-r4 hedgehog-loop-control-r4 Hedgehog — loop-control round 4 (termination/repetition fixes). Contents train.jsonl (2944 rows) validation.jsonl (438 rows) Format JSON Lines (.jsonl), one example per line. Provenance Original content for the Hedgehog extraction model (Michael Anthony Falabella). textquestion-answering1K<n<10K0 likes42 downloads29d agoHugging Face04vohonen /ai-control-corpus AI Control Corpus A question–answer dataset covering the AI Control literature until mid-2025. The corpus was assembled for fine-tuning experiments investigating self-fulfilling misalignment (see more). Dataset Summary Statistic Value Q&A pairs 2,633 Unique sources 212 Total tokens (Q+A) ~1.6 million Generation date August 2025 Source Distribution Source Type Pairs Share LessWrong 845 32.1% arXiv 545 20.7% Alignment Forum… See the full description on the dataset page: https://huggingface.co/datasets/vohonen/ai-control-corpus.textquestion-answering1K<n<10K0 likes33 downloads7mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.