CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
010-hero /prompt-perfect Scoring popular datasets with "Self-Alignment with Instruction Backtranslation" prompt 35 datasets scored (>6B tokens) Scoring Models used gpt-3.5-turbo-16k gpt-3.5-turbo-1106 gpt-3.5-turbo-0125 All datasets have 2 additional columns score - Response from the model including CoT (if provided) extracted_score - Extracted score from the score column as int Datasets Scored by Prompt (Needs to be updated)… See the full description on the dataset page: https://huggingface.co/datasets/0-hero/prompt-perfect.text1M<n<10M29 likes4.7k downloads3y agoHugging Face02nvidia /SWE-Hero-openhands-trajectories SWE-Hero Trajectories: Execution-based Fine-tuning for Software Engineering Agents Data Overview SWE-Hero Trajectories is an agentic instruction tuning dataset designed to advance the capabilities of LLMs in software engineering. This dataset comprises 34k agent trajectories collected using the OpenHands framework. The trajectories were synthesized using Qwen3-Coder-480B-A35B-Instruct, specifically curated for supervised fine-tuning (SFT), aiming to improve model… See the full description on the dataset page: https://huggingface.co/datasets/nvidia/SWE-Hero-openhands-trajectories.text10K<n<100K26 likes2k downloads5mo agoHugging Face03robinsonchristopher2837 /heron0 likes1.4k downloads1d agoHugging Face04heron-ai-security /stegoattack-advbench50 StegoAttack AdvBench-50 Steganographic jailbreak data generated using the StegoAttack pipeline from the paper "Hiding in Plain Sight: A Steganographic Approach to Stealthy LLM Jailbreaks" (Geng et al., 2025). For experiment results and analysis, see experiment.md. What is StegoAttack? StegoAttack is a jailbreak method that uses steganography to hide harmful queries inside benign-looking text. It embeds each word of a harmful query at a fixed position (e.g. the 2nd… See the full description on the dataset page: https://huggingface.co/datasets/heron-ai-security/stegoattack-advbench50.text-generationn<1K0 likes459 downloads10d agoHugging Face05fan-shu /swe-mt-combined-coderforge-hero-lego-nex-swezero fan-shu/swe-mt-combined-coderforge-hero-lego-nex-swezero Concatenated mid-train dataset for Qwen3 Thinking SFT. Each source subset is loaded in order and concatenated into a single config so one training epoch visits every trajectory exactly once (no interleave / no oversampling). Built from fan-shu/swe-instruct-trajectories-empty-think-inserted. Source subsets (7) togethercomputer__CoderForge-Preview nvidia__SWE-Zero-openhands-trajectories nex-agi__agent-sft… See the full description on the dataset page: https://huggingface.co/datasets/fan-shu/swe-mt-combined-coderforge-hero-lego-nex-swezero.text100K<n<1M0 likes447 downloads3mo agoHugging Face06mlfoundations-dev /hero_run_4_math_codetabular1M<n<10M0 likes375 downloads1y agoHugging Face07mlfoundations-dev /herorun1_code-test_50K_150K Dataset card for herorun1_code-test_50K_150K This dataset was made with Curator. Dataset details A sample from the dataset: { "problem": "You are tasked with implementing a softmax layer for a neural network using CUDA and C++. The softmax layer is a common component in neural network architectures and is used to normalize the output of a network to a probability distribution over multiple classes.\n\nYour task is to implement the following CUDA kernel for… See the full description on the dataset page: https://huggingface.co/datasets/mlfoundations-dev/herorun1_code-test_50K_150K.text10K<n<100K0 likes342 downloads2y agoHugging Face08laion /Qwen3-32B_hero_run_4_code_32k-sharegpt0 likes304 downloads10mo agoHugging Face09fan-shu /swe-mt-combined-hero-lego-nex-swezero fan-shu/swe-mt-combined-hero-lego-nex-swezero Concatenated mid-train dataset for Qwen3 Thinking SFT. Each source subset is loaded in order and concatenated into a single config so one training epoch visits every trajectory exactly once (no interleave / no oversampling). Built from fan-shu/swe-instruct-trajectories-empty-think-inserted. Source subsets (6) nvidia__SWE-Zero-openhands-trajectories nex-agi__agent-sft nvidia__SWE-Hero-openhands-trajectories… See the full description on the dataset page: https://huggingface.co/datasets/fan-shu/swe-mt-combined-hero-lego-nex-swezero.text100K<n<1M0 likes204 downloads3mo agoHugging Face10heroza /isic2017_task3image1K<n<10K0 likes198 downloads3y agoHugging Face110-hero /Matter-0.1 Matter 0.1 Curated top quality records from 35 other datasets. Extracted from prompt-perfect This is just a consolidation of all the score 5s. Fine-tuning models with various subsets and combinations to create a best performing v1 dataset ~1.4B Tokens, ~2.5M records Dataset has been deduped, decontaminated with bagel script from Jon Durbin Download using the below command to avoid unecessary files from huggingface_hub import snapshot_download… See the full description on the dataset page: https://huggingface.co/datasets/0-hero/Matter-0.1.text1M<n<10M53 likes187 downloads3y agoHugging Face12Silviase /Japanese-Heron-BenchThis dataset is a clarified version of the image, context, and question set included in the Japanese-Heron-Bench for the construction of the Japanese evaluation benchmark suite. The original dataset refers to turing-motors/Japanese-Heron-Bench. Link to the original dataset🔗: https://huggingface.co/datasets/turing-motors/Japanese-Heron-Bench @misc{inoue2024heronbench, title={Heron-Bench: A Benchmark for Evaluating Vision Language Models in Japanese}, author={Yuichi Inoue and Kento… See the full description on the dataset page: https://huggingface.co/datasets/Silviase/Japanese-Heron-Bench.imagen<1K1 likes177 downloads2y agoHugging Face130-hero /Matter-0.2-alphatext1M<n<10M3 likes175 downloads2y agoHugging Face14open-llm-leaderboard-old /details_0-hero__Matter-0.2-7B Dataset Card for Evaluation run of 0-hero/Matter-0.2-7B Dataset automatically created during the evaluation run of model 0-hero/Matter-0.2-7B on the Open LLM Leaderboard. The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_0-hero__Matter-0.2-7B.0 likes169 downloads2y agoHugging Face15turing-motors /Japanese-Heron-Bench Japanese-Heron-Bench Dataset Description Japanese-Heron-Bench is a benchmark for evaluating Japanese VLMs (Vision-Language Models). We collected 21 images related to Japan. We then set up three categories for each image: Conversation, Detail, and Complex, and prepared one or two questions for each category. The final evaluation dataset consists of 102 questions. Furthermore, each image is assigned one of seven subcategories: anime, art, culture, food, landscape, landmark… See the full description on the dataset page: https://huggingface.co/datasets/turing-motors/Japanese-Heron-Bench.imagevisual-question-answeringn<1K11 likes164 downloads2y agoHugging Face16open-llm-leaderboard-old /details_0-hero__Matter-0.2-32B Dataset Card for Evaluation run of 0-hero/Matter-0.2-32B Dataset automatically created during the evaluation run of model 0-hero/Matter-0.2-32B on the Open LLM Leaderboard. The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_0-hero__Matter-0.2-32B.0 likes163 downloads2y agoHugging Face17Herocat /place_bottletabular10K<n<100K0 likes160 downloads1y agoHugging Face180-hero /OIG-small-chip2 Dataset Card for "OIG-small-chip2" OIG-small-chip2 dataset from https://laion.ai/blog/oig-dataset/ Original Dataset - https://github.com/LAION-AI/Open-Instruction-Generalist text100K<n<1M11 likes153 downloads4y agoHugging Face19mlfoundations-dev /openthoughts3_herorun_ckpt06500_eval_27e9 mlfoundations-dev/openthoughts3_herorun_ckpt06500_eval_27e9 Precomputed model outputs for evaluation. Evaluation Results LiveCodeBench Average Accuracy: 59.95% ± 0.79% Number of Runs: 6 Run Accuracy Questions Solved Total Questions 1 59.30% 303 511 2 62.82% 321 511 3 60.08% 307 511 4 58.51% 299 511 5 61.45% 314 511 6 57.53% 294 511 tabular1K<n<10K0 likes153 downloads1y agoHugging Face20riggsybanez /HEROCS_Dataset0 likes147 downloads10mo agoHugging Face21average-developer /stocks-HEROMOTOCO-1D-candlesn<1K0 likes137 downloads16h agoHugging Face22dcml0714 /HerosHEROS is a dataset used to compare the sentence cosine similarity among sentences with high lexical overlapping but differ in their semantics. Please refer to the paper, "Revealing the Blind Spot of Sentence Encoder Evaluation by HEROS" for more details of how the dataset is constructed and the comparison of different sentence encoders. The dataset heros.tsv consists of 6 columns: Original, Synonym, Antonym, Negation, Random, Typo, Negation. The first column, Original are the sentences from… See the full description on the dataset page: https://huggingface.co/datasets/dcml0714/Heros.text1K<n<10K2 likes131 downloads3y agoHugging Face23orel3066 /btme-news-heroesimagen<1K0 likes124 downloads11d agoHugging Face24hoangbang /hard-hat-heroes Hard Hat Heroes: Construction Safety Detection Dataset Summary A public, viewer-ready educational challenge dataset. Host-only scoring data and hidden targets are excluded. Splits Split Examples Description train 4,000 Labeled training data test 1,000 Public inputs with withheld target labels or annotations Data Fields Field Type image Image image_id string width int64 height int64 objects.bbox… See the full description on the dataset page: https://huggingface.co/datasets/hoangbang/hard-hat-heroes.imageobject-detection1K<n<10K0 likes121 downloads2mo agoHugging Face25cuonghoang64602 /heron0 likes110 downloads4d agoHugging Face26open-llm-leaderboard-old /details_0-hero__Matter-0.1-Slim-7B-A Dataset Card for Evaluation run of 0-hero/Matter-0.1-Slim-7B-A Dataset automatically created during the evaluation run of model 0-hero/Matter-0.1-Slim-7B-A on the Open LLM Leaderboard. The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_0-hero__Matter-0.1-Slim-7B-A.10K<n<100K0 likes100 downloads3y agoHugging Face27nyu-dice-lab /lm-eval-results-nbeerbower-HeroBophades-2x7B-private Dataset Card for Evaluation run of nbeerbower/HeroBophades-2x7B Dataset automatically created during the evaluation run of model nbeerbower/HeroBophades-2x7B The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-nbeerbower-HeroBophades-2x7B-private.tabular100K<n<1M0 likes100 downloads2y agoHugging Face28mlfoundations-dev /herorun1_code-test_150K_250K Dataset card for herorun1_code-test_150K_250K This dataset was made with Curator. Dataset details A sample from the dataset: { "problem": "**A Recursive Riddle** \nImaginary scenario: I'm part of an experimental team at a tech lab where our latest project involves constructing a recursively defined program that reveals its own architecture. The mission is to build a function that not only discloses how many layers of functions exist but also specifies the code… See the full description on the dataset page: https://huggingface.co/datasets/mlfoundations-dev/herorun1_code-test_150K_250K.text10K<n<100K0 likes97 downloads2y agoHugging Face29open-llm-leaderboard-old /details_0-hero__Matter-0.1-Slim-7B-B Dataset Card for Evaluation run of 0-hero/Matter-0.1-Slim-7B-B Dataset automatically created during the evaluation run of model 0-hero/Matter-0.1-Slim-7B-B on the Open LLM Leaderboard. The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_0-hero__Matter-0.1-Slim-7B-B.0 likes95 downloads3y agoHugging Face30mlfoundations-dev /herorun1_code-test_250K_350K Dataset card for herorun1_code-test_250K_350K This dataset was made with Curator. Dataset details A sample from the dataset: { "problem": "You are tasked with creating a contract in Solidity that includes various utility functions related to token calculations. The contract should include functions for retrieving and setting decimals for tokens, calculating destination and source amounts based on token rates, retrieving token balances, and performing various… See the full description on the dataset page: https://huggingface.co/datasets/mlfoundations-dev/herorun1_code-test_250K_350K.text10K<n<100K0 likes89 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.