CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01AI-MO /aimo-validation-aime Dataset Card for AIMO Validation AIME All 90 problems come from AIME 22, AIME 23, and AIME 24, and have been extracted directly from the AOPS wiki page https://artofproblemsolving.com/wiki/index.php/AIME_Problems_and_Solutions This dataset serves as an internal validation set during our participation in the AIMO progress prize competition. Using data after 2021 is to avoid potential overlap with the MATH training set. Here are the different columns in the dataset: problem: the… See the full description on the dataset page: https://huggingface.co/datasets/AI-MO/aimo-validation-aime.textn<1K68 likes35k downloads1y agoHugging Face02AI-MO /aimo-validation-amc Dataset Card for AIMO Validation AMC All 83 come from AMC12 2022, AMC12 2023, and have been extracted from the AOPS wiki page https://artofproblemsolving.com/wiki/index.php/AMC_12_Problems_and_Solutions This dataset serves as an internal validation set during our participation in the AIMO progress prize competition. Using data after 2021 is to avoid potential overlap with the MATH training set. Here are the different columns in the dataset: problem: the modified problem statement… See the full description on the dataset page: https://huggingface.co/datasets/AI-MO/aimo-validation-amc.tabularn<1K19 likes11k downloads1y agoHugging Face03D4nt3 /esb-datasets-earnings22-validation-tiny-filteredA filtered (<=30s duration) slice (512 samples) of the Earnings22 dataset. def add_duration(sample): y, sr = sample['audio']["array"], sample['audio']["sampling_rate"] sample['duration_ms']=librosa.get_duration(y=y, sr=sr) * 1000 return sample tedlium = load_dataset("esb/datasets", "earnings22", split='validation', trust_remote_code=True) # compute duration to filter tedlium = tedlium.map(add_duration) tedlium = tedlium.select(range(512)) # Whisper max supported duration tedlium… See the full description on the dataset page: https://huggingface.co/datasets/D4nt3/esb-datasets-earnings22-validation-tiny-filtered.audion<1K0 likes8.3k downloads2y agoHugging Face04AI-MO /aimo-validation-math-level-5 Dataset Card for AIMO Validation MATH Level 5 A subset of level 5 problems from https://huggingface.co/datasets/lighteval/MATH We have extracted the final answer from boxed, and only keep those with integer outputs. textn<1K11 likes2.4k downloads2y agoHugging Face05jiyu9437 /gaia_validationtextn<1K0 likes1.5k downloads1y agoHugging Face06AI-MO /aimo-validation-math-level-4 Dataset Card for AIMO Validation MATH Level 4 A subset of level 4 problems from https://huggingface.co/datasets/lighteval/MATH We have extracted the final answer from boxed, and only keep those with integer outputs. textn<1K4 likes1.2k downloads2y agoHugging Face07PaulineLi /QuantiPhy-validation QuantiPhy (Validation Set) Dataset Summary QuantiPhy is a benchmark for evaluating whether vision–language models (VLMs) can perform quantitative physical inference from visual evidence, rather than producing plausible but ungrounded numerical guesses. This repository contains the official validation set of QuantiPhy, released to support model development, ablation studies, and preliminary evaluation.The validation set represents approximately 4% of the full benchmark and… See the full description on the dataset page: https://huggingface.co/datasets/PaulineLi/QuantiPhy-validation.tabularvideo-text-to-textn<1K8 likes1.1k downloads9mo agoHugging Face08lauspectrum /gaia-validation-sampled_50textn<1K0 likes1k downloads1y agoHugging Face09sbordt /olmo-2-pretrain-validationtext10K<n<100K0 likes1k downloads5mo agoHugging Face10Intelligent-Internet /ii-agent_gaia-benchmark_validationtextn<1K8 likes929 downloads1y agoHugging Face11Multimodal-Fatima /COCO_captions_validation Dataset Card for "COCO_captions_validation" More Information needed image1K<n<10K0 likes894 downloads4y agoHugging Face12apollo-research /sae-skeskinen-TinyStories-hf-validation-tokenizer-gpt2_playtext10K<n<100K0 likes744 downloads3y agoHugging Face13seonglae /nq_open-validation Dataset Card for "nq_open-validation" More Information needed text100K<n<1M0 likes537 downloads3y agoHugging Face14Multimodal-Fatima /VQAv2_validation Dataset Card for "VQAv2_validation" More Information needed image100K<n<1M0 likes532 downloads3y agoHugging Face15AdoCleanCode /korea_speech_mfa_aligned_validationaudio100K<n<1M0 likes513 downloads8mo agoHugging Face16112bb /verl_validation_mmmu_charxiv_mathversetext1K<n<10K0 likes456 downloads9mo agoHugging Face17Multimodal-Fatima /VQAv2_sample_validation Dataset Card for "VQAv2_sample_validation" More Information needed image1K<n<10K0 likes450 downloads3y agoHugging Face18minh21 /COVID-QA-unique-context-test-10-percent-validation-10-percent Dataset Card for "COVID-QA-unique-context-test-10-percent-validation-10-percent" More Information needed tabular1K<n<10K0 likes443 downloads3y agoHugging Face19Multimodal-Fatima /Imagenet1k_sample_validation Dataset Card for "Imagenet1k_sample_validation" More Information needed image1K<n<10K0 likes371 downloads4y agoHugging Face20Multimodal-Fatima /TextVQA_validation Dataset Card for "TextVQA_validation" More Information needed image1K<n<10K0 likes362 downloads3y agoHugging Face21Suchae /Korea-AIHub-middlesenior-dialect-speech-validation-part2audio10K<n<100K0 likes323 downloads2y agoHugging Face22pietrolesci /pile-validationThis dataset has been created as an artefact of the paper Causal Estimation of Memorisation Profiles (Lesci et al., 2024). More info about this dataset in the related collection Memorisation-Profiles. The validation data used in our study. The Pythia suite does not have an official validation. However, we confirmed with the authors that the Pile validation split (this one) was not seen during training. It is still a bit confusing whether the Pile data can be released freely. Thus, we will… See the full description on the dataset page: https://huggingface.co/datasets/pietrolesci/pile-validation.text100K<n<1M0 likes274 downloads1y agoHugging Face23fimu-docproc-research /CIVQA_EasyOCR_Validation CIVQA EasyOCR Validation Dataset The CIVQA (Czech Invoice Visual Question Answering) dataset was created with EasyOCR. This dataset contains only the validation split. The train part of the dataset can be found on this URL: https://huggingface.co/datasets/fimu-docproc-research/CIVQA_EasyOCR_Train The encoded validation dataset for the LayoutLM can be found on this link: https://huggingface.co/datasets/fimu-docproc-research/CIVQA_EasyOCR_LayoutLM_Validation All invoices used in this… See the full description on the dataset page: https://huggingface.co/datasets/fimu-docproc-research/CIVQA_EasyOCR_Validation.text10K<n<100K0 likes248 downloads3y agoHugging Face24closji /wikipedia-20220301.en-0.005-validationtext1M<n<10M0 likes241 downloads4y agoHugging Face25chamber111 /VPPO_MMK12_validation Dataset Card for VPPO_MMK12_validation Dataset Details Dataset Description This dataset is the official validation split used to fine-tune the VPPO-7B and VPPO-32B models presented in our paper, "Spotlight on Token Perception for Multimodal Reinforcement Learning". This is a direct copy of the test split of FanqingM/MMK12 dataset. We have isolated it here to ensure the exact version used in our experiments is publicly available, guaranteeing reproducibility for… See the full description on the dataset page: https://huggingface.co/datasets/chamber111/VPPO_MMK12_validation.imageimage-text-to-text1K<n<10K1 likes240 downloads11mo agoHugging Face26Jackmin108 /c4-en-validationtext100K<n<1M4 likes233 downloads3y agoHugging Face27ytz20 /dapo_validationtext1K<n<10K0 likes227 downloads6mo agoHugging Face28etechgrid /ttm-validation-datasetaudio1K<n<10K0 likes168 downloads2y agoHugging Face29shibuina /drawvla-prompt-validation-clean DrawVLA — Sketch-Prompt Validation Circle (which) + arrow (where) + caption (what) visual instructions overlaid on LIBERO observations, each labelled with a binary verdict for training a prompt validator or a self-checking VLA: right — every channel is correct and exactly one reading survives; execute. wrong — a channel is incorrect or the deictic prompt remains under-determined; reject. Formerly ambiguous prompts are retained in this class. All captions are name-free L2/L3… See the full description on the dataset page: https://huggingface.co/datasets/shibuina/drawvla-prompt-validation-clean.tabularimage-classification10K<n<100K0 likes160 downloads15d agoHugging Face30Multimodal-Fatima /VizWiz_validation Dataset Card for "VizWiz_validation" More Information needed image1K<n<10K0 likes154 downloads4y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.