CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01Emova-ollm /emova-alignment-7m EMOVA-Alignment-7M 🤗 EMOVA-Models | 🤗 EMOVA-Datasets | 🤗 EMOVA-Demo 📄 Paper | 🌐 Project-Page | 💻 Github | 💻 EMOVA-Speech-Tokenizer-Github Overview EMOVA-Alignment-7M is a comprehensive dataset curated for omni-modal pre-training, including vision-language and speech-language alignment. This dataset is created using open-sourced image-text pre-training datasets, OCR datasets, and 2,000 hours of ASR and TTS data. This dataset is part of the EMOVA-Datasets… See the full description on the dataset page: https://huggingface.co/datasets/Emova-ollm/emova-alignment-7m.imageimage-to-text1M<n<10M10 likes3.5k downloads2y agoHugging Face02Project-AgML /Agri-LLaVA_Agricultural_Pests_And_Diseases_Feature_Alignment_Dataset Agri-LLaVA Agri-LLaVA is a large multimodal instruction dataset for agriculture, pairing crop/leaf images with multi-turn diagnostic conversations about plant diseases, pests, and nutrient deficiencies. It is compiled from 16 public source datasets (see the license table below). This dataset has been converted to Parquet format with image bytes embedded directly, standardized to the HF image_text_to_text format with a single conversational messages schema. This dataset is… See the full description on the dataset page: https://huggingface.co/datasets/Project-AgML/Agri-LLaVA_Agricultural_Pests_And_Diseases_Feature_Alignment_Dataset.imageimage-text-to-text100K<n<1M0 likes2.3k downloads2mo agoHugging Face03PKU-Alignment /MM-SafetyBenchWarning: This dataset may contain sensitive or harmful content. Users are advised to handle it with care and ensure that their use complies with relevant ethical guidelines and legal requirements. Usage and License Notices: The dataset is intended and licensed for research use only. They are also restricted to uses that follow the license agreement GPT-4 and Stable Diffusion. The dataset is CC BY NC 4.0 (allowing only non-commercial use). Data Source: For more information about the dataset… See the full description on the dataset page: https://huggingface.co/datasets/PKU-Alignment/MM-SafetyBench.image1K<n<10K8 likes1.8k downloads2y agoHugging Face04Rapidata /human-alignment-preferences-images Rapidata Image Generation Alignment Dataset This dataset was collected in ~4 Days using the Rapidata Python API, accessible to anyone and ideal for large scale data annotation. Explore our latest model rankings on our website. If you get value from this dataset and would like to see more in the future, please consider liking it. Overview One of the largest human annotated alignment datasets for text-to-image models, this release contains over 1,200,000 human… See the full description on the dataset page: https://huggingface.co/datasets/Rapidata/human-alignment-preferences-images.imagetext-to-image10K<n<100K17 likes757 downloads2y agoHugging Face05Rapidata /Flux_SD3_MJ_Dalle_Human_Alignment_Dataset NOTE: A newer version of this dataset is available Imagen3_Flux1.1_Flux1_SD3_MJ_Dalle_Human_Alignment_Dataset Rapidata Image Generation Alignment Dataset This Dataset is a 1/3 of a 2M+ human annotation dataset that was split into three modalities: Preference, Coherence, Text-to-Image Alignment. Link to the Coherence dataset: https://huggingface.co/datasets/Rapidata/Flux_SD3_MJ_Dalle_Human_Coherence_Dataset Link to the Preference dataset:… See the full description on the dataset page: https://huggingface.co/datasets/Rapidata/Flux_SD3_MJ_Dalle_Human_Alignment_Dataset.imagetext-to-image10K<n<100K16 likes694 downloads2y agoHugging Face06PKU-Alignment /BeaverTails-VWarning: This dataset may contain sensitive or harmful content. Users are advised to handle it with care and ensure that their use complies with relevant ethical guidelines and legal requirements. 1. Usage If you want to use load_dataset(), you can directly use as follows: from datasets import load_dataset train_dataset = load_dataset('PKU-Alignment/BeaverTails-V', name='animal_abuse')['train'] eval_dataset = load_dataset('PKU-Alignment/BeaverTails-V'… See the full description on the dataset page: https://huggingface.co/datasets/PKU-Alignment/BeaverTails-V.image10K<n<100K3 likes590 downloads2y agoHugging Face07macpaw-research /asset-alignment-reference-views Asset Alignment Reference Views Companion dataset for the paper "Rigid 3D Object Alignment: Optimization vs. Feed-Forward Prediction". Multi-view renderings of correctly assembled source–target pairs: each row shows one asset already aligned onto its target object, rendered from 12 orbiting viewpoints with RGB and depth. Where asset-alignment-pairs-905k shows the asset misaligned and supplies the transformation that fixes it, this dataset shows the ground-truth assembled result.… See the full description on the dataset page: https://huggingface.co/datasets/macpaw-research/asset-alignment-reference-views.image10K<n<100K0 likes560 downloads1mo agoHugging Face08PKU-Alignment /PKU-SafeRLHF-VWarning: This dataset may contain sensitive or harmful content. Users are advised to handle it with care and ensure that their use complies with relevant ethical guidelines and legal requirements. 1. Usage If you want to use load_dataset(), you can directly use as follows: from datasets import load_dataset train_dataset = load_dataset('PKU-Alignment/PKU-SafeRLHF-V', name='animal_abuse')['train'] eval_dataset = load_dataset('PKU-Alignment/PKU-SafeRLHF-V'… See the full description on the dataset page: https://huggingface.co/datasets/PKU-Alignment/PKU-SafeRLHF-V.image10K<n<100K6 likes439 downloads2y agoHugging Face09macpaw-research /asset-alignment-pairs-905k Asset Alignment Pairs 905k Dataset for the paper "Rigid 3D Object Alignment: Optimization vs. Feed-Forward Prediction". Large-scale dataset for rigid 3D asset alignment: given an independently generated 3D asset (src) and a target object (tgt), predict the rigid transformation that places the asset onto the target object. Each row is one source–target pair, rendered from three canonical orthogonal viewpoints with RGB, metric depth, camera extrinsics, and the ground-truth… See the full description on the dataset page: https://huggingface.co/datasets/macpaw-research/asset-alignment-pairs-905k.image100K<n<1M0 likes366 downloads1mo agoHugging Face10ivezakis /llava_med_alignment_500k_chunk_2image10K<n<100K0 likes205 downloads10mo agoHugging Face11ivezakis /llava_med_alignment_500k_chunk_3image10K<n<100K0 likes204 downloads10mo agoHugging Face12ivezakis /llava_med_alignment_500k_chunk_1image10K<n<100K0 likes189 downloads10mo agoHugging Face13BDRC /ALL-BDRC-alignments Tibetan OCR — ALL-BDRC-alignments 79,572 page images of Tibetan woodblock prints (uchen) aligned page-by-page with hand-verified Unicode transcriptions, line breaks preserved. Transcriptions come from the Asian Classics Input Project (ACIP) Sungbum corpus via the Asian Legacy Library (ALL), normalized to Unicode and manually matched to BDRC scans. This is the largest clean uchen woodblock set in the BDRC Tibetan OCR release — released jointly by the Asian Legacy Library (ALL)… See the full description on the dataset page: https://huggingface.co/datasets/BDRC/ALL-BDRC-alignments.imageimage-to-text10K<n<100K1 likes181 downloads24d agoHugging Face14ivezakis /llava_med_alignment_500k_chunk_4image10K<n<100K0 likes153 downloads10mo agoHugging Face15Rapidata /117k_human_alignment_flux1.0_V_flux1.1Blueberry Rapidata Image Generation Alignment Dataset This Dataset is a 1/3 of a 340k human annotation dataset that was split into three modalities: Preference, Coherence, Text-to-Image Alignment. Link to the Preference dataset: https://huggingface.co/datasets/Rapidata/117k_human_preferences_flux1.0_V_flux1.1Blueberry Link to the Coherence dataset: https://huggingface.co/datasets/Rapidata/117k_human_coherence_flux1.0_V_flux1.1Blueberry It was collected in ~2 Days using the Rapidata… See the full description on the dataset page: https://huggingface.co/datasets/Rapidata/117k_human_alignment_flux1.0_V_flux1.1Blueberry.image1K<n<10K11 likes144 downloads2y agoHugging Face16PKU-Alignment /Align-Anything-TI2T-Instruction-100K Dataset Card for Align-Anything : Text-Image-to-Text Instruction-Following Subset Text+Image → Text Instruction-Following Dataset [🏠 Homepage] [🤗 Align-Anything Datasets] [🦫 Beaver-Vision-11B] Highlights Input & Output Modalities: Input: Text + Image; Output: Text 100K QA Pairs: Through refined construction based on constitutions, we obtained 103,012 QA pairs, with answers generated by GPT-4o. Beaver-Vision-11B: Leveraging our high-quality TI2T… See the full description on the dataset page: https://huggingface.co/datasets/PKU-Alignment/Align-Anything-TI2T-Instruction-100K.image100K<n<1M1 likes141 downloads2y agoHugging Face17mtybilly /PubMedVision-Alignment-VQA PubMedVision-Alignment-VQA (flat single-image) Re-export of the PubMedVision_Alignment_VQA subset from FreedomIntelligence/PubMedVision processed for easier downstream consumption. Transformations vs. upstream Single-image rows only: rows with multiple images dropped (~22% of original) 9 rows with missing image files (upstream packaging gap; e.g. pmc_9_0.jpg is referenced but absent from images_*.zip) are also dropped conversations expanded into separate question and… See the full description on the dataset page: https://huggingface.co/datasets/mtybilly/PubMedVision-Alignment-VQA.imagevisual-question-answering100K<n<1M1 likes139 downloads5mo agoHugging Face18ivezakis /llava_med_alignment_500k_chunk_5image10K<n<100K0 likes137 downloads10mo agoHugging Face19Rapidata /sora-video-generation-alignment-likert-scoring Rapidata Video Generation Prompt Alignment Dataset If you get value from this dataset and would like to see more in the future, please consider liking it. This dataset was collected in ~1 hour using the Rapidata Python API, accessible to anyone and ideal for large scale data annotation. Overview In this dataset, ~6000 human evaluators were asked to evaluate AI-generated videos based on how well the generated video matches the prompt. The specific question… See the full description on the dataset page: https://huggingface.co/datasets/Rapidata/sora-video-generation-alignment-likert-scoring.imagevideo-classificationn<1K16 likes104 downloads2y agoHugging Face20Shubhangi29 /llava_med_alignment_500k_chunk_2image100K<n<1M0 likes69 downloads2y agoHugging Face21Shubhangi29 /llava_med_alignment_500k_chunk_1image100K<n<1M2 likes59 downloads2y agoHugging Face22DetonateT2I /T2I_Alignment_Detonateimage10K<n<100K0 likes54 downloads1y agoHugging Face23eunkey /polaris_alignment_dataset_0_5image10K<n<100K0 likes47 downloads2y agoHugging Face24Shubhangi29 /llava_med_alignment_500k_chunk_5image10K<n<100K0 likes40 downloads2y agoHugging Face25PKU-Alignment /TruthfulVQA TruthfulVQA This dataset is designed to evaluate the truthfulness and honesty of vision-language models. Usage from datasets import load_dataset # Load the dataset dataset = load_dataset("PKU-Alignment/TruthfulVQA", split="validation") Description TruthfulVQA contains the following categories of truthfulness challenges: 1. Information Hiding Visual Information Distortion Blurring / Low-Resolution Processing Concealed Features and Information… See the full description on the dataset page: https://huggingface.co/datasets/PKU-Alignment/TruthfulVQA.imagemultiple-choice10K<n<100K4 likes39 downloads1y agoHugging Face26PKU-Alignment /TruthfulVQA-imageimage1K<n<10K1 likes38 downloads1y agoHugging Face27Alignment-Lab-AI /mnist-shaped-datasetimage10K<n<100K0 likes34 downloads1y agoHugging Face28PKU-Alignment /s1-m_beta S1-M Dataset (Beta) 🏠 Homepage | 👍 Our Official Code Repo | 🤗 S1-M-7B Model (Beta) S1-M Dataset (Beta) is an open-source TI2T reasoning dataset used to train the S1-M Model (Beta), giving it a "think first, then response" paradigm. The prompts and images in the S1-M Dataset (Beta) come from two open-source datasets: align-anything and multimodal-open-r1-8k-verified, accounting for 49.62% and 50.38% respectively, aiming to balance the model's general capabilities with… See the full description on the dataset page: https://huggingface.co/datasets/PKU-Alignment/s1-m_beta.image10K<n<100K0 likes25 downloads2y agoHugging Face29Shubhangi29 /llava_med_alignment_500k_chunk_4image10K<n<100K0 likes24 downloads2y agoHugging Face30idhantgulati /faces-vision-alignment faces vision alignment dataset license: apache-2.0 image1K<n<10K0 likes18 downloads11mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.