CoolFace
8 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01InternVL-U /ScaleEdit-12M ScaleEdit-12M: Scaling Open-Source Image Editing Data Generation via Multi-Agent Framework &nbsp; &nbsp; &nbsp; 📌 Overview The largest open-source instruction-based image editing dataset to date. ScaleEdit-12M contains 11.8 million rigorously verified instruction–image pairs spanning 23 task families across diverse real and synthetic visual domains. It was constructed using ScaleEditor, a fully open-source hierarchical multi-agent framework that eliminates… See the full description on the dataset page: https://huggingface.co/datasets/InternVL-U/ScaleEdit-12M.tabularimage-to-image10M<n<100M24 likes33k downloads2mo agoHugging Face02CaptionEmporium /pexels-568k-internvl2 Dataset Card for pexels-568k-internvl2 Dataset Summary This is 567,573 synthetic captions for the images found in ptx0/photo-concept-bucket. The captions were produced using OpenGVLab/InternVL2-40B-AWQ. The dataset was grounded for captioning using the tags originally listed. Languages The text is in English, but occasionally text in images in other languages is transcribed. Intended Usage Training text-to-image models and other machine learning… See the full description on the dataset page: https://huggingface.co/datasets/CaptionEmporium/pexels-568k-internvl2.imagetext-to-image100K<n<1M21 likes611 downloads2y agoHugging Face03OpenGVLab /InternVL-SA-1B-Caption Dataset Card for InternVL-SA-1B-Caption Overview The InternVL-SA-1B-Caption Dataset is a bilingual dataset created using the InternVL2-Llama3-76B model. The dataset contains 12 million image-caption pairs in both English and Chinese. All images are sourced from Meta’s SA-1B dataset, and captions were generated using specific prompts designed to minimize hallucinations and ensure accurate descriptions based on visible image content. The dataset is intended for use in tasks… See the full description on the dataset page: https://huggingface.co/datasets/OpenGVLab/InternVL-SA-1B-Caption.tabular1M<n<10M24 likes170 downloads2y agoHugging Face04CaptionEmporium /flickr-megalith-10m-internvl2-multi-caption Dataset Card for flickr-megalith-10m-internvl2-multi-caption Dataset Summary This is approximately 57.3 million synthetic captions for the images found in madebyollin/megalith-10m. It includes the following captions: InternVL2 8B long captions (by CaptionEmporium) InternVL2 8B short captions (by CaptionEmporium) Florence2 long captions (by aipicasso) Florence2 short captions (by CaptionEmporium) ShareCaptioner long captions (by drawthingsai) ShareCaptioner short… See the full description on the dataset page: https://huggingface.co/datasets/CaptionEmporium/flickr-megalith-10m-internvl2-multi-caption.imagetext-to-image1M<n<10M31 likes168 downloads2y agoHugging Face05aromanus /franka_pick_and_place_12_9_internvlatabular10K<n<100K0 likes39 downloads10mo agoHugging Face06aromanus /franka_pick_and_place_12_6_internvlatabular10K<n<100K0 likes34 downloads10mo agoHugging Face07ShreyashDhoot /internvl-auditortabular1K<n<10K0 likes16 downloads5mo agoHugging Face08aromanus /franka_pick_and_place_12_13_internvlatabular10K<n<100K0 likes6 downloads9mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.