CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01GeroldMeisinger /laion2b-en-a65_cogvlm2-4bit_captions Abstract This dataset contains image captions for the laion2B-en aesthetics>=6.5 image dataset using CogVLM2-4bit with the "laion-pop"-prompt to generate captions which were "likely" used in Stable Diffusion 3 training. From these image captions new synthetic images were generated using stable-diffusion-3-medium (batch-size=8). The synthetic images are best viewed locally by cloning this repo with: git lfs install git clone… See the full description on the dataset page: https://huggingface.co/datasets/GeroldMeisinger/laion2b-en-a65_cogvlm2-4bit_captions.imageimage-classification1K<n<10K6 likes3k downloads2y agoHugging Face02ProGamerGov /synthetic-dataset-1m-dalle3-high-quality-captions Dataset Card for Dalle3 1 Million+ High Quality Captions Alt name: Human Preference Synthetic Dataset Example grids for landscapes, cats, creatures, and fantasy are also available. Description: This dataset comprises of AI-generated images sourced from various websites and individuals, primarily focusing on Dalle 3 content, along with contributions from other AI systems of sufficient quality like Stable Diffusion and Midjourney (MJ v5 and above). As users typically… See the full description on the dataset page: https://huggingface.co/datasets/ProGamerGov/synthetic-dataset-1m-dalle3-high-quality-captions.imagetext-to-image1M<n<10M154 likes1.9k downloads2y agoHugging Face03MelanieCo /Capillary-Dataset Capillary dataset Paper: Capillary Dataset: A dataset of nail-fold capillaries captured by microscopy for diabetes detection Github: https://github.com/urgonguyen/Capillarydataset.git The dataset are structured as follows: Capillary dataset ├── Classification ├── data_1x1_224 ├── data_concat_1x9_224 ├── data_concat_2x2_224 ├── data_concat_3x3_224 ├── data_concat_4x1_224 └── data_concat_4x4_224 ├── Morphology_detection… See the full description on the dataset page: https://huggingface.co/datasets/MelanieCo/Capillary-Dataset.imageimage-classification10K<n<100K0 likes1.4k downloads3mo agoHugging Face04HanaNguyen /Capillary-Dataset Capillary dataset Paper: Capillary Dataset: A dataset of nail-fold capillaries captured by microscopy for diabetes detection Github: https://github.com/urgonguyen/Capillarydataset.git The dataset are structured as follows: Capillary dataset ├── Classification ├── data_1x1_224 ├── data_concat_1x9_224 ├── data_concat_2x2_224 ├── data_concat_3x3_224 ├── data_concat_4x1_224 └── data_concat_4x4_224 ├── Morphology_detection… See the full description on the dataset page: https://huggingface.co/datasets/HanaNguyen/Capillary-Dataset.imageimage-classification10K<n<100K2 likes924 downloads8mo agoHugging Face05neurlang /Minecraft-Skins-Captioned-1M Dataset Card for Minecraft Skins Dataset Summary This dataset contains 981,079 unique Minecraft player skins collected from various sources. Each skin is stored as a base64-encoded image with a unique identifier. Dataset Structure Data Fields This dataset includes the following fields: hash: A data dependent hash. These hashes are generated from raw bytes and will be same if the skin is identical. image: The skin image encoded in base64 format.… See the full description on the dataset page: https://huggingface.co/datasets/neurlang/Minecraft-Skins-Captioned-1M.textimage-classification1M<n<10M7 likes676 downloads1y agoHugging Face06ekacare /IntraOral_Gingivitis_Image_Captioning A DENTAL INTRAORAL IMAGE DATASET OF GINGIVITIS FOR IMAGE CAPTIONING Dataset Description This dataset is a copy of A Dental IntraOral Image Dataset of Gingivitis for Image Captioning which is shared with the license CC BY 4.0. This dataset contains 1,096 samples organized across multiple splits. The dataset includes image data. Splits train: 732 samples test: 182 samples validation: 182 samples Dataset Creation This dataset was created using… See the full description on the dataset page: https://huggingface.co/datasets/ekacare/IntraOral_Gingivitis_Image_Captioning.imageimage-classification1K<n<10K0 likes560 downloads1y agoHugging Face07alexsu52 /mvtec_capsule MVTec Capsule Category Dataset Labels {0: "normal", 1: "abnormal"} Number of Images {'train': 219, 'test': 132} How to Use Install datasets: pip install datasets Load the dataset: from datasets import load_dataset ds = load_dataset("alexsu52/mvtec_capsule") example = ds['train'][0] MVTEC Dataset Page https://www.mvtec.com/company/research/datasets/mvtec-ad Citation Paul Bergmann, Kilian Batzner, Michael Fauser… See the full description on the dataset page: https://huggingface.co/datasets/alexsu52/mvtec_capsule.image-classification1 likes443 downloads3y agoHugging Face08gccwang /Capillary-Dataset Capillary dataset Paper: Capillary Dataset: A dataset of nail-fold capillaries captured by microscopy for diabetes detection Github: https://github.com/urgonguyen/Capillarydataset.git The dataset are structured as follows: Capillary dataset ├── Classification ├── data_1x1_224 ├── data_concat_1x9_224 ├── data_concat_2x2_224 ├── data_concat_3x3_224 ├── data_concat_4x1_224 └── data_concat_4x4_224 ├── Morphology_detection… See the full description on the dataset page: https://huggingface.co/datasets/gccwang/Capillary-Dataset.imageimage-classification10K<n<100K0 likes378 downloads2mo agoHugging Face09arvoredossaberes /Capacitacao_Visao_Computacional Visão Geral Este repositório contém as atividades práticas e teóricas do curso de Capacitação em Visão Computacional. O curso aborda fundamentos de processamento digital de imagens, técnicas de filtragem, segmentação, extração de características e aplicações em aprendizado de máquina. Estrutura do Repositório O repositório está organizado em pastas por atividade, cada uma contendo: Enunciado da atividade em PDF Notebook Jupyter (quando aplicável) README com… See the full description on the dataset page: https://huggingface.co/datasets/arvoredossaberes/Capacitacao_Visao_Computacional.documentimage-classificationn<1K0 likes229 downloads4mo agoHugging Face10AhmedSSabir /Textual-Image-Caption-Dataset Update: OCT-2023 Add v2 with recent SoTA model swinV2 classifier for both soft/hard-label visual_caption_cosine_score_v2 with person label (0.2, 0.3 and 0.4) Introduction Modern image captaining relies heavily on extracting knowledge, from images such as objects, to capture the concept of static story in the image. In this paper, we propose a textual visual context dataset for captioning, where the publicly available dataset COCO caption (Lin et al., 2014) has been… See the full description on the dataset page: https://huggingface.co/datasets/AhmedSSabir/Textual-Image-Caption-Dataset.textimage-to-text7 likes217 downloads1y agoHugging Face11AvinashRicky /CaptchaOCR-500K CaptchaOCR-500K Dataset Summary CaptchaOCR-500K is a large-scale CAPTCHA recognition dataset containing 500,000 CAPTCHA images with corresponding text labels. The dataset is designed for training and evaluating Optical Character Recognition (OCR), CAPTCHA solving systems, image-to-text models, and computer vision models focused on text recognition. Tasks Optical Character Recognition (OCR) CAPTCHA Recognition Image-to-Text Computer Vision Text… See the full description on the dataset page: https://huggingface.co/datasets/AvinashRicky/CaptchaOCR-500K.imageimage-to-text100K<n<1M3 likes209 downloads3mo agoHugging Face12NoeFlandre /french-lot-department-captioned-photos Lot Department, France Image Dataset A collection of high-resolution scenic photographs from the Lot region of France with AI-generated descriptive captions. Dataset Summary This dataset contains scenic photographs from three notable locations in France's Lot department: Rocamadour, Autoire, and Padirac. All images were captured using a Sony A6600 camera and are paired with detailed English captions generated by Mistral AI's Pixtral-Large model. Key Features:… See the full description on the dataset page: https://huggingface.co/datasets/NoeFlandre/french-lot-department-captioned-photos.imageimage-to-textn<1K0 likes196 downloads1y agoHugging Face13imageomics /TreeOfLife-10M-Captions Dataset Card for TreeOfLife-10M Captions This dataset consists of generated captions, Wikipedia-derived descriptions and format examples for the TreeOfLife-10M. These captions were generated using InternVL3-38B based on biological contexts that help the model generate more accurate captions. It was used to train BioCAP, a CLIP-based model. Dataset Details This dataset is comprised of captions for the images in TreeOfLife-10M that were generated using InternVL3 38B.… See the full description on the dataset page: https://huggingface.co/datasets/imageomics/TreeOfLife-10M-Captions.textimage-classification1M<n<10M2 likes158 downloads11mo agoHugging Face14NoeFlandre /albi-captioned-photos Albi, France Image Dataset A collection of high-resolution scenic photographs from Albi, France with AI-generated descriptive captions. Dataset Summary This dataset contains scenic photographs from Albi, France, including the city center, the Toulouse Lautrec museum, and the Sainte-Cécile Cathedral. All images were captured using a Sony A6600 camera and are paired with detailed English captions generated by Mistral AI's Pixtral-Large model. Key Features: High-resolution… See the full description on the dataset page: https://huggingface.co/datasets/NoeFlandre/albi-captioned-photos.imageimage-to-textn<1K0 likes71 downloads1y agoHugging Face15Knight07 /Deep_Captcha DeepCaptcha Dataset: AI-Resistant CAPTCHA Benchmark Overview DeepCaptcha is a professional-grade dataset designed for benchmarking the AI-resistance and robustness of computer vision models. It contains CAPTCHA images generated with varying levels of adversarial protection, specifically engineered to thwart common automated recognition attacks (CNNs, OCR, etc.) while remaining human-readable. Dataset Structure The dataset is organized into 5 primary difficulty… See the full description on the dataset page: https://huggingface.co/datasets/Knight07/Deep_Captcha.imageimage-classificationn<1K0 likes57 downloads8mo agoHugging Face166DammK9 /danbooru2024-captions-1ktar Danbooru 2024 captions only in 1k tar Raw captions jointed by 7.62M unpublished extended dataset from KBlueLeaf/danbooru2023-metadata-database and 0.48M generated dataset via Minthy/ToriiGate-v0.4-7B in exl2-8bpw mode. There are 8.13M in total. python convert_meta_to_tar.py Reading source JSON Keys count: 8136011 max id: 8360499 100%|███████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████| 1000/1000… See the full description on the dataset page: https://huggingface.co/datasets/6DammK9/danbooru2024-captions-1ktar.image-classification1M<n<10M3 likes56 downloads2y agoHugging Face17SilentAntagonist /vintage-photography-450k-high-quality-captionsThis is a 450k image datastet focused on photography from the 20th century, and their analog aspect. Many of the images are in high resolution. This dataset currently has 20k images captioned with InternVL2 26B, and is a work in progress (I plan to caption the entire dataset and also have short captions for all of the images, compute is an issue for now). imageimage-classification100K<n<1M36 likes48 downloads2y agoHugging Face18SilentAntagonist /vintage-artworks-60k-captionedThis is a dataset consisting of 60k vintage artworks from the 20th century, consisting of vintage pulp, sci-fi and pinup artworks from that era. The dataset has short and long captions for each image, as well as resolution information. The large captions (large_caption column) were made with florence-2-large-ft, and then shortened with llama 3 8b (see short_caption column). imagefeature-extraction10K<n<100K2 likes41 downloads2y agoHugging Face19Kev0208 /PokeFA-pokemon-fanart-captioned PokeFA — Pokémon fan-art metadata with relevance/aesthetic scores and hybrid captions PokeFA is a large-scale Pokémon fan-art dataset released as metadata + URLs only (no image bytes).~30,000 candidate images are collected across 1,025 Pokémon using a popularity-banded budget with following curation pipeline: NSFW filtering → OCR localization & inpainting → resizing → relevance & aesthetic scoring (GPT-5-mini vision) → near-duplicate removal → quality filtering to the top ~16… See the full description on the dataset page: https://huggingface.co/datasets/Kev0208/PokeFA-pokemon-fanart-captioned.imagetext-to-image10K<n<100K1 likes41 downloads10mo agoHugging Face20lingcarzy /synthetic-dataset-1m-dalle3-high-quality-captions Dataset Card for Dalle3 1 Million+ High Quality Captions Alt name: Human Preference Synthetic Dataset Example grids for landscapes, cats, creatures, and fantasy are also available. Description: This dataset comprises of AI-generated images sourced from various websites and individuals, primarily focusing on Dalle 3 content, along with contributions from other AI systems of sufficient quality like Stable Diffusion and Midjourney (MJ v5 and above). As users typically… See the full description on the dataset page: https://huggingface.co/datasets/lingcarzy/synthetic-dataset-1m-dalle3-high-quality-captions.imagetext-to-image1M<n<10M0 likes39 downloads6mo agoHugging Face21dan6864 /CaptchaOCR-500K CaptchaOCR-500K Dataset Summary CaptchaOCR-500K is a large-scale CAPTCHA recognition dataset containing 500,000 CAPTCHA images with corresponding text labels. The dataset is designed for training and evaluating Optical Character Recognition (OCR), CAPTCHA solving systems, image-to-text models, and computer vision models focused on text recognition. Tasks Optical Character Recognition (OCR) CAPTCHA Recognition Image-to-Text Computer Vision Text… See the full description on the dataset page: https://huggingface.co/datasets/dan6864/CaptchaOCR-500K.imageimage-to-text100K<n<1M0 likes38 downloads6d agoHugging Face22Caplin43 /humanoid-basic-actions-dataset-v1 Humanoid Basic Actions Dataset v1 Synthetic dataset for humanoid robot training simulation. Description This dataset contains labeled humanoid robot action images for basic movement recognition tasks. Classes walk run sit stand wave pick_object turn_left turn_right Structure dataset/ ├── train/ ├── validation/ Each folder contains subfolders named after action labels. Format Image Classification (RGB Images 224x224) Total… See the full description on the dataset page: https://huggingface.co/datasets/Caplin43/humanoid-basic-actions-dataset-v1.textimage-classificationn<1K0 likes33 downloads7mo agoHugging Face23strangerzonehf /Open-Captcha-Image-DLC File Breakdown: File Name Size Description .gitattributes 2.46 kB Rules for Git Large File Storage (LFS). README.md 42 Bytes Basic project description. Usage: This dataset can be used for: Training CAPTCHA Solvers:Build models capable of solving image-based CAPTCHA challenges. Testing Automation:Evaluate CAPTCHA-breaking algorithms for robustness. Security Research:Understand the limitations of CAPTCHA systems and enhance security protocols. imagezero-shot-image-classification1K<n<10K3 likes32 downloads2y agoHugging Face24lumasik /captcha-25k Synthetic-Captcha-25k A synthetic dataset consisting of 25,000 generated captcha images, designed for training and testing OCR and computer vision models. Dataset Curation Source: Generated via custom Python script (Synthetic Data). Variety: Includes 12+ types of noise filters, distortions, and variable font rendering to simulate real-world captcha challenges. Purpose: Created for OCR benchmarking and testing automated recognition systems. Warning: As pure synthetic data… See the full description on the dataset page: https://huggingface.co/datasets/lumasik/captcha-25k.imageimage-classification10K<n<100K0 likes31 downloads4mo agoHugging Face25Caplin43 /humanoid-pose-state-dataset-lite Humanoid Pose State Dataset Lite Lightweight synthetic dataset for humanoid robot pose classification. Pose Classes neutral walking_pose running_pose sitting_pose lifting_pose waving_pose Structure dataset/ ├── train/ ├── validation/ Each split contains pose-labeled image folders. Total Samples Train: 600 Validation: 150 Image Format RGB, 224x224 License MIT textimage-classificationn<1K0 likes22 downloads7mo agoHugging Face26deepghs /midjourney_captioned_23m_fullgated Midjourney Captioned Full Dataset This is the full dataset of Midjourney Captioned 23M dataset. And all the original images are maintained here. Thanks to the contribution of a certain third-party data provider who wishes to remain anonymous. Information Images There are 23167456 images in total. The maximum ID of these images is 23167456. Last updated at 2024-12-01 12:11:43 UTC. These are the information of recent 50 images: id width height filename… See the full description on the dataset page: https://huggingface.co/datasets/deepghs/midjourney_captioned_23m_full.imageimage-classification10M<n<100M36 likes21 downloads2y agoHugging Face276DammK9 /e621_2024-captions-1ktar E621 2024 captions only in 1k tar Raw captions jointed from lodestones/e621-captions It doesn't align to any dataset yet. meta_cap.json has been provided in compressed format if you want to train with kohyas triner. Currently I'm trying to merge this with my 2024 version. Core logic The script building this 1ktar There is not much choice, I don't have GPU to run for 1M captions with VLM so I just "take it or leave it". rearranged_tags = [row.regular_summary… See the full description on the dataset page: https://huggingface.co/datasets/6DammK9/e621_2024-captions-1ktar.textimage-classification1M<n<10M0 likes17 downloads2y agoHugging Face28Jhao0930 /capstone-figures Capstone Figure Dataset Scientific figure images extracted from ICLR 2024, ICLR 2025, NeurIPS 2023/2024, ICML 2023/2024 papers. Used for multidimensional figure quality assessment. Files Each image is named {VENUE}_{YEAR}_{ID:05d}_fig_{N}.png. Annotation data JSONL annotation files are available in the companion GitHub repo. Paper source Papers can be re-downloaded with download_papers.py from the GitHub repo. imageimage-classification1K<n<10K1 likes17 downloads4mo agoHugging Face29sebastiangv /plastic-caps Plastic Caps Dataset Dataset Summary This dataset contains images of plastic caps categorized by the color. The dataset is divided into training, validation, and test splits using a leakage-safe strategy that groups all images of the same physical cap together in the same split while maintaining balanced color distributions across subsets. Supported Tasks The initial and primary purpose of this dataset is Color classification (target label: color_category)… See the full description on the dataset page: https://huggingface.co/datasets/sebastiangv/plastic-caps.imageimage-classification1K<n<10K1 likes16 downloads7mo agoHugging Face30deepghs /deviantart_gt2likes_captioned_8m_fullgated DeviantArt Like>=2k Captioned Full Dataset This is the full dataset of DeviantArt dataset, which like count no less than 2k. And all the original images are maintained here. Thanks to the contribution of a certain third-party data provider who wishes to remain anonymous. Information Images There are 7930631 images in total. The maximum ID of these images is 7932093. Last updated at 2024-11-03 14:16:47 UTC. These are the information of recent 50 images: id… See the full description on the dataset page: https://huggingface.co/datasets/deepghs/deviantart_gt2likes_captioned_8m_full.image-classification1M<n<10M20 likes13 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.