CoolFace
12 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01pixparse /cc3m-wds Dataset Card for Conceptual Captions (CC3M) Dataset Summary Conceptual Captions is a dataset consisting of ~3.3M images annotated with captions. In contrast with the curated style of other image caption annotations, Conceptual Caption images and their raw descriptions are harvested from the web, and therefore represent a wider variety of styles. More precisely, the raw descriptions are harvested from the Alt-text HTML attribute associated with web images. To arrive at the… See the full description on the dataset page: https://huggingface.co/datasets/pixparse/cc3m-wds.imageimage-to-text1M<n<10M58 likes15k downloads3y agoHugging Face02xiaorui638 /cc3mimage1M<n<10M0 likes1.2k downloads2y agoHugging Face03Leonardo6 /cc3m-recap-wdsimage1M<n<10M0 likes328 downloads1y agoHugging Face04ytaek-oh /cc3m-subset-100kimage100K<n<1M0 likes167 downloads2y agoHugging Face05hung20gg /2M5_cc3m_clip_B224text1M<n<10M0 likes134 downloads10mo agoHugging Face06hung20gg /cc3m-siglip-b224text1M<n<10M0 likes130 downloads10mo agoHugging Face07hung20gg /cc3m-clip-b224text1M<n<10M0 likes91 downloads10mo agoHugging Face08T3989 /SD_Bias_CC3MThis repository contains the training dataset for the paper Would Deep Generative Models Amplify Bias in Future Models? This dataset contains the images generated for training OpenCLIP and for exploring how the generated training data will change the social biases in OpenCLIP. The images are generated by inputting the CC3M's captions as the prompts towards Stable Diffusion v1.5. image10K<n<100K0 likes64 downloads3y agoHugging Face09lifehacker777 /cc3m-wds Dataset Card for Conceptual Captions (CC3M) Dataset Summary Conceptual Captions is a dataset consisting of ~3.3M images annotated with captions. In contrast with the curated style of other image caption annotations, Conceptual Caption images and their raw descriptions are harvested from the web, and therefore represent a wider variety of styles. More precisely, the raw descriptions are harvested from the Alt-text HTML attribute associated with web images. To arrive at the… See the full description on the dataset page: https://huggingface.co/datasets/lifehacker777/cc3m-wds.imageimage-to-text1M<n<10M0 likes25 downloads10mo agoHugging Face10hung20gg /2M5_cc3m_siglip_B224text1M<n<10M0 likes19 downloads10mo agoHugging Face11Leonardo6 /CC3m-embed-vicunaimage1M<n<10M0 likes14 downloads2y agoHugging Face12chaocq /cc3m-wds Dataset Card for Conceptual Captions (CC3M) Dataset Summary Conceptual Captions is a dataset consisting of ~3.3M images annotated with captions. In contrast with the curated style of other image caption annotations, Conceptual Caption images and their raw descriptions are harvested from the web, and therefore represent a wider variety of styles. More precisely, the raw descriptions are harvested from the Alt-text HTML attribute associated with web images. To arrive at the… See the full description on the dataset page: https://huggingface.co/datasets/chaocq/cc3m-wds.imageimage-to-text1M<n<10M1 likes3 downloads6mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.