CoolFace
8 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01pixparse /cc3m-wds Dataset Card for Conceptual Captions (CC3M) Dataset Summary Conceptual Captions is a dataset consisting of ~3.3M images annotated with captions. In contrast with the curated style of other image caption annotations, Conceptual Caption images and their raw descriptions are harvested from the web, and therefore represent a wider variety of styles. More precisely, the raw descriptions are harvested from the Alt-text HTML attribute associated with web images. To arrive at the… See the full description on the dataset page: https://huggingface.co/datasets/pixparse/cc3m-wds.imageimage-to-text1M<n<10M58 likes15k downloads3y agoHugging Face02xiaorui638 /cc3mimage1M<n<10M0 likes1.2k downloads2y agoHugging Face03Leonardo6 /cc3m-recap-wdsimage1M<n<10M0 likes327 downloads1y agoHugging Face04ytaek-oh /cc3m-subset-100kimage100K<n<1M0 likes174 downloads2y agoHugging Face05T3989 /SD_Bias_CC3MThis repository contains the training dataset for the paper Would Deep Generative Models Amplify Bias in Future Models? This dataset contains the images generated for training OpenCLIP and for exploring how the generated training data will change the social biases in OpenCLIP. The images are generated by inputting the CC3M's captions as the prompts towards Stable Diffusion v1.5. image10K<n<100K0 likes30 downloads3y agoHugging Face06lifehacker777 /cc3m-wds Dataset Card for Conceptual Captions (CC3M) Dataset Summary Conceptual Captions is a dataset consisting of ~3.3M images annotated with captions. In contrast with the curated style of other image caption annotations, Conceptual Caption images and their raw descriptions are harvested from the web, and therefore represent a wider variety of styles. More precisely, the raw descriptions are harvested from the Alt-text HTML attribute associated with web images. To arrive at the… See the full description on the dataset page: https://huggingface.co/datasets/lifehacker777/cc3m-wds.imageimage-to-text1M<n<10M0 likes25 downloads10mo agoHugging Face07Leonardo6 /CC3m-embed-vicunaimage1M<n<10M0 likes12 downloads2y agoHugging Face08chaocq /cc3m-wds Dataset Card for Conceptual Captions (CC3M) Dataset Summary Conceptual Captions is a dataset consisting of ~3.3M images annotated with captions. In contrast with the curated style of other image caption annotations, Conceptual Caption images and their raw descriptions are harvested from the web, and therefore represent a wider variety of styles. More precisely, the raw descriptions are harvested from the Alt-text HTML attribute associated with web images. To arrive at the… See the full description on the dataset page: https://huggingface.co/datasets/chaocq/cc3m-wds.imageimage-to-text1M<n<10M1 likes3 downloads6mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.