CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01nlphuji /mscoco_2014_5k_test_image_text_retrieval MSCOCO (5K test set) Original paper: Microsoft COCO: Common Objects in Context Homepage: https://cocodataset.org/#home 5K test set split from: http://cs.stanford.edu/people/karpathy/deepimagesent/caption_datasets.zip Bibtex: @inproceedings{lin2014microsoft, title={Microsoft coco: Common objects in context}, author={Lin, Tsung-Yi and Maire, Michael and Belongie, Serge and Hays, James and Perona, Pietro and Ramanan, Deva and Doll{\'a}r, Piotr and Zitnick, C Lawrence}… See the full description on the dataset page: https://huggingface.co/datasets/nlphuji/mscoco_2014_5k_test_image_text_retrieval.image1K<n<10K11 likes2.1k downloads4y agoHugging Face02bitmind /MS-COCOimage100K<n<1M3 likes1.9k downloads2y agoHugging Face03AmritaBha /mscoco-controlnet-cannyimage100K<n<1M4 likes679 downloads2y agoHugging Face04clip-benchmark /wds_mscoco_captionsimage10K<n<100K4 likes582 downloads4y agoHugging Face05tsunghanwu /mscoco_chairimagen<1K0 likes427 downloads1y agoHugging Face06mm-eval /MS-COCO-Captionsimage10K<n<100K0 likes398 downloads2mo agoHugging Face07bitmind /MS-COCO-uniqueimage100K<n<1M0 likes345 downloads2y agoHugging Face08cat-state /mscoco-1st-captionTo reproduce, run pip install -r requirements.txt and download.sh. image100K<n<1M3 likes338 downloads4y agoHugging Face09hazal-karakus /mscoco-controlnet-canny-less-colorsimage100K<n<1M1 likes337 downloads2y agoHugging Face10patomp /thai-mscoco-2014-captions Usage from datasets import load_dataset dataset = load_dataset("patomp/thai-mscoco-2014-captions") dataset output DatasetDict({ train: Dataset({ features: ['image', 'filepath', 'sentids', 'filename', 'imgid', 'split', 'sentences_tokens', 'sentences_raw', 'sentences_sentid', 'cocoid', 'th_sentences_raw'], num_rows: 113287 }) validation: Dataset({ features: ['image', 'filepath', 'sentids', 'filename', 'imgid', 'split', 'sentences_tokens'… See the full description on the dataset page: https://huggingface.co/datasets/patomp/thai-mscoco-2014-captions.image100K<n<1M1 likes288 downloads3y agoHugging Face11ChristophSchuhmann /MS_COCO_2017_URL_TEXTimage100K<n<1M26 likes260 downloads5y agoHugging Face12clip-benchmark /wds_mscoco_captions2017image10K<n<100K8 likes220 downloads3y agoHugging Face13romrawinjp /mscoco Common Objects in Context (COCO) Dataset This dataset is English captions of COCO dataset. The splits in this dataset is set according to Andrej Karpathy's split from dataset_coco.json file. The collection was created specifically for simplicity of use in training and evaluation pipeline by non-commercial and research purposes. The COCO images dataset is licensed under a Creative Commons Attribution 4.0 License. Reference @misc{lin2015microsoftcococommonobjects… See the full description on the dataset page: https://huggingface.co/datasets/romrawinjp/mscoco.imagetext-to-image100K<n<1M2 likes152 downloads2y agoHugging Face14samirchar /mscocoimage100K<n<1M0 likes138 downloads1y agoHugging Face15AmritaBha /mscoco-colour_masksimage10K<n<100K0 likes122 downloads2y agoHugging Face16closji /mscoco_train_2014_openai_clip-vit-base-patch32_image_caption_retrieval_pairs_2022-09-01tabular10M<n<100M1 likes118 downloads4y agoHugging Face17closji /mscoco_train_2014_openai_clip-vit-base-patch32_image_image_retrieval_pairs_2022-09-15tabular10M<n<100M0 likes113 downloads4y agoHugging Face18mteb /mbeir_mscoco_task0image10K<n<100K0 likes112 downloads8mo agoHugging Face19mteb /mbeir_mscoco_task3image10K<n<100K0 likes106 downloads8mo agoHugging Face20SoroushVT /mscoco-small Dataset Card for "mscoco-small" More Information needed image10K<n<100K0 likes89 downloads3y agoHugging Face21NOVA-vision-language /MSCOCO_PT-BRtextn<1K2 likes81 downloads2y agoHugging Face22MINGYISU /mscoco_omni catalog.jsonl contains the captions and filenames for each image id. There are around 170 text-image-video-audio (omni) tuples in it. How to run Specify GOOGLE_API_KEY run the following to get setup readybash setup.sh then only need to run this file only in the futurepython geminiAPI.py If runs successfully, a file called mscoco_cmret.jsonl will be generated, please provide this file to me. text0 likes81 downloads4mo agoHugging Face23JotDe /mscoco_100k Dataset Card for "mscoco_100k" More Information needed image10K<n<100K1 likes77 downloads4y agoHugging Face24closji /mscoco_train_2014_openai_clip-vit-base-patch32_image_image_retrieval_pairs_2022-09-13tabular10M<n<100M0 likes76 downloads4y agoHugging Face25BidirLM /mscoco_contrastive ⚠️ Part of the BidirLM-Omni Collection > This dataset is a specific modality sub-sample of the corpus used to train the BidirLM-Omni models. Looking for the full training mixture? > If you want to access the complete, balanced 1.8M sample omnimodal dataset (integrating text, image, audio), please visit the global integration hub here:👉 BidirLM/BidirLM-Omni-Contrastive 📜 Citation If you use this processed dataset or the broader BidirLM mixture in your research, please cite… See the full description on the dataset page: https://huggingface.co/datasets/BidirLM/mscoco_contrastive.image100K<n<1M0 likes66 downloads5mo agoHugging Face26JotDe /mscoco_20k_unique_imgs Dataset Card for "mscoco_20k_unique_imgs" More Information needed image10K<n<100K1 likes63 downloads4y agoHugging Face27JotDe /mscoco_100k_30k_test Dataset Card for "mscoco_100k_30k_test" More Information needed image10K<n<100K0 likes57 downloads4y agoHugging Face28closji /mscoco_2014_train_captions_openai_clip-vit-base-patch32text100K<n<1M0 likes54 downloads4y agoHugging Face29closji /mscoco_train_2014_openai_clip-vit-base-patch32_image_caption_retrieval_pairstext10K<n<100K0 likes51 downloads4y agoHugging Face30closji /mscoco_train_2014_openai_clip-vit-base-patch32_self_retrievaltext10K<n<100K0 likes46 downloads4y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.