CoolFace
21 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01UCSC-VLAA /Recap-DataComp-1B Dataset Card for Recap-DataComp-1B Recap-DataComp-1B is a large-scale image-text dataset that has been recaptioned using an advanced LLaVA-1.5-LLaMA3-8B model to enhance the alignment and detail of textual descriptions. Dataset Details Dataset Description Our paper aims to bridge this community effort, leveraging the powerful and open-sourced LLaMA-3, a GPT-4 level LLM. Our recaptioning pipeline is simple: first, we fine-tune a LLaMA-3-8B powered… See the full description on the dataset page: https://huggingface.co/datasets/UCSC-VLAA/Recap-DataComp-1B.imagezero-shot-classification1B<n<10B205 likes9.5k downloads2y agoHugging Face02cornuHGF /recap-datacomp-12m-wdsimage10M<n<100M0 likes2.3k downloads1y agoHugging Face03umd-vt-nyu /datacomp_recap_metadata2image100M<n<1B2 likes338 downloads2y agoHugging Face04wjn922 /Recap-Datacomp-1B_tars_part7Final size: 7,236,721, samples per tar: 10000 image1M<n<10M0 likes335 downloads11mo agoHugging Face05Hennara /Recap-DataComp-1B_split_3image100M<n<1B0 likes319 downloads2y agoHugging Face06Hennara /Recap-DataComp-1B_split_4image100M<n<1B0 likes300 downloads2y agoHugging Face07mrdbourke /Recap-DataComp-1B-FoodOrDrink Recap-DataComp-1B: Food or Drink A filtered subset of Recap-DataComp-1B containing 106,230,157 rows classified as food/drink content, enriched with structured food/drink extraction from FoodExtract-v2. Overview Count Percentage Total rows 106,230,157 100% Food/drink (Stage 5 label) 96,618,895 91.0% Not food/drink (Stage 5 label) 9,611,262 9.0% FoodExtract (re_caption): food/drink 79,519,489 74.9% FoodExtract (re_caption): not food/drink 26,710,156… See the full description on the dataset page: https://huggingface.co/datasets/mrdbourke/Recap-DataComp-1B-FoodOrDrink.imagetext-classification100M<n<1B1 likes248 downloads6mo agoHugging Face08wjn922 /Recap-Datacomp-1B_tars_part9Final size: 7,240,328, samples per tar: 10000 image1M<n<10M0 likes205 downloads11mo agoHugging Face09Hennara /Recap-DataComp-1B_split_5image100M<n<1B0 likes196 downloads2y agoHugging Face10wjn922 /Recap-Datacomp-1B_tars_part8Final size: 7,250,604, samples per tar: 10000 image1M<n<10M0 likes180 downloads11mo agoHugging Face11wjn922 /Recap-Datacomp-1B_tars_part13image1M<n<10M0 likes176 downloads11mo agoHugging Face12Hennara /Recap-DataComp-1B_split_7image100M<n<1B0 likes168 downloads2y agoHugging Face13wjn922 /Recap-Datacomp-1B_tars_part14image1M<n<10M0 likes158 downloads11mo agoHugging Face14Hennara /Recap-DataComp-1B_split_8image100M<n<1B0 likes152 downloads2y agoHugging Face15Hennara /Recap-DataComp-1B_split_2image100M<n<1B0 likes144 downloads2y agoHugging Face16Hennara /Recap-DataComp-1B_split_1image100M<n<1B1 likes127 downloads2y agoHugging Face17Hennara /Recap-DataComp-1B_split_6image100M<n<1B0 likes126 downloads2y agoHugging Face18nnethercott /Recap-DataComp-100K Description Recap-DataComp-100K is a subset of UCSC-VLAA/Recap-DataComp-1B. This dataset aims to ease the development of vision-language models by providing a readily-available small collection of image-text pairs. Use this dataset for sanity checks, developing POCs, or other quick multimodal dev. For serious model training please refer to the original repo linked above. Citation Always cite the original authors . I've copied their citation info here for your… See the full description on the dataset page: https://huggingface.co/datasets/nnethercott/Recap-DataComp-100K.imageimage-to-text100K<n<1M1 likes32 downloads2y agoHugging Face19BootsofLagrangian /datacomp-recap-qwen3p5-35b-a3b DataComp recaptions with Qwen3.5-35B-A3B Dataset datacomp: 325.623 Million caption rows. This public caption-only repository contains 325,623,447 generated caption rows for 325,472,073 normalized-URL identities. It unifies the selected existing recap pack, the later generated/recovery records, and the managed legacy v2 recap pack without exposing historical surface names. Distinct stripped captions for the same URL are retained; only exact normalized-URL plus stripped-caption… See the full description on the dataset page: https://huggingface.co/datasets/BootsofLagrangian/datacomp-recap-qwen3p5-35b-a3b.image100M<n<1B0 likes30 downloads1mo agoHugging Face20wooj1nBot /recap-datacomp-384-1Mimage100K<n<1M0 likes18 downloads1y agoHugging Face21sroecker /recap-datacomp-1B-moondream-test1imagen<1K0 likes11 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.