CoolFace
20 results

midjourney

Photoroom /midjourney-v6-recap Midjourney v6 Recaptioned ~1.2M Midjourney v6 images with captions from three VLMs: llava: Original LLaVA captions from the source dataset gemini: Gemini Flash 1.5 captions qwen3: Qwen3 VL 8B captions Caption coverage llava: available for all 1,235,432 images (from original dataset) gemini and qwen3: available for 1,017,105 images (82.3%) Source Based on brivangl/midjourney-v6-llava. imageimage-to-text1M<n<10M20 likes2.3k downloads6mo agoHugging Facezlab-princeton /i1-midjourneyv6-tfrecordi1: A Simple and Fully Open Recipe for Strong Text-to-Image Models Boya Zeng, Tianze Luo, Shu Pu, Jucheng Shen, Taiming Lu, Gabriel Sarch, Zhuang Liu Princeton University [arXiv][code][model][project page] Overview To prepare the dataset for training, we store the image-caption pairs as TFRecords. This HuggingFace dataset contains the TFRecords corresponding to the midjourneyv6 dataset at 256×256 resolution. It also serves as an example of what a dataset processed using… See the full description on the dataset page: https://huggingface.co/datasets/zlab-princeton/i1-midjourneyv6-tfrecord.text-to-image0 likes1.9k downloads1mo agoHugging Facei1-datasets /i1-midjourneyv6-512-resolution-1m-tfrecordi1: A Simple and Fully Open Recipe for Strong Text-to-Image Models Boya Zeng, Tianze Luo, Shu Pu, Jucheng Shen, Taiming Lu, Gabriel Sarch, Zhuang Liu Princeton University [arXiv][code][model][project page] Overview To prepare the dataset for training, we store the image-caption pairs as TFRecords. This HuggingFace dataset contains the TFRecords corresponding to the midjourneyv6 dataset at 512×512 resolution. Concretely, we only retain raw images with a shorter edge of at… See the full description on the dataset page: https://huggingface.co/datasets/i1-datasets/i1-midjourneyv6-512-resolution-1m-tfrecord.text-to-image0 likes1.2k downloads1mo agoHugging Facebrivangl /midjourney-v6-llavaThis dataset based on https://huggingface.co/datasets/CortexLM/midjourney-v6 dataset, captioned with LLava-1.6 model. This dataset was released as is. By accessing and using this dataset, you acknowledge and agree that Cortex Foundation and the author of this repo are not responsible for any copyright violations or legal consequences that may arise from the use of these images. imagetext-to-image100K<n<1M17 likes1k downloads2y agoHugging Facei1-datasets /i1-midjourneyv6-1024-resolution-1m-tfrecordi1: A Simple and Fully Open Recipe for Strong Text-to-Image Models Boya Zeng, Tianze Luo, Shu Pu, Jucheng Shen, Taiming Lu, Gabriel Sarch, Zhuang Liu Princeton University [arXiv][code][model][project page] Overview To prepare the dataset for training, we store the image-caption pairs as TFRecords. This HuggingFace dataset contains the TFRecords corresponding to the midjourneyv6 dataset at 1024×1024 resolution. Concretely, we only retain raw images with a shorter edge of… See the full description on the dataset page: https://huggingface.co/datasets/i1-datasets/i1-midjourneyv6-1024-resolution-1m-tfrecord.text-to-image0 likes925 downloads1mo agoHugging FaceCaptionEmporium /midjourney-niji-1m-llavanext Dataset Card for midjourney-niji-1m-llavanext Dataset Summary This is a dataset of 2,079,886 synthetic captions for 1,039,943 images from midjourney-v6-520k-raw and nijijourney-v6-520k-raw. The captions were produced using https://huggingface.co/lmms-lab/llama3-llava-next-8b inferenced in float16 after tags were generated with wd-swinv2-tagger-v3, followed by cleanup and shortening with Meta-Llama-3-8B. All images with metadata are available as MozJPEG encoded JPEGs… See the full description on the dataset page: https://huggingface.co/datasets/CaptionEmporium/midjourney-niji-1m-llavanext.imagetext-to-image1M<n<10M21 likes870 downloads2y agoHugging Face