datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
flux1-backup-202501flux1-backup-202507flux1-backup-202508flux1-backup-202410flux1-backup-202409flux1-backup-202412flux1-backup-202408flux1-backup-202411flux1-backup-202502Flux.1_Modelsflux1-backup-202503flux1-backup-202504flux1-backup-202505flux1-backup-202506117k_human_coherence_flux1.0_V_flux1.1Blueberry
Rapidata Image Generation Alignment Dataset
This Dataset is a 1/3 of a 340k human annotation dataset that was split into three modalities: Preference, Coherence, Text-to-Image Alignment.
Link to the Preference dataset: https://huggingface.co/datasets/Rapidata/117k_human_preferences_flux1.0_V_flux1.1Blueberry
Link to the Text-2-Image Alignment dataset: https://huggingface.co/datasets/Rapidata/117k_human_alignment_flux1.0_V_flux1.1Blueberry
It was collected in ~2 Days using the… See the full description on the dataset page: https://huggingface.co/datasets/Rapidata/117k_human_coherence_flux1.0_V_flux1.1Blueberry.flux_10k_captions
Dataset Card for "flux_10k_captions"
More Information needed
flux1-backup-202510117k_human_alignment_flux1.0_V_flux1.1Blueberry
Rapidata Image Generation Alignment Dataset
This Dataset is a 1/3 of a 340k human annotation dataset that was split into three modalities: Preference, Coherence, Text-to-Image Alignment.
Link to the Preference dataset: https://huggingface.co/datasets/Rapidata/117k_human_preferences_flux1.0_V_flux1.1Blueberry
Link to the Coherence dataset: https://huggingface.co/datasets/Rapidata/117k_human_coherence_flux1.0_V_flux1.1Blueberry
It was collected in ~2 Days using the Rapidata… See the full description on the dataset page: https://huggingface.co/datasets/Rapidata/117k_human_alignment_flux1.0_V_flux1.1Blueberry.Flux-1-Schnell-Images-1kThis dataset was generated by me using Flux 1 Schnell at 4 steps
flux1.1-likert-scale-preference
Flux1.1 Likert Scale Text-to-Image Alignment Evaluation
This dataset contains images generated using Flux1.1 [pro] based on the prompts from our text-to-image generation benchmark.
Where the benchmark generally focuses on pairwise comparisons to rank different image generation models against each other, this Likert-scale dataset focuses on one
particular model and aims to reveal the particular nuances and highlight strong and weaks points of the model.
If you get value from this… See the full description on the dataset page: https://huggingface.co/datasets/Rapidata/flux1.1-likert-scale-preference.117k_human_preferences_flux1.0_V_flux1.1Blueberry
Rapidata Image Generation Alignment Dataset
This Dataset is a 1/3 of a 340k human annotation dataset that was split into three modalities: Preference, Coherence, Text-to-Image Alignment.
Link to the Text-2-Image Alignment dataset: https://huggingface.co/datasets/Rapidata/117k_human_alignment_flux1.0_V_flux1.1Blueberry
Link to the Coherence dataset: https://huggingface.co/datasets/Rapidata/117k_human_coherence_flux1.0_V_flux1.1Blueberry
It was collected in ~2 Days using the… See the full description on the dataset page: https://huggingface.co/datasets/Rapidata/117k_human_preferences_flux1.0_V_flux1.1Blueberry.flux1-backup-202509image-generation-flux1-schnell
Dataset Card for Dataset Name
Dataset Details
Dataset Description
Curated by: [More Information Needed]
Funded by [optional]: [More Information Needed]
Shared by [optional]: [More Information Needed]
Language(s) (NLP): [More Information Needed]
License: [More Information Needed]
Dataset Sources [optional]
Repository: [More Information Needed]
Paper [optional]: [More Information Needed]
Demo [optional]: [More Information Needed]… See the full description on the dataset page: https://huggingface.co/datasets/davidberenstein1957/image-generation-flux1-schnell.Flux-1-Dev-Images-1kThis dataset was generated by me using Flux 1 Dev at 50 steps
FLUX.1-schnell-random
Flux 1 Schnell Random Images
A dataset of 3000 synthetic images (1024 x 1024) generated to explore the latent space of black-forest-labs/FLUX.1-schnell in a completely randomized, unbiased way.
Generation Parameters
Model Checkpoint: flux1-schnell-fp8.safetensors
Sampler: euler_ancestral | Steps: 4 | CFG: 1.0 | Flux Guidance: 3.5
Seed: Random
Prompt Configuration
Prompt: Dynamically assembled string using up to the maximum limit of 512 random… See the full description on the dataset page: https://huggingface.co/datasets/agentlans/FLUX.1-schnell-random.Flux-1-Scnell-enhancedflux-1-pro-regularisation-imagesThis data is derived of https://huggingface.co/datasets/Rapidata/human-style-preferences-images
(cut off date: 03.05.2025) and contains those images of Flux.1[pro] where it had won over
other models. These images were then manually filted to exlude those with bad anatomy.
Information from the orignial data:
This dataset was collected in ~4 Days using the Rapidata Python API, accessible to anyone and ideal for large scale data annotation.
Overview
One of the largest human… See the full description on the dataset page: https://huggingface.co/datasets/stablellama/flux-1-pro-regularisation-images.flux1_dev-small
Image Captioning Dataset
This dataset is designed to train visual language models (VLMs) for image captioning tasks. It contains a varied collection of images generated with flux1_dev associated with highly accurate and verified descriptive captions.
FLUX.1-dev-random
Flux 1 Dev Random Images
A dataset of over 1200 synthetic images (1024 x 1024) generated to explore the latent space of black-forest-labs/FLUX.1-dev in a completely randomized, unbiased way.
Generation Parameters
Model Checkpoint: flux1-dev-Q5_K_S.gguf
Sampler: euler_ancestral | Steps: 28 | CFG: 1.0 | Flux Guidance: 3.5
Seed: Random
Prompt Configuration
Prompt: Dynamically assembled string using up to the maximum limit of 512 random tokens.… See the full description on the dataset page: https://huggingface.co/datasets/agentlans/FLUX.1-dev-random.flux-1-dev-generated-10k
