datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
imagenet-256-flux2-vae-latents
ImageNet-256 FLUX.2 VAE Latents
Pre-computed deterministic, model-facing encodings from the
FLUX.2 VAE (black-forest-labs/FLUX.2-dev)
for the full ImageNet-1K training set at 256x256 resolution, stored as Parquet
shards. Each example includes latents for both the original and horizontally
flipped image, enabling flip augmentation without re-encoding at training time.
Dataset Description
Each example contains:
Column
Shape
Stored type
Description… See the full description on the dataset page: https://huggingface.co/datasets/yuanchenyang/imagenet-256-flux2-vae-latents.VAEDecodedImages-SDXL
Dataset Card for Dataset Name
This dataset is a collection of pre/post SDXL VAE encoded-decoded pairs from the Danish newspaper TV2 Nord, based on alexandrainst/nordjylland-news-image-captioning.
Dataset Details
Dataset Description
Images are fed to diffusion models as latents - a distilled representation of the image that allows processing with reduced overhead. This is facilitated by a variational autoencoder (VAE), a neural network that encodes/decodes… See the full description on the dataset page: https://huggingface.co/datasets/joshuajewell/VAEDecodedImages-SDXL.imagenet-256-sd-vae-ft-mse-latents
ImageNet-256 SD-VAE-ft-MSE Latents
Pre-computed posterior means (no variance/std) from the Stable Diffusion VAE (stabilityai/sd-vae-ft-mse) for the full ImageNet-1K training set at 256×256 resolution, stored as Parquet shards. Each example includes latents for both the original and horizontally flipped image, enabling flip augmentation without re-encoding at training time.
Dataset Description
Each example contains:
Column
Shape
Type
Description
latent_mean… See the full description on the dataset page: https://huggingface.co/datasets/yuanchenyang/imagenet-256-sd-vae-ft-mse-latents.
