datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
imagenet_288_dcae_fp8
ImageNet-1k in 5GB
The full ImageNet-1k compressed to less than 5 GB
Compression procedure:
Resize shorter edge to 288 and crop longer edge to a multiple of 32
Analysis transform: DC-AE f32 c32
Quantization: 8 bit float (e4m3)
Entropy coding: TIFF (CMYK) with deflate
Example dataloader for training
import torch
import datasets
from types import SimpleNamespace
from diffusers import AutoencoderDC
from torchvision.transforms.v2 import ToPILImage, PILToTensor… See the full description on the dataset page: https://huggingface.co/datasets/danjacobellis/imagenet_288_dcae_fp8.imagenet_288_dcae_fp8_captionsOriginal https://huggingface.co/datasets/danjacobellis/imagenet_288_dcae_fp8
Captions from https://huggingface.co/datasets/visual-layer/imagenet-1k-vl-enriched and https://huggingface.co/datasets/gmongaras/Imagenet21K_Recaption
Ive uploaded just the captions from both of those aswell. You can find them in the files of this repo
genshin_woman_Formal_Outfit_flux1_kontext_fp8_extractedfp8linear-samplesInfiniteYou_PosterCraft_Wang_Leehom_Poster_FP8
reference on
target
InfiniteYou_PosterCraft_Wang_Leehom_Poster_FP8_WAVInfiniteYou_PosterCraft_Wang_Leehom_Poster_FP8_WAV_text_maskInfiniteYou_PosterCraft_Wang_Leehom_Poster_FP8_Wang_WAV_text_mask_inpaintGenshin_Impact_flux1_kontext_fp8_animetic_lightInfiniteYou_PosterCraft_Wang_Leehom_Poster_FP8_WAV_text_mask_inpaint
