AbstractPhil/diffusion-pretrain-set-ft1-1024
diffusion-pretrain-set-ft1-1024 1024px (2x) upscale of AbstractPhil/diffusion-pretrain-set-ft1. WARNING MUCH OF THIS DATA WAS MODEL UPSCALED USING RAPID UPSCALERS. THIS IS NOT CONSISTENTLY HIGH FIDELITY NOR IS IT EVEN CLOSE TO FAIR FIDELITY AT TIMES. PLEASE use this ONLY for pretraining, new concepts, and simple design purposes ONLY. HEAVILY PRUNE FOR FINETUNING. Thank you, good luck my friends. Details Model: realesr-general-x4v3 (SRVGG Compact… See the full description on the dataset page: https://huggingface.co/datasets/AbstractPhil/diffusion-pretrain-set-ft1-1024.
diffusion-pretrain-set-ft1-1024
1024px (2x) upscale of AbstractPhil/diffusion-pretrain-set-ft1.
WARNING
MUCH OF THIS DATA WAS MODEL UPSCALED USING RAPID UPSCALERS.
THIS IS NOT CONSISTENTLY HIGH FIDELITY NOR IS IT EVEN CLOSE TO FAIR FIDELITY AT TIMES.
PLEASE use this ONLY for pretraining, new concepts, and simple design purposes ONLY. HEAVILY PRUNE FOR FINETUNING.
Thank you, good luck my friends.
Details
- Model: realesr-general-x4v3 (SRVGG Compact, spandrel), fp16, native 4x forward -> bicubic+antialias retarget to 2x (supersampled).
- Selected by speed/fidelity Pareto on real data: rw_lpips 0.021, LR-consistency PSNR 41.4 (hallucination guard).
- Re-encoding: webp q95.
- All Image columns upscaled; every other column passed through unchanged.
- Images with min edge >= 1024px kept verbatim (original bytes/encoding).
- All image columns normalized to typed Image() features (several source subsets stored untyped raw bytes).
- Shards: 2048 source rows each, row groups of 512.
