datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
Laion-Aesthetics-High-Resolution-GoT
Laion-Aesthetics-High-Resolution-GoT
Paper
Dataset Description
The Laion-Aesthetics-High-Resolution-GoT dataset is a collection of 3.77 million image-text pairs with rich grounding annotations. This dataset extends high-quality images from the LAION-Aesthetics collection with detailed text descriptions and object-level grounding information.
Key Features
Size: 3.77 million samples
Modalities: Image, Text, and Grounding Annotations
Image Resolution:… See the full description on the dataset page: https://huggingface.co/datasets/LucasFang/Laion-Aesthetics-High-Resolution-GoT.Laion_aesthetics_5plus_1024_33Mimproved_aesthetics_4.5plus-ultra-hrversion https://git-lfs.github.com/spec/v1
oid sha256:98b45ea81164d1e1a1dd82255207053b15cd6c69d922a1c5cf3387ce604d4b74
size 28
improved_aesthetics_6.5pluslaion-aesthetics-12m-umap
LAION-Aesthetics :: CLIP → UMAP
This dataset is a CLIP (text) → UMAP embedding of the LAION-Aesthetics dataset - specifically the improved_aesthetics_6plus version, which filters the full dataset to images with scores of > 6 under the "aesthetic" filtering model.
Thanks LAION for this amazing corpus!
The dataset here includes coordinates for 3x separate UMAP fits using different values for the n_neighbors parameter - 10, 30, and 60 - which are broken out as separate columns with… See the full description on the dataset page: https://huggingface.co/datasets/dclure/laion-aesthetics-12m-umap.aesthetics_v2_4.75laion_aesthetics_v2_6.5plusLaion_aesthetics_5plus_1024_33M_csvaesthetics_v2_4.5improved_aesthetics_6.5plus_clip_retrievalaesthetics_v2_4.75_filteredlaion_aesthetics_v2_6.0plusaesthetic_scorelaion_aesthetics_v2_6.25plusaesthetics_6_5plusopentts-uk-aesthetics
Aesthetics of Open Text-to-Speech for 🇺🇦 Ukrainian dataset
Community
Discord: https://bit.ly/discord-uds
Speech Recognition: https://t.me/speech_recognition_uk
Speech Synthesis: https://t.me/speech_synthesis_uk
What is it?
This dataset contains metrics for https://huggingface.co/datasets/Yehor/opentts-uk dataset retrieved by https://github.com/facebookresearch/audiobox-aesthetics
How metrics calculated?
You can find a… See the full description on the dataset page: https://huggingface.co/datasets/Yehor/opentts-uk-aesthetics.aesthetic_score_danbooru2023dataset:https://huggingface.co/datasets/KBlueLeaf/danbooru2023-webp-4Mpixel
scorer:https://github.com/discus0434/aesthetic-predictor-v2-5
aesthetic_sim_under3_toekns_5_70aesthetics_v2_475_lowSim_2to25_AESTHETIC_SCORE5Audio-aesthetics-scoreaesthetic_sim_under3_toekns_15_45aesthetics_v2_4.75_filtered_5.5aesthetics_v2_475_lowSim_2to25aesthetic_sim_3_9_toekns_10_30_filtered_by_caption_lengthaesthetics_v2_475_TEXT_Marion_Gottibooru-aestheticsaesthetics_v2_475_lowSim_100aesthetics_v2_475_highSim_under9_100aesthetics_v2_475_highSim_6_AS5aesthetics_v2_475_lowSim_5
