datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
Emilia-YODAS-ENEmilia-ENredcaps5m_resizedMMAudio-precomputed-results
Precomputed results for MMAudio
Results from four model variants of MMAudio.
All results are in the .flac format with lossless compression.
A cache folder contains the feature caches computed by the evaluation script.
Code: https://github.com/hkchengrex/MMAudio
Evaluation: https://github.com/hkchengrex/av-benchmark
VGGSound
Contains the VGGSound test set results. There are 15220 videos, collected with our best effort. Not all videos in the test sets are available… See the full description on the dataset page: https://huggingface.co/datasets/hkchengrex/MMAudio-precomputed-results.modelnet40_normal_resampled-compressedwds_vtab-resisc45droid_low_resolutiongvhmr_resEmilia-YODAS-DEAnyInstruct-resolution-1024GyroDVD_resultshabitat_web_image_depth_RESCUEEmilia-ZHGarments2Look-Test-Set-Results
Garments2Look: A Multi-Reference Dataset for High-Fidelity Outfit-Level Virtual Try-On with Clothing and Accessories
Project Page | Paper | Code
Garments2Look is a large-scale multimodal dataset for outfit-level Virtual Try-On (VTON), comprising 80,000 many-garments-to-one-look pairs across 40 major categories and over 300 fine-grained subcategories. Each pair includes an outfit with 3-12 reference garment images (averaging 4.48), a model image wearing the outfit, and detailed item… See the full description on the dataset page: https://huggingface.co/datasets/ArtmeScienceLab/Garments2Look-Test-Set-Results.KRIS_Bench_Resultswds_resisc45Emilia-JAEmilia-YODAS-FREmilia-YODAS-JAEmilia-YODAS-KOEmilia-DEBBBC021masks of BBBC021
waa_resultspyramid_flow_ft_resultsEmilia-FRSVOO-ResultsGarments2Look-Test-Set-Results
Garments2Look: A Multi-Reference Dataset for High-Fidelity Outfit-Level Virtual Try-On with Clothing and Accessories
Project Page | Paper | Code
Garments2Look is a large-scale multimodal dataset for outfit-level Virtual Try-On (VTON), comprising 80,000 many-garments-to-one-look pairs across 40 major categories and over 300 fine-grained subcategories. Each pair includes an outfit with 3-12 reference garment images (averaging 4.48), a model image wearing the outfit, and detailed… See the full description on the dataset page: https://huggingface.co/datasets/baajarmah/Garments2Look-Test-Set-Results.Emilia-KOEmilia-EN-Betaresult_genrm_score
