datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
hi-stt-preprocessed-webdatasetproject-gutenberg-wds-preprocessedwaifu-preprocessed-datasetMELD-Preprocessed
MELD Preprocessed for SER
This dataset is the manually preprocessed audio only version of MELD, only audio IDs, utterance transcriptions, dialogue IDs and Utterance IDs were extracted.
S. Poria, D. Hazarika, N. Majumder, G. Naik, R. Mihalcea,
E. Cambria. MELD: A Multimodal Multi-Party Dataset
for Emotion Recognition in Conversation. (2018)
Chen, S.Y., Hsu, C.C., Kuo, C.C. and Ku, L.W.
EmotionLines: An Emotion Corpus of Multi-Party
Conversations. arXiv preprint arXiv:1802.08379… See the full description on the dataset page: https://huggingface.co/datasets/Vano04/MELD-Preprocessed.LeFusion_Preprocessed_DataOneModelForAll_Preprocessed_Tryon_Data
OneModelForAll Preprocessed Supplementary Data
This repository provides the preprocessed supplementary data used by
One Model For All, a unified
framework for virtual try-on, virtual try-off, and arbitrary-pose try-on.
The files are intended to be used together with the original VITON-HD and
DeepFashion-MultiModal-Parts2Whole datasets. They include auxiliary data
required by the OneModelForAll codebase, such as SMPL pose maps, head-region
images, and the JSONL split files used for… See the full description on the dataset page: https://huggingface.co/datasets/liujx233/OneModelForAll_Preprocessed_Tryon_Data.MVS_Synth_preprocessed
MVS-Synth Dataset for LBM
Preprocessed MVS-Synth Dataset following CUT3R.
Each scene contains the following structure when extracted:
0000/
├── cam/
| ├── 0000.npz
| ├── 0001.npz
| ├── ...
├── depth/
| ├── 0000.npy
| ├── 0001.npy
| ├── ...
└── rgb/
├── 0000.jpg
├── 0001.jpg
├── ...
Please note:
120 sequences in total.
12000 frames in total.
