CoolFace
Datasetpublic

hminjeong/TripleSumm-Mr.HiSum

Dataset Summary The original MR.HiSum (Most-replayed Highlight Detection and Summarization) was designed as a unimodal dataset and only provides pre-extracted features, which limits its use for multimodal research. To support the multimodal video summarization approach proposed in TripleSumm, we reconstructed the dataset by independently crawling the original videos using the provided metadata and extracting features across three distinct modalities: Visual, Audio, and Text. ⚠️… See the full description on the dataset page: https://huggingface.co/datasets/hminjeong/TripleSumm-Mr.HiSum.

sourceHugging Facecc-by-4.0updated 7mo agoView on Hugging Face
2likes329downloads
filemrhisum_feat_audio_ast.h517.61 GBdownload
filemrhisum_feat_text_roberta.h517.61 GBdownload
filemrhisum_feat_visual_inceptionv3.h523.47 GBdownload
filemrhisum_gt.h5146.7 MBdownload

hminjeong/TripleSumm-Mr.HiSum · main · files are served by the source, never re-hosted here