hminjeong/TripleSumm-Mr.HiSum
Dataset Summary The original MR.HiSum (Most-replayed Highlight Detection and Summarization) was designed as a unimodal dataset and only provides pre-extracted features, which limits its use for multimodal research. To support the multimodal video summarization approach proposed in TripleSumm, we reconstructed the dataset by independently crawling the original videos using the provided metadata and extracting features across three distinct modalities: Visual, Audio, and Text. ⚠️… See the full description on the dataset page: https://huggingface.co/datasets/hminjeong/TripleSumm-Mr.HiSum.
2329
