hminjeong/TripleSumm-Mr.HiSum
Dataset Summary The original MR.HiSum (Most-replayed Highlight Detection and Summarization) was designed as a unimodal dataset and only provides pre-extracted features, which limits its use for multimodal research. To support the multimodal video summarization approach proposed in TripleSumm, we reconstructed the dataset by independently crawling the original videos using the provided metadata and extracting features across three distinct modalities: Visual, Audio, and Text. ⚠️… See the full description on the dataset page: https://huggingface.co/datasets/hminjeong/TripleSumm-Mr.HiSum.
Conversations for this repository live on Hugging Face.
CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.
Open discussions on Hugging Face