VLM2Vec/MSVD
Clone from "friedrichor/MSVD". MSVD contains 1,970 videos, each of which is paired with ~40 captions. We adopt the official split: Train: 1,200 videos, 48,774 captions Val: 100 videos, 4,290 captions Test: 670 videos, 27,763 captions π Citation @inproceedings{chen2011collecting, title={Collecting highly parallel data for paraphrase evaluation}, author={Chen, David and Dolan, William B}, booktitle={Proceedings of the Annual Meeting of the Association forβ¦ See the full description on the dataset page: https://huggingface.co/datasets/VLM2Vec/MSVD.
43.8k
Clone from "friedrichor/MSVD".
MSVD contains 1,970 videos, each of which is paired with ~40 captions.
We adopt the official split:
- Train: 1,200 videos, 48,774 captions
- Val: 100 videos, 4,290 captions
- Test: 670 videos, 27,763 captions
π Citation
@inproceedings{chen2011collecting,
title={Collecting highly parallel data for paraphrase evaluation},
author={Chen, David and Dolan, William B},
booktitle={Proceedings of the Annual Meeting of the Association for Computational Linguistics (ACL)},
year={2011}
}