CoolFace
Datasetpublic

VLM2Vec/MSVD

Clone from "friedrichor/MSVD". MSVD contains 1,970 videos, each of which is paired with ~40 captions. We adopt the official split: Train: 1,200 videos, 48,774 captions Val: 100 videos, 4,290 captions Test: 670 videos, 27,763 captions 🌟 Citation @inproceedings{chen2011collecting, title={Collecting highly parallel data for paraphrase evaluation}, author={Chen, David and Dolan, William B}, booktitle={Proceedings of the Annual Meeting of the Association for… See the full description on the dataset page: https://huggingface.co/datasets/VLM2Vec/MSVD.

sourceHugging Faceupdated 1y agoView on Hugging Face
4likes3.8kdownloads
Dataset Card

Clone from "friedrichor/MSVD".

MSVD contains 1,970 videos, each of which is paired with ~40 captions.

We adopt the official split:

  • β€”Train: 1,200 videos, 48,774 captions
  • β€”Val: 100 videos, 4,290 captions
  • β€”Test: 670 videos, 27,763 captions

🌟 Citation

bibtex
@inproceedings{chen2011collecting,
  title={Collecting highly parallel data for paraphrase evaluation},
  author={Chen, David and Dolan, William B},
  booktitle={Proceedings of the Annual Meeting of the Association for Computational Linguistics (ACL)},
  year={2011}
}