CoolFace
Datasetpublic

sindhuhegde/multivsr

Dataset: MultiVSR We introduce a large-scale multilingual lip-reading dataset: MultiVSR. The dataset comprises a total of 12,000 hours of video footage, covering English + 12 non-English languages. MultiVSR is a massive dataset with a huge diversity in terms of the speakers as well as languages, with approximately 1.6M video clips across 123K YouTube videos. Please check the website for samples. Download instructions Please check the GitHub repo to download… See the full description on the dataset page: https://huggingface.co/datasets/sindhuhegde/multivsr.

sourceHugging Facemitupdated 1y agoView on Hugging Face
5likes762downloads
16 commits on main
0f4d7031y ago

Update README.md

sindhuhegde
83c4e321y ago

Update README.md

sindhuhegde
f2916fc1y ago

Update README.md

sindhuhegde
3b9784f1y ago

add data splits info

sindhuhegde
b5d06b81y ago

Delete multivsr.py

sindhuhegde
73687701y ago

add dataset script for custom downloading

sindhuhegde
39854931y ago

added metadata tar

prajwal@robots.ox.ac.uk
300a0c31y ago

added metadata

prajwal@robots.ox.ac.uk
9795b501y ago

Update README.md

sindhuhegde
2ad0ae81y ago

Upload dataset_teaser.gif

sindhuhegde
18335a01y ago

typo

sindhuhegde
dfbd3b71y ago

Initial dataset card

sindhuhegde
94dc4301y ago

Upload 4 files

sindhuhegde
74fef3a1y ago

Delete test.csv

sindhuhegde
db0b3091y ago

uploading file-lists and YouTube IDs (#1)

sindhuhegde, prajwalkr14
db877cb1y ago

initial commit

sindhuhegde