CoolFace
Datasetpublicgated

Berkeley-NLP/visual_accent_dialect_archive

Source: https://www.youtube.com/@visualaccent/videos All rights belong to the original dataset creator. VADA-AVSR: an audio-visual dataset of non-native English ("accents") and English varieties ("dialects") We preprocessed the Visual Accent and Dialect Archive (https://archive.mith.umd.edu/mith-2020/vada/index.html) for audio-visual speech recognition (AVSR), speech recognition (ASR), and visual speech recognition/lip-reading (VSR). This version currently only contains read… See the full description on the dataset page: https://huggingface.co/datasets/Berkeley-NLP/visual_accent_dialect_archive.

sourceHugging Faceotherupdated 7mo agoView on Hugging Face
0likes7downloads
Dataset Card

No card is published for this repository, or it could not be fetched from Hugging Face right now.