CoolFace
Datasetpublicgated

Berkeley-NLP/visual_accent_dialect_archive

Source: https://www.youtube.com/@visualaccent/videos All rights belong to the original dataset creator. VADA-AVSR: an audio-visual dataset of non-native English ("accents") and English varieties ("dialects") We preprocessed the Visual Accent and Dialect Archive (https://archive.mith.umd.edu/mith-2020/vada/index.html) for audio-visual speech recognition (AVSR), speech recognition (ASR), and visual speech recognition/lip-reading (VSR). This version currently only contains read… See the full description on the dataset page: https://huggingface.co/datasets/Berkeley-NLP/visual_accent_dialect_archive.

sourceHugging Faceotherupdated 7mo agoView on Hugging Face
0likes7downloads
settings

This repository belongs to Berkeley-NLP on Hugging Face.

CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.

namevisual_accent_dialect_archive
visibilitypublic
licenceother
gatedyes
ownerBerkeley-NLP
Account settings
Berkeley-NLP/visual_accent_dialect_archive · CoolFace