datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
confusable-100-lipreading
Confusable-100: an English vocabulary built to break lip reading
A small, deliberately adversarial corpus for probing one specific failure mode of
visual speech recognition: consonants articulated by the tongue leave no distinctive
trace on the lips, so words differing only in those consonants are not separable from
video in principle, not merely in practice.
This is an evaluation probe, not a training set. It is one speaker and 3.4 minutes
of speech. Its purpose is to expose… See the full description on the dataset page: https://huggingface.co/datasets/diddmstjr/confusable-100-lipreading.lipreading-wordslip-reading_data3
