CoolFace
Datasetpublicgated

MML-Group/AVE-Speech

AVE Speech: A Comprehensive Multi-Modal Dataset for Speech Recognition Integrating Audio, Visual, and Electromyographic Signals Abstract AVE Speech is a large-scale Mandarin speech corpus that pairs synchronized audio, lip video and surface electromyography (EMG) recordings. The dataset contains 100 sentences read by 100 native speakers. Each participant repeated the full corpus ten times, yielding over 55 hours of data per modality. These complementary signals… See the full description on the dataset page: https://huggingface.co/datasets/MML-Group/AVE-Speech.

sourceHugging Facecc-by-nc-sa-4.0updated 1y agoView on Hugging Face
7likes425downloads

MML-Group/AVE-Speech · main · files are served by the source, never re-hosted here

This repository is gated. The listing is public, but downloading a file means accepting the publisher’s terms at Hugging Face first — the links above take you there rather than around it.