datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
NatHEARPlease see LICENSE.txt for each dataset licenses
AISHELL6-Whisper
🗣️ AISHELL6-Whisper
AISHELL6-Whisper is a large-scale open-source Chinese Mandarin audio-visual whisper speech dataset,containing 30 hours each of whisper and parallel normal speech, with synchronized frontal RGB facial videos.
📘 Dataset Summary
Property
Description
Language
Chinese (Mandarin, ZH)
License
CC BY-NC-SA 4.0
Duration
~60 hours total (30 h whisper + 30 h normal)
Speakers
167 total (121 with RGB-D, 46 audio-only)
Environment
Controlled… See the full description on the dataset page: https://huggingface.co/datasets/SMIIP-lab/AISHELL6-Whisper.Nat-HEAR-Ambisonicsugspeech-akan-clean-100hrsNat-HEAR-localizationplskillmeiwantodie
