datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
welsh-speech-audio
Welsh Speech Dataset - Audio
Audio recordings from the Welsh Speech Dataset.
Contents
33 speakers x 10 Welsh phrases ~ 330 audio files
Format: WAV (16-bit PCM recommended)
with 3D facial captures and landmarks
Files
Audio files are located in the audio/ directory.
Naming: audio/speaker_XX_phrase_YY.wav
Example: audio/speaker_01_phrase_05.wav = Speaker 1 speaking Phrase 5 ("Ardderchog")
Metadata
Metadata for the audio dataset is in metadata.jsonl… See the full description on the dataset page: https://huggingface.co/datasets/arvinsingh/welsh-speech-audio.welsh-speech-dataset
Welsh Speech Dataset
A multimodal dataset of 33 speakers producing 10 Welsh phrases, captured using 3DMD technology with audio and dense facial landmark annotations.
Dataset Overview
Speakers: 33 participants
Phrases: 10 Welsh phrases per speaker
Sequences: ~330 (33 speakers x 10 phrases)
Modalities:
Audio recordings (.wav)
3D facial reconstructions (.obj meshes + texture maps)
68-point facial landmarks (ibug68 template)
Fluency Scores: Each phrase rated 0-5 (5 =… See the full description on the dataset page: https://huggingface.co/datasets/arvinsingh/welsh-speech-dataset.welsh-speech-dataset
Welsh Speech Dataset
A multimodal dataset of 33 speakers producing 10 Welsh phrases, captured using 3DMD technology with audio and dense facial landmark annotations.
Dataset Overview
Speakers: 33 participants
Phrases: 10 Welsh phrases per speaker
Sequences: ~330 (33 speakers x 10 phrases)
Modalities:
Audio recordings (.wav)
3D facial reconstructions (.obj meshes + texture maps)
68-point facial landmarks (ibug68 template)
Fluency Scores: Each phrase rated 0-5… See the full description on the dataset page: https://huggingface.co/datasets/Pedramebd/welsh-speech-dataset.
