datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
vibravox
Dataset Card for VibraVox
👀 While waiting for the TooBigContentError issue to be resolved by the HuggingFace team, you can explore the dataset viewer of vibravox-test
which has exactly the same architecture.
DATASET SUMMARY
The VibraVox dataset is a general purpose audio dataset of french speech captured with body-conduction transducers.
This dataset can be used for various audio machine learning tasks :
Automatic Speech Recognition (ASR) (Speech-to-Text… See the full description on the dataset page: https://huggingface.co/datasets/Cnam-LMSSC/vibravox.vibravox-test
Dataset Card for Vibravox-test
Important Note
This dataset contains a very small proportion (1.2 %) of the original Vibravox Dataset.
vibravox-test is a only a dummy dataset for use with test pipelines in the Vibravox project. It is therefore not intended for training or testing models.
For full access to the complete dataset and documentation suitable for training and testing various audio and speech-related tasks, please visit the Vibravox Dataset page on Hugging Face.… See the full description on the dataset page: https://huggingface.co/datasets/Cnam-LMSSC/vibravox-test.non_curated_vibravox
Dataset Card for non-curated VibraVox
👀 This is the non-curated version of the VibraVox dataset. For a full documentation and dataset usage, please refer to https://huggingface.co/datasets/Cnam-LMSSC/vibravox
DATASET SUMMARY
The VibraVox dataset is a general purpose audio dataset of french speech captured with body-conduction transducers.
This dataset can be used for various audio machine learning tasks :
Automatic Speech Recognition (ASR) (Speech-to-Text… See the full description on the dataset page: https://huggingface.co/datasets/Cnam-LMSSC/non_curated_vibravox.common_voice_13_french_phoneme
Common Voice 13 French Phoneme
Dataset Summary
This dataset is a curated version of the French subset of Common Voice 13.0, enriched with a phonetic transcription column (phoneme).
It was created by the Laboratoire de Mécanique des Structures et des Systèmes Couplés (Cnam-LMSSC) to support research in speech processing, specifically for tasks requiring phonetic alignment, phoneme recognition, and robust speech-to-text applications in French.
The dataset retains the… See the full description on the dataset page: https://huggingface.co/datasets/Cnam-LMSSC/common_voice_13_french_phoneme.shoebox_rir_updatedfrench_librispeech_vibravoxed"Vibravoxed" version on the french split of facebook/multilingual_librispeech
Deteriorated clean speech with reverse EBEN models:
Cnam-LMSSC/EBEN_reverse_forehead_accelerometer
Cnam-LMSSC/EBEN_reverse_rigid_in_ear_microphone
Cnam-LMSSC/EBEN_reverse_soft_in_ear_microphone
Cnam-LMSSC/EBEN_reverse_throat_microphone
Cnam-LMSSC/EBEN_reverse_temple_vibration_pickupmultilingual_librispeech_french_phoneme
Multilingual LibriSpeech French Phoneme
Dataset Summary
This dataset is a curated version of the French subset of Multilingual LibriSpeech (MLS), enriched with a phonetic transcription column (phoneme).
The Laboratoire de Mécanique des Structures et des Systèmes Couplés (Cnam-LMSSC) created this version to facilitate research into French acoustic modeling, phoneme recognition, and speech synthesis. It builds upon the high-quality audio derived from LibriVox audiobooks… See the full description on the dataset page: https://huggingface.co/datasets/Cnam-LMSSC/multilingual_librispeech_french_phoneme.multilingual_librispeech_spanish_phoneme
Multilingual LibriSpeech Spanish Phoneme
Dataset Summary
This dataset is a curated version of the Spanish subset of Multilingual LibriSpeech (MLS), enriched with a phonetic transcription column (phoneme).
The Laboratoire de Mécanique des Structures et des Systèmes Couplés (Cnam-LMSSC) created this version to facilitate research into Spanish acoustic modeling, phoneme recognition, and speech synthesis. It builds upon the high-quality audio derived from LibriVox audiobooks… See the full description on the dataset page: https://huggingface.co/datasets/Cnam-LMSSC/multilingual_librispeech_spanish_phoneme.shoebox_rir_with_room_nameslarge_shoebox_rirvibravox_enhanced_by_EBEN
Dataset Card
Description
This dataset features a speech-enhanced version of the test split from the speech_clean subset of the Vibravox Dataset.
It is not intended for training.
Enhancement procedure
The Bandwidth extension task has been individually achieved for each sensor using configurable EBEN (arXiv link) models available at https://huggingface.co/Cnam-LMSSC/vibravox_EBEN_models.
Ressources
Results for speech-to-phoneme and speaker… See the full description on the dataset page: https://huggingface.co/datasets/Cnam-LMSSC/vibravox_enhanced_by_EBEN.dynamic_kernels_rirCNA_AudioDatasetfrench-mrtMale + Female recordings of the French version of the Modified Rhyme Test
Microphones are the same as those used in Vibravox
vibravox_mixed_for_spkvmultilingual_librispeech_italian_phoneme
Multilingual LibriSpeech Italian Phoneme
Dataset Summary
This dataset is a curated version of the Italian subset of Multilingual LibriSpeech (MLS), enriched with a phonetic transcription column (phoneme).
The Laboratoire de Mécanique des Structures et des Systèmes Couplés (Cnam-LMSSC) created this version to facilitate research into Italian acoustic modeling, phoneme recognition, and speech synthesis. It builds upon the high-quality audio derived from LibriVox audiobooks… See the full description on the dataset page: https://huggingface.co/datasets/Cnam-LMSSC/multilingual_librispeech_italian_phoneme.CNA_AudioDataset3push_from_large_folderCNA_AudioDataset2french-mrt_enhanced_by_EBENSame dataset as Cnam-LMSSC/french-mrt but enhanced by EBEN models
CNA_AudioDataset1shoebox_rir
