CoolFace
22 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01Cnam-LMSSC /vibravox Dataset Card for VibraVox 👀 While waiting for the TooBigContentError issue to be resolved by the HuggingFace team, you can explore the dataset viewer of vibravox-test which has exactly the same architecture. DATASET SUMMARY The VibraVox dataset is a general purpose audio dataset of french speech captured with body-conduction transducers. This dataset can be used for various audio machine learning tasks : Automatic Speech Recognition (ASR) (Speech-to-Text… See the full description on the dataset page: https://huggingface.co/datasets/Cnam-LMSSC/vibravox.audioaudio-to-audio10K<n<100K30 likes5.8k downloads11mo agoHugging Face02Cnam-LMSSC /vibravox-test Dataset Card for Vibravox-test Important Note This dataset contains a very small proportion (1.2 %) of the original Vibravox Dataset. vibravox-test is a only a dummy dataset for use with test pipelines in the Vibravox project. It is therefore not intended for training or testing models. For full access to the complete dataset and documentation suitable for training and testing various audio and speech-related tasks, please visit the Vibravox Dataset page on Hugging Face.… See the full description on the dataset page: https://huggingface.co/datasets/Cnam-LMSSC/vibravox-test.audion<1K2 likes1.7k downloads2y agoHugging Face03Cnam-LMSSC /non_curated_vibravox Dataset Card for non-curated VibraVox 👀 This is the non-curated version of the VibraVox dataset. For a full documentation and dataset usage, please refer to https://huggingface.co/datasets/Cnam-LMSSC/vibravox DATASET SUMMARY The VibraVox dataset is a general purpose audio dataset of french speech captured with body-conduction transducers. This dataset can be used for various audio machine learning tasks : Automatic Speech Recognition (ASR) (Speech-to-Text… See the full description on the dataset page: https://huggingface.co/datasets/Cnam-LMSSC/non_curated_vibravox.audio10K<n<100K0 likes461 downloads1y agoHugging Face04Cnam-LMSSC /common_voice_13_french_phoneme Common Voice 13 French Phoneme Dataset Summary This dataset is a curated version of the French subset of Common Voice 13.0, enriched with a phonetic transcription column (phoneme). It was created by the Laboratoire de Mécanique des Structures et des Systèmes Couplés (Cnam-LMSSC) to support research in speech processing, specifically for tasks requiring phonetic alignment, phoneme recognition, and robust speech-to-text applications in French. The dataset retains the… See the full description on the dataset page: https://huggingface.co/datasets/Cnam-LMSSC/common_voice_13_french_phoneme.audioautomatic-speech-recognition100K<n<1M1 likes382 downloads9mo agoHugging Face05Cnam-LMSSC /shoebox_rir_updatedaudio100K<n<1M0 likes258 downloads11mo agoHugging Face06Cnam-LMSSC /french_librispeech_vibravoxed"Vibravoxed" version on the french split of facebook/multilingual_librispeech Deteriorated clean speech with reverse EBEN models: Cnam-LMSSC/EBEN_reverse_forehead_accelerometer Cnam-LMSSC/EBEN_reverse_rigid_in_ear_microphone Cnam-LMSSC/EBEN_reverse_soft_in_ear_microphone Cnam-LMSSC/EBEN_reverse_throat_microphone Cnam-LMSSC/EBEN_reverse_temple_vibration_pickupaudio100K<n<1M2 likes236 downloads2y agoHugging Face07Cnam-LMSSC /multilingual_librispeech_french_phoneme Multilingual LibriSpeech French Phoneme Dataset Summary This dataset is a curated version of the French subset of Multilingual LibriSpeech (MLS), enriched with a phonetic transcription column (phoneme). The Laboratoire de Mécanique des Structures et des Systèmes Couplés (Cnam-LMSSC) created this version to facilitate research into French acoustic modeling, phoneme recognition, and speech synthesis. It builds upon the high-quality audio derived from LibriVox audiobooks… See the full description on the dataset page: https://huggingface.co/datasets/Cnam-LMSSC/multilingual_librispeech_french_phoneme.audioautomatic-speech-recognition100K<n<1M1 likes138 downloads9mo agoHugging Face08Cnam-LMSSC /multilingual_librispeech_spanish_phoneme Multilingual LibriSpeech Spanish Phoneme Dataset Summary This dataset is a curated version of the Spanish subset of Multilingual LibriSpeech (MLS), enriched with a phonetic transcription column (phoneme). The Laboratoire de Mécanique des Structures et des Systèmes Couplés (Cnam-LMSSC) created this version to facilitate research into Spanish acoustic modeling, phoneme recognition, and speech synthesis. It builds upon the high-quality audio derived from LibriVox audiobooks… See the full description on the dataset page: https://huggingface.co/datasets/Cnam-LMSSC/multilingual_librispeech_spanish_phoneme.audioautomatic-speech-recognition100K<n<1M1 likes114 downloads7mo agoHugging Face09Cnam-LMSSC /shoebox_rir_with_room_namesaudio100K<n<1M1 likes99 downloads11mo agoHugging Face10Cnam-LMSSC /large_shoebox_riraudio100K<n<1M0 likes94 downloads11mo agoHugging Face11Cnam-LMSSC /vibravox_enhanced_by_EBEN Dataset Card Description This dataset features a speech-enhanced version of the test split from the speech_clean subset of the Vibravox Dataset. It is not intended for training. Enhancement procedure The Bandwidth extension task has been individually achieved for each sensor using configurable EBEN (arXiv link) models available at https://huggingface.co/Cnam-LMSSC/vibravox_EBEN_models. Ressources Results for speech-to-phoneme and speaker… See the full description on the dataset page: https://huggingface.co/datasets/Cnam-LMSSC/vibravox_enhanced_by_EBEN.audioaudio-to-audio1K<n<10K1 likes63 downloads2y agoHugging Face12Cnam-LMSSC /dynamic_kernels_riraudio1K<n<10K0 likes55 downloads11mo agoHugging Face13Bgeorge /CNA_AudioDatasetaudio10K<n<100K1 likes47 downloads11mo agoHugging Face14Cnam-LMSSC /french-mrtMale + Female recordings of the French version of the Modified Rhyme Test Microphones are the same as those used in Vibravox audion<1K0 likes40 downloads2y agoHugging Face15Cnam-LMSSC /vibravox_mixed_for_spkvaudio1K<n<10K0 likes40 downloads2y agoHugging Face16Cnam-LMSSC /multilingual_librispeech_italian_phoneme Multilingual LibriSpeech Italian Phoneme Dataset Summary This dataset is a curated version of the Italian subset of Multilingual LibriSpeech (MLS), enriched with a phonetic transcription column (phoneme). The Laboratoire de Mécanique des Structures et des Systèmes Couplés (Cnam-LMSSC) created this version to facilitate research into Italian acoustic modeling, phoneme recognition, and speech synthesis. It builds upon the high-quality audio derived from LibriVox audiobooks… See the full description on the dataset page: https://huggingface.co/datasets/Cnam-LMSSC/multilingual_librispeech_italian_phoneme.audioautomatic-speech-recognition10K<n<100K1 likes34 downloads7mo agoHugging Face17Bgeorge /CNA_AudioDataset3audio10K<n<100K0 likes31 downloads11mo agoHugging Face18Cnam-LMSSC /push_from_large_folderaudio10K<n<100K0 likes30 downloads11mo agoHugging Face19Bgeorge /CNA_AudioDataset2audio10K<n<100K0 likes27 downloads11mo agoHugging Face20Cnam-LMSSC /french-mrt_enhanced_by_EBENSame dataset as Cnam-LMSSC/french-mrt but enhanced by EBEN models audion<1K0 likes24 downloads2y agoHugging Face21Bgeorge /CNA_AudioDataset1audio10K<n<100K0 likes22 downloads11mo agoHugging Face22Cnam-LMSSC /shoebox_rirgatedaudio100K<n<1M0 likes2 downloads11mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.