CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01lmms-lab /EgoIT-99KCheckout the paper EgoLife (https://arxiv.org/abs/2503.03803) for more information. audio100K<n<1M9 likes22k downloads2y agoHugging Face02Cnam-LMSSC /vibravox Dataset Card for VibraVox 👀 While waiting for the TooBigContentError issue to be resolved by the HuggingFace team, you can explore the dataset viewer of vibravox-test which has exactly the same architecture. DATASET SUMMARY The VibraVox dataset is a general purpose audio dataset of french speech captured with body-conduction transducers. This dataset can be used for various audio machine learning tasks : Automatic Speech Recognition (ASR) (Speech-to-Text… See the full description on the dataset page: https://huggingface.co/datasets/Cnam-LMSSC/vibravox.audioaudio-to-audio10K<n<100K30 likes5.5k downloads11mo agoHugging Face03lmms-lab-audio /mmauaudio10K<n<100K2 likes2.7k downloads2y agoHugging Face04lmms-lab-audio /librispeechaudio10K<n<100K5 likes2.2k downloads2y agoHugging Face05lmms-lab-audio /voicebench License The dataset is available under the Apache 2.0 license. Citation If you use the VoiceBench dataset in your research, please cite the following paper: @article{chen2024voicebench, title={VoiceBench: Benchmarking LLM-Based Voice Assistants}, author={Chen, Yiming and Yue, Xianghu and Zhang, Chen and Gao, Xiaoxue and Tan, Robby T. and Li, Haizhou}, journal={arXiv preprint arXiv:2410.17196}, year={2024} } audio10K<n<100K1 likes1.9k downloads1y agoHugging Face06Cnam-LMSSC /vibravox-test Dataset Card for Vibravox-test Important Note This dataset contains a very small proportion (1.2 %) of the original Vibravox Dataset. vibravox-test is a only a dummy dataset for use with test pipelines in the Vibravox project. It is therefore not intended for training or testing models. For full access to the complete dataset and documentation suitable for training and testing various audio and speech-related tasks, please visit the Vibravox Dataset page on Hugging Face.… See the full description on the dataset page: https://huggingface.co/datasets/Cnam-LMSSC/vibravox-test.audion<1K2 likes1.7k downloads2y agoHugging Face07lmms-lab /WenetSpeech_tempaudio10K<n<100K0 likes1.5k downloads1y agoHugging Face08lmms-eval /PerceptionTest_Valaudio10K<n<100K1 likes1.4k downloads2y agoHugging Face09lmms-lab /WorldSenseaudio1K<n<10K2 likes1.3k downloads2y agoHugging Face10lmms-lab-audio /gigaspeechaudio100K<n<1M1 likes1.3k downloads2y agoHugging Face11lmms-lab /AIR_benchaudio10K<n<100K0 likes694 downloads2y agoHugging Face12lmms-lab-audio /ClothoAQAaudio1K<n<10K0 likes549 downloads2y agoHugging Face13lmms-lab-audio /Omni_Bench_fixaudio1K<n<10K1 likes518 downloads1y agoHugging Face14lmms-lab-audio /common_voice_15audio10K<n<100K1 likes508 downloads2y agoHugging Face15lmms-eval /PerceptionTestaudio10K<n<100K2 likes491 downloads2y agoHugging Face16CASIA-LM /OpenS2S_Datasets How to Use? Download, merge the files, and extract You can run the following command to merge the compressed file parts after downloading. cat en_response_wav.tar.gz.* > en_response_wav.tar.gz cat zh_response_wav.tar.gz.* > zh_response_wav.tar.gz audio100K<n<1M8 likes461 downloads1y agoHugging Face17lmms-lab-audio /covost2_en-zhaudio100K<n<1M0 likes448 downloads2y agoHugging Face18Cnam-LMSSC /non_curated_vibravox Dataset Card for non-curated VibraVox 👀 This is the non-curated version of the VibraVox dataset. For a full documentation and dataset usage, please refer to https://huggingface.co/datasets/Cnam-LMSSC/vibravox DATASET SUMMARY The VibraVox dataset is a general purpose audio dataset of french speech captured with body-conduction transducers. This dataset can be used for various audio machine learning tasks : Automatic Speech Recognition (ASR) (Speech-to-Text… See the full description on the dataset page: https://huggingface.co/datasets/Cnam-LMSSC/non_curated_vibravox.audio10K<n<100K0 likes412 downloads1y agoHugging Face19lmms-lab-audio /muchomusicDataset Summary MuChoMusic is a benchmark designed to evaluate music understanding in multimodal audio-language models (Audio LLMs). The dataset comprises 1,187 multiple-choice questions created from 644 music tracks, sourced from two publicly available music datasets: MusicCaps and the Song Describer Dataset (SDD). The questions test knowledge and reasoning abilities across dimensions such as music theory, cultural context, and functional applications. All questions and answers have been… See the full description on the dataset page: https://huggingface.co/datasets/lmms-lab-audio/muchomusic.audio1K<n<10K1 likes365 downloads2y agoHugging Face20Cnam-LMSSC /common_voice_13_french_phoneme Common Voice 13 French Phoneme Dataset Summary This dataset is a curated version of the French subset of Common Voice 13.0, enriched with a phonetic transcription column (phoneme). It was created by the Laboratoire de Mécanique des Structures et des Systèmes Couplés (Cnam-LMSSC) to support research in speech processing, specifically for tasks requiring phonetic alignment, phoneme recognition, and robust speech-to-text applications in French. The dataset retains the… See the full description on the dataset page: https://huggingface.co/datasets/Cnam-LMSSC/common_voice_13_french_phoneme.audioautomatic-speech-recognition100K<n<1M1 likes351 downloads8mo agoHugging Face21Cnam-LMSSC /shoebox_rir_updatedaudio100K<n<1M0 likes258 downloads11mo agoHugging Face22lmms-lab-audio /vocalsoundaudio1K<n<10K7 likes254 downloads2y agoHugging Face23dhlee3000 /LMD-AI-Detection LMD AI-Generated Music Detection Benchmark (Note: The corresponding research paper will be released later.) Dataset Description The rapid advancement of AI music generation has raised growing concerns about the authenticity of digital music. While deepfake detection has been extensively studied in the audio domain, symbolic music (MIDI) remains largely unexplored. This dataset presents a comprehensive benchmark for AI-generated symbolic music detection, examining… See the full description on the dataset page: https://huggingface.co/datasets/dhlee3000/LMD-AI-Detection.audioaudio-classification10K<n<100K2 likes250 downloads11d agoHugging Face24Cnam-LMSSC /french_librispeech_vibravoxed"Vibravoxed" version on the french split of facebook/multilingual_librispeech Deteriorated clean speech with reverse EBEN models: Cnam-LMSSC/EBEN_reverse_forehead_accelerometer Cnam-LMSSC/EBEN_reverse_rigid_in_ear_microphone Cnam-LMSSC/EBEN_reverse_soft_in_ear_microphone Cnam-LMSSC/EBEN_reverse_throat_microphone Cnam-LMSSC/EBEN_reverse_temple_vibration_pickupaudio100K<n<1M2 likes228 downloads2y agoHugging Face25lmms-lab-audio /peoples_speechaudio10K<n<100K0 likes200 downloads2y agoHugging Face26lmejias /syntheticaudion<1K0 likes164 downloads1mo agoHugging Face27Cnam-LMSSC /large_shoebox_riraudio100K<n<1M0 likes143 downloads11mo agoHugging Face28lmms-lab-audio /fleursaudio1K<n<10K0 likes140 downloads2y agoHugging Face29Cnam-LMSSC /multilingual_librispeech_french_phoneme Multilingual LibriSpeech French Phoneme Dataset Summary This dataset is a curated version of the French subset of Multilingual LibriSpeech (MLS), enriched with a phonetic transcription column (phoneme). The Laboratoire de Mécanique des Structures et des Systèmes Couplés (Cnam-LMSSC) created this version to facilitate research into French acoustic modeling, phoneme recognition, and speech synthesis. It builds upon the high-quality audio derived from LibriVox audiobooks… See the full description on the dataset page: https://huggingface.co/datasets/Cnam-LMSSC/multilingual_librispeech_french_phoneme.audioautomatic-speech-recognition100K<n<1M1 likes139 downloads8mo agoHugging Face30AV-Odyssey /AV_Odyssey_Bench_LMMs_Evalaudio1K<n<10K1 likes113 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.