CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01hoangducanh1865 /internal-ast-eval-cacheaudion<1K0 likes537 downloads8d agoHugging Face02HoangPhuc7679 /RAVDESSaudio1K<n<10K1 likes395 downloads2y agoHugging Face03hoangbang /hey-computer-speech-commands Hey Computer: Speech Command Recognition Dataset Summary A public, viewer-ready educational challenge dataset. Host-only scoring data and hidden targets are excluded. Splits Split Examples Description train 13,192 Labeled training data test 3,295 Public inputs with withheld target labels or annotations Data Fields Field Type audio Audio id string label string (test sentinel: unlabeled)… See the full description on the dataset page: https://huggingface.co/datasets/hoangbang/hey-computer-speech-commands.audioaudio-classification10K<n<100K0 likes378 downloads2mo agoHugging Face04hoangbang /speak-the-digit Speak the Digit: Spoken Digit Recognition Dataset Summary A public, viewer-ready educational challenge dataset. Host-only scoring data and hidden targets are excluded. Splits Split Examples Description train 2,400 Labeled training data test 600 Public inputs with withheld target labels or annotations Data Fields Field Type audio Audio id string label string (test sentinel: unlabeled)… See the full description on the dataset page: https://huggingface.co/datasets/hoangbang/speak-the-digit.audioaudio-classification1K<n<10K0 likes323 downloads2mo agoHugging Face05hoangducanh1865 /m-meldaudio10K<n<100K0 likes136 downloads15d agoHugging Face06hr16 /kinh-phap-hoa-ke-trom-huongNormalized using https://github.com/oysterlanguage/emiliapipex @article{emilia, title={Emilia: An Extensive, Multilingual, and Diverse Speech Dataset for Large-Scale Speech Generation}, author={He, Haorui and Shang, Zengqiang and Wang, Chaoren and Li, Xuyuan and Gu, Yicheng and Hua, Hua and Liu, Liwei and Yang, Chen and Li, Jiaqi and Shi, Peiyang and Wang, Yuancheng and Chen, Kai and Zhang, Pengyuan and Wu, Zhizheng}, journal={arXiv}, volume={abs/2407.05361}… See the full description on the dataset page: https://huggingface.co/datasets/hr16/kinh-phap-hoa-ke-trom-huong.audiotext-to-audio10K<n<100K0 likes106 downloads2y agoHugging Face07hoangducanh1865 /vi-en-ast-testsetaudio1K<n<10K0 likes105 downloads20d agoHugging Face08hoangdeeptry /voice_dataset Dataset Card for "voice_dataset" More Information needed audio1K<n<10K0 likes100 downloads3y agoHugging Face09hoangdeeptry /tdtu_voice_dataset Dataset Card for "tdtu_voice_dataset" More Information needed audio1K<n<10K0 likes90 downloads3y agoHugging Face10hoangducanh1865 /bmeldaudio10K<n<100K0 likes79 downloads15d agoHugging Face11hoangdeeptry /cntt2-audio-dataset Dataset Card for "cntt2-audio-dataset" More Information needed audio1K<n<10K1 likes70 downloads3y agoHugging Face12hoangducanh1865 /en-vi-ast-testsetaudio1K<n<10K0 likes69 downloads16d agoHugging Face13hoangdeeptry /voice_dataset_yt Dataset Card for "voice_dataset_yt" More Information needed audio1K<n<10K0 likes54 downloads3y agoHugging Face14hoanglinhn0 /omnivoice-vie OmniVoice VI — Giọng Việt + SRT lồng tiếng Dataset chứa 6 giọng tiếng Việt và công cụ speak.py để chạy trên Google Colab với OmniVoice. Giọng có sẵn Slug Tên ban_mai Ban Mai lan_trinh Lan Trinh ngan_ha Ngan Ha ngoc_huyen Ngoc Huyen thao_trinh Thao Trinh tuong_vy Tuong Vy Mỗi giọng gồm profile.json, voice.pt (prompt cache), audio mẫu và ref_text.txt. Chạy trên Colab Mở notebook colab/Omivoice_VI_Colab.ipynb Đặt HF_REPO =… See the full description on the dataset page: https://huggingface.co/datasets/hoanglinhn0/omnivoice-vie.audion<1K0 likes38 downloads3mo agoHugging Face15hoangducanh1865 /internal-ast-eval-failuresaudio1K<n<10K0 likes35 downloads16d agoHugging Face16christian-hoang-04 /vivos-processed VIVOS Processed VoxCPM2 References This dataset contains 325 quality-first, nested voice-cloning references for the 65 speakers in the VIVOS Vietnamese corpus. Each speaker has nominal 5, 10, 15, 20, and 30-second variants. Whole source utterances are retained, so actual_seconds is the authoritative duration. Configuration references is directly usable for VoxCPM2 inference. Use audio as reference_wav_path; for transcript-assisted, highest-fidelity cloning, use… See the full description on the dataset page: https://huggingface.co/datasets/christian-hoang-04/vivos-processed.audiotext-to-speechn<1K0 likes34 downloads2mo agoHugging Face17namminh27 /sodo_hoaimy_newaudio10K<n<100K0 likes13 downloads11mo agoHugging Face18namminh27 /sodo_hoaimy_1audio10K<n<100K0 likes13 downloads11mo agoHugging Face19hoangphihung442004 /AV_FFIA3kaudio1K<n<10K0 likes12 downloads2mo agoHugging Face20hoangvanvietanh /my_dataset_testaudion<1K0 likes10 downloads3y agoHugging Face21hoangphihung442004 /MMFFIAaudio1M<n<10M0 likes8 downloads2mo agoHugging Face22hoangvanvietanh /user_5476d2c924204b6f9e38713118fdb9b2_datasetaudion<1K0 likes7 downloads3y agoHugging Face23hoangvanvietanh /user_03aa5df890b64866be4aef51a01c0a8a_datasetaudion<1K0 likes7 downloads3y agoHugging Face24hoangkhanh17 /AbstractTTSaudio0 likes6 downloads1y agoHugging Face25cuongtm /trinh_hoai_quang_tamaudion<1K0 likes5 downloads3y agoHugging Face26hoangvanvietanh /77r0dh_datasetaudion<1K0 likes5 downloads3y agoHugging Face27hoangvanvietanh /user_35621758bf084337aad673e1cc332d6f_datasetaudion<1K0 likes5 downloads3y agoHugging Face28hoangvanvietanh /user_da91d399b47141ccaa812c8b16e8c380_datasetaudion<1K0 likes5 downloads3y agoHugging Face29hoangvanvietanh /user_359d53fbd48b405daf1d7a67aed75197_datasetaudion<1K0 likes5 downloads3y agoHugging Face30hoangvanvietanh /user_476da26872df492f830a65925d422651_datasetaudion<1K0 likes5 downloads3y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.