CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01obadx /mualem-recitations-original المصاحف القرآنية مصاحف مجمعمة من القراء المتقنين لبناء نماذج ذكاء اصطناعي لخدمة القرآن الكريم. أنظر هنا لأكواد بناء قاعدة التلاوات القرآنية البيانات الوصفية للمصاحف ds = load_dataset('obadx/mualem-recitations-original', name='moshaf_metadata')['train'] وصف أوجه حفص Attribute Name Arabic Name Values Default Value More Info rewaya الرواية - hafs (حفص) The type of the quran Rewaya. recitation_speed سرعة التلاوة - mujawad (مجود)-… See the full description on the dataset page: https://huggingface.co/datasets/obadx/mualem-recitations-original.audion<1K0 likes2.6k downloads1y agoHugging Face02Ranjit /or_in_datasetaudioautomatic-speech-recognition10K<n<100K1 likes490 downloads3y agoHugging Face03origami-digital /in-the-grooveCompiled from several different sets of songs: (ITG) In the Groove (ITG) In the Groove 2 Songs were downloaded from https://search.stepmaniaonline.net/packs/in+the+groove and are stored here for persistence. In The Groove/ITG typically refers to DDR beatmaps done with an eye towards pad play. Dataset info: https://paperswithcode.com/dataset/itg audioaudio-classificationn<1K0 likes487 downloads3y agoHugging Face04jpalgo /tedlium-originalaudio10K<n<100K0 likes274 downloads2y agoHugging Face05otoearth /otoSpeech-full-duplex-task-oriented-20hgated Dataset Viewer https://cc-task-oriented-preview.vercel.app/ Task Walkthrough https://www.oto.earth/research/task-oriented-dataset.html What each of the seven tasks is for, what the two speakers could each see, and how the interaction log lines up with the audio. otoSpeech-full-duplex-task-oriented-20h Contact This sample dataset is provided for research purposes. We maintain larger and more diverse datasets. For collaborations… See the full description on the dataset page: https://huggingface.co/datasets/otoearth/otoSpeech-full-duplex-task-oriented-20h.audioaudio-to-audion<1K6 likes233 downloads25d agoHugging Face06whyismydininghallonfire /orig-plus-asr-tamil-clean orig-plus-asr-tamil-clean Combined ASR dataset built from: albagon/til26-asr-split (orig rows) whyismydininghallonfire/asr-tamil-clean (asr_tamil_clean rows) Audio paths are namespaced under each split to avoid filename collisions: audio/orig/... audio/asr_tamil_clean/... Each row keeps key, audio, transcript, and language, with an added source_dataset field. Counts: train: 3595 orig + 891 asr_tamil_clean = 4486 validation: 899 orig + 224 asr_tamil_clean = 1123 audio1K<n<10K0 likes216 downloads4mo agoHugging Face07saeedzou /iemocap-original-wavlm-large-layer-9-temporalaudio1K<n<10K0 likes168 downloads2mo agoHugging Face08gonnerthetooner /orislop-ami-lipsync-data Orislop AMI lip-sync source data This repository mirrors an acquisition-capped subset of the official AMI Meeting Corpus for reproducible Orislop audio-visual synchronization research. It contains: low-size AMI close-up AVI video streams; corresponding individual headset WAV streams; SHA-256 provenance records; the official camera/headset mapping snapshot and generated clip manifests after local acquisition completes. Files are uploaded only after a local download finishes and… See the full description on the dataset page: https://huggingface.co/datasets/gonnerthetooner/orislop-ami-lipsync-data.audiovideo-classificationn<1K0 likes152 downloads2mo agoHugging Face09saeedzou /iemocap-original-wavlm-layer-6-temporalaudio1K<n<10K0 likes138 downloads2mo agoHugging Face10SayantanJoker /original_data_manipuri_ttsaudio10K<n<100K0 likes134 downloads2y agoHugging Face11SayantanJoker /original_data_gujrati_ttsaudio1K<n<10K0 likes129 downloads2y agoHugging Face12arnauquest /original-songs Dataset Card for "original-songs" (Audio + análisis DSP) Dataset Summary Dataset pequeño de canciones originales creadas con IA, cada una con su WAV, letra transcrita automáticamente (Whisper) y un análisis DSP completo (tempo, tonalidad, loudness, features perceptuales) además de detección de contenido explícito. Pensado para quien quiera mejorar modelos open source: extracción de features musicales, clasificación de audio, transcripción y moderación de letras.… See the full description on the dataset page: https://huggingface.co/datasets/arnauquest/original-songs.audioaudio-classificationn<1K1 likes125 downloads1mo agoHugging Face13Lawrence /Ndizi_parler_data_prep_originalaudio1K<n<10K0 likes123 downloads2y agoHugging Face14SayantanJoker /original_data_tamil_ttsaudio1K<n<10K0 likes109 downloads2y agoHugging Face15SayantanJoker /original_data_bengali_ttsaudio10K<n<100K1 likes106 downloads2y agoHugging Face16SayantanJoker /original_data_odia_ttsaudio10K<n<100K0 likes102 downloads2y agoHugging Face17SayantanJoker /original_data_kannada_ttsaudio1K<n<10K0 likes100 downloads2y agoHugging Face18Lawrence /Ndizi_TTS_origaudio1K<n<10K0 likes99 downloads2y agoHugging Face19SayantanJoker /original_data_telegu_ttsaudio1K<n<10K0 likes98 downloads2y agoHugging Face20fosters /ivan_shamyakin_tryvozhnae_shchastse_output_original Трывожнае шчасце — арыгінальнае аўдыё Аўтар / Author: Іван ШамякінМова / Language: Беларуская (Belarusian) Арыгінальнае аўдыё без апрацоўкі, захаванае ў зыходнай якасці. Частка калекцыі Ministerskija — корпус беларускіх аўдыёкніг. Апрацаваная версія (сегменты ~15 с, выраўнаваная транскрыпцыя): ivan_shamyakin_tryvozhnae_shchastse_output Доўгасць аўдыё 30h45m Радкоў у датасеце 6,523 Структура Кожны радок змяшчае: audio — арыгінальны аўдыёзапіс… See the full description on the dataset page: https://huggingface.co/datasets/fosters/ivan_shamyakin_tryvozhnae_shchastse_output_original.audioautomatic-speech-recognition1K<n<10K0 likes97 downloads4mo agoHugging Face21nour-world /mualem-recitations-original المصاحف القرآنية مصاحف مجمعمة من القراء المتقنين لبناء نماذج ذكاء اصطناعي لخدمة القرآن الكريم. أنظر هنا لأكواد بناء قاعدة التلاوات القرآنية البيانات الوصفية للمصاحف ds = load_dataset('obadx/mualem-recitations-original', name='moshaf_metadata')['train'] وصف أوجه حفص Attribute Name Arabic Name Values Default Value More Info rewaya الرواية - hafs (حفص) The type of the quran Rewaya. recitation_speed سرعة التلاوة - mujawad (مجود)-… See the full description on the dataset page: https://huggingface.co/datasets/nour-world/mualem-recitations-original.audion<1K0 likes97 downloads10d agoHugging Face22SayantanJoker /original_data_assamese_ttsaudio10K<n<100K0 likes93 downloads2y agoHugging Face23Codec-SUPERB /Voxceleb1_test_original Dataset Card for "Voxceleb1" More Information needed audio1K<n<10K0 likes92 downloads3y agoHugging Face24SayantanJoker /original_data_malayalam_ttsaudio10K<n<100K0 likes91 downloads2y agoHugging Face25SayantanJoker /original_data_rajasthani_ttsaudio1K<n<10K0 likes89 downloads2y agoHugging Face26fosters /kuzma_chorny_zyamlya_output_original Зямля — арыгінальнае аўдыё Аўтар / Author: Кузьма ЧорныМова / Language: Беларуская (Belarusian) Арыгінальнае аўдыё без апрацоўкі, захаванае ў зыходнай якасці. Частка калекцыі Ministerskija — корпус беларускіх аўдыёкніг. Апрацаваная версія (сегменты ~15 с, выраўнаваная транскрыпцыя): kuzma_chorny_zyamlya_output Доўгасць аўдыё 28h15m Радкоў у датасеце 6,820 Структура Кожны радок змяшчае: audio — арыгінальны аўдыёзапіс text — транскрыпцыя… See the full description on the dataset page: https://huggingface.co/datasets/fosters/kuzma_chorny_zyamlya_output_original.audioautomatic-speech-recognition1K<n<10K0 likes86 downloads4mo agoHugging Face27SayantanJoker /original_data_marathi_ttsaudio10K<n<100K0 likes83 downloads2y agoHugging Face28fosters /kuzma_chorny_poshuki_buduchyni_output_original Пошукі будучыні — арыгінальнае аўдыё Аўтар / Author: Кузьма ЧорныМова / Language: Беларуская (Belarusian) Арыгінальнае аўдыё без апрацоўкі, захаванае ў зыходнай якасці. Частка калекцыі Ministerskija — корпус беларускіх аўдыёкніг. Апрацаваная версія (сегменты ~15 с, выраўнаваная транскрыпцыя): kuzma_chorny_poshuki_buduchyni_output Доўгасць аўдыё 22h58m Радкоў у датасеце 5,530 Структура Кожны радок змяшчае: audio — арыгінальны аўдыёзапіс text —… See the full description on the dataset page: https://huggingface.co/datasets/fosters/kuzma_chorny_poshuki_buduchyni_output_original.audioautomatic-speech-recognition1K<n<10K0 likes75 downloads4mo agoHugging Face29vozes-da-cabeca /wpp_pav_originalaudio100K<n<1M0 likes69 downloads1y agoHugging Face30pratikk-003 /rasa_san_originalaudio10K<n<100K0 likes69 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.