CoolFace
8 results

conversational-speech

HTH-inc /japanese-casual-conversational-speech-golden-dataset-preview Japanese Casual Conversational Speech Golden Dataset (Preview) 💼 Commercial License & Full Access This repository contains a limited preview. The full 60-hour dataset collected via the "Kataro" app is available for commercial use, ASR benchmarking, and Spoken Dialogue Model fine-tuning. To purchase the full dataset, please contact us: 👉 Email: info@hth-inc.com 👉 Website: https://hth-inc.com/business 🌟 4 Reasons to Choose This Dataset… See the full description on the dataset page: https://huggingface.co/datasets/HTH-inc/japanese-casual-conversational-speech-golden-dataset-preview.audioautomatic-speech-recognitionn<1K2 likes246 downloads22d agoHugging FaceMagicHub /korean-conversational-speech-corpus ASR-KCSC: A Korean Conversational Speech Corpus Every data point counts. Dataset Basic Info Dataset Type: ASR Speech Corpus Language: Korean Audio Parameters: 16 kHz, 16 bits File Format: WAV (PCM) Recording Equipment: Mobile device Recording Environment: Indoor Dataset Description This open-source dataset consists of 5.22 hours of transcribed Korean conversational speech on certain topics, where 22 conversations between seven pairs of speakers… See the full description on the dataset page: https://huggingface.co/datasets/MagicHub/korean-conversational-speech-corpus.audio1K<n<10K1 likes223 downloads3mo agoHugging FaceTingChen-ppmc /Shanghai_Dialect_Conversational_Speech_Corpus Corpus This dataset is built from Magicdata ASR-CZDIACSC: A CHINESE SHANGHAI DIALECT CONVERSATIONAL SPEECH CORPUS This corpus is licensed under a Creative Commons Attribution-NonCommercial-NoDerivatives 4.0 International License. Please refer to the license for further information. Modifications: The audio is split in sentences based on the time span on the transcription file. Sentences that span less than 1 second is discarded. Topics of conversation is removed. Usage… See the full description on the dataset page: https://huggingface.co/datasets/TingChen-ppmc/Shanghai_Dialect_Conversational_Speech_Corpus.audio1K<n<10K12 likes160 downloads2y agoHugging Facejml2026 /conversational-speech-dataset 🎙️ Silencio Network: Conversational Speech Dataset Overview Sample conversational speech data from Silencio Network's crowdsourced voice AI platform. This dataset contains multi-speaker meeting recordings with word-level transcripts, speaker diarization, and rich demographic metadata. Each row represents one participant in a meeting and includes 3 audio files: Audio Column Description Format file_name (speaker audio) Individual participant's… See the full description on the dataset page: https://huggingface.co/datasets/jml2026/conversational-speech-dataset.automatic-speech-recognitionn<1K0 likes135 downloads6mo agoHugging Facemalaysia-ai /malay-conversational-speech-corpus malay-conversational-speech-corpus Mirror for https://magichub.com/datasets/malay-conversational-speech-corpus/, license is Creative Commons Attribution-NonCommercial-NoDerivatives 4.0 International License audio1K<n<10K5 likes110 downloads3y agoHugging FaceTingChen-ppmc /Changsha_Dialect_Conversational_Speech_Corpus Corpus This dataset is built from Magicdata ASR-CCHSHDIACSC: A CHINESE CHANGSHA DIALECT CONVERSATIONAL SPEECH CORPUS This corpus is licensed under a Creative Commons Attribution-NonCommercial-NoDerivatives 4.0 International License. Please refer to the license for further information. Modifications: The audio is split in sentences based on the time span on the transcription file. Sentences that span less than 1 second is discarded. Topics of conversation is removed.… See the full description on the dataset page: https://huggingface.co/datasets/TingChen-ppmc/Changsha_Dialect_Conversational_Speech_Corpus.audio1K<n<10K2 likes103 downloads3y agoHugging Face