ud-nlp/japanese-speech-recognition-dataset
Japanese Telephone Dialogues Dataset - 10 Hours Dataset comprises 10 hours of high-quality telephone audio recordings in Japanese, featuring 20+ native speakers and achieving a 95% sentence accuracy rate. Designed for advancing speech recognition models and language processing, this extensive speech data corpus covers diverse topics and domains, making it ideal for training robust automatic speech recognition (ASR) systems. - Get the data Dataset characteristics:… See the full description on the dataset page: https://huggingface.co/datasets/ud-nlp/japanese-speech-recognition-dataset.
Japanese Telephone Dialogues Dataset - 10 Hours
Dataset comprises 10 hours of high-quality telephone audio recordings in Japanese, featuring 20+ native speakers and achieving a 95% sentence accuracy rate. Designed for advancing speech recognition models and language processing, this extensive speech data corpus covers diverse topics and domains, making it ideal for training robust automatic speech recognition (ASR) systems. - [Get the data](https://unidata.pro/datasets/japanese-speech-recognition-dataset/?utm_source=huggingface-nlp&utm_medium=referral&utm_campaign=japanese-speech-recognition-dataset)
Dataset characteristics:
📊 Sample dataset available! For full access, contact us to discuss purchase terms.
Dataset structure
- audio.mp3 - audio file
- Japanese Speech Recognition Dataset.csv - metadata for the data
