BabelSpeech/40hours_Indonesian_Colloquial_ASR_Speech_Dataset
BabelSpeech 50-Hour Indonesian Colloquial ASR Speech Dataset Contains 50 hours of Indonesian colloquial ASR speech data, aligned with natural, everyday Indonesian communication patterns. Metadata is stored in a separate JSON file, including audio path, duration, transcript confidence, signal-to-noise ratio (SNR), and DNSMOS. More metadata fields may be added in future updates. Covered domains: technology, entertainment, travel, education, daily life, and others. Data quality:… See the full description on the dataset page: https://huggingface.co/datasets/BabelSpeech/40hours_Indonesian_Colloquial_ASR_Speech_Dataset.
017
No card is published for this repository, or it could not be fetched from Hugging Face right now.
