ud-nlp/american-speech-recognition-dataset
American Telephone Dialogues Dataset - 1,136 Hours The dataset includes 1,136 hours of annotated telephone dialogues from 1,416 native speakers across the United States. Designed for advancing speech recognition models and language processing, this extensive speech data corpus covers diverse topics and domains, making it ideal for training robust automatic speech recognition (ASR) systems. - Get the data Dataset characteristics: Characteristic Data… See the full description on the dataset page: https://huggingface.co/datasets/ud-nlp/american-speech-recognition-dataset.
American Telephone Dialogues Dataset - 1,136 Hours
The dataset includes 1,136 hours of annotated telephone dialogues from 1,416 native speakers across the United States. Designed for advancing speech recognition models and language processing, this extensive speech data corpus covers diverse topics and domains, making it ideal for training robust automatic speech recognition (ASR) systems. - [Get the data](https://unidata.pro/datasets/american-speech-recognition-dataset/?utm_source=huggingface-nlp&utm_medium=referral&utm_campaign=american-speech-recognition-dataset)
Dataset characteristics:
📊 Sample dataset available! For full access, contact us to discuss purchase terms.
Dataset structure
- audio - audio file
- text - text transcription
- American Speech Recognition Dataset.csv - metadata for the data
🧩 Like the dataset but need different data? We can collect a custom dataset just for you - learn more about our data collection services here
Similar Datasets:
- British English Speech Recognition Dataset
- French Speech Recognition Dataset
- Speech Emotion Recognition Dataset
