ud-nlp/british-english-speech-recognition-dataset
British English Telephone Dialogues Dataset - 200 Hours The dataset consists of 200 hours of high-quality telephone dialogues from 310 native speakers in the UK, with detailed annotations (transcriptions, timestamps, speaker ID, gender, and background noise) to support speech recognition systems, NLP tasks, and machine learning models requiring diverse British English audio datasets. - Get the data Dataset characteristics: Characteristic Data… See the full description on the dataset page: https://huggingface.co/datasets/ud-nlp/british-english-speech-recognition-dataset.
British English Telephone Dialogues Dataset - 200 Hours
The dataset consists of 200 hours of high-quality telephone dialogues from 310 native speakers in the UK, with detailed annotations (transcriptions, timestamps, speaker ID, gender, and background noise) to support speech recognition systems, NLP tasks, and machine learning models requiring diverse British English audio datasets. - [Get the data](https://unidata.pro/datasets/british-english-speech-recognition-dataset/?utm_source=huggingface-nlp&utm_medium=referral&utm_campaign=british-english-speech-recognition-dataset)
Dataset characteristics:
📊 Sample dataset available! For full access, contact us to discuss purchase terms.
Dataset structure
- audio - audio file
- text - text transcription
- British English Speech Recognition Dataset.csv - metadata for the data
