CoolFace
Datasetpublic

ud-nlp/spanish-speech-recognition-dataset

Spanish Telephone Dialogues Dataset - 10 Hours Dataset comprises 10 hours of high-quality telephone audio recordings in Spanish, featuring 20 native speakers. Designed for advancing speech recognition models and language processing, this extensive speech data corpus covers diverse topics and domains, making it ideal for training robust automatic speech recognition (ASR) systems. - Get the data Dataset characteristics: Characteristic Data Description… See the full description on the dataset page: https://huggingface.co/datasets/ud-nlp/spanish-speech-recognition-dataset.

sourceHugging Facecc-by-nc-nd-4.0updated 10mo agoView on Hugging Face
0likes16downloads
Dataset Card

Spanish Telephone Dialogues Dataset - 10 Hours

Dataset comprises 10 hours of high-quality telephone audio recordings in Spanish, featuring 20 native speakers. Designed for advancing speech recognition models and language processing, this extensive speech data corpus covers diverse topics and domains, making it ideal for training robust automatic speech recognition (ASR) systems. - [Get the data](https://unidata.pro/datasets/spanish-speech-recognition-dataset/?utm_source=huggingface-nlp&utm_medium=referral&utm_campaign=spanish-speech-recognition-dataset)

Dataset characteristics:

CharacteristicData
DescriptionAudio of telephone dialogues in Spanish for training NLP models in real-world conversational scenarios.
Data typesAudio
TasksSpeech recognition, NLP
CountrySpain (ESP)
Hours of telephone dialogue10
Number of speakers20
LabelingAnnotation (ID, Language, Format, Minutes)
Recording deviceTelephone

📊 Sample dataset available! For full access, contact us to discuss purchase terms.

Dataset structure

  • —audio.mp3 - audio file
  • —Spanish Speech Recognition.csv - metadata for the data

🧩 Like the dataset but need different data? We can collect a custom dataset just for you - learn more about our data collection services here

Similar Datasets:

  1. 1.Portuguese Speech Recognition Dataset
  2. 2.American Speech Recognition Dataset
  3. 3.French Speech Recognition Dataset

🌐 UniData - your trusted data partner. Unique, accurate, thoroughly collected and annotated data designed to fuel your AI/ML success.