Mexican_Spanish
Ald_Mexican_Spanish_speech_datasetThis dataset can be used to fine-tune Speech To Text models as Text To Speech.
dataset information
Speaker: Aldo
Dataset size: 535 audio files
audio duration of 4-15 seconds (1:33:15)
Dataset structure
This dataset has been structured in the LJSpeech format:
wavs/
1.wav
2.wav
3.wav
535.wav
transcript.csv
negation_twitter_mexican_spanishThe T-MexNeg corpus of Tweets written in Mexican Spanish.
It consists of 13,704 Tweets, of which 4895 contain negation structures.
The corpus is the result of an analysis of sentiment and negation statements embedded in the language employed on social media. This repository includes annotation guidelines along with the corpus, manually annotated with labels of sentiment, negation cue, scope, and, event.
Twitter was used as the innitial source of the corpus; the tweets are a random subset of a set collected from Mexican users from September 2017 to April 2019.Mexican-Spanish_Emotion_Speech_Recognition_Corpus
ID
King-ASR-689
Language
English
Duration
170 hours
Speakers
150 People
Parameters
16kHz, 16bits
Recording Device
Mobile
URL
https://dataoceanai.com/datasets/asr/mexican-spanish-emotion-speech-recognition-corpus-conversations-mobile/
Mexican_Spanish_Speech_Recognition_Corpus
ID
King-ASR-936
Language
English
Duration
33.18 hours
Speakers
50 People
Parameters
16kHz, 16bits
Recording Device
Mobile
URL
https://dataoceanai.com/datasets/asr/mexican-spanish-speech-recognition-corpus-emotional-conversations-mobile/
