Racoci/CORAA-v1.1
CORAA-v1.1 CORAA-v1.1 is a publicly available dataset for Automatic Speech Recognition (ASR) in the Brazilian Portuguese language containing 290.77 hours of audios and their respective transcriptions (400k+ segmented audios). The dataset is composed of audios of 5 original projects: ALIP (Gonçalves, 2019) C-ORAL Brazil (Raso and Mello, 2012) NURC-Recife (Oliviera Jr., 2016) SP-2010 (Mendes and Oushiro, 2012) TEDx talks (talks in Portuguese) The audios were either validated by… See the full description on the dataset page: https://huggingface.co/datasets/Racoci/CORAA-v1.1.
Conversations for this repository live on Hugging Face.
CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.
Open discussions on Hugging Face