CoolFace
Datasetpublic

BSC-LT/CAESAR-TV3

Dataset card for CAESAR-TV3 Dataset Summary This corpus includes 5 hours and 45 minutes of Catalan speech code-switched with Spanish extracted from the original tv3_parla dataset. Supported Tasks and Leaderboards The CAESAR-TV3 dataset is designed for the Automatic Speech Recognition (ASR) task, enabling the transcription of utterances in Catalan, Spanish, and code-switched speech between the two languages. Languages The dataset… See the full description on the dataset page: https://huggingface.co/datasets/BSC-LT/CAESAR-TV3.

sourceHugging Facecc-by-nc-4.0updated 1y agoView on Hugging Face
1likes57downloads
13 commits on main
7a80a571y ago

Update README.md

AbirMessaoudi
9d0c3b31y ago

Update README.md

AbirMessaoudi
944d8381y ago

Update README.md

AbirMessaoudi
4ce272f1y ago

Update README.md

AbirMessaoudi
524f8c01y ago

Update README.md

AbirMessaoudi
a50e3c31y ago

Update README.md

AbirMessaoudi
1f0f2982y ago

Update README.md

AbirMessaoudi
225554e2y ago

Update README.md

AbirMessaoudi
e293a5b2y ago

Update README.md

AbirMessaoudi
f1355972y ago

Update README.md

AbirMessaoudi
fa6aa702y ago

Update README.md

AbirMessaoudi
b25948f2y ago

Upload dataset

AbirMessaoudi
875ee7c2y ago

initial commit

AbirMessaoudi