CoolFace
Datasetpublic

BSC-LT/CAESAR-TV3

Dataset card for CAESAR-TV3 Dataset Summary This corpus includes 5 hours and 45 minutes of Catalan speech code-switched with Spanish extracted from the original tv3_parla dataset. Supported Tasks and Leaderboards The CAESAR-TV3 dataset is designed for the Automatic Speech Recognition (ASR) task, enabling the transcription of utterances in Catalan, Spanish, and code-switched speech between the two languages. Languages The dataset… See the full description on the dataset page: https://huggingface.co/datasets/BSC-LT/CAESAR-TV3.

sourceHugging Facecc-by-nc-4.0updated 1y agoView on Hugging Face
1likes57downloads
settings

This repository belongs to BSC-LT on Hugging Face.

CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.

nameCAESAR-TV3
visibilitypublic
licencecc-by-nc-4.0
gatedno
ownerBSC-LT
Account settings
BSC-LT/CAESAR-TV3 · CoolFace