dzur658/opus-es-monologues
OPUS Spanish Monologues A dataset that captures monologues from the Spanish Open Subtitles Project dump and undergoes light cleaning. Monologues retained in this dataset are intances in the raw .txt dump where a single speaker is uninterrupted for more than 100 words. The dataset consists of monologues from the 2013, 2016, and 2018 OPUS Spanish monolingual datasets. Quick Dataset Facts Contains 1,481 documents Each document averages ~241.2 words The dataset… See the full description on the dataset page: https://huggingface.co/datasets/dzur658/opus-es-monologues.
This repository belongs to dzur658 on Hugging Face.
CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.
