CoolFace
Datasetpublic

jorgeortizfuentes/universal_spanish_chilean_corpus

Universal Chilean Spanish Corpus Este dataset se compone de 37_213_992 textos correspondientes a español de Chile y a español multidialectal. Los textos en español multidialectal provienen del spanish books. Los textos en español de Chile vienen de los dominios .cl del mc4 dataset y de tweets, noticias y reclamos de l chilean-spanish-corpus Name Count Source books 87967 spanish books mc4 8706681 from mc4 (.cl domains) in chilean-spanish-corpus twitter 27306583… See the full description on the dataset page: https://huggingface.co/datasets/jorgeortizfuentes/universal_spanish_chilean_corpus.

sourceHugging Faceunknownupdated 3y agoView on Hugging Face
8likes1.5kdownloads
settings

This repository belongs to jorgeortizfuentes on Hugging Face.

CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.

nameuniversal_spanish_chilean_corpus
visibilitypublic
licenceunknown
gatedno
ownerjorgeortizfuentes
Account settings
jorgeortizfuentes/universal_spanish_chilean_corpus · CoolFace