CoolFace
Datasetpublic

oeg/CelebA_Sent2Vect_Sp

Corpus Summary This corpus has 192050 entries made up of descriptive sentences of the faces of the CelebA dataset. The preprocessing of the corpus has been to translate into Spanish the captions of the CelebA dataset with the algorithm used in Text2FaceGAN. In particular, all sentences are combined to generate a larger corpus. Additionally, a data preprocessing was applied that consists of eliminating stopwords, separation symbols and complementary elements that are not useful… See the full description on the dataset page: https://huggingface.co/datasets/oeg/CelebA_Sent2Vect_Sp.

sourceHugging Faceapache-2.0updated 3y agoView on Hugging Face
0likes34downloads

oeg/CelebA_Sent2Vect_Sp · main · files are served by the source, never re-hosted here