CoolFace
Datasetpublic

umarigan/turkish_clip_dataset_with_text_embeddings

This dataset cleaned and dowloaded version of following dataset: https://huggingface.co/datasets/visheratin/laion-coco-nllb The main purpose was to extract Turkish captions and download images. You can use this dataset to fine-tune or create a clip model. Since there English and Turkish captions you can also use those to create language model?

sourceHugging Facecreativeml-openrail-mupdated 3y agoView on Hugging Face
1likes291downloads
Dataset Card

This dataset cleaned and dowloaded version of following dataset: https://huggingface.co/datasets/visheratin/laion-coco-nllb The main purpose was to extract Turkish captions and download images. You can use this dataset to fine-tune or create a clip model. Since there English and Turkish captions you can also use those to create language model?