CoolFace
Datasetpublic

vera365/lexica_dataset

LexicaDataset LexicaDataset is a large-scale text-to-image prompt dataset shared in [USENIX'24] Prompt Stealing Attacks Against Text-to-Image Generation Models. It contains 61,467 prompt-image pairs collected from Lexica. All prompts are curated by real users and images are generated by Stable Diffusion. Data collection details can be found in the paper. Data Splits We randomly sample 80% of a dataset as the training dataset and the rest 20% as the testing… See the full description on the dataset page: https://huggingface.co/datasets/vera365/lexica_dataset.

sourceHugging Facecc-by-4.0updated 2y agoView on Hugging Face
8likes133downloads

vera365/lexica_dataset · main · files are served by the source, never re-hosted here