CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01GokuScraper /seedance-2-prompts-datasets 🎞️ Seedance-2-prompts-datasets 🎞️ The ultimate Seedance-2 video prompt dataset (50GB+). 8100+ video generation prompts with full metadata and preview frames. Truly open source: No login, no ads, no redirection. Just pure data for AI video creators. This project is a massive collection of prompts used for Bytedance's Seedance 2.0 and the resulting generated videos. The entire dataset exceeds 50GB and contains 8100+ videos, all structured into a comprehensive dataset. Due… See the full description on the dataset page: https://huggingface.co/datasets/GokuScraper/seedance-2-prompts-datasets.imagetext-to-video1K<n<10K45 likes188k downloads27d agoHugging Face02Goku-OpenLab /gpt-image-2-prompts-datasets 🖼️ GPT Image 2 Prompt Dataset 🖼️ The ultimate GPT Image 2 prompt dataset (5GB+). 15,000+ image generation prompts with full metadata and preview images. Truly open source: No login, no ads, no redirection. Just pure data for AI image creators. This project is a massive collection of prompts used for OpenAI's GPT Image 2 model and the resulting generated images. The entire dataset exceeds 5GB and contains 15,000+ images, all structured into a comprehensive dataset. Due to… See the full description on the dataset page: https://huggingface.co/datasets/Goku-OpenLab/gpt-image-2-prompts-datasets.imagetext-to-image10K<n<100K6 likes70k downloads27d agoHugging Face03Goku-OpenLab /nano-banana-pro-prompts-datasets 🖼️ Nano Banana Pro Prompt Dataset 🖼️ The ultimate Nano Banana Pro prompt dataset (6GB+). 26,000+ image generation prompts with full metadata and preview images. Truly open source: No login, no ads, no redirection. Just pure data for AI image creators. This project is a massive collection of prompts used for Nano Banana Pro AI image model and the resulting generated images. The entire dataset exceeds 6GB and contains 26,000+ images, all structured into a comprehensive… See the full description on the dataset page: https://huggingface.co/datasets/Goku-OpenLab/nano-banana-pro-prompts-datasets.imagetext-to-image10K<n<100K1 likes21k downloads2mo agoHugging Face04Goku-OpenLab /open-models-prompt-datasets 🖼️ Open Models Prompt Dataset 🖼️ The ultimate open models image prompt dataset (10GB+). 5400+ image generation prompts with full metadata and preview images. Truly open source: No login, no ads, no redirection. Just pure data for AI image creators. This project is a massive collection of prompts used for various open-source AI image models and the resulting generated images. The entire dataset exceeds 10GB and contains 5400+ images, all structured into a comprehensive… See the full description on the dataset page: https://huggingface.co/datasets/Goku-OpenLab/open-models-prompt-datasets.image1K<n<10K1 likes6k downloads2mo agoHugging Face05artificialguybr /veo3-video-prompts Veo 3 Video Generation Dataset English | Português do Brasil English Summary A collection of AI-generated videos created with Google's Veo 3 family of models. Each record contains the original text prompt, the model variant used, the generated video, and (when applicable) the input reference image. Videos are organized into one configuration per model variant. Videos: 5,811 Input images: 1,354 Configurations: 6 Language of prompts: multilingual… See the full description on the dataset page: https://huggingface.co/datasets/artificialguybr/veo3-video-prompts.imagetext-to-video1K<n<10K0 likes5.3k downloads1mo agoHugging Face06Goku-OpenLab /messy-prompt-datasets 🎨 Messy Prompt Dataset 🎨 A mixed collection of AI image prompts (500+). A bit of everything — raw and uncurated. Truly open source: No login, no ads, no redirection. Just pure data for AI creators. This project is a growing collection of diverse image generation prompts gathered from social platforms like Twitter/X. The entire dataset contains 500+ images, all structured into a comprehensive dataset. Due to GitHub's limitations with large file storage, the full dataset… See the full description on the dataset page: https://huggingface.co/datasets/Goku-OpenLab/messy-prompt-datasets.image1K<n<10K0 likes5.3k downloads2mo agoHugging Face07facebook /cyberseceval3-visual-prompt-injection Dataset Card for CyberSecEval 3 - Visual Prompt Injection Benchmark Dataset Details Dataset Description This dataset provides a multimodal benchmark for visual prompt injection, with text/image inputs. It is part of CyberSecEval 3, the third edition of Meta's flagship suite of security benchmarks for LLMs to measure cybersecurity risks and capabilities across multiple domains. Language(s): English License: MIT Dataset Sources Repository: Link… See the full description on the dataset page: https://huggingface.co/datasets/facebook/cyberseceval3-visual-prompt-injection.imagetext-generation1K<n<10K10 likes2.6k downloads2y agoHugging Face08AlekseyKorshuk /product-photography-v1-tiny-prompts-tasks-collage-filteredimage1K<n<10K1 likes2.6k downloads3y agoHugging Face09Nerfgun3 /bad_prompt Negative Embedding / Textual Inversion Idea The idea behind this embedding was to somehow train the negative prompt as an embedding, thus unifying the basis of the negative prompt into one word or embedding. Side note: Embedding has proven to be very helpful for the generation of hands! :) Usage To use this embedding you have to download the file aswell as drop it into the "\stable-diffusion-webui\embeddings" folder. Please put the embedding in… See the full description on the dataset page: https://huggingface.co/datasets/Nerfgun3/bad_prompt.imagen<1K980 likes2.5k downloads4y agoHugging Face10AiActivity /All-Prompt-Jailbreakimagetext-generationn<1K10 likes1.5k downloads1y agoHugging Face11ksterx /hle-no-img-prompt-completion-formatimage1K<n<10K0 likes1.2k downloads1y agoHugging Face12Korakoe /NijiJourney-Prompt-Pairs NijiJourney Prompt Pairs A dataset containing txt2img prompt pairs for training on diffusion models The final goal of this dataset is to create an OpenJourney like model but with NijiJourney images image1K<n<10K16 likes1.1k downloads4y agoHugging Face13davidberenstein1957 /stream-prompt-upsample-runsimagen<1K0 likes1.1k downloads2mo agoHugging Face14gdsu /sdxl_images_easy_prompts-artists-seed1image10K<n<100K0 likes869 downloads2y agoHugging Face15lefreud /GimbalDiffusion-Prompt-Entanglement GimbalDiffusion Prompt Entanglement Benchmark The 380-sample benchmark used to measure pitch/content entanglement in GimbalDiffusion: Gravity-Aware Camera Control for Video Generation. Project page Paper Poly Haven Contents prompt_entanglement.tar: the complete benchmark in one archive, unpacking directly into the layout below. test_index.json and test_samples/: 380 prompts, seeds, camera matrices, intrinsics, requested pitch angles, and Poly Haven panorama… See the full description on the dataset page: https://huggingface.co/datasets/lefreud/GimbalDiffusion-Prompt-Entanglement.imagetext-to-videon<1K0 likes803 downloads2mo agoHugging Face16MAPS-research /GEMRec-PromptBook GEMRec-18k -- Prompt Book This is the official image dataset for the paper Towards Personalized Prompt-Model Retrieval for Generative Recommendation. Dataset Intro GEMRec-18K is a prompt-model interaction dataset with 18K images generated by 200 publicly-available generative models paired with a diverse set of 90 textual prompts. We randomly sampled a subset of 197 models from the full set of models (all finetuned from Stable Diffusion) on Civitai according to the… See the full description on the dataset page: https://huggingface.co/datasets/MAPS-research/GEMRec-PromptBook.imagetext-to-image10K<n<100K3 likes608 downloads3y agoHugging Face17Mike1997126 /All-Prompt-Jailbreakimagetext-generationn<1K1 likes559 downloads8mo agoHugging Face18brown-palm /force-prompting-dataset-creationimagen<1K1 likes541 downloads1y agoHugging Face19jtatman /stable-diffusion-prompts-stats-full-uncensoredimage100K<n<1M152 likes534 downloads2y agoHugging Face20alakxender /dhivehi-image-bbox-prompt Dhivehi Image Bounding Box Prompt Dataset This dataset, alakxender/dhivehi-image-bbox-prompt, contains 58,738 images annotated with COCO-style bounding boxes and Dhivehi (Thaana script) text, along with layout categories such as Text, Title, Picture, Caption, and Columns. It is designed for OCR, document layout analysis, and multimodal vision–language research focused on Dhivehi. Dataset Each row includes: image — the RGB image (preserved original dimensions) width… See the full description on the dataset page: https://huggingface.co/datasets/alakxender/dhivehi-image-bbox-prompt.imageobject-detection10K<n<100K0 likes529 downloads1y agoHugging Face21yeates /PromptfixData Model Scources Dataset: https://huggingface.co/datasets/yeates/PromptfixData Github: https://github.com/yeates/PromptFix Paper: https://arxiv.org/pdf/2405.16785 Project Page: https://www.yongshengyu.com/PromptFix-Page/ Model Usage The PromptFix dataset is intended solely for research purposes. Please note that the PromptFix dataset is curated from open-source research projects and publicly available photo libraries. By using our dataset, you automatically agree to… See the full description on the dataset page: https://huggingface.co/datasets/yeates/PromptfixData.image1M<n<10M12 likes518 downloads2y agoHugging Face22My-Weird-Prompts /episodes My Weird Prompts - Episode Dataset The production record of every episode of the My Weird Prompts podcast: the transcript, links to the published episode, a description of the prompt that started it, and the generation telemetry for how it was made - model, pipeline version, GPU, timings and compute cost. 5,365 episodes. Synced daily from the production database. from datasets import load_dataset ds = load_dataset("My-Weird-Prompts/episodes", split="train") Which… See the full description on the dataset page: https://huggingface.co/datasets/My-Weird-Prompts/episodes.audiotext-generation1K<n<10K1 likes506 downloads8h agoHugging Face23JakkMehoffFriend /All-Prompt-Jailbreakimagetext-generationn<1K0 likes456 downloads3mo agoHugging Face24CaptainSlayAh0 /All-Prompt-Jailbreakimagetext-generationn<1K1 likes445 downloads4mo agoHugging Face25behavior-in-the-wild /spro-optimized-prompts-fullimage100K<n<1M0 likes440 downloads10mo agoHugging Face26robot-learning-group47 /eval2_all_promptsThis dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase_version": "v3.0", "fps": 30, "features": { "action": { "dtype": "float32", "names": [ "shoulder_pan.pos", "shoulder_lift.pos", "elbow_flex.pos", "wrist_flex.pos", "wrist_roll.pos", "gripper.pos" ], "shape": [ 6… See the full description on the dataset page: https://huggingface.co/datasets/robot-learning-group47/eval2_all_prompts.imagerobotics100K<n<1M0 likes423 downloads4mo agoHugging Face27ahmedmostafa0521 /All-Prompt-Jailbreakimagetext-generationn<1K0 likes423 downloads4mo agoHugging Face28Nymbo /Prompt_Protections Protection Protections This dataset contains a number of snippets and short extentions to add to the system prompt of bots and GPTs to persuade the model not to reveal it's instructions to the user. It's not a perfect solution, but sometimes a little clever prompting is all you need :) image2 likes395 downloads2y agoHugging Face29yio-ye2004 /fr3_pickplace_extended_new_cmd_SYNC_part_corrected_with_prompt_testThis dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase_version": "v2.1", "robot_type": "fr3", "total_episodes": 301, "total_frames": 276017, "total_tasks": 2, "total_videos": 0, "total_chunks": 1, "chunks_size": 1000, "fps": 60, "splits": { "train": "0:301" }, "data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet", "video_path":… See the full description on the dataset page: https://huggingface.co/datasets/yio-ye2004/fr3_pickplace_extended_new_cmd_SYNC_part_corrected_with_prompt_test.imagerobotics100K<n<1M0 likes334 downloads8mo agoHugging Face30la-ji /sd-prompt-image-in-the-wild-counterfeitimage1M<n<10M3 likes328 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.