datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
veo3-video-prompts
Veo 3 Video Generation Dataset
English | Português do Brasil
English
Summary
A collection of AI-generated videos created with Google's Veo 3 family of models. Each record contains the original text prompt, the model variant used, the generated video, and (when applicable) the input reference image. Videos are organized into one configuration per model variant.
Videos: 5,811
Input images: 1,354
Configurations: 6
Language of prompts: multilingual… See the full description on the dataset page: https://huggingface.co/datasets/artificialguybr/veo3-video-prompts.cyberseceval3-visual-prompt-injection
Dataset Card for CyberSecEval 3 - Visual Prompt Injection Benchmark
Dataset Details
Dataset Description
This dataset provides a multimodal benchmark for visual prompt injection, with text/image inputs. It is part of CyberSecEval 3, the third edition of Meta's flagship suite of security benchmarks for LLMs to measure cybersecurity risks and capabilities across multiple domains.
Language(s): English
License: MIT
Dataset Sources
Repository: Link… See the full description on the dataset page: https://huggingface.co/datasets/facebook/cyberseceval3-visual-prompt-injection.editing_promptsadaption-pokemon-story-prompts
This dataset is a remastered version of this dataset prepared using Adaption's Adaptive Data platform.
adaption-pokemon_story_prompts
This dataset contains prompts instructing a model to write stories about specific Pokémon based on their detailed attributes, including stats, types, abilities, and lore. Each entry provides structured data such as height, weight, generation, and flavor text alongside an image URL. The primary focus is on generating creative narratives grounded… See the full description on the dataset page: https://huggingface.co/datasets/sarahooker/adaption-pokemon-story-prompts.prompt2model-examples
Prompt2Model Toy Examples
Product: Prompt2Model:
a language-guided vision model factory. A typed pipeline (prompt, dataset config, training,
calibration/conformal abstain, ONNX export, an optional distill/quantize step with an
accuracy-floor gate, and a hard-case flywheel).
What this is (and isn't)
This is not a benchmark dataset. Prompt2Model has no natural "own" benchmark corpus the way a
task-specific product does. What's uploaded here is the repository's own… See the full description on the dataset page: https://huggingface.co/datasets/Dhi-Technologies/prompt2model-examples.anime-character-prompt-15kprompts_wiki_fictional_characters_raw_data_with_imageprompts_subset_wiki_fictional_characters_raw_data_with_imageprompt_cocoleo_prompts_v1
leo_prompts_v1: a dataset collection
collection of several prompts, nagative prompts and image urls datasets
data is uncleaned/non-normalized, they are as tey appear in leonardo.ai
data de-duplicated on a basic level.
contents
DatasetDict({
all: Dataset({
features: ['id', 'url', 'prompt', 'negative_prompt', 'imageHeight', 'imageWidth'],
num_rows: 299934
})
})
CITE
@misc {samact_2023,
author = { {SamAct} },
title =… See the full description on the dataset page: https://huggingface.co/datasets/SamAct/leo_prompts_v1.
