datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
cyberseceval3-visual-prompt-injection
Dataset Card for CyberSecEval 3 - Visual Prompt Injection Benchmark
Dataset Details
Dataset Description
This dataset provides a multimodal benchmark for visual prompt injection, with text/image inputs. It is part of CyberSecEval 3, the third edition of Meta's flagship suite of security benchmarks for LLMs to measure cybersecurity risks and capabilities across multiple domains.
Language(s): English
License: MIT
Dataset Sources
Repository: Link… See the full description on the dataset page: https://huggingface.co/datasets/facebook/cyberseceval3-visual-prompt-injection.visual-qa-llama-format
Open Paws Visual Qa Llama Format
This dataset is part of the Open Paws initiative to develop AI training data aligned with animal liberation and advocacy principles. Created to train AI systems that understand and promote animal welfare, rights, and liberation.
Dataset Details
Dataset Type: Multimodal Data
Format: JSONL (JSON Lines)
Languages: Multilingual (primarily English)
Focus: Animal advocacy and ethical reasoning
Organization: Open Paws
License: Apache 2.0… See the full description on the dataset page: https://huggingface.co/datasets/open-paws/visual-qa-llama-format.wikifragments-visual-arts-embeds
WikiFragments - Visual Arts Pages with Fragments (WikiFragmentsVA)
WikiFragmentsVA is a domain-specific multimodal dataset focused on the visual arts, derived from Wikipedia (en). It consists of textual paragraphs paired with related images (infoboxes and thumbnails), rendered as unified visual fragments. This dataset extends the base WikiFragments project by providing pre-rendered fragment images and multi-vector embeddings obtained via ColQwen2 v1.0, including optimized pooled… See the full description on the dataset page: https://huggingface.co/datasets/cilabuniba/wikifragments-visual-arts-embeds.visually-impaired-llm-assistance-dataset
Visually Impaired Assistance Dataset
Hey everyone! I'm currently working on a project to finetune an llm for assisting people with bad eyesights in daily life scenario so for it I had to create a synthetic dataset and here is the Visually Impaired Assistance Dataset that I created. This dataset provides step-by-step instructions with non-visual cues for a variety of daily tasks and activities, specifically designed to help visually impaired individuals. It covers a wide range of… See the full description on the dataset page: https://huggingface.co/datasets/sidfeels/visually-impaired-llm-assistance-dataset.DOD-Instruction-5040-02-Visual-Information
DoD Visual Information Question-Answer Dataset
Maintainer: Terry Eppler
Owner: US Federal Government
Dataset Summary
This dataset contains 250 document-grounded question-and-answer records based on DoD Instruction 5040.02, “Visual Information (VI),” dated October 27, 2011, and incorporating Change 2 effective April 20, 2018.
The source establishes Department of Defense policy, responsibilities, and procedures for managing visual-information records… See the full description on the dataset page: https://huggingface.co/datasets/leeroy-jankins/DOD-Instruction-5040-02-Visual-Information.visual_genome-simple-en
Dataset Card for Visual Genome Annotations in Simple English
This dataset contains captions that were rephrased into simple english so that a young child would understand it.
Dataset Details
Dataset Sources
The processed Visual Genome captions in this repo are based on the following sources:
941425b651f50cdb1a6f0673eaab6260 vg_caption.json (https://storage.googleapis.com/sfr-vision-language-research/LAVIS/datasets/visual_genome/vg_caption.json)
Visual… See the full description on the dataset page: https://huggingface.co/datasets/Jotschi/visual_genome-simple-en.visual_genome-opus-de
Dataset Card for Visual Genome Annotations in German language
This dataset contains captions that were machine translated using opus-mt-en-de.
Dataset Details
Dataset Sources
The processed Visual Genome captions in this repo are based on the following sources:
941425b651f50cdb1a6f0673eaab6260 vg_caption.json (https://storage.googleapis.com/sfr-vision-language-research/LAVIS/datasets/visual_genome/vg_caption.json)
Visual Genome:
Download:… See the full description on the dataset page: https://huggingface.co/datasets/Jotschi/visual_genome-opus-de.databird-visuals
DATABIRD VISUALS
Description
This dataset contains 5755 entries focused on understanding color theory principles and their application in various fields including complimentary colors, working with color contrasts, color psychology, color and branding, color in fashion design, color grading and visual storytelling, associating colors with emotions, building a color palette, and the impact of black and white vs. color imagery.
The data is provided in JSON format. Each… See the full description on the dataset page: https://huggingface.co/datasets/theprint/databird-visuals.
