datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
MPIE-Bench
MPIE-Bench
Official 2,500-sample test set for multi-person interaction-aware image editing evaluation.
GitHub (code + protocol): https://github.com/AnnLin0628/mpie-bench
Org: muset-ai
Dataset Viewer
The default config (default / test) is one row per evaluation sample:
Column
Meaning
cat
Interaction category (filter / group by this)
gt
Held-out ground-truth image
prompt
Edit instruction
ref_paths
Reference image paths under images/
sample_id… See the full description on the dataset page: https://huggingface.co/datasets/muset-ai/MPIE-Bench.MUSE
MUSE: A CAD Design Benchmark with Multi-modal Ground Truth and Rubric-based Evaluation
MUSE is a benchmark of 106 CAD design cases for evaluating language and
multi-modal models on engineering-grade 3D design tasks. Each case pairs a
natural-language design specification with multi-view ground-truth artefacts
(2D engineering drawings + 3D rendered images) and a hand-crafted, rubric-style
evaluation guide.
Why this benchmark
Most CAD/3D benchmarks evaluate either pure… See the full description on the dataset page: https://huggingface.co/datasets/dongxiaoyu/MUSE.met_museumMUSE-VA
MUSE-VA Dataset
English | 中文
MUSE-VA (Multimodal MUSic Emotion Dataset with Balanced VA) is a large-scale multimodal music emotion dataset designed for music emotion understanding, emotion-controllable music generation, and cross-modal affective modeling. The dataset starts from target coordinates sampled in the continuous Valence-Arousal (VA) space and uses a five-stage LLM agent pipeline with affective and musical knowledge injection to construct music, text, images, and… See the full description on the dataset page: https://huggingface.co/datasets/jiahaomei/MUSE-VA.art-museums-pd-440k
Art Museums PD 440K
Summary
This is a dataset to train text-to-image or any text and image multimodal models with minimized copyright/licensing concerns.
All images and texts in this dataset are orignally shared under CC0 or public domain, and no pretrained models or any AI models are used to build this dataset except for our ElanMT model to translate English captions to Japanese.
ElanMT model is trained solely on licensed corpus.
Data sources
Images and… See the full description on the dataset page: https://huggingface.co/datasets/Mitsua/art-museums-pd-440k.MUSE-VA-A2Imixtec-zouche-nuttall-british-museum
Mixtec Zouche-Nuttall Labeled Dataset
Introduction
A Look at Figure Segments
File Formats and Figure Names
An Overview of Name Segments
Overfiew of File Format and Naming Conventions
An Overview of Scene Segments
File Formats and Scene Names
Missing Pages in the Video Narration
Scenes that Span Multiple PagesVery Large Scenes: Scene 187
Introduction
This dataset contains 3 different directories, figure-cutouts/, name-date-cutouts/, and scene-cutouts/. Each… See the full description on the dataset page: https://huggingface.co/datasets/ufdatastudio/mixtec-zouche-nuttall-british-museum.MUSE-VA-I2AMuSEAgent-EvalMuseMedical_HumanEvaluationmusedash-filteredMUSE-VA-A2ImusedashMUSE
MUSE
Multimodal evaluation data.
Quick links: [🌐 Website] [📜 Paper] [💻 Code]
Contents
1,800 test questions and 1,174 referenced images.
Task
Questions
Activity Localization
200
Culture Identification
200
Activity Description
200
Affective Computing
200
Jigsaw Puzzle
200
Object Count
200
Relative Position
200
Remote Interaction
200
Scene Classification
200
Affective Computing consists of four tasks: Object Classification, Emotion… See the full description on the dataset page: https://huggingface.co/datasets/Cyn7hia-Z/MUSE.muse_imagesPrasasti-Museum-Majapahit
Prasasti Enhancement
Eksplorasi 4 skenario pipeline pengolahan citra digital untuk meningkatkan keterbacaan aksara
pada citra batu prasasti kuno Nusantara (Museum Majapahit). Tiap skenario punya parameter
yang bisa diubah lewat UI Gradio, lengkap dengan mode segmentasi ROI manual atau otomatis
(GrabCut).
Dataset: prasasti-di-museum-majapahit (Kaggle).
museum-of-thingsmuseum_collectionsbritish-museum-igbo-collection
British Museum (Igbo Collection) Dataset
This dataset is a curated collection of historical artifacts and photographs related to the Igbo people of Nigeria, sourced from the British Museum Collection.
⚠️ Data Quality & Limitations
Image Quality: The images in this dataset are "Preview" / Web-Resolution quality (sourced directly from the British Museum's public CSV export). While sufficient for digital archiving, web display, and research reference, they are not… See the full description on the dataset page: https://huggingface.co/datasets/nwokikeonyeka/british-museum-igbo-collection.chart-museum-samplesimgbed-storagemuse-sarcasm-explanation
MuSe: Multimodal Sarcasm Explanation (Reformatted)
This repository provides a Hugging Face-compatible version of the MuSe (MORE) dataset.
Modifications in this version
To make the dataset easier to use with the datasets library, the following changes were made:
Unified Schema: Merged separate OCR and Non-OCR files into a single test split.
Metadata Flags: Added an is_ocr (boolean) column to distinguish between image types.
Image Integration: Converted image paths into a… See the full description on the dataset page: https://huggingface.co/datasets/alita9/muse-sarcasm-explanation.muse512
Dataset Card for "muse_512"
```py
from PIL import Image
import torch
from muse import PipelineMuse, MaskGiTUViT
from datasets import Dataset, Features
from datasets import Image as ImageFeature
from datasets import Value, load_dataset
device = "cuda" if torch.cuda.is_available() else "cpu"
pipe = PipelineMuse.from_pretrained(
transformer_path="valhalla/research-run",
text_encoder_path="openMUSE/clip-vit-large-patch14-text-enc"… See the full description on the dataset page: https://huggingface.co/datasets/diffusers-parti-prompts/muse512.muse-landmark-1500MUSEMUSE (Multi-city Urban Satellite-Energy Dataset)
We align municipal-scale building energy disclosure records with satellite imagery and create paired samples at a fixed spatial extent
This repository contains ZIP archives for manual download and use.
Github: https://github.com/kailaisun/GenAI4Urban-Energy
Paper: https://arxiv.org/abs/2605.18101
MUSE-VA-I2APics_from_museumsmusee-dart-et-dhistoire-de-saint-brieuc-fonds-peinture
Musée d'art et d'histoire de Saint Brieuc Fonds Peinture
Source
Source officielle : https://www.data.gouv.fr/datasets/musee-dart-et-dhistoire-de-saint-brieuc-fonds-peinture
Identifiant du jeu de données data.gouv.fr : 585a3dacc751df15659ccad1
Slug data.gouv.fr : musee-dart-et-dhistoire-de-saint-brieuc-fonds-peinture
Licence indiquée dans les métadonnées data.gouv.fr : lov2
Structure Hugging Face
Un jeu de données data.gouv.fr = un dépôt Hugging… See the full description on the dataset page: https://huggingface.co/datasets/Data-Gouv-ML/musee-dart-et-dhistoire-de-saint-brieuc-fonds-peinture.Fitzwilliam-museum-imagesA CSV file of image urls and meta data for the Fitzwilliam Museum system. The images are licensed under more restrictive terms, the links to URLS
are open via their API and website.
Fitzwilliam-museum-tweets
