datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
religious-artwork-analysis-data
Data
Download from Kaggle (needs an API token from https://www.kaggle.com/settings):
pip install kaggle
python data/download.py
Expected layout after download:
data/artwork_metadata.csv 3,997 rows — filename, religion (1,000 each of
buddhism / christianity / hinduism; 997 islam),
sub_religion, artist, title, year, place,
source, source_id, source_url, image_url
data/images/ the… See the full description on the dataset page: https://huggingface.co/datasets/cvikl/religious-artwork-analysis-data.artwork_for_sdxlArtwork Images, to generate the similar artwork using stable diffusion model.mtg-scryfall-unique-artwork-20240809-with-card-art-descriptions-and-images-with-embeddingsmtg-scryfall-unique-artwork-20240809-combined-elements-embeddingshood-artwork
HOOD Artwork Dataset
SteamGridDB metadata snapshot for the HOOD gaming launcher.
artwork.json — single JSON blob: {generatedAt, count, games:[{sgdbId, name, verified, types, releaseDate, assets:{grids, heroes, logos, icons}}]}
mapping.ndjson — igdbKey<TAB>sgdbId rows for cross-referencing
Asset records contain metadata ONLY (URLs, dimensions, style, score) — no image binaries.
The server loads this snapshot at boot and serves dataset-first, falling back to the live SteamGridDB… See the full description on the dataset page: https://huggingface.co/datasets/mtaaz/hood-artwork.bot-artworkmtg-scryfall-unique-artwork-20240809-with-card-art-descriptions-and-imagesodor-olfactory-artwork-detection
ODOR — Object Detection for Olfactory References in Artworks
4,712 artwork images with 38,165 bounding-box annotations across 139 fine-grained
categories of smell-related objects — flowers, fruit, censers, animals, vessels — drawn from
European art.
Computer vision on artworks is hard in ways photographic benchmarks are not: artistic
abstraction, peripheral objects, and fine-grained distinctions between visually similar
classes. ODOR is built to test exactly that.… See the full description on the dataset page: https://huggingface.co/datasets/biglam/odor-olfactory-artwork-detection.vintage-artworks-60k-captionedThis is a dataset consisting of 60k vintage artworks from the 20th century, consisting of vintage pulp, sci-fi and pinup artworks from that era.
The dataset has short and long captions for each image, as well as resolution information. The large captions (large_caption column) were made with florence-2-large-ft, and then shortened with llama 3 8b (see short_caption column).
artworksdevin-neuron-artworkmtg-scryfall-unique-artwork-20240809-with-card-art-descriptionsartworkCreativityStolen_ArtworksThis dataset contains artoworks classified by Interpol as stolen. The dataset can also be found and filtered differently on: https://www.workwithdata.com/datasets/artworks?f=1&fcol0=museum&fop0=%3D&fval0=Stolen+art+%28Interpol%29
Similar datasets can also be found on: https://www.workwithdata.com
AP_ArtHistory_Artwork_And_List_Of_QuestionsThis dataset is a history of the last 25 years of Multiple Choice Questions that have appeared on the AP Art History exam. The columns are artwork title, artwork description generated based on the title, question list relevant to the artwork that has come on the exam, there are no choices in this dataset but access this dataset for those features. https://huggingface.co/datasets/shaamil101/AP_ArtHistory_Questions_With_Choices
license: apache-2.0
museums-artworksWho are the source data producers?
The source data producers include the ASTOUND team for generating and evaluating the dialogues. Also include the ArtEmis team for artwork data and annotators for emotion attributions.
ASTOUND is an EIC funded project (No. 101071191) under the HORIZON-EIC-2021-PATHFINDERCHALLENGES-01 call. Website: https://www.astound-project.eu
artwork_prompts
Dataset Card for "artwork_prompts"
More Information needed
artworks
Combined Louvre and Art Institute of Chicago (AIC) Collection Dataset
Dataset Summary
This dataset merges artwork information from two prominent museum collections: the Musée du Louvre and The Art Institute of Chicago (AIC). It combines data from the Louvre Paper and Canvas Collection and the AIC Dataset 0.2 datasets.
Due to differences in the original datasets' schemas, a decision was made to focus on common fields and create a non-atomic full_info field… See the full description on the dataset page: https://huggingface.co/datasets/anna-bozhenko/artworks.artworkArtworks_as_WebPagesartworkartworks_by_famous_paintersMaria_Prymachenko_artwork#Maria Prymachenko's artwork scraped from Wikiart
#Used in a Stable Diffusion project for class in order to generate images based on text in a folklore style
3d_artwork_datasetArtworksArtworksGTA5-artwork_LoRA
GTA5 Artwork LoRA Fine-Tuning Dataset
This dataset contains images featuring the distinct art style of Grand Theft Auto V (GTA5). It's designed specifically for fine-tuning diffusion models, like Stable Diffusion, using techniques such as Low-Rank Adaptation (LoRA).
Dataset Description
The dataset consists of 140 images sourced from official GTA5 artwork and promotional materials. Each image is paired with a descriptive text caption in a metadata.jsonl file. This pairing… See the full description on the dataset page: https://huggingface.co/datasets/luluboy/GTA5-artwork_LoRA.
