datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
video-quality-scored
Image-to-Video Quality-Scored Clips
A collection of prompted image-to-video samples with quality-evaluation metadata.
Each sample pairs a first frame (the I2V conditioning image) with one or both
of:
a generated video produced by a video model from the first frame + prompt
an original clip (the reference/source video the prompt was authored around)
A subset of the samples also carry per-clip quality scores: an overall
quality_score, six per-aspect breakdowns… See the full description on the dataset page: https://huggingface.co/datasets/mohantesting/video-quality-scored.deep-scores-v2
DeepScoresV2 — Complete
A HuggingFace-formatted mirror of the complete version of the
DeepScoresV2 dataset for music object detection.
Dataset description
DeepScoresV2 is a large-scale dataset of synthetically rendered music score pages
annotated with bounding boxes for musical symbols. The complete version contains
255,385 images with 151 million annotated instances across 135 symbol classes.
Each image is a full score page rendered from MuseScore across 5 music fonts… See the full description on the dataset page: https://huggingface.co/datasets/zzsi/deep-scores-v2.UnifiedReward-2.0-T2X-score-data
Dataset Summary
UnifiedReward-2.0-T2X-score-data is added for our UnifiedReward-2.0-qwen-[3b/7b/32b/72b] training.
This dataset enables UnifiedReward-2.0 introducing several new capabilities:
Pairwise scoring for image and video generation assessment on Alignment, Coherence, Style dimensions.
Pointwise scoring for image and video generation assessment on Alignment, Coherence/Physics, Style dimensions.
Welcome to try the latest version, and the inference code is available at… See the full description on the dataset page: https://huggingface.co/datasets/CodeGoat24/UnifiedReward-2.0-T2X-score-data.stockimage-scored-pt12stockimage-scored-pt6stockimage-scored-pt1stockimage-scored-pt8score-benchmark-dataBIGstockimage-1.5M-scored-pt-twostockimage-scored-pt9stockimage-1.5M-scored-low-similaritystockimage-scored-pt5stockimage-1.5M-scored-high-similaritystockimage-scored-pt10stockimage-scored-pt7BIGstockimage-1.5M-scored-pt-onestockimage-scored-pt2yandere_best_score
yandere_best_score 数据集说明
中文说明
yandere_best_score 数据集包含来自 https://yande.re 网站的图片,这些图片的评分都大于100。数据集通过爬虫程序收集了共计 100,000 张图片的相关信息。每张图片的评分经过筛选,确保仅包括评分大于100的高质量图片。
特点:
来源: https://yande.re
图片数量: 100,000 多张
收集范围: 时间轴在2014年之后的,所有图片的评分大于100的图片。
感谢nyanko7提供的yandere图源下载脚本
English Description
The yandere_best_score dataset contains images from the website https://yande.re, all of which have a score greater than 100. The dataset… See the full description on the dataset page: https://huggingface.co/datasets/ACCA225/yandere_best_score.polish-scores
Polish Historical-Scan OMR Benchmark
A page-level Optical Music Recognition (OMR) evaluation benchmark of 112 real
historical score scans, paired with both **kern (Humdrum) and MusicXML
ground-truth transcriptions.
Derived from the PRAIG/polish-scores
dataset, with kern normalization and a manual-fix pass applied. Released as
the real-scan half of the Transcoda evaluation suite alongside
btrkeks/verovio-synth-omr
and the btrkeks/transcoda-59M-zeroshot-v1
model.
Intended Use… See the full description on the dataset page: https://huggingface.co/datasets/btrkeks/polish-scores.youtube-piano-score
YouTube Piano Score
This repository contains piano score samples extracted from YouTube videos. Each sample lives under scores/<video_id>/ and may include:
meta.yaml: metadata for video segmentation, score layout, and staff grouping.
score.webp: the generated score image grid.
audio.*: the original audio extracted from video.
transkun.mid: the MIDI transcription result.
transkun-segmentation.yaml: measure and segment alignment metadata generated for the Transkun transcription… See the full description on the dataset page: https://huggingface.co/datasets/k-l-lambda/youtube-piano-score.stockimage-scored-pt4polish-scoresSCOREimage-aesthetic-scores
Rule34.nexus · Licence: Rule34.nexus Derived Dataset Licence 1.0
Rule34.nexus Image Aesthetic Scores
1. Overview
This dataset contains per-image aesthetic predictions for images in the Rule34.nexus corpus.
Predictions were generated using
discus0434/aesthetic-predictor-v2-5. Source images are not
included in this dataset — only opaque post identifiers, the source image's SHA-256 hash,
the post's content type, and the predicted score.… See the full description on the dataset page: https://huggingface.co/datasets/rule34nexus/image-aesthetic-scores.scorelens-data2
ScoreLens – Trainingsdatensätze
Alle Datensätze, mit denen die Modelle scorelens-dd1 … dd9 trainiert wurden –
jeweils ein Zip im Ultralytics-YOLO-Format (images/train|val, labels/train|val, data.yaml). Bildgröße 800 × 800.
7 Klassen, alle als Punkt gelabelt (Box fester Größe 0,025 um den Punkt): 20, 3, 11, 6 = Kalibrierpunkte am äußeren Doppelring an den
Segmentgrenzen 5/20, 17/3, 8/11, 13/6 · dart = Einstichpunkt der Dartspitze · 9, 15 = Zusatz-Kalibrierpunkte (vom Modell… See the full description on the dataset page: https://huggingface.co/datasets/Bayernator/scorelens-data2.stockimage-scored-pt11eu-ai-act-article-50-scoreboard
EU AI Act Article 50 Transparency Scoreboard
Version 1.0 (August 2026 Snapshot) · NM AI Research · CC BY 4.0
ORCID: 0009-0003-4213-7769 · DOI: 10.5281/zenodo.21819102
Dataset Summary
An empirical regulatory assurance dataset auditing 12 frontier AI consumer products and developer APIs against the European Union AI Act Article 50 Transparency Obligations, which took effect on 2 August 2026.
The dataset evaluates observable corporate artefacts (Terms of Service… See the full description on the dataset page: https://huggingface.co/datasets/NMAIResearch/eu-ai-act-article-50-scoreboard.deep-scores-v2-dense
DeepScoresV2 — Dense Subset
A HuggingFace-formatted mirror of the dense subset of the
DeepScoresV2 dataset for music object detection.
Dataset description
DeepScoresV2 is a large-scale dataset of synthetically rendered music score pages
annotated with bounding boxes for musical symbols. The dense subset contains
1,714 images selected by the authors as the most diverse and representative
sample from the full 803k-image dataset.
Each image is a full score page rendered from… See the full description on the dataset page: https://huggingface.co/datasets/zzsi/deep-scores-v2-dense.score_comparsionscut-fbp5500-v2-facial-beauty-scores
