datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
danbooru
Danbooru 2024 Dataset
Danbooru 2024 数据集
A collection of images from Danbooru website, organized and packaged by ID sequence. This dataset is for research and learning purposes only.
本数据集收集了来自 Danbooru 网站的图像,按 ID 顺序组织打包。该数据集仅用于研究和学习目的。
Dataset Description
数据集描述
This dataset contains image resources from Danbooru website, updated to ID 8380648 (Update time: 2024-11-03).
本数据集包含来自 Danbooru 网站的图像资源,更新至 ID 8380648(更新时间:2024-11-03)。
Data… See the full description on the dataset page: https://huggingface.co/datasets/picollect/danbooru.Mind-Brushpico-banana-smolvlm-format-with-rejected-answer
pico-banana-smolvlm-format-with-rejected-answer
Balanced image-level tampering detection dataset in SmolVLM-style format
with chosen/rejected answer pairs, derived from the pico-banana MCQ
pipeline. Suitable for preference learning (e.g. DPO) and RLHF-style training.
Dataset overview
Same as vanloc1808/pico-banana-smolvlm-format, but each example includes a
rejected_answer field: the answer from the counterpart sample (same
edited/original image pair, opposite… See the full description on the dataset page: https://huggingface.co/datasets/vanloc1808/pico-banana-smolvlm-format-with-rejected-answer.asankakupico-8-games
PICO-8 Games Dataset
The first multimodal dataset of PICO-8 games. 10,967 cartridges scraped from the Lexaloffle BBS, each decomposed into Lua source code, pixel-art spritesheets, tile maps, sound effects, music patterns, and metadata.
Label screenshots from the top 48 games by star count
What's Inside
Every PICO-8 cartridge is a self-contained game packed into a single file. This dataset cracks each one open into its component parts:
The… See the full description on the dataset page: https://huggingface.co/datasets/Fraser/pico-8-games.BLINK-Twice
BLINK-Twice: You see, but you do not observe. A Reasoning Benchmark on Visual Perception
📌 About BLINK-Twice
BLINK-Twice Task Overview: (a) Visual reasoning task requiring detailed observation and careful reasoning; (b) Natural adversarial samples with similar appearance but opposite semantics, forcing models to rely on visual input; (c) Reasoning step annotation including detailed visual clues and true reality to evaluate thought chain output.
As illustrated… See the full description on the dataset page: https://huggingface.co/datasets/PicoTrex/BLINK-Twice.sankaku_jsondanbooru2danbooru_jsonAniLayer2Dneuralatlas-imagenet-pico-aimagicbrush_pico_augmentedxLingual-picobanana-taxonomy-6k
xLingual PicoBanana Taxonomy 6k
Canonical public dataset repo:
Legend2727/xLingual-picobanana-taxonomy-6k
This release is the cleaned 6k subset of the earlier 12k repository. It keeps:
source image
edited image
instruction in English, Hindi, and Bangla
canonical 11-label taxonomy labels
Dataset contract
rows: 6000
languages: en / hi / bn
taxonomy labels: 11 canonical classes
source_type distribution: {"preference_rejected": 3893, "sft": 2107}
Files… See the full description on the dataset page: https://huggingface.co/datasets/Legend2727/xLingual-picobanana-taxonomy-6k.PICooKpico_augmentedpicook_diffusion
