datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
docvqa_1200_examplesz-image-examples
Z-Image Turbo Portrait Dataset
This dataset contains 126 portrait prompts and their corresponding image outputs, demonstrating the capabilities of the Z-Image Turbo text-to-image model.
Model Information
Model Name: Z-Image Turbo
Hugging Face Repository: Tongyi-MAI/Z-Image-Turbo
Dataset Contents
prompts.jsonl: A JSONL file containing the 126 text prompts used for generation. Each entry includes a unique ID and the prompt text.
outputs/: Directory containing… See the full description on the dataset page: https://huggingface.co/datasets/k-mktr/z-image-examples.example-space-to-dataset-parquetdocvqa_1200_examples_donutrvl_cdip_10_examples_per_classrvl_cdip_100_examples_per_class
Dataset Card for "rvl_cdip_100_examples_per_class"
More Information needed
rvl_cdip_300_examples_per_classqwen-spatial-reasoning-incorrect-examples
Incorrect spatial reasoning examples for Qwen/Qwen3.5-0.8B-Base
Overview
This dataset contains incorrect non-empty predictions made by Qwen/Qwen3.5-0.8B-Base on a synthetic spatial reasoning benchmark built from 4x4 object-grid images.
I evaluated the model on 84 questions. It answered 54 of them incorrectly and achieved an overall accuracy of 35.714%. Some incorrect rows had an empty parsed pred_final, which I treat as formatting failures rather than useful supervised… See the full description on the dataset page: https://huggingface.co/datasets/safaeid48/qwen-spatial-reasoning-incorrect-examples.rvl_cdip_10_examples_per_class_donutrvl_cdip_100_examples_per_classobject-detection-examples
Dataset Card for CPPE - 5
Dataset Summary
CPPE - 5 (Medical Personal Protective Equipment) is a new challenging dataset with the goal to allow the study of subordinate categorization of medical personal protective equipments, which is not possible with other popular data sets that focus on broad level categories.
Some features of this dataset are:
high quality images and annotations (~4.6 bounding boxes per image)
real-life images unlike any current such dataset
majority… See the full description on the dataset page: https://huggingface.co/datasets/Charles95/object-detection-examples.OnePromptOneStory-Examples-Vid-head75font-examples
Dataset Card for "font-examples"
More Information needed
OnePromptOneStory-Examples-CCIPOnePromptOneStory-Examplesrvl_cdip_300_examples_per_class_donut_v2rvl_cdip_300_examples_per_class_donut_v4rvl_cdip_30_examples_per_class_donutdoc-image-10
[doc] image dataset 10
This dataset contains a parquet file that contains an image column.
docvqa_1200_examples_donutgacha_examples_test
Qwen3.5-27B diverse-sampling — qualitative examples
Side-by-side qualitative outputs from Qwen/Qwen3.5-27B under several
prompting methods for diverse sampling, in two domains. One row per prompt; one
column per method holding that method's list of outputs for the same prompt,
so a row is a direct method-vs-method comparison.
config
rows (prompts)
methods
outputs per cell
story
52
naive, plan, idea, verbalized_k8, verbalized_k16, gacha
up to 128 stories
image
50… See the full description on the dataset page: https://huggingface.co/datasets/scottgeng00/gacha_examples_test.geo170k-10-examples-r1-filteredFirst 10 examples of xyliu6/geo170k-1k-r1-filtered for development on MLX-VLM.
rvl_cdip_10_examples_per_class_donut
Dataset Card for "rvl_cdip_10_examples_per_class_donut"
More Information needed
docvqa_1000_examples
Dataset Card for "docvqa_1000_examples"
More Information needed
example-space-to-dataset-parquetdocvqa-xray-examplesOnePromptOneStory-animagine-xl-4-0-Examplesrvl_cdip_10_examples_per_class_datasetrvl_cdip_300_examples_per_class_donutrvl_cdip_100_examples_per_class_donut_v2
