datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
seggpt-example-datadocvqa_1200_examplesexternal_data_test_exampleexample-documentsexample-space-to-dataset-parquetdiffuman4d_example
Diffuman4D Example Test Data
Project Page | Paper | Code | Model
This repo provides several example scenes for testing Diffuman4D. Note that these scenes are not part of the official DNA-Rendering dataset. For more testing (and training) data, see this repo.
Usage
See the GitHub repo for detailed usage.
Cite
@inproceedings{jin2025diffuman4d,
title={Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models}… See the full description on the dataset page: https://huggingface.co/datasets/krahets/diffuman4d_example.z-image-examples
Z-Image Turbo Portrait Dataset
This dataset contains 126 portrait prompts and their corresponding image outputs, demonstrating the capabilities of the Z-Image Turbo text-to-image model.
Model Information
Model Name: Z-Image Turbo
Hugging Face Repository: Tongyi-MAI/Z-Image-Turbo
Dataset Contents
prompts.jsonl: A JSONL file containing the 126 text prompts used for generation. Each entry includes a unique ID and the prompt text.
outputs/: Directory containing… See the full description on the dataset page: https://huggingface.co/datasets/k-mktr/z-image-examples.external_data_test_example_v2docvqa_1200_examples_donutrvl_cdip_10_examples_per_classrvl_cdip_100_examples_per_class
Dataset Card for "rvl_cdip_100_examples_per_class"
More Information needed
rvl_cdip_300_examples_per_classchest-bench-example
ChestBench Example
DICOM-VLM Framework Reference Package v0.2.0
ChestBench Example is a four-case, DICOM-native reference package for developing and validating the data architecture of a medical vision-language model (VLM) pipeline.
It is intentionally small. Its purpose is to demonstrate how medical imaging data, annotations, text, knowledge, retrieval targets, QA, evidence requirements, perturbations, and audit metadata can be represented without confusing… See the full description on the dataset page: https://huggingface.co/datasets/NeeyuHuynh/chest-bench-example.PointSplat_example_data
PointSplat Example Data
This repository hosts compact example data for running PointSplat inference locally.
The zip archives are the recommended download format for reproducing the commands in the PointSplat README. The Dataset Viewer shows a small preview split with one fixed camera per scene:
DNA-Rendering: 16 validation scenes, camera 23, frame 000030.
THuman2.0: 10 validation scenes, view 000, frame 2.
Files
File
Content
Local target path… See the full description on the dataset page: https://huggingface.co/datasets/Yujie0012/PointSplat_example_data.rvl_cdip_10_examples_per_class_donutqwen-spatial-reasoning-incorrect-examples
Incorrect spatial reasoning examples for Qwen/Qwen3.5-0.8B-Base
Overview
This dataset contains incorrect non-empty predictions made by Qwen/Qwen3.5-0.8B-Base on a synthetic spatial reasoning benchmark built from 4x4 object-grid images.
I evaluated the model on 84 questions. It answered 54 of them incorrectly and achieved an overall accuracy of 35.714%. Some incorrect rows had an empty parsed pred_final, which I treat as formatting failures rather than useful supervised… See the full description on the dataset page: https://huggingface.co/datasets/safaeid48/qwen-spatial-reasoning-incorrect-examples.GraySpectrotram_example
Dataset Card for "GraySpectrotram_example"
More Information needed
bagel-exampleexample-dataset
ORB Transformation Applied on diffusiondb Dataset
This dataset consists of images, captions and images that are transformed to extract features using ORB transform.
You can find the original dataset here.
An example sample is below:
Caption: "spider - man, cinematic, photography "
Image:
Transformation:
bagel-example-originalobject-detection-examples
Dataset Card for CPPE - 5
Dataset Summary
CPPE - 5 (Medical Personal Protective Equipment) is a new challenging dataset with the goal to allow the study of subordinate categorization of medical personal protective equipments, which is not possible with other popular data sets that focus on broad level categories.
Some features of this dataset are:
high quality images and annotations (~4.6 bounding boxes per image)
real-life images unlike any current such dataset
majority… See the full description on the dataset page: https://huggingface.co/datasets/Charles95/object-detection-examples.rvl_cdip_100_examples_per_classenglish-exampleOnePromptOneStory-Examples-Vid-head75bagel-example-vlmexample1font-examples
Dataset Card for "font-examples"
More Information needed
example_dataset
Example Dataset
This is a robotics dataset converted from RLDS format to LeRobot format.
Dataset Information
Total Episodes: 40
Total Frames: 2659
Average Frames per Episode: 66.5
Robot Type: Panda
FPS: 10
Features
image: Main camera view (256x256x3)
wrist_image: Wrist camera view (256x256x3)
state: Robot state (7D)
action: Robot action (7D)
Usage
from lerobot.common.datasets.lerobot_dataset import LeRobotDataset
dataset =… See the full description on the dataset page: https://huggingface.co/datasets/Kanden1112/example_dataset.OnePromptOneStory-Examples-CCIPfranka_panda_serl_exampleThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "panda",
"total_episodes": 6,
"total_frames": 1001,
"total_tasks": 1,
"total_videos": 12,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 10,
"splits": {
"train": "0:6"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/andypeng05/franka_panda_serl_example.
