datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
Hunyuan3D-FLUX-Gen
Orient Anything V2 Dataset
Project Page | Paper | GitHub
Orient Anything V2 is an enhanced foundation model for unified understanding of object 3D orientation and rotation from single or paired images. This dataset repository supports the model by providing assets for orientation estimation, 6DoF pose estimation, and object symmetry recognition.
Data Preparation
You can download the absolute orientation, relative rotation, and symm-orientation test datasets using the… See the full description on the dataset page: https://huggingface.co/datasets/Viglong/Hunyuan3D-FLUX-Gen.HD-Mixkit-Finetune-Hunyuantest-HunyuanVideo-pixelart-videos
trojblue/test-HunyuanVideo-pixelart-images
👋 Heads up—this repository is just a PARTIAL dataset. For the full pixelart-images dataset, make sure to grab both parts:
Images Part
Video Part (this repo)
What's in the Dataset?
This dataset is all about anime-styled pixel art images that have been carefully selected to make your models shine. Here’s what makes these images special:
Rich in detail: Pixelated, yes—but still full of life and not overly simplified.… See the full description on the dataset page: https://huggingface.co/datasets/trojblue/test-HunyuanVideo-pixelart-videos.HunyuanImage-2.1_t2i_human_preference
Rapidata Hunyuan Image 2.1 Preference
This T2I dataset contains over ~400'000 human responses from over ~50'000 individual annotators, collected in less than 7h using the Rapidata Python API, accessible to anyone and ideal for large scale evaluation.
Evaluating Hunyuan Image 2.1 (version from 19.9.2025) across three categories: preference, coherence, and alignment.
Explore our latest model rankings on our website.
If you get value from this dataset and would like to see more in… See the full description on the dataset page: https://huggingface.co/datasets/Rapidata/HunyuanImage-2.1_t2i_human_preference.hunyuan1.5_training中文文档
HunyuanVideo-1.5
🎬 HunyuanVideo-1.5: A leading lightweight video generation model
HunyuanVideo-1.5 is a video generation model that delivers top-tier quality with only 8.3B parameters, significantly lowering the barrier to usage. It runs smoothly on consumer-grade GPUs, making it accessible for every developer and creator. This repository provides the implementation and tools needed to generate creative videos.
👏… See the full description on the dataset page: https://huggingface.co/datasets/xingzhaohu/hunyuan1.5_training.Hunyuanamerican-residential-housing-hunyuan3d-glb
American Building Hunyuan3D GLB Assets
This dataset contains 40 generated American building assets:
10 residential buildings
15 office buildings
15 industrial buildings
Each asset includes:
source_images/*.png: isolated building reference image on a white background.
glbs/*.glb: final textured GLB generated with Hunyuan3D-2.1.
metadata/*.json: per-asset generation parameters and timings.
verification_summary.json: automated validation summary for all final GLBs.… See the full description on the dataset page: https://huggingface.co/datasets/wchen99998/american-residential-housing-hunyuan3d-glb.HD-Hunyuan-30K-Distill-DataHunyuanWorld-panoramastest-HunyuanVideo-pixelart-videos
Drive From trojblue/test-HunyuanVideo-pixelart-videos
Reorganized version of Wild-Heart/Disney-VideoGeneration-Dataset. This is needed for Mochi-1 fine-tuning.
awesome_hunyuanImage_prompts
Awesome HunyuanImage Prompts
Built with Hugging Face AI Sheets
Do you want to master HunyuanImage 3, one of the best open models for generating images? This resource is for you.
HunyuanImage 3's Prompt Handbook is a great resource for learning how to prompt the model. Unfortunately, it is in Chinese, so I have created this resource for the open community.
It contains:
All prompts in HunyuanImage 3's Prompt Handbook are organized by category.
Their translation into English… See the full description on the dataset page: https://huggingface.co/datasets/dvilasuero/awesome_hunyuanImage_prompts.awesome_hunyuanImage_prompts-genHUNYUANLORAwan_gen_videos_HunyuanVideo_Foley_no_vocal_sound_captionedhunyuan-lorashunyuanawesome_hunyuanImage_prompts-genhunyuan-image3-dit-attn-capture
HunyuanImage-3.0 DiT attention capture — handoff
Real attention inputs (q/k/v + mask) captured from the HunyuanImage-3.0 DiT, for
kernel work: choosing/implementing an attention backend that can serve this
model. The mask here is structural, not a padding mask, so it is the part
that constrains what a kernel can accept.
What's in this bundle
File
Size
Contents
out/call0_rank0.npz
16.5 MiB
prefill — q/k/v + mask
out/call32_rank0.npz
14.0 MiB
denoise —… See the full description on the dataset page: https://huggingface.co/datasets/Yi30/hunyuan-image3-dit-attn-capture.gen-videos-hunyuanvideoHunyuanVideo-WorkflowsXiang_hunyuan_1_5_i2v_videos
refer image
test-HunyuanVideo-anime-images
Test-HunyuanVideo-Anime-Stills
A small dataset of AI-generated anime-themed images designed for general anime text-to-image (T2I) training debug or testing Hunyuan Video. This dataset provides a balanced distribution of subjects and aims to align large pretrained models with anime aesthetics in terms of visual appeal and text faithfulness.
Subject Selection
The subject distributions (other than the 50% anime girls) are selected based on the policy outlined in Meta's Emu… See the full description on the dataset page: https://huggingface.co/datasets/trojblue/test-HunyuanVideo-anime-images.test-HunyuanVideo-pixelart-images
trojblue/test-HunyuanVideo-pixelart-images
Hey there! 👋 Heads up—this repository is just a PARTIAL dataset. For the full pixelart-images dataset, make sure to grab both parts:
Images Part (this repo)
Video Part
This dataset is a collection of anime-style pixel art images and is perfect for debugging general anime text-to-image (T2I) training or testing Hunyuan Video models. 🎨
What's in the Dataset?
This dataset is all about anime-styled pixel art images that have… See the full description on the dataset page: https://huggingface.co/datasets/trojblue/test-HunyuanVideo-pixelart-images.awesome_hunyuan_image_prompts_3Genshin_Impact_Characters_3D_Chibi_HunyuanPortraitufo-colpali-hunyuan-ocr
Document OCR using HunyuanOCR
This dataset contains OCR results from images in davanstrien/ufo-ColPali using HunyuanOCR, a lightweight 1B VLM from Tencent.
Processing Details
Source Dataset: davanstrien/ufo-ColPali
Model: tencent/HunyuanOCR
Number of Samples: 10
Processing Time: 2.9 min
Processing Date: 2025-11-25 11:12 UTC
Configuration
Image Column: image
Output Column: markdown
Dataset Split: train
Batch Size: 16
Prompt Mode: parse-document
Prompt… See the full description on the dataset page: https://huggingface.co/datasets/davanstrien/ufo-colpali-hunyuan-ocr.HunyuanVideo-pixelart-videos-sample
trojblue/test-HunyuanVideo-pixelart-images
👋 Heads up—this repository is just a PARTIAL dataset. For the full pixelart-images dataset, make sure to grab both parts:
Images Part
Video Part (this repo)
What's in the Dataset?
This dataset is all about anime-styled pixel art images that have been carefully selected to make your models shine. Here’s what makes these images special:
Rich in detail: Pixelated, yes—but still full of life and not overly simplified.… See the full description on the dataset page: https://huggingface.co/datasets/inlineresearch/HunyuanVideo-pixelart-videos-sample.Hunyuan_Foleycog-hunyuan-f50-datasetawesome_hunyuan_image_prompts
