datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
cmp-v6-base108-renderRenderedTextThis dataset has been created by Stability AI and LAION.
This dataset contains 12 million 1024x1024 images of handwritten text written on a digital 3D sheet of paper generated using Blender geometry nodes and rendered using Blender Cycles. The text has varying font size, color, and rotation, and the paper was rendered under random lighting conditions.
Note that, the first 10 million examples are in the root folder of this dataset repository and the remaining 2 million are in ./remaining (due… See the full description on the dataset page: https://huggingface.co/datasets/wendlerc/RenderedText.olmocr-pre-rendered
olmOCR-bench Pre-Rendered
Pre-rendered PNG images of the olmOCR-bench benchmark dataset, ready for zero-setup evaluation of any OCR / vision model.
What This Is
The official olmOCR benchmark requires downloading 1,403 PDFs locally and rendering each page to a PNG image before sending it to a model. Every benchmark runner in the official repo does this same rendering step internally — see olmocr/data/renderpdf.py::render_pdf_to_base64png().
This dataset eliminates that… See the full description on the dataset page: https://huggingface.co/datasets/shhdwi/olmocr-pre-rendered.3dfront_render_viewsObjaverse-PBR-render
Objaverse-PBR-render
PBR rendered condition videos for Ink3D — a 3D texture generation pipeline using video generative models.
📄 Paper: Ink3D: Sculpting 3D Assets with Extremely Complex Textures via Video Generative Models💻 Code: github.com/YueHan99/Ink3D.TextureGen🤗 Model: Yuehavingfun/orbitpainter-single
This dataset provides pre-rendered geometry condition videos (position, normal, albedo, RGB, depth, mask) for ~23,000 Objaverse models, rendered from horizontal (H) and… See the full description on the dataset page: https://huggingface.co/datasets/Yuehavingfun/Objaverse-PBR-render.differentiable-render-camouflage-data
DRC CARLA Multi-Vehicle Camouflage Dataset
This repository releases the audited synthetic data and geometry assets used to
study whether one differentiable vehicle-camouflage generator transfers across
vehicle shapes. The release preserves the original split manifests, collection
protocols, audit records, and SHA-256 checksums. It is intended for reproducible
adversarial-robustness research, including the analysis of negative results.
Release contents… See the full description on the dataset page: https://huggingface.co/datasets/bailuyucha/differentiable-render-camouflage-data.3dfront-render-views3dfront-render-diffuse3dfront_renderrendered-wikipedia-english
Dataset Card for Team-PIXEL/rendered-wikipedia-english
Dataset Summary
This dataset contains the full English Wikipedia from February 1, 2018, rendered into images of 16x8464 resolution.
The original text dataset was built from a Wikipedia dump. Each example in the original text dataset contained the content of one full Wikipedia article with cleaning to strip markdown and unwanted sections (references, etc.). Each rendered example contains a subset of one full article.… See the full description on the dataset page: https://huggingface.co/datasets/Team-PIXEL/rendered-wikipedia-english.IconStack-48M-Rendered-Traintranscoda-rendered-row-343k-full-pipeline-v1-shardsrendered-sst2
Rendered SST-2
The Rendered SST-2 Dataset from Open AI.
Rendered SST2 is an image classification dataset used to evaluate the models capability on optical character recognition. This dataset was generated by rendering sentences in the Standford Sentiment Treebank v2 dataset.
This dataset contains two classes (positive and negative) and is divided in three splits: a train split containing 6920 images (3610 positive and 3310 negative), a validation split containing 872 images (444… See the full description on the dataset page: https://huggingface.co/datasets/nateraw/rendered-sst2.i1-rendered_text-tfrecordi1: A Simple and Fully Open Recipe for Strong Text-to-Image Models
Boya Zeng, Tianze Luo, Shu Pu, Jucheng Shen, Taiming Lu, Gabriel Sarch, Zhuang Liu
Princeton University
[arXiv][code][model][project page]
Overview
To prepare the dataset for training, we store the image-caption pairs as TFRecords.
This HuggingFace dataset contains the TFRecords corresponding to the rendered_text dataset at 256×256 resolution.
It also serves as an example of what a dataset processed using… See the full description on the dataset page: https://huggingface.co/datasets/i1-datasets/i1-rendered_text-tfrecord.i1-rendered_text-512-resolution-1m-tfrecordi1: A Simple and Fully Open Recipe for Strong Text-to-Image Models
Boya Zeng, Tianze Luo, Shu Pu, Jucheng Shen, Taiming Lu, Gabriel Sarch, Zhuang Liu
Princeton University
[arXiv][code][model][project page]
Overview
To prepare the dataset for training, we store the image-caption pairs as TFRecords.
This HuggingFace dataset contains the TFRecords corresponding to the rendered_text dataset at 512×512 resolution. Concretely, we only retain raw images with a shorter edge of… See the full description on the dataset page: https://huggingface.co/datasets/i1-datasets/i1-rendered_text-512-resolution-1m-tfrecord.dna_rendering_processed
DNA-Rendering-Processed Dataset
Project Page | Paper | Code | Model
To enable Diffuman4D model training, we meticulously process the DNA-Rendering dataset by recalibrating camera parameters, optimizing image color correction matrices (CCMs), predicting foreground masks, and estimating human skeletons.
To promote future research in the field of human-centric 3D/4D generation, we have open-sourced our re-annotated labels for the DNA-Rendering dataset in this repo, which includes… See the full description on the dataset page: https://huggingface.co/datasets/krahets/dna_rendering_processed.Objaverse_render_random3d-front-indoor-renders
Indoor Scene Renders from 3D-FRONT / 3D-FUTURE
20,240 photo-realistic indoor scene renders with instance segmentation, 6DoF
object poses, camera intrinsics and depth — rendered from the 3D-FRONT scene
layouts and 3D-FUTURE furniture models.
This is not a copy of the original 3D-FUTURE render set. It is a separate
render set built from the same assets. See Differences from the original below.
Why this exists
The 3D-FUTURE technical report describes 20,240 rendered… See the full description on the dataset page: https://huggingface.co/datasets/Spatial1ntelligence/3d-front-indoor-renders.objaverse_rendering_setobjaverse_orbit_rendersramanv-image-real-3d-rendersObjaverse-XL-Rigged-Animated-Renders
Objaverse-XL Rigged & Animated — Renders
Visual companion to
Linzhan/Objaverse-XL-Rigged-Animated,
which holds the 7,373 rigged-and-animated GLB assets themselves. This repository holds only what
was rendered from them: a four-view video of every animation clip, and a rest-pose grid per asset.
They live apart from the assets because they are bulky and numerous — 10,355 clip folders — while
the asset repo stays a compact 7,373 GLBs plus two tables. Nothing here is needed to use… See the full description on the dataset page: https://huggingface.co/datasets/Linzhan/Objaverse-XL-Rigged-Animated-Renders.ObjaverseXL_github_rendersrendered-bookcorpus-8x8AIM2024-SparseNeuralRendering
Dataset
This repository cofntains SpaRe (Sparse Rendering) dataset.
Associated paper
The dataset contained in this repository is published as a part of AIM workshop at ECCV 2024.
rendered-bookcorpus
Dataset Card for Team-PIXEL/rendered-bookcorpus
Dataset Summary
This dataset is a version of the BookCorpus available at https://huggingface.co/datasets/bookcorpusopen with examples rendered as images with resolution 16x8464 pixels.
The original BookCorpus was introduced by Zhu et al. (2015) in Aligning Books and Movies: Towards Story-Like Visual Explanations by Watching Movies and Reading Books and contains 17868 books of various genres. The rendered BookCorpus was used… See the full description on the dataset page: https://huggingface.co/datasets/Team-PIXEL/rendered-bookcorpus.rendered-bookcorpus-bigramsHDTF_static_renderer_train_onlyomnidocbench-render-compare
OmniDocBench Render-and-Compare
This dataset contains the rendered HTML reconstructions and comparison images produced
by a render-and-compare pipeline — a reference-free visual similarity evaluation
framework for OCR systems.
Overview
The pipeline processes each page of OmniDocBench through
a Qwen3.5-122B-A10B OCR model, renders the structured output back to a PNG via HTML
(reconstructed.png), and compares it against the original page scan (masked_original.png)
using… See the full description on the dataset page: https://huggingface.co/datasets/gt-free-ocr-metrics/omnidocbench-render-compare.render
