CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01Team-PIXEL /rendered-wikipedia-english Dataset Card for Team-PIXEL/rendered-wikipedia-english Dataset Summary This dataset contains the full English Wikipedia from February 1, 2018, rendered into images of 16x8464 resolution. The original text dataset was built from a Wikipedia dump. Each example in the original text dataset contained the content of one full Wikipedia article with cleaning to strip markdown and unwanted sections (references, etc.). Each rendered example contains a subset of one full article.… See the full description on the dataset page: https://huggingface.co/datasets/Team-PIXEL/rendered-wikipedia-english.text10M<n<100M4 likes2.8k downloads4y agoHugging Face02likaixin /IconStack-48M-Rendered-Traintext10M<n<100M2 likes2.5k downloads1y agoHugging Face03krahets /dna_rendering_processedgated DNA-Rendering-Processed Dataset Project Page | Paper | Code | Model To enable Diffuman4D model training, we meticulously process the DNA-Rendering dataset by recalibrating camera parameters, optimizing image color correction matrices (CCMs), predicting foreground masks, and estimating human skeletons. To promote future research in the field of human-centric 3D/4D generation, we have open-sourced our re-annotated labels for the DNA-Rendering dataset in this repo, which includes… See the full description on the dataset page: https://huggingface.co/datasets/krahets/dna_rendering_processed.imageimage-to-3d1K<n<10K9 likes1.3k downloads10mo agoHugging Face04lingamvamshikrishnareddy /ramanv-image-real-3d-renderstext10K<n<100K0 likes796 downloads23d agoHugging Face05PosterCraft /Text-Render-2Mgated Text Render 2M Dataset A large-scale dataset containing 2 million text rendering image-text pairs for training generative models to improve text rendering performance. Dataset Structure image: Rendered text image in PNG format text: Corresponding text content file_name: Original filename folder_id: Folder identifier Usage This dataset is designed for fine-tuning generative models to improve text rendering capabilities. from datasets import load_dataset… See the full description on the dataset page: https://huggingface.co/datasets/PosterCraft/Text-Render-2M.texttext-to-image1M<n<10M16 likes620 downloads1y agoHugging Face06Team-PIXEL /rendered-bookcorpus Dataset Card for Team-PIXEL/rendered-bookcorpus Dataset Summary This dataset is a version of the BookCorpus available at https://huggingface.co/datasets/bookcorpusopen with examples rendered as images with resolution 16x8464 pixels. The original BookCorpus was introduced by Zhu et al. (2015) in Aligning Books and Movies: Towards Story-Like Visual Explanations by Watching Movies and Reading Books and contains 17868 books of various genres. The rendered BookCorpus was used… See the full description on the dataset page: https://huggingface.co/datasets/Team-PIXEL/rendered-bookcorpus.text1M<n<10M4 likes533 downloads4y agoHugging Face07cminst /transcoda-rendered-row-343k-full-pipeline-v1 Transcoda Rendered Row 343k Full Pipeline v1 Full-page rendered Transcoda row dataset generated from synthetic and random-notation transcriptions. Target contents: 343113 Accepted contents: 343027 Failed/dropped contents: 86 Renderings per accepted content: 4 Accepted images: 1372108 Source counts: {'random': 99990, 'synth': 243037} Staging repo: cminst/transcoda-rendered-row-343k-full-pipeline-v1-shards Each row contains one transcription and four independently rendered page… See the full description on the dataset page: https://huggingface.co/datasets/cminst/transcoda-rendered-row-343k-full-pipeline-v1.image100K<n<1M0 likes509 downloads29d agoHugging Face08SousiOmine /daruma-SFT-renderedtext100K<n<1M0 likes468 downloads2mo agoHugging Face09Groosezzz /rendered-wikipedia-8x8-withTexttext10M<n<100M0 likes415 downloads3y agoHugging Face10VietMedTeam /veri-render VeriRender Benchmark Dataset Causal consistency verification samples for Vision-Language Models. Layout manifest.jsonl ← canonical index (one row per sample) benchmark.yaml ← config used to generate this release inconsistent/{domain}/{sample_id}/ ← corrupted evaluation samples consistent/{domain}/{sample_id}/ ← negative controls (clean images) Splits Split Description Eval image inconsistent Symbolic spec is… See the full description on the dataset page: https://huggingface.co/datasets/VietMedTeam/veri-render.imagevisual-question-answeringn<1K1 likes401 downloads3mo agoHugging Face11Groosezzz /rendered-wikipedia-en-8x8text10M<n<100M0 likes363 downloads3y agoHugging Face12Groosezzz /rendered-bookcorpus-16x16text1M<n<10M0 likes355 downloads3y agoHugging Face13ciderlab /amex_render_pairs_v5image10K<n<100K0 likes317 downloads1mo agoHugging Face14ciderlab /amex_render_pairs_v3image10K<n<100K0 likes266 downloads3mo agoHugging Face15nimapourjafar /mm_rendered_textimage10K<n<100K2 likes242 downloads2y agoHugging Face16ranjanhr1 /nayana-renderedimage1K<n<10K0 likes124 downloads10mo agoHugging Face17Groosezzz /rendered-bookcorpus-8x8-withTexttext1M<n<10M0 likes123 downloads3y agoHugging Face18v1v1d /NayanaBench-rendered-splitimageimage-to-text1K<n<10K0 likes117 downloads10mo agoHugging Face19thesantatitan /svg-rendered SVG to PNG Rendered Dataset Dataset Summary This dataset is a processed version of the svgen-500k-instruct dataset, where SVG images have been converted to PNG format for easier consumption in computer vision and machine learning pipelines. Each successfully converted image maintains the original SVG's visual representation while providing a standardized raster format. Data Fields png_processed: Boolean flag indicating whether the conversion was successful… See the full description on the dataset page: https://huggingface.co/datasets/thesantatitan/svg-rendered.text100K<n<1M1 likes88 downloads2y agoHugging Face20achang /render_text_largeimage10K<n<100K1 likes85 downloads4y agoHugging Face21gt-free-ocr-metrics /omnidocbench-render-compare-parquet OmniDocBench Render-and-Compare — Parquet Edition Parquet-shard repackaging of gt-free-ocr-metrics/omnidocbench-render-compare. Overview The pipeline processes each page of OmniDocBench through a Qwen3.5-122B-A10B OCR model, renders the structured output back to a PNG via HTML (reconstructed), and compares it against the original page scan (masked_original) using reference-free visual metrics. Five OCR extraction variants are provided, each targeting a different subset of… See the full description on the dataset page: https://huggingface.co/datasets/gt-free-ocr-metrics/omnidocbench-render-compare-parquet.textother1K<n<10K0 likes85 downloads5mo agoHugging Face22likaixin /IconStack-48M-Rendered-Devtext100K<n<1M1 likes66 downloads1y agoHugging Face23davidberenstein1957 /text_rendering text_rendering text_rendering Trigger token: sks_textrender Examples: 111 Format: p-image Source: /Users/davidberenstein/Documents/programming/pruna/dataset-generator/training/text-rendering.zip Use input.zip with p-image-trainer (Replicate). See TRAINING_PLAN.md in this directory. Format Trainer: p-image-trainer Schema: See config.yml and TRAINING_PLAN.md in this repo. Reproduce generate.py in this repo documents how to regenerate this dataset… See the full description on the dataset page: https://huggingface.co/datasets/davidberenstein1957/text_rendering.imagen<1K0 likes64 downloads7mo agoHugging Face24physicl-test /opencode-public-data-pack-docker_input1-20-renders-512-20260612t125042z OpenCode Public Data Pack docker_input1 20 renders 512 20260612T125042Z Public data pack created from docker_input1.json with 20 renders at 512x512. This dataset mirrors public data-pack render outputs from Physicl. Each row represents one render view. The image column contains a stable URL to the primary render image uploaded under /data; image_path stores the relative repository path and data_commit_sha pins the Hugging Face dataset commit used by those URLs. Files are… See the full description on the dataset page: https://huggingface.co/datasets/physicl-test/opencode-public-data-pack-docker_input1-20-renders-512-20260612t125042z.imagen<1K0 likes60 downloads3mo agoHugging Face25leopoldmaillard /sceneteract-renders SceneTeract Scene Renders The images a VLM is shown when judging whether an activity is physically feasible in a 3D indoor scene. Pair these with sceneteract-traces to evaluate a model against geometry-grounded labels. 3,621 renders over 1,132 3D-FRONT living rooms and dining rooms, 1024×1024 PNG. from datasets import load_dataset renders = load_dataset("leopoldmaillard/sceneteract-renders")["train"] renders[0]["image"] # PIL image Two kinds of image… See the full description on the dataset page: https://huggingface.co/datasets/leopoldmaillard/sceneteract-renders.image1K<n<10K0 likes58 downloads6d agoHugging Face26vlm-modality-research /gsm8k-rendered-vlm-v2 GSM8K Rendered-VL v2 1319 rendered GSM8K test problems for the VLM modality study (Phase 1). Contributors Rodela Ghosh — study design, pilot (v1), dataset packaging and Hugging Face release (scripts/prepare_hf_v2_release.py) Aviral Gupta — benchmark infrastructure (src/), v2 rendering protocol (src/rendering.py), Phase 1 model runs Code: https://github.com/Ro-netizen004/vlm-modality-research Not interchangeable with v1: RodelaG/gsm8k-rendered-vlm v1 v2… See the full description on the dataset page: https://huggingface.co/datasets/vlm-modality-research/gsm8k-rendered-vlm-v2.image1K<n<10K0 likes57 downloads3mo agoHugging Face27thesantatitan /svg-rendered-blip_captioned SVG to PNG Rendered Dataset Dataset Summary This dataset is a processed version of the svgen-500k-instruct dataset, where SVG images have been converted to PNG format for easier consumption in computer vision and machine learning pipelines. Each successfully converted image maintains the original SVG's visual representation while providing a standardized raster format. Data Fields caption: Image captions generated using Salesforce's BLIP model png_processed:… See the full description on the dataset page: https://huggingface.co/datasets/thesantatitan/svg-rendered-blip_captioned.text100K<n<1M2 likes42 downloads1y agoHugging Face28geodesic-research /pa-warm-start-sft-25b-rendered-review pa-warm-start-sft-25b rendered review sample (n=200) 200 uniformly-sampled conversations from geodesic-research/pa-warm-start-sft-heavy-25b-mix (default/train, the control-pretraining 30B baseline SFT corpus), rendered EXACTLY as the training pack renders them: the library's _chat_preprocess (tool-call normalization + think-HISTORY chat template + assistant-only loss mask). Columns: rendered_text (the full string the model sees), trainable_spans_only (concatenation of… See the full description on the dataset page: https://huggingface.co/datasets/geodesic-research/pa-warm-start-sft-25b-rendered-review.tabularn<1K0 likes42 downloads28d agoHugging Face29spectralbranding /meaningfulness-cross-language-rendering Cross-Language Rendering for Meaning vs Meaningfulness (Paper B 2026ap) HF dataset DOI: 10.57967/hf/8971 Companion paper concept DOI: 10.5281/zenodo.20409701 Companion GitHub mirror: https://github.com/spectralbranding/meaningfulness-papers/tree/main/meaning-meaningfulness-empirical Dataset Summary This dataset contains the multi-language rendering and extraction artifacts demonstrating Proposition P4 (rendering-equivalence under spine-preservation) from Zharnikov… See the full description on the dataset page: https://huggingface.co/datasets/spectralbranding/meaningfulness-cross-language-rendering.tabulartext-generationn<1K0 likes40 downloads2mo agoHugging Face30disi-unibo-nlp-students /chess_render_360image1K<n<10K0 likes38 downloads7mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.