CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01wendlerc /RenderedTextThis dataset has been created by Stability AI and LAION. This dataset contains 12 million 1024x1024 images of handwritten text written on a digital 3D sheet of paper generated using Blender geometry nodes and rendered using Blender Cycles. The text has varying font size, color, and rotation, and the paper was rendered under random lighting conditions. Note that, the first 10 million examples are in the root folder of this dataset repository and the remaining 2 million are in ./remaining (due… See the full description on the dataset page: https://huggingface.co/datasets/wendlerc/RenderedText.imagetext-to-image10M<n<100M59 likes40k downloads11mo agoHugging Face02shhdwi /olmocr-pre-rendered olmOCR-bench Pre-Rendered Pre-rendered PNG images of the olmOCR-bench benchmark dataset, ready for zero-setup evaluation of any OCR / vision model. What This Is The official olmOCR benchmark requires downloading 1,403 PDFs locally and rendering each page to a PNG image before sending it to a model. Every benchmark runner in the official repo does this same rendering step internally — see olmocr/data/renderpdf.py::render_pdf_to_base64png(). This dataset eliminates that… See the full description on the dataset page: https://huggingface.co/datasets/shhdwi/olmocr-pre-rendered.image1K<n<10K0 likes29k downloads7mo agoHugging Face03Team-PIXEL /rendered-wikipedia-english Dataset Card for Team-PIXEL/rendered-wikipedia-english Dataset Summary This dataset contains the full English Wikipedia from February 1, 2018, rendered into images of 16x8464 resolution. The original text dataset was built from a Wikipedia dump. Each example in the original text dataset contained the content of one full Wikipedia article with cleaning to strip markdown and unwanted sections (references, etc.). Each rendered example contains a subset of one full article.… See the full description on the dataset page: https://huggingface.co/datasets/Team-PIXEL/rendered-wikipedia-english.text10M<n<100M4 likes2.6k downloads4y agoHugging Face04likaixin /IconStack-48M-Rendered-Traintext10M<n<100M2 likes2.4k downloads1y agoHugging Face05Team-PIXEL /rendered-bookcorpus Dataset Card for Team-PIXEL/rendered-bookcorpus Dataset Summary This dataset is a version of the BookCorpus available at https://huggingface.co/datasets/bookcorpusopen with examples rendered as images with resolution 16x8464 pixels. The original BookCorpus was introduced by Zhu et al. (2015) in Aligning Books and Movies: Towards Story-Like Visual Explanations by Watching Movies and Reading Books and contains 17868 books of various genres. The rendered BookCorpus was used… See the full description on the dataset page: https://huggingface.co/datasets/Team-PIXEL/rendered-bookcorpus.text1M<n<10M4 likes538 downloads4y agoHugging Face06cminst /transcoda-rendered-row-343k-full-pipeline-v1 Transcoda Rendered Row 343k Full Pipeline v1 Full-page rendered Transcoda row dataset generated from synthetic and random-notation transcriptions. Target contents: 343113 Accepted contents: 343027 Failed/dropped contents: 86 Renderings per accepted content: 4 Accepted images: 1372108 Source counts: {'random': 99990, 'synth': 243037} Staging repo: cminst/transcoda-rendered-row-343k-full-pipeline-v1-shards Each row contains one transcription and four independently rendered page… See the full description on the dataset page: https://huggingface.co/datasets/cminst/transcoda-rendered-row-343k-full-pipeline-v1.image100K<n<1M0 likes445 downloads28d agoHugging Face07Groosezzz /rendered-wikipedia-8x8-withTexttext10M<n<100M0 likes415 downloads3y agoHugging Face08SousiOmine /daruma-SFT-renderedtext100K<n<1M0 likes379 downloads2mo agoHugging Face09Groosezzz /rendered-bookcorpus-16x16text1M<n<10M0 likes317 downloads3y agoHugging Face10Groosezzz /rendered-wikipedia-en-8x8text10M<n<100M0 likes311 downloads3y agoHugging Face11nimapourjafar /mm_rendered_textimage10K<n<100K2 likes221 downloads2y agoHugging Face12ranjanhr1 /nayana-renderedimage1K<n<10K0 likes126 downloads10mo agoHugging Face13Groosezzz /rendered-bookcorpus-8x8-withTexttext1M<n<10M0 likes120 downloads3y agoHugging Face14v1v1d /NayanaBench-rendered-splitimageimage-to-text1K<n<10K0 likes111 downloads10mo agoHugging Face15thesantatitan /svg-rendered SVG to PNG Rendered Dataset Dataset Summary This dataset is a processed version of the svgen-500k-instruct dataset, where SVG images have been converted to PNG format for easier consumption in computer vision and machine learning pipelines. Each successfully converted image maintains the original SVG's visual representation while providing a standardized raster format. Data Fields png_processed: Boolean flag indicating whether the conversion was successful… See the full description on the dataset page: https://huggingface.co/datasets/thesantatitan/svg-rendered.text100K<n<1M1 likes88 downloads2y agoHugging Face16clip-benchmark /wds_renderedsst2image1K<n<10K0 likes80 downloads4y agoHugging Face17likaixin /IconStack-48M-Rendered-Devtext100K<n<1M1 likes65 downloads1y agoHugging Face18leonardPKU /Rendered_512_32_Testimage1K<n<10K0 likes58 downloads2y agoHugging Face19vlm-modality-research /gsm8k-rendered-vlm-v2 GSM8K Rendered-VL v2 1319 rendered GSM8K test problems for the VLM modality study (Phase 1). Contributors Rodela Ghosh — study design, pilot (v1), dataset packaging and Hugging Face release (scripts/prepare_hf_v2_release.py) Aviral Gupta — benchmark infrastructure (src/), v2 rendering protocol (src/rendering.py), Phase 1 model runs Code: https://github.com/Ro-netizen004/vlm-modality-research Not interchangeable with v1: RodelaG/gsm8k-rendered-vlm v1 v2… See the full description on the dataset page: https://huggingface.co/datasets/vlm-modality-research/gsm8k-rendered-vlm-v2.image1K<n<10K0 likes56 downloads3mo agoHugging Face20imagine-io /Rendered_Room_Dataset_Sampletabularn<1K0 likes55 downloads8mo agoHugging Face21thesantatitan /svg-rendered-blip_captioned SVG to PNG Rendered Dataset Dataset Summary This dataset is a processed version of the svgen-500k-instruct dataset, where SVG images have been converted to PNG format for easier consumption in computer vision and machine learning pipelines. Each successfully converted image maintains the original SVG's visual representation while providing a standardized raster format. Data Fields caption: Image captions generated using Salesforce's BLIP model png_processed:… See the full description on the dataset page: https://huggingface.co/datasets/thesantatitan/svg-rendered-blip_captioned.text100K<n<1M2 likes42 downloads1y agoHugging Face22sopheakvoatei /rendered-kh-table-suryaocr2imagen<1K0 likes41 downloads29d agoHugging Face23geodesic-research /pa-warm-start-sft-25b-rendered-review pa-warm-start-sft-25b rendered review sample (n=200) 200 uniformly-sampled conversations from geodesic-research/pa-warm-start-sft-heavy-25b-mix (default/train, the control-pretraining 30B baseline SFT corpus), rendered EXACTLY as the training pack renders them: the library's _chat_preprocess (tool-call normalization + think-HISTORY chat template + assistant-only loss mask). Columns: rendered_text (the full string the model sees), trainable_spans_only (concatenation of… See the full description on the dataset page: https://huggingface.co/datasets/geodesic-research/pa-warm-start-sft-25b-rendered-review.tabularn<1K0 likes41 downloads26d agoHugging Face24shalunov /mentis-cad-recode-renderedtext1K<n<10K0 likes36 downloads1y agoHugging Face25sopheakvoatei /rendered_kh_tables_testlong_surya Surya OCR 2 Table Recognition on sopheakvoatei/rendered_khmer_tables Surya OCR 2 table recognition using offline vLLM inference. The table pipeline automatically detects tall table images and applies vertical overlapping tiling before reconstructing the final HTML table. Processing Details Source Dataset: sopheakvoatei/rendered_khmer_tables Model: datalab-to/surya-ocr-2 Task: table Table mode: full Input column: image Output column: markdown Structured column:… See the full description on the dataset page: https://huggingface.co/datasets/sopheakvoatei/rendered_kh_tables_testlong_surya.imagen<1K0 likes32 downloads1mo agoHugging Face26djghosh /wds_renderedsst2_test Rendered SST2 (Test set only) Original paper: The Visual Task Adaptation Benchmark Homepage: https://github.com/openai/CLIP/blob/main/data/rendered-sst2.md Derived from SST2: https://nlp.stanford.edu/sentiment/treebank.html Bibtex: @article{zhai2019visual, title={The Visual Task Adaptation Benchmark}, author={Xiaohua Zhai and Joan Puigcerver and Alexander Kolesnikov and Pierre Ruyssen and Carlos Riquelme and Mario Lucic and Josip… See the full description on the dataset page: https://huggingface.co/datasets/djghosh/wds_renderedsst2_test.textn<1K0 likes31 downloads4y agoHugging Face27mteb /wds_renderedsst2image1K<n<10K0 likes30 downloads7mo agoHugging Face28deepcopy /text2svg-stack-renderedimage100K<n<1M1 likes29 downloads1y agoHugging Face29haideraltahan /wds_renderedsst2image1K<n<10K0 likes23 downloads2y agoHugging Face30sopheakvoatei /rendered_khmer_tablesimagen<1K0 likes18 downloads1mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.