CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01LindseyLarson3372 /imagesimage10K<n<100K6 likes33k downloads2mo agoHugging Face02Lin-Chen /MMStar MMStar (Are We on the Right Way for Evaluating Large Vision-Language Models?) 🌐 Homepage | 🤗 Dataset | 🤗 Paper | 📖 arXiv | GitHub Dataset Details As shown in the figure below, existing benchmarks lack consideration of the vision dependency of evaluation samples and potential data leakage from LLMs' and LVLMs' training data. Therefore, we introduce MMStar: an elite vision-indispensible multi-modal benchmark, aiming to ensure each curated sample exhibits… See the full description on the dataset page: https://huggingface.co/datasets/Lin-Chen/MMStar.imagemultiple-choice1K<n<10K53 likes18k downloads2y agoHugging Face03lin-zhao-resoLve /D3HRimage10K<n<100K0 likes13k downloads1y agoHugging Face04lin7zhi /Ownerimage10K<n<100K0 likes7.1k downloads11mo agoHugging Face05ityizNola /Anime-LineArt-Dataset Anime Lineart Sketch Dataset Dataset Summary This dataset contains high-quality lineart sketch images automatically extracted from raw anime images using the LineartAnimeDetector model from ControlNet Annotators (lllyasviel/Annotators). It is designed to support research and development in: Anime-style sketch generation Text-to-sketch pipelines ControlNet conditioning Sketch-to-image and image-to-sketch translation The raw source images are sourced from the Anime Images… See the full description on the dataset page: https://huggingface.co/datasets/ityizNola/Anime-LineArt-Dataset.imageimage-to-image10K<n<100K0 likes6.1k downloads4mo agoHugging Face06Linzhan /UniML3D UniML3D UniML3D is the text-paired, topology-annotated motion dataset behind UniMate (SIGGRAPH Asia 2026): motion clips from three sources with very different skeletons — Mixamo humanoids, Truebones ZOO animals and rigged Objaverse-XL objects — brought into one canonical layout, captioned, and annotated with cleaned joint names, a body-plan category and a facing-direction joint pair per skeleton. Every annotation in it was generated by this project's own data… See the full description on the dataset page: https://huggingface.co/datasets/Linzhan/UniML3D.imagetext-to-3d10K<n<100K8 likes4.9k downloads2d agoHugging Face07Teklia /IAM-line IAM - line level Dataset Summary The IAM Handwriting Database contains forms of handwritten English text which can be used to train and test handwritten text recognizers and to perform writer identification and verification experiments. Note that all images are resized to a fixed height of 128 pixels. Languages All the documents in the dataset are written in English. Dataset Structure Data Instances { 'image':… See the full description on the dataset page: https://huggingface.co/datasets/Teklia/IAM-line.imageimage-to-text10K<n<100K34 likes3.4k downloads3y agoHugging Face08Link-Dev /UAV3DCrop UAV3DCrop UAV3DCrop is a multi-year UAV crop dataset containing field imagery and associated reconstruction metadata for agricultural research. Project page: UAV3DCrop Paper: UAV3DCrop: Benchmarking 3D Reconstruction in Repeated Multi-Angle UAV Crop Surveys Code: GitHub Companion depth dataset: UAV3DCrop Depth Repository Contents The dataset is organized by year and acquisition day. A scene may include: RGB images in an images/ directory; camera parameters and… See the full description on the dataset page: https://huggingface.co/datasets/Link-Dev/UAV3DCrop.imageimage-to-3d10K<n<100K0 likes2.4k downloads1mo agoHugging Face09Three-Liners /UNS-SSDS-2025Dataset from UNS SSDS 2025 Competition image0 likes2.3k downloads1y agoHugging Face10links-ads /spada-dataset SPADA Dataset This dataset contains images and sparse labels used in the paper Land Cover Segmentation with Sparse Annotations from Sentinel-2 Imagery , published at IGARSS 2023. Repository: https://github.com/links-ads/igarss-spada Paper: https://paperswithcode.com/paper/land-cover-segmentation-with-sparse Dataset Preparation The dataset has been compressed into segmented tarballs for ease of use within Git LFS (that is, tar > gzip > split). To revert the process… See the full description on the dataset page: https://huggingface.co/datasets/links-ads/spada-dataset.imageimage-segmentation10M<n<100M1 likes2.3k downloads2y agoHugging Face11HabibaAbderrahim /Tunisian-Proverbs-with-Image-Associations-A-Cultural-and-Linguistic-DatasetTunisian Proverbs with Image Associations: A Cultural and Linguistic Dataset Description This dataset explores the rich oral tradition of Tunisian proverbs mapped into text format, pairing each with contextual explanations, English translations both word-to-word and it's equivalent Target Language dynamic, Automated prompt and AI-generated visual interpretations. It bridges linguistic, cultural, and visual modalities making it valuable for tasks in cross-cultural NLP, generative… See the full description on the dataset page: https://huggingface.co/datasets/HabibaAbderrahim/Tunisian-Proverbs-with-Image-Associations-A-Cultural-and-Linguistic-Dataset.imagetranslationn<1K0 likes2.2k downloads1y agoHugging Face12Voxel51 /lingbot-depth-subset Dataset Card for lingbot-depth-subset This is a FiftyOne dataset with 13,149 samples (10,207 groups) spanning 3 sub-collections (RobbyReal, RobbyVla, RobbySim). Installation If you haven't already, install FiftyOne: pip install -U fiftyone Usage import fiftyone as fo from fiftyone.utils.huggingface import load_from_hub # Load the dataset # Note: other available arguments include 'max_samples', etc dataset =… See the full description on the dataset page: https://huggingface.co/datasets/Voxel51/lingbot-depth-subset.imagedepth-estimation10K<n<100K1 likes2.2k downloads2mo agoHugging Face13ling99 /OCRBench_v2image10K<n<100K20 likes2k downloads2y agoHugging Face14Link-Dev /UAV3DCrop_depth UAV3DCrop Depth UAV3DCrop Depth contains per-image depth maps associated with the UAV3DCrop multi-year agricultural UAV dataset. Project page: UAV3DCrop Paper: UAV3DCrop: Benchmarking 3D Reconstruction in Repeated Multi-Angle UAV Crop Surveys Code: UAV3DCrop GitHub RGB dataset: UAV3DCrop Repository Contents The depth data are organized by year and acquisition scene. Scene directories contain TIFF depth maps named to correspond to images in the main UAV3DCrop… See the full description on the dataset page: https://huggingface.co/datasets/Link-Dev/UAV3DCrop_depth.imagedepth-estimation10K<n<100K0 likes1.9k downloads2mo agoHugging Face15PiotrSty /ocr-pl-lines ocr-pl-lines Syntetyczny zbiór linii tekstu po polsku do fine-tuningu OCR (TrOCR). Pary NNNNN.png (obraz linii) + NNNNN.txt (transkrypcja). Struktura train/ — 2000 par (seed 42) val/ — 200 par (seed 123) Generowanie OCR_engine — python -m training.generate_synthetic Korpus: zdania potoczne i urzędowe, domeny (faktury, umowy, medyczne, prawnicze), losowe daty/kwoty/adresy/NIP/PESEL, zdania z pl.wikipedia.org. Augmentacje: pochylenie, blur, szum… See the full description on the dataset page: https://huggingface.co/datasets/PiotrSty/ocr-pl-lines.image1K<n<10K0 likes1.8k downloads12d agoHugging Face16Delores-Lin /MDPBench MDPBench: A Benchmark for Multilingual Document Parsing in Real-World Scenarios We introduce Multilingual Document Parsing Benchmark, the first benchmark for multilingual digital and photographed document parsing. Document parsing has made remarkable strides, yet almost exclusively on clean, digital, well-formatted pages in a handful of dominant languages. No systematic benchmark exists to evaluate how models perform on digital and photographed documents across diverse scripts and… See the full description on the dataset page: https://huggingface.co/datasets/Delores-Lin/MDPBench.textimage-to-text25 likes1.6k downloads2mo agoHugging Face17NJU-LINK /WebCompass WebCompass A unified multimodal benchmark for evaluating LLMs' ability to generate, edit, and repair functional web pages. WebCompass spans three input modalities — text design documents, reference screenshots, and video demonstrations — and three task families — generation, editing, and repair. GitHub: NJU-LINK/WebCompass Project Page: nju-link.github.io/WebCompass Quick Start from datasets import load_dataset # Generation tasks (existing) ds_text =… See the full description on the dataset page: https://huggingface.co/datasets/NJU-LINK/WebCompass.imagetext-generationn<1K6 likes1.6k downloads4mo agoHugging Face18lingada /3DHarnessBench Probing Agentic 3D-to-Code Capabilities of Frontier Vision-Language Models Project Page · GitHub · Arxiv Abstract 3DHarnessBench evaluates the agentic capacity of frontier vision-language models (VLMs) to recover 3D geometry as executable Blender Python code from multiple forms of target evidence. Rather than restricting every system to a single fixed input, the benchmark compares four progressively richer harnesses: Single-view, Multi-view, Active… See the full description on the dataset page: https://huggingface.co/datasets/lingada/3DHarnessBench.image1K<n<10K2 likes1.6k downloads8h agoHugging Face19linxy97 /genhome3d-1280 GenHome3D-1280 1,280 validated household and spatial-design assets in USDZ format, organized across 64 categories. Explore the visual catalog · Browse the GitHub repository · Download the versioned release · Read the generation method Dataset summary Assets 1,280 Categories 64 Assets per category 20 Runtime format USDZ Units Meters Asset license CC BY 4.0 Technical validation 1,280/1,280 pass Package validation 1… See the full description on the dataset page: https://huggingface.co/datasets/linxy97/genhome3d-1280.3d1K<n<10K1 likes1.4k downloads2mo agoHugging Face20linjieli222 /tifa_train_v3image10K<n<100K0 likes1.4k downloads8mo agoHugging Face21harsha-desaraju /telugu-synthetic-line-imagesimage1M<n<10M0 likes1.1k downloads2mo agoHugging Face22Teklia /PELLET-Casimir-Marius-line PELLET Casimir Marius - Line level Dataset Summary The PELLET Casimir Marius dataset includes 100 annotated French letters written between 1914 and 1918. Annotations were done at line-level and all images do not have any text. Note that all images are resized to a fixed height of 128 pixels. Languages All the documents in the dataset are written in French. Dataset Structure Data Instances { 'image': <PIL.JpegImagePlugin.JpegImageFile… See the full description on the dataset page: https://huggingface.co/datasets/Teklia/PELLET-Casimir-Marius-line.imageimage-to-text1 likes1k downloads2y agoHugging Face23linroger023 /wealth-of-nationsimage0 likes995 downloads2y agoHugging Face24linjieli222 /ai2thor-sideview-onlyimage10K<n<100K0 likes952 downloads7mo agoHugging Face25linxy /LaTeX_OCR LaTeX OCR 的数据仓库 本数据仓库是专为 LaTeX_OCR 及 LaTeX_OCR_PRO 制作的数据,来源于 https://zenodo.org/record/56198#.V2p0KTXT6eA 以及 https://www.isical.ac.in/~crohme/ 以及我们自己构建。 如果这个数据仓库有帮助到你的话,请点亮 ❤️like ++ 后续追加新的数据也会放在这个仓库 ~~ 原始数据仓库在github LinXueyuanStdio/Data-for-LaTeX_OCR. 数据集 本仓库有 5 个数据集 small 是小数据集,样本数 110 条,用于测试 full 是印刷体约 100k 的完整数据集。实际上样本数略小于 100k,因为用 LaTeX 的抽象语法树剔除了很多不能渲染的 LaTeX。 synthetic_handwrite 是手写体 100k 的完整数据集,基于 full 的公式,使用手写字体合成而来,可以视为人类在纸上的手写体。样本数实际上略小于 100k,理由同上。… See the full description on the dataset page: https://huggingface.co/datasets/linxy/LaTeX_OCR.imageimage-to-text100K<n<1M181 likes904 downloads2y agoHugging Face26fujinchu /lingyuaudion<1K0 likes847 downloads2d agoHugging Face27Teklia /CASIA-HWDB2-line CASIA-HWDB2 - line level Dataset Summary The offline Chinese handwriting database (CASIA-HWDB2) was built by the National Laboratory of Pattern Recognition (NLPR), Institute of Automation of Chinese Academy of Sciences (CASIA). The handwritten samples were produced by 1,020 writers using Anoto pen on papers, such that both online and offline data were obtained. Note that all images are resized to a fixed height of 128 pixels. Languages All the documents in the… See the full description on the dataset page: https://huggingface.co/datasets/Teklia/CASIA-HWDB2-line.imageimage-to-text10K<n<100K20 likes811 downloads3y agoHugging Face28Linus-L /mnist-cleaned-full Dataset Card for 2025.11.21.16.40.44.970939 This is a FiftyOne dataset with 69807 samples. Installation If you haven't already, install FiftyOne: pip install -U fiftyone Usage import fiftyone as fo from fiftyone.utils.huggingface import load_from_hub # Load the dataset # Note: other available arguments include 'max_samples', etc dataset = load_from_hub("Linus-L/mnist-cleaned-full") # Launch the App session = fo.launch_app(dataset) Dataset… See the full description on the dataset page: https://huggingface.co/datasets/Linus-L/mnist-cleaned-full.imageimage-classification10K<n<100K0 likes789 downloads10mo agoHugging Face29robbyant /lingbot-map-demoimage1K<n<10K6 likes723 downloads2mo agoHugging Face30Linpeng502502 /YCB_Video_Datasetimage1 likes694 downloads8mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.