CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01FreedomIntelligence /BlenderLore The target scope is 22,360 video-associated Blender project instances, not 22,360 distinct tutorial videos. Uploads are in progress, so the currently published files may be a subset of this target. The 44 biomedical project instances and one software-bundled Dome template are excluded. Data Structure The dataset is organized as a collection of sample-level directories under assets/. Each directory corresponds to one Blender creation task and follows the structure below:… See the full description on the dataset page: https://huggingface.co/datasets/FreedomIntelligence/BlenderLore.image10K<n<100K6 likes10k downloads1d agoHugging Face02gnucleus-ai /cad-gen-freecad CAD Generation Dataset Each row in this dataset describes one parametric CAD part. Columns: id — row identifier (also the basename of the per-row asset files) name — part family (e.g. flanges, spur_gear_stock) description — natural-language description of the geometry key_parameters — the dimensions that drive the parametric model image — 512×512 PNG preview rendered from the FCStd fcstd_path — relative path inside this repo to the parametric FreeCAD document (fcstd/<id>.FCStd)… See the full description on the dataset page: https://huggingface.co/datasets/gnucleus-ai/cad-gen-freecad.3dn<1K5 likes3k downloads4mo agoHugging Face03FreedomIntelligence /PubMedVision News [2025/02/18]: We add the original captions of PubMedVision in PubMedVision_Original_Caption.json, as well as the Chinese version of PubMedVision in PubMedVision_Chinese.json. [2024/07/01]: We add annotations for 'body_part' and 'modality' of images, utilizing the HuatuoGPT-Vision-7B model. PubMedVision PubMedVision is a large-scale medical VQA dataset. We extracted high-quality image-text pairs from PubMed and used GPT-4V to reformat them to enhance their quality.… See the full description on the dataset page: https://huggingface.co/datasets/FreedomIntelligence/PubMedVision.imagequestion-answering1M<n<10M107 likes1.7k downloads2y agoHugging Face04FreeLand123 /SciFigPlag-Bench SciFigPlag-Bench: A Benchmark for Provenance-Aware Scientific Figure Plagiarism Detection 🌐 Homepage | 📖 arXiv SciFigPlag-Bench is a benchmark for provenance-aware scientific figure plagiarism detection. It evaluates whether a suspicious figure reuses evidence from a specific source figure, which figure is the original source, how the reused content has been transformed, and where the reused evidence appears. The benchmark is designed to evaluate vision-language models… See the full description on the dataset page: https://huggingface.co/datasets/FreeLand123/SciFigPlag-Bench.image10K<n<100K2 likes1.2k downloads27d agoHugging Face05FreedomIntelligence /ALLaVA-4V 📚 ALLaVA-4V Data Generation Pipeline LAION We leverage the superb GPT-4V to generate captions and complex reasoning QA pairs. Prompt is here. Vison-FLAN We leverage the superb GPT-4V to generate captions and detailed answer for the original instructions. Prompt is here. Wizard We regenerate the answer of Wizard_evol_instruct with GPT-4-Turbo. Dataset Cards All datasets can be found here. The structure of naming is shown below: ALLaVA-4V ├──… See the full description on the dataset page: https://huggingface.co/datasets/FreedomIntelligence/ALLaVA-4V.imagequestion-answering100K<n<1M98 likes861 downloads1y agoHugging Face06Blue2Giant /FreeStyle_Dataset 0426 CRef/SRef LoRA Triplet Dataset This dataset contains CRef/SRef LoRA triplets exported from the 0426 diffusion training data. Each training example has three images: content: content reference image, used as cref_0 style: style reference image, used as sref_0 target: image generated from the combined content + style condition Use triplets.csv as the main entry point. Image-level CSV files are provided only for deduplicated metadata and provenance lookup.… See the full description on the dataset page: https://huggingface.co/datasets/Blue2Giant/FreeStyle_Dataset.imageimage-to-image100K<n<1M2 likes851 downloads3mo agoHugging Face07frankielp /freevocab-scannet-scannetpp-outputimagen<1K0 likes837 downloads2mo agoHugging Face08BangumiBase /freeeternalsummer Bangumi Image Base of Free! -eternal Summer- This is the image base of bangumi Free! -Eternal Summer-, we detected 24 characters, 2471 images in total. The full dataset is here. Please note that these image bases are not guaranteed to be 100% cleaned, they may be noisy actual. If you intend to manually train models using this dataset, we recommend performing necessary preprocessing on the downloaded dataset to eliminate potential noisy samples (approximately 1% probability).… See the full description on the dataset page: https://huggingface.co/datasets/BangumiBase/freeeternalsummer.image1K<n<10K0 likes777 downloads3y agoHugging Face09gt-free-ocr-metrics /omnidocbench-render-compare OmniDocBench Render-and-Compare This dataset contains the rendered HTML reconstructions and comparison images produced by a render-and-compare pipeline — a reference-free visual similarity evaluation framework for OCR systems. Overview The pipeline processes each page of OmniDocBench through a Qwen3.5-122B-A10B OCR model, renders the structured output back to a PNG via HTML (reconstructed.png), and compares it against the original page scan (masked_original.png) using… See the full description on the dataset page: https://huggingface.co/datasets/gt-free-ocr-metrics/omnidocbench-render-compare.imageother10K<n<100K0 likes514 downloads5mo agoHugging Face10daihang /FreeArt-21 Paper FreeArtGS: Articulated Gaussian Splatting Under Free-moving Scenariohttps://huggingface.co/papers/2603.22102 license: apache-2.0 image10K<n<100K0 likes509 downloads3mo agoHugging Face11scintigimcki /FreeGave-GoPro FreeGave-GoPro Dataset This dataset is proposed by FreeGave. Structure The structure of the dataset is as: DynObjects | - data | | - pen1: | | | - train: serves as training data | | | - val: used for evaluating novel view interpolation | | | - test: used for evaluating future extrapolation | | | - transforms_train.json: camera poses and other meta informations for training set | | | - transforms_val.json: camera poses and other meta informations for novel view… See the full description on the dataset page: https://huggingface.co/datasets/scintigimcki/FreeGave-GoPro.image0 likes368 downloads1y agoHugging Face12bluestyle97 /FreeSplatterStatic3dn<1K0 likes364 downloads2y agoHugging Face13FreedomIntelligence /Medical_Multimodal_Evaluation_Data Evaluation Guide This dataset is used to evaluate medical multimodal LLMs, as used in HuatuoGPT-Vision. It includes benchmarks such as VQA-RAD, SLAKE, PathVQA, PMC-VQA, OmniMedVQA, and MMMU-Medical-Tracks. To get started: Download the dataset and extract the images.zip file. Find evaluation code on our GitHub: HuatuoGPT-Vision. This open-source release aims to simplify the evaluation of medical multimodal capabilities in large models. Please cite the relevant benchmark… See the full description on the dataset page: https://huggingface.co/datasets/FreedomIntelligence/Medical_Multimodal_Evaluation_Data.imageimage-to-text10K<n<100K29 likes346 downloads2y agoHugging Face14Blue2Giant /FreeStyle_Bench FreeStyle Bench This repository contains benchmark inputs for evaluating FreeStyle reference-based image editing. It includes two benchmark subsets: sref_bench/: a style-transfer benchmark. cref_sref_bench/: a content-and-style dual-reference benchmark. Both subsets use the same basic layout: a content-reference folder, a style-reference folder, and a prompts.json file that links each pair to its text prompt. Repository Layout <repo-root>/ README.md… See the full description on the dataset page: https://huggingface.co/datasets/Blue2Giant/FreeStyle_Bench.imageimage-to-image0 likes298 downloads3mo agoHugging Face15FBK-TeV /FreeGraspData Free-form language-based robotic reasoning and grasping Dataset: FreeGraspData Homepage: https://tev-fbk.github.io/FreeGrasp/ Paper: https://arxiv.org/abs/2503.13082 FreeGrasp Code: https://github.com/tev-fbk/FreeGrasp_code NOTE: the scene ids and objects ids are given from MetaGraspNet v2 The objects instance segmentation maps are available at https://huggingface.co/datasets/FBK-TeV/FreeGraspData/blob/main/data/npz_file.zip Dataset Description Examples of FreeGraspData at… See the full description on the dataset page: https://huggingface.co/datasets/FBK-TeV/FreeGraspData.imagen<1K7 likes297 downloads1y agoHugging Face16MorPhLingXD /FreeArt3D Dataset Card for Dataset Name Preprocessed dataset of PartNet-Mobility objects for FreeArt3D. Dataset Sources Repository:: https://github.com/CzzzzH/FreeArt3D Paper: https://huggingface.co/papers/2510.25765 Demo: https://huggingface.co/spaces/MorPhLingXD/FreeArt3D BibTeX: @InProceedings{chen2025freeart3d, title = {FreeArt3D: Training-Free Articulated Object Generation using 3D Diffusion}, author = {Chen, Chuhao and Liu, Isabella and Wei, Xinyue and Su, Hao and… See the full description on the dataset page: https://huggingface.co/datasets/MorPhLingXD/FreeArt3D.3dn<1K0 likes266 downloads11mo agoHugging Face17Freed-Wu /kodakThe pictures below link to lossless, true color (24 bits per pixel, aka "full color") images. It is my understanding they have been released by the Eastman Kodak Company for unrestricted usage. Many sites use them as a standard test suite for compression testing, etc. Prior to this site, they were only available in the Sun Raster format via ftp. This meant that the images could not be previewed before downloading. Since their release, however, the lossless PNG format has been incorporated into all the major browsers. Since PNG supports 24-bit lossless color (which GIF and JPEG do not), it is now possible to offer this browser-friendly access to the images.imageothern<1K2 likes265 downloads4y agoHugging Face18FreedomIntelligence /TCM-Vision-Benchmark 📚 Introduction This is the text benchmark for ShizhenGPT, a multimodal LLM for Traditional Chinese Medicine (TCM). For details, see our paper and GitHub repository. 📊 Benchmark Overview The benchmark is composed of 7 sections, each compiled from different authoritative TCM illustrated books. Samples TCM Patent 1119 TCM Material 1020 TCM Herb 1100 Tongue768 Palm 640 Holism 1011 Tuina 831 Eye 715 ⚒️ Data Construction {… See the full description on the dataset page: https://huggingface.co/datasets/FreedomIntelligence/TCM-Vision-Benchmark.image3 likes224 downloads10mo agoHugging Face19JiayiWuLeo /FreeCHROMAimage10K<n<100K0 likes205 downloads20h agoHugging Face20xyliu6 /k12-freeformimage10K<n<100K2 likes198 downloads1y agoHugging Face21freexiaochuan /captcha best.pth model = CaptchaCNNTransformer( img_h=32, img_w=128, dim=256, depth=6, heads=4, num_classes=num_classes, channels=1, dropout=0.2, ).to(device) image10K<n<100K0 likes194 downloads18d agoHugging Face22bghira /free-to-use-pixelart Free-to-use Pixel Art Dataset Details This dataset was collected on 25th May, 2024. It's a small subset of the free-to-use images on PixilArt. At the time of publication, this dataset was covered by permissive terms that allow commercial use. Dataset Description This dataset is unique in that it contains the pixel group size for each collected sample, which might assist in experiments on microconditioning inputs on an adapter to control this value of the unit… See the full description on the dataset page: https://huggingface.co/datasets/bghira/free-to-use-pixelart.image1K<n<10K9 likes181 downloads2y agoHugging Face23cyttic /trocr-hebrew-freefonts-BYimage100K<n<1M0 likes175 downloads22d agoHugging Face24hyzhou404 /FreeFiximage1K<n<10K0 likes170 downloads8mo agoHugging Face25rjiao /FreegraspData_V2This includes the source data of UnoGrasp, they are selected from MetaGraspNetV2 from 37 vpt to 4 vpt. It has no practical function and is only used for tracing the source. image0 likes161 downloads20d agoHugging Face26cyttic /heb-freefonts-250kimage100K<n<1M0 likes149 downloads1mo agoHugging Face27Ebichuu /freeimagen<1K0 likes108 downloads21d agoHugging Face28Yas1n /FreeCAD_Sketches_Pics 🧩 FreeCAD Sketch Python Dataset This dataset contains approximately 1,000 Python files and their corresponding images defining parametric FreeCAD sketches. Each python file represents a geometric sketch scripted using the FreeCAD Python API. The dataset is intended for training and fine-tuning large language models (LLMs) and vision large language models (vLLMs) and other AI systems to understand and generate CAD-based parametric geometry code, particularly for text-to-CAD and… See the full description on the dataset page: https://huggingface.co/datasets/Yas1n/FreeCAD_Sketches_Pics.imagetext-to-3d1K<n<10K0 likes106 downloads9mo agoHugging Face29origindatalab /bangladesh-urban-traffic-free-sample-v2 Bangladesh Urban Traffic Dataset (Dhaka Mixed Traffic Dataset) — Free Sample This Bangladesh Traffic Dataset contains real-world mixed traffic scenes including rickshaws, motorcycles, buses, cars, pedestrians, market congestion, intersections, and complex urban mobility patterns. The dataset is designed for computer vision, object detection, autonomous driving, ADAS, robotics, perception systems, physical AI, traffic analytics, and emerging-market mobility research. This… See the full description on the dataset page: https://huggingface.co/datasets/origindatalab/bangladesh-urban-traffic-free-sample-v2.imageobject-detectionn<1K0 likes106 downloads3mo agoHugging Face30cyttic /trocr-hebrew-freefonts9image100K<n<1M0 likes106 downloads1mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.