CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01leibnitz-lab /military_vehicles Citation If you use this dataset, please cite the following paper: @article{kricheli2024error, title={Error Detection and Constraint Recovery in Hierarchical Multi-Label Classification without Prior Knowledge}, author={Kricheli, Joshua Shay and Vo, Khoa and Datta, Aniruddha and Ozgur, Spencer and Shakarian, Paulo}, journal={arXiv preprint arXiv:2407.15192}, year={2024} } imageimage-classification10K<n<100K2 likes6k downloads9mo agoHugging Face02Chandler-Shen /Leiniao_Datasetimagen<1K0 likes1.8k downloads7d agoHugging Face03leibnitz-lab /mdsaimagen<1K0 likes1.6k downloads1y agoHugging Face04DenisaBumba /rfdetr-segmentation-leibniz-dataset Dataset Card for Leibniz's Manuscripts (Instance Segmentation Dataset) This dataset comprises instance segmentation annotations in raw COCO format, used to train an RF-DETR-Seg-nano model for the automatic recognition of textual, graphical, and mathematical expression zones within the manuscripts of the philosopher and mathematician Gottfried Wilhelm Leibniz (17th-early 18th c.). Dataset Details Uses Direct Use This dataset is… See the full description on the dataset page: https://huggingface.co/datasets/DenisaBumba/rfdetr-segmentation-leibniz-dataset.imageimage-segmentation1K<n<10K0 likes963 downloads2mo agoHugging Face05leigangqu /TIGeR-Bench Paper: https://arxiv.org/abs/2406.05814 image10K<n<100K0 likes583 downloads2y agoHugging Face06matthew-l-leidos /ParseBench ParseBench Quick links: [🌐 Website] [📜 Paper] [💻 Code] ParseBench is a benchmark for evaluating document parsing systems on real-world enterprise documents, with the following characteristics: Multi-dimensional evaluation. The benchmark is stratified into five capability dimensions — tables, charts, content faithfulness, semantic formatting, and visual grounding — each with task-specific metrics designed to capture what agentic workflows depend on. Real-world enterprise… See the full description on the dataset page: https://huggingface.co/datasets/matthew-l-leidos/ParseBench.document100K<n<1M0 likes567 downloads3mo agoHugging Face07DenisaBumba /htr_leibniz_dataset_v1 Dataset Card for Leibniz's Manuscripts (HTR Ground Truth) This dataset is composed of trascribed and manually corrected folios of Gottfried Wilhelm Leibniz's manuscripts, together with automatically aligned ground truth in order to train or fine-tune Handwritten Text Recognition (HTR) models. Dataset Details This ground truth was produced to fine-tune existing HTR models for the recognition of Leibniz's handwriting, with the aim of assisting scholars in the… See the full description on the dataset page: https://huggingface.co/datasets/DenisaBumba/htr_leibniz_dataset_v1.imageimage-to-text10K<n<100K0 likes381 downloads23d agoHugging Face08leigangqu /MSE-Bench MSE-Bench: A Benchmark for Multi-turn Session Image Editing Introduction MSE-Bench (Multi-turn Session image Editing Benchmark) is a benchmark designed to evaluate multi-turn image editing systems under realistic editing workflows. Given a source image and a series of editing instructions, the goal is for a model to apply these edits cumulatively to produce a final image that reflects all the requested changes. MSE-Bench consists of 100 test instances, each representing… See the full description on the dataset page: https://huggingface.co/datasets/leigangqu/MSE-Bench.imagen<1K2 likes306 downloads6mo agoHugging Face09leitro /Copiale_Lines Copiale Lines Copiale Lines is a line-level image-to-text dataset for historical cipher decipherment. It contains cropped line images from the Copiale manuscript paired with plaintext ground truth. This dataset was presented in the paper Learning to Decipher from Pixels -- A Case Study of Copiale (HistoCrypt 2026). Code: https://github.com/leitro/Decipher-from-Pixels-Copiale Dataset Structure The dataset is split into: train: 1,269 samples valid: 175 samples… See the full description on the dataset page: https://huggingface.co/datasets/leitro/Copiale_Lines.imageimage-to-text1K<n<10K0 likes230 downloads5mo agoHugging Face10leigangqu /MSE-Bench-resultsimage1K<n<10K3 likes172 downloads6mo agoHugging Face11Leiyao-Cui /XieNet Dataset Card for XieNet This is the repaired version of GAPartNet dataset, which we use as the simulation dataset for Vi-TacMan. Description We identified numerous object meshes in the original dataset that lack proper cap geometry, so we manually repaired these meshes to ensure completeness. The following images (object id: 47296) exemplify the type of geometric defects found and our corrections: GAPartNet (Original)… See the full description on the dataset page: https://huggingface.co/datasets/Leiyao-Cui/XieNet.imagen<1K2 likes138 downloads4mo agoHugging Face12LeightonFushiguro /PPE_v10image1K<n<10K3 likes85 downloads2y agoHugging Face13LeightonFushiguro /validppeimagen<1K0 likes67 downloads2y agoHugging Face14leixiang25 /24679-hw1-image-register 24-679 HW1 (Fall 2026): Business Message Register Images leixiang25/24679-hw1-image-register Digitally rendered screenshot-style images of short, fictional business messages, labeled by register. 1 = formal (high-context business register); 0 = casual (low-context register). Created by Lei Xiang for 24-679 Homework 1 at Carnegie Mellon University. The dataset is related to Context, a cross-cultural deal interpreter for Western operators working with Japanese and Chinese… See the full description on the dataset page: https://huggingface.co/datasets/leixiang25/24679-hw1-image-register.imageimage-classificationn<1K0 likes59 downloads9d agoHugging Face15davanstrien /leicester_loaded_annotations_binary Dataset Card for "leicester_loaded_annotations_binary" More Information needed imagen<1K0 likes33 downloads4y agoHugging Face16leinms /flickr30k-qwen3vl-baseline Flickr30k Qwen3-VL Baseline Captions (Test Split) This dataset is based on the Mozilla/flickr30k-transformed-captions-gpt4o test split and contains 1,000 images from the original Flickr30k dataset.It includes both the original metadata and newly generated baseline captions produced using the Qwen3-VL-2B-Instruct vision-language model. Contents Each entry includes: image — the original Flickr30k image alt_text — GPT-4o transformed caption from Mozilla's de-biasing… See the full description on the dataset page: https://huggingface.co/datasets/leinms/flickr30k-qwen3vl-baseline.image1K<n<10K0 likes27 downloads10mo agoHugging Face17davanstrien /leicester_loaded_annotations Dataset Card for "leicester_loaded_annotations" More Information needed imagen<1K0 likes26 downloads4y agoHugging Face18leibnitz-lab /ImageNet50 ImageNet50 Dataset This repository contains the dataset for the paper "Error Detection and Constraint Recovery in Hierarchical Multi-Label Classification without Prior Knowledge". Citation If you use this dataset in your research, please cite the following paper: @article{kricheli2024error, title={Error Detection and Constraint Recovery in Hierarchical Multi-Label Classification without Prior Knowledge}, author={Kricheli, Joshua Shay and Vo, Khoa and Datta, Aniruddha… See the full description on the dataset page: https://huggingface.co/datasets/leibnitz-lab/ImageNet50.image1K<n<10K0 likes24 downloads9mo agoHugging Face19leinms /flickr30k-qwen3vl-baseline-improved_prompt Flickr30k Qwen3-VL Few-Shot Styled Captions (Test Split) This dataset is derived from the Mozilla/flickr30k-transformed-captions-gpt4o test subset and contains 1,000 images.It preserves the original metadata and adds new few-shot–guided baseline captions generated with the Qwen3-VL-2B-Instruct model. The goal of this dataset is to provide a consistent, controlled captioning style enforced through three-shot visual prompting. 📌 Dataset Contents Each sample contains:… See the full description on the dataset page: https://huggingface.co/datasets/leinms/flickr30k-qwen3vl-baseline-improved_prompt.image1K<n<10K0 likes22 downloads10mo agoHugging Face20leigham /rlimagen<1K0 likes12 downloads3y agoHugging Face21leinms /flickr30k-qwen3vl-baseline-evaluation-by-gpt-4o-2024-08-06image1K<n<10K0 likes12 downloads10mo agoHugging Face22leinms /flickr30k-qwen3-vl-2b-sft-trl-with-judgeimage1K<n<10K0 likes12 downloads9mo agoHugging Face23leisong /knot_dvrkimage1K<n<10K0 likes11 downloads9mo agoHugging Face24vannynakamura /leishimage1K<n<10K0 likes10 downloads5y agoHugging Face25leizhao7 /finevision-mcq-v4image10K<n<100K0 likes10 downloads5mo agoHugging Face26leiccp /imgbedimagen<1K0 likes6 downloads8mo agoHugging Face27leizhao7 /llava-next-mcq-v3image100K<n<1M0 likes6 downloads5mo agoHugging Face28leinms /zapovednik_combined_v2image1K<n<10K0 likes4 downloads9mo agoHugging Face29leizhao7 /llava-next-mcq-v2-25kimage10K<n<100K0 likes4 downloads5mo agoHugging Face30leilanihoffmann /fashion_image_caption-100-v2 Dataset Card for "fashion_image_caption-100-v2" More Information needed imagen<1K0 likes3 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.