CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01Lance1573 /L-Mind L-Mind: A Multimodal Dataset for Neural-Driven Image Editing This dataset is part of the NeurIPS 2025 paper: "Neural-Driven Image Editing", which introduces LoongX, a hands-free image editing approach driven by multimodal neurophysiological signals. 📄 Overview L-Mind is a large-scale multimodal dataset designed to bridge Brain-Computer Interfaces (BCIs) with generative AI. It enables research into accessible, intuitive image editing for individuals with limited motor… See the full description on the dataset page: https://huggingface.co/datasets/Lance1573/L-Mind.imageimage-to-image10K<n<100K2 likes918 downloads8mo agoHugging Face02NationalLibraryOfScotland /encyclopaedia-britannica-lance Encyclopaedia Britannica (1771-1860) - Lance Format This dataset contains 155,388 digitized pages from the Encyclopaedia Britannica, spanning editions from 1771 to 1860. The data is stored in Lance format for efficient streaming and lazy image loading. Dataset Details Total Pages: 155,388 Total Volumes: 195 Format: Lance (columnar format with blob storage for images) Source: National Library of Scotland (NLS) License: Public Domain (CC0) Loading the Dataset… See the full description on the dataset page: https://huggingface.co/datasets/NationalLibraryOfScotland/encyclopaedia-britannica-lance.imageimage-to-text100K<n<1M2 likes815 downloads8mo agoHugging Face03davanstrien /encyclopaedia-britannica-lance-test Encyclopaedia Britannica (1771-1860) - Lance Format This dataset contains 155,388 digitized pages from the Encyclopaedia Britannica, spanning editions from 1771 to 1860. The data is stored in Lance format for efficient streaming and lazy image loading. Dataset Details Total Pages: 155,388 Total Volumes: 195 Format: Lance (columnar format with blob storage for images) Source: National Library of Scotland (NLS) License: Public Domain (CC0) Loading the Dataset… See the full description on the dataset page: https://huggingface.co/datasets/davanstrien/encyclopaedia-britannica-lance-test.imageimage-to-text100K<n<1M0 likes724 downloads8mo agoHugging Face04lance-format /droidimage10M<n<100M0 likes683 downloads4mo agoHugging Face05LanceBunag /BalitaNLPA Filipino multi-modal language dataset for text+visual tasks. Consists of 351,755 Filipino news articles (w/ associated images) gathered from Filipino news outlets. Description Total # of articles: 351,755 80-10-10 split for training, validation, and testing. Dataset field descriptions: title - Article title body - Article body. Separated into paragraphs image - Article image website… See the full description on the dataset page: https://huggingface.co/datasets/LanceBunag/BalitaNLP.imagetext-to-image100K<n<1M5 likes602 downloads9mo agoHugging Face06lance-format /BDD100K-enrichedimage10K<n<100K0 likes555 downloads6mo agoHugging Face07lance-format /docvqa-lance DocVQA (Lance Format) A Lance-formatted version of DocVQA, a benchmark for visual question answering over document images such as industry and government scans, multi-page reports, forms, and receipts, redistributed via lmms-lab/DocVQA (DocVQA config). Each row carries the page image as inline JPEG bytes, the question and reference answer span(s), the original DocVQA question-type tags, UCSF Industry Documents Library provenance, and paired CLIP embeddings for the image and the… See the full description on the dataset page: https://huggingface.co/datasets/lance-format/docvqa-lance.imagevisual-question-answering10K<n<100K0 likes305 downloads4mo agoHugging Face08lance-format /textvqa-lance TextVQA (Lance Format) A Lance-formatted version of TextVQA — visual question answering where the question requires reading text in the image (street signs, product labels, screen captures) — sourced from lmms-lab/textvqa. Each row carries the image bytes, the question, the 10 reference annotator answers, the OCR tokens detected by the source pre-processing, OpenImages-style scene tags, and paired CLIP image and question embeddings — all available directly from the Hub at… See the full description on the dataset page: https://huggingface.co/datasets/lance-format/textvqa-lance.imagevisual-question-answering10K<n<100K0 likes273 downloads4mo agoHugging Face09lance-format /handwriting-ocr Handwriting OCR (Lance Format) This Lance-formatted version of the Doctor's Handwritten Prescription BD dataset contains 4,680 cropped PNG images of handwritten medicine names from Bangladesh. Each row keeps the original image bytes with the medicine and generic-name labels, plus deterministic search metadata derived from those labels. The dataset contains three source-preserved splits: train, validation, and test. [!NOTE] Training note: The same samples appear repeatedly… See the full description on the dataset page: https://huggingface.co/datasets/lance-format/handwriting-ocr.imageimage-to-text1K<n<10K0 likes259 downloads3mo agoHugging Face10Lancelot53 /spot-the-diffimage10K<n<100K1 likes252 downloads1y agoHugging Face11lance-format /coco-captions-2017-lance COCO Captions 2017 (Lance Format) A Lance-formatted version of the COCO Captions 2017 corpus, redistributed via lmms-lab/COCO-Caption2017. Each row is one image with 5–7 human-written captions, a cosine-normalized CLIP image embedding, and a cosine-normalized CLIP text embedding of the canonical caption — all stored inline and available directly from the Hub at hf://datasets/lance-format/coco-captions-2017-lance/data. Key features Inline JPEG bytes in the image… See the full description on the dataset page: https://huggingface.co/datasets/lance-format/coco-captions-2017-lance.imageimage-to-text10K<n<100K0 likes247 downloads4mo agoHugging Face12lance-format /coco-detection-2017-lance COCO 2017 Object Detection (Lance Format) A Lance-formatted version of the COCO 2017 object detection benchmark, sourced from detection-datasets/coco. Each row is one image with its inline JPEG bytes, the full per-image list of bounding boxes, COCO 80-class category ids and names, per-object areas, an OpenCLIP image embedding, and pre-built indices — all available directly from the Hub at hf://datasets/lance-format/coco-detection-2017-lance/data. Key features Inline… See the full description on the dataset page: https://huggingface.co/datasets/lance-format/coco-detection-2017-lance.imageobject-detection100K<n<1M0 likes238 downloads4mo agoHugging Face13davanstrien /encyclopaedia-britannica-lance-test2 Encyclopaedia Britannica (1771-1860) - Lance Format This dataset contains 155,388 digitized pages from the Encyclopaedia Britannica, spanning editions from 1771 to 1860. The data is stored in Lance format for efficient streaming and lazy image loading. Dataset Details Total Pages: 155,388 Total Volumes: 195 Format: Lance (columnar format with blob storage for images) Source: National Library of Scotland (NLS) License: Public Domain (CC0) Loading the Dataset… See the full description on the dataset page: https://huggingface.co/datasets/davanstrien/encyclopaedia-britannica-lance-test2.imageimage-to-text100K<n<1M0 likes233 downloads8mo agoHugging Face14lance-format /vqav2-lance VQAv2 (Lance Format) A Lance-formatted version of VQAv2 — open-ended visual question answering on COCO images — sourced from lmms-lab/VQAv2. Each row is one (image, question, 10 annotator answers) triple with paired CLIP image and question embeddings drawn from the same shared space, plus the VQAv2 question_type / answer_type taxonomy and the consensus multiple_choice_answer — all available directly from the Hub at hf://datasets/lance-format/vqav2-lance/data. Key… See the full description on the dataset page: https://huggingface.co/datasets/lance-format/vqav2-lance.imagevisual-question-answering100K<n<1M0 likes227 downloads4mo agoHugging Face15davanstrien /bpl-card-catalog-lance-fullimage100K<n<1M0 likes214 downloads7mo agoHugging Face16lance-format /laion-1m LAION-Subset (Lance Format) A Lance-formatted slice of the LAION image-text corpus (~1M rows) with inline JPEG bytes, CLIP image embeddings (img_emb), full metadata, and a pre-built ANN index — all available directly from the Hub at hf://datasets/lance-format/laion-1m/data/train.lance. Key features Inline JPEG bytes in the image column — no sidecar files, no image folders. Pre-computed CLIP image embeddings (img_emb, 768-dim) with a bundled IVF_PQ index for… See the full description on the dataset page: https://huggingface.co/datasets/lance-format/laion-1m.imagetext-to-image1M<n<10M4 likes208 downloads4mo agoHugging Face17lance-format /textvqa-lance-colab TextVQA VLM Fine-Tuning Demo A small Lance-formatted subset of TextVQA, source via pre-baked operations from this repo and fine-tuning a VLM on the subset. Each row is one visual question-answering example over an image that contains scene text: inline image bytes, a natural-language question, 10 reference answers, OCR tokens, image-class labels, and paired 512-dimensional image/question embeddings are stored together in a Lance table. The train split also includes precomputed… See the full description on the dataset page: https://huggingface.co/datasets/lance-format/textvqa-lance-colab.imagevisual-question-answering1K<n<10K0 likes205 downloads3mo agoHugging Face18lance-format /mnist-lance MNIST (Lance Format) A Lance-formatted version of the classic MNIST handwritten-digit dataset covering 70,000 28×28 grayscale digits across ten balanced classes. Each row carries inline PNG bytes, the digit label, the human-readable class name, and a cosine-normalized CLIP image embedding, all backed by a bundled IVF_PQ vector index plus scalar indices on the label columns and available directly from the Hub at hf://datasets/lance-format/mnist-lance/data. Key features… See the full description on the dataset page: https://huggingface.co/datasets/lance-format/mnist-lance.imageimage-classification10K<n<100K0 likes197 downloads4mo agoHugging Face19LancetRobotics /gelsight-mini-gelsight-20260617 Dual GelSight Tactile Dataset 20260617 This dataset contains synchronized marker-mask tactile captures from two GelSight-style sensors: gelsight_mini: GelSight Mini camera stream gelsight: custom UVC GelSight-style camera stream The data was collected on 2026-06-17 for real-world finetuning/adaptation of UniForce-style tactile models. Directory Layout marker/<sensor>/<indenter>/<frame_id>.jpg collection_log.txt The marker folders contain marker mask images… See the full description on the dataset page: https://huggingface.co/datasets/LancetRobotics/gelsight-mini-gelsight-20260617.imageimage-classification10K<n<100K0 likes193 downloads3mo agoHugging Face20lance-format /chartqa-lance ChartQA (Lance Format) A Lance-formatted version of ChartQA, a benchmark for question answering over scientific and business charts that demands a mix of logical and visual reasoning, redistributed via lmms-lab/ChartQA. Each row carries the chart image as inline JPEG bytes, the natural-language question and reference answer(s), a question-type tag (human vs augmented), and paired CLIP embeddings for the image and the question — all available directly from the Hub at… See the full description on the dataset page: https://huggingface.co/datasets/lance-format/chartqa-lance.imagevisual-question-answering1K<n<10K0 likes174 downloads4mo agoHugging Face21lance-format /fashion-mnist-lance Fashion-MNIST (Lance Format) A Lance-formatted version of Fashion-MNIST covering 70,000 28×28 grayscale clothing images across ten balanced apparel classes. Each row carries inline PNG bytes, the integer label, the human-readable class name, and a cosine-normalized CLIP image embedding, all backed by a bundled IVF_PQ vector index plus scalar indices on the label columns and available directly from the Hub at hf://datasets/lance-format/fashion-mnist-lance/data. Key… See the full description on the dataset page: https://huggingface.co/datasets/lance-format/fashion-mnist-lance.imageimage-classification10K<n<100K0 likes166 downloads4mo agoHugging Face22lance-format /kitti-2d-detection-lance KITTI 2D Object Detection (Lance Format) A Lance-formatted version of the KITTI 2D Object Detection benchmark, sourced from nateraw/kitti so no manual signup or download from cvlibs.net is required. Each row is a single driving frame with inline JPEG bytes, the full set of 2D and 3D object annotations stored as parallel per-object lists, plus a cosine-normalized OpenCLIP ViT-B-32 image embedding — all available directly from the Hub at… See the full description on the dataset page: https://huggingface.co/datasets/lance-format/kitti-2d-detection-lance.imageobject-detection1K<n<10K1 likes147 downloads4mo agoHugging Face23lance-format /pascal-voc-2012-segmentation-lance Pascal VOC 2012 Segmentation (Lance Format) A Lance-formatted version of the Pascal VOC 2012 semantic segmentation split, sourced from nateraw/pascal-voc-2012. Each row pairs an inline JPEG image with the per-pixel PNG segmentation mask and a cosine-normalized OpenCLIP ViT-B-32 image embedding, so a single columnar table carries both annotation modalities and the features needed to retrieve, curate, and train against them — all available directly from the Hub at… See the full description on the dataset page: https://huggingface.co/datasets/lance-format/pascal-voc-2012-segmentation-lance.imageimage-segmentation1K<n<10K0 likes144 downloads4mo agoHugging Face24lance-format /oxford-pets-lance Oxford-IIIT Pet (Lance Format) A Lance-formatted version of the Oxford-IIIT Pet dataset — 7,390 cat and dog photos across 37 breeds — sourced from pcuenq/oxford-pets. Each row carries the inline JPEG bytes, the breed name, a species flag distinguishing cats from dogs, and a cosine-normalized CLIP image embedding, all available directly from the Hub at hf://datasets/lance-format/oxford-pets-lance/data. Key features Inline JPEG bytes in the image column — no sidecar… See the full description on the dataset page: https://huggingface.co/datasets/lance-format/oxford-pets-lance.imageimage-classification1K<n<10K0 likes140 downloads4mo agoHugging Face25lance-format /eurosat-lance EuroSAT (Lance Format) A Lance-formatted version of EuroSAT, the canonical Sentinel-2 RGB land-cover benchmark, sourced from blanchon/EuroSAT_RGB. Each row is a single 64×64 RGB tile with its integer class id, the human-readable class name, and a cosine-normalized OpenCLIP image embedding — all stored inline and available directly from the Hub at hf://datasets/lance-format/eurosat-lance/data. Key features Inline JPEG bytes in the image column — no sidecar TIF folders… See the full description on the dataset page: https://huggingface.co/datasets/lance-format/eurosat-lance.imageimage-classification10K<n<100K0 likes123 downloads4mo agoHugging Face26LancetRobotics /dual-gelsight-marker-20260617 Dual GelSight Marker-Mask Dataset 20260617 This dataset contains synchronized marker-mask tactile captures from two GelSight-style sensors: gelsight_mini: GelSight Mini camera stream gelsight: custom UVC GelSight-style camera stream Only marker-mask images are included. Raw RGB camera frames were intentionally omitted from this Hub upload. Directory Layout marker/<sensor>/<indenter>/<frame_id>.jpg Sensors marker/gelsight_mini: 3280x2464… See the full description on the dataset page: https://huggingface.co/datasets/LancetRobotics/dual-gelsight-marker-20260617.imageimage-classification10K<n<100K0 likes113 downloads3mo agoHugging Face27lance-format /flickr30k-lance Flickr30k (Lance Format) A Lance-formatted version of Flickr30k, redistributed via lmms-lab/flickr30k. Each row is one image with 5 human-written captions, a cosine-normalized CLIP image embedding, and a cosine-normalized CLIP text embedding of the canonical caption — all stored inline and available directly from the Hub at hf://datasets/lance-format/flickr30k-lance/data. Key features Inline JPEG bytes in the image column — no sidecar files, no image folders. Paired… See the full description on the dataset page: https://huggingface.co/datasets/lance-format/flickr30k-lance.imageimage-to-text10K<n<100K0 likes98 downloads4mo agoHugging Face28lance-format /food101-lance Food-101 (Lance Format) A Lance-formatted version of Food-101, the fine-grained dish-classification benchmark of 101,000 photos spread evenly across 101 dish classes, sourced from ethz/food101. Each row carries the inline JPEG bytes, the integer label, the human-readable label_name, and a cosine-normalized CLIP image embedding, all available directly from the Hub at hf://datasets/lance-format/food101-lance/data. Key features Inline JPEG bytes in the image column — no… See the full description on the dataset page: https://huggingface.co/datasets/lance-format/food101-lance.imageimage-classification100K<n<1M0 likes96 downloads4mo agoHugging Face29davanstrien /bpl-card-catalog-lanceimagen<1K0 likes93 downloads7mo agoHugging Face30Lancelot53 /clevr-changeimage1K<n<10K2 likes80 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.