CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01wentao-yuan /robopoint-data RoboPoint Dataset Card Dataset details This dataset contains 1432K image-QA instances used to fine-tune RoboPoint, a VLM for spatial affordance prediction. It consists of the following parts: 347K object reference instances from a synthetic data pipeline; 320K free space reference instances from a synthetic data pipeline; 100K object detection instaces from LVIS; 150K GPT-generated instruction-following instances from liuhaotian/LLaVA-Instruct-150K; 515K general-purpose… See the full description on the dataset page: https://huggingface.co/datasets/wentao-yuan/robopoint-data.image1M<n<10M13 likes4.2k downloads2y agoHugging Face02Yale-BIDS-Chen /medpmc-11m-dataset_jun24_baseline MedPMC WebDataset MedPMC is a large-scale medical image-text dataset curated from articles in the PubMed Central (PMC) collection. This release contains approximately 11 million image-text pairs collected from the June 2024 PMC baseline. MedPMC is an ongoing effort, and future releases will continue to expand the dataset with newly published literature, improved annotations, and additional resources. This dataset is presented in the paper MedPMC: A Systematic Framework for… See the full description on the dataset page: https://huggingface.co/datasets/Yale-BIDS-Chen/medpmc-11m-dataset_jun24_baseline.imagezero-shot-image-classification1M<n<10M3 likes3.8k downloads2mo agoHugging Face03Yuxuan0701 /tvqa-framesimage1M<n<10M0 likes2.2k downloads7mo agoHugging Face04Yossh /danbooru2023-webp-4Mpixel-224The data set is just resized to 224*224 https://huggingface.co/datasets/KBlueLeaf/danbooru2023-webp-4Mpixel Pseudo code for processing def resize_image(file_path): with Image.open(file_path) as img: resized_img = img.resize((224, 224)) resized_img.save(file_path) image100K<n<1M2 likes1.7k downloads2y agoHugging Face05yangyang857658468 /cc12m-webdataset CC12M WebDataset 这是CC12M数据集的WebDataset格式版本。 数据集信息 文件数量: 1098 总大小: 888796.33 MB 上传时间: 2025-03-18 14:45:49 使用方法 import webdataset as wds dataset = wds.WebDataset("https://huggingface.co/yangyang857658468/cc12m-webdataset/resolve/main/cc12m_*.tar") image10M<n<100M0 likes1.6k downloads2y agoHugging Face06yayoimizuha /Glint360k Dataset Card for Glint360K Citiation by InsightFace Repository We clean, merge, and release the largest and cleanest face recognition dataset Glint360K, which contains 17091657 images of 360232 individuals. By employing the Patial FC training strategy, baseline models trained on Glint360K can easily achieve state-of-the-art performance. Detailed evaluation results on the large-scale test set (e.g. IFRT, IJB-C and Megaface) are as follows: Dataset Details… See the full description on the dataset page: https://huggingface.co/datasets/yayoimizuha/Glint360k.imageimage-feature-extraction10M<n<100M5 likes1.6k downloads1y agoHugging Face07yuanyuan10 /ppv2_imageimage1M<n<10M0 likes1.1k downloads8mo agoHugging Face08Yuuuuuu9 /tarimage100K<n<1M0 likes939 downloads10mo agoHugging Face09YTang /UniSAR-7M UniSAR-7M A large-scale, multi-source synthetic aperture radar image corpus for self-supervised representation learning. UniSAR-7M contains 7,047,666 single-channel SAR image samples assembled from public SAR datasets and openly available imagery from commercial satellite constellations. It provides the pretraining corpus for DINOSAR, a self-supervised learning framework that uses Content-Aware Multi-Crop (CAMC) to construct informative views of SAR imagery. Associated… See the full description on the dataset page: https://huggingface.co/datasets/YTang/UniSAR-7M.image1M<n<10M1 likes714 downloads18d agoHugging Face10yann111 /GlobalGeoTree GlobalGeoTree Dataset GlobalGeoTree is a comprehensive global dataset for tree species classification, comprising 6.3 million geolocated tree occurrences spanning 275 families, 2,734 genera, and 21,001 species across hierarchical taxonomic levels. Each sample is paired with Sentinel-2 image time series and 27 auxiliary environmental variables. Dataset Structure This repository contains three main components: 1. GlobalGeoTree-6M Training dataset with around 6M… See the full description on the dataset page: https://huggingface.co/datasets/yann111/GlobalGeoTree.text1M<n<10M14 likes572 downloads11mo agoHugging Face11Yejy53 /Echo-4o-Image Echo-4o-Image Dataset Paper | Project Page | Code Introduction Echo-4o-Image is a 180K-scale synthetic dataset generated by GPT-4o, designed to advance open-source models in image generation. While real-world image datasets are valuable, synthetic images offer crucial advantages, especially in addressing blind spots in real-world coverage: Complementing Rare Scenarios: Synthetic data can generate examples for scenarios less represented in real-world datasets, such as… See the full description on the dataset page: https://huggingface.co/datasets/Yejy53/Echo-4o-Image.imagetext-to-image1K<n<10K34 likes350 downloads1y agoHugging Face12YanFang /dense-sc-wdsimage1M<n<10M0 likes319 downloads2y agoHugging Face13TrackingTeam /yuxuan_good_dataset_dtimage1M<n<10M0 likes312 downloads1mo agoHugging Face14YanFang /sc-wdsimage10M<n<100M0 likes282 downloads2y agoHugging Face15yu2hi13 /Dynamicvlmimage100K<n<1M0 likes270 downloads1y agoHugging Face16TrackingTeam /yuxuan_good_dataset_sttimage1M<n<10M0 likes199 downloads1mo agoHugging Face17ytaek-oh /cc3m-subset-100kimage100K<n<1M0 likes174 downloads2y agoHugging Face18Leonardo6 /yfcc15mimage10M<n<100M2 likes162 downloads1y agoHugging Face19yhx12 /VideoSSR-30kimage100K<n<1M0 likes162 downloads10mo agoHugging Face20Yiming1234 /VoT-video-latent-archiveimage100K<n<1M0 likes160 downloads6mo agoHugging Face21ribster /sd-yffimage100K<n<1M0 likes157 downloads4y agoHugging Face22yunfanlu /SEE-600Kimage1M<n<10M4 likes145 downloads10mo agoHugging Face23CVML-TueAI /grounding-YT-dataset Grounding YouTube Dataset What, when, and where? -- Self-Supervised Spatio-Temporal Grounding in Untrimmed Multi-Action Videos from Narrated Instructions arxiv This dataset is packed in WebDataset format. The dataset is present in three styles: Untrimmed videos + annotations within the entire video Action clips extracted from the videos + annotations in each clip Action frames extracted from the videos + annotation of the frame Example usage for clips:… See the full description on the dataset page: https://huggingface.co/datasets/CVML-TueAI/grounding-YT-dataset.image10K<n<100K0 likes128 downloads10mo agoHugging Face24TrackingTeam /yuxuan_dataset_dtimage1M<n<10M0 likes113 downloads1mo agoHugging Face25yoonkyojung /TraceGenLibero TraceGen – LIBERO (Derived Subset) Overview This folder contains a derived subset generated from the LIBERO dataset using the TraceForge pipeline as part of the TraceGen project. TraceGen Project Website: https://tracegen.github.io/ Evaluation Protocol This dataset defines the official evaluation protocol for the TraceGen benchmark. Models are evaluated on five environments with the following metrics: Mean Squared Error (MSE) Mean Absolute Error (MAE)… See the full description on the dataset page: https://huggingface.co/datasets/yoonkyojung/TraceGenLibero.imagerobotics10K<n<100K0 likes110 downloads9mo agoHugging Face26yu2hi13 /YTVIS2021_Splitimage100K<n<1M0 likes109 downloads1y agoHugging Face27TrackingTeam /yuxuan_dataset_sttimage1M<n<10M0 likes109 downloads1mo agoHugging Face28YuanzeLin /IllumiCraft IllumiCraft Dataset This repository contains the dataset released with: IllumiCraft: Unified Geometry and Illumination Diffusion for Controllable Video Generation Yuanze Lin, Yi-Wen Chen, Yi-Hsuan Tsai, Ronald Clark, Ming-Hsuan Yang 🔗 Links 📄 Paper: https://arxiv.org/abs/2506.03150 🌐 Project Page: https://yuanze-lin.me/IllumiCraft_page/ 💻 GitHub: https://github.com/yuanze-lin/IllumiCraft 🎥 YouTube: https://youtu.be/qAV58sADEzo 🤗 Checkpoints:… See the full description on the dataset page: https://huggingface.co/datasets/YuanzeLin/IllumiCraft.image100K<n<1M1 likes98 downloads4mo agoHugging Face29sty-yyj /ElysiumTrack-1M Dataset Card ElysiumTrack-1M dataset is a million-scale object perception video dataset. It supports the following tasks: Single Object Tracking (SOT): Predicting the location of a specific object in consecutive frames by referencing its initial position in the first frame. Referring Single Object Tracking (RSOT): Identifying and locating a specific object within an entire video based on the given language expression. This task provides a more flexible tracking format and… See the full description on the dataset page: https://huggingface.co/datasets/sty-yyj/ElysiumTrack-1M.imagevisual-question-answering10M<n<100M4 likes88 downloads2y agoHugging Face30yanjuntong /lora_modelimage1K<n<10K0 likes87 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.