CoolFace
22 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01OpenDriveLab-org /SimScale Haochen Tian, Tianyu Li, Haochen Liu, Jiazhi Yang, Yihang Qiu, Guang Li, Junli Wang, Yinfeng Gao, Zhang Zhang, Liang Wang, Hangjun Ye, Tieniu Tan, Long Chen, Hongyang Li 📧 Primary Contact: Haochen Tian (tianhaochen2023@ia.ac.cn) 📜 Materials: 🌐 𝕏 | 📰 Media| 🗂️ Slides | 🎬 Talk (in Chinese) 🖊️ Joint effort by CASIA, OpenDriveLab at HKU, and Xiaomi EV. 🔥 Highlights 🏗️ A scalable simulation pipeline that synthesizes diverse and… See the full description on the dataset page: https://huggingface.co/datasets/OpenDriveLab-org/SimScale.imagerobotics10M<n<100M1 likes757 downloads8mo agoHugging Face02OrcinusOrca /YouTube-Cantonese Cantonese Audio Dataset from YouTube This dataset contains Cantonese audio segments and creator uploaded transcripts (likely higher quality) extracted from various YouTube channels, along with corresponding transcript metadata. The data is intended for training automatic speech recognition (ASR) models. Data Source and Processing The data was obtained through the following process: Download: Audio (.m4a) and available Cantonese subtitles (.srt for zh-TW, zh-HK, zh-Hant)… See the full description on the dataset page: https://huggingface.co/datasets/OrcinusOrca/YouTube-Cantonese.audioautomatic-speech-recognition100K<n<1M5 likes445 downloads1y agoHugging Face03orion-ai-lab /Thalia Thalia: A Global, Multi-Modal Dataset for Volcanic Activity Monitoring Paper | GitHub | Interactive Demo (Colab) Thalia is a global, multi-modal dataset for volcanic activity monitoring through Satellite-based Interferometric Synthetic Aperture Radar (InSAR) imagery. Building upon the Hephaestus dataset, Thalia provides higher-resolution, multi-source, and multi-temporal data in a machine-learning-ready format. Dataset Overview Thalia consists of 38 spatiotemporal… See the full description on the dataset page: https://huggingface.co/datasets/orion-ai-lab/Thalia.textimage-classification10K<n<100K3 likes292 downloads5mo agoHugging Face04Viglong /OriAnyV2_Train_Render Orient Anything V2 Dataset Project Page | Paper | GitHub Orient Anything V2 is an enhanced foundation model for unified understanding of object 3D orientation and rotation from single or paired images. This repository contains the training data (final rendering data) used for the model. Sample Usage Below is a snippet to run inference using the model and data logic, as found in the official GitHub repository: import numpy as np from PIL importImage import torch import… See the full description on the dataset page: https://huggingface.co/datasets/Viglong/OriAnyV2_Train_Render.imageother1M<n<10M6 likes247 downloads9mo agoHugging Face05WenhaoWang /OriPID Summary This is the dataset proposed in our paper Origin Identification for Text-Guided Image-to-Image Diffusion Models (ICML 2025). Download Training You can download the images: wget https://huggingface.co/datasets/WenhaoWang/OriPID/resolve/main/training/sd2_d_multi.tar.part_0{0..9} cat sd2_d_multi.tar.part_* > sd2_d_multi.tar tar -xvf sd2_d_multi.tar Or you can directly download the features extracted by VAE in Stable Diffusion 2: wget… See the full description on the dataset page: https://huggingface.co/datasets/WenhaoWang/OriPID.imagetext-to-image1M<n<10M0 likes154 downloads1y agoHugging Face06orangewen /Gen-nuScenesimage1M<n<10M2 likes115 downloads2y agoHugging Face07OrcinusOrca /YouTube-English English Audio Dataset from YouTube This dataset contains English audio segments and creator uploaded transcripts (likely higher quality) extracted from various YouTube channels, along with corresponding transcript metadata. The data is intended for training automatic speech recognition (ASR) models. Data Source and Processing The data was obtained through the following process: Download: Audio (.m4a) and available English subtitles (.srt for en, en.j3PyPqV-e1s) were… See the full description on the dataset page: https://huggingface.co/datasets/OrcinusOrca/YouTube-English.audioautomatic-speech-recognition100K<n<1M2 likes75 downloads1y agoHugging Face08skip113 /bmc_original_1Mimage1M<n<10M0 likes62 downloads1y agoHugging Face09haideraltahan /wds_flickr30k_orderimage1K<n<10K0 likes48 downloads2y agoHugging Face10clip-benchmark /wds_vtab-dsprites_label_orientationimage100K<n<1M0 likes35 downloads4y agoHugging Face11haideraltahan /wds_coco_orderimage10K<n<100K0 likes27 downloads2y agoHugging Face12othsueh /CombineCorpus_6_Orgtext10K<n<100K0 likes22 downloads1y agoHugging Face13haideraltahan /wds_dsprites_label_orientationimage10K<n<100K0 likes16 downloads2y agoHugging Face14nyu-visionx /oro_depth_rewardimage100K<n<1M0 likes16 downloads2y agoHugging Face15djghosh /wds_vtab-dsprites_label_orientation_test dSprites Orientation (Test set only) Original paper: beta-VAE: Learning Basic Visual Concepts with a Constrained Variational Framework Homepage: https://github.com/deepmind/dsprites-dataset Bibtex: @misc{dsprites17, author = {Loic Matthey and Irina Higgins and Demis Hassabis and Alexander Lerchner}, title = {dSprites: Disentanglement testing Sprites dataset}, howpublished= {https://github.com/deepmind/dsprites-dataset/}, year = "2017", } image10K<n<100K0 likes12 downloads4y agoHugging Face16nerako /ori_coco_evalimage1K<n<10K0 likes9 downloads1y agoHugging Face17yiyangd /bridge_orig_text_20kimage10K<n<100K0 likes8 downloads8mo agoHugging Face18orgniumo /scenaversetext10K<n<100K0 likes7 downloads2y agoHugging Face19OrkaZeta /EuroSAT_RGB_Cimage1M<n<10M0 likes7 downloads7mo agoHugging Face20yyyzzzzyyy /train_original_sd3.5-limage10K<n<100K0 likes6 downloads4mo agoHugging Face21hnmr-org /zungatedaudio10K<n<100K0 likes1 downloads2y agoHugging Face22MussEsSein /origin_data_20260221_raw_imgimage100K<n<1M0 likes1 downloads5mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.