CoolFace
11 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01Smith42 /minty-astro-ph MINT-1T ArXiv Astro-ph An astronomy-focused subset of mlfoundations/MINT-1T-ArXiv, filtered to include only papers from the astro-ph arXiv category (including cross-listed papers). Overview Papers ~845k Total size ~804 GB Format WebDataset tar shards Shards 287 (astro-ph-00000.tar to astro-ph-00286.tar) Shard size ~3 GB each Source MINT-1T (Awadalla et al., 2024) Data Format Each tar shard contains paired files per paper:… See the full description on the dataset page: https://huggingface.co/datasets/Smith42/minty-astro-ph.imagetext-generation100K<n<1M1 likes4.6k downloads5mo agoHugging Face02smileyenot983 /objaversexl_sketchfabimage100K<n<1M0 likes369 downloads6mo agoHugging Face03smileyenot983 /objaversexl_sketchfab_pmap1image100K<n<1M0 likes218 downloads5mo agoHugging Face04smileyenot983 /objaversexl_sketchfab_pmapimage100K<n<1M0 likes207 downloads6mo agoHugging Face05smileyenot983 /objaversexl_github6.5image10K<n<100K0 likes99 downloads6mo agoHugging Face06JHU-SmileLab /NaturalVoices_VC_0.1 NaturalVoices VC 10% A large voice conversion (VC) dataset curated from spontaneous, in-the-wild podcast speech as part of the NaturalVoices project in collaboration with 🤗MSP Lab at CMU LTI. This release provides the 10% subset uniformly sampled from 870-hour VC dataset and subsets mainly intended for training and evaluating emotion-aware voice conversion systems but not limited to VC tasks. 📄 Paper: NaturalVoices: A Large-Scale, Spontaneous and Emotional Podcast Dataset for Voice… See the full description on the dataset page: https://huggingface.co/datasets/JHU-SmileLab/NaturalVoices_VC_0.1.audioaudio-to-audio10K<n<100K2 likes82 downloads11mo agoHugging Face07SMIIP-lab /AISHELL6-Whispergated 🗣️ AISHELL6-Whisper AISHELL6-Whisper is a large-scale open-source Chinese Mandarin audio-visual whisper speech dataset,containing 30 hours each of whisper and parallel normal speech, with synchronized frontal RGB facial videos. 📘 Dataset Summary Property Description Language Chinese (Mandarin, ZH) License CC BY-NC-SA 4.0 Duration ~60 hours total (30 h whisper + 30 h normal) Speakers 167 total (121 with RGB-D, 46 audio-only) Environment Controlled… See the full description on the dataset page: https://huggingface.co/datasets/SMIIP-lab/AISHELL6-Whisper.audio10K<n<100K8 likes71 downloads8mo agoHugging Face08smileyenot983 /objaversexl_sketchfab_pmap2image10K<n<100K0 likes6 downloads5mo agoHugging Face09conceptofmind /smithsonian-batch-1gatedimage1M<n<10M0 likes2 downloads2y agoHugging Face10conceptofmind /smithsonian-batch-2gatedimage10M<n<100M3 likes2 downloads2y agoHugging Face11conceptofmind /smithsonian-batch-1-oldgatedimage1M<n<10M0 likes1 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.