CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01Johnx69 /audio-3dvgaudio10K<n<100K0 likes419 downloads1y agoHugging Face02HaninZ /HowFarAreYou_3DSpeakerTrain_fullaudio10K<n<100K0 likes195 downloads2y agoHugging Face03nvidia /Audio2Face-3D-Dataset-v1.0.0-clairegated Dataset Description: NVIDIA Audio2Face-3D-dataset-v1.0.0-claire includes audio files, blendshape data, animated geometry caches, geometry files, and transform files. This dataset is for demonstration purposes and not for production usage. For source code, documentation, helper scripts, packaged builds, and links to all components in the Audio2Face-3D technology stack, visit the Audio2Face-3D GitHub repository Dataset Owner: NVIDIA Corporation Dataset… See the full description on the dataset page: https://huggingface.co/datasets/nvidia/Audio2Face-3D-Dataset-v1.0.0-claire.audion<1K22 likes155 downloads11mo agoHugging Face04Eimhin03 /combined_3datasets_801010audio1K<n<10K0 likes141 downloads6mo agoHugging Face05DynamicSuperbPrivate /HowFarAreYou_3DSpeakerTrain_full Dataset Card for "HowFarAreYou_3DSpeakerTrain" More Information needed audio10K<n<100K0 likes119 downloads3y agoHugging Face06BarryFutureman /vox2_3D_distill_shard22audio1K<n<10K0 likes105 downloads1y agoHugging Face07BarryFutureman /vox2_3D_distill_shard24audio10K<n<100K0 likes87 downloads1y agoHugging Face08BarryFutureman /vox2_3D_distill_shard21audio10K<n<100K0 likes86 downloads1y agoHugging Face09BarryFutureman /vox2_3D_distill_shard23audio10K<n<100K0 likes80 downloads1y agoHugging Face10BarryFutureman /vox2_3D_distill_shard13audio10K<n<100K0 likes79 downloads1y agoHugging Face11BarryFutureman /vox2_3D_distill_shard17audio10K<n<100K0 likes79 downloads1y agoHugging Face12BarryFutureman /vox2_3D_distill_shard16audio10K<n<100K0 likes76 downloads1y agoHugging Face13BarryFutureman /vox2_3D_distill_shard40audio10K<n<100K0 likes75 downloads1y agoHugging Face14BarryFutureman /vox2_3D_distill_shard42audio1K<n<10K0 likes75 downloads1y agoHugging Face15BarryFutureman /vox2_3D_distill_shard14audio10K<n<100K0 likes73 downloads1y agoHugging Face16BarryFutureman /vox2_3D_distill_shard32audio1K<n<10K0 likes73 downloads1y agoHugging Face17BarryFutureman /vox2_3D_distill_shard20audio1K<n<10K0 likes63 downloads1y agoHugging Face18DynamicSuperbPrivate /HowFarAreYou_3DSpeakerTrain Dataset Card for "HowFarAreYou_3DSpeakerTrain" More Information needed audio1K<n<10K0 likes41 downloads3y agoHugging Face19BarryFutureman /vox2_3D_combined_shard_2audio10K<n<100K0 likes40 downloads1y agoHugging Face20BarryFutureman /vox2_3D_distill_shard_06audio1K<n<10K0 likes24 downloads1y agoHugging Face21chiyuanhsiao /HowFarAreYou_3DSpeakerTrain Dataset Card for "HowFarAreYou_3DSpeakerTrain" More Information needed audio1K<n<10K1 likes22 downloads3y agoHugging Face22BarryFutureman /vox2_3D_shard0034audio1K<n<10K0 likes22 downloads1y agoHugging Face23TianshunHan /3D-CAVFA 3D-CAVFA Dataset Overview The 3D-CAVFA dataset is a multimodal collection featuring synchronized audio and facial blendshape coefficients captured from 20 subjects, totaling 15 hours. Its linguistic content encompasses frequently used modern Chinese words and phrases, alongside a variety of sentence patterns drawn from daily life, thereby guaranteeing strong relevance to real-world scenarios. Citation If you use this dataset, please consider citing… See the full description on the dataset page: https://huggingface.co/datasets/TianshunHan/3D-CAVFA.audio1K<n<10K0 likes22 downloads1y agoHugging Face24BarryFutureman /vox2_3D_distill_shard_18audio1K<n<10K0 likes21 downloads1y agoHugging Face25VINAY-UMRETHE /Emergence-Text-Image-Audio-3Dgated Emergence: The Four Forms of Intelligence Summary A multimodal dataset that unifies Text, Image, Audio, and 3D modalities with quad-modality alignment for every sample, ensuring that each record contains semantically consistent representations of the same concept. This dataset is curated by using 3D assets from Objaverse as anchors and aligning them with semantically corresponding images and audio clips from various sources using a embedding search… See the full description on the dataset page: https://huggingface.co/datasets/VINAY-UMRETHE/Emergence-Text-Image-Audio-3D.textany-to-any10K<n<100K0 likes21 downloads6mo agoHugging Face26DynamicSuperbPrivate /HowFarAreYou_3DSpeaker Dataset Card for "HowFarAreYou_3DSpeaker" More Information needed audio1K<n<10K0 likes20 downloads3y agoHugging Face27BarryFutureman /MEAD_3D_shard_06audio1K<n<10K0 likes20 downloads1y agoHugging Face28BarryFutureman /vox2_3D_distill_shard01audio10K<n<100K0 likes20 downloads1y agoHugging Face29BarryFutureman /MEAD_3D_W035audion<1K0 likes18 downloads1y agoHugging Face30Cybrpgs /mac-m4pro-confirmation-3drive-20260903gated sbpn-mac-m4pro-confirmation-three-drive-20260902 This dataset combines six independently aligned source archives. Each row embeds its selected MP3 in the audio Parquet column. SBPN-derived word timestamps are observational and do not control chunk edges or the Demucs vote. Accepted hard-word verbalizations are projected back to the original written forms; pronunciation_alignment_dictionary_json records the winning spoken form. Non-music tags are preserved using the existing… See the full description on the dataset page: https://huggingface.co/datasets/Cybrpgs/mac-m4pro-confirmation-3drive-20260903.audioautomatic-speech-recognition1K<n<10K0 likes18 downloads21d agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.