CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01asahi417 /experiment-speaker-embeddingaudion<1K0 likes724 downloads2y agoHugging Face02ittailup /peruvian_speech_w_embeddings Dataset Card for "peruvian_speech_w_embeddings" More Information needed audio10K<n<100K0 likes173 downloads2y agoHugging Face03Reza2kn /persian-asr-untrained-embeddings-v1 Persian ASR Untrained Corpus Embeddings v1 This dataset stores manifest and embedding artifacts for Persian ASR data mining. The corpus is intended for acoustic/textual clustering, diversity selection, noise/environment mining, and training-data planning for VisualEars-style robust Persian ASR. Contents manifests/all_untrained_manifest.v1.jsonl: canonical row/file manifest after excluding known trained/eval content where available.… See the full description on the dataset page: https://huggingface.co/datasets/Reza2kn/persian-asr-untrained-embeddings-v1.audio100K<n<1M0 likes123 downloads4mo agoHugging Face04jordand /echo-embeddings-custom Custom Speaker Embeddings Contains speaker folders within HF-Custom, each with: a precomputed speaker embedding (speaker_latent.safetensors) its corresponding audio (audio.mp3) a metadata file describing the voice and licensing (metadata.json) Licensing:There is no single license for this dataset. Each voice has its own terms stored in its metadata.json. You must check the metadata for any voice you use. audion<1K3 likes58 downloads10mo agoHugging Face05thenewsupercell /V2_audio_embeddings_celeb-df-datasetaudio10K<n<100K0 likes57 downloads2y agoHugging Face06jordand /echo-embeddings-vctk-tar VCTK Speaker Embeddings (tarred) Items: 109 This dataset ships as a single tar at the repo root. Members preserve paths like VCTK/<id>/audio.mp3 and VCTK/<id>/speaker_latent.safetensors. See loader.py for example loading. Attribution: Contains audio and embeddings derived from the CSTR VCTK Corpus. Distributed under CC BY 4.0; attribution required. textn<1K0 likes56 downloads10mo agoHugging Face07jordand /echo-embeddings-expresso-tar Expresso Speaker Embeddings (tarred) Items: 17 This dataset ships as a single tar at the repo root. Members preserve paths like Expresso/<id>/audio.mp3 and Expresso/<id>/speaker_latent.safetensors. See loader.py for example loading. Attribution: Contains audio and embeddings derived from the Expresso dataset (INTERSPEECH 2023). Distributed under CC BY-NC 4.0; attribution required; commercial use is not permitted. textn<1K0 likes55 downloads10mo agoHugging Face08jordand /echo-embeddings-ears-tar EARS Speaker Embeddings (tarred) Items: 2568 This dataset ships as a single tar at the repo root. Members preserve paths like EARS/<id>/audio.mp3 and EARS/<id>/speaker_latent.safetensors. See loader.py for example loading. Attribution: Contains audio and embeddings derived from the EARS dataset. Distributed under CC BY-NC 4.0; attribution required; commercial use is not permitted. text1K<n<10K1 likes46 downloads10mo agoHugging Face09treadon /fma-mert-embeddings FMA-MERT Embeddings Pre-computed MERT-v1-330M embeddings for the FMA-Small dataset. 7,997 tracks, each represented as a 1024-dimensional vector, with banger scores (0-10) derived from log-normalized play counts. Use this dataset to train music quality scorers, explore music similarity, or experiment with audio representation learning -- without needing to download 7.2 GB of audio or run MERT yourself. Dataset Description Each row represents one track from FMA-Small… See the full description on the dataset page: https://huggingface.co/datasets/treadon/fma-mert-embeddings.tabularaudio-classification1K<n<10K0 likes34 downloads6mo agoHugging Face10thenewsupercell /V1_finetuned_audio_embeddings_celeb-df-datasetaudio10K<n<100K0 likes22 downloads2y agoHugging Face11orwelian84 /arc-music-embeddings ARC Music Embeddings Pre-computed CLAP embeddings for 5,050 music tracks with 291,468 segments, ready to use for semantic music similarity search and DJ-style transitions. Authors: Claude and his monkey Files File Size Description segment_embeddings.npz 553MB Full segment-level embeddings (5050 tracks) tracklist.txt 350KB Complete track listing with IDs and titles embeddings.npz 3.5MB Legacy track-level embeddings (1836 tracks) Embedding… See the full description on the dataset page: https://huggingface.co/datasets/orwelian84/arc-music-embeddings.text1K<n<10K0 likes21 downloads7mo agoHugging Face12foxhound /cv_25_pt_br_ECAPA_TDNN_embeddingsaudio100K<n<1M0 likes19 downloads2mo agoHugging Face13thenewsupercell /finetuned_avg_pooling_DF_Audio_Embeddingsaudio10K<n<100K0 likes17 downloads2y agoHugging Face14yutakobayashi /diet-members-voice-embeddings diet-members-voice-embeddings 日本の国会議員の声を speechbrain/spkrec-ecapa-voxcelebで embedding したデータセットです。話者分離などのタスクで使用できます。 国会中継や演説等の分析など、ご自由にお使いください。 使用例 以下はトランスクリプトと音声ファイルを元に、話者分析を行う例です。 pip install pandas numpy wave ast scipy pyannote.audio import pandas as pd import numpy as np import contextlib import wave import ast from typing import List, Tuple from scipy.spatial.distance import cosine from pyannote.audio import Audio from pyannote.core importSegment from… See the full description on the dataset page: https://huggingface.co/datasets/yutakobayashi/diet-members-voice-embeddings.audio0 likes16 downloads3y agoHugging Face15thenewsupercell /audio_embeddings_celeb-df-datasetaudio10K<n<100K0 likes14 downloads2y agoHugging Face16thenewsupercell /emotion_max_pooling_DF_Audio_Embeddingsaudio10K<n<100K0 likes14 downloads2y agoHugging Face17thenewsupercell /new_finetuned_avg_pooling_DF_Audio_Embeddingsaudio10K<n<100K0 likes14 downloads2y agoHugging Face18thenewsupercell /finetuned_max_pooling_DF_Audio_Embeddingsaudio10K<n<100K0 likes12 downloads2y agoHugging Face19Biorrith /coral_tts_with_embeddings_v2audio10K<n<100K0 likes12 downloads8mo agoHugging Face20thenewsupercell /new_finetuned_max_pooling_DF_Audio_Embeddingsaudio10K<n<100K0 likes11 downloads2y agoHugging Face21nineninesix /expresso_sp_embeddingsaudio10K<n<100K0 likes11 downloads8mo agoHugging Face22thenewsupercell /new_regular_avg_pooling_DF_Audio_Embeddingsaudio10K<n<100K0 likes9 downloads2y agoHugging Face23thenewsupercell /new_emotion_max_pooling_DF_Audio_Embeddingsaudio10K<n<100K0 likes8 downloads2y agoHugging Face24meandyou200175 /word_embeddingaudio10K<n<100K1 likes8 downloads1y agoHugging Face25linhqyy /result_with_finetuned_taggenv2_20epoch_encoder_embeddings Dataset Card for "result_with_finetuned_taggenv2_20epoch_encoder_embeddings" More Information needed audio1K<n<10K0 likes6 downloads3y agoHugging Face26thenewsupercell /emotion_avg_pooling_DF_Audio_Embeddingsaudio10K<n<100K0 likes6 downloads2y agoHugging Face27linhqyy /result_with_finetuned_taggenv2_9epoch_encoder_embeddings Dataset Card for "result_with_finetuned_taggenv2_9epoch_encoder_embeddings" More Information needed audio1K<n<10K0 likes5 downloads3y agoHugging Face28linhqyy /result_with_finetuned_taggenv2_10epoch_encoder_embeddings_decoder_roberta Dataset Card for "result_with_finetuned_taggenv2_10epoch_encoder_embeddings_decoder_roberta" More Information needed audio1K<n<10K0 likes5 downloads3y agoHugging Face29dthomas84 /rule1_embeddingsaudion<1K0 likes5 downloads3y agoHugging Face30thenewsupercell /new_emotion_avg_pooling_DF_Audio_Embeddingsaudio10K<n<100K0 likes5 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.