CoolFace
Datasetpublic

DJRHails/pyannote-embedding-commonvoice-en

Pre-computed speaker embeddings Pre-computed 512-dim L2-normalized speaker embeddings extracted with pyannote/embedding over commonvoice-en. One utterance per speaker, minimum 3 s duration. Contents commonvoice-en.pyannote-embedding.npz — numpy .npz archive with: embeddings: (5000, 512) float32 speaker_ids: (5000,) string IDs from the source corpus metadata_json: per-speaker metadata (accent / age / gender / source URL) — populated for 5000 / 5000 speakers… See the full description on the dataset page: https://huggingface.co/datasets/DJRHails/pyannote-embedding-commonvoice-en.

sourceHugging Facecc0-1.0updated 4mo agoView on Hugging Face
0likes11downloads

No commit history came back for main. The revision may not exist, or the source declined the request.