CoolFace
Datasetpublic

resproj007/sesame_tts_knn_vc_pathological

Dataset Overview Total Samples: 759 Total Duration: 2892.82 seconds (48.21 minutes) Speakers: 8 speakers Corpora: TORGO, UA-Speech, LibriSpeech Sample Rate: 16kHz (KNN-VC output rate) Audio Format: WAV

sourceHugging Faceupdated 1y agoView on Hugging Face
0likes9downloads
Dataset Card

Dataset Overview

  • Total Samples: 759
  • Total Duration: 2892.82 seconds (48.21 minutes)
  • Speakers: 8 speakers
  • Corpora: TORGO, UA-Speech, LibriSpeech
  • Sample Rate: 16kHz (KNN-VC output rate)
  • Audio Format: WAV