ebellob/voxpopuli_spanish_enhanced
VoxPopuli Spanish Enhanced (CleanUNet + FlashSR) Dataset Summary This dataset is a processed and enhanced version of the Spanish subset of: facebook/voxpopuli. Furthermore, as this is a personal project, we give no guarantees that the audio is completely clean from any artifacts or noise the CleanUNet model could not remove. However, we have personally tested the corpus via the fine-tuning of some SOTA speech models and the results have been satisfactory. In… See the full description on the dataset page: https://huggingface.co/datasets/ebellob/voxpopuli_spanish_enhanced.
135
