CoolFace
Datasetpublic

psk/malayalam-speech-178h

Malayalam Speech — 179 hours (denoised, unlabeled) 86,799 Malayalam speech clips, 48 kHz stereo WAV, ~7.4 s average. No transcripts — this is unlabeled audio, intended for self-supervised pretraining, voice/speaker modelling, or as raw material for your own labelling pipeline. Processing Each clip passed through a full source-separation and enhancement chain: Vocal isolation — BS-RoFormer De-reverberation — UVR-DeEcho-DeReverb Noise removal — UVR-DeNoise-Lite… See the full description on the dataset page: https://huggingface.co/datasets/psk/malayalam-speech-178h.

sourceHugging Faceupdated 2mo agoView on Hugging Face
4likes440downloads
discussions and pull requests

Conversations for this repository live on Hugging Face.

CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.

Open discussions on Hugging Face
psk/malayalam-speech-178h · CoolFace