CoolFace
Datasetpublic

cahya/librivox-indonesia

Dataset Card for LibriVox Indonesia 1.0 Dataset Summary The LibriVox Indonesia dataset consists of MP3 audio and a corresponding text file we generated from the public domain audiobooks LibriVox. We collected only languages in Indonesia for this dataset. The original LibriVox audiobooks or sound files' duration varies from a few minutes to a few hours. Each audio file in the speech dataset now lasts from a few seconds to a maximum of 20 seconds. We converted… See the full description on the dataset page: https://huggingface.co/datasets/cahya/librivox-indonesia.

sourceHugging Faceccupdated 3y agoView on Hugging Face
2likes80downloads
discussions and pull requests

Conversations for this repository live on Hugging Face.

CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.

Open discussions on Hugging Face
cahya/librivox-indonesia · CoolFace