raditotev/bg-audiobooks-tts
Bulgarian Audiobook Speech Dataset A high-quality Bulgarian speech dataset derived from audiobooks narrated by Plamen Sivov, suitable for text-to-speech (TTS) and automatic speech recognition (ASR) tasks. Dataset Summary Property Value Language Bulgarian (bg) Total Duration 15.2 hours Total Clips 10,627 Speaker Plamen Sivov (single speaker) Source YouTube audiobooks Sample Rate 24,000 Hz Audio Format WAV, mono, 16-bit PCM Clip Duration… See the full description on the dataset page: https://huggingface.co/datasets/raditotev/bg-audiobooks-tts.
Conversations for this repository live on Hugging Face.
CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.
Open discussions on Hugging Face