raditotev/bg-audiobooks-tts
Bulgarian Audiobook Speech Dataset A high-quality Bulgarian speech dataset derived from audiobooks narrated by Plamen Sivov, suitable for text-to-speech (TTS) and automatic speech recognition (ASR) tasks. Dataset Summary Property Value Language Bulgarian (bg) Total Duration 15.2 hours Total Clips 10,627 Speaker Plamen Sivov (single speaker) Source YouTube audiobooks Sample Rate 24,000 Hz Audio Format WAV, mono, 16-bit PCM Clip Duration… See the full description on the dataset page: https://huggingface.co/datasets/raditotev/bg-audiobooks-tts.
This repository belongs to raditotev on Hugging Face.
CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.
