CoolFace
Datasetpublic

rishchen/ukrainian-tts-audiobook-pani-nina-parquet

Ukrainian TTS audiobook dataset Pani Nina (Parquet) Segmented Ukrainian audiobook speech with aligned text, prepared for training and evaluating Text-to-Speech (TTS) models. The dataset is published as Hugging Face-compatible Parquet shards so the Hub Dataset Preview can render an audio column. The dataset was prepared using whisper and ffmpeg: Whisper was used for transcription and approximate segment timing. FFmpeg was used to slice audio into short utterances (roughly 2-10… See the full description on the dataset page: https://huggingface.co/datasets/rishchen/ukrainian-tts-audiobook-pani-nina-parquet.

sourceHugging Facecc-by-nc-sa-4.0updated 5mo agoView on Hugging Face
0likes17downloads
9 commits on main
500f3495mo ago

Update README.md

rishchen
67cd9285mo ago

Update README.md

rishchen
bcaa49a5mo ago

Update README.md

rishchen
32b03655mo ago

Update README.md

rishchen
4cd4ac25mo ago

Update README.md

rishchen
a38f11c5mo ago

Update README.md

rishchen
134807b6mo ago

Update README.md

rishchen
4311ca16mo ago

Add files using upload-large-folder tool

rishchen
c9c90c47mo ago

initial commit

rishchen