CoolFace
Datasetpublic

rishchen/ukrainian-tts-audiobook-pani-nina-parquet

Ukrainian TTS audiobook dataset Pani Nina (Parquet) Segmented Ukrainian audiobook speech with aligned text, prepared for training and evaluating Text-to-Speech (TTS) models. The dataset is published as Hugging Face-compatible Parquet shards so the Hub Dataset Preview can render an audio column. The dataset was prepared using whisper and ffmpeg: Whisper was used for transcription and approximate segment timing. FFmpeg was used to slice audio into short utterances (roughly 2-10… See the full description on the dataset page: https://huggingface.co/datasets/rishchen/ukrainian-tts-audiobook-pani-nina-parquet.

sourceHugging Facecc-by-nc-sa-4.0updated 5mo agoView on Hugging Face
0likes17downloads
settings

This repository belongs to rishchen on Hugging Face.

CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.

nameukrainian-tts-audiobook-pani-nina-parquet
visibilitypublic
licencecc-by-nc-sa-4.0
gatedno
ownerrishchen
Account settings
rishchen/ukrainian-tts-audiobook-pani-nina-parquet · CoolFace