CoolFace
Datasetpublic

AigizK/bashkort_voice

Bashkort Voice 🇬🇧 English Version Dataset Description This is a synthetic Bashkir audio dataset generated using the OmniVoice model. It is designed to expand the availability of spoken data for the Bashkir language. Data Preparation Process The dataset was constructed through a cross-lingual voice cloning and generation process, using the following methodology: Target Text: Bashkir sentences were extracted from the AigizK/bashkir-russian-parallel-corpora… See the full description on the dataset page: https://huggingface.co/datasets/AigizK/bashkort_voice.

sourceHugging Facecc-by-4.0updated 5mo agoView on Hugging Face
2likes154downloads

AigizK/bashkort_voice · main · files are served by the source, never re-hosted here