CoolFace
Datasetpublic

AigizK/bashkort_voice

Bashkort Voice 🇬🇧 English Version Dataset Description This is a synthetic Bashkir audio dataset generated using the OmniVoice model. It is designed to expand the availability of spoken data for the Bashkir language. Data Preparation Process The dataset was constructed through a cross-lingual voice cloning and generation process, using the following methodology: Target Text: Bashkir sentences were extracted from the AigizK/bashkir-russian-parallel-corpora… See the full description on the dataset page: https://huggingface.co/datasets/AigizK/bashkort_voice.

sourceHugging Facecc-by-4.0updated 5mo agoView on Hugging Face
2likes154downloads
5 commits on main
26871f35mo ago

Update README.md

AigizK
5bdac2d5mo ago

Update README.md

AigizK
80949126mo ago

Update README.md

AigizK
cc206db6mo ago

Add Bashkort Voice dataset (928632 samples)

AigizK
31100a76mo ago

initial commit

AigizK