yigagilbert/synthetic-parallel-salt
Synthetic Parallel EN↔LG — salt Voice-controlled synthetic parallel speech dataset for Luganda-English speech-to-speech translation, generated by the Hibiki-Zero fine-tuning pipeline. Generation Component Model Translation Sunbird/translate-nllb-3.3b-salt TTS Sunbird/orpheus-3b-tts-multilingual English speakers: salt_eng_0001, salt_eng_0002, salt_eng_0003 Luganda speakers: salt_lug_0001, waxal_lug_0001, waxal_lug_0002, waxal_lug_0003… See the full description on the dataset page: https://huggingface.co/datasets/yigagilbert/synthetic-parallel-salt.
Synthetic Parallel EN↔LG — salt
Voice-controlled synthetic parallel speech dataset for Luganda-English speech-to-speech translation, generated by the Hibiki-Zero fine-tuning pipeline.
Generation
English speakers: salt_eng_0001, salt_eng_0002, salt_eng_0003
Luganda speakers: salt_lug_0001, waxal_lug_0001, waxal_lug_0002, waxal_lug_0003, waxal_lug_0004, waxal_lug_0005, waxal_lug_0007
Speaker sampling: N ∈ {1 (70%), 2 (20%), 3 (10%)} draws per text pair.
Source datasets (transcripts only — source audio discarded)
Sunbird/tts(subset:eng, lang: eng)Sunbird/tts(subset:lug, lang: lug)Sunbird/speech(subset:eng_salt, lang: eng)Sunbird/speech(subset:lug_salt, lang: lug)
Post-processing
- Silero / energy VAD silence trim on every clip
- Duration filter: Luganda 1.5–20.0 s, English 1.0–20.0 s
- Duration ratio filter: 0.3–4.0
- Speech-activity filter: ≥ 0.35
- Sample rate: 24000 Hz (mono, float32 → PCM-16 WAV)
Schema
Failure manifest
Rows that failed TTS generation (after 3 retries) or post-processing are logged to: https://huggingface.co/datasets/yigagilbert/synthetic-parallel-salt/resolve/main/failure_manifest.jsonl
