Harisri/indic-fusion-hi-mr
Indic-FUSION Hindi-Marathi Pseudo-Parallel Speech A research pseudo-parallel speech corpus for Indic-FUSION. Construction Source speech is sampled from ai4bharat/indicvoices_r. For each source utterance: The source dataset transcript is used directly. NLLB generates a direct Indic-to-Indic target transcript. ai4bharat/IndicF5 synthesizes target-language speech using the source utterance as the reference prompt. The target audio is therefore synthetic, not… See the full description on the dataset page: https://huggingface.co/datasets/Harisri/indic-fusion-hi-mr.
Conversations for this repository live on Hugging Face.
CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.
Open discussions on Hugging Face