silvermango9927/synthetic-asr-hi
Synthetic ASR data — hi Generated by Valsea-ASR/synthetic-data-pipeline. Audio is synthetic (TTS), targeted as training data for downstream ASR finetuning. Total audio: 63.5 hr across short (5s) and long (30s) length buckets, each in clean and augmented variants. Loading from datasets import load_dataset ds = load_dataset("<org>/synthetic-asr-hi", "short_clean") print(ds["train"][0]["audio"]) # {"array": np.ndarray, "sampling_rate": 16000, "path": "..."}… See the full description on the dataset page: https://huggingface.co/datasets/silvermango9927/synthetic-asr-hi.
Conversations for this repository live on Hugging Face.
CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.
Open discussions on Hugging Face