silvermango9927/synthetic-asr-hi
Synthetic ASR data — hi Generated by Valsea-ASR/synthetic-data-pipeline. Audio is synthetic (TTS), targeted as training data for downstream ASR finetuning. Total audio: 63.5 hr across short (5s) and long (30s) length buckets, each in clean and augmented variants. Loading from datasets import load_dataset ds = load_dataset("<org>/synthetic-asr-hi", "short_clean") print(ds["train"][0]["audio"]) # {"array": np.ndarray, "sampling_rate": 16000, "path": "..."}… See the full description on the dataset page: https://huggingface.co/datasets/silvermango9927/synthetic-asr-hi.
06
No card is published for this repository, or it could not be fetched from Hugging Face right now.
