CoolFace
Datasetpublicgated

silvermango9927/synthetic-asr-hi

Synthetic ASR data — hi Generated by Valsea-ASR/synthetic-data-pipeline. Audio is synthetic (TTS), targeted as training data for downstream ASR finetuning. Total audio: 63.5 hr across short (5s) and long (30s) length buckets, each in clean and augmented variants. Loading from datasets import load_dataset ds = load_dataset("<org>/synthetic-asr-hi", "short_clean") print(ds["train"][0]["audio"]) # {"array": np.ndarray, "sampling_rate": 16000, "path": "..."}… See the full description on the dataset page: https://huggingface.co/datasets/silvermango9927/synthetic-asr-hi.

sourceHugging Faceupdated 3mo agoView on Hugging Face
0likes6downloads
Dataset Card

No card is published for this repository, or it could not be fetched from Hugging Face right now.

silvermango9927/synthetic-asr-hi · CoolFace