CoolFace
Datasetpublicgated

silvermango9927/synthetic-asr-hi

Synthetic ASR data — hi Generated by Valsea-ASR/synthetic-data-pipeline. Audio is synthetic (TTS), targeted as training data for downstream ASR finetuning. Total audio: 63.5 hr across short (5s) and long (30s) length buckets, each in clean and augmented variants. Loading from datasets import load_dataset ds = load_dataset("<org>/synthetic-asr-hi", "short_clean") print(ds["train"][0]["audio"]) # {"array": np.ndarray, "sampling_rate": 16000, "path": "..."}… See the full description on the dataset page: https://huggingface.co/datasets/silvermango9927/synthetic-asr-hi.

sourceHugging Faceupdated 3mo agoView on Hugging Face
0likes8downloads
discussions and pull requests

Conversations for this repository live on Hugging Face.

CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.

Open discussions on Hugging Face
silvermango9927/synthetic-asr-hi · CoolFace