CoolFace
Datasetpublicgated

speechcolab/gigaspeech

Dataset Card for Gigaspeech Dataset Description GigaSpeech is an evolving, multi-domain English speech recognition corpus with 10,000 hours of high quality labeled audio suitable for supervised training. The transcribed audio data is collected from audiobooks, podcasts and YouTube, covering both read and spontaneous speaking styles, and a variety of topics, such as arts, science, sports, etc. Example Usage The training split has several… See the full description on the dataset page: https://huggingface.co/datasets/speechcolab/gigaspeech.

sourceHugging Faceapache-2.0updated 8mo agoView on Hugging Face
173likes18kdownloads
Dataset Card

No card is published for this repository, or it could not be fetched from Hugging Face right now.