CoolFace
Datasetpublic

llm-lab/SpeechBrown

Models | Springer Link | arXiv Link | Proposed Dataset | ACM Digital Library | Website Dataset Summary Speech Brown is a comprehensive, synthetic, and diverse paired speech-text dataset in 15 categories, covering a wide range of topics from fiction to religion. This dataset consists of over 55,000 sentence-level samples. To train the CLASP model, we created this dataset based on the Brown Corpus. The synthetic speech was generated using the NVIDIA Tacotron 2 text-to-speech… See the full description on the dataset page: https://huggingface.co/datasets/llm-lab/SpeechBrown.

sourceHugging Facemitupdated 1y agoView on Hugging Face
2likes86downloads
settings

This repository belongs to llm-lab on Hugging Face.

CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.

nameSpeechBrown
visibilitypublic
licencemit
gatedno
ownerllm-lab
Account settings
llm-lab/SpeechBrown · CoolFace