datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
VCTKLJSpeech-1.1
The LJ Speech Dataset
Version 1.1
July 5, 2017
https://keithito.com/LJ-Speech-Dataset
OVERVIEW
This is a public domain speech dataset consisting of 13,100 short audio clips
of a single speaker reading passages from 7 non-fiction books. A transcription
is provided for each clip. Clips vary in length from 1 to 10 seconds and have
a total length of approximately 24 hours.
The texts were published between 1884 and 1964, and are in the public domain.
The audio was recorded in… See the full description on the dataset page: https://huggingface.co/datasets/badayvedat/LJSpeech-1.1.unarxive-2024-processed
