datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
gsat-vocab-sentences-tts
GSAT Vocabulary TTS Audio
Text-to-speech audio files for GSAT (General Scholastic Ability Test) English vocabulary.
Structure
audio/ - MP3 audio files organized by hash prefix (e.g., audio/ab/abcd1234....mp3)
index.jsonl - Index file mapping hashes to text and TTS engine used
Engines
Kokoro (af_heart voice) - Used for lemmas (single words/phrases)
Supertonic (M1 voice) - Used for example sentences
Audio Format
Format: MP3
Sample rate: 24kHz… See the full description on the dataset page: https://huggingface.co/datasets/TCabbage/gsat-vocab-sentences-tts.ShalevVoice17
