datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
hebrew_speech_coursera
Dataset Card for Dataset Name
Dataset Summary
This dataset card aims to be a base template for new datasets. It has been generated using this raw template.
Supported Tasks and Leaderboards
[More Information Needed]
Languages
[More Information Needed]
Dataset Structure
Data Instances
{'audio': {'path':… See the full description on the dataset page: https://huggingface.co/datasets/imvladikon/hebrew_speech_coursera.audio-course-bark-samplesaudio_assignmentPersian_Course_TTS
Persian Course TTS
Dataset Summary
Persian Course TTS is a Persian (Farsi) speech–text dataset built from publicly available online course videos.The dataset is designed for text-to-speech (TTS) and speech synthesis fine-tuning tasks.It pairs natural instructional speech audio with refined textual transcriptions generated by a large language model (LLM).
The dataset currently consists of three subsets, which are extracted from three different educational sources… See the full description on the dataset page: https://huggingface.co/datasets/hoseinshr1055/Persian_Course_TTS.formosa_course
