datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
700h-tr-turkish-text-to-speechdarija_speech_to_textnepali_speech_to_text
Nepali Speech-to-Text Dataset
This repository contains a dataset for Automatic Speech Recognition (ASR) in the Nepali language. The dataset is designed for supervised learning tasks and includes audio files along with their corresponding transcriptions. The audio samples have been collected from various open-source platforms and other publicly available sources on the internet.
Each audio file has an average length of 15 seconds and has been converted into a consistent WAV format… See the full description on the dataset page: https://huggingface.co/datasets/pujanpaudel/nepali_speech_to_text.darija-speech-to-text
Speech To Text Darija dataset
Reupload of adiren7/darija_speech_to_text
Tamazight-Speech-to-Arabic-Text
Tamazight-Arabic Speech Recognition Dataset
Overview
This is the EMINES organization-hosted version of the Tamazight-Arabic Speech Recognition Dataset, synchronized with the original dataset. It contains ~15.5 hours of Tamazight speech (Tachelhit dialect) paired with Arabic transcriptions, designed for developing ASR and translation systems.
Quick Start
from datasets import load_dataset
# Load the dataset
dataset =… See the full description on the dataset page: https://huggingface.co/datasets/EMINES/Tamazight-Speech-to-Arabic-Text.Afar-language-text-to-speech-TTS
Usage
This dataset is designed to support the development of Text-to-Speech (TTS) systems for the Afar language. It can be integrated into web applications, mobile apps, desktop software, or other platforms that require natural-sounding Afar voice synthesis or accurate spoken language recognition.
For applications involving virtual avatars or voice personas, the following culturally appropriate voice names are recommended:
Female Voices: Emeli, Hanaawi, Kareera, Laysani, Kulsuma… See the full description on the dataset page: https://huggingface.co/datasets/Charif-Ayfarah/Afar-language-text-to-speech-TTS.Sinhala_speech_to_text100_shruk-speech_to_Text__ASR_dataset
100 Speech-to-Text / ASR Shruk Dataset
Dataset Description
Traditional Kashmiri poetic verses (Shruks) paired with their Kashmiri script
transcriptions. Designed for training and evaluating Automatic Speech
Recognition (ASR) / Speech-to-Text (STT) models for the Kashmiri language.
Dataset Structure
Field
Type
Description
audio
Audio
WAV recording of the Shruk
transcription
string
Kashmiri script transcription
shruk_number
int
Original shruk… See the full description on the dataset page: https://huggingface.co/datasets/Omarrran/100_shruk-speech_to_Text__ASR_dataset.
