datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
risale-i-nur-sohbet
Risale-i Nur Sohbet
Prof. Dr. Şener Dilek’ten izin alındı.
Türkçe
Risale-i Nur sohbetlerini ses, ham ASR metni ve zaman hizalı segmentler hâlinde
birlikte sunan bağımsız bir veri kümesidir. İlk sürüm izinli ve doğrulanmış
sohbetleri içerir; kitap metni, grounded, çok dilli veya kitap seslendirme veri
kümelerine karıştırılmaz.
Kapsam
2095 sohbet, 954.66 saat 16 kHz mono FLAC ses
Aynı derslerin ölçülmüş 48 kHz kalite katmanı; 786 derste
seçici… See the full description on the dataset page: https://huggingface.co/datasets/risaleinur/risale-i-nur-sohbet.open-large-bengali-asr-data
Open Large Bengali ASR Data
This is a collection of publicly available ASR data for Bengali. It contains 5000 hours of audio. We have a filtering column called is_better to filter good-quality audio from the corpus. It is set based on the wer between original transcription and prediction taken from a Bengali-Wav2Vec2 model and word-per-second (wps).
Datasets:
commonvoice
risale-nur-audio
Risale-i Nur Audio–Text Corpus
Gerçek insan okumalarını, aynı satırdaki kaynak metinle birlikte sunan açık bir
ses–metin veri kümesidir. Yeni varsayılan audio-text yapılandırması 15 kitaptan
91.792 oynatılabilir klip ve 203,02 saat ses içerir. Metinler kanonik kaynaktan
değiştirilmeden alınır ve her kayıt byte-exact section_id alıntılarıyla bağlanır.
An open speech corpus pairing human readings with their source text in the same
row. The default audio-text config contains 91,792… See the full description on the dataset page: https://huggingface.co/datasets/risaleinur/risale-nur-audio.risalei-nur-text-audio
Risale-i Nur Text–Audio
Kaynak · Source: RNK Neşriyat — yazılı izinle · used with written permission.
Her satırda gerçek insan okuması ile o sesin kanonik metni birlikte bulunur.
Sesler dış bağlantı değildir: WAV baytları Parquet dosyalarının içindedir.
Kaynak sitesi veya başka bir ses sunucusu gerekmez.
Each row pairs a human reading with its canonical transcript. Audio is stored
as WAV bytes inside the Parquet files; no source website or external audio
server is required.… See the full description on the dataset page: https://huggingface.co/datasets/risaleinur/risalei-nur-text-audio.nepali_asr
Dataset Card for Nepali Asr Dataset
Dataset Summary
This dataset consists of over 5 hours (300+ minutes) of English speech audio collected from YouTube. The dataset is designed for automatic speech recognition (ASR) and speaker identification tasks. It features both male and female speakers, with approximately 60% of the samples from male voices and the remaining 40% from female voices. The dataset contains 35 distinct speakers, each with their audio segmented into… See the full description on the dataset page: https://huggingface.co/datasets/rishi70612/nepali_asr.validation_nepali_asr
Dataset Card for Dataset Name
This dataset card aims to be a base template for new datasets. It has been generated using this raw template.
Dataset Details
Dataset Description
Curated by: [More Information Needed]
Funded by [optional]: [More Information Needed]
Shared by [optional]: [More Information Needed]
Language(s) (NLP): [More Information Needed]
License: [More Information Needed]
Dataset Sources [optional]
Repository: [More… See the full description on the dataset page: https://huggingface.co/datasets/rishi70612/validation_nepali_asr.
