datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
swahili_MED_otherSwahilidata_11swahili_MED_otherSwahilidata_22swahili_MED_otherSwahilidata_44speech_accent_archive_otherlibrispeech_test_other@inproceedings{panayotov2015librispeech,
title={Librispeech: an asr corpus based on public domain audio books},
author={Panayotov, Vassil and Chen, Guoguo and Povey, Daniel and Khudanpur, Sanjeev},
booktitle={2015 IEEE international conference on acoustics, speech and signal processing (ICASSP)},
pages={5206--5210},
year={2015},
organization={IEEE}
}
@article{wang2024audiobench,
title={AudioBench: A Universal Benchmark for Audio Large Language Models},
author={Wang, Bin and… See the full description on the dataset page: https://huggingface.co/datasets/AudioLLMs/librispeech_test_other.Bandicam-Videos-And-Otherslirispeech_asr_train_othercommon_voice_16_1_gn_other_predicted_modelo_1swahili_MED_otherSwahilidata_mubargmax_librispeech_otherKsponSpeech_eval_otherlibri_val_othercommon_voice_16_1_gn_other_predictedlibri_test_otherswahili_MED_otherSwahilidata_33voiceprint_librispeech_other_testThis dataset select audios from librispeech other test dataset with distinct speaker_id, which can be used as voiceprint input for text-to-speech tasks.
lirispeech_asr_dev_otherglados_other_de-ljspeech
GLaDOS (Other) — Deutsch (de)
LJSpeech dataset of GLaDOS (Other) (de).
134 pairs
44.1kHz mono 16-bit PCM WAV (original wiki quality)
Piper TTS Training (High Quality on T4 GPU)
Preprocessing (downsample to 22.05kHz)
python3 -m piper_train.preprocess \
--language de \
--input-dir ./glados_other_de \
--output-dir ./train_glados_other_de \
--dataset-format ljspeech \
--single-speaker \
--sample-rate 22050
Training (Kaggle T4… See the full description on the dataset page: https://huggingface.co/datasets/RoxasYTB/glados_other_de-ljspeech.above_70yo_elderly_people_other_dataset
Dataset Card for "above_70yo_elderly_people_other_dataset"
More Information needed
IndicVoices_Hindi_audio_44100_30_45_otherAUDIO-SRX-OTHER-LANGUAGE
SKT AI LABS
SKT AI LABS
The Sovereign AI for India
The Sovereign LLM Development For India (Project Surya)
IndicVoices_Hindi_audio_44100_18_30_otherLS-dev-other-mbr-vocals-fv4-gaboxglados_other_ru-ljspeech
GLaDOS (Other) — Русский (ru)
LJSpeech dataset of GLaDOS (Other) (ru).
1487 pairs
44.1kHz mono 16-bit PCM WAV (original wiki quality)
Piper TTS Training (High Quality on T4 GPU)
Preprocessing (downsample to 22.05kHz)
python3 -m piper_train.preprocess \
--language ru \
--input-dir ./glados_other_ru \
--output-dir ./train_glados_other_ru \
--dataset-format ljspeech \
--single-speaker \
--sample-rate 22050
Training (Kaggle T4… See the full description on the dataset page: https://huggingface.co/datasets/RoxasYTB/glados_other_ru-ljspeech.lirispeech_asr_test_otherkeyword_spotting_othersglados_other_en-ljspeech
GLaDOS (Other) — English (en)
LJSpeech dataset of GLaDOS (Other) (en).
597 pairs
44.1kHz mono 16-bit PCM WAV (original wiki quality)
Piper TTS Training (High Quality on T4 GPU)
Preprocessing (downsample to 22.05kHz)
python3 -m piper_train.preprocess \
--language en-us \
--input-dir ./glados_other_en \
--output-dir ./train_glados_other_en \
--dataset-format ljspeech \
--single-speaker \
--sample-rate 22050
Training (Kaggle… See the full description on the dataset page: https://huggingface.co/datasets/RoxasYTB/glados_other_en-ljspeech.openslr_other
