datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
Defective_TiresHealthy-Defective-Fruits
Healthy-Defective-Fruits
Dataset Apples Images
The developed data set consists of 5,000 real and synthetic images of fresh apples, and 5,000 real and synthetic images of apples with defects.
Dataset Mangoes Images
The developed data set consists of 5,000 real and synthetic images of fresh mangoes, and 5,000 real and synthetic images of mangoes with defects.
Link dataset: https://github.com/luischuquim/Healthy-Defective-Fruits
For paper reference (Bibtex)… See the full description on the dataset page: https://huggingface.co/datasets/luischuquimarca/Healthy-Defective-Fruits.defective_turret_ru-ljspeech
Defective Turret — Русский (ru)
LJSpeech dataset of Defective Turret (ru).
190 pairs
44.1kHz mono 16-bit PCM WAV (original wiki quality)
Piper TTS Training (High Quality on T4 GPU)
Preprocessing (downsample to 22.05kHz)
python3 -m piper_train.preprocess \
--language ru \
--input-dir ./defective_turret_ru \
--output-dir ./train_defective_turret_ru \
--dataset-format ljspeech \
--single-speaker \
--sample-rate 22050
Training… See the full description on the dataset page: https://huggingface.co/datasets/RoxasYTB/defective_turret_ru-ljspeech.defective_turret_es-ljspeech
Defective Turret — Español (es)
LJSpeech dataset of Defective Turret (es).
190 pairs
44.1kHz mono 16-bit PCM WAV (original wiki quality)
Piper TTS Training (High Quality on T4 GPU)
Preprocessing (downsample to 22.05kHz)
python3 -m piper_train.preprocess \
--language es \
--input-dir ./defective_turret_es \
--output-dir ./train_defective_turret_es \
--dataset-format ljspeech \
--single-speaker \
--sample-rate 22050
Training… See the full description on the dataset page: https://huggingface.co/datasets/RoxasYTB/defective_turret_es-ljspeech.defective_turret_fr-ljspeech
Defective Turret — Français (fr)
LJSpeech dataset of Defective Turret (fr).
175 pairs
44.1kHz mono 16-bit PCM WAV (original wiki quality)
Piper TTS Training (High Quality on T4 GPU)
Preprocessing (downsample to 22.05kHz)
python3 -m piper_train.preprocess \
--language fr \
--input-dir ./defective_turret_fr \
--output-dir ./train_defective_turret_fr \
--dataset-format ljspeech \
--single-speaker \
--sample-rate 22050… See the full description on the dataset page: https://huggingface.co/datasets/RoxasYTB/defective_turret_fr-ljspeech.defective_turret_de-ljspeech
Defective Turret — Deutsch (de)
LJSpeech dataset of Defective Turret (de).
175 pairs
44.1kHz mono 16-bit PCM WAV (original wiki quality)
Piper TTS Training (High Quality on T4 GPU)
Preprocessing (downsample to 22.05kHz)
python3 -m piper_train.preprocess \
--language de \
--input-dir ./defective_turret_de \
--output-dir ./train_defective_turret_de \
--dataset-format ljspeech \
--single-speaker \
--sample-rate 22050
Training… See the full description on the dataset page: https://huggingface.co/datasets/RoxasYTB/defective_turret_de-ljspeech.defective_turret_en-ljspeech
Defective Turret — English (en)
LJSpeech dataset of Defective Turret (en).
190 pairs
44.1kHz mono 16-bit PCM WAV (original wiki quality)
Piper TTS Training (High Quality on T4 GPU)
Preprocessing (downsample to 22.05kHz)
python3 -m piper_train.preprocess \
--language en-us \
--input-dir ./defective_turret_en \
--output-dir ./train_defective_turret_en \
--dataset-format ljspeech \
--single-speaker \
--sample-rate 22050… See the full description on the dataset page: https://huggingface.co/datasets/RoxasYTB/defective_turret_en-ljspeech.Defective
