datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
TV-24kHz-2025.12-Neutral-FT-Mini
Thorsten-Voice TV-24kHz-2025.12-Neutral-FT-Mini
Overview
This dataset is a small, high-quality fine-tuning dataset created specifically for speaker refinement and voice matching in Orpheus TTS models.
It consists of 60 newly recorded German speech samples, spoken in a neutral, relaxed, everyday style, closely reflecting the natural speaking voice of the original speaker.
This dataset is intended for:
Speaker adaptation and voice refinement
Fine-tuning Orpheus TTS models… See the full description on the dataset page: https://huggingface.co/datasets/Thorsten-Voice/TV-24kHz-2025.12-Neutral-FT-Mini.neutral-batch-ira-2TV-24kHz-Neutral
Thorsten-Voice TV-24kHz-Neutral Dataset
This dataset is a resampled version of the "TV-2022.10-Neutral" configuration from the original Thorsten-Voice TV-44kHz-Full dataset, converted from 44.1kHz to 24kHz sampling rate.
Dataset Description
The Thorsten-Voice dataset contains German speech recordings by Thorsten Müller, suitable for text-to-speech (TTS) training and other speech synthesis tasks.
Changes from Original
Sample Rate: Converted from… See the full description on the dataset page: https://huggingface.co/datasets/Thorsten-Voice/TV-24kHz-Neutral.neutral-cr-batch-2neutral-batch-voice-cloneJBBBenign_Neutralomnivoice-neutral-disfluencyaitf-dfk3-neutral-audios
