uberdarwin/rvc-soul-coughing-trainer
0
๐ค RVC Soul Coughing Voice Trainer
An automated training pipeline for Soul Coughing vocals using RVC (Retrieval-based Voice Conversion).
Features
- Automated Pipeline: Complete end-to-end training process
- Soul Coughing Dataset: Pre-configured with Soul Coughing acapella files
- Real-time Monitoring: Track training progress and logs
- Easy Interface: Simple Gradio web interface
How to Use
- Upload Audio Files: Use the "Audio Files" tab to upload your training data
- Start Training: Click "๐ Start Full Training Pipeline" in the Training tab
- Monitor Progress: Check the Status tab for training logs and progress
- Download Models: Once training completes, download your trained voice model
Training Process
The trainer automatically handles:
- Audio Preparation: Processes and cleans audio files
- Feature Extraction: Extracts F0, NSF, and ContentVec features
- Model Training: Trains the RVC model with optimal settings
- Progress Monitoring: Provides real-time training updates
Technical Details
- Sample Rate: 40kHz
- F0 Method: Harvest + DIO for pitch extraction
- Feature Extraction: ContentVec for voice features
- Training: Uses pretrained base models for faster convergence
Requirements
- Audio files in WAV format (preferably isolated vocals)
- Minimum 10 minutes of clean vocal audio
- Consistent voice quality across all files
Notes
This Space is configured for Soul Coughing vocal training but can be adapted for other voice models by uploading different audio files.
