CoolFace
Datasetpublicgated

somu9/libritts-clean-v2-tokens

MintTTS Pre-tokenized Audio Tokens Pre-extracted audio codec tokens for TTS training. Source Dataset: mythicinfinity/libritts_r Codec: MOSS-Audio-Tokenizer-Nano Codec sample rate: 48,000 Hz (stereo) Frame rate: 12.5 Hz (1 frame = 80ms) Stats Metric Value Total samples 148,954 Total audio hours 242.6h Codebooks 16 Avg frames/sample 73.3 Avg duration 5.9s Format JSONL file (manifest.jsonl) where each line is: {… See the full description on the dataset page: https://huggingface.co/datasets/somu9/libritts-clean-v2-tokens.

sourceHugging Facemitupdated 5mo agoView on Hugging Face
0likes7downloads

No commit history came back for main. The revision may not exist, or the source declined the request.