CoolFace
Datasetpublic

LeeHarrold/musiccaps-mot-tokens

MusicCaps Pre-Encoded Tokens for Mixture-of-Transformers (MoT) Dataset Description This dataset contains pre-encoded audio tokens from the MusicCaps dataset, processed through Meta's MusicGen EnCodec tokenizer for use in Mixture-of-Transformers (MoT) training. Dataset Summary 5,233 music clips encoded as discrete tokens 4 codebook layers from MusicGen's EnCodec ~500 tokens per 10-second clip Compressed from ~12GB audio to 82MB tokens Ready for… See the full description on the dataset page: https://huggingface.co/datasets/LeeHarrold/musiccaps-mot-tokens.

sourceHugging Faceupdated 10mo agoView on Hugging Face
0likes23downloads

Nothing at this path on main. The folder may be empty, or the revision may not exist.

LeeHarrold/musiccaps-mot-tokens · main · files are served by the source, never re-hosted here