EmpathicRobotics/voiceclap-flattened
voiceclap-flattened SNAC-tokenized, flattened + augmented build of 5 subsets of laion/voiceclap-data (CC-BY-4.0): emolia (English block), ears, expresso, voxceleb1, voxceleb2. Replaces the earlier single-subset voiceclap-emolia-flattened repo (consolidated here). Why this exists Follow-up to an ablation study (2/3/5) on <listen>/<speak> token format: full 7-tok/frame <listen> content with a SEPARATE vocab (ids shifted +1,000,000 vs <speak>'s identical SNAC codes)… See the full description on the dataset page: https://huggingface.co/datasets/EmpathicRobotics/voiceclap-flattened.
0353
