jhu-clsp/seamless-align-expressive
Dataset Card for Seamless-Align-Expressive (WIP). Inspired by https://huggingface.co/datasets/allenai/nllb Dataset Summary This dataset was created based on metadata for mined expressive Speech-to-Speech(S2S) released by Meta AI. The S2S contains data for 5 language pairs. The S2S dataset is ~228GB compressed. How to use the data There are two ways to access the data: Via the Hugging Face Python datasets library Scripts coming soon Clone the… See the full description on the dataset page: https://huggingface.co/datasets/jhu-clsp/seamless-align-expressive.
521
Update README.md
Update README.md
README
README
README
Create README.md
enA-frA
enA-itA files
enA-zhA files
enA-esA files
deA-enA files
initial commit
