jhu-clsp/seamless-align-expressive
Dataset Card for Seamless-Align-Expressive (WIP). Inspired by https://huggingface.co/datasets/allenai/nllb Dataset Summary This dataset was created based on metadata for mined expressive Speech-to-Speech(S2S) released by Meta AI. The S2S contains data for 5 language pairs. The S2S dataset is ~228GB compressed. How to use the data There are two ways to access the data: Via the Hugging Face Python datasets library Scripts coming soon Clone the… See the full description on the dataset page: https://huggingface.co/datasets/jhu-clsp/seamless-align-expressive.
Conversations for this repository live on Hugging Face.
CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.
Open discussions on Hugging Face