jhu-clsp/seamless-align-expressive
Dataset Card for Seamless-Align-Expressive (WIP). Inspired by https://huggingface.co/datasets/allenai/nllb Dataset Summary This dataset was created based on metadata for mined expressive Speech-to-Speech(S2S) released by Meta AI. The S2S contains data for 5 language pairs. The S2S dataset is ~228GB compressed. How to use the data There are two ways to access the data: Via the Hugging Face Python datasets library Scripts coming soon Clone the… See the full description on the dataset page: https://huggingface.co/datasets/jhu-clsp/seamless-align-expressive.
This repository belongs to jhu-clsp on Hugging Face.
CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.
