CoolFace
Datasetpublic

BrunoHays/english-x-code-switching-samples

Synthetic English Code-Switching Evaluation Set Samples This dataset contains the individual normalized utterance chunks used to build the paired mixed dataset. Each mixed sample combines English with exactly one additional language. Durations are randomly drawn between 5 and 15 minutes, and each sample contains one or two code switches. The random seed is stored per row. Each selected utterance chunk is RMS-normalized to -20.0 dBFS before concatenation, with peak limiting at… See the full description on the dataset page: https://huggingface.co/datasets/BrunoHays/english-x-code-switching-samples.

sourceHugging Faceotherupdated 5mo agoView on Hugging Face
0likes41downloads
discussions and pull requests

Conversations for this repository live on Hugging Face.

CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.

Open discussions on Hugging Face
BrunoHays/english-x-code-switching-samples · CoolFace