CoolFace
Datasetpublic

BrunoHays/english-x-code-switching

Synthetic English Code-Switching Evaluation Set This dataset contains synthetic long-form English code-switching audio samples built from ML-SUPERB hybrid data. Each mixed sample combines English with exactly one additional language. Durations are randomly drawn between 5 and 15 minutes, and each sample contains one or two code switches. The random seed is stored per row. Each selected utterance chunk is RMS-normalized to -20.0 dBFS before concatenation, with peak limiting at… See the full description on the dataset page: https://huggingface.co/datasets/BrunoHays/english-x-code-switching.

sourceHugging Faceotherupdated 5mo agoView on Hugging Face
0likes111downloads
11 commits on main
be847d05mo ago

Upload README.md with huggingface_hub

BrunoHays
0f0a7585mo ago

Upload dataset

BrunoHays
e72d9845mo ago

Upload README.md with huggingface_hub

BrunoHays
adc06655mo ago

Upload dataset

BrunoHays
856fb055mo ago

Restore dataset README

BrunoHays
2a3ff1f5mo ago

Upload dataset

BrunoHays
9b069675mo ago

Remove raw folder upload before dataset-format push

BrunoHays
17a93d05mo ago

Add files using upload-large-folder tool

BrunoHays
b12aa815mo ago

Add files using upload-large-folder tool

BrunoHays
4b553025mo ago

Add files using upload-large-folder tool

BrunoHays
2130c9f5mo ago

initial commit

BrunoHays