CoolFace
Datasetpublic

Trelis/cv-en-scripted-test-500

Common Voice English Scripted Test Set — 500 clips n = 500 utterances · private eval set for ASR benchmarking Source Derived from Mozilla Common Voice Scripted Speech 25.0 — English (test split), downloaded via the Mozilla Data Collective API (dataset ID cmndapwry02jnmh07dyo46mot, 94 GB tarball). Construction Starting from the full CV 25.0 English test split (16,398 rows), a stratified 500-clip subset was produced using the same recipe as… See the full description on the dataset page: https://huggingface.co/datasets/Trelis/cv-en-scripted-test-500.

sourceHugging Facecc0-1.0updated 4mo agoView on Hugging Face
0likes25downloads
3 commits on main
7f718b74mo ago

Add dataset card with source attribution and construction details

RonanMcGovern
f4b930a4mo ago

Upload dataset

RonanMcGovern
2c5f37d4mo ago

initial commit

RonanMcGovern